Introducing Evaluation Cards: A Live Interpretive Layer for Understanding the AI Evaluations Ecosystem
• 1
We’re building a research coalition on evaluating evaluations (EvalEval)! Hosted by Hugging Face, University of Edinburgh, and EleutherAI.
Every Eval Ever: A Unifying Schema and Community Repository for AI Evaluation Results
Evaluation Cards: An Interpretive Layer for AI Evaluation Reporting
Standardized evaluation cards for AI models and benchmarks
Agentic search over the EEE datastore for your use case
Summarize schema discussions and add comments
Receive and process benchmark data via webhook