Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing

  • Log In
  • Sign Up

EvalEval Coalition

Team
community
https://evalevalai.com/
evaluatingevals
evaleval
Activity Feed Request to join this org

AI & ML interests

We’re building a research coalition on evaluating evaluations (EvalEval)! Hosted by Hugging Face, University of Edinburgh, and EleutherAI.

Recent Activity

evijit  new activity about 6 hours ago
evaleval/general-eval-card:Harmonize the reproducibility-fields allowlist between the signal and the per-row card
j-chim  updated a bucket about 10 hours ago
evaleval/entity-registry-storage
evijit  updated a bucket about 22 hours ago
evaleval/general-eval-card-storage
View all activity

Articles

AI evals are becoming the new compute bottleneck

7 days ago
•
26

Yacine Jernite's profile pictureIrene Solaiman's profile pictureCanyu Chen's profile pictureFelix Friedrich's profile pictureAlina Leidinger's profile pictureMargaret Mitchell's profile pictureJennifer Mickel's profile pictureUsman Gohar's profile pictureLevent Sagun's profile pictureShubham Singh's profile pictureAvijit Ghosh's profile pictureLeshem Choshen's profile pictureAurélien-Morgan CLAUDON's profile pictureAmita Shukla's profile picturePrajna Soni's profile pictureAnshuman Suri's profile pictureJoseph [open/acc] Pollack's profile pictureMowafak Allaham's profile picturewave's profile pictureAli El Filali's profile pictureAndrew Tran's profile pictureMonojit's profile pictureKevin Wei's profile pictureJan Batzner's profile pictureJenny Chim's profile pictureMubashara Akhtar's profile pictureSree Harsha Nelaturu's profile pictureHossein A. (Saeed) Rahmani's profile pictureAbdul Muhsin Hameed's profile pictureSrishti's profile pictureJoshua Noble's profile pictureEvalEval Bot's profile pictureDamian Stachura's profile pictureŠimon Podhajský's profile pictureAnastassia Kornilova's profile pictureInge V's profile pictureAris's profile pictureSriram Mohan's profile pictureTommaso Cerruti's profile pictureImamaShehzad's profile pictureMarek Suppa's profile pictureYifan Mai's profile pictureGeorgia Channing's profile pictureAsaf Yehudai's profile pictureHarsh's profile picture

evaleval 's Spaces 4

Running
4

Eval Cards

📋

Standardized evaluation cards for AI models and benchmarks

about 2 hours ago
Running

eval-card-registry

🗂

1 day ago
Running

BenchmarkCard Webhook

📋

Receive and process benchmark data via webhook

6 days ago
Running

README

🤗

Sep 9, 2025
Company
TOS Privacy About Careers
Website
Models Datasets Spaces Pricing Docs