One Schema for Every Eval
The EvalEval Coalition has launched a unified, open data format and public dataset for AI evaluation results, aiming to address fragmentation and enable trust and comparability across frameworks. The …
The EvalEval Coalition has launched a unified, open data format and public dataset for AI evaluation results, aiming to address fragmentation and enable trust and comparability across frameworks. The …
At ACM FAccT 2026 in Montreal, Mozilla.ai demonstrated that AI guardrails require the same rigorous evaluation as the models they govern, presenting a tutorial on contextual evaluation of LLM guardrai…
The EvalEval Coalition announced new grant support from Founders Pledge (Global Catastrophic Risks Fund), the Survival and Flourishing Fund, and the Weizenbaum Project funded by the German Federal Min…
Hugging Face and the EvalEval Coalition launched an integration that allows contributors to submit standardized evaluation results (EEE schema) to Hugging Face Community Evals, consolidating scattered…