cd/entity/Fleiss' kappaΒ· homeβ€Ί entitiesβ€Ί Fleiss' kappa
grep -l @fleiss' kappa /news/*.json | wc -l β†’ 1

Fleiss' kappa

mentions 1 type Person feed RSS

// recent coverage 1 mentions

14:49
2026-07-22
arize.com
large-language-models

How to measure human-LLM judge alignment

A new guide breaks down evaluation alignment between human experts and LLM judges into three measurable questions: human reliability on the rubric, LLM-human agreement relative to human-human agreemen…

// co-occurs with top 4 entities
// topics top 2 topics