cd/entity/Cohen's kappa· home entities Cohen's kappa
grep -l @cohen's kappa /news/*.json | wc -l → 2

Cohen's kappa

mentions 2 type Person feed RSS

// recent coverage 2 mentions

14:49
2026-07-22
arize.com
large-language-models

How to measure human-LLM judge alignment

A new guide breaks down evaluation alignment between human experts and LLM judges into three measurable questions: human reliability on the rubric, LLM-human agreement relative to human-human agreemen…

// co-occurs with top 7 entities
// topics top 5 topics