04:00
2026-08-12
machinebrief.com
artificial-intelligence
Toward Human Rights Benchmarking for LLMs: A Pilot Methodology
Researchers have developed HumRightsBench, the first expert-validated, scenario-based benchmark for evaluating whether large language models can reason correctly about international human rights law, โฆ