00:00
2026-07-13
aclanthology.org
large-language-models
When Do LLMs Need Human Experts? Evidence for Social Science from Jurisprudential Classification
Even frontier large language models such as GPT-5.2 and leading open-weight alternatives consistently underperform fine-tuned BERT on a challenging legal reasoning task, according to a study by Caroliβ¦