04:00
2026-08-11
arxiv.org
artificial-intelligence
TREAT: Evaluating Access to Formal Knowledge across Equivalent Mathematical Representations
Researchers introduced TREAT, a benchmark with 737 theorem identities and 29,480 transformed rows, to test whether large language models can recognize known theorems from equivalence-preserving formulβ¦