cd/entity/MATH· home entities MATH
grep -l @math /news/*.json | wc -l → 8

MATH

mentions 8 type Organization feed RSS

// recent coverage 8 mentions

05:53
2026-07-10
github.com
artificial-intelligence

TinyToT – Tree of Thoughts Inference Server

TinyToT, a lightweight inference server compatible with Ollama, achieves 97% accuracy on a 35-question benchmark spanning graduate-level science, medicine, law, finance, and software engineering witho…

16:07
2026-06-17
danlevy.net
large-language-models

LLM benchmarks are answering someone else's question

LLM benchmarks like MMLU and HumanEval are irrelevant for most businesses building AI products, as they measure generic performance rather than specific system tasks. Teams should instead build custom…

// co-occurs with top 8 entities
// topics top 6 topics