cd/entity/MMLU-Pro· home entities MMLU-Pro
grep -l @mmlu-pro /news/*.json | wc -l → 13

MMLU-Pro

mentions 13 type Organization feed RSS

// recent coverage 13 mentions

04:00
2026-08-17
machinebrief.com
artificial-intelligence

KV Cache Compression Through the Lens of Transform Coding

Researchers introduced Attention-Aware Transform Coding (AATC), a KV cache compression method that reduces memory use by approximately 5.8x while maintaining near-lossless accuracy on models like Llam…

05:28
2026-08-13
lesswrong.com
ai-safety

LLMs have the capacity for self-imposed steganography

A BlueDot Impact Technical AI Safety project demonstrated that large language models (LLMs) can be trained to perform steganography, embedding hidden information in their outputs to evade monitoring. …

11:01
2026-07-23
promptcube3.com
artificial-intelligence

Gemma 4 vs Gemini 3.1 Flash-Lite: The Hybrid Win

Google's Gemma 4 model uses a 68k-parameter probe layer that predicts decoding errors by reading hidden states, achieving a 0.79-0.88 AUROC on audio benchmarks despite being trained on zero audio data…

04:00
2026-07-13
arxiv.org
large-language-models

HALO: Hybrid Adaptive Latent Reasoning for Language Models

A new method called HALO (Hybrid Adaptive Latent Reasoning) improves frozen pretrained language models with adaptive extra computation, achieving the best overall average on MMLU-Pro and GPQA-Diamond …

11:24
2026-06-04
huggingface.co
artificial-intelligence

Task-Seeded Synthetic Q&A Generation for Nemotron Pretraining

NVIDIA researchers developed a task-seeded synthetic Q&A generation workflow for Nemotron-family pretraining that uses public task training splits as capability seeds to generate new task-aligned exam…

// co-occurs with top 8 entities
// topics top 6 topics