cd/entity/arXiv· home› entities› arXiv
grep -l @arxiv /news/*.json | wc -l → 3626

arXiv

mentions 3626 type Organization page 61/182 feed RSS

// recent coverage 3626 mentions

04:00
2026-09-11
machinebrief.com
large-language-models

Negative Self-Distillation: Learning to Reason by Avoiding Flaws

A new arXiv paper (2609.11699v1) introduces Negative Self-Distillation (NSD), a framework that improves large language model reasoning by diverging from self-generated flawed reasoning rather than imi…

01:22
2026-09-11
arxiv.org
large-language-models

Sky sphere representation in language models (29 Jul 2026)

A 29 July 2026 arXiv paper by Aleksandr Berdnikov reports that most open-source language models of roughly 100B parameters hold a decodable representation of the night sky map in their residual stream…

15:03
2026-09-10
amazon.science
machine-learning

Why don’t machine learning research agents overfit?

A new paper, "What fits (into few tokens) doesn't overfit: Compression and generalization in ML research agents," argues that LLM-based research agents avoid overfitting on heavily reused benchmarks b…

12:57
2026-09-10
devnavigator.com
artificial-intelligence

AI Verifiers: 3 Powerful Layers Making Autonomous AI More Reliable

A framework called LLM-as-a-Verifier evaluates complete agent trajectories without additional model training by computing an expected score from the probability distribution across possible scoring to…

12:53
2026-09-10
arxiv.org
large-language-models

What Makes Rotary Positional Encodings Useful?

A paper by Federico Barbero and co-authors, revised 13 May 2025 as v3 on arXiv, argues that Rotary Positional Encodings (RoPE) in Transformer-based large language models are not primarily useful for d…

← prev page 61 / 182 next →
// co-occurs with top 8 entities
// topics top 6 topics