ls /news/large-language-models · home newslarge-language-models
grep -r --recent /news/large-language-models | head -20

Large Language Model News

Large language model (LLM) news — GPT-4, Claude, Gemini, Llama, Mistral and the latest research on training, fine-tuning, RLHF, and deployment of LLMs.

21229 articles page 776 of 1062 0 sources 30 min sync cycle updated 2026-06-15

// latest articles 21229 indexed

04:00
2026-06-15
arxiv.org
large-language-models · 1m read · neu

The Culture Funnel: You Can't Align What isn't in the Data

Researchers at CohereLabs argue that large language models suffer from a 'cultural data funnel,' where cultural signals decline sharply during post-training while geographically concentrated data dominates. They release …

04:00
2026-06-15
arxiv.org
neural-networks · 1m read · neu

The Weight Norm Sets the Grokking Timescale: A Causal Delay Law

Researchers at arXiv have causally demonstrated that the weight norm sets the grokking timescale in neural networks, settling a dispute over whether weight norm causes the delayed generalization. By intervening on the no…

04:00
2026-06-15
arxiv.org
machine-learning · 1m read ↑ pos

Diffusion Policy Optimization without Drifting Apart

Researchers identified the double-drift phenomenon causing instability in diffusion policy-gradient methods and proposed DiPOD, a framework that interleaves self-distillation with policy-improving gradient updates to mai…

04:00
2026-06-15
arxiv.org
machine-learning · 1m read · neu

Uncertainty Estimation and Generalization Bounds for Modern Deep Learning

A new thesis investigates how Bayesian principles can improve understanding of modern deep learning systems, introducing the Deep Variational Implicit Process (DVIP) and post-hoc methods VaLLA and FMGP for uncertainty es…

← prev page 776 / 1062 next →
LIVE [news/large-language] indexed:21229 page:776/1062 en · ua 2026-05-20 ·