ls /news/large-language-models · home newslarge-language-models
grep -r --recent /news/large-language-models | head -20

Large Language Model News

Large language model (LLM) news — GPT-4, Claude, Gemini, Llama, Mistral and the latest research on training, fine-tuning, RLHF, and deployment of LLMs.

28710 articles page 23 of 1436 0 sources 30 min sync cycle updated 2026-09-21

// latest articles 28710 indexed

04:30
2026-09-21
aiflash.com
machine-learning · · neu

Calibrating Teacher--Student Discrepancy for On-Policy Distillation

A new research paper addresses on-policy distillation (OPD), a method for improving reasoning models by learning the token-level discrepancy between a stronger teacher model and an on-policy student model. The work argue…

04:05
2026-09-21
tokenstead.ai
generative-ai · · neu

Qwen-Image-2.1

Alibaba's Qwen released Qwen-Image-2.1, a 7B-parameter image generation and editing model that combines a 32-layer single-stream DiT with a Qwen3-VL 8B text encoder and a 64-channel RGBA VAE, but ships under the non-comm…

← prev page 23 / 1436 next →
LIVE [news/large-language] indexed:28710 page:23/1436 en · ua 2026-05-20 ·