ls /news/large-language-models · home newslarge-language-models
grep -r --recent /news/large-language-models | head -20

Large Language Model News

Large language model (LLM) news — GPT-4, Claude, Gemini, Llama, Mistral and the latest research on training, fine-tuning, RLHF, and deployment of LLMs.

28721 articles page 24 of 1437 0 sources 30 min sync cycle updated 2026-09-21

// latest articles 28721 indexed

04:30
2026-09-21
aiflash.com
machine-learning · · neu

Calibrating Teacher--Student Discrepancy for On-Policy Distillation

A new research paper addresses on-policy distillation (OPD), a method for improving reasoning models by learning the token-level discrepancy between a stronger teacher model and an on-policy student model. The work argue…

04:05
2026-09-21
tokenstead.ai
generative-ai · · neu

Qwen-Image-2.1

Alibaba's Qwen released Qwen-Image-2.1, a 7B-parameter image generation and editing model that combines a 32-layer single-stream DiT with a Qwen3-VL 8B text encoder and a 64-channel RGBA VAE, but ships under the non-comm…

← prev page 24 / 1437 next →
LIVE [news/large-language] indexed:28721 page:24/1437 en · ua 2026-05-20 ·