ls /news/large-language-models · home newslarge-language-models
grep -r --recent /news/large-language-models | head -20

Large Language Model News

Large language model (LLM) news — GPT-4, Claude, Gemini, Llama, Mistral and the latest research on training, fine-tuning, RLHF, and deployment of LLMs.

19985 articles page 604 of 1000 0 sources 30 min sync cycle updated 2026-06-20

// latest articles 19985 indexed

07:07
2026-06-20
dev.to
artificial-intelligence · 1m read · neu

AI Agents Explained: the Thought-Action-Observation Loop

An engineer demonstrates how AI agents use a Thought-Action-Observation loop to solve multi-step tasks by calling tools like calculators and search. The agent iterates until completion, with each real observation fed bac…

07:06
2026-06-20
dev.to
large-language-models · 1m read · neu

Temperature and Sampling: the LLM Creativity Dial

A developer created an interactive tool demonstrating how temperature and sampling parameters control the creativity and reproducibility of large language model outputs. The tool shows that low temperature (near 0) produ…

07:05
2026-06-20
dev.to
large-language-models · 1m read · neu

The Context Window: an LLM's Short-Term Memory, Explained

A developer explains that large language models (LLMs) are stateless and their 'memory' is limited to a fixed context window. When the window fills, the oldest messages are dropped and cannot be recalled. The post demons…

06:58
2026-06-20
news.ycombinator.com
large-language-models · 1m read · neu

Ask HN: Will we start seeing tools for LLM use?

A Hacker News user asks whether the community will develop tools that structure command-line output specifically for LLM consumption, noting that current compression methods save tokens but may increase interaction turns…

05:57
2026-06-20
dev.to
large-language-models · 7m read · neu

Nobody Knows Why It Said That

An engineer building with AI daily reveals that even the creators of large language models at Anthropic, Google DeepMind, and OpenAI do not fully understand why their models work. The post explains that LLMs encode knowl…

04:45
2026-06-20
narracomm.com
artificial-intelligence · 2m read · neu

High End Brand

Anthropic confirmed deprecation timelines for several Claude models, with developers needing to migrate before older 4-series models are retired. xAI's Colossus 2 supercluster in Memphis is now operational with GB200/GB3…

04:33
2026-06-20
tomshardware.com
artificial-intelligence · 1m read · neu

China will have a Fable 5-class AI model before next year

China is expected to release a Fable 5-class AI model before next year, according to industry sources. The development signals China's accelerating progress in artificial intelligence and its ambition to compete with lea…

← prev page 604 / 1000 next →
LIVE [news/large-language] indexed:19985 page:604/1000 en · ua 2026-05-20 ·