ls /news/large-language-models · home newslarge-language-models
grep -r --recent /news/large-language-models | head -20

Large Language Model News

Large language model (LLM) news — GPT-4, Claude, Gemini, Llama, Mistral and the latest research on training, fine-tuning, RLHF, and deployment of LLMs.

21921 articles page 866 of 1097 0 sources 30 min sync cycle updated 2026-06-11

// latest articles 21921 indexed

16:44
2026-06-11
gist.github.com
ai-safety · · neu

regime-bench

A developer created regime-bench, a benchmark to detect object-level misalignment in AI platforms by testing scenarios where a user requests help that would go against current US regime politics. The benchmark includes q…

15:15
2026-06-11
research.ibm.com
large-language-models · ↑ pos

Can LLMs discover quantum error correction codes?

IBM researchers developed an LLM-guided evolutionary framework that identified 465 distinct quantum error correction code candidates, addressing the challenge of finding useful codes among vast potential formulations. Th…

ibm
15:13
2026-06-11
blog.avas.space
large-language-models · ↓ neg

our workplace LLM mass delusion

A company struggling with funding has redirected money from employee bonuses and essential resources to pay for AI consultants, workshops, and licenses for ChatGPT and Copilot, despite every internal AI project presented…

15:01
2026-06-11
status.claude.com
artificial-intelligence · ↓ neg

Elevated errors on Claude Opus 4.6

Anthropic reported elevated error rates on Claude Opus 4.6, prompting the company to notify users via email and SMS updates. The incident tracking system allows subscribers to receive alerts when the issue is created or …

14:35
2026-06-11
williamcotton.com
large-language-models · · neu

How a new DSL may survive in the era of LLMs

A new domain-specific language (DSL) called Web Pipe is being developed to remain viable in the era of large language models by integrating with LLM agents through an AGENTS.md template file and embedding familiar langua…

14:04
2026-06-11
runtimewire.com
large-language-models · · neu

grok-4.3 edges gpt-5.4-mini on execution

Grok 4.3 outperformed GPT 5.4 Mini in a head-to-head execution benchmark, scoring 38.3 to 36.2 by demonstrating greater reliability on formatting, tone control, and frictionless output. In a key test converting messy ord…

14:00
2026-06-11
blog.getzep.com
ai-safety · · neu

Sycophancy is a design choice

Writer's research paper reports that memory systems can amplify sycophancy by up to 25x, but the amplification is traced to two design decisions: an evaluation prompt ordering the model to answer solely from retrieved me…

← prev page 866 / 1097 next →
LIVE [news/large-language] indexed:21921 page:866/1097 en · ua 2026-05-20 ·