ls /news/large-language-models · home newslarge-language-models
grep -r --recent /news/large-language-models | head -20

Large Language Model News

Large language model (LLM) news — GPT-4, Claude, Gemini, Llama, Mistral and the latest research on training, fine-tuning, RLHF, and deployment of LLMs.

21669 articles page 843 of 1084 0 sources 30 min sync cycle updated 2026-06-12

// latest articles 21669 indexed

06:44
2026-06-12
marginalrevolution.com
ai-research · 2m read · neu

Again, the paper format will be dying out

A new paper co-authored by 37 researchers from Stanford, Carnegie Mellon, and the University of Michigan argues that the traditional paper format is obsolete in the AI era, citing "narrative tax" and "engineering tax" as…

06:37
2026-06-12
gist.github.com
large-language-models · 9m read · neu

Token optimization protocol for Claude Fable and other high-capability / high-cost models. Apply this skill at the START of any session involving iterative code builds, multi-file projects, design…

A developer has created "fable-economy," a token optimization protocol for high-capability AI models like Claude Fable that aims to reduce output token spend by 40–70% without sacrificing quality. The protocol enforces f…

06:26
2026-06-12
github.com
large-language-models · 1m read ↑ pos

Doclang-Project/Doclang

The DocLang Project has released DocLang, an AI-native markup format for unstructured content that maps to LLM tokens while preserving structure, semantics, layout, and geometry. The repository hosts the normative specif…

06:22
2026-06-12
dev.to
large-language-models · 9m read · neu

8GB to 70B: A Real Hardware Guide for Local LLMs

A developer found that running a 70B parameter LLM locally with only 8GB of VRAM is possible but requires significant optimizations and trade-offs. While a 70B model in FP16 format demands up to 140GB of VRAM, 4-bit quan…

05:56
2026-06-12
github.com
large-language-models · 6m read ↑ pos

LLM for the ESP32-S3

Two ESP32-S3 microcontrollers running a Llama-architecture language model have achieved the first multi-chip pipelined LLM inference on ESP32-class hardware, splitting layers across two boards connected by three jumper w…

05:17
2026-06-12
lesswrong.com
ai-safety · 1m read · neu

PSA: Almost nobody is working on alignment

A large fraction of the AI safety community is not working on alignment, the effort to ensure superintelligent AIs follow human values and instructions. Most researchers instead focus on indirect work such as capability …

← prev page 843 / 1084 next →
LIVE [news/large-language] indexed:21669 page:843/1084 en · ua 2026-05-20 ·