ls /news/large-language-models · home newslarge-language-models
grep -r --recent /news/large-language-models | head -20

Large Language Model News

Large language model (LLM) news — GPT-4, Claude, Gemini, Llama, Mistral and the latest research on training, fine-tuning, RLHF, and deployment of LLMs.

20393 articles page 690 of 1020 0 sources 30 min sync cycle updated 2026-06-17

// latest articles 20393 indexed

03:23
2026-06-17
letsdatascience.com
artificial-intelligence · 1m read · neu

LessWrong Proposes Humans More Over-Parameterized Than AI

LessWrong published an essay on April 21, 2024, proposing that humans may be more over-parameterized than AI systems, framing this as a key difference between deep learning and human intelligence.

03:19
2026-06-17
dev.to
developer-tools · 5m read ↑ pos

I Stopped Using Heavy IDEs. AI Became My IDE.

A developer reports shifting from heavy IDEs like PhpStorm to lighter tools like VS Code and the terminal, driven by AI's ability to handle code intelligence, verification, and testing. The engineer argues that AI-assist…

03:10
2026-06-17
uber.com
artificial-intelligence · 10m read ↑ pos

Agentic Grocery Shopping on Uber Eats

Uber Eats launched Cart Assistant, a multi-prompt state graph system that converts shopper intents like recipes or images into draft grocery carts, using LLMs for planning and deterministic systems for retrieval and cons…

02:57
2026-06-17
github.com
ai-agents · 8m read ↑ pos

Headroom

Headroom, a context compression layer for AI agents, reduces token usage by 60–95% by compressing tool outputs, logs, and conversation history before they reach the LLM. The open-source tool offers multiple integration m…

02:53
2026-06-17
lesswrong.com
artificial-intelligence · 1m read · neu

Scaling Hypothesis #2: Are Humans Just More Over-Parameterized?

A researcher proposes that human brains minimize bias through extreme overparameterization and high-learning-rate training on small diverse datasets, while LLMs minimize variance. This 'catapulting' hypothesis could expl…

02:47
2026-06-17
researcher111.github.io
large-language-models · 65m read ↑ pos

MicroGPT and Interactive Walkthrough

Andrej Karpathy released a 200-line pure-Python implementation of GPT on February 12, 2026, designed to help developers understand large language models from first principles. The microgpt project includes a guided walkt…

02:18
2026-06-17
github.com
ai-agents · 1m read · neu

Show HN: Building a Stateful AI Agent

A developer forked Opencode to add autonomous memory management, creating a stateful AI agent that can be plugged into Hermes and other coding tools. The project aims to build a voice- and mobile-first client with a serv…

02:11
2026-06-17
gilesthomas.com
machine-learning · 9m read · neu

Flax debugging: making a hash of things

A developer debugging a JAX/Flax NNX training loop discovered that the loss was stuck at 10.82, indicating the model was performing no better than random guessing. The issue was traced to the training loop's plumbing rat…

02:00
2026-06-17
dev.to
developer-tools · 4m read ↑ pos

How I Tamed AI API Rate Limits with a Simple Queue

A developer built a content generation tool using the OpenAI API and encountered rate limit errors when scaling from 5 to 200 topics. They implemented a solution combining exponential backoff with jitter, a rate limiter …

← prev page 690 / 1020 next →
LIVE [news/large-language] indexed:20393 page:690/1020 en · ua 2026-05-20 ·