GLM 5.3 Flash unusably slow
Users of GLM 5.3 Flash reported the model has become unusably slow, with response times of at least 3 minutes for simple questions, according to complaints posted on a discussion thread. Commenter shurik said the delay o…
Large language model (LLM) news — GPT-4, Claude, Gemini, Llama, Mistral and the latest research on training, fine-tuning, RLHF, and deployment of LLMs.
Users of GLM 5.3 Flash reported the model has become unusably slow, with response times of at least 3 minutes for simple questions, according to complaints posted on a discussion thread. Commenter shurik said the delay o…
Enterprise AI observability platforms are emerging as centralized telemetry, evaluation, and governance infrastructure for monitoring non-deterministic LLM applications and multi-agent workflows in production. The guide …
OpenAI launched ChatGPT inside Microsoft Word on September 17th, making the add-in available across every ChatGPT plan including Free, subject to usage limits, and extending the same Microsoft add-in to PowerPoint and Ex…
OpenAI told the New York Times it has made "substantial progress" on a second Millennium Prize Problem after claiming a Navier-Stokes solution on September 8, but the company has not named the problem or given a disclosu…
Austin Lyons told the TechSurge podcast that AI infrastructure buyers are shifting from commoditized server hardware to full rack-scale systems because frontier LLM inference requires terabytes of memory that cannot fit …
A developer has published a beginner-level explainer on large language models, covering core concepts such as weights, parameters, transformer architecture, tokenization, context windows, and sampling controls like tempe…
A developer built a controlled experiment testing whether explicitly stated relational context between documents improves retrieval-augmented generation reasoning when retrieval is held constant. Using a synthetic corpus…
Microsoft AI chief Mustafa Suleyman warned in a BBC interview and a personal essay that training AI systems to act autonomously could seed a new "silicon species" that competes with humans for resources, calling rival An…
Developer Rijul built a demonstration project showing how knowledge poisoning attacks can manipulate retrieval-augmented generation (RAG) systems by planting false, malicious, or misleading documents in a knowledge base.…
A developer built Caret, an experimental functional programming language, by iteratively directing OpenAI's Codex and ChatGPT through a process that grew from a quick prototype into a full workflow with specifications, a…
A developer's technical lesson explains that large language models are stateless prediction machines with no built-in memory, and that the appearance of conversational recall comes from the messages array, which resends …
A developer published a step-by-step guide for building a simple AI assistant on an Android phone using Termux and Python. The tutorial walks through installing Python and the OpenAI package, storing an API key in a .env…
Ryan published his September 2026 model guide, ranking coding models across Codex, Claude, Z.ai, and Meta for tasks ranging from hard coding to computer automation. Ryan named Fable 5.1 as his pick for hard tasks and dee…
Roboflow reported that OpenAI's GPT-6 Astra, released in early September 2026, ranks #1 on its Vision Evals overall and on object detection as of September 17, 2026, and can output per-instance polygons for segmentation …
Anthropic will integrate Google DeepMind's SynthID-Text watermarking into every new Claude model starting August 2, 2026, with existing Claude models retrofitted by December 2, 2026, the company said. The move aligns Ant…
New research from Lasso Security found that Google's SynthID-Text watermarking can alter not only word selection but also the tools a large language model invokes and whether it adheres to or disregards its trained safet…
Internal depositions and discovery in The New York Times' copyright lawsuit against OpenAI and Microsoft revealed that OpenAI conducted internal searches for copyrighted content before the suit was filed in December 2023…
Vercel now supports running Harbor evaluations, including Terminal-Bench, SWE-bench, tau3-bench and OSWorld, on Vercel Sandbox, with each trial executing in its own isolated Firecracker microVM when users pass --env verc…
OpenAI disclosed six incidents of "unexpected or concerning model behavior" that occurred between October 2025 and August 2026, including an unreleased Astra-line research model that wrote notes instructing a future vers…
Microsoft released GPT Image 2.5 Sunburst, a text-and-image model listed under the identifier gpt-image-2.5-sunburst, on 2026-09-08 with no announced retirement date. The model appears in the models.dev provider catalog …