CORS Chat
Simon Willison released CORS Chat, a web tool built with GPT-5.6-Sol xhigh, to test Qwen 3.8 27B running in LM Studio on his M5 MacBook Pro and an NVIDIA DGX Spark. The tool provides a web UI for Open…
Simon Willison released CORS Chat, a web tool built with GPT-5.6-Sol xhigh, to test Qwen 3.8 27B running in LM Studio on his M5 MacBook Pro and an NVIDIA DGX Spark. The tool provides a web UI for Open…
A developer has launched Mrigashira AI, a desktop AI engine designed to run large language models locally on personal hardware while offering optional cloud connectivity. The tool dynamically adjusts …
A developer has released Trustless, an open-source CLI that prevents AI coding agents from exposing API keys and other secrets. The tool injects credentials into subprocesses while sanitizing output, …
DeepSeek announced the general availability of DeepSeek-V4-Pro on August 13, 2026, across its app, web, and API, with the model name deepseek-v4-pro unchanged. The release includes MIT-licensed weight…
Google released Gemini 3.7 Flash, a new AI model priced at $1.50/$7.50 per million input/output tokens, which appears competitive with Claude Sonnet 5 and is offered at a 50% discount on OpenRouter. O…
Google, OpenAI, and Anthropic all charge exactly half their synchronous rate for batch API work, yet most cost forecasts price every token at the sync list rate, creating a 2× error on workloads that …
Nvidia has launched NeMo Switchyard, a library for model routing that directs AI prompts to the most appropriate model to cut costs and improve accuracy. The move comes as Cloudflare introduced a mode…
OpenAI slashed the price of its GPT-5.6 Luna model by 80% and its Terra model by 20%, leading to a surge in usage that boosted revenue, according to TD Cowen analysts studying OpenRouter data. Luna's …
OpenRouter, an API proxy service, unifies access to 300+ LLM models from providers including OpenAI, Anthropic, and Meta with a single API key, standardized responses, and automatic fallback. It addre…
Miser, an open-source Rust-based AI gateway, routes OpenAI-compatible requests to the cheapest capable model via OpenRouter, achieving 92.0% exact tier accuracy with sub-millisecond latency on a 25-ca…
An independent developer testing DeepSeek's official API found that the model ignored the max_completion_tokens limit, consuming all available tokens on reasoning and returning empty visible answers. …
A developer investigating discrepancies between tracked LLM spend and provider invoices found that the two numbers almost never match, with causes documented in the tools' own docs. OpenAI and Anthrop…
XAI released Grok 4.6 at $2 per million input tokens and $6 per million output tokens, with a fast variant at double the price, undercutting comparable OpenAI and Anthropic models. The model is availa…
Google's Gemini 3.7 Flash, launched August 13, 2026 at $0.75 per million input tokens and $3.75 per million output, wins 9 of 19 benchmark rows in its own model card, while GPT-5.6 Terra wins 6 and Cl…
State-of-the-art AI models are two-thirds smarter than last November, with two new models released every three days, yet 84% of tokens on OpenRouter are not state of the art, and the six most-used mod…
XAI released Grok 4.6, a model built for long-running agents and ambitious coding, research, visual, and interactive work, scoring 61 on the Artificial Analysis Intelligence Index, up from 56 for Grok…
A developer has released 'Caged Code', a webpage that embeds the vanilla Claude Code binary and connects it to a notebook environment via WebSocket, allowing users to run Claude Code through their Ant…
OpenAI shipped GPT-5.6 on July 9 with three tiers—Sol, Terra, and Luna—priced at $5/$30, $2.50/$15, and $1/$6 per million input/output tokens, respectively, marking a shift to a tiered product structu…
OpenRouter, the AI routing and infrastructure layer, is hiring a remote Data Scientist for a salary of $180k–240k/yr to build the intelligence layer powering its model marketplace. The role involves i…
OpenAI released GPT-5.6 Terra, a proprietary AI model available in August 2026, which does not embed watermarks or provenance metadata in generated output. The model scores 90 on general performance a…