llm-openrouter 0.7.1
Llm-openrouter 0.7.1 was released on 2nd September 2026, according to the project's changelog entry. The release is a version update to the llm-openrouter plugin, which connects the LLM command-line t…
Llm-openrouter 0.7.1 was released on 2nd September 2026, according to the project's changelog entry. The release is a version update to the llm-openrouter plugin, which connects the LLM command-line t…
In a recent experiment, three Claude sessions acting as product owners communicated with each other over a Unix domain socket while working on separate areas of a codebase, with Fable 5.1 orchestratin…
A study of 7,534 citations from Perplexity's AI models across 380 software categories found that 59.8% of sources cited are outside the top 100,000 most-visited websites, and 23.4% are not in the top …
A Dask maintainer and staff software engineer at OpenTeams criticizes ArtificialAnalysis's intelligence-vs-cost plot for using a logarithmic cost scale, official API pricing, and datacenter pricing fo…
Z.ai has released GLM-5.3-Flash, an open-source, MIT-licensed model that is aggressively cheap at $0.15 per million input tokens and runs on Chinese-made silicon, signaling a shift in the AI hardware …
Mst3k-anything, an open-source tool that automatically generates Mystery Science Theater 3000-style robot heckling for any video, is now available, using yt-dlp, ffmpeg, sherpa-onnx with Parakeet 110M…
Anthropic released Claude Fable 5.1 and Claude Mythos 5.1 on September 1, 2026, at unchanged prices of $10/$50 per million tokens, with cache reads cut to $0.25, while Google released Gemini 3.8 Flash…
Doublespeed AI, the model powering the social media platform doublespeed, achieved 61.0% accuracy in a benchmark of 200 real TikTok post pairs, outperforming frontier models including Claude and GPT 5…
The z-ai/glm-5.3 endpoint on Modal was retired from B3IT monitoring on 2026-09-08 after the tracker could not find enough border inputs, while its LT logprob probe also failed to return usable logprob…
OpenRouter's tracking dashboard shows the deepseek/deepseek-v4-flash endpoint at atlas-cloud/fp4 recorded a B3IT total-variation level shift of 0.55, while its LT logprob tracking remains impossible b…
Inception released Mercury 2.5, a diffusion large language model (dLLM) that generates tokens in parallel, achieving 1,107 tokens/sec on standard GPUs and a 10+ point intelligence jump over Mercury 2,…
A developer compared managed multi-provider routers, self-hosted gateways, and thin in-app adapters for property-management moderation, concluding that raw token rates alone cannot determine the cheap…
A New York City resident built an AI-powered bird sprinkler system at his parents' Florida home, using a Raspberry Pi 4, an IP camera, and OpenRouter's AI to detect birds and trigger a sprinkler. The …
DeepSeek V3 outperformed Claude 3.5 Sonnet in a six-hour coding stress test, achieving 95% logic accuracy versus 92% for Claude, while costing about $0.28 per 1k tokens compared to Claude's $15.00, ac…
On June 4, 2026, NVIDIA released Nemotron 3 Ultra, a 550 billion parameter open-weight hybrid Mamba-Attention Mixture-of-Experts model that activates about 55 billion parameters per token, designed fo…
Brilliant, a new design tool that lets AI agents edit real vector canvases, launched on Hacker News, positioning itself as a professional design canvas that Claude Code, Cursor, or any MCP agent can d…
Verba, an open-source language-learning app, lets users practice eight languages through conversations with an AI that runs entirely offline on their machine, with no account or server required. The a…
A developer reports that running OpenCode with DeepSeek V4 models, both Pro and Flash, delivers coding performance comparable to more expensive models like Claude or GPT at a fraction of the cost. The…
Zhipu AI's GLM-5.3-Flash, an MIT open-source 320B-A18B MoE model, is priced at roughly 1/40th of Claude Opus 4.8's per-token rate while matching its intelligence on the Artificial Analysis Intelligenc…
OpenRouter has tracked the deepseek/deepseek-v4-flash-0731 endpoint at streamlake/fp8 since 2026-08-13, recording $0.032 in total spend across 66,386 queries. The B3IT border-input detector measured a…