I shipped an AI side panel in 2 weekends
A developer built a Chrome extension that lets users highlight text and send it to multiple AI models via a side panel in two weekends. The extension supports OpenAI, Claude, Gemini, and DeepSeek, but…
A developer built a Chrome extension that lets users highlight text and send it to multiple AI models via a side panel in two weekends. The extension supports OpenAI, Claude, Gemini, and DeepSeek, but…
Model routing, which directs simple tasks to cheap models and complex ones to frontier models, can cut AI costs by 40-70% in 2026 while mitigating vendor dependency risks highlighted by Fable 5's susp…
Developer Simon Willison replaced Anthropic's Claude Code with OpenCode after losing trust in Anthropic due to the Fable 5 silent downgrade controversy and export-control issues. He cites open standar…
GNOME-aligned AI assistant Newelle released version 1.4.5, adding integrated AI image generation support via Stable Diffusion and cloud models from OpenAI, Pollinations, and OpenRouter. The update als…
ByteDance open-sourced DeerFlow 2.0, a long-horizon agent runtime that orchestrates sub-agents, sandboxes, persistent memory, and an extensible skill system. The project reached 74,960 GitHub stars an…
Hermes Agent, an open-source AI agent, has achieved 188,000 GitHub stars and processes 224 billion daily tokens on OpenRouter, making it the fastest-growing open-source agent of 2026, yet its underlyi…
Z.ai released GLM-5.2, an open-weight AI model that performs within one percentage point of Anthropic's Opus 4.8 on agentic benchmarks while costing roughly one-fifth as much, according to CNBC. The c…
On June 12, the US government ordered Anthropic to disable Fable 5 globally within six hours, leaving developers without access and highlighting the risk of regulatory takedowns. The following day, GP…
OpenClaw Launch introduces a managed AI agent service deployable in 30 seconds without installation or coding, offering integrations with email, calendar, Slack, and support for multiple AI providers …
A developer tested connecting an OpenAI-compatible API to WordPress AI using the Koneek plugin. The test successfully integrated OpenRouter as a provider, enabling WordPress AI features through config…
A data science student built a calculator to compare LLM API costs after finding that agentic coding sessions on OpenRouter could vary from cents to dollars per loop. The cheapest paid models in mid-2…
Nous Research's open-source Hermes Agent, using Mixture of Agents presets, outperformed Anthropic's Claude Opus 4.8 and OpenAI's GPT-5.5 on SWE-bench Pro benchmarks. The framework strings multiple lan…
Nous Research released Mixture of Agents presets as virtual models in Hermes Agent, allowing users to select multi-model workflows like any other model. The company claims its MoA presets outperform i…
Developer Scott Spencer released My-Pi Coding-Agent, a curated distribution of the Pi coding agent CLI with prewired extensions including MCP, LSP, skills, recall, redaction, telemetry, team mode, and…
Developers are adopting local proxy routers like Weave and 9Router to cut API costs and bypass the context re-read tax in AI coding assistants such as Claude Code and Cursor. These routers intercept A…
Weave released a smart model routing proxy that works with Claude, Codex, and Cursor, ranking #1 on the RouterArena leaderboard. The open-source tool uses an on-box embedder to select the best model p…
A developer built Aantraa, an AI-powered platform for audio and video translation, caption generation, and viral short clip creation, in one week. The platform relies heavily on AI APIs, using OpenRou…
Per-token prices for large language models are collapsing, but AI bills are exploding as reasoning models consume far more tokens per task. Uber burned through a year's AI budget in four months, and M…
AI search engine Exa raised $250 million in Series C funding at a $2.2 billion valuation, led by a16z, to power AI agents with high-quality web search. Exa already serves over 5,000 companies includin…
NeuralWatt, a US-based AI inference provider, introduced energy-based metering for LLM inference, charging by kilowatt-hour instead of per token. A user reported an average 82.9% cost reduction compar…