Evaluating LLM models for DBA tasks
Percona Lab released dbaai_bench, an open-source harness that evaluates large language models on real database administration tasks against remote hosts, and reported that almost all tested models com…
Percona Lab released dbaai_bench, an open-source harness that evaluates large language models on real database administration tasks against remote hosts, and reported that almost all tested models com…
Weekly token consumption on OpenRouter surged more than 25,000 percent since January 2025, rising from 0.5 trillion to 126.2 trillion tokens, according to The Decoder. The publication attributes the i…
An undisclosed multimodal AI model named 'Union Alpha' has emerged claiming performance comparable to Claude Fable 5.1 and GPT-6 Astra on the Terminal-Bench 4.0 benchmark at less than $2 per task, acc…
A developer's side-by-side test found that Google's Gemini 3.8 Flash cloud model produced a faster Rust parsing function than the local Gemma 4 model (gemma-4-26B_q4_0-it.gguf), cutting starts_with ca…
OmniRoute, an MIT-licensed self-hosted AI gateway from developer diegosouzapw, routes coding-agent sessions across 352 providers — including 150+ free ones — from a single OpenAI-compatible endpoint a…
Nous Research's Hermes Agent, an MIT-licensed open-source AI agent that remembers context across sessions, uses tools, and runs scheduled jobs, can be put to work through five free self-hosted setups,…
TypeSafe's Jev Ultrafast browser agent completed a natural-language Google Flights search from Zürich to London in 7.1 seconds, including text generation and loading waits, according to the project's …
Developer ohkariku-boop released echodot v0.1, a free, MIT-licensed, open-source desktop app that drafts replies in a user's personal writing style via a global hotkey (⌘⇧E / Ctrl+Shift+E) that reads …
A former OpenAI memory engineer launched FastRecall, an API that stores and recalls context across different AI models, with free recalls and paid plans starting at $2 per month. FastRecall respects p…
An unidentified multimodal AI model called Union Alpha surfaced unannounced on model-routing platforms including OpenRouter with a 262K token context window and no lab claiming it. On a software engin…
OpenRouter's tracking page for the z-ai/glm-5.3-flash endpoint on Cloudflare reports a B3IT total-variation level shift of 0.21, with the endpoint under active monitoring through its border inputs. Th…
VideoRouter launched as an OpenRouter-style routing service for video and image generation APIs, charging a flat 2% platform fee versus OpenRouter's 5%. The service negotiates volume pricing with GPU-…
Kilo Gateway, Kilo's OpenAI-compatible inference gateway at https://api.kilo.ai/api/gateway, provides one API key across 500+ models from 60+ providers with pay-as-you-go billing at exact provider rat…
Rubric's team published a numbered list of agent design patterns drawn from building agents internally and for clients including Graphite and Albertsons, recommending system prompts stay under 5,000 t…
Monid.ai launched a router for agent tools that gives AI agents access to more than 2,000 tools across 72+ providers through one base URL and one API key, with usage metered per call. The connector la…
OpenCode and OpenRouter launched Union Alpha, a free "stealth" coding model, on September 16 with a 262,144-token context window, text and image input support, and a zero-retention data policy. The mo…
A developer's audit of pi's session logs against OpenRouter's generation API found a supposedly cheap DeepSeek V4.1 Flash setup paying about three times what it should, because OpenRouter's default Au…
OpenRouter listed a new stealth model called Union Alpha on September 16, 2026, developed and operated by an anonymous third-party provider. Union Alpha is a multimodal model built for research, codin…
A developer compiled a curated list of more than 35 text-generation LLM API providers that offer free, self-replenishing quotas, cataloging rate limits and daily token allowances for 47 services inclu…
A developer analysis of OpenRouter's LLM routing gateway warns that the abstraction between model and provider leaks in production, with the same model ID yielding significantly different performance,…