cd/entity/RouteLLM· home entities RouteLLM
grep -l @routellm /news/*.json | wc -l → 17

RouteLLM

mentions 17 type Organization feed RSS

// recent coverage 17 mentions

09:00
2026-09-09
infoworld.com
artificial-intelligence

The five important tools for controlling AI costs

InfoWorld reports that AI bill shock is spreading across the industry, driven by opaque large language model (LLM) costs and poor attribution of spend to business value. The article recommends five to…

00:00
2026-08-11
mindstudio.ai
artificial-intelligence

What Is NVIDIA SwitchYard? The Open-Source Local AI Model Router

NVIDIA has released SwitchYard, an open-source routing library that automatically directs AI agent tasks to the most appropriate local or frontier model, claiming it can deliver 50% faster responses a…

00:00
2026-08-11
mindstudio.ai
artificial-intelligence

Set Up a Local AI Router With SwitchYard and Nemotron Lightning

NVIDIA's open-source SwitchYard router, built on RouteLLM, can be configured with a locally hosted Nemotron 3.5 Lightning model to cut API costs without losing task quality. On DGX Spark hardware, Nem…

21:08
2026-07-31
sourcefeed.dev
ai-infrastructure

The LLM Router Was a 2024 Idea

Manifest, an open-source development platform, shipped an LLM router in March 2024, saw 7,000 cloud users use it for four months, and deprecated it in June, with full shutdown on September 1. The comp…

23:52
2026-07-29
techstrong.ai
large-language-models

LLM Routers Have Become a Service Category of Their Own

LLM routers have evolved from a niche infrastructure trick into a mainstream product category, with gateways like OpenRouter, LiteLLM, and Portkey offering API unification and failover, while smart ro…

12:00
2026-07-29
letsdatascience.com
artificial-intelligence

Microsoft Outlines a Three-Layer Agent Routing Stack for AKS

Microsoft's Azure Kubernetes Service engineering team published a reference implementation on June 29 that separates agent-request routing into three layers: semantic model selection via RouteLLM, gat…

16:00
2026-07-27
spectrocloud.com
ai-infrastructure

When does local inference pay for itself? - Spectro Cloud

Spectro Cloud released a TCO calculator for its PaletteAI Inference Launchpad, a turnkey appliance that runs open models on local GPUs with intelligent routing, claiming it can cut inference costs by …

16:01
2026-07-18
pub.towardsai.net
artificial-intelligence

The Model Is Not the Product

Microsoft Copilot now runs on Anthropic's Claude in some features, but users still perceive Claude as smarter because the model is not the product, according to AI researcher Louiza Boujida. Boujida e…

22:39
2026-07-12
avriz.io
machine-learning

We taught our platform to learn its own pricing decisions

Avriz, a cloud platform for coding agents, has built a contextual bandit system that learns to route LLM calls to the cheapest capable model without risking user experience. The system trains exclusiv…

06:20
2026-06-24
github.com
large-language-models

I built an LLM router that doesn't use an LLM

Developer Lore released Wayfinder, an open-source LLM router that determines whether to send a prompt to a local or cloud model by analyzing structural features like length, headings, and code, withou…

05:17
2026-06-18
github.com
large-language-models

Maslul – Smart LLM router – one call, the right model

Maslul, a new open-source Python library, provides smart LLM routing and provider normalization across Anthropic, Gemini, Grok, and OpenAI, allowing developers to route each request to the right model…

// co-occurs with top 8 entities
// topics top 6 topics