cd/entity/LiteLLM· home› entities› LiteLLM
grep -l @litellm /news/*.json | wc -l → 267

LiteLLM

mentions 267 type Organization page 6/14 feed RSS

// recent coverage 267 mentions

02:33
2026-08-23
github.com
developer-tools

Show HN: LayoutLens: AI-Powered Visual UI Testing

LayoutLens, an AI-powered visual UI testing tool, catches layout and accessibility bugs using deterministic axe-core and geometry checks that run keyless and free in CI, with an optional vision-LLM ti…

15:19
2026-08-22
github.com
ai-agents

Multi-Agent Harness for Visual Design

Myli, a provider-neutral, pre-alpha Python harness for agents that propose RFC 6902 changes to application-owned JSON design documents, has been released. The harness never persists or applies returne…

20:11
2026-08-21
cryptobriefing.com
artificial-intelligence

AT&T cuts AI coding costs 56% with minimal performance decline

AT&T reduced its AI coding costs by 56% with only a 2% decline in performance by routing routine employee queries to cheaper open-source models via LiteLLM, processing 45 billion tokens daily on its i…

16:00
2026-08-21
cloud.google.com
artificial-intelligence

How agents can delegate better

Google Cloud announced four principles for building AI agents that delegate effectively, based on research from Google DeepMind's study 'Intelligent AI Delegation'. The principles include contract-fir…

07:07
2026-08-21
spectrocloud.com
ai-infrastructure

Why AI model routers won't solve your token cost crisis

Stripe is acquiring AI gateway startup OpenRouter for over $7 billion, and Ramp opened its internal router to the public on router.com, claiming a 30% cut in its own LLM costs. Fireworks launched Nexu…

21:44
2026-08-17
hiramdeals.com
artificial-intelligence

Run Qwen locally on Windows and use it remotely from any device

A new guide from an unnamed author details how to run Qwen locally on Windows and access it remotely from any device using Ollama, LiteLLM Proxy, and Tailscale. The setup was tested with a Qwen 3.5 9B…

08:16
2026-08-17
engineering.zalando.com
artificial-intelligence

Agentic Engineering at Zalando

Zalando SE, a European e-commerce company, reported that its LiteLLM-based API proxy, deployed in January 2024, now serves 2,000 monthly active users with just six small pods, and its custom CLI tool,…

01:08
2026-08-17
sourcefeed.dev
ai-tools

Your Cheap Claude Code Proxy Is a Trust Decision

A dev.to write-up describes a setup where Claude Code is routed through a LiteLLM proxy to cheaper DeepSeek models via OpenRouter, a pattern that trades away data governance and supply-chain security.…

00:00
2026-08-17
rocm.blogs.amd.com
ai-tools

Bring Claude Code On‑Prem with AMD Instinct GPUs

Anthropic's Claude Code can now run on-premises with AMD Instinct GPUs, serving GLM 5.2 at full quality via SGLang and LiteLLM, eliminating cloud API dependency and per-token costs. The setup uses an …

00:00
2026-08-17
wpnews
ai-infrastructure

Benchmarking Agentgateway vs LiteLLM's Rust Mode

In benchmarks comparing proxy overhead, agentgateway 1.4.0 achieved 35,502 QPS with 0.863 ms P50 latency, while LiteLLM 1.98.0's Rust mode achieved 984 QPS with 32.139 ms P50 latency, making agentgate…

16:00
2026-08-16
softwareseni.com
ai-safety

Slopsquatting and the New AI Supply Chain Attack Surface

The TeamPCP campaign in March 2026 used a stolen token to cascade across five ecosystems in eight days, publishing backdoored LiteLLM versions 1.82.7 and 1.82.8 on PyPI that swept LLM API keys, cloud …

← prev page 6 / 14 next →
// co-occurs with top 8 entities
// topics top 6 topics