cd/entity/LiteLLM· home entities LiteLLM
grep -l @litellm /news/*.json | wc -l → 160

LiteLLM

mentions 160 type Organization page 5/8 feed RSS

// recent coverage 160 mentions

08:02
2026-07-11
gist.github.com
large-language-models

Config-NaN - configuracion NaN Builders para LLMs

NaN Builders has published a comprehensive configuration guide for its LLM API, detailing setup instructions for clients, IDEs, agents, and SDKs. The guide covers available models including deepseek-v…

11:36
2026-07-10
justindfuller.com
large-language-models

Accretive Editing

Accretive editing is a failure mode of current AI tools where they add parentheticals or addendums instead of correcting text, as seen when Claude updated a project description to include both old and…

17:09
2026-07-08
dev.to
ai-infrastructure

AI Gateway Fees Compared: Who Marks Up Your Tokens?

A developer compared AI gateway pricing models in 2026, finding that none of the major gateways—including LLM Gateway, OpenRouter, Vercel AI Gateway, Cloudflare AI Gateway, Eden AI, Portkey, and LiteL…

17:08
2026-07-08
dev.to
ai-tools

8 Best AI Gateways in 2026 (Compared)

A developer evaluated eight AI gateways based on provider coverage, pricing transparency, self-hosting, observability, and ease of setup. The top pick is LLM Gateway, an open-source solution that rout…

02:01
2026-07-07
devashish.me
large-language-models

Owning Inference - Qwen3.6 on DGX Spark for real coding

A developer successfully runs the Qwen3.6-27B-FP8 model locally on an Nvidia DGX Spark, achieving reasoning, tool use, and multi-token prediction at 256K context, and uses it to ship real code for an …

00:00
2026-07-06
mbgsec.com
ai-safety

Attackers Don’t Buy Tokens. They Steal Yours.

Security researchers at MBG Security built a global network of honeypots with exposed AI inference and agent endpoints, observing attackers actively scanning for and exploiting vulnerabilities like CV…

10:32
2026-07-04
dev.to
large-language-models

Solving the GPU Pinning Saga and Gemma's Meta-Commentary

Glad Labs fixed a GPU pinning issue where LiteLLM 1.89.2's global api_base override prevented per-model routing, causing vision tasks to cold-load onto the wrong GPU. The team also hardened content gu…

20:44
2026-07-03
tigera.io
ai-agents

Six AI agent SDKs for enterprise Kubernetes, compared

Six AI agent SDKs—LangGraph, CrewAI, Google ADK, and others—are compared for enterprise Kubernetes deployment, with most being model-agnostic and containerizable for on-premise use, though Anthropic's…

14:07
2026-06-30
letsdatascience.com
ai-infrastructure

Zenity Labs Reveals AI Infrastructure Weaponization

Zenity Labs revealed that attackers are weaponizing AI infrastructure, targeting exposed model gateways and unmanaged LLM endpoints. The firm's sensors detected thousands of real-world attacks, includ…

11:07
2026-06-30
dev.to
large-language-models

That 200 OK From Your LLM Gateway Probably Means Nothing

A developer warns that HTTP 200 responses from LLM gateways do not guarantee correct output, as gateways like LiteLLM, Portkey, and OpenRouter only check transport-level success. The developer propose…

14:20
2026-06-28
stefan.schueller.net
artificial-intelligence

Show HN: Making a label printer work under Linux using agentic AI

A developer used agentic AI to decompile a Chinese label printer's Android app and create a Go script that prints PDFs via Bluetooth on Linux, after struggling with poor print quality under CUPS. The …

← prev page 5 / 8 next →
// co-occurs with top 8 entities
// topics top 6 topics