cd/entity/Qwen3-32B· home entities Qwen3-32B
grep -l @qwen3-32b /news/*.json | wc -l → 54

Qwen3-32B

mentions 54 type Organization page 1/3 feed RSS

// recent coverage 54 mentions

03:01
2026-08-19
dev.to
artificial-intelligence

How I Cut AI API Costs 95% — A Data Scientist's Field Guide

A data scientist at an unnamed company cut AI API costs by 95% by analyzing six months of logs and implementing a model-routing pipeline that matches each request to the cheapest adequate model. The a…

04:13
2026-07-15
dev.to
large-language-models

I Was Shocked I'm Overpaying for AI by 40x as a Bootcamp Grad

A bootcamp graduate discovered they were overpaying for AI by up to 40x after analyzing API costs. By switching from GPT-4o to models like DeepSeek V4 Flash via Global API, the developer reduced a $50…

14:55
2026-07-14
dev.to
artificial-intelligence

How I Cut My OpenAI Bill by 97% — The Full Migration Guide

A developer cut their OpenAI bill by 97% by migrating to Global API, which offers models like DeepSeek V4 Flash at 40× lower cost than GPT-4o. The migration required changing only two lines of code—th…

02:52
2026-07-12
sourcefeed.dev
artificial-intelligence

OpenAI Drop-ins Are Easy. Production Is Not.

OpenAI's wire protocol has become the interop layer for multiple inference providers, making model swaps cheap via base URL changes. However, production readiness requires quality gates, fallbacks, an…

00:14
2026-07-12
dev.to
large-language-models

Migrating Off OpenAI: A Backend Engineer's Notes From Production

A backend engineer migrated three production services from OpenAI to DeepSeek V4 Flash via a Global API endpoint, reducing monthly costs from $500 to approximately $12.50—a 40× price difference—while …

11:58
2026-07-01
dev.to
large-language-models

I Cut My AI Bill 97.5% in One Afternoon — And You Can Too

A developer cut their monthly AI bill from $487.92 to $12.50 by switching from OpenAI's GPT-4o to DeepSeek V4 Flash via the Global API, achieving a 97.5% cost reduction. The migration required changin…

00:02
2026-06-30
dev.to
large-language-models

From $500 to $12.50: My Real Migration Off OpenAI in 2026

A developer migrated from OpenAI's GPT-4o to DeepSeek V4 Flash via a global API provider, reducing monthly costs from $487 to $12.50 with only two lines of code changed. The switch required only alter…

08:07
2026-06-29
dev.to
large-language-models

A Better LLM Judge? The Rubric Made My Small Model Worse

A developer found that improving the rubric for a small LLM judge (Qwen2.5-1.5B) did not increase its agreement with human votes, which remained around 43%. However, swapping to a larger model (DeepSe…

14:01
2026-06-27
dev.to
large-language-models

The Developer's Guide to Trimming AI API Costs Without Crying

A backend engineer at an unnamed company slashed their team's LLM API costs from $11,400 to $1,830 per month by switching to cheaper models for most tasks and implementing tiered routing. The team rep…

09:37
2026-06-27
dev.to
large-language-models

Cutting OpenAI Costs From Scratch: What Nobody Tells You

A B2B SaaS startup cut its LLM inference costs by 97% by switching from GPT-4o to cheaper alternatives like DeepSeek V4 Flash, reducing a $14,200 monthly OpenAI bill to an estimated $355. The develope…

07:01
2026-06-27
dev.to
large-language-models

I Tracked Every API Dollar Across 184 Models: Here's The Data

A developer tracked API costs across 184 models over 18 months, spending $340,000 in credits. The data reveals that direct provider pricing can be 40x cheaper than GPT-4o, but operational friction and…

16:20
2026-06-26
dev.to
large-language-models

How I Cut Our AI API Bill by 95%: What Actually Worked

A developer cut their company's AI API bill by 95% from $11,000 to under $400 per month by implementing per-request model routing and tiered escalation. The team replaced expensive GPT-4o calls with c…

page 1 / 3 next →
// co-occurs with top 8 entities
// topics top 6 topics