GPT-5.6 Price Drop: What OpenAI's Pricing Shift Reveals OpenAI has cut pricing on GPT-5.6 by 25–30% across input and output tokens, responding to enterprise pushback against high AI costs. The reduction targets production workloads where per-token burn rates had driven teams to consider self-hosted alternatives, according to a developer who reported monthly bills for 50k daily API calls were 'eye-watering'. GPT-5.6 Price Drop: What OpenAI's Pricing Shift Reveals Enterprises are pushing back hard on AI costs, and OpenAI just blinked first with a price cut on GPT-5.6. The move doesn't surprise me — I've been hearing from more and more teams who bulk at the per-token burn rate for production workloads. I started using GPT-5.6 in a real-time summarization pipeline about three months ago. For the volume we needed — roughly 50k API calls a day — the monthly bill was eye-watering. We actually considered switching to a self-hosted alternative because the margins on our service were getting squeezed. That's the kind of pressure that seems to be driving this change. OpenAI's response is a reduction in both input and output token pricing. Roughly 25–30% depending on the tier, from what I can gather from the updated pricing page. The cut specifically targets GPT-5.6, Next How I Fixed My GPT‑2 Reproducibility Nightmare Part 2 → /en/news/4450/ All Replies (2) Q N