Nobody knows where their AI budget is going
Gartner forecasts worldwide AI spending will reach $2.59 trillion in 2026, a 47% increase from 2025, as organizations struggle to track where their AI budgets are going. The FinOps Foundation reports …
Gartner forecasts worldwide AI spending will reach $2.59 trillion in 2026, a 47% increase from 2025, as organizations struggle to track where their AI budgets are going. The FinOps Foundation reports …
A new analysis from InfoWorld warns that data preparation and context quality, not GPU costs, are the hidden drivers of AI spending, citing a FinOps Foundation survey finding that 73% of enterprises r…
GPU cloud provider developer relations staff report that AI teams are repeating cloud-cost mistakes, with GPU spending treated as a bold bet rather than an operating cost, leading to waste. One team r…
An enterprise internal developer platform typically run by a ~60-person product organisation costs roughly $7.5 million a year, and organizations must assess agent-native readiness and ROI before rebu…
Stanford's 2025 AI Index Report found that inference cost for a system performing at GPT-3.5's level dropped more than 280-fold between November 2022 and October 2024, yet the FinOps Foundation's Stat…
JetBrains has centralized its AI usage through a shared access and accounting layer called Central CLI after development-related AI spending increased roughly tenfold in six months. The company report…
Meta Platforms Inc. has dismantled its internal AI token leaderboard and issued memos to curb employee AI usage after realizing internal AI costs were on track for billions of dollars, according to re…
The Linux Foundation launched the Tokenomics Foundation with 29 founding members, including JPMorgan Chase, IBM, Accenture, and Oracle, to standardize how enterprises measure and manage AI token costs…
Oumi cofounder and CEO Manos Koukoumidis is building tools to help enterprises cut AI costs by using smaller, task-specific models instead of expensive frontier systems, a market that also includes co…
A Cast AI report finds 69% of Kubernetes clusters are CPU-overprovisioned in 2026, up from 40% in 2024, with average CPU utilization at 8%, and the company advocates for admission-time policies like R…
The Linux Foundation announced the intent to launch the Tokenomics Foundation, a new open-source initiative to establish standards and benchmarks for AI infrastructure cost management. The foundation …
The average Kubernetes cluster uses just 8% of the CPU it pays for, according to the Cast AI 2026 State of Kubernetes Optimization Report, highlighting the need for Kubernetes FinOps to close the gap …
Costory rebuilt its FinOps MCP server after discovering that LLMs like Claude ignored dedicated tools like query_cost_diff in favor of calling query_cost twice and doing subtraction themselves. The te…
A new Kubernetes cost optimization checklist reveals that average CPU utilization across production clusters is only 8% in 2026, down from 10% the prior year, with 69% of clusters over-provisioning CP…
Snowflake is embedding AI into its cost management tools to help FinOps teams govern AI-driven workloads, addressing the top priority of managing AI spend as 98% of FinOps teams now oversee AI costs. …
Kubernetes cost optimization removes waste through technical actions like rightsizing pods and adopting Spot instances, while cost management provides ongoing visibility, allocation, and governance. A…
The Linux Foundation announced on June 3, 2026, its intent to launch the Tokenomics Foundation, dedicated to open standards for AI cost management, with support from Google, Microsoft, Oracle, JPMorga…
Uber burned through its entire 2026 AI budget in four months, driven by widespread use of AI coding tools like Claude Code across the organization, not by a single moonshot project. Leaked audio from …
Major corporations including Uber, Amazon, and Walmart are rationing AI tool usage after inference costs consumed annual budgets months ahead of schedule. Uber exhausted its entire 2026 AI budget by A…
Token prices for large language models dropped roughly 80% between 2025 and 2026, but engineering teams are seeing AI bills explode due to the Jevons paradox—cheaper tokens drive much higher usage, es…