The Plunging Price of Thought The cost of achieving a given level of AI performance has fallen about 47% per quarter, or 13× per year, since 2023, according to an analysis by researcher David Roodman published with data and code on GitHub. The analysis found the decline is four times faster than DNA sequencing, six times faster than compute, 18 times faster than lithium batteries, and 54 times faster than electricity in the century up to 1973, with cost falling 66% per quarter (75× per year) for performance that has just debuted as state of the art and 32% per quarter (4.7× per year) two years later. Roodman cited OpenAI's o3, released January 31, 2025, achieving a 75% score on GPQA Diamond at an estimated 30 cents per question, versus GPT-5.6 Luna scoring the same just under 18 months later at $0.0004 per question — a 725-fold drop. Key takeaways - AI has gotten cheaper more quickly than any other transformative technology in history. The cost of achieving a given level of AI performance has fallen about 47% per quarter since 2023, or 13× per year. That price drop is four times faster than DNA sequencing, six times faster than compute, 18 times faster than lithium batteries, and in the century up to 1973 54 times faster than electricity. - Coarser evidence suggests that the price of thought has been falling at least this fast since the dawn of commercial LLM inference in November 2021, when OpenAI fully released GPT-3. - The speed of price drop varies by domain: slower on game-based puzzles, at 39–43% per quarter, faster on math problems, at 50–52% per quarter. - The cost of a given level of performance often falls fastest right after that level is first achieved, that is, when it is state of the art SOTA . We see this pattern on three of our five main benchmarks of AI capability. Averaging across all five, cost falls 66% per quarter 75× per year for performance that has just debuted as SOTA. Two years later, prices fall half as fast, at 32% per quarter 4.7× per year . Data and code are on GitHub https://github.com/droodman/inference-cost/ . An overlay page https://droodman.github.io/inference-cost/ has many plots and tables to explore. Overview The “GPT” in “ChatGPT” stands for “Generative Pre-trained Transformer,” a technical description of how the AI inside it works. Surely, though, the creators of OpenAI’s GPT models were nodding to an older meaning of the initialism: general-purpose technology