We Got the Prompt Cache Working. Our Pipeline Got Slower.
An engineer at a small AI startup benchmarked OpenAI's Codex app-server against the standard subprocess approach and found that prompt caching made the pipeline 39% more expensive in raw tokens and slower overall. The ca…