00:06
2026-10-01
dev.to
ai-agents
Token Compression for Coding Agents: Fine-Tuned Middleware Cuts Codex Costs by 30%
A developer released an open-source local proxy that wraps Codex and intercepts tool-call results before they return to the model, using a fine-tuned Qwen compression model to cut input tokens by 29.6…