cd /news/artificial-intelligence/codec-gauge-learning-compression-fri… · home topics artificial-intelligence article
[ARTICLE · art-71402] src=arxiv.org ↗ pub= topic=artificial-intelligence verified=true sentiment=↑ positive

Codec-Gauge: Learning Compression-Friendly Gauges for Transformer KV Caches

A new post-training method called Codec-Gauge, introduced in arXiv:2607.20538v1, learns small orthogonal channel transforms for Transformer KV caches to improve compression fidelity. Across six models at 3, 4, and 6 bits per value, Codec-Gauge reduces zfp KL divergence by 44.0% on average compared to raw coordinates, outperforming random, Hadamard, DCT, and PCA/KLT controls without changing model weights or attention semantics.

read1 min views1 publishedJul 24, 2026

arXiv:2607.20538v1 Announce Type: new Abstract: Long-context Transformer inference increasingly relies on KV-cache compression or quantization. Prior rotation and transform-coding results suggest that the channel basis of each key/value vector affects how faithfully a fixed backend preserves model behavior. We introduce Codec-Gauge, a post-training cache-coordinate layer that learns small orthogonal channel transforms around existing compression and quantization backends. Its frequency-distribution objective combines a token-channel DCT spectral-centroid loss with a smooth rate proxy to concentrate KV energy in low-frequency codec-facing layouts. We evaluate actual compression and decompression using measured bytes and rolling compressed-history scoring. Across six models at $3$, $4$, and $6$ bits/value, learned gauges reduce zfp KL divergence by $44.0%$ on average relative to raw coordinates and outperform random, Hadamard, DCT, and PCA/KLT controls. The same gauges improve quality preservation for block-uniform and KIVI-style quantization. Experiments on a 27B model and long-context task prompts reproduce the quality trend, while serial storage and timing measurements validate the implemented compressed-cache paths. These results establish cache-coordinate geometry as a practical post-training variable for improving compression fidelity without changing model weights, attention semantics, or backend coding rules.

── more in #artificial-intelligence 4 stories · sorted by recency
── more on @codec-gauge 3 stories trending now
sponsored brought to you by zahid.host 4,200+ EU-deployed projects
reading about agents? ship yours in a single git push.

Run your AI side-project on zahid.host

EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.

$git push zahid main
Live at https://your-agent.zahid.host
Get free account → Pricing
from €0/mo · no card required
LIVE [news/codec-gauge-learning…] indexed:0 read:1min 2026-07-24 ·