Technical synopsis
The paper is built around a precise cache-compression story with explicit maths and implementation detail.
A shorter, crisper manuscript if you want the idea in a more compact form.
- Adaptive mixed-precision cache compression
- Recency-aware bit allocation
- Sink-aware reservation for stable anchors
- Head/layer coupling for finer precision control
- Explicit metadata for invertibility and auditing
- Forward-pass wrapping for drop-in use