Technical synopsis
The paper is built around a precise cache-compression story with explicit maths and implementation detail.
An earlier deep write-up, useful for tracing how the mathematics and implementation story evolved.
- Adaptive mixed-precision cache compression
- Recency-aware bit allocation
- Sink-aware reservation for stable anchors
- Head/layer coupling for finer precision control
- Explicit metadata for invertibility and auditing
- Forward-pass wrapping for drop-in use