What it is
A quick, honest read on the project and why it belongs in your public portfolio.
Drop-in KV cache compressor for local LLM inference - Run 70B models on 8GB RAM
- Original project developed under Aman Sachan.
- Primary language: Python.
- GitHub repository: https://github.com/AmSach/kvquant.
- External project link: https://github.com/AmSach/kvquant.