Pegasus logo
Aman Sachan papers • projects • research index
Project detail

kvquant

Drop-in KV cache compressor for local LLM inference - Run 70B models on 8GB RAM

What it is

A quick, honest read on the project and why it belongs in your public portfolio.

Drop-in KV cache compressor for local LLM inference - Run 70B models on 8GB RAM

  • Original project developed under Aman Sachan.
  • Primary language: Python.
  • GitHub repository: https://github.com/AmSach/kvquant.
  • External project link: https://github.com/AmSach/kvquant.

Why it matters

A short narrative that makes the page feel like a real portfolio entry, not a bare repo link.

Core idea

Drop-in KV cache compressor for local LLM inference - Run 70B models on 8GB RAM

Public surface

This page gives the project a clean landing spot for GitHub Pages, search, and sharing. It also keeps the repository discoverable alongside your papers.

Portfolio role

This repository is an original project, and is surfaced as a primary portfolio item.

Associated papers

Direct links between the project and its publication track.