Taleef7/turboquant

View on GitHub

Evidence-first KV-cache compression for Qwen2.5: packed low-bit KV blocks, PyTorch references, Triton kernels, and reproducible GPU benchmarks.

Stars
0
Forks
0
Open beginner issues
0
Indexed issues
5
Dominant language
Python
License
MIT
Last GitHub push
Aug 2, 2026
Latest indexed
Sep 20, 2026
Contributing guide
Contributing guide
Code of conduct
No code of conduct
Beginner labels
No beginner labels indexed
PR merge metrics
No merged PRs in 30d
5 open issues indexed Loading issues

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.