Taleef7/turboquant
View on GitHubEvidence-first KV-cache compression for Qwen2.5: packed low-bit KV blocks, PyTorch references, Triton kernels, and reproducible GPU benchmarks.
- Stars
- 0
- Forks
- 0
- Open beginner issues
- 0
- Indexed issues
- 5
- Dominant language
- Python
- License
- MIT
- Last GitHub push
- Aug 2, 2026
- Latest indexed
- Sep 20, 2026
- Contributing guide
- Contributing guide
- Code of conduct
- No code of conduct
- Beginner labels
- No beginner labels indexed
- PR merge metrics
- No merged PRs in 30d
5 open issues indexed
Loading issues
-
documentation
Difficulty 4/5 3-5 days Newbie friendliness 48/100
Taleef7/turboquant#5 · 2 comments ·
-
enhancement testing
Difficulty 5/5 Over a week Newbie friendliness 20/100
Taleef7/turboquant#1 · 1 comment ·
-
enhancement optimization
Difficulty 5/5 Over a week Newbie friendliness 35/100
Taleef7/turboquant#2 ·
-
enhancement optimization
Difficulty 5/5 Over a week Newbie friendliness 25/100
Taleef7/turboquant#3 ·
-
enhancement performance
Difficulty 5/5 Over a week Newbie friendliness 35/100
Taleef7/turboquant#4 ·