vllm-project/vllm
View on GitHub[Feature]: Use kernel abstraction for fp4 scaled mm
Open
#31823 opened on Jan 6, 2026
feature requesthelp wantedunstale
Description
🚀 The feature, motivation and pitch
See MPLinearKernel and ScaledMMLinearKernel.
Alternatives
No response
Additional context
No response
Before submitting a new issue...
- Make sure you already searched for relevant issues, and asked the chatbot living at the bottom right corner of the documentation page, which can answer lots of frequently asked questions.