Hacktoberfest 2026:メンテナが10月に向けて印を付けた、オープンで初心者向けの issue。 Hacktoberfest の issue を見る

Carry per-group slot mappings into multi-KV forwards

オープン
#3,109 コメント 0 件 リアクション 0 件 担当者 0 名 GitHub で見る

メンテナーはふだん 1 日以内に返信

まだ誰も着手していません。

評価

難易度
4/5
見積もり時間
3〜5日
初心者へのやさしさ
48/100
issue の種類
バグ
明瞭さ
おおむね明確
活発さ
活発
技術スタック
cpp

調査の方向性

共有の Multi-KV forward フローから開始し、StepInputs::slot_mapping を CommonAttentionMetadata と MultiKvCacheIndex を通して追跡する。forward 用の runner 所有のマッピングがどこに存在するか、また登録された DeepSeek V4 パスがそれらをどのように消費するかを特定する。各 KV グループが実際のマッピングを受け取り、そのパスを通じて非恒等マッピングが mutation によって証明されれば完了とする。

索引モデルが issue の本文から書いたものです。

説明

Row: MODEL-MM-deepseek-v4-deepseek-v4-for-causal-lm

DeepSeek V4 Vision W4 found that StepInputs::slot_mapping[g] is computed for every KV group, but the runner forwards only the full-attention group's mapping through CommonAttentionMetadata. MultiKvCacheIndex carries per-group block tables but no parallel per-group slot mappings.

That makes a production multi-cache model validate the real cache groups and then either invent identity slots or ignore non-primary cache writes. W4 needs the actual SWA, compressed-latent, attention-compressor, indexer-key, and indexer-compressor slots.

Fix in the same flow: add a borrowed per-group slot-mapping channel to the shared multi-KV index, populate it from runner-owned step inputs for the forward lifetime, and mutation-prove a non-identity mapping through the registered DeepSeek V4 path. The pull request for #2411 closes this issue too.

主要言語
C++
スター
440
フォーク
55
平均マージ
1日 4時間
マージ済み PR(30日)
366

環境構築

はじめの一歩

  1. issue を最後まで読み、次にプロジェクトのコントリビューションガイドを読みます。
  2. 着手することを issue にコメントします — 二人が同じ作業をするのを防げます。
  3. リポジトリをフォークし、ブランチを切って変更します。
  4. issue 番号を参照したプルリクエストを送ります。

mudler/vllm.cpp のほかの issue

mudler/vllm.cpp の issue をすべて見る

似ている issue

C++ の issue をもっと見る

新しい issue をメールで受け取る

初心者向けの GitHub issue を短くまとめたダイジェスト。