[FEA]: Support CU_LAUNCH_ATTRIBUTE_PREFERRED_CLUSTER_DIMENSION in LaunchConfig
Los mantenedores suelen responder en 1 día
@lijinf2 ya está trabajando en esto.
Desde el 18/8/2026.
Evaluación
Este issue todavía no se ha evaluado.
Descripción
Is this a duplicate?
- I confirmed there appear to be no duplicate issues for this request and that I agree to the Code of Conduct
Area
cuda.core
Is your feature request related to a problem? Please describe.
LaunchConfig does not expose CU_LAUNCH_ATTRIBUTE_PREFERRED_CLUSTER_DIMENSION, which sets preferred thread-block cluster dimensions for a launch (preferred size must be a multiple of the minimum cluster dimension).
Parent tracking: #496.
Describe the solution you'd like
- Users can set
CU_LAUNCH_ATTRIBUTE_PREFERRED_CLUSTER_DIMENSIONviaLaunchConfig. - Mapping to the native launch attribute is tested.
Describe alternatives you've considered
Use cuda.bindings.driver (CUlaunchAttribute / CU_LAUNCH_ATTRIBUTE_PREFERRED_CLUSTER_DIMENSION) and cuLaunchKernelEx directly.
Downside: leaves cuda.core users without a first-class LaunchConfig API; they must drop to low-level bindings, manage the attribute array themselves, and lose consistency with other LaunchConfig launch attributes (is_cooperative, programmatic_stream_serialization, etc.).
Additional context
Incremental LaunchConfig attribute coverage under #496 (same approach as #1334 for programmatic_stream_serialization).
No example code in the design notes for this attribute.
- Lenguaje dominante
- Cython
- Estrellas
- 3.4k
- Forks
- 334
- Merge medio
- 1 d 17 h
- PR fusionados (30 d)
- 123
Preparar el entorno
- Sin Dockerfile ni archivo de Docker Compose
- Tiene una plantilla de pull request
- Leer la guía de contribución
Primeros pasos
- Lee el issue completo y luego la guía de contribución del proyecto.
- Comenta en el issue que vas a ocuparte — evita que dos personas hagan lo mismo.
- Haz un fork del repositorio y trabaja en una rama.
- Abre un pull request que haga referencia al número del issue.
Más de NVIDIA/cuda-python
-
triage
Dificultad 2/5 1-3 horas Aptitud para principiantes 70/100
NVIDIA/cuda-python#3015 · 2 comentarios ·
Los mantenedores suelen responder en 1 día
-
triage
Dificultad 2/5 1-3 horas Aptitud para principiantes 84/100
NVIDIA/cuda-python#2952 ·
Los mantenedores suelen responder en 1 día
-
[DOC]: cuda.core 1.1.1 note misstates program cache permissionsPosiblemente ocupada @leofang la tomó hace 11 días. Abiertodocumentation P1
Dificultad 1/5 Menos de una hora Aptitud para principiantes 88/100
NVIDIA/cuda-python#2717 · 1 asignado ·
Los mantenedores suelen responder en 1 día
-
[DOC]: `PinnedMemoryResource.allocate` documents no parametersPosiblemente ocupada @Andy-Jost la tomó hace 11 días. Abiertocuda.core documentation P1
Dificultad 1/5 1-3 horas Aptitud para principiantes 90/100
NVIDIA/cuda-python#2712 · 1 asignado ·
Los mantenedores suelen responder en 1 día
-
[BUG]: LocatedHeaderDir is mutable, so callers can poison the cached header-directory lookupPosiblemente ocupada @rwgk la tomó hace 11 días. Abiertotriage
Dificultad 2/5 1-3 horas Aptitud para principiantes 82/100
NVIDIA/cuda-python#2646 · 1 reacción ·
Los mantenedores suelen responder en 1 día