Hacktoberfest 2026: le issue che i maintainer hanno segnato per ottobre, aperte e adatte ai principianti. Sfoglia le issue Hacktoberfest

Add support for experimental wheel variants (i.e., wheelnext)

Aperta
#2,092 0 commenti 0 reazioni 0 assegnatari Vedi su GitHub

I maintainer di solito rispondono entro 1 giorno

Nessuno ha ancora preso questa issue.

Valutazione

Difficoltà
4/5
Tempo stimato
3-5 giorni
Idoneità per principianti
45/100
Tipo di issue
Funzionalità
Chiarezza
Abbastanza chiara
Stato di attività
Ferma
Stack tecnologico
python

Direzione di ricerca

La issue riguarda la modifica del processo di build e pubblicazione delle wheel. Inizia esaminando gli script di build del progetto, probabilmente in setup.py o pyproject.toml, e i workflow CI/CD. Studia la specifica WheelNext e il modo in cui progetti come PyTorch implementano i metadati delle varianti. L'obiettivo è produrre wheel specifiche per il backend (CUDA, ROCm, Metal) con i metadati corretti, assicurandosi che le wheel CPU rimangano disponibili come fallback. I test consisteranno nella compilazione locale delle wheel e nella verifica dei metadati.

Scritto dal modello di indicizzazione a partire dal testo della issue.

Descrizione

Is your feature request related to a problem? Please describe.
Today, installing llama-cpp-python on machines with different GPU backends (CUDA, ROCm, Metal, etc.) requires separate package names, custom extra indexes, or installer-level logic to select the correct wheel. This creates friction for downstream tooling (CLIs, orchestrators, and packaging systems) that want to provide a “just works” experience, especially when users don’t know which backend they need. Even a simple developer-driven install might require picking precisely the correct wheel.

Describe the solution you'd like
Add support for WheelNext-compatible experimental wheel variants when building and publishing wheels.

This would allow llama-cpp-python to produce a single package version that provides multiple backend-aware binary wheels, each annotated with variant metadata (e.g., GPU type, CUDA version, ROCm version).

Installers that understand the WheelNext spec (now used experimentally by PyTorch, uv, and others) can automatically select the correct backend wheel based on the system’s hardware/software configuration without a need for custom index URLs, separate packages, or manual backend flags.

Key pieces:

  • Generate wheels with variant metadata following the experimental WheelNext (wheel variants) conventions.
  • Publish per-backend wheels using the standardized naming + metadata fields.
  • Ensure that CPU-only wheels remain available as fallback.

This would significantly simplify installation for all users and remove backend-selection logic from downstream tools. Wheel variants are fully backward-compatible so existing workflows won't be disrupted.

Describe alternatives you've considered

  • Separate package names per backend (e.g., llama-cpp-python-cuda): fragments packaging and forces manual selection.
  • Extras for backend variants (pip install llama-cpp-python[cuda]): still requires external detection and doesn’t integrate with hardware-aware installer selection.
  • Custom index URLs for backend wheels: brittle and requires orchestration logic outside Python packaging.
  • CLI-backed installation routing (what many downstream projects do currently): it’s reinventing the wheel and provides an inconsistent experience for end users.

All of these solutions put the burden on downstream tooling rather than on standardized wheel metadata.

Additional context

Lingua principale
Python
Stelle
10.6k
Fork
1.5k
Merge medio
3h 57m
PR unite (30g)
4

Preparare l'ambiente

Come iniziare

  1. Leggi tutta la issue e poi la guida ai contributi del progetto.
  2. Commenta sulla issue per dire che te ne occupi tu — evita che due persone facciano lo stesso lavoro.
  3. Fai un fork del repository e lavora su un branch.
  4. Apri una pull request che faccia riferimento al numero della issue.

Altre issue di abetlen/llama-cpp-python

Tutte le issue di abetlen/llama-cpp-python

Issue simili

Altre issue su Python

Ricevi le nuove issue nella tua casella

Un breve riepilogo di issue GitHub adatte ai principianti.