sgl-project/sglang

[Feature] Detailed Break down of time spend on Launching SGLang Diffusion

Offen

#19.087 geöffnet am 20.02.2026

 (10 Kommentare) (0 Reaktionen) (1 zugewiesene Person)Python (6.216 Forks)auto 404
good first issue

Repository-Metriken

Stars
 (28.442 Sterne)
PR-Merge-Metriken
 (Durchschn. Merge 2T 1h) (1.000 gemergte PRs in 30 T)

Beschreibung

Checklist

Motivation

Diffusion and LLM have huge differences in compute characteristics. We want to have a detailed optimization of the launch time spent on SGLang Diffusion.

In this sense, to optimize the launch time, we should have a detailed breakdown of what is actually taking time when we launch our models. Please use Qwen-Image as an example, and try to break down the time spent. Then let's see whether we shall spend our time on optimize the launching time.

Related resources

No response

Contributor Guide