vllm.v1.core.sched.diffusion_scheduler
¶
Canvas-width handling and async read deferral for diffusion requests.
Functions:
-
diffusion_canvas_width–The canvas width a diffusion request asked for, else the served one.
_read_in_flight(request, width)
¶
True when a read-only request has all its denoise steps in flight.
The request emits its canvas on the last step and ends, so a further step is discarded.
Source code in vllm/v1/core/sched/diffusion_scheduler.py
diffusion_canvas_width(request, canvas_length)
¶
The canvas width a diffusion request asked for, else the served one.