vllm_omni.diffusion.sched.base_scheduler ¶
BatchSamplingParamsKey module-attribute ¶
BatchSamplingParamsKey = (
StepBatchSamplingParamsKey
| RequestBatchSamplingParamsKey
)
BaseScheduler ¶
Bases: ABC
Shared queue/state bookkeeping for diffusion schedulers.
finish_requests ¶
finish_requests(
request_ids: str | list[str],
status: DiffusionRequestStatus,
) -> None
get_admission_wait_decision ¶
Return the admission-delay policy for the next scheduling wave.
get_diffusion_kv_cleanup_targets ¶
initialize ¶
initialize(
od_config: OmniDiffusionConfig,
*,
kv_cache_config: KVCacheConfig | None = None,
scheduler_block_size: int | None = None,
hash_block_size: int | None = None,
kv_vllm_config: VllmConfig | None = None,
) -> None
native_kv_poll_output ¶
native_kv_poll_output(
*, drain_request_ids: list[str] | None = None
) -> DiffusionSchedulerOutput | None
Poll after compute; cancellation/close may wait for selected loads.
pending_finished_request_ids ¶
Finished requests whose state the engine has not consumed yet.
release_kv_drains ¶
Called only after all ranks completed and Worker row cleanup succeeded.
should_end_admission_wait ¶
should_end_admission_wait(
decision: _AdmissionWaitDecision,
*,
now: float,
stable_since: float,
) -> bool
Return whether an active admission delay should end.
update_from_output abstractmethod ¶
update_from_output(
sched_output: DiffusionSchedulerOutput,
output: BaseRunnerOutput,
) -> set[str]
update_kv_connector_output ¶
SchedulerInterface ¶
Bases: BaseScheduler
Deprecated compatibility base for custom scheduler injection.
Prefer subclassing :class:BaseScheduler directly. Subclassing this name still works but emits a :class:DeprecationWarning.