enable_chunked_prefill
SchedulerConfig flag enabling chunking of prefill requests based on the remaining max_num_batched_tokens budget.
SettingvLLM v0.23.0verified Jul 9
Claims (0)
Availability
enable_chunked_prefill is exposed on SchedulerConfig in the v0.23.0 source snapshot.vllm.config.scheduler_configcomplete_inventorySchedulerConfig dataclass field surface in vllm/config/scheduler.py