Schema.ai· vLLM
vLLM official site ↗vLLM v0.23.0API up

enable_chunked_prefill

SchedulerConfig flag enabling chunking of prefill requests based on the remaining max_num_batched_tokens budget.

SettingvLLM v0.23.0verified Jul 9

Claims (0)

No claims about this entity yet.

Availability

enable_chunked_prefill is exposed on SchedulerConfig in the v0.23.0 source snapshot.
vllm.config.scheduler_configcomplete_inventorySchedulerConfig dataclass field surface in vllm/config/scheduler.py

vLLM v0.23.0 config scheduler.py · vllm/config/scheduler.py:84-90