All versions · Version range: v0.4.2
vLLM v0.4.2 Performance and Tuning
- Type
- docs_page
- Publisher
- vllm-project
- Relationship
- first_party
- Version
- vLLM v0.4.2
- Retrieved
- May 26
- Local snapshot
- corpus/domains/vllm/core-v0/evidence/snapshots/vllm-docs-v042-performance.html
- SHA-256
- b72ea4b3dd80da23…
Claims citing this source (1)
| Claim | Locator | Verified |
|---|---|---|
The vLLM v0.4.2 performance docs state that 512 is the default max_num_batched_tokens when enable_chunked_prefill=True. | vllm-docs-v042-performance.html:419-424 | 1 day ago |