Schema.ai
· vLLM
vLLM official site
Log in
All versions · Version range: v0.4.2–v0.23.0
vllm
/
settings
Settings
Search
3 entities matching "source.vllm.code.config.v042"
Name
Description
Version
Verified
Sources
max_num_batched_tokens
Maximum number of tokens to process in a scheduler iteration.
v0.4.2, v0.23.0
1 day ago
snapshots/vllm-code-v042-config.py
config/scheduler.py
snapshots/vllm-docs-v042-performance.html
configuration/optimization.md
max_num_seqs
Maximum number of sequences to process in a scheduler iteration.
v0.4.2, v0.23.0
1 day ago
snapshots/vllm-code-v042-config.py
config/scheduler.py
configuration/optimization.md
max_model_len
Maximum sequence length, including prompt and output tokens, used by vLLM configuration and scheduler validation.
v0.4.2, v0.23.0
Jul 9
snapshots/vllm-code-v042-config.py
config/scheduler.py