Schema.ai· vLLM· Schema 0.2.0
vLLM official site vLLM v0.25.0Log in

Scheduler sequence capacity

seed slicenot evaluated at v0.25.0legacy evidence: v0.4.2legacy evidence: main-2026-07-21evidence: source code and documentation

Coverage

Coverage for max_num_seqs scheduler capacity, configuration exposure, constraints, and KV-cache pressure guidance.

Strong for seeded v0.4.2 SchedulerConfig exposure and constraints; partial for operational tuning guidance.

This topic connects max_num_seqs facts to scheduler capacity and KV-cache pressure, but it is not a full vLLM memory-management model.

Coverage at vLLM v0.25.0

Schema has not evaluated this topic at the selected version.

0 claims0 history facts0/1 entities

Claims in this topic (0)

No claims at this version scope.

Surfaces

Known gaps

The corpus does not yet include runtime telemetry tying max_num_seqs changes to specific preemption rates.

Schema.ai can surface the documented tuning direction but should not prescribe an exact value for a workload.

Check: Add telemetry-backed examples or benchmark evidence before value-specific recommendations.

Adjacent features such as speculative decoding are not covered by this topic.

Questions about those features should return guidance-only until source-grounded claims are added.

Check: Capture and promote source snapshots for each adjacent scheduler feature before including it in catalog coverage.