Schema.ai
· vLLM
vLLM v0.23.0
API up
vllm
/
topics
Topics
Search
Topic
Summary
Claims
Surfaces
Comprehensiveness
Prefix cache effectiveness metrics
Coverage for the vLLM v0.23.0 first-party prefix-cache counters and how to compute and interpret the prefix cache hit rate.
5
0
seed slice
Scheduler and KV pressure metrics
Coverage for a small vLLM v0.23.0 metric slice that exposes scheduler queue state, KV cache usage, and preemption signals.
8
1
seed slice