vllm:num_requests_running
Gauge for the number of requests in model execution batches.
MetricPrometheusvLLM v0.22.1
Semantic coverage
gap · Availability & namegap · Structural identitygap · Measurement semanticsgap · Causal/runtime semanticsgap · Interpretation & compositiongap · Guarded guidance