Live, readable telemetry from every vLLM-like Docker workload discovered on AI01. Pick a container to inspect its runtime details and watch the log flow by subsystem.
connecting
HOST AI01
LAST SAMPLE waiting
log refresh every 3s
State
---
waiting for Docker inventory
Model
---
served model
Image
---
Docker image
Log lines
0
bounded tail
Latest observed telemetry
Performance
waiting for metrics
Prefill / prompt
not reported
latest prompt throughput
Decode / generation
not reported
latest generation throughput
GPU KV cache
not reported
current cache utilization
Prefix cache hit
not reported
internal and external hit rate
MTP / speculative
not reported
depth and acceptance
LMCache tokens
not reported
hit, computed, and load
CKV / cache transfer
not reported
CKV or transfer activity
Requests
not reported
running and waiting
Values are the newest matching measurements in the fetched log tail.