Requests (24h)
18,330
all models · all tenants+6.2%
Throughput, latency, and spend across owned models and providers.
Requests (24h)
18,330
Throughput
764 / hr
P95 latency
842 ms
Error rate
0.6%
Tokens (24h)
6.1M
Cost (24h)
$22.61
UTC · today
24h
| Model | Reqs | Avg | P95 | Err | Cost |
|---|---|---|---|---|---|
| slm-support-v2 | 6,200 | 280ms | 540ms | 0.3% | $2.40 |
| general-llm-v3 | 4,100 | 420ms | 780ms | 0.4% | $6.80 |
| slm-apiheal-v1 | 3,180 | 210ms | 390ms | 0.2% | $1.60 |
| claude-sonnet | 2,850 | 510ms | 920ms | 0.8% | $8.40 |
| gpt-4o | 2,000 | 390ms | 710ms | 0.5% | $3.41 |