AI Accuracy (weighted avg.)
96.4%
1.4pt above target · trending up
Classification
97.8%
Generative
94.9%
Multimodal
95.6%
Updated 2 minutes ago · 30-day rolling window
AI Accuracy (weighted avg.)
1.4pt above target · trending up
Classification
97.8%
Generative
94.9%
Multimodal
95.6%
Updated 2 minutes ago · 30-day rolling window
Model Performance Score
7-day trend
Training Jobs (30d)
GPU Cluster Usage
Across 4 clusters · 32 GPUs
Inference Requests
+14.2% vs last week
6,214 req/min avg.
Generated today, 09:14 AM · based on last 30 days of platform telemetry
Platform-wide accuracy climbed to 96.4% this month, led by the support-tone-v3 fine-tune rollout. However, GPU cluster node-01 is running at 92% utilization with rising thermals, and Llama-3.1-70B shows a 2.3pt accuracy regression since the last dataset refresh. Token spend is trending to exceed the monthly budget by Aug 3 unless routing is optimized. Overall system risk remains low, and confidence in current model rankings is high given consistent evaluation coverage.
Reroute Vision traffic off node-01 before thermal throttling impacts latency SLA.
Shift summarization traffic to Claude-Haiku-4 to cut costs by ~$620/mo with minimal quality loss.
Consider rolling back Llama-3.1-70B to the pre-refresh checkpoint pending root-cause review.
Top by accuracy
Claude-Sonnet
Top by speed
GPT-4o-mini
Top by cost
Gemini-1.5
Requests per day, last 14 days
By workload type, this month
Jobs by outcome, last 6 weeks
| Week | Done | Failed | Running |
|---|---|---|---|
| W1 | 12 | 2 | 3 |
| W2 | 15 | 1 | 4 |
| W3 | 18 | 3 | 2 |
| W4 | 14 | 2 | 5 |
| W5 | 19 | 1 | 4 |
| W6 | 22 | 2 | 5 |
Milliseconds per request
Fastest model
Embed-3-lg
Slowest model
Whisper-lg
| Job Name | Base Model | Progress | Started | Status |
|---|---|---|---|---|
| support-tone-v4 | GPT-4o-mini | 78% |
2h ago | Training |
| legal-doc-summarizer-v2 | Claude-Haiku-4 | Queued |
— | Queued |
| churn-risk-classifier-v1 | text-embedding-3-large | 100% |
1d ago | Complete |
| sentiment-tuned-v3 | GPT-4o-mini | Failed |
3h ago | Failed · OOM |
| onboarding-assistant-v1 | Claude-Sonnet-4.5 | 34% |
48m ago | Training |
| fraud-detector-v6 | Gemini-1.5-Pro | 100% |
2d ago | Deployed |
Vectors
42.7M
Namespaces
3
Index size
218 GB
Recall@10
98.2%
GPU node-01 thermal throttling risk
92% utilization · 2 min ago
Llama-3.1-70B accuracy regression
-2.3pt since dataset refresh · 1h ago
Token spend trending over budget
Projected overage by Aug 3 · 3h ago
Critical
1
Warning
2
Resolved (7d)
14
Last 14 days
down from 1.8% · 2 weeks ago
"Can you walk me through the refund process for..."
Resolved"How do I connect my Stripe account to the..."
In progress"My invoice total doesn't match what was quoted..."
EscalatedQuery funnel · last 24 hours · 214,880 queries
100% of total volume
-5.4% dropped · empty index results
-9.1% dropped · below similarity cutoff
82.0% end-to-end retrieval yield
Recall@10
98.2%
Avg. chunks/query
6.4
Index freshness
11m
Multi-step agent runs by outcome, last 24 hours
Avg. steps/run
4.7
Avg. run time
38s
Cost saved (24h)
$1,240