Service Telemetry & Operations Dashboard

Data Source: Production Gateway Cluster · Streaming Mode: Active
Total Volume
128,940
↑ +14.2% vs baseline period
SLA Availability
99.8%
SLA Target Met
Active Alerts
3 items
2 pending, 1 auto-recovering
Avg Latency (P95)
42ms
↓ -8ms stable post-optimization
24h Cluster Request Throughput (TPS)
Avg: 1,420 TPS · Peak: 2,890 TPS
+14.8%
2,890 00:00 08:00 16:00 23:59
Core Service Latency Distribution (ms)
P95 Benchmark across 6 services
6 Nodes
24 Gateway 48 Auth 92 Order 35 Payment 18 Storage 64 Search
Item ID Service / Description Cluster Module Throughput Updated Status Action
ITEM-2026-001 Core Engine API Cluster (Primary) Core Gateway 1.28M req/s Just now Healthy
ITEM-2026-002 Async Queue Worker Group (Spike) Message Bus 45.2K msg/s 3m ago Warning
ITEM-2026-003 Vector DB Cache Partition (98% hit) Cache Tier 12.8 GB 12m ago Healthy
ITEM-2026-004 Batch Data Ingestion Pipeline (Daily) Sync Service 850 GB total 2h ago Finished
Create Service Entity