Platform/Health
Platform Observability
Metrics, logs, traces, synthetic checks, queue lag, data freshness and cost of the observability platform itself.
App health API
…
DB latency
—
DB connected
No
Internal service health
| Service | Status | Signal |
|---|---|---|
| API gateway | healthy | 12 ms |
| Webhook ingestion | healthy | 48 ms |
| Streaming ingestion | degraded | 210 ms |
| Normalization workers | healthy | 90 ms |
| Evaluation queue | healthy | lag 4s |
| Search index | healthy | lag 12s |
| Object storage | healthy | 35 ms |
| Postgres primary | healthy | 2 ms |
| Redis | healthy | 1 ms |
| Billing meter | healthy | lag 1s |
Detect
Missing provider eventsIngestion backlogFailed webhooksDelayed call processingEvaluation failureStorage failureSearch-index lagIntegration-token expiryData-quality degradation
Service objectives (configurable)
Webhook ack P95
configurable · validate under load
Live dashboard update
near-real-time
Post-call processing
async pipeline
Search common queries
P95 objective
Dashboard load
common views
Tenant isolation
hard requirement
Evaluator outage
queue + skip policy