3
Incident4 Inc17 WarnUpd 1s
Platform/Health

Platform Observability

Metrics, logs, traces, synthetic checks, queue lag, data freshness and cost of the observability platform itself.

App health API

…

DB latency

—

DB connected

No

Internal service health

ServiceStatusSignal
API gatewayHealthyHealthy12 ms
Webhook ingestionHealthyHealthy48 ms
Streaming ingestionDegradedDegraded210 ms
Normalization workersHealthyHealthy90 ms
Evaluation queueHealthyHealthylag 4s
Search indexHealthyHealthylag 12s
Object storageHealthyHealthy35 ms
Postgres primaryHealthyHealthy2 ms
RedisHealthyHealthy1 ms
Billing meterHealthyHealthylag 1s

Detect

Missing provider eventsIngestion backlogFailed webhooksDelayed call processingEvaluation failureStorage failureSearch-index lagIntegration-token expiryData-quality degradation

Service objectives (configurable)

Webhook ack P95

configurable · validate under load

< 300ms

Live dashboard update

near-real-time

< 5s

Post-call processing

async pipeline

shortly after end

Search common queries

P95 objective

< 2s

Dashboard load

common views

< 3s

Tenant isolation

hard requirement

zero cross-tenant

Evaluator outage

queue + skip policy

graceful degrade