Log Aggregation & APM
updated 20:51:19

Service Health

One screen tells you what is broken right now — across logs, metrics and traces.

Services up
1/3
Request rate
35.30/s
Error rate
4.7%
Worst p95
1.10s
Request rate
req/s by service · 1h
Error rate
% of 5xx · 1h
Log volume
lines/min · all sources · 1h
Services
status · throughput · errors · latency · resources
Metrics →
ServiceStatusReq/sErrorsp50p95CPUMem
apiDegraded5.675.0%437ms1.10s0.0494 MB
paymentsDegraded11.717.5%397ms930ms0.06113 MB
worker8× restartOperational17.922.8%191ms568ms0.06121 MB
Active incidents
auto-detected · 6h
All →
  • worker restarted 19×
    Process start time changed 19 time(s) in the last 6h.
    5s ago
  • postgres-yckw0wkg8k4co0osg8kwco0k-145449955773 crashed
    2026-08-08 20:51:11.364 UTC [4164848] FATAL: database "cid_user" does not exist
    8s ago
  • postgres-yckw0wkg8k4co0osg8kwco0k-145449955773 crashed
    2026-08-08 20:50:56.162 UTC [4164826] FATAL: database "cid_user" does not exist
    23s ago
  • postgres-yckw0wkg8k4co0osg8kwco0k-145449955773 crashed
    2026-08-08 20:49:55.306 UTC [4164739] FATAL: database "cid_user" does not exist
    1m ago
  • postgres-yckw0wkg8k4co0osg8kwco0k-145449955773 crashed
    2026-08-08 20:48:59.476 UTC [4164659] FATAL: database "cid_user" does not exist
    2m ago
  • worker crashed
    FATAL: container out of memory (OOM) — killing process
    3m ago