Skip to content

Observability

How I know what my services are doing in production.

ToolStatusWhen I reach for it
PrometheusdailyScraping and storing time-series metrics from every service
GrafanadailyDashboards and alerts on top of Prometheus and Loki
pprofdailyGo CPU and memory profiling during performance investigations
SentrydailyError tracking and stack traces for fast incident triage
OpenTelemetrylearningVendor-neutral instrumentation for traces, metrics, and logs
LokilearningLog aggregation that integrates natively with Grafana
JaegerfutureDistributed tracing UI when full OTel trace visualisation is needed

Edit to match your stack.