New: Debug encrypted microservice traffic with Speedscale's eBPF collector Read the announcement

Observability guides

API observability: find the problem and prove the fix

Metrics, logs, and traces tell you where to look. API traffic shows the request, response, and dependency behavior that caused the failure. These guides connect the two.

Start here

How the layers of production visibility fit together

Choose what you need to do

All observability articles

Reliability Engineering in the AI Era

See why reliability engineering must span code, testing, telemetry, and incidents as AI agents erase the boundary between pre-production and production.

The Pod Was Cheaper. The Service Wasn’t.

Use OpenCost and proxymock to prove Kubernetes rightsizing lowers cost per successful request without hiding behavior or throughput regressions in testing.

Flamegraphs Find It. Replay Proves It.

Use Grafana Pyroscope and proxymock with an AI coding agent to find a Go CPU hotspot, preserve API behavior, and verify the performance fix.

Go from alert to reproducible test

Capture the API traffic behind a production issue, replay it safely, and verify the fix against the same behavior.