We ran into this around session cookies when the cache went cold. I expected a tooling problem. It was an assumption in the schema. EXPLAIN and a short rollback window saved the afternoon. Writes slowed first, then a handful of reads timed out. Dropping the unused index and batching the backfill unblocked the queue. I now review write paths with the same care as the happy-path query.
We ran into this around session cookies during a traffic spike. Incidents were a scavenger hunt across three dashboards. We needed traces and logs without a platform project. OpenTelemetry was the shape. The vendor was the real decision. A weekly budget cap and one starter dashboard got the team using it. Juniors could follow a request without asking who owned the graphs.
We ran into this around session cookies after we split the monolith. I expected a tooling problem. It was an assumption in the schema. EXPLAIN and a short rollback window saved the afternoon. Writes slowed first, then a handful of reads timed out. Dropping the unused index and batching the backfill unblocked the queue. I now review write paths with the same care as the happy-path query.
We ran into this around session cookies while rolling back payments. Incidents were a scavenger hunt across three dashboards. We needed traces and logs without a platform project. OpenTelemetry was the shape. The vendor was the real decision. A weekly budget cap and one starter dashboard got the team using it. Juniors could follow a request without asking who owned the graphs.
We ran into this around session cookies on a quiet Sunday incident. I expected a tooling problem. It was an assumption in the schema. EXPLAIN and a short rollback window saved the afternoon. Writes slowed first, then a handful of reads timed out. Dropping the unused index and batching the backfill unblocked the queue. I now review write paths with the same care as the happy-path query.
We ran into this around session cookies after the replica failover. Incidents were a scavenger hunt across three dashboards. We needed traces and logs without a platform project. OpenTelemetry was the shape. The vendor was the real decision. A weekly budget cap and one starter dashboard got the team using it. Juniors could follow a request without asking who owned the graphs.
We ran into this around session cookies during a Friday deploy. I expected a tooling problem. It was an assumption in the schema. EXPLAIN and a short rollback window saved the afternoon. Writes slowed first, then a handful of reads timed out. Dropping the unused index and batching the backfill unblocked the queue. I now review write paths with the same care as the happy-path query.