CTRL ALTSCtrl Alt Solvereal-world developer knowledge
SearchJoin
Ctrl Alt Solve — Don't just ask what to do. Learn what developers experienced when they did it.
M

Meena Sri

@meenasri

SDE-2

0 followers1 following0 expertise score6 posts

Weekend open-source contributor. Weekday production warrior. (telugu)

ReactNode.jsMongoDB

Experiences shared

Problem solving9d ago·1 entry

broke in background jobs during a Friday deploy

We ran into this around background jobs during a Friday deploy. Incidents were a scavenger hunt across three dashboards. We needed traces and logs without a platform project. OpenTelemetry was the shape. The vendor was the real decision. A weekly budget cap and one starter dashboard got the team using it. Juniors could follow a request without asking who owned the graphs.

0 views0 likes0 comments0 bookmarks
#background jobs
MMeena Sri
Read →
Problem solving18d ago·1 entry

a silent 502 that only hit 2% of traffic

No spike in CPU. Error budgets looked fine at a glance. Users still reported blank pages in a thin slice of traffic. Logs only showed upstream resets with no clear application exception. The culprit was a stale keep-alive timeout between nginx and the app. Aligning idle timeouts stopped the intermittent 502s within an hour. We also added a dashboard for upstream reset reasons so the next page is faster.

337 views0 likes0 comments0 bookmarks
#debugging#nginx#networking
Meena Sri
Case study19d ago·1 entry

a monolith endpoint without a big-bang rewrite

We extracted one high-churn billing endpoint behind a strangler facade. Dual-writes ran for two weeks while we compared totals nightly. A feature flag controlled read traffic so we could roll back instantly. The hardest part was matching edge-case rounding in legacy invoices. Cutover finished with no customer-facing downtime and a smaller blast radius. We kept the facade until three more endpoints followed the same path.

18 views0 likes0 comments0 bookmarks
#billing#migration#architecture
Meena Sri
Comparison20d ago·1 entry

vs Postgres for short-lived job locks

We needed locks so queue workers did not process the same job twice. Redis SET NX was faster under load and easy to expire automatically. Postgres advisory locks were simpler operationally for our small team. Failover behavior mattered more than raw latency in our case. We chose Postgres first, then moved hot paths to Redis later. Pick the lock store you can operate confidently at 3am.

367 views0 likes0 comments0 bookmarks
#queues#postgres#redis
MMeena Sri
Discussion21d ago·1 entry

typed API clients reduce bugs more than OpenAPI docs?

Curious how teams balance generated clients and hand-written SDKs. OpenAPI docs help humans, but drift still sneaks into multi-repo setups. Generated clients catch breaking changes in CI before they hit prod. They can also create noisy diffs when schemas change often. Plain fetch wrappers stay flexible but hide contract mismatches. What has actually reduced production bugs on your teams?

167 views0 likes0 comments0 bookmarks
#api#typescript#dx
Meena Sri
Recommendation22d ago·1 entry

for a practical observability stack for a 6-person team

We need traces and logs without hiring a full-time SRE. Right now we stitch screenshots from three tools during incidents. OpenTelemetry looks right, but the vendor choice is unclear. We ship weekly and cannot afford a six-month platform project. What has worked for small teams that still sleep at night? Especially interested in cost ceilings and onboarding time for juniors.

352 views0 likes0 comments0 bookmarks
#sre#opentelemetry#observability
Meena Sri
M
Read →
M
Read →
Read →
M
Read →
M
Read →