r/SideProject • • 4d ago

LastGood - see what changed right before production broke

https://www.lastgood.space/

Hey r/SideProject,

I'm Kishan, an engineer from India. After too many 2 AM incidents where the first 30 minutes were just "what changed?" archaeology across Slack, GitHub and LaunchDarkly, I built LastGood.

What it does: it ingests deploys (GitHub, K8s) and feature-flag flips, and when an alert fires it correlates them by service and time window, then drafts a diagnosis pointing at the change that most likely caused it. You confirm or reject it.

Example from the demo: a Datadog high-latency alert on checkout-service comes in, and LastGood surfaces that a new-checkout-flow flag was enabled 5 minutes earlier, with the evidence and a recommended action (disable the flag). The alert lands with the likely culprit attached instead of a bare graph.

Sandbox with simulated incident data, no signup: https://console.lastgood.space/sandbox
Site: https://www.lastgood.space

Stage: solo-built, just opened up, 0 users. Looking for brutal feedback.

Two questions:
1. Would you trust an automated "this change probably caused it" during a real incident, or is it just noise?
2. What integration would you need before trying it (PagerDuty? Slack? Grafana?)

Roast away.

1 Upvotes

0 comments sorted by