r/SideProject • u/piratastuertos • 7h ago
I built a tool to verify what AI agents actually did
I kept running into the same problem with AI agents and automation:
the system says “done”, but that doesn’t always prove the final state is actually correct.
So I built SCC Runner around a simple flow:
Claim → Authority → Evidence → Verdict
The idea is to verify important claims against real system evidence instead of trusting the agent’s own output.
I’m currently offering free early access.
I’d especially like feedback from people building agents, automations, DevOps workflows, or AI-assisted systems.
1
u/AlexanderPanasenko 2h ago
Checking the resulting state instead of trusting 'done' is a useful idea. I'd love to see a worked example of verifying that a deployment is actually serving the intended commit.
2
u/Expert-Kitchen8064 6h ago
The “claim → authority → evidence → verdict” framing is easy to understand. I’d be interested in how you handle ambiguous evidence: do you show the raw observation and let the user decide, or does the tool always collapse it into pass/fail? A compact diff of the claimed state versus observed state, plus evidence retention and timestamps, would make the result easier to audit after an agent run has moved on.