AI agents make CI pass, but a green check does not prove the failure was fixed. Sutura is a GitHub Action and CLI that reproduces the failure in an isolated sandbox, separates flakes from real failures, searches bounded repairs, rejects green-wash (deleted tests, weakened assertions, relaxed config), then puts the winner through an adversarial audit by NVIDIA Nemotron with GPT-6 Astra and TypeSafe Jev as veto-only second opinions. It opens an evidence-backed PR for a human. Never auto-merges.
Hi Product Hunt, Juan here. A few honest notes before you click around.
What runs where. The repair model is NVIDIA Nemotron on Nebius Token Factory, which is the hackathon this was built for. GPT-6 Astra sits in the audit gate as a second opinion: it re-runs the same adversarial question and can only reject a repair, never approve one. A third voice, TypeSafe's Jev, answers the same question as a typed choice with a calibrated probability and confidence, and it can only veto too. All three are visible as rows in every result.
What the Case Lab is. Five fixed cases you can read right now, each with a deterministic replay. Live runs against the public demo repository are capped at 4 per hour and 24 per day; when the cap is hit you still get the recorded result. Today's live smoke result, with all three audit voices on the record: https://sutura-case-lab.vercel.a...
What the numbers are. The v0.3.1 release benchmark ran all 51 Placebo cases on the release commit: zero false approvals, 17 of 19 green-wash traps refused, 10 of 18 repairable failures fixed, 10 of 10 flaky cases correctly left unpatched, USD 3.77 total. Two provider infrastructure stops happened during the run and are disclosed. Evidence files and every workflow URL: https://github.com/juan294/sutur...
What the GPT-6 Astra Challenge was like. Astra was the coding model this project started with, and that is why it ended up inside the product too: it is now the second auditor on every repair, wired in this week and running live on the public demo. Its job is to disagree with Nemotron when Nemotron is wrong, and the evidence shows when it does. Built with Astra, and running Astra.
Ask me anything about the verification gates, the benchmark, or why it refuses to auto-merge.
No comment highlights available yet. Please check back later!
About Sutura on Product Hunt
“Verified self-healing CI that proves the fix”
Sutura launched on Product Hunt on September 18th, 2026 and earned 0 upvotes and 1 comments, placing #20 on the daily leaderboard. AI agents make CI pass, but a green check does not prove the failure was fixed. Sutura is a GitHub Action and CLI that reproduces the failure in an isolated sandbox, separates flakes from real failures, searches bounded repairs, rejects green-wash (deleted tests, weakened assertions, relaxed config), then puts the winner through an adversarial audit by NVIDIA Nemotron with GPT-6 Astra and TypeSafe Jev as veto-only second opinions. It opens an evidence-backed PR for a human. Never auto-merges.
Sutura was featured in Open Source (68.8k followers), Developer Tools (519.6k followers), GitHub (41.4k followers) and OpenAI Day (38 followers) on Product Hunt. Together, these topics include over 130.3k products, making this a competitive space to launch in.
Who hunted Sutura?
Sutura was hunted by Juan González Ponce. A “hunter” on Product Hunt is the community member who submits a product to the platform — uploading the images, the link, and tagging the makers behind it. Hunters typically write the first comment explaining why a product is worth attention, and their followers are notified the moment they post. Around 79% of featured launches on Product Hunt are self-hunted by their makers, but a well-known hunter still acts as a signal of quality to the rest of the community. See the full all-time top hunters leaderboard to discover who is shaping the Product Hunt ecosystem.
Want to see how Sutura stacked up against nearby launches in real time? Check out the live launch dashboard for upvote speed charts, proximity comparisons, and more analytics.