This product was not featured by Product Hunt yet.
It will not be visible on their landing page and won't be ranked (cannot win product of the day regardless of upvotes).

Product upvotes vs the next 3

Waiting for data. Loading

Product comments vs the next 3

Waiting for data. Loading

Product upvote speed vs the next 3

Waiting for data. Loading

Product upvotes and comments

Waiting for data. Loading

Product vs the next 3

Loading

nuhuh

Your AI agent said "Done." nuhuh runs the experiment

Coding agents end most tasks with "Done! All tests pass." Sometimes it's false. nuhuh is a Stop hook that extracts every claim from the agent's final message and re-runs reality, fresh tests, real exit codes, actual files, then bounces a false Done back with the evidence. Local, deterministic, no API key, MIT.

Top comment

Last week my agent told me "All tests pass." One test failed. It had reported the result of a run it never made. So I built a test runner wearing a Stop hook.

nuhuh treats the final message as a list of hypotheses and re-runs each one. The whole suite in a clean process, files checked on disk, localhost actually probed. A false Done gets rejected and the evidence goes straight back to the agent, which returns to work.

It ships with a benchmark whose ground truth doesn't know the tool exists, so it catches nuhuh's own mistakes too. Across 102 runs per model, Codex falsely declared Done 2.1% of the time, Haiku 6.0%, frontier Claude 0%. And every false Done contained zero checkable claims, just confident tone, which is exactly why the benchmark exists.

That finding became a feature. NUHUH_STRICT=1 bounces any completion declaration that carries no checkable claim. Replayed against all 306 runs it catches 7 of 7 false dones, and the honest price is one extra bounce on roughly half the true dones, which is why it's opt-in and aimed at unattended lanes like CI and batch runs.

Try it without installing anything. npx nuhuh demo stages a lie and catches it in ten seconds.

Ask me anything about the false accusations we caught our own tool making, there are six and each one is now a regression test.

About nuhuh on Product Hunt

Your AI agent said "Done." nuhuh runs the experiment

nuhuh was submitted on Product Hunt and earned 2 upvotes and 1 comments, placing #156 on the daily leaderboard. Coding agents end most tasks with "Done! All tests pass." Sometimes it's false. nuhuh is a Stop hook that extracts every claim from the agent's final message and re-runs reality, fresh tests, real exit codes, actual files, then bounces a false Done back with the evidence. Local, deterministic, no API key, MIT.

On the analytics side, nuhuh competes within Open Source, Developer Tools, Artificial Intelligence and GitHub — topics that collectively have 1.1M followers on Product Hunt. The dashboard above tracks how nuhuh performed against the three products that launched closest to it on the same day.

Who hunted nuhuh?

nuhuh was hunted by JinHyuk Sung. A “hunter” on Product Hunt is the community member who submits a product to the platform — uploading the images, the link, and tagging the makers behind it. Hunters typically write the first comment explaining why a product is worth attention, and their followers are notified the moment they post. Around 79% of featured launches on Product Hunt are self-hunted by their makers, but a well-known hunter still acts as a signal of quality to the rest of the community. See the full all-time top hunters leaderboard to discover who is shaping the Product Hunt ecosystem.

For a complete overview of nuhuh including community comment highlights and product details, visit the product overview.