This product was not featured by Product Hunt yet.
It will not be visible on their landing page and won't be ranked (cannot win product of the day regardless of upvotes).

Product upvotes vs the next 3

Waiting for data. Loading

Product comments vs the next 3

Waiting for data. Loading

Product upvote speed vs the next 3

Waiting for data. Loading

Product upvotes and comments

Waiting for data. Loading

Product vs the next 3

Loading

Silene Bench

A primary care benchmark you can play yourself

Can an AI agent run a primary care clinic over time? SileneBench is an experimental, playable benchmark exploring that question. Patients return, resources are limited, and rare diseases can hide among everyday cases. Run the clinic yourself or watch GPT-6 Astra make clinical and operational decisions. This hackathon demo is a first step toward a year-long simulation, with support for connecting your own agent harness planned.

Top comment

Hey everyone! I’m Adrian, a researcher with a PhD background in machine learning for healthcare. My previous work includes using ML to support the diagnosis of rare diseases.

With SileneBench, I’m exploring a question: can an AI agent make good clinical and operational decisions while running a primary care practice over time?

Patients return, appointment slots are limited, and earlier decisions affect what happens next. I’m particularly interested in whether an agent can manage everyday care while picking up on rare diseases as evidence gradually emerges.

You can play the demo yourself or watch GPT-6 Astra run the clinic. I wanted humans to be able to play the same environment from the start, with the longer-term goal of comparing human and agent performance.

This is an early hackathon demo. The goal is a roughly year-long simulation, including chronic disease management and consequences that unfold over months. There’s still work ahead on calibration with doctors and evaluation. Support for connecting your own agent harness is planned, but isn’t available yet.

If you give it a try, I’d love to hear whether the simulation feels too easy and what would make it more challenging. Would you be interested in connecting your own agent harness to SileneBench? And does the idea of a visual benchmark, where you can watch an agent’s decisions play out or try it yourself, appeal to you?

About Silene Bench on Product Hunt

A primary care benchmark you can play yourself

Silene Bench was submitted on Product Hunt and earned 4 upvotes and 1 comments, placing #157 on the daily leaderboard. Can an AI agent run a primary care clinic over time? SileneBench is an experimental, playable benchmark exploring that question. Patients return, resources are limited, and rare diseases can hide among everyday cases. Run the clinic yourself or watch GPT-6 Astra make clinical and operational decisions. This hackathon demo is a first step toward a year-long simulation, with support for connecting your own agent harness planned.

On the analytics side, Silene Bench competes within Artificial Intelligence, Games, Medical and OpenAI Day — topics that collectively have 581.5k followers on Product Hunt. The dashboard above tracks how Silene Bench performed against the three products that launched closest to it on the same day.

Who hunted Silene Bench?

Silene Bench was hunted by Adrian Michalski. A “hunter” on Product Hunt is the community member who submits a product to the platform — uploading the images, the link, and tagging the makers behind it. Hunters typically write the first comment explaining why a product is worth attention, and their followers are notified the moment they post. Around 79% of featured launches on Product Hunt are self-hunted by their makers, but a well-known hunter still acts as a signal of quality to the rest of the community. See the full all-time top hunters leaderboard to discover who is shaping the Product Hunt ecosystem.

For a complete overview of Silene Bench including community comment highlights and product details, visit the product overview.