Talk to an AI Agent with a real face and voice, in real time
Human AI Agents have a real face and a real voice, and run live conversation rather than turn-based exchange. Interrupt mid-sentence, trail off, talk over it, and the endpointing holds. Setup is one still photo, a persona and a voice. No rig, no capture session, no script. Two face models behind a single API. Portrait for scale, Presence for expressiveness. Both drop into Pipecat and LiveKit. Try to interrupt it. Most demos cannot survive that, and it is the fastest way to judge this one.
I have sat through a lot of AI demos where the face looks perfect but the conversation is unstable.
You say something, it waits. You pause to think, it talks over you. You interrupt, and it finishes its sentence anyway.
Everyone nods and nobody says the obvious thing, which is that this is not a conversation.
So we built Human AI Agents.
What it is
One still photo, a persona written in plain language, and a voice. You get an agent you can talk to and interrupt.
What runs underneath
Two face models behind a single API. Portrait for speed and scale, Presence for expressiveness. Both stream over WebSocket and drop into Pipecat or LiveKit.
What we actually spent the time on
Turn-taking. Knowing when someone has finished a sentence rather than paused to think. It is the unglamorous part and it is most of the product.
Try to break it. Interrupt it mid-sentence, talk over it, trail off, use an accent, most of all have fun!
I'll be around all day.
About Ojin on Product Hunt
“Talk to an AI Agent with a real face and voice, in real time”
Ojin launched on Product Hunt on August 27th, 2026 and earned 105 upvotes and 20 comments, placing #16 on the daily leaderboard. Human AI Agents have a real face and a real voice, and run live conversation rather than turn-based exchange. Interrupt mid-sentence, trail off, talk over it, and the endpointing holds. Setup is one still photo, a persona and a voice. No rig, no capture session, no script. Two face models behind a single API. Portrait for scale, Presence for expressiveness. Both drop into Pipecat and LiveKit. Try to interrupt it. Most demos cannot survive that, and it is the fastest way to judge this one.
On the analytics side, Ojin competes within Artificial Intelligence — topics that collectively have 477.4k followers on Product Hunt. The dashboard above tracks how Ojin performed against the three products that launched closest to it on the same day.
Who hunted Ojin?
Ojin was hunted by Mio. A “hunter” on Product Hunt is the community member who submits a product to the platform — uploading the images, the link, and tagging the makers behind it. Hunters typically write the first comment explaining why a product is worth attention, and their followers are notified the moment they post. Around 79% of featured launches on Product Hunt are self-hunted by their makers, but a well-known hunter still acts as a signal of quality to the rest of the community. See the full all-time top hunters leaderboard to discover who is shaping the Product Hunt ecosystem.
Hi Product Hunt. I am Mio, founder of Ojin.
I have sat through a lot of AI demos where the face looks perfect but the conversation is unstable.
You say something, it waits. You pause to think, it talks over you. You interrupt, and it finishes its sentence anyway.
Everyone nods and nobody says the obvious thing, which is that this is not a conversation.
So we built Human AI Agents.
What it is
One still photo, a persona written in plain language, and a voice. You get an agent you can talk to and interrupt.
What runs underneath
Two face models behind a single API. Portrait for speed and scale, Presence for expressiveness. Both stream over WebSocket and drop into Pipecat or LiveKit.
What we actually spent the time on
Turn-taking. Knowing when someone has finished a sentence rather than paused to think. It is the unglamorous part and it is most of the product.
Try to break it. Interrupt it mid-sentence, talk over it, trail off, use an accent, most of all have fun!
I'll be around all day.