Human AI Agents have a real face and a real voice, and run live conversation rather than turn-based exchange. Interrupt mid-sentence, trail off, talk over it, and the endpointing holds. Setup is one still photo, a persona and a voice. No rig, no capture session, no script. Two face models behind a single API. Portrait for scale, Presence for expressiveness. Both drop into Pipecat and LiveKit. Try to interrupt it. Most demos cannot survive that, and it is the fastest way to judge this one.
I have sat through a lot of AI demos where the face looks perfect but the conversation is unstable.
You say something, it waits. You pause to think, it talks over you. You interrupt, and it finishes its sentence anyway.
Everyone nods and nobody says the obvious thing, which is that this is not a conversation.
So we built Human AI Agents.
What it is
One still photo, a persona written in plain language, and a voice. You get an agent you can talk to and interrupt.
What runs underneath
Two face models behind a single API. Portrait for speed and scale, Presence for expressiveness. Both stream over WebSocket and drop into Pipecat or LiveKit.
What we actually spent the time on
Turn-taking. Knowing when someone has finished a sentence rather than paused to think. It is the unglamorous part and it is most of the product.
Try to break it. Interrupt it mid-sentence, talk over it, trail off, use an accent, most of all have fun!
I'll be around all day.
Report
Congrats on the launch@iammio quick question on the silence detection did background hum/mic noise cause false triggers during testing?
@vikramp7470 Thanks Vikram! The hardest part wasn't tuning any single detector - it was finding one silence-detection approach that held up across every scenario we throw at it. We tested on-device VAD, cloud-based VAD, and external providers like ai-coustics and Deepgram, and each shined in some conditions and fell apart in others. No single method was robust everywhere. What finally worked was a combination of them rather than picking a winner - leaning on different signals depending on the context. Getting that blend to behave consistently was the real grind.
Report
Congrats on the launch! ✨ Got some nice recommendations from Sofia for travelling and it continues the discussions pretty smoothly when interrupted or when I interject something randomly (and knows to respond to that also). It still is obvious that it's an AI, but the experience is pretty good.
@mad94 This made my day, thank you. Sofia handling travel recommendations and then picking up a random interjection is the exact thing we spent the longest on, so hearing it hold up in the wild is worth more to me than any number on the board today.
You are right that it still reads as AI. If you remember what gave it away, the voice, the face, or the timing between them, I would love to know.
Report
Congrats on the launch! I use LemonSlice right now but am always open to new providers. Do you have a differentiator?
Happy to have contributed to this as a product engineer. It's been a wild few months of building, and it's great to finally see people trying what we made. Huge respect to the team, genuinely talented and great people to build with. Let's go! 🚀
I'm very excited to finally have Ojin released!! A lot of engineering efforts went into making this product possible and I'm so proud of the team. Can't wait for people to try it out!
Ojin
Hi Product Hunt. I am Mio, founder of Ojin.
I have sat through a lot of AI demos where the face looks perfect but the conversation is unstable.
You say something, it waits. You pause to think, it talks over you. You interrupt, and it finishes its sentence anyway.
Everyone nods and nobody says the obvious thing, which is that this is not a conversation.
So we built Human AI Agents.
What it is
One still photo, a persona written in plain language, and a voice. You get an agent you can talk to and interrupt.
What runs underneath
Two face models behind a single API. Portrait for speed and scale, Presence for expressiveness. Both stream over WebSocket and drop into Pipecat or LiveKit.
What we actually spent the time on
Turn-taking. Knowing when someone has finished a sentence rather than paused to think. It is the unglamorous part and it is most of the product.
Try to break it. Interrupt it mid-sentence, talk over it, trail off, use an accent, most of all have fun!
I'll be around all day.
Congrats on the launch@iammio quick question on the silence detection did background hum/mic noise cause false triggers during testing?
Ojin
@vikramp7470 Thanks Vikram! The hardest part wasn't tuning any single detector - it was finding one silence-detection approach that held up across every scenario we throw at it. We tested on-device VAD, cloud-based VAD, and external providers like ai-coustics and Deepgram, and each shined in some conditions and fell apart in others. No single method was robust everywhere. What finally worked was a combination of them rather than picking a winner - leaning on different signals depending on the context. Getting that blend to behave consistently was the real grind.
Congrats on the launch! ✨ Got some nice recommendations from Sofia for travelling and it continues the discussions pretty smoothly when interrupted or when I interject something randomly (and knows to respond to that also). It still is obvious that it's an AI, but the experience is pretty good.
Ojin
@mad94 This made my day, thank you. Sofia handling travel recommendations and then picking up a random interjection is the exact thing we spent the longest on, so hearing it hold up in the wild is worth more to me than any number on the board today.
You are right that it still reads as AI. If you remember what gave it away, the voice, the face, or the timing between them, I would love to know.
Congrats on the launch! I use LemonSlice right now but am always open to new providers. Do you have a differentiator?
Ojin
Happy to have contributed to this as a product engineer. It's been a wild few months of building, and it's great to finally see people trying what we made. Huge respect to the team, genuinely talented and great people to build with. Let's go! 🚀
Ojin
@aaron_jablonski The turn-taking work was the hard part and most of it was yours, Thanks Aaron.
Congrats on the launch! Had a lot of fun with the API portal so far.
So happy to see Ojin out today after all the hard work! Proud to work with this team 🔥
Ojin
@seema_chhokar Thank you Seema, long road to today.
I'm very excited to finally have Ojin released!! A lot of engineering efforts went into making this product possible and I'm so proud of the team. Can't wait for people to try it out!