Sesame, headed by Oculus co-founder Brendan Iribe, exits stealth with an undisclosed amount of funding, unveiling an AI speech model and plans for AI glasses
Here's an embedded 13-minute audio clip of my first attempt! — www.theverge.com/news/621022/ ... X: @sesame : At Sesame, we believe in a future where computers are lifelike. Today we are unveiling an early glimpse of our expressive voice technology, highlighting our focus on lifelike interactions and our vision for all-day wearable voice companions. https://www.sesame.com/... [video] Tobi Lutke / @tobi : Man, sesame's voice model is absolutely insane. You have to try this demo. GG @brendaniribe https://www.sesame.com/... Fabrizio Rinaldi / @linuz90 : Ok this is insane, and I'd even say this could pass the Turing test for me, and surely for a lot of people. The voice quirks especially really sell it. I might start telling my parents to never trust they're speaking to a human on the phone anymore. https://www.sesame.com/... Zach Tratar / @zachtratar : The new Sesame voice model feels like a ChatGPT moment for voice. It's just that good. Nikunj Kothari / @nikunj : This is the most “her” moment I've felt in the last few years.. It's truly uncanny - recommend trying this out now https://www.sesame.com/... Josh Dance / @joshdance : Sesame's voice model is insane. Full stop. Try it out. https://www.sesame.com/... The most realistic AI voice I have heard yet. Tobi Lutke / @tobi : These AIs are definitely better conversationalists than I am. @kimmonismus : Holy SH*t! This sounds absolutely human and natural. We crossed the threshold where it's indistinguishable if something is AI generated or human. And much earlier than expected. Sesame makes AI sounds real like never before. I am shocked. [video] @zachweinberg : Holy shit moments in tech are rare. This is one of them. Wow. Blown away. @a16z : We're thrilled to announce our investment in Sesame AI, a company redefining computing with audio-first AI glasses. For decades, computing has evolved, yet we're still stuck on screens. While AR glasses have focused on visuals, AI-powered speech interfaces have remained [image] Guillermo Rauch / @rauchg : Absolutely astonishing voice ai demo. The whole site experience is 🔥 https://www.sesame.com/... LinkedIn: Jenny He : When we first invested in Sesame two years ago, I knew Brendan and Ankit were building something extraordinary. … Forums: Hacker News : Crossing the uncanny valley of conversational voice r/ChatGPT : Sesame AI's Voice is INSANE - OpenAI needs to catch up! r/singularity : Crossing the uncanny valley of conversational voice r/artificial : Sesame's new text to voice model is insane. Inflections, quirks, pauses r/LocalLLaMA : “Crossing the uncanny valley of conversational voice” post by Sesame - realtime conversation audio model rivalling OpenAI
Context & Ripple Effects
Sesame enters the market with backing from Andreessen Horowitz and a founder whose Oculus history gives its wearable-computing ambition added relevance. Its initial product emphasis is expressive speech rather than a general-purpose assistant alone.
The launch quickly became a test of whether realistic voice can create stronger user attachment: the early speech demo drew unusually strong reactions, while OpenAI expanded its own speech API later that month. Subsequent coverage shows Sesame moving from a demo toward an agent, mobile beta, and smart-glasses program.
First-order effects
- Sesame can use its speech model as the entry point for an all-day voice companion, tying its near-term product identity to natural, conversational audio and eventual AI glasses.
- The public demo gives Sesame a mechanism to recruit users, partners, and investors around a tangible capability; later reports of fundraising discussions above $200 million indicate that the launch strengthened that positioning.
Second-order effects
- Voice-model providers and assistant makers face a higher qualitative bar: realism and conversational presence become product differentiators alongside transcription accuracy and latency.
- Building glasses around a voice companion makes hardware distribution central to Sesame's strategy, increasing the importance of integrating the model, agent behavior, and wearable interface rather than selling speech as a standalone feature.
Third-order effects
- If voice companions prove useful over extended daily use, AI interaction could shift from discrete text prompts toward ambient, hands-free interfaces—a market where control of the device and user relationship matters as much as the underlying model.
- The pattern may also intensify scrutiny of synthetic voices and user trust: more lifelike systems increase the value of consent, disclosure, and safeguards, even as providers compete on naturalness.
The trend: Sesame is one data point in the shift from voice as an AI feature to voice as the primary interface for ambient, wearable agents.