Cartesia
Ultra-low-latency, real-time voice AI and text-to-speech built on state space models
Lifelike conversational voice AI companions and ambient intelligence, from the team behind viral voices Maya and Miles.
One of the most convincing conversational voice demos to date, but Sesame is still an early-stage preview and research product rather than a finished, broadly supported app. Treat access as limited and evolving.
Sesame is a voice AI company behind the viral Maya and Miles conversational agents, known for unusually natural, expressive and interruptible speech. Users can try the voices free via a browser research preview and a mobile preview app, while developers can self-host the open-source CSM-1B model. Sesame is also building intelligent eyewear for hands-free voice agents, targeted for 2027.
Sesame AI is a Bay Area company creating personal voice agents designed to feel natural rather than robotic. Its research preview voices, Maya and Miles, went viral in early 2025 for speech that includes breaths, pauses, disfluencies and the ability to be interrupted mid-sentence, powered by the company's Conversational Speech Model (CSM). Beyond the browser demo, Sesame is rolling out a mobile preview app and is developing lightweight intelligent eyewear that lets people talk to their agent hands-free, targeted for 2027. Founded by Oculus co-founder Brendan Iribe, the company frames its mission as bringing ambient, voice-first intelligence into everyday life.
Sesame is a voice AI company behind Maya and Miles, conversational agents that went viral for their unusually natural, interruptible speech. Users can try the voices free through a browser research preview and a mobile preview app on iOS and Android. Developers can self-host the open-source CSM-1B speech model, and Sesame is building intelligent eyewear targeted for 2027. There is no publicly announced consumer pricing, and the product remains an early preview.
Sesame AI Inc. is a Bay Area company founded by Brendan Iribe, co-founder and former CTO of Oculus, alongside co-founders including Ankit Kumar and Ryan Brown. It describes itself as an interdisciplinary team of artists, makers and engineers focused on bringing ambient, voice-first intelligence into daily life.
The company gained wide attention in February 2025 when its demo voices, Maya and Miles, went viral for sounding strikingly human. Sesame has since expanded from a browser demo into a mobile preview and is developing intelligent eyewear as a hardware platform for its agents.
Sesame's core technology is its Conversational Speech Model (CSM), which produces speech with natural intonation, breaths, disfluencies and emotional expression, and supports real-time, interruptible dialogue. The consumer-facing experience centers on personal voice agents, notably Maya and Miles, accessible through a browser research preview and a mobile preview app.
For developers, Sesame open-sourced its base model CSM-1B under the Apache 2.0 license, allowing self-hosted speech generation. Looking ahead, the company is building lightweight intelligent eyewear with quality audio so users can converse with their agent hands-free, planned for 2027.
Sesame targets curious early adopters who want the most natural-sounding, voice-first AI experience, along with AI and voice researchers and developers drawn to its open-source model. Its future eyewear extends the target market to mainstream consumers seeking ambient, hands-free assistance, though today the audience is primarily preview users and technologists rather than enterprise buyers.
Tech-curious individuals and early adopters who enjoy natural spoken conversation with AI, want to think out loud, or seek an ambient voice companion for everyday moments.
Currently there is no paid consumer plan; end users simply access the free preview. Future buyers would be consumers purchasing Sesame's intelligent eyewear once available.
AI enthusiasts, voice-tech commentators, developers experimenting with the open-source CSM model, and media covering conversational AI and wearables.
An early-adopter consumer or developer who values best-in-class, natural voice interaction, is comfortable using preview-stage products, and is interested in voice-first, hands-free AI rather than a stable, enterprise-grade commercial platform.
Sesame raised a Series A of approximately $47.5 million led by Andreessen Horowitz (a16z) in early 2025, followed by a reported $250 million Series B in late 2025 led by Sequoia Capital and Spark Capital. Investors and figures may be updated as official disclosures change.
Sesame is a conversational voice AI company building lifelike personal agents. Maya and Miles are its two demo voices, one warmer and female-sounding, one male-sounding, that went viral in early 2025 for speech that feels remarkably natural, complete with breaths, pauses and the ability to be interrupted mid-sentence.
Yes. You can talk to Maya and Miles for free in your browser through the Research Preview at app.sesame.com, and download the Sesame mobile preview app on iOS and Android. It is still an early preview, so features and availability may be limited and change over time.
Partly. Sesame open-sourced its base Conversational Speech Model, CSM-1B, under the Apache 2.0 license, so developers can download and self-host it from GitHub or Hugging Face. The full Maya and Miles product experience, however, is not open source.
Both offer real-time, interruptible spoken conversation, but Sesame is focused specifically on voice presence and personality, aiming for maximally natural, emotionally expressive speech. ChatGPT is a broader general-purpose assistant, while Sesame is a specialized voice AI company also building dedicated hardware.
Yes. Sesame is developing lightweight intelligent eyewear with high-quality audio so users can talk to their agent hands-free throughout the day. The company has said this eyewear is coming in 2027.
Side-by-side pages for pricing, features, and best-fit use cases.
Ultra-low-latency, real-time voice AI and text-to-speech built on state space models
Emotionally intelligent voice AI with an empathic interface and expression measurement, built for developers.
Developer platform for building, testing, and deploying real-time AI voice agents
Build, test, and deploy production-grade AI voice agents that handle inbound and outbound phone calls.