The next interface between humans and machines.
Every leap in computing was accommodated by a new interface: computers with the keyboard and mouse, the internet with the browser, mobile with the touchscreen. AI is the biggest leap yet, and it is still stuck behind those old interfaces. We are building what comes next.
Talk to AI without making a sound.
Hundreds of millions of people already use voice to interact with AI every week. It is faster and more natural than typing. But you cannot speak out loud for most of your day, not in an office, a meeting, or a train. So the most natural way to use AI goes quiet exactly where you are.
Silent Speech fixes that. You mouth the words silently, and our model reads how your face moves and turns it into text and commands for any AI. Voice when you can speak. Silent when you cannot. On the Mac and phone you already own.
No audio, only how your face moves. The best voice models get full audio and hit ~3% word error rate. We use pixels alone and are already under 10%, our best is 4%.
Join Silent Speech.
We will invite people gradually as the research preview expands.
We're building the interface for the next era of human-machine life.
Intelligence is no longer the hard part. Computers can already think. The question that defines the next decade is how we live with that intelligence. We believe the answer is presence: AI that dissolves into your world instead of sitting behind a screen, that is with you, aware of your life, acting on your behalf, reachable as naturally as a thought. The interface, not the model, is what turns intelligence into something that changes how you live.
That interface is what interfaces.inc exists to build. Silent Speech is our first step.
Bring Silent Speech to the moments voice cannot reach.
Building something voice-first? Silent Speech works everywhere your users cannot talk. License our model, API and SDK.
Request accessVoice is the future. But what happens when you cannot use it?
Voice is becoming the most natural way to interact with AI. The world is already reorganizing around it, including the places where speaking out loud makes no sense.
I want to interview someone who actually wears one of these soundproof beaks to talk to AI in the office.
voice-to-text has gone too far
agreed feels big, i want a new kind of computer
I don’t think you understand what ChatGPT Voice unlocks... Work can be done ANYWHERE now.
Voice is a very big deal.
Voice AI is absolutely going to be game changing. Typing is fine, but when you just talk to your device and get an instant answer back, it lowers the barrier to asking questions about anything. It’s the closest thing to sci-fi we’ve had in AI. 0 latency is the key though.
Voice AI will be fairly game changing for interacting with AI Agents to take actions for you on mobile. Texting agents will often be less convenient than simply using a GUI. But talking to agents can be far more efficient than GUIs. This opens up entirely new software use cases.
Voice is one of the most frequent and information-rich form of human communication — and for the first time, AI is making it programmable at scale.
Ever since I started working with @natfriedman and @danielgross in 2022, we had a core belief that voice would be a massive unlock for the future of HCI. We were on the lookout for the team that would cross the uncanny valley and build a generational company in voice AI.
Excited to announce that @simpleailab has raised a $14M seed round led by @firstharmonic. For the past year, we’ve been building AI voice agents to transform direct-to-consumer sales. We fundamentally believe that voice AI is the future of all B2C calls.
voice is the future
go try this out! voice is the future of how we interact with superintelligence, and gpt-live-1 is a step function improvement in voice intelligence, latency, and safety. twas a blast working with the stacked team on this one
absolutely fucking love love this. okay sorry openai, you are doing amazingly cool stuff when it comes to interfaces/modalities. i love voice & it is my primary input method for computers now (therefore it deserves a dedicated mechanic to activate).
Voice won’t be the next interface. It will make the interface disappear. AI that understands accents, noise, jargon and fragments removes the need to translate work into fields and forms. Today: transcription. Next: organizational memory.
“The keyboard and mouse are slowly dying.” Sector Head at Coatue Max Cook: “The keyboard and mouse, you should throw them away. We’re going to natural language interacting with agentic models.”

