Home / Founders / Mati Staniszewski
Mati Staniszewski
Co-founder and CEO
ElevenLabs
Mati Staniszewski is co-founder and CEO of ElevenLabs, the voice AI company known for expressive text-to-speech, dubbing, and conversational agents. He met co-founder Piotr Dabkowski in high school in Warsaw, hacked on side projects while working at Palantir and Google, and turned a shared annoyance with monotonous Polish movie dubs into a research-and-product company focused only on audio.
Warsaw high school and hack weekends
They met in an IB class in Warsaw about fifteen years ago, bonded over mathematics, and stayed friends through living, studying, and traveling together. While Piotr was at Google and Mati at Palantir they ran weekend hack projects — recommenders, crypto risk tools, then an audio coach that analyzed how you speak. On Sequoia's Training Data podcast he traced the founding spark to late 2021: Piotr tried to watch a film with a girlfriend who did not speak English, switched to Polish, and hit the childhood experience of every character voiced by one flat narrator. "It's a horrible experience," Mati said. "We think this will change."
Staying focused on audio
ElevenLabs launched into a world that expected foundation-model labs to crush vertical voice startups. Staniszewski's answer on Training Data was focus: "staying focused and staying focused in our case on audio." Little research attention had gone to speech compared with text and images; applying transformers and diffusion to audio, plus building product around the models, is how they competed. Data, he argued, is the hard part — not only transcripts of what was said, but labels for how it was said, emotion, and nonverbals.
Voice as the fundamental interface
On the same podcast he stated the long bet plainly: "voice will fundamentally be the interface for interacting with technology." It carries emotion, intonation, and imperfection that text does not. That thesis pushed ElevenLabs from narration and dubbing into agents for healthcare outreach, customer support, education, and consumer experiments — always with quality, latency, and reliability as the three things customers buy.
On the voice Turing test
He has set an ambitious bar: prove human-level conversational voice is possible soon, whether through cascading speech-to-text / LLM / text-to-speech stacks or true duplex speech-to-speech models. Prosumer releases come first so unexpected use cases surface before enterprise packaging. The through-line is specialized audio research wrapped in products people can ship.
Index score breakdown
Overall 77 · Rank 5 on the AI Founders to Watch Index
| Factor | Score |
|---|---|
| Innovation | 9 |
| Impact | 9 |
| Company success | 8 |
| Vision clarity | 8 |
| Credibility | 8 |
| Momentum | 7 |
| Independence of signal | 2 |
