Home / Founders / Mati Staniszewski

Mati Staniszewski

Founders to Watch · Fall 2026

Mati Staniszewski

Co-founder and CEO

ElevenLabs

Series C Voice, speech & multimodal
Why watch Staniszewski built ElevenLabs so synthetic voices sound close enough to people that creators use them without embarrassment.

Mati Staniszewski is co-founder and CEO of ElevenLabs, the voice AI company known for expressive text-to-speech, dubbing, and conversational agents. He met co-founder Piotr Dabkowski in high school in Warsaw, hacked on side projects while working at Palantir and Google, and turned a shared annoyance with monotonous Polish movie dubs into a research-and-product company focused only on audio.

Warsaw high school and hack weekends

They met in an IB class in Warsaw about fifteen years ago, bonded over mathematics, and stayed friends through living, studying, and traveling together. While Piotr was at Google and Mati at Palantir they ran weekend hack projects — recommenders, crypto risk tools, then an audio coach that analyzed how you speak. On Sequoia's Training Data podcast he traced the founding spark to late 2021: Piotr tried to watch a film with a girlfriend who did not speak English, switched to Polish, and hit the childhood experience of every character voiced by one flat narrator. "It's a horrible experience," Mati said. "We think this will change."

Staying focused on audio

ElevenLabs launched into a world that expected foundation-model labs to crush vertical voice startups. Staniszewski's answer on Training Data was focus: "staying focused and staying focused in our case on audio." Little research attention had gone to speech compared with text and images; applying transformers and diffusion to audio, plus building product around the models, is how they competed. Data, he argued, is the hard part — not only transcripts of what was said, but labels for how it was said, emotion, and nonverbals.

Voice as the fundamental interface

On the same podcast he stated the long bet plainly: "voice will fundamentally be the interface for interacting with technology." It carries emotion, intonation, and imperfection that text does not. That thesis pushed ElevenLabs from narration and dubbing into agents for healthcare outreach, customer support, education, and consumer experiments — always with quality, latency, and reliability as the three things customers buy.

On the voice Turing test

He has set an ambitious bar: prove human-level conversational voice is possible soon, whether through cascading speech-to-text / LLM / text-to-speech stacks or true duplex speech-to-speech models. Prosumer releases come first so unexpected use cases surface before enterprise packaging. The through-line is specialized audio research wrapped in products people can ship.

Index score breakdown

Overall 77 · Rank 5 on the AI Founders to Watch Index

FactorScore
Innovation9
Impact9
Company success8
Vision clarity8
Credibility8
Momentum7
Independence of signal2

Links

← All founders