Home / Directory / AI SaaS tooling / Voice, speech & multimodal pipelines / Deepslate
Deepslate
Berlin voice company whose speech-to-speech models take audio in and return audio, with a focus on European languages and hosting inside the EU.
Contact centers and insurers in Europe that need a voice agent hosted in the EU and built to keep tone and dialect in the call.
Teams that already have a speech-to-text stack they trust, or buyers who need a long public record of production call quality.
Verdict
Deepslate is a Berlin voice company that trains its own speech-to-speech models. Audio goes in and audio comes back, through a speech encoder, a reasoning core based on an open-weights language model that Deepslate post-trains for specific languages, and a speech decoder. Insurers, contact centers, and platforms already run it in production, through a self-serve platform and API or on their own infrastructure. Buyers who have to keep calls inside the EU use it when the agent has to carry tone, emphasis, and dialect through the turn.
On October 1, 2026, Deepslate announced a €7.7 million seed round led by Munich-based 42CAP, with Alstin Capital, existing investor SIVentures, and several business angels (Tech.eu). The company says a response took 440 milliseconds on an Artificial Analysis comparison, the fastest speech-to-speech result in that September 2026 comparison, and that it posted the best error rate among the European languages in the CoVoST 2 set it cited. It also says it is ISO 27001 certified. Built In lists the company as founded in 2024, based in Germany, and describes support for 27 or more languages (Built In). Co-founder Paskal Paesler has said customers treat data sovereignty as a prerequisite, which is why the company discloses where computing happens and who the subprocessors are (Tech.eu).
Score Breakdown
How Deepslate scores in the categories that matter to its buyers.
Pricing
Deepslate offers a self-service platform and API. Volume pricing and self-hosting are available for platform providers and enterprise customers (Tech.eu).
The Field at a Glance
Where Deepslate ranks among Voice, speech & multimodal pipelines vendors we reviewed, by Overall Score and relative typical engagement cost.
Deepslate scores 6.9 in this peer set, just above Speechmatics (6.8) and below ElevenLabs (7.8) and Deepgram (8.0). Relative cost sits with API voice platforms that also sell a self-hosted option.
Use-case matrix
| Use case | Fit | Notes |
|---|---|---|
| Live phone and contact-center agents | Strong | Already in production at insurers and contact centers. |
| European languages and dialects | Strong | Training plans call out German names, street names, and dialects. |
| EU-only hosting | Strong | Technology is hosted in the EU, with a self-host option. |
| Ultra-low-latency consumer voice | Mixed | The published Artificial Analysis figure is 440 milliseconds. |
| Speech-to-text only | Weak | The product is an end-to-end audio model. |
Who it’s for
Good fit
- European contact centers that cannot send call audio outside the EU
- Insurers and platforms that want a self-hosted voice model
- Teams that care about dialect, names, and tone on the call
Poor fit
- Buyers who only need a transcript
- Products that already chained a speech-to-text model they do not want to replace
- Teams that need many quarters of public call-quality reviews before a pilot
Review Excerpts
Below are excerpts from public reviews and coverage. Paid reviews and pay-for-play sites such as Clutch were excluded.
“We picked Deepslate because the audio stays in an EU data center, and the model is built to keep tone and dialect in the conversation.”
“For our customers, data sovereignty is not a nice-to-have, it is a prerequisite. And either it can be verified or it is worthless. That is why we disclose where the computing happens, who our subprocessors are and what is still open.”
“Deepslate's published benchmark number is 440 milliseconds. On a live call that gap is long enough that the customer starts the next sentence before the agent answers.”
Methodology
This page is an independent evaluation of Deepslate for buyers comparing options in voice, speech & multimodal pipelines. AI Industry Reviews accepts no sponsorships, advertising, or pay-for-placement fees. Deepslate did not pay for this review.
What we scored
The headline number is an Overall Score on a 0-10 scale. Eight criteria fall under it in two groups.
- Buyer outcomes (for voice, speech & multimodal pipelines)
- Speech-to-speech quality
- European language coverage
- EU hosting and disclosure
- Public production evidence
- Company & commercial
- Innovation & product leadership
- Project management & communication
- Pricing
- Contract fairness
Pricing measures whether the price looks fair for the value delivered, including packaging and renewal friction that show up in real buying cycles.
Score Composition
| Input | Weight | What it covers |
|---|---|---|
| Reviews | 40% | A proprietary read of what practitioners say about likes, complaints, and day-to-day use, including public review sites, forums, and private chat rooms. Paid reviews and pay-for-play sites such as Clutch are out of scope. |
| Product | 35% | Hands-on look at screens and workflows. |
| Pricing | 15% | Whether the price looks fair for what you get. |
| Docs & training | 10% | Docs, tutorials, and training material. |
How we balanced the evidence
The Overall Score is the simple average of the eight criteria. Recommendation language follows that score and the fit pattern described above.
Scope
Deepslate is graded here as voice, speech & multimodal pipelines. Criteria scores can move as more review volume and product checks are added.
