Speech recognition that adapts to your language in seconds
Built for production
When accuracy, reliability and privacy are non-negotiable.
Accurate on domain‑specific language
Medication names, order numbers, addresses, or case references - all transcribed accurately. Adapts from a text upload or in real-time, no audio samples or fine tuning.
Native-grade transcription across European languages
English, Dutch, French, German, Italian, Polish, (Brazilian) Portuguese, Spanish, Swedish, Frisian.
An endpoint for any audio
Healthcare voice engine
No tradeoffs between
accuracy and speed
Generic speech engines fail on domain‑specific language.
Reson8 gets the words right that others get wrong.


Reson8 adapts to your business context in real-time or with a text upload. No audio files or finetuning required.
Accurate transcription of order codes, reference numbers, licence plates, invoice numbers, flight numbers, and more.
Smaller, faster models give you infrastructure which works in the real world.
PRICING
From pilots to production deployments
Free
- 600 credits / month (10 h prerecorded)
- 1 concurrent connection
- Basic support
Pro
- 6,000 credits / month (100 h prerecorded)
- 20 concurrent connections
- Standard support
- €0.006 per credit overage
Growth
- 30,000 credits / month (500 h prerecorded)
- 100 concurrent connections
- Priority support
- €0.004 per credit overage
Enterprise
- Custom credit volume & pricing
- Custom concurrency
- Dedicated support
- Custom overage rates
- On-prem / VPC deployment
Frequently asked questions
Reson8 is built for European languages and gets the words right that matter most. Generic speech models mishear medication names, order numbers, addresses and case references. Reson8’s speech-to-text adapts to your domain-specific terminology in real time or from a single text upload, no fine-tuning or audio samples needed.
We have full support for 10 languages: English, German, Spanish, French, Dutch, Italian, Polish, Portuguese, Swedish, Frisian.
Reson8 routinely leads speech-to-text benchmarks from third-parties such as Hugging Face, Artificial Analysis, Coval, and Voice Arena.
Yes, a single stream can switch dynamically between languages mid-conversation. Language detection is automatic and out-of-the-box, no manual language flagging necessary. Find more information on language detection here.
Yes, diarization is supported across all endpoints, ensuring transcripts accurately distinguish between speakers no matter the complexity of the audio. Find more information on diarization here.
Each endpoint is applicable for a different use case:
- Prerecorded: for complete audio files, uploaded over REST. You get the full transcript back in a single response.
- Realtime: for live audio streamed over WebSocket, when you want to show live transcripts while speaking.
- Turns: also a WebSocket stream, but it detects when a speaker starts and stops talking, returning the transcript of each turn separately. Use this endpoint for cases like voice agents, when you need low latency end-of-turn detection and native barge-in.
Check out our guide on choosing the right endpoint here.
No, we don't and never will. Audio is processed on our own GPU infrastructure, which is fully built and operated by Reson8. Find more information on our trust center here.
Reson8 retains zero data. All audio is streamed pass-through and never stored; this also applies to files and transcripts.
Yes, Reson8 is fully ISO 27001 certified and GDPR compliant. We own and operate our full GPU infrastructure ourselves in the Netherlands, Reson8 is not reliant on any third party subprocessors of audio, making Reson8 uniquely suitable to regulated industries like healthcare, legal, finance, and more. Find our DPA here.
Sign up today! Get an API key through console.reson8.dev, talk to our team or check our documentation to get started!



