Speech-to-speech synthesis with Alex Serdiuk, CEO, Respeecher

Picture of Kane Simms
Kane Simms

More ways to listen and watch

About this episode

Emmy-Award winning Respeecher join us to share the future of synthesised speech.

Supporting Ukraine

The VoiceLunch Foundation is taking donations to help support the voice lunch and voice technology community in Ukraine. VUX World has, of course, donated. I plead with you to donate too.

Donate here

Presented by Deepgram and Symbl.ai

Presented by DeepgramDeepgram is a Speech Company whose goal is to have every voice heard and understood. We have revolutionized speech-to-text (STT) with an End-to-End Deep Learning platform. This AI architectural advantage means you don’t have to compromise on speed, accuracy, scalability, or cost to build the next big idea in voice. Our easy-to-use SDKs and APIs allow developers to quickly test and embed our STT solution into their voice products. For more information, visit:

Presented by Symbl.aiSee how easy it is to add simple but powerful call coaching and call tracking functionality to your customer experience solutions with Symbl.ai’s customizable Conversation Intelligence APIs. From calls to videos to text conversations — apply best in class contextual AI in no time by getting started for free.

Speech-to-speech

Emmy Award-winning Respeecher is changing the speech synthesis game. Move over TTS and SSML, and enter Speech to Speech.

From voice preservation, to accessibility, to voiceovers to film studios, the uses for speech to speech are endless.

Rather than programming machines to read text out loud, like most speech synthesis systems (text-to-speech), Respeecher uses it deep learning to sonically reproduce voices at film studio levels of fidelity. With Respeecher, a voice actor (or anyone) can simply speak and have their voice transformed into synthesised speech in next to real time. It reproduces all of the character and delivery in the voice, so that the resulting synthesised speech is exactly like the original source. Whether you shout, whisper or sing, the speech to speech technology will replicate everything. It’s truly ground breaking and has to be hear to be believed.

And you can hear it in the intro, as I introduce the episode using four different Respeecher voices.

Respeecher CEO, Alex Serdiuk, joins us to share more.

Timestamps

00:00 Intro and presenting Deepgram and Symbl.ai
04:05 Welcome Alex and closing the Ukraine airspace
08:40 About Respeecher
12:40 Sourcing voices
14:10 Nixon and winning an Emmy
17:37 Process of creating speech to speech
25:34 Limitations of TTS for long for audio
29:00 The future of voice acting
34:25 Voice marketplace
38:00 Pricing of speech to speech voices
42:00 How to achieve higher quality voices
43:30 Accessibility
48:50 Endless use cases
51:00 Ethics
55:54 Outro and more information

Links

Learn more at https://respeecher.com

Share

Weekly newsletter

The latest in AI-powered customer experience. Make better strategic decisions with the help of our weekly newsletter.

Host

Kane Simms

A strategic AI advisor who, for the past decade, has helped business leaders and product owners transform customer experience using conversational and generative AI.

Share

✓   Link copied

Weekly newsletter

The latest in AI-powered customer experience. Make better strategic decisions with the help of our weekly newsletter.

Top articles and podcasts

Related content

The best AI use cases are right in front of your eyes

Loading the Elevenlabs Text to Speech AudioNative Player… CCW published a report recently that asked contact centre leaders how much impact AI has had on

How to build reliable AI agents: Test Driven Development is back!

Loading the Elevenlabs Text to Speech AudioNative Player… One of the top questions on the lips of teams building AI agents is ‘when is it

The two things contact centres have been missing until now
Cresta CEO Ping Wu unpacks the two things contact centres have always lacked, and how AI changes that.
Our deep analysis of the Gartner Magic Quadrant for Conversational AI Platforms 2026
Discover where all 14 vendors landed, why voice is only an optional criterion and the question the report doesn't ask.

Meet VUX at CCW Europe, Amsterdam, October 5-7

Register now: Why your contact centre playbook is obsolete