I’m joined by Endre Davids from Seamly to explore what really happens when businesses take an existing chatbot and move it into the voice channel.
We discuss why latency often starts deeper in the conversational platform than teams expect. Endre explains how companies can spend time optimising speech recognition and text-to-speech while the chatbot itself is responsible for much bigger delays.
The conversation gets into why voice needs a different interaction model, because people speak differently from how they type. That puts far more pressure on getting the conversation design right.
We also cover how to manage perceived latency, when it makes sense to reuse existing chatbot infrastructure and how sentiment, multimodality and customer context could shape richer voice experiences.
Endre and I keep landing on the same point. The success of voice AI depends heavily on conversation design, and voice punishes the shortcuts that chat lets you get away with.