
016: Introducing Pheme, the speech generation model built to scale.
About this episode
In this insightful episode, host Kylie Whitehead converses with Dr. Ivan Vulić, a Senior Scientist at PolyAI and a Principal Research Associate at the University of Cambridge. They discuss the development and advantages of 'Pheme', a new, more efficient model for voice generation developed by PolyAI. Unlike existing Text-To-Speech (TTS) models, Pheme is designed to generate more conversational and natural sounding speech, which can be tailored to the unique needs of different businesses and used for brand voices. They also touch on the balance between performance and quality in building conversational systems, and the ethical considerations surrounding voice synthesis.
Get every episode summarized
Each time Deep Learning with PolyAI publishes, we email you a written briefing from the transcript — the topics, who appeared, and any specific claims, with the ad reads skipped.
Email me new episodesFree for 3 shows. No card needed.
Hosts & guests
No transcript yet
This episode has not been transcribed. Request it and it moves to the front of the queue.
More episodes
More from Deep Learning with PolyAI

What happens if your AI agents get 1% better every day?
Deep Learning with PolyAI

Who's coordinating your army of AI agents?
Deep Learning with PolyAI

Why should CX leaders care about MCP?
Deep Learning with PolyAI

Can AI really hear a call the way a person does?
Deep Learning with PolyAI