Search
AI News ARCHIVE

Robots Can Now Talk like Humans with Thanks to Google’s DeepMind Project

Google’s DeepMind researchers have been at it again creating technology that will help change the world as we know it and this time it comes in the form of WaveNet.

Daniel Okafor
Daniel OkaforSenior AI Reporter
2 min read
Robots Can Now Talk like Humans with Thanks to Google’s DeepMind Project

Google’s DeepMind researchers have been at it again creating technology that will help change the world as we know it and this time it comes in the form of WaveNet. WaveNet is a convolutional neural network that is even closer to mimicking both US English and Mandarin Chinese. The network can also switch between different voices and create unique musical fragments.


This type of text-to-speech system (TTS) is different to those currently in use today as is trained on raw audio waveform from multiple speakers.  It then uses the network to generate synthetic utterances and sends the sample back to the network to generate the next sample.  One of DeepMind’s researcher’s comments, “As well as yielding more natural-sounding speech, using raw waveforms means that WaveNet can model any kind of audio, including music.”

WaveNet also has the ability to learn different characteristics of various voices (female and male) including breathing and mouth gestures. The researcher’s state, “To make sure [WaveNet] knew which voice to use for any given utterance; we conditioned the network on the identity of the speaker.  Interestingly, we found that training on many speakers made it better at modeling a single speaker than training on that speaker alone, suggesting a form of transfer learning.”


One area that WaveNet mat struggle with is it will be limited to Google products, at least for the time being as the approach requires so much data and computing power.  The processing power needed is said to be around 16,000 samples per second to create realistic speech sounds.  But, WaveNet can still be used to remodel music audio and speech recognition and is something we will see much more of soon.

Related Links;


More News To Read

  • Meet AliceX – Your Very Own Virtual Reality Girlfriend
  • New Disney Robot Uses Air-Water Actuators to Move Around Smoothly
  • Deep Web vs. Dark Web – Is There A Difference?
  • Futuristic Propulsion That Can Bring Us to Mars In a Few Weeks Becoming a…
  • This Hybrid Electric Sport Car to Give Tesla a Run for Its Money

Related Coverage

Google · Preferred Sources

Don't miss new tech stories on Google

Add TrendinTech once in the Google app and our stories appear in your news suggestions.

Add Now
Daniel Okafor

Daniel Okafor

Senior AI Reporter

Daniel Okafor is the Senior AI Reporter at TrendinTech, where he covers large language models, machine learning research and the practical use of artificial intelligence across business and government. He previously reported on artificial intelligence for MIT Technology Review, covering the labs behind the current generation of frontier models and the policy debates in Washington and Brussels. Daniel holds a Master of Science in Machine Learning from Carnegie Mellon University and follows the research community closely, attending NeurIPS and ICML each year to speak with the people behind the papers. He has a particular interest in evaluation: how models are benchmarked, where those benchmarks fail and what that means for the companies betting on them.

All stories by Daniel Okafor (316)