RealTalk: Speech Synthesis Model Recreates a Human Voice Perfectly
RealTalk: This Speech Synthesis Model Our Engineers Built Recreates a Human Voice Perfectly 15
…and it’s the voice of Joe Rogan. This includes the breathes, ‘um’s and ‘ah’s, and all other noises.The replica of Rogan’s voice the team created was produced using a text-to-speech deep learning system they developed called RealTalk, which generates life-like speech using only text inputs. But in the next few years (or even sooner), we’ll see the technology advance to the point where only a few seconds of audio are needed to create a life-like replica of anyone’s voice on the planet.
Source: medium.com