6 ms·
>"Most models are trained on the original speaker's voice, but maybe only a little bit." Really cool that you got this to work. I used to work on TTS (a few y
by tmoney1818 6y ago
>"Most models are trained on the original speaker's voice, but maybe only a little bit."
Really cool that you got this to work. I used to work on TTS (a few years ago, now), and we trained on celebrity voices, but used full audiobooks. https://github.com/Kyubyong/tacotron https://github.com/Kyubyong/tacotron
Here are some of our Nick Offerman samples: https://soundcloud.com/kyubyong-park/sets/tacotron_nick_215k https://soundcloud.com/kyubyong-park/sets/tacotron_nick_215k .
- echelon 6y agoHey! I've seen your results! Really fantastic work! Thanks for making this so open and accessible.