hecko
08/02/2021, 9:57 AMmetaphysician
08/02/2021, 9:57 AMhecko
08/02/2021, 9:57 AMhecko
08/02/2021, 9:58 AMhecko
08/02/2021, 9:59 AMhecko
08/02/2021, 10:00 AMhecko
08/02/2021, 10:02 AMApple. please turn that into the training data we have"
can't remember if the training notebook asks for the tacotron model as inputmetaphysician
08/02/2021, 10:04 AMhecko
08/02/2021, 10:09 AMhecko
08/02/2021, 10:09 AMmetaphysician
08/02/2021, 10:11 AMmetaphysician
08/02/2021, 10:12 AMhecko
08/02/2021, 10:13 AMhecko
08/02/2021, 10:14 AMhecko
08/02/2021, 10:14 AMhecko
08/02/2021, 10:14 AMmetaphysician
08/02/2021, 10:17 AMhecko
08/02/2021, 10:17 AMhecko
08/02/2021, 10:18 AMhecko
08/02/2021, 10:25 AMreal time voice cloning that's different, i believe it's actually trained to do speaker recognition (as in "who's speaking that")
and the final output for the recognition is a thousand neurons each corresponding to a speaker, but on the layer before that there's like 100 neurons that together encode the characteristics of a voice
so they ignore the final layer and take the 100 neurons, and then they use them to synthesize voice hecko
08/02/2021, 10:26 AMmetaphysician
08/02/2021, 10:29 AMhecko
08/02/2021, 10:29 AMhecko
08/02/2021, 10:30 AMmetaphysician
08/02/2021, 10:32 AMhecko
08/02/2021, 10:34 AMhecko
08/02/2021, 10:34 AMmetaphysician
08/02/2021, 10:36 AMhecko
08/02/2021, 10:40 AMhecko
08/02/2021, 10:40 AM