impossible → routine
Speech built one sample at a time
Generating human-sounding speech directly as a raw audio waveform.
Impossible
DeepMind publishes WaveNet and calls building up samples one step at a time computationally expensive.
source1 year
Routine
The production version generates the Google Assistant voices for US English and Japanese on all platforms.
50 ms of compute per second of speech, 1,000x faster than the research model
source