impossible → routine

Speech built one sample at a time

Generating human-sounding speech directly as a raw audio waveform.

Impossible

DeepMind publishes WaveNet and calls building up samples one step at a time computationally expensive.

source

1 year

Routine

The production version generates the Google Assistant voices for US English and Japanese on all platforms.

50 ms of compute per second of speech, 1,000x faster than the research model

source

All dated pairs