The story
It started with a question.
The question
Could speech itself adapt to the person listening, live, and still sound like the one speaking?
What we saw
We came from speech research at Google, where the work was teaching machines to hear. The person listening was still doing the hard part: the learner replaying the podcast, the audience at the session nobody could afford to interpret, the patient nodding along to the doctor. The speaker, however willing, cannot help each of them at once.
What we set out to build
Speech that meets every listener: interpreted into their language, said again at their level, kept in the speaker's own voice, and run wherever the audio is allowed to be. We started with the live case, because it is the hardest, and we put the results on this site as they happen.