
Members of the team work on research and develop audio representations that best captures both the semantic and acoustic information within the audio signal to enable Gemini models to understand audio inputs and also generates natural audio sounds. This builds the foundation for audio-to-audio dialog models, text-to-speech, speech-to-speech translation, duplex audio dialogs.
Artificial intelligence will be one of humanity’s most transformative inventions. At Google DeepMind, we are a pioneering AI lab with exceptional interdisciplinary teams focused on advancing AI development to solve complex global challenges and accelerate high-quality product innovation for billions of users. We use our technologies for widespread public benefit and scientific discovery, ensuring safety and ethics are always our highest priority.

A problem isn't truly solved until it's solved for all. Googlers build products that help create opportunities for everyone, whether down the street or across the globe. Bring your insight, imagination and a healthy disregard for the impossible. Bring everything that makes you unique. Together, we can build for everyone.
Check out our career opportunities at goo.gle/3DLEokh