Google reports its technologies now support over 300 languages for more than 7 billion people, introducing new tools to improve speech understanding, data collection for underrepresented languages, and accessibility.

  • Gemini 3.5 Live Translate enables real-time spoken translation across 70 languages and 2,000+ pairs, capturing code-switching and emotional cues.
  • The Universal Speech Model, trained on 12 million hours of audio, uses cross-lingual transfer learning to support the goal of covering the world’s 1,000 most-spoken languages.
  • Open-data partnerships including WAXAL, Project Vaani, and the Amplify Initiative have collected extensive speech data for Sub-Saharan African and Indian languages.
  • TranslateGemma provides lightweight open translation models that run efficiently on-device without internet connectivity.
  • Sign Language-to-Text (SL2T) powers sign-to-text dictation in Gboard and Live Transcribe for over 50 sign languages, starting with American Sign Language.

These innovations aim to bridge the digital divide by making language technology accessible offline, supporting low-resource dialects, and ensuring cultural nuance is preserved in AI systems.