Google has released two new speech-to-speech models, Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking, which are similar in shape to OpenAI's GPT-Live models.

  • The release includes Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking.
  • A web UI allows users to select a model and voice preset, enter an optional system prompt, and start a voice conversation through the browser.
  • Users can interrupt the model while it is talking, with transcripts capturing these interruptions.
  • The implementation connects to a WebSocket endpoint and uses the Web Audio API for capture and playback without external libraries.