Google has released two new speech-to-speech models, Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking, which are similar in shape to OpenAI's GPT-Live models.
- The release includes Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking.
- A web UI allows users to select a model and voice preset, enter an optional system prompt, and start a voice conversation through the browser.
- Users can interrupt the model while it is talking, with transcripts capturing these interruptions.
- The implementation connects to a WebSocket endpoint and uses the Web Audio API for capture and playback without external libraries.