Gemini’s new voice model can switch between 97 languages mid-conversation
Google announced Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking on September 15. Both support 97 languages and detect and switch language on their own mid-sentence, and the model processes video in near real time. Instead of going quiet while it fetches data, the Extended Thinking version narrates progress with verbal cues like “let me check that”. The models are rolling out across Search Live, Workspace, AI Studio and developer APIs.
Sources
- Google, “Introducing Gemini 3.8 Live and 3.8 Live Extended Thinking”, (blog.google)
About this story
This story was posted on Instagram by @jarrus.tech on Sept. 22, 2026.
Spotted an error in this story? [email protected] · Instagram
Short link: thejarrus.com/en/gemini-live-languages
This story in Turkish: Gemini’ın yeni ses modeli sohbetin ortasında 97 dil arasında geçiş yapabiliyor
On the same topic
OpenAI’s image model posts a near-perfect score on text-dense images
On UltraText Bench, a dense-text test led by Westlake University, OpenAI’s GPT Image 2 ranked first of 24 configurations with 99.35 out of 100.
AWS adds Chinese lab Z.ai’s GLM 5.3 model to Amazon Bedrock
AWS added GLM 5.3, a 753-billion-parameter model from Beijing-based Z.ai (formerly Zhipu AI), to Amazon Bedrock for eligible enterprise customers on October 5.
Microsoft releases three new voice models, one transcribing 60 languages in real time
On October 1, Microsoft AI announced MAI-Transcribe-2-Streaming, which transcribes 60 languages in real time, and the 23-language MAI-Voice-2.1 and Flash.