< Back to all clusters
[TECHNOLOGY] · United States · 24 sources

started · updated

Google launches Gemini 3. 8 Live AI models for real-time voice interaction

Google has announced the release of its most advanced speech-to-speech AI models, Gemini 3. 8 Live and Gemini 3. 8 Live Extended Thinking. These models are designed to facilitate natural, low-latency voice interactions by allowing the AI to reason and speak simultaneously. While the standard Gemini 3. 8 Live focuses on scalability and cost-efficiency, the Extended Thinking variant is optimized for complex, multi-step reasoning tasks.

Key features include the ability to process visual information in near real-time, automatic detection and switching between 97 supported languages, and the capacity to execute API calls and tool functions in the background without interrupting the conversation. To combat misinformation, Google is implementing SynthID, an invisible digital watermark, on all audio generated by these models.

In addition to the core models, Google is updating Gemini Notebook (formerly NotebookLM) with new learning tools. These updates include real-time voice interaction in nearly 100 languages, an integrated audio recorder for capturing lectures, and the ability to generate interactive learning overviews, such as quizzes and flashcards, based on user-provided documents.

Entities

Artificial Analysis · Gemini · Google · Google DeepMind · OpenAI

Claims

What the coverage asserts, and how many sources carry each claim.

Sources

about 13 hours ago
about 8 hours ago
about 16 hours ago
about 9 hours ago
about 13 hours ago
about 18 hours ago
about 6 hours ago