Observed Signal · Sep 16, 2026 · Product Launch · Source: OnlineMarketing.de · Impact: 4/5 · Sentiment: Positive
Google Launches Gemini 3.8 Live and Extended Thinking Audio Models
Google has introduced two new AI audio models, Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking, designed for near-zero-latency, real-time conversations and complex multi-step tasks. The models support parallel processing of speech, visual inputs, and function calls, ensuring uninterrupted interactions. They are available via the Gemini Live API and Google AI Studio, with pricing at $0.005 per minute for audio input and $0.018 for output. Gemini 3.8 Live is integrated into Search Live for camera-based interaction, while Extended Thinking is incorporated into Gemini Live and Workspace (Docs, Gmail, Keep) for Google AI subscribers. All generated audio is watermarked with SynthID for transparency. Independent tests show strong performance, with Extended Thinking scoring 82.6 on the Speech-to-Speech Index. Additionally, Google highlighted Gemini 3.5 Transcribe, a streaming speech-to-text model with a 4.0% word error rate.
Google's launch of advanced real-time audio models (Gemini 3.8 Live and Extended Thinking) is a significant leap in conversational AI, with potential to transform user interaction paradigms in search, workspaces, and potentially advertising contexts. As a major platform update, it merits high importance.
Track Google Signals & Market Shifts in Real-Time
Polaris7 autonomous intelligence agents track regulatory filings, primary sources, executive changes, and deal flow 24/7. Create your free Explorer workspace to monitor these entities.
Key Takeaways & Evidence Grounding
- Google launched two new AI audio models: Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking for real-time voice interactions.
- Gemini 3.8 Live supports real-time camera and speech processing, auto-switching between 97 languages, and handles parallel API calls and tool use.
- Gemini 3.8 Live Extended Thinking is integrated into Gemini Live and Workspace (Docs, Gmail, Keep) for Google AI subscribers, scoring 82.6 on the Speech-to-Speech Quality Index.
- Pricing for the models is $0.005 per minute for audio input and $0.018 per minute for audio output via the Gemini Live API.
- All audio generated is watermarked with SynthID, and Gemini 3.5 Transcribe was highlighted with a 4.0% word error rate for streaming transcription.
Connected Companies & Entities
1 Entity mapped“Google has introduced two new AI audio models, Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking....”
Ontology Mapping & Concepts
Related Market Signals & Shifts
Recent verified developments and strategic activity across this market segment.
Google Launches Gemini 3.8 Live with Real-Time Avatars for Enterprises
Google has introduced Gemini 3.8 Live with a new Live Avatar feature for enterprise customers. The technology combines speech dialogue with near-real-time generated video, providing digital assistants with a visible, animated presence. Available immediately through Gemini Enterprise, the feature processes audio and visual inputs simultaneously to generate speech, facial expressions, and lip movements of a digital avatar. It supports native speech-to-speech communication in 97 languages, with avatar lip-sync and expressions adapting dynamically. Google positions the technology for customer service, advisory, and interactive digital companionship. Enterprises can also create custom avatars from reference images (initially via allowlist) and all outputs are watermarked with SynthID for authenticity. The system supports asynchronous tool execution during conversations, enabling complex business processes without interrupting the dialogue.
Google Gemini 3 Elevates Search and App AI
Google announces Gemini 3, its latest AI model, integrated directly into the Google Gemini App and activated in Search via AI Mode. The model promises multimodal understanding, nuanced responses, and generative layouts, with agentic capabilities that support multi-step tasks and coding. Gemini 3 is supported by a new Antigravity platform for Vibe Coding, enabling autonomous agent actions and app-level workflows; developers can access Gemini 3 through Google AI Studio, Vertex AI, Gemini CLI, and the new Agentric-Building platform Antigravity. The rollout emphasizes deep integration across Google services (Maps, Canvas, Chrome) and introduces enhanced AI capabilities for developers and end users. Google notes Gemini 3's broad reach (650 million monthly active users; 13 million developers) and cites performance benchmarks from internal assessments. Availability begins in the US for Pro/Ultra users, with broader deployment including Germany planned later. The company positions Gemini 3 as a major step in its AI strategy against competitors like OpenAI, Meta, and Anthropic.
Google Expands Gemini Notebook with Visible Thinking Steps
Google has renamed and expanded NotebookLM into Gemini Notebook, a more integrated research workspace that connects the Gemini app, Google Search and Google Workspace. New user-facing features include expandable "thinking steps" for eligible Google AI Ultra and Pro subscribers, in-notebook code execution, and continued citation-backed outputs (Audio Overviews, Mind Maps). The rollout is phased (initial access for Google AI Ultra and Workspace AI Ultra Access, then Pro on the web) with mobile expansion planned. Google says uploaded sources remain private and are not used to train models; the company frames thinking steps as experimental and an inspection aid rather than proof of correctness.
Track Real-Time Market Signals & Shifts
Set up custom watchlists to receive automated, evidence-grounded executive digests whenever material signals or shifts occur across your tracked landscape.
