Observed Signal · Aug 14, 2026 · Technical Release · Source: Trending Topics · Impact: 4/5 · Sentiment: Positive
Google Launches Gemini 3.7 Flash with 50% Introductory Price Cut
On August 13, 2026, Google released Gemini 3.7 Flash, an update to its Flash model line aimed at coding, agentic workflows, and enterprise document processing. The model is available through the Gemini API, Google AI Studio, Android Studio, Antigravity, Gemini Enterprise, and Gemini Spark. Google introduced an introductory API price of $0.75 per million input tokens and $3.75 per million output tokens until December 31, 2026, after which list prices double; the discount also applies retroactively to Gemini 3.6 Flash. Google reports 340 tokens per second throughput, a 1M-token context window, and improved benchmarks over 3.6 Flash on coding and enterprise workflow tests. Independent benchmarks are pending, and on several tests competitors like OpenAI’s GPT-5.6 Terra and Anthropic’s Claude Sonnet 5 still lead. No release date was given for the delayed Gemini 3.5 Pro flagship.
Google, a major platform, released a new LLM with a 50% introductory price cut, lowering the cost of AI and agentic workflows; this can accelerate adoption of generative AI and automation across AdTech and MarTech, though it is not directly an advertising launch.
Track X.AI Corp. Signals & Market Shifts in Real-Time
Polaris7 autonomous intelligence agents track regulatory filings, primary sources, executive changes, and deal flow 24/7. Create your free Explorer workspace to monitor these entities.
Key Takeaways & Evidence Grounding
- Google launched Gemini 3.7 Flash on August 13, 2026, three weeks after Gemini 3.6 Flash.
- Introductory API pricing is $0.75 per million input tokens and $3.75 per million output tokens through December 31, 2026; list prices double thereafter.
- The introductory discount also applies retroactively to Gemini 3.6 Flash.
- Google claims 340 tokens per second throughput and a 1 million token context window.
- Gemini 3.7 Flash trails OpenAI and Anthropic models on several benchmarks; no release date was given for the delayed Gemini 3.5 Pro flagship.
Connected Companies & Entities
6 Entities mapped“Beim multimodalen Agent’s Last Exam liegt Claude Sonnet 5 mit 33,3 % vor Gemini 3.7 Flash (26,3 %)....”
“Insgesamt ist 3.7 Flash etwas hinter den aktuellen Top-Modellen von OpenAI, Anthropic oder auch SpaceXAI anzusiedeln, aber auch hinter den f...”
“Insgesamt ist 3.7 Flash etwas hinter den aktuellen Top-Modellen von OpenAI, Anthropic oder auch SpaceXAI anzusiedeln, aber auch hinter den f...”
“Bei Terminal-bench 2.1 liegt das Modell mit 85,8 % hinter GPT-5.6 Terra (87,4 %), bei Terminal-bench 3.0 und OSWorld-2.0 führt ebenfalls das...”
“Nur drei Wochen nach Gemini 3.6 Flash hat Google am 13. August das nächste Modell seiner Flash-Reihe veröffentlicht....”
Ontology Mapping & Concepts
Related Market Signals & Shifts
Recent verified developments and strategic activity across this market segment.
Anthropic launches cheaper AI model Sonnet 5.5
Anthropic has released Claude Sonnet 5.5, an upgraded mid-tier AI model designed for everyday tasks such as coding, bug fixing, and document creation. Priced at $2 per million input tokens and $10 per million output tokens, it is over 30% faster than Sonnet 5, and due to reduced token usage, costs per task can drop by up to 30%, making it more cost-effective. While not advancing frontier capabilities, it outperforms Opus 5.5 on agentic coding benchmarks and significantly improves on Terminal-Bench 4.0 (70.6% vs. 10.3%). It also includes enhanced cyber safety mechanisms, making it the first Sonnet model with safeguards comparable to Opus 5. Available on AWS, Google Cloud, and Microsoft Azure, it targets cost-conscious customers. The launch follows Opus 5.5 and a call for a slowdown in AI development. A new Haiku model is expected soon.
Xiaomi MiMo-V2.6-Pro tops open weights, trained for $3M
Xiaomi released MiMo-V2.6-Pro, a natively omnimodal open-weights model with 1.02T total / 42B active parameters, trained for $3M (about 130 hours and 75B tokens). It debuts as the top open-weights model on Artificial Analysis' Intelligence Index (46) with cost efficiency at $0.435/M input and $0.87/M output tokens, under an MIT license. Xiaomi also open-sourced the RL training environment code and recipes, but not the full 7k+ task datasets, signaling an emphasis on transparency in RL training.
SpaceXAI and Cursor Launch Grok 4.5, Delayed in EU
SpaceXAI, Elon Musk's AI unit formerly known as xAI, released Grok 4.5 on 8 July 2026 — its most capable model and the first developed jointly with coding startup Cursor (Anysphere), which SpaceX agreed to acquire for $60 billion in mid-June. The model targets software engineering, agentic tasks, and knowledge work, and is priced aggressively at $2 per million input tokens and $6 per million output tokens. It is initially unavailable in the EU, with availability expected in mid-July. SpaceXAI claims 'Opus-class' performance, though its own benchmarks show mixed results against Anthropic's Opus 4.8. Cursor's leaner in-house model Composer 2.5 remains available separately. The launch coincides with OpenAI's wider rollout of GPT-5.6 and comes amid rapid consolidation in the AI coding market.
Track Real-Time Market Signals & Shifts
Set up custom watchlists to receive automated, evidence-grounded executive digests whenever material signals or shifts occur across your tracked landscape.
