Observed Signal · Jul 21, 2026 · Product Launch · Source: CNBC Technology · Impact: 4/5 · Sentiment: Positive

Google expands Gemini with cheaper models, Mythos rival

Executive Signal Summary

Google DeepMind released three new Gemini models—Gemini 3.6 Flash, 3.5 Flash‑Lite and 3.5 Flash Cyber—aimed at users building and running AI agents with improvements in efficiency, latency and reliability. Gemini 3.6 Flash is positioned as the workhorse: it reportedly uses about 17% fewer tokens than its predecessor, boosts coding, multimodal and knowledge‑work performance, and Google says it is cheaper per task than some competing offerings. Flash‑Lite is the fastest, most cost‑efficient option for high‑volume or cost‑sensitive workloads. Flash Cyber is fine‑tuned to find and patch security vulnerabilities and will initially be available only to governments and trusted partners via a limited pilot. Google is also testing Gemini 3.5 Pro with partners (launch delayed for performance fixes), developing a specialized chip to run Gemini more efficiently, and has begun a major pretraining run for Gemini 4.

Polaris7 AgentPolaris7 Strategic Assessment
High Confidence

Major-platform technical model launches from Google/Alphabet affect model cost, competitive positioning vs. Anthropic/OpenAI/Chinese rivals, and have implications for infrastructure and AI serving economics.

SIGNAL RADAR

Track Google Signals & Market Shifts in Real-Time

Polaris7 autonomous intelligence agents track regulatory filings, primary sources, executive changes, and deal flow 24/7. Create your free Explorer workspace to monitor these entities.

Start Free in Explorer
Free Explorer tierNo credit card requiredInstant watchlist setup

Key Takeaways & Evidence Grounding

  • Google released three new Gemini models: 3.6 Flash, 3.5 Flash‑Lite and 3.5 Flash Cyber.
  • Gemini 3.6 Flash cuts token usage by about 17%, improves coding, multimodal and knowledge‑work performance, and is claimed to be cheaper per task than some rivals (e.g., OpenAI, Moonshot AI, Alibaba).
  • Gemini 3.5 Flash‑Lite is positioned as the fastest, most cost‑efficient option for high‑volume or cost‑sensitive workloads.
  • Gemini 3.5 Flash Cyber is tailored to detect and patch software vulnerabilities and will be available initially only to governments and trusted partners via a limited pilot.
  • Google is testing Gemini 3.5 Pro with partners (launch delayed due to performance issues), is developing a specialized chip to run Gemini more efficiently, and has started a major pretraining run for Gemini 4.

Connected Companies & Entities

7 Entities mapped

“Google is also launching Gemini 3.6 Flash, which improves coding, multimodal and knowledge-work performance while using up to 17% fewer toke...”

“Alphabet is releasing three new Gemini models on Tuesday, including its clearest answer yet to Anthropic’s lead in cybersecurity, as the com...”

“A Google Cloud spokesperson told CNBC in a statement that its teams are 'constantly researching and experimenting with new innovations to de...”

“The new model could help Google narrow its cybersecurity gap with Anthropic, which has built an early lead in automated code defense with it...”

“Artificial Analysis data shows Gemini Flash already undercuts comparable models from Anthropic, OpenAI and Chinese rivals on cost....”

“Moonshot AI’s Kimi K3 drew enough demand that the company limited new subscriptions and API access because of capacity constraints....”

“Alibaba is teasing Qwen 3.8 Max, which it said trails only Anthropic’s Fable 5 in overall performance....”

Primary Source Grounding & Direct Attribution
Direct Origin Attribution
Primary Reporting: CNBC Technology•Published: Jul 21, 2026
Original Coverage Title: “Google expands Gemini lineup with cheaper models and new Mythos rival”

Related Market Signals & Shifts

Recent verified developments and strategic activity across this market segment.

AI Model ReleaseSep 3, 2026

Google Launches Gemini 3.8 Flash and Cyber AI Models

Google launched Gemini 3.8 Flash and Gemini 3.8 Flash Cyber on September 3, 2026, both derived from the same foundation model. Flash is optimized for coding, reasoning, and AI agents, while Flash Cyber specializes in defensive cybersecurity and vulnerability patching. The models lead the HLE-Verified benchmark (54.9%) and perform strongly on DeepSWE v1.1, with Flash Cyber scoring 86.2% on Cybergym (beating GPT-5.5 Cyber) and 47.2% Pass@1 on CWE-Bench. Pricing remains $0.75 per million input tokens and $3.75 per million output tokens through 2026, then doubles in 2027, undercutting rivals like OpenAI and Anthropic. Cyber access is restricted via the Fairwind Program to select partners (e.g., CrowdStrike, Snowflake, government, critical infrastructure). Models are available across Google platforms, with MrBeast promoting them.

Read assessment
Large Language Models (LLM) & AIMar 4, 2026

Gemini 3.1 Flash-Lite Arrives: Faster, Cheaper

Google unveils Gemini 3.1 Flash-Lite, a cost-efficient variant designed for speed and enterprise use. The model is described as 2.5 times faster than Gemini 2.5 Flash and offers lower costs, with pricing of 0.25 USD per million input tokens and 1.50 USD per million output tokens. It features dynamic Thinking Levels that let users tune the model's reasoning depth. Gemini 3.1 Flash-Lite is available now as a Preview in the Gemini API via Google AI Studio and to enterprises on Vertex AI. Google also notes a 45% improvement in output tempo. In benchmarks, it achieved around 86.9% on the GPQA Diamond test. Google showcases deployment scenarios ranging from translations and content moderation to dashboards and CRM processes, including a Retail Business Agent that can plan and execute multi-step tasks like reporting and dashboard automation.

Read assessment
Large Language Models (LLM) & AIDec 18, 2025

Gemini 3 Flash Becomes Default AI Mode Model

Google announces Gemini 3 Flash as the default model for its Gemini App, AI Mode in Google Search, and related AI workflows, emphasizing speed and efficiency. The model introduces features such as Agent CC for Gmail and a Disco Browser, and is positioned as faster and more token-efficient than Gemini 2.5. In benchmarks, Gemini 3 Flash reportedly outs as fast or faster than competing models, with a 33.7% score on Humanity’s Last Exam and improved output latency. The system is described as using about 30% fewer tokens on average than Gemini 2.5 and being three times faster, with costs cited at roughly $0.50 per million input tokens and $3 per million output tokens. Access for enterprises is via Vertex AI and Gemini Enterprise, while developers can use the Gemini API in Google AI Studio, Gemini CLI, and the new Google Antigravity platform. The rollout is described as global, establishing Gemini 3 Flash as a foundational AI capability for search and apps.

Read assessment

Track Real-Time Market Signals & Shifts

Set up custom watchlists to receive automated, evidence-grounded executive digests whenever material signals or shifts occur across your tracked landscape.