TLDR
- Google launched three new Gemini AI models: Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber
- Gemini 3.6 Flash reduces output token usage by 17% versus its predecessor and is priced at $1.50/1M input tokens
- Gemini 3.5 Flash Cyber is a security-focused model available only to governments and trusted partners via CodeMender
- 3.5 Flash-Lite runs at 350 output tokens per second, priced at $0.30/1M input tokens
- The launches come ahead of Alphabet’s quarterly earnings, with GOOGL stock down 0.66% to $349.66
Alphabet launched three new Gemini AI models on Tuesday, just days before the company is due to report quarterly earnings. GOOGL stock was trading at $349.66, down 0.66% on the day.
The new models — Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber — are all aimed at developers and enterprises building AI-powered applications at scale.
Gemini 3.6 Flash is the headline release. Google says it outperforms its predecessor, 3.5 Flash, on coding, knowledge work, and multimodal tasks, while using 17% fewer output tokens. It’s priced at $1.50 per million input tokens and $7.50 per million output tokens.
Today we’re expanding the Gemini family with three new models built to be faster, more token efficient, and reliable at scale.
Meet the new Gemini models ↓ pic.twitter.com/hVYNQqxgjd
— Google (@Google) July 21, 2026
On coding benchmarks, 3.6 Flash scores 49% on DeepSWE, up from 37% for 3.5 Flash. On MLE Bench, it jumps to 63.9% from 49.7%. Computer use performance on OSWorld-Verified also improved to 83.0% from 78.4%.
Customers including Figma, Harvey, Hebbia, and JetBrains are already using the model. Both Hebbia and Harvey flagged its multimodal strengths — particularly document parsing and chart analysis.
Gemini 3.5 Flash-Lite: Speed at Low Cost
The 3.5 Flash-Lite model is built for high-throughput tasks where speed matters. It runs at 350 output tokens per second and is priced at $0.30/1M input tokens and $2.50/1M output tokens.
Flash-Lite significantly outperforms its predecessor, 3.1 Flash-Lite, on agentic and coding tasks. On Terminal-Bench 2.1, it scores 54% versus 31%. On SWE-Bench Pro, it scores 54.2% versus 49.6% for the older 3 Flash model.
Palo Alto Networks and Ramp are among the early customers, pointing to its speed and cost efficiency for scaling workflows.
Cybersecurity Gets Its Own Model
Gemini 3.5 Flash Cyber is a purpose-built security model, fine-tuned from 3.5 Flash to find and fix software vulnerabilities. It operates within Google’s CodeMender agent framework, where multiple Flash Cyber agents work together to produce a combined security report.
Google is taking a cautious approach here. The model will only be available to governments and trusted partners through a limited-access pilot, citing the dual-use risks of the technology.
On the CyberGym benchmark, Flash Cyber reaches what Google describes as competitive frontier performance.
The cybersecurity push puts Google in more direct competition with OpenAI and Anthropic, both of which have recently released tools aimed at security researchers and vulnerability detection.
Google’s flagship Gemini 3.5 Pro model is still in testing with partners. It was originally expected in June but has no confirmed launch date. The company also confirmed it has begun pre-training for Gemini 4.
Both 3.6 Flash and 3.5 Flash-Lite are available now via Google AI Studio, Android Studio, Gemini Enterprise, and the Gemini app. Flash-Lite is also rolling out inside Google Search.
Stop guessing and start investing with confidence. KnockoutStocks gives you the AI insights, market intelligence, and stock research you need to spot opportunities, cut through the noise, and make smarter investment decisions — all in one powerful platform.
Sign up today and get 50% OFF full access to our premium stock picks.
Simply use coupon code SPECIAL50 at checkout to claim your exclusive discount.







