The Tech ArchiveThe Tech ArchiveThe Tech Archive
Small BusinessMarketingDevelopers
ArticlesTopicsSeriesAbout

Get the practical AI brief

Verified, no-hype AI tips you can actually use - in your inbox. Free.

No spam. We verify what we send. Unsubscribe anytime.

The Tech ArchiveThe Tech Archive

The Tech Archive

AI news, analysis & explainers

AboutSmall BusinessMarketingDevelopersArticlesTopicsSeriesMethodologyAI DisclosureCorrections

© 2026 All rights reserved.

Back to home
0 readers reading
  1. Home
  2. Articles
  3. Artificial Intelligence
  4. Google's July 2026 Gemini Updates: 3 New Models, an Always-On Agent, and Smarter Video Explained for Small Business

Contents

Google's July 2026 Gemini Updates: 3 New Models, an Always-On Agent, and Smarter Video Explained for Small Business
Artificial Intelligence

Google's July 2026 Gemini Updates: 3 New Models, an Always-On Agent, and Smarter Video Explained for Small Business

Google shipped Gemini 3.6 Flash, Flash-Lite, Flash-Cyber, Gemini Spark, Notebook Collections, and Gemini Omni in July 2026. Here's what each actually does and which ones help your business.

Sham

Sham

AI Engineer & Founder, The Tech Archive

14 min read
0 views
July 30, 2026

Google's summer of 2026 wasn't about building the smartest AI on paper — it was about building the AI that gets work done fastest, cheapest, and in the background while you sleep. Between July 16 and July 21, 2026, Google shipped three new Gemini models, a 24/7 AI agent that runs without your laptop open, a research organization feature, and a text-to-video tool — all aimed at the same proposition: stop paying for raw intelligence you don't use, and start paying for the right tool on the right job.

The headline release is Gemini 3.6 Flash (with 3.5 Flash-Lite and the restricted 3.5 Flash-Cyber), launched July 21, 2026. It uses ~17% fewer output tokens than the model it replaces, costs $7.50 per million output tokens (down from $9.00), and posted higher scores on coding, long-context, and computer-use benchmarks. The same day, Google shipped Flash-Lite at $0.30/$2.50 per million tokens for high-volume repetitive work, and Flash-Cyber for government vulnerability finding. Gemini Spark — your always-on personal agent — expanded to the $20/month Google AI Pro tier that same week. And Gemini Notebook Collections rolled out to tame your research pile.

Here's what each one does, what it costs, and whether it matters for how you work.

Last verified: July 30, 2026 · Pricing and availability change quickly — verify at Google's official pages before purchasing.

  • Gemini 3.6 Flash: new default workhorse, $1.50/$7.50 per M tokens, 17% fewer output tokens
  • Gemini 3.5 Flash-Lite: cheapest tier, $0.30/$2.50 per M tokens, 350 tokens/sec
  • Gemini 3.5 Flash-Cyber: security-only, restricted to governments/CodeMender — no public access
  • Gemini Spark: 24/7 personal agent, now on AI Pro ($19.99/mo) and AI Ultra ($99.99+/mo)
  • Gemini Notebook Collections: organize notebooks into flexible groups (rolling out)
  • Google Vids + Gemini Omni: generate and edit videos by typing what you want changed

What Did Google Release on July 21, 2026? Three New Gemini Models

Google released three models on a single day: Gemini 3.6 Flash, Gemini 3.5 Flash-Lite, and Gemini 3.5 Flash Cyber, according to Google's official announcement (Confirmed).

Model Input / Output (per 1M tokens) Best for Available to
Gemini 3.6 Flash $1.50 / $7.50 General coding, writing, multimodal, computer use All developers + consumers
Gemini 3.5 Flash-Lite $0.30 / $2.50 High-volume repetitive tasks, data extraction All developers + consumers
Gemini 3.5 Flash-Cyber Not published (pilot) Finding + patching security vulnerabilities Governments + trusted partners only

Notably, Gemini 3.5 Pro was not released. Google confirmed it's "testing with partners" and that pre-training has already begun on Gemini 4. The signal is deliberate: the workhorse Flash tier keeps iterating while the flagship Pro tier waits.

What makes Gemini 3.6 Flash different?

Gemini 3.6 Flash is the new default in the Gemini family — the model Google wants everyone using by default for everyday work. The model card from Google DeepMind confirms a 1 million-token context window and a March 2026 knowledge cutoff.

The win isn't raw intelligence — it's token efficiency. Google's own figures say 3.6 Flash uses roughly 17% fewer output tokens than its predecessor (3.5 Flash) while scoring higher across DeepSWE (49% vs 37%), MLE-Bench (63.9% vs 49.7%), and OSWorld-Verified computer use (83.0% vs 78.4%), per AI Release Tracker and AIToolsReview. For your budget, that means fewer tokens spent to reach the same answer.

When to use Gemini 3.5 Flash-Lite instead?

Think of Flash-Lite as the model for the work you do a thousand times a day — sorting incoming messages, reading documents, answering simple questions, extracting product features from listings. Google measured it at 350 output tokens per second, per The Keyword announcement. It roughly doubled 3.1 Flash-Lite's Terminal-Bench 2.1 score (54% vs 31%), per AIToolsReview.

You don't need a surgeon to refill your water glass. Flash-Lite is fast, reliable, available all day — the model you point at volume work, not complex reasoning.

What is Gemini 3.5 Flash-Cyber?

Flash-Cyber is a security-specialised model fine-tuned on top of 3.5 Flash to find, validate, and patch software vulnerabilities. It only runs inside Google's CodeMender agent, where multiple Cyber-model sub-agents collaborate to produce a single vulnerability report. It's not available to the public, and is gated to governments and trusted partners under a limited-access pilot program, per Google DeepMind's model page.

It still matters for context: Google is building specialised AI for serious problems — not just chatty assistants. If your business is consumer-facing, you'll never touch Flash-Cyber directly, but the defensive security it enables could protect the platforms you build on.

How Much Do the New Gemini Models Cost?

Pricing is the real story of this release. For developers building production apps on the Gemini API:

Model Input (per 1M tokens) Cached input Output (per 1M tokens) Context
Gemini 3.6 Flash $1.50 $0.15 $7.50 1M
Gemini 3.5 Flash (superseded) $1.50 $0.15 $9.00 1M
Gemini 3.5 Flash-Lite $0.30 — $2.50 1M

Figures from Google's official Gemini API pricing page and verified against FelloAI's pricing tracker as of July 2026. The output price for 3.6 Flash dropped from $9.00 (3.5 Flash) to $7.50 — a 17% cost reduction on the same job.

For consumer subscriptions:

  • Google AI Plus: $4.99/mo (cut from $7.99 on June 8, 2026)
  • Google AI Pro: $19.99/mo (now includes Gemini Spark access, expanded July 2026)
  • Google AI Ultra: from $99.99/mo; the top Ultra tier is $199.99/mo with up to 20x Pro limits

Consumer pricing verified from Suprmind's pricing aggregator and FelloAI, July 2026.

What Is Gemini Spark? The 24/7 Agent That Keeps Working After You Close Your Laptop

Gemini Spark is Google's personal AI agent — not a chat window, but an agent that keeps working after you walk away. CEO Sundar Pichai described it as "your personal AI agent in Gemini app that helps you navigate your digital life, taking action on your behalf and under your direction," per Google's I/O 2026 keynote transcript (Confirmed).

The architecture matters:

  • Runs on dedicated virtual machines on Google Cloud, so it stays on whether your laptop is open or not
  • Powered by Gemini 3.5 and the Google Antigravity harness, which handles long-horizon tasks in the background
  • Integrates with Google's own tools, with third-party integrations coming through MCP (Model Context Protocol)
  • Available via the Gemini app, and soon through email and chat

The shift is the language itself. You don't tell Spark "do this one thing" — you tell it "watch this," "keep checking," "handle this every week without me asking again." That's the agentic shift everyone in AI keeps talking about: from commands to delegations.

What does Gemini Spark cost and who can access it?

  • Started May 19, 2026 for trusted testers at Google I/O 2026
  • U.S. beta opened to Google AI Ultra subscribers the week of May 25
  • As of July 2026, expanded to Google AI Pro subscribers ($19.99/mo) — confirmed by MegaMobile Content
  • Region availability: Supported everywhere Gemini Apps are available except the EEA, Nigeria, Switzerland, and the United Kingdom — verified against Google's Gemini Help Center (AI Agents Library)

Note: One source (Digital Applied) reported Spark launched on AI Pro and Ultra together at I/O; Google's Help Center and most reporting place Pro access as a July 2026 expansion. Treat the July expansion as Confirmed and the May overlap as Reported.

What Are Gemini Notebook Collections? The Organization Fix You Asked For

If you've used Gemini Notebook (formerly NotebookLM) heavily, you know the pain: dozens of notebooks with no real order, one flat chronological list. Collections fixes that.

Started rolling out July 21, 2026, per Android Authority, Collections let you group related notebooks the way you'd make a music playlist. A single notebook can belong to multiple Collections. No rigid folders — just flexible, semantically sensible groups.

The practical value: build a Collection for "content ideas," drop in every testimonial, case study, and call recap, then ask Gemini to synthesise them into one clean summary you can turn into next week's post. The feature is rolling out gradually and is currently limited to personal organisation — no support for sharing Collections yet.

For more on the broader NotebookLM platform, see our full guides on Gemini Notebook (formerly NotebookLM) in 2026 and how to organize Gemini Notebook Collections.

What Is Google Vids + Gemini Omni? Video Generation by Typing

Google Vids added Gemini Omni on July 16, 2026, giving you the ability to generate and edit high-quality video clips from a simple text prompt, optionally seeded with an image or sketch as a reference. Per Google's Keyword announcement and Google Workspace Updates:

  • Generate a video by typing what you want to see, optionally adding a photo or rough sketch as a reference
  • Edit by typing: "fix the color-grading," "restyle the visuals in anime," or "remove that New York siren in the background"
  • The model supports step-by-step edits — you tweak without starting from scratch
  • Every AI-generated clip carries an invisible SynthID watermark for content transparency

Availability: Google AI Pro and AI Ultra consumers, plus Workspace Business and Enterprise customers. Rollout is gradual (up to 15 days for visibility). Editing non-AI videos with Omni is not available at launch in the EEA, Switzerland, the UK, Texas, or Illinois.

If you want to embed AI video in your small business content without licensing separate AI tools, this brings video generation into the Workspace apps you already pay for. For a wider look at free AI video options, see our guide on how to use premium AI video models for free in 2026.

The Big Pattern: Efficiency Over Raw Intelligence

Step back and the July 2026 releases tell one story: AI's new bragging rights aren't about being smarter, they're about how little it takes to get the job done.

A year ago, every new model release was a brag about raw intelligence. Now the pitch is token efficiency (17% fewer output tokens), price-per-token cuts ($9 → $7.50 output), throughput (350 tokens/sec on Flash-Lite), and always-on persistence (Spark runs without you). That's a fundamentally different kind of race.

For a business owner, this maps to three concrete choices:

  1. High-volume repetitive work → Flash-Lite at $0.30/$2.50 per M tokens. Don't pay frontier prices for sorting and filtering.
  2. Coding, writing, multimodal → 3.6 Flash at $1.50/$7.50. Same context window, better benchmarks, 17% cheaper output.
  3. Things you keep forgetting to follow up on → Gemini Spark, the always-on agent. Pay $19.99/mo for delegation, not for a smarter answer to a one-shot question.

What This Means for You

You don't need the most powerful model available — you need the right tool for the right job. The three model tiers and Spark give you the dials:

  • Fast and cheap (Flash-Lite) for the tasks you do a thousand times a day.
  • Sharper reasoning (3.6 Flash) for the work that actually needs thinking.
  • A background agent (Spark) for the follow-ups and monitoring you'd otherwise drop.

Pick one update from this release. Not all of them — one. Test it on one real workflow this week. The honest part: tools move fast and feel overwhelming. New models and agents ship almost weekly. You don't need to master all of it overnight. You just need to apply one thing consistently. For deeper reads on where Google's agentic platform is heading, see our guides on Google Antigravity for business automation, Gemini 3.6 Flash for AI agents, Gemini in Google Workspace, and setting up AI agents for productivity in 2026.


FAQ

Q: What did Google release on July 21, 2026? A: Three new Gemini models on a single day: Gemini 3.6 Flash (the new default workhorse, $1.50/$7.50 per million tokens), Gemini 3.5 Flash-Lite (the cheapest tier at $0.30/$2.50 per million tokens), and Gemini 3.5 Flash-Cyber (a security-specialised model restricted to governments and trusted partners through Google's CodeMender agent). Gemini 3.5 Pro was not released and remains in partner testing.

Q: How much does Gemini 3.6 Flash cost? A: Gemini 3.6 Flash costs $1.50 per million input tokens and $7.50 per million output tokens, per Google's official API pricing. That is a 17% reduction from the output price of the model it replaces (Gemini 3.5 Flash at $9.00 per million output tokens). It uses roughly 17% fewer output tokens on the same queries while scoring higher on coding, long-context, and computer-use benchmarks.

Q: What is Gemini Spark and how do I access it? A: Gemini Spark is Google's 24/7 personal AI agent that runs on dedicated Google Cloud virtual machines and continues working after you close your laptop. It was announced May 19, 2026 at Google I/O and expanded to Google AI Pro subscribers ($19.99/mo) in July 2026. Access requires being 18+, a personal Google Account, Keep Activity enabled, and using Spark in English through the Gemini web, mobile, or macOS app. It is not available in the EEA, Nigeria, Switzerland, or the United Kingdom.

Q: What are Gemini Notebook Collections? A: Collections is a new organizational feature rolling out to Gemini Notebook (formerly NotebookLM) starting July 21, 2026 that lets you group related notebooks the way you would build a music playlist. A single notebook can belong to multiple Collections. It's currently limited to personal organisation and does not support sharing yet — the rollout is gradual over up to 15 days.

Q: Can I use Google Vids and Gemini Omni for free? A: Google Vids with Gemini Omni is available to Google AI Pro ($19.99/mo) and AI Ultra subscribers, and to Google Workspace Business and Enterprise customers. It is not available on a free tier. Editing non-AI videos with Omni is not available at launch in the EEA, Switzerland, the UK, Texas, or Illinois. Every AI-generated clip carries an invisible SynthID digital watermark for transparency.

Q: Was Gemini 3.5 Pro released in July 2026? A: No. Google confirmed in its July 21, 2026 announcement that Gemini 3.5 Pro is "currently testing with partners" with no general availability date. The same announcement stated that pre-training has already begun on Gemini 4. The signal is that Google is iterating on its workhorse Flash tier while the flagship Pro tier waits until it can compete with the current frontier.


Sources
  1. Google — "Introducing Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber" (July 21, 2026): https://blog.google/innovation-and-ai/models-and-research/gemini-models/gemini-3-6-flash-3-5-flash-lite-3-5-flash-cyber/
  2. Google DeepMind — Gemini 3.6 Flash model card: https://deepmind.google/models/model-cards/gemini-3-6-flash/
  3. Google DeepMind — Gemini 3.5 Flash Cyber: https://deepmind.google/models/gemini/cyber/
  4. Google AI for Developers — Gemini API release notes (changelog): https://ai.google.dev/gemini-api/docs/changelog
  5. Google AI for Developers — Gemini API pricing: https://ai.google.dev/gemini-api/docs/pricing
  6. Google — "I/O 2026: Welcome to the agentic Gemini era" (Sundar Pichai keynote, May 19, 2026): https://blog.google/innovation-and-ai/sundar-pichai-io-2026/
  7. Google — "Google Vids gets powerful upgrades with Gemini Omni": https://blog.google/products-and-platforms/products/workspace/gemini-omni-personal-avatars/
  8. Google Workspace Updates — "Generate higher quality AI video clips and edit any video with Gemini Omni in Vids" (July 16, 2026): https://workspaceupdates.googleblog.com/2026/07/generate-higher-quality-ai-video-clips-and-edit-any-video-with-Gemini-Omni-in-Vids.html
  9. Google Gemini Help Center — Gemini Spark availability: https://support.google.com/gemini/answer/17094507
  10. Android Authority — "Google adds handy Collections feature to Gemini Notebook" (July 21, 2026): https://www.androidauthority.com/gemini-notebook-collections-feature-3689444/
  11. MegaMobile Content — "Gemini Spark Drops to $20 as Google Expands AI Agent Access" (July 24, 2026): https://www.megamobilecontent.com/news/2026/07/24/gemini-spark-google-ai-pro-rollout/
  12. AI Release Tracker — "Gemini 3.6 Flash — Benchmarks, Specs & Release Date": https://aireleasetracker.com/model/google/gemini-3.6-flash
  13. AIToolsReview — "Gemini 3.6 Flash & 3.5 Flash-Lite: Benchmarks, Pricing" (July 23, 2026): https://aitoolsreview.co.uk/insights/gemini-3-6-flash-3-5-flash-lite
  14. FelloAI — "Gemini Pricing 2026: Complete Guide to Plans, API Costs": https://felloai.com/gemini-pricing/
  15. AI Agents Library — "Gemini Spark Availability: Countries & Requirements": https://www.aiagentslibrary.com/blog/gemini-spark-availability/

Updates & Corrections
  • 2026-07-30 — Initial publication. All facts verified against primary sources (Google's blog, Google DeepMind, Google Gemini API pricing/changelog, Google Workspace Updates, Google Gemini Help Center) as of July 30, 2026. Pricing and availability are volatile; flagged accordingly.

Get the practical AI brief

Verified, no-hype AI tips you can actually use - in your inbox. Free.

No spam. We verify what we send. Unsubscribe anytime.

Tags

#Google Gemini#["AI agents"#"small business AI"]#"Gemini Spark"#"gemini 3.6 flash"

Discussion

0 comments
Sham

Sham

AI Engineer & Founder, The Tech Archive

AI engineer (Azure AI-102/AI-900). Writes practical, tested, hype-free guides on using AI for real work and small business at The Tech Archive.

Related Articles

View all
How to Automate Your Lead Pipeline With an AI Agent in 2026: The 6-Step System
Artificial Intelligence

How to Automate Your Lead Pipeline With an AI Agent in 2026: The 6-Step System

16 min
How to Audit Your Prompts for Claude 5: The Context Engineering Workflow Anthropic Used to Cut 80%
Artificial Intelligence

How to Audit Your Prompts for Claude 5: The Context Engineering Workflow Anthropic Used to Cut 80%

13 min
Buzz by Block: Is Jack Dorsey's Free AI Agent Workspace Actually Usable in 2026?
Artificial Intelligence

Buzz by Block: Is Jack Dorsey's Free AI Agent Workspace Actually Usable in 2026?

17 min
How to Build a Multi-Agent AI Team on One Screen in 2026 (Free Orchestration Stack)
Artificial Intelligence

How to Build a Multi-Agent AI Team on One Screen in 2026 (Free Orchestration Stack)

14 min
Anthropic's Open Weights Position Explained: What Dario Amodei Actually Wants (and Who's Pushing Back)
Artificial Intelligence

Anthropic's Open Weights Position Explained: What Dario Amodei Actually Wants (and Who's Pushing Back)

15 min
How to Make Vox-Style AI Videos With Claude Code and Higgsfield MCP (2026)
Artificial Intelligence

How to Make Vox-Style AI Videos With Claude Code and Higgsfield MCP (2026)

21 min