Skip to content
ATAI Today Brief
HomeNewsConceptsGuidesToolbox
AboutSubscribeUA
Subscribe

AI Today Brief

The daily AI-engineering brief. Built in public. EN · UA.

XTelegramLinkedInYouTubeRSS

Follow AI Today Brief on LinkedIn for daily AI-engineering updates and the weekly “5 shifts that changed how developers work” PDF.

Explore

NewsDigestsConceptsGuides

Company

SubscribeAdvertiseAbout

Legal

Editorial policyAI disclosurePrivacyTerms

© 2026 AI Today Brief. All rights reserved.

  1. Home/
  2. News/
  3. Tools & releases/
  4. Google Releases Gemini 3.6 Flash and 3.5 Flash-Lite for Agentic Workflows
Tools & releases

Google Releases Gemini 3.6 Flash and 3.5 Flash-Lite for Agentic Workflows

July 22, 2026· 4 min read
OKCurated by Oleksandr Kuzmenko, AI Product Engineer·Updated July 22, 2026·Sources cited on every story
AI-assisted · editor-reviewed·How we use AI
Google Releases Gemini 3.6 Flash and 3.5 Flash-Lite for Agentic Workflows

Google introduced Gemini 3.6 Flash and Gemini 3.5 Flash-Lite, designed to lower latency and token usage in agentic workflows. Gemini 3.6 Flash reduces output tokens by 17% while costing $1.50/1M input and $7.50/1M output, and adds native client-side computer use capabilities.

Impact: High

Why it matters

Developers can immediately swap 3.5 Flash for 3.6 Flash or 3.5 Flash-Lite to reduce token costs and improve multi-step agent performance.

TL;DR

  • 01Gemini 3.6 Flash delivers a 17% reduction in output token usage and costs $1.50/$7.50 per 1M tokens.
  • 02Gemini 3.5 Flash-Lite achieves 350 tokens/sec at $0.30/$2.50 per 1M tokens with SWE-Bench Pro score of 54.2%.
  • 03Computer use is now available natively as a client-side API tool across Gemini Flash models.

Key facts

3.6 Flash Input Price$1.50 / 1M tokens
3.6 Flash Output Price$7.50 / 1M tokens
3.5 Flash-Lite Input Price$0.30 / 1M tokens
3.5 Flash-Lite Output Price$2.50 / 1M tokens
3.6 Flash Input Price
$1.50 / 1M tokens
3.6 Flash Output Price
$7.50 / 1M tokens
3.5 Flash-Lite Speed
350 tokens/sec (self-reported)
3.5 Flash-Lite Input Price
$0.30 / 1M tokens
3.5 Flash-Lite Output Price
$2.50 / 1M tokens
3.6 Flash DeepSWE Benchmark
49% (self-reported)

Model Specifications and Pricing

Google's Gemini 3.6 Flash is priced at $1.50 per 1M input tokens and $7.50 per 1M output tokens. Gemini 3.5 Flash-Lite operates at $0.30 per 1M input tokens and $2.50 per 1M output tokens, achieving an execution speed of 350 output tokens per second.

Benchmark Improvements and Token Efficiency

  • Gemini 3.6 Flash achieves 49% accuracy on DeepSWE (up from 37% in 3.5 Flash) and 63.9% on MLE Bench.
  • On OSWorld-Verified, 3.6 Flash scores 83.0% with built-in client-side computer use support.
  • Gemini 3.5 Flash-Lite reaches 54% on Terminal-Bench 2.1, 72.2% on GDM-MRCR v2, and 54.2% on SWE-Bench Pro.

Specialized Domain Models

Google also announced Gemini 3.5 Flash Cyber, fine-tuned specifically for vulnerability detection and automated patching within the CodeMender framework.

Try it in 2 minutes

curl https://generativelanguage.googleapis.com/v1beta/models/gemini-3.6-flash:generateContent?key=$GEMINI_API_KEY

bash

✓ When to use

  • Use Gemini 3.6 Flash as the execution engine in multi-agent orchestration to cut token costs by up to 17%
  • Use Gemini 3.5 Flash-Lite for high-throughput sub-agent tasks like search, document parsing, and visual web design generation

✕ When NOT to use

  • Avoid using Flash-Lite as the primary planning or reasoning master agent for complex multi-step architecture design

What to do today

  • →Update Gemini SDK to test 3.6 Flash in agent execution pipelines to reduce output token consumption.
  • →Benchmark 3.5 Flash-Lite as a high-throughput worker for document parsing and agentic search.

What the community says

  • “Fable or gpt5.6 sol for planning. Gemini 3.6 Flash for executing. Wow, Google is onto something here.”

    — ernestrc on Hacker News

  • “This is a very fast model. I was already impressed by how fast 3.5 Flash was.”

    — bjackman on Hacker News

#Gemini#Gemini API#CodeMender

Sources

  • Introducing Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber
ShareShare on XShare on LinkedIn
Next story →NVIDIA Isaac Lab 3.0 Decouples Simulation Engine for Physical AI Training

Related stories

  • Tools & releasesAndrew Ng Releases OpenWorker for Local Desktop AI Deliverables
  • Tools & releasesMoonshot AI Releases Kimi Code CLI Terminal Agent with Agent Client Protocol Support
  • Tools & releasesRunway Launches Media Router API for Dynamic Generative Model Selection
  • Tools & releasesNVIDIA Isaac Lab 3.0 Decouples Simulation Engine for Physical AI Training

Email digest

Get the morning AI brief

One email a day — the stories that matter for engineers, founders and tech leads. Human-edited, with links to primary sources.

  • ✓120+ sources scanned daily
  • ✓Edited by a human
  • ✓1 email per day
  • ✓EN + UA

By subscribing you agree to the privacy policy.