Close Menu
GeekPlanet

    Subscribe to Updates

    Be Geeky and subscribe to GeekPlanet for Technology, Security and Gadgets.

    What's Hot

    Tips to Secure Your Online Banking in India

    Android Privacy Controls You Should Enable Right Now

    Upcoming Budget Smartwatches in India This Quarter

    Facebook X (Twitter) Instagram
    • Privacy & Policy
    • Terms & Conditions
    • Contact US
    Facebook X (Twitter) Instagram YouTube
    GeekPlanetGeekPlanet
    AtlasVpn
    • Home
    • Gadgets
    • Entertainment
    • Cyber Security
    • How To’s & Guides
    • Reviews
    • Python
    GeekPlanet
    Home - AI - Google Gemini 3.6 Flash Launches With 17% Better Token Efficiency
    AI

    Google Gemini 3.6 Flash Launches With 17% Better Token Efficiency

    Geek PlanetBy Geek Planet4 Mins Read
    Facebook Twitter Pinterest LinkedIn Telegram Tumblr Email
    Google Gemini 3.5 and 3.6 Flash AI models announcement
    Share
    Facebook Twitter LinkedIn Pinterest Email
    Gemini 3.6 Flash benchmark results

    Google announced Gemini 3.6 Flash on July 21, 2026, along with two other new models in the Flash family. The update targets developers building production AI agents who need better token efficiency, lower latency, and more reliable performance at scale.

    What Gemini 3.6 Flash Brings to the Table

    Gemini 3.6 Flash is Google’s new workhorse model for coding, knowledge work, and multimodal tasks. According to the Artificial Analysis Index, it uses 17% fewer output tokens than Gemini 3.5 Flash while delivering better quality across the board.

    The pricing sits at $1.50 per 1 million input tokens and $7.50 per 1 million output tokens. Google claims this makes it cheaper per task than GPT-5.6 Terra Max, Kimi K3, and Qwen 3.7 Max.

    Key Benchmark Improvements

    • DeepSWE: 49% accuracy vs. 37% for 3.5 Flash (software engineering tasks)
    • MLE Bench: 63.9% vs. 49.7% (machine learning research)
    • OSWorld-Verified: 83.0% vs. 78.4% (computer use capabilities)
    • GDPval-AA v2: 1421 vs. 1349 (knowledge work and document analysis)

    Customers including Figma, Harvey, Hebbia, and JetBrains have already tested 3.6 Flash and report improvements in document parsing, chart analysis, report drafting, and code generation workflows.

    Gemini 3.5 Flash-Lite: Speed at Scale

    The second model, Gemini 3.5 Flash-Lite, runs at 350 output tokens per second according to Artificial Analysis. It’s priced at $0.30 per 1 million input tokens and $2.50 per 1 million output tokens, making it ideal for high-volume agentic workflows.

    Flash-Lite outperforms the older Gemini 3 Flash on several benchmarks:

    • SWE-Bench Pro: 54.2% vs. 49.6%
    • OSWorld-Verified: 74.0% vs. 65.1%
    • Terminal-Bench 2.1: 54% vs. 31%

    The model includes configurable thinking levels, from minimal (low latency) to higher levels for complex multi-step tasks. Computer use is now a built-in tool in the Gemini API.

    Gemini 3.5 Flash Cyber: Security-Focused Model

    The third release, Gemini 3.5 Flash Cyber, is fine-tuned specifically for finding and fixing cybersecurity vulnerabilities. It powers Google’s CodeMender agent, which uses multiple Flash Cyber agents working together to produce security reports.

    Google calls this the “first truly competitive AI cyber defense system.” On the CyberGym benchmark, CodeMender with Flash Cyber reaches frontier-level performance at a fraction of the cost of larger models.

    Due to the dual-use nature of this technology, Flash Cyber will be available exclusively to governments and trusted partners through CodeMender as a limited-access pilot program.

    Gemini 3.5 Pro: Still in Testing

    Google also confirmed that Gemini 3.5 Pro is currently testing with partners and will be made broadly available when ready. Meanwhile, the team has started what they call their “most ambitious pre-training run yet” for Gemini 4.

    Both 3.6 Flash and 3.5 Flash-Lite are available starting today through the Gemini API and Google AI Studio.

    Frequently Asked Questions

    How much does Gemini 3.6 Flash cost?

    Gemini 3.6 Flash costs $1.50 per 1 million input tokens and $7.50 per 1 million output tokens. This is cheaper than 3.5 Flash on a per-task basis due to the 17% reduction in output token usage.

    Is Gemini 3.6 Flash better than GPT-5.6?

    According to Google’s claims, 3.6 Flash is cheaper per task than GPT-5.6 Terra Max, Kimi K3, and Qwen 3.7 Max. Independent benchmarks show it performing at or slightly below Grok 4.5, but at five times the speed.

    What is Gemini 3.5 Flash Cyber used for?

    Flash Cyber is fine-tuned for finding and fixing cybersecurity vulnerabilities. It powers Google’s CodeMender agent and will be available to governments and trusted partners through a limited pilot program.

    When will Gemini 3.5 Pro be released?

    Gemini 3.5 Pro is currently in testing with partners. Google has not announced a specific release date but says it will be available “as soon as it’s ready.”

    Can I use Gemini 3.6 Flash for coding?

    Yes. 3.6 Flash is optimized for code generation and software engineering tasks. It scored 49% on the DeepSWE benchmark compared to 37% for 3.5 Flash, showing significant improvement in coding workflows.

    Share. Facebook Twitter Pinterest LinkedIn Tumblr Email
    Geek Planet
    • Website
    • Facebook
    • X (Twitter)
    • Instagram

    Hello, tech enthusiasts! I'm Devender, your guide through the ever-evolving world of technology. With a passion for innovation and a knack for breaking down complex concepts into digestible bits, I'm here to help you navigate the digital frontier.

    Related Posts

    Generative AI Tools Helping Indian Startups Scale Fast

    AI: Latest Updates Every Tech User Should Know August 2026

    Generative AI Tools Helping Indian Startups Scale Fast

    Efficient AI Writing Tools for Indian Small Business Owners

    Local AI Model Deployment: Practical Guide for Small Business

    AI Developments Changing Indian Tech: What Users Need to Know August 2026

    Leave A Reply Cancel Reply

    Top Posts

    Tips to Secure Your Online Banking in India

    Data Structures and Their Functions in Python

    How do I edit a sent message on WhatsApp?

    Don't Miss

    Tips to Secure Your Online Banking in India

    Follow these expert tips to keep your bank accounts safe from hackers.

    Android Privacy Controls You Should Enable Right Now

    Upcoming Budget Smartwatches in India This Quarter

    Generative AI Tools Helping Indian Startups Scale Fast

    Stay In Touch
    • Facebook
    • YouTube
    • TikTok
    • WhatsApp
    • Twitter
    • Instagram
    TPS4
    Most Popular

    Tips to Secure Your Online Banking in India

    Data Structures and Their Functions in Python

    How do I edit a sent message on WhatsApp?

    Our Picks

    Tips to Secure Your Online Banking in India

    Android Privacy Controls You Should Enable Right Now

    Upcoming Budget Smartwatches in India This Quarter

    Subscribe to Updates

    Be Geeky and subscribe to GeekPlanet for Technology, Security and Gadgets.

    Facebook X (Twitter) Instagram Pinterest
    © 2026 GeekPlanet.in Managed by MyAdsMantra Global.

    Type above and press Enter to search. Press Esc to cancel.

    Ad Blocker Enabled!
    Ad Blocker Enabled!
    GeekPlanet is a safe place for every tech lover and is made possible by displaying online advertisements to our visitors. Please support us by disabling your Ad Blocker.