:

GOOGLE CUTS GEMINI 3.6 FLASH PRICING BELOW 3.5

AI DESK2 MIN READ
TUE, JUL 21, 2026

■ AI-SUMMARIZED FROM 5 SOURCES ▸ TIMELINE

Google released Gemini 3.6 Flash at lower prices than its predecessor while launching two additional models. The new flagship costs $1.50 per million input tokens and $7.50 per million output tokens.

Google announced three new Gemini models Tuesday, reshaping its AI pricing structure with more aggressive positioning against competitors. Gemini 3.6 Flash undercuts the previous generation with pricing at $1.50/1M input tokens and $7.50/1M output tokens—a notable reduction from 3.5 Flash's rates. The model delivers improved performance on coding, knowledge work, and multimodal tasks while reducing output token usage by up to 17% versus 3.5 Flash. Gemini 3.5 Flash-Lite targets cost-sensitive applications at $0.30/1M input tokens and $2.50/1M output tokens, positioning itself as Google's most affordable option for lighter workloads. Google also introduced Gemini 3.5 Flash Cyber, a specialized model addressing cybersecurity use cases where Anthropic has established early market presence. The pricing moves reflect intensifying competition in the AI model market, where cost efficiency has become a primary differentiator. By reducing prices on improved models, Google aims to accelerate adoption among developers building AI agents and applications requiring high throughput at scale. Looking ahead, Google confirmed it has begun pre-training for Gemini 4, described as its "most ambitious pre-training run yet." The company also teased an upcoming Gemini 3.5 Pro, though no release date or pricing has been announced. These announcements arrive as major AI labs compete on both capability and economics. Google's strategy prioritizes efficiency and cost reduction alongside performance improvements, signaling confidence in its ability to deliver stronger models at lower price points. The expanded model lineup—ranging from specialized cybersecurity variants to ultra-low-cost options—allows Google to serve different customer segments simultaneously. Developers can access the new models through Google's AI API and cloud services, with Gemini 3.6 Flash available immediately.

■ SOURCES

Ars TechnicaTechmemeTechmemeTechmemeHacker News

■ SUMMARY WRITTEN BY AI FROM THE LINKS ABOVE

■ MORE FROM THE AI DESK

Anthropic's Claude Cowork desktop app now enables users to record their screen while performing a task, add voice commentary, and have Claude convert the recording into a reusable skill.

JUST NOWAI Desk

A field experiment with 1,559 Pakistani judges found that an AI assistant called JudgeGPT increased case resolution by 6.3 percent. Training proved critical—judges who received hands-on instruction saw gains, while untrained judges showed negligible improvements.

JUST NOWAI Desk

OpenAI has opened ChatGPT to advertisers, creating a new revenue stream and expanding the AI chatbot's commercial applications. The move marks a significant shift in how the company monetizes its flagship product.

JUST NOWAI Desk

Alibaba's Qwen team unveiled Qwen-Image-3.0, an image generator capable of creating detailed infographics and multi-language documents with legible text at ten-pixel sizes. The model processes prompts up to 4,500 tokens and supports twelve languages natively.

2H AGOIndustry Desk

■ SUBSCRIBE TO THE DAILY BRIEF

ONE EMAIL, 5 STORIES, 06:00 UTC. UNSUBSCRIBE ANYTIME.