:

GOOGLE CUTS GEMINI 3.6 FLASH PRICING BELOW 3.5

AI DESK2 MIN READ
TUE, JUL 21, 2026

■ AI-SUMMARIZED FROM 5 SOURCES ▸ TIMELINE

Google released Gemini 3.6 Flash at lower prices than its predecessor while launching two additional models. The new flagship costs $1.50 per million input tokens and $7.50 per million output tokens.

Google announced three new Gemini models Tuesday, reshaping its AI pricing structure with more aggressive positioning against competitors. Gemini 3.6 Flash undercuts the previous generation with pricing at $1.50/1M input tokens and $7.50/1M output tokens—a notable reduction from 3.5 Flash's rates. The model delivers improved performance on coding, knowledge work, and multimodal tasks while reducing output token usage by up to 17% versus 3.5 Flash. Gemini 3.5 Flash-Lite targets cost-sensitive applications at $0.30/1M input tokens and $2.50/1M output tokens, positioning itself as Google's most affordable option for lighter workloads. Google also introduced Gemini 3.5 Flash Cyber, a specialized model addressing cybersecurity use cases where Anthropic has established early market presence. The pricing moves reflect intensifying competition in the AI model market, where cost efficiency has become a primary differentiator. By reducing prices on improved models, Google aims to accelerate adoption among developers building AI agents and applications requiring high throughput at scale. Looking ahead, Google confirmed it has begun pre-training for Gemini 4, described as its "most ambitious pre-training run yet." The company also teased an upcoming Gemini 3.5 Pro, though no release date or pricing has been announced. These announcements arrive as major AI labs compete on both capability and economics. Google's strategy prioritizes efficiency and cost reduction alongside performance improvements, signaling confidence in its ability to deliver stronger models at lower price points. The expanded model lineup—ranging from specialized cybersecurity variants to ultra-low-cost options—allows Google to serve different customer segments simultaneously. Developers can access the new models through Google's AI API and cloud services, with Gemini 3.6 Flash available immediately.

■ SOURCES

Ars TechnicaTechmemeTechmemeTechmemeHacker News

■ SUMMARY WRITTEN BY AI FROM THE LINKS ABOVE

■ MORE FROM THE AI DESK

Streaming services are abandoning their specialized formats as artificial intelligence makes content creation, organization, and recommendations simpler. Spotify, Netflix, YouTube, and TikTok are converging into all-purpose entertainment destinations.

JUST NOWAI Desk

Wealth managers earning upwards of $500,000 annually are deploying AI to handle routine tasks, redirecting their focus toward client advisory work. The technology shift marks an early adaptation phase as the industry grapples with automation's broader implications.

JUST NOWAI Desk

Substack is rolling out an AI detection feature powered by Pangram that lets readers identify potentially AI-generated content on the platform. The tool scans posts, notes, replies, and comments to estimate how much text may be AI-written or AI-assisted.

JUST NOWAI Desk

Poolside has launched Laguna S 2.1, an 118-billion parameter open-weight model designed for agentic coding and long-horizon tasks. The company claims it competes with significantly larger open models.

1H AGOAI Desk

■ SUBSCRIBE TO THE DAILY BRIEF

ONE EMAIL, 5 STORIES, 06:00 UTC. UNSUBSCRIBE ANYTIME.