August 2026 · Google release

Gemini 3.5 Flash API Pricing & Benchmarks

Google launched Gemini 3.5 Flash in August 2026 — a faster, cheaper flash model with 1M context, vision support, and improved reasoning over Gemini 3.1 Flash. Full pricing and API access details.

What is Gemini 3.5 Flash?

Gemini 3.5 Flash is Google's latest flash-tier model, released in August 2026 as an upgrade to Gemini 3.1 Flash. It features a 1,000,000-token context window, vision support for image inputs, and improved instruction following. The model is designed for high-throughput workloads where speed and cost matter more than maximum reasoning depth.

Compared to Gemini 3.1 Flash, the 3.5 version offers better performance on coding, math, and instruction-following benchmarks while maintaining a similar cost profile. It's a strong fit for chatbots, content generation, and light coding tasks.

Official API pricing

Gemini 3.5 Flash costs $1.50 per million input tokens and $9.00 per million output tokens at standard rates. This positions it above Gemini 3.1 Flash but below Gemini 3.1 Pro, offering a balance of capability and cost for flash-tier workloads.

ModelInputOutputContext
Gemini 3.5 Flash (gemini-3.5-flash)$1.50$9.001,000,000

Source: Google official pricing as of August 2026. Pricing is per million tokens.

How 3.5 Flash compares to other models

Gemini 3.5 Flash sits in the mid-range of flash models. At $1.50/$9.00 per MTok, it costs more than DeepSeek V4 Flash ($0.14/$0.43) and MiMo V2.5 ($0.10/$0.40) but offers stronger Google-specific integrations and multimodal capabilities. It's cheaper than Claude Sonnet 5 ($2/$10 promo) and GPT-5.4 ($2/$8).

The 1M context window and vision support make it a good fit for document processing, image analysis, and multi-turn conversations where the model needs to reference earlier context. For pure text tasks at lower cost, DeepSeek V4 Flash or MiMo V2.5 may be more economical.

API access

Gemini 3.5 Flash is available through the Google AI API and Vertex AI. It supports the OpenAI chat completions format through compatible providers, so existing integrations can switch by changing the model ID to gemini-3.5-flash.

APItokendeal provides OpenAI-compatible access to Gemini 3.5 Flash through a single API key. Teams can use token packs with prepaid credits — at 77% off official pricing, the effective cost drops to $0.34/$1.71 per MTok.

Access Gemini 3.5 Flash and other models

APItokendeal provides OpenAI-compatible API access to Gemini 3.5 Flash, GPT, Claude, DeepSeek, Kimi, and more. One API key, one dashboard, prepaid token packs with no surprise bills.