Google launched Gemini 3.5 Flash in August 2026 — a faster, cheaper flash model with 1M context, vision support, and improved reasoning over Gemini 3.1 Flash. Full pricing and API access details.
Gemini 3.5 Flash is Google's latest flash-tier model, released in August 2026 as an upgrade to Gemini 3.1 Flash. It features a 1,000,000-token context window, vision support for image inputs, and improved instruction following. The model is designed for high-throughput workloads where speed and cost matter more than maximum reasoning depth.
Compared to Gemini 3.1 Flash, the 3.5 version offers better performance on coding, math, and instruction-following benchmarks while maintaining a similar cost profile. It's a strong fit for chatbots, content generation, and light coding tasks.
Gemini 3.5 Flash costs $1.50 per million input tokens and $9.00 per million output tokens at standard rates. This positions it above Gemini 3.1 Flash but below Gemini 3.1 Pro, offering a balance of capability and cost for flash-tier workloads.
| Model | Input | Output | Context |
|---|---|---|---|
| Gemini 3.5 Flash (gemini-3.5-flash) | $1.50 | $9.00 | 1,000,000 |
Source: Google official pricing as of August 2026. Pricing is per million tokens.
Gemini 3.5 Flash sits in the mid-range of flash models. At $1.50/$9.00 per MTok, it costs more than DeepSeek V4 Flash ($0.14/$0.43) and MiMo V2.5 ($0.10/$0.40) but offers stronger Google-specific integrations and multimodal capabilities. It's cheaper than Claude Sonnet 5 ($2/$10 promo) and GPT-5.4 ($2/$8).
The 1M context window and vision support make it a good fit for document processing, image analysis, and multi-turn conversations where the model needs to reference earlier context. For pure text tasks at lower cost, DeepSeek V4 Flash or MiMo V2.5 may be more economical.
Gemini 3.5 Flash is available through the Google AI API and Vertex AI. It supports the OpenAI chat completions format through compatible providers, so existing integrations can switch by changing the model ID to gemini-3.5-flash.
APItokendeal provides OpenAI-compatible access to Gemini 3.5 Flash through a single API key. Teams can use token packs with prepaid credits — at 77% off official pricing, the effective cost drops to $0.34/$1.71 per MTok.
APItokendeal provides OpenAI-compatible API access to Gemini 3.5 Flash, GPT, Claude, DeepSeek, Kimi, and more. One API key, one dashboard, prepaid token packs with no surprise bills.