Gemini 3.5 Flash

4.65
Gemini 3.5 Flash delivers near-Pro intelligence at Flash-tier cost and speed: Pro-level coding proficiency, parallel agentic execution, all at the same price point as a Flash model.
Advertisement 728 × 90
CompanyGoogle
Context1M
Released2026-05
Updated2026-08-17

Gemini 3.5 Flash Overview

Gemini 3.5 Flash is Google’s high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimized for coding proficiency and parallel agentic execution…

Gemini 3.5 Flash Pricing

PlanPriceDescription
Standard API Usage Input $1.50 / Output $9.00 per 1M Tokens Supports up to a 1M-token context window. Output pricing includes reasoning tokens. Cached input is approximately $0.15 per 1M Tokens. Suitable for everyday AI application development.
Batch / Flex API Mode Input $0.75 / Output $4.50 per 1M Tokens Designed for batch processing and non-real-time workloads. Pricing is approximately 50% lower than standard API usage, making it suitable for large-scale content processing and automation tasks.
Priority API Mode Input $2.70 / Output $16.20 per 1M Tokens Provides higher-priority resource access with faster response performance. Priced at approximately 1.8× the standard API rate, ideal for applications requiring higher reliability and lower latency.
Free API Tier Free (Limited Usage) Google AI Studio provides free access for testing, learning, and lightweight development with usage limits.

Gemini 3.5 Flash Key Features

1.Agentic Execution:Autonomously executes multi-step workflows, scoring 83.6% on MCP Atlas and 78.4% on OSWorld-Verified, supporting hours-long unattended operation.

2.Code Generation & Programming:Scores 76.2% on Terminal-Bench 2.1 and 55.1% on SWE-Bench Pro. Delivers a full application in ~10 minutes, significantly faster than competing models.

3.Multimodal Understanding:Processes text, image, audio, and video holistically. Achieves 84.2% on CharXiv Reasoning and 83.6% on MMMU-Pro.

4.Long-Context Processing:Natively supports 128k+ tokens, scoring 77.3% on MRCR v2, ideal for large-scale document analysis and long conversation histories.

Summary

Developers & Programmers:Excels at code generation, debugging, and refactoring. Delivers a full application in ~10 minutes, achieving significant efficiency gains over competing models. Ideal for daily coding assistance and review support.

Enterprises Building Agent Workflows:Specializes in multi-step automation tasks (83.6% on MCP Atlas). Applicable to software pipelines, financial processing, customer onboarding, OCR data extraction, and tax workflows.

Large-Scale Deployments Prioritizing Cost & Speed:Priced at $1.50/1M tokens input and $9.00/1M tokens output—roughly one-third the cost of competitors. Output speed of ~289 tokens/sec is 4x faster, making it ideal for production use cases demanding high throughput and low latency.

Users Handling Multimodal Content:Processes text, images, audio, video, and PDFs holistically. Excels in academic literature reasoning (84.2% on CharXiv Reasoning), suitable for mixed-information processing involving charts, screenshots, and documents.

General Users:Available for free in the Gemini app and Google Search AI Mode. Ideal for daily Q&A, document organization, and content summarization tasks.

Comments (0)

Leave a comment

Advertisement 728 × 90

Google Model Comparison

Model Context Pricing API Released Global Heat
Gemini 3.5 Flash
1M Paid YES 2026-05
93/100
1M YES 2026-08
95/100
1M YES 2026-07
96/100
1M YES 2026-07
74/100
131K Paid YES 2026-06
80/100

Similar Models

Gemini 3.6 Flash
96
Google
Gemini 3.1 Pro
95
Google DeepMind’s latest Gemini Pro model delivers advanced reasoning, multimodal understanding, coding support, and enterprise AI capabilities for professional applications.
google
Gemini 3.7 Flash
95
Just three weeks after the last update, Google has released another new model. Gemini 3.7 Flash is more than a minor refresh. Its coding, Agent, and automation capabilities have all improved noticeably, while API pricing is cut in half through the end of 2026. It may not be the most powerful model available, but for developers, the value proposition is hard to ignore.
Google
Nano Banana 2
80
Google's latest image model delivers pro-grade quality at blazing speed, excels at complex generation and iterative editing, and combines deep understanding with high cost-effectiveness.
Google
Nano Banana Pro
79
Google's most powerful image model, featuring precise text rendering, multi-image fusion, and 4K output, built for professional design.
Google
Gemini 3.1 Flash Lite
77
Google's lightweight multimodal model, supporting text, image, audio, video, and PDF. Delivers low latency and high throughput for large-scale agentic workloads.
Google
Gemini 3.5 Flash Lite
74
Google’s lightweight multimodal AI model optimized for high-volume tasks, AI Agent workflows, and cost-efficient AI applications.
Google
Nano Banana 2 Lite
69
Google's fastest, most cost-efficient image generation model — 4-second output, low cost, built for high-concurrency development and scaled visual applications.
Google

Related Tools

Related News