Google

Gemini 3.5 Flash — Benchmark Scores, Pricing & Performance Analysis

Input Cost
$1.5/1M
Output Cost
$9.0/1M
Context Window
1.0M
Max Output
66K
Knowledge Cutoff
Jan 2025
Accepts
text, image, video, audio, pdf

Gemini 3.5 Flash is an AI model by Google. View detailed benchmark scores, pricing data, and performance metrics on Serenities AI Models.

Specifications last checked against the provider on . Benchmark scores carry their own source, linked beside each number.

General Benchmarks

AA Intelligence Index
3325th of 66

Artificial Analysis composite intelligence score across 10 sub-benchmarks

Artificial Analysis (via OpenRouter)

Coding Benchmarks

Design Arena ELO
127520th of 120

Human preference ELO for building things — websites, UI, games, charts, SVG

Design Arena (via OpenRouter)
AA Coding Index
70.123rd of 96

Artificial Analysis composite coding score

Artificial Analysis (via OpenRouter)
Website Arena ELO
127525th of 108

Human preference ELO for building a working web page from a brief

Design Arena (via OpenRouter)

Reasoning Benchmarks

AA Agentic Index
27.328th of 70

Artificial Analysis composite score for multi-step tool-using tasks

Artificial Analysis (via OpenRouter)
GPQA Diamond
92.8%7th of 140

Graduate-level science Q&A by domain experts

Papers
τ²-Bench Airline
74.7%28th of 93

Multi-turn service agent making tool calls under strict policy constraints

OpenRouter (measured)

Cost Benchmarks

Input Cost
$1.5321st of 413

Cost per 1M input tokens

OpenRouter API
Output Cost
$9.0333rd of 413

Cost per 1M output tokens

OpenRouter API
Cached Input Cost
$0.1578th of 136

Cost per 1M cached input tokens — the price that actually applies to a long agent conversation

OpenRouter API

Context Benchmarks

Context Length
1.0M18th of 451

Maximum context window size

OpenRouter API

Available from 33 providers

The same model costs different amounts depending on who serves it. You can bring your own key for 4 of these — connect it here.

ProviderInput / 1MOutput / 1MCached inContextYour key
Kenari$0$01.0M
UnoRouter$0.1857$1.11421.0M
Kilo Gateway$0.75$4.5$0.0751.0M
302.AI$1.5$91.0M
Abacus$1.5$9$0.151.0M
AIHubMix$1.5$9$1.51.0M
CrossModel$1.5$9$0.151.0M
DevPass (LLM Gateway)$1.5$9$0.151.0M
Eden AI$1.5$9$0.151.0M
FastRouter$1.5$91.0M
GitHub Copilot$1.5$9$0.15200K
Google$1.5$9$0.151.0MSupported
Impossibl$1.5$9$0.151.0M
LLM Gateway$1.5$9$0.151.0M
Merge Gateway$1.5$9$0.151.0M
NanoGPT$1.5$9$0.151.0M
NEAR AI Cloud$1.5$9$0.151.0M
Neon$1.5$9$0.151.0M
Ofox$1.5$9$0.151.0M
OpenCode Zen$1.5$9$0.151.0M
OpenRouter$1.5$9$0.151.0MSupported
Opper$1.5$9$0.151.0M
OrcaRouter$1.5$9$0.151.0M
Pioneer$1.5$9$0.151.0M
SAP AI Core$1.5$9$0.151.0M
Vercel AI Gateway$1.5$9$0.151.0MSupported
Vertex$1.5$9$0.151.0MSupported
ZenMux$1.5$9$0.151.0M
Poe$1.5152$9.0909$0.15151.0M
Venice AI$1.55$9.45$0.1551.0M
Xpersona$1.55$12.2$0.1551.0M
Cortecs$1.649$9.899$0.1651.0M
Requesty$1.65$9.9$0.1651.0M

Gemini 3.5 Flash — Benchmark Scores Overview

Scores normalized to percentage scale for visual comparison. ELO scores mapped to 0-100 range (1100-1500).

Gemini 3.5 Flash — Frequently Asked Questions

How much does Gemini 3.5 Flash cost?

Gemini 3.5 Flash costs $1.5 per 1M input tokens and $9.0 per 1M output tokens. This is mid-range pricing for its capability level.

What is the context window of Gemini 3.5 Flash?

Gemini 3.5 Flash has a context window of 1.0M tokens. This determines how much text, conversation history, and code the model can process in a single request.

Who created Gemini 3.5 Flash?

Gemini 3.5 Flash was created by Google. It is classified as a mid model in our catalogue.

Is Gemini 3.5 Flash open source?

No, Gemini 3.5 Flash is a proprietary model. It is available through Google's API and compatible providers.