// google

Gemini 3 Flash Preview

google/gemini-3-flash-preview

reasoningtool callingimage inputfile inputaudio inputvideo inputprompt caching
OpenRouter + LiveBenchAll modelsFull leaderboard

Gemini 3 Flash Preview costs $0.500 per 1M input tokens and $3.00 per 1M output tokens with a 1.0M-token context window.

A reasoning, tool-calling model from google, released Dec 17, 2025. LiveBench hasn't published a run for it, so this page carries measured price and specs only — no benchmark numbers.

Input / 1M
$0.500
Output / 1M
$3.00
Context
1.0M
LiveBench overall
Not evaluated

Specs and pricing

Input price

USD per 1M prompt tokens

$0.500
Output price

USD per 1M completion tokens, reasoning included

$3.00
Blended price

3:1 input:output — #124 most expensive of 343 listed models

$1.13
Cached input read

Blank means not priced separately — not free

$0.050
Cache write$0.083
Context window1,048,576 tokens
Max output65,536 tokens
Input modalitiestext, image, file, audio, video
Tool callingYes
Extended reasoningYes
Knowledge cutoff
ListedDec 17, 2025
All Google API pricing

Also listed as google/gemini-3-flash-preview:batch ($0.563 blended)— the same model at a different price or rate limit.

What Gemini 3 Flash Preview costs to run

Monthly list cost across five workload shapes, using the published cached-input rate where there is one. A floor, not a quote — batch discounts and cache writes aren't included.

WorkloadPer month
Support chatbot

1.2K in / 400 out × 200K requests

$327.60/mo
RAG assistant

8K in / 600 out × 100K requests

$400.00/mo
Coding agent

40K in / 4K out × 20K requests

$388.00/mo
Document extraction

20K in / 1.5K out × 50K requests

$702.50/mo
Bulk classification

500 in / 20 out × 5M requests

$1,325/mo
Price your own workload

Compare Gemini 3 Flash Preview with

Benchmarked models at a similar price.

Gemini 3 Flash Preview FAQ

How much does Gemini 3 Flash Preview cost?

Gemini 3 Flash Preview lists at $0.500 per 1M input tokens and $3.00 per 1M output tokens, with cached input reads at $0.050 per 1M. A RAG assistant handling 100K requests a month (8K tokens in, 600 out) comes to about $400.00 at list price.

What is the context window of Gemini 3 Flash Preview?

Gemini 3 Flash Preview accepts up to 1,048,576 tokens of context and can return up to 65,536 output tokens per response.

Does Gemini 3 Flash Preview support tool calling?

Yes — Gemini 3 Flash Preview supports tool (function) calling and exposes extended reasoning. It accepts text, image, file, audio, video input.

How does Gemini 3 Flash Preview score on benchmarks?

LiveBench has not published a run for Gemini 3 Flash Preview yet. Until it does, compare it on price and specs, or run your own evaluation on your own tasks.

What is the API model ID for Gemini 3 Flash Preview?

On OpenRouter the model ID is "google/gemini-3-flash-preview". It was listed on Dec 17, 2025.

More from google

How these numbers are produced

  • Price and specs — provider list data from OpenRouter, refreshed every 15 minutes. “Blended” is a 3:1 input:output mix.
  • Scores and ranksLiveBench release 2026-06-25. Ranks count only models that have a published run; a blank means “not evaluated”, never “bad”.
  • Cost per point — the measured dollars LiveBench spent on the run, divided by the score it earned.

Published benchmarks rank models on someone else's tasks. Before committing, see LLM & agent evaluation for building an eval on your own.