Gemini 3 Flash

Googletextimageaudiovideo

Lower-cost Flash tier; one of the cheapest 1M-context multimodal options from a major lab.

Gemini 3 Flash strengths

  • Very low cost
  • 1M context
  • Multimodal
  • Fast latency

Pricing & context

Context window1M tokens
Input price /1M$0.50
Output price /1M$3.00
Modalitiestext, image, audio, video

Cost guide: a typical call of about 10K input + 2K output tokens costs roughly $0.011 at list prices. Worth modelling against cheaper tiers before committing high-volume traffic.

When to choose Gemini 3 Flash

Gemini 3 Flash is best for cost-sensitive multimodal and long-context tasks at scale. If your workload is more cost-sensitive, weigh it against gpt-oss-120b (≈$0.03 input /1M) first.

Gemini 3 Flash FAQ

How much does Gemini 3 Flash cost?

Gemini 3 Flash is priced at $0.50 per 1M input tokens and $3.00 per 1M output tokens (public API list price), with a 1M tokens context window. A typical call of about 10K input and 2K output tokens costs roughly $0.011.

What is Gemini 3 Flash best for?

Gemini 3 Flash by Google is best for cost-sensitive multimodal and long-context tasks at scale.

How does Gemini 3 Flash pricing compare to GPT-5.4 mini?

Gemini 3 Flash input costs $0.50 per 1M tokens versus $0.75 for GPT-5.4 mini, roughly 1.5x less expensive on input. Output is $3.00 vs $4.50.

Is Gemini 3 Flash multimodal?

Gemini 3 Flash supports text, image, audio, video.

Other models

All models →
01Claude Fable 5Anthropic$10.00
02Claude Opus 5Anthropic$5.00
03GPT-5.6 SolOpenAI$5.00
04GPT-5.5OpenAI$5.00