Gemini 3.5 Flash

Googletextimageaudiovideo

Launched at Google I/O 2026 (May 19). Fast multimodal model balancing speed and capability. Superseded July 21, 2026 by Gemini 3.6 Flash, which drops the output rate from $9.00 to $7.50 per 1M and uses roughly 17% fewer output tokens for the same task.

Gemini 3.5 Flash strengths

  • Fast multimodal processing
  • 1M context
  • Good value for agentic use
  • Native audio/video understanding

Pricing & context

Context window1M tokens
Input price /1M$1.50
Output price /1M$9.00
Modalitiestext, image, audio, video

Cost guide: a typical call of about 10K input + 2K output tokens costs roughly $0.033 at list prices. Worth modelling against cheaper tiers before committing high-volume traffic.

When to choose Gemini 3.5 Flash

Gemini 3.5 Flash is best for high-volume multimodal apps and agents needing speed with large context. If your workload is more cost-sensitive, weigh it against gpt-oss-120b (≈$0.03 input /1M) first.

Gemini 3.5 Flash FAQ

How much does Gemini 3.5 Flash cost?

Gemini 3.5 Flash is priced at $1.50 per 1M input tokens and $9.00 per 1M output tokens (public API list price), with a 1M tokens context window. A typical call of about 10K input and 2K output tokens costs roughly $0.033.

What is Gemini 3.5 Flash best for?

Gemini 3.5 Flash by Google is best for high-volume multimodal apps and agents needing speed with large context.

How does Gemini 3.5 Flash pricing compare to Sakana Fugu Ultra?

Gemini 3.5 Flash input costs $1.50 per 1M tokens versus $5.00 for Sakana Fugu Ultra, roughly 3.3x less expensive on input. Output is $9.00 vs $30.00.

Is Gemini 3.5 Flash multimodal?

Gemini 3.5 Flash supports text, image, audio, video.

Tools that use Gemini 3.5 Flash

Other models

All models →
01Claude Fable 5Anthropic$10.00
02Claude Opus 5Anthropic$5.00
03GPT-5.6 SolOpenAI$5.00
04GPT-5.5OpenAI$5.00