Gemini 3.1 Pro

Googletextimageaudiovideo

Google's flagship in 2026 with a 2M context window; tiered pricing above 200K tokens. Context caching can cut cached reads to ~$0.20-$0.40/1M.

Gemini 3.1 Pro strengths

  • Largest mainstream context (2M)
  • Strong multimodal reasoning
  • Native Google ecosystem integration
  • Competitive pricing

Pricing & context

Context window2M tokens
Input price /1M$2.00 (under 200K; $4.00 above)
Output price /1M$12.00 (under 200K; $18.00 above)
Modalitiestext, image, audio, video

Cost guide: a typical call of about 10K input + 2K output tokens costs roughly $0.044 at list prices. Worth modelling against cheaper tiers before committing high-volume traffic.

When to choose Gemini 3.1 Pro

Gemini 3.1 Pro is best for long-document and multimodal workloads, RAG over huge corpora, and Google Cloud-native apps. If your workload is more cost-sensitive, weigh it against gpt-oss-120b (≈$0.03 input /1M) first.

Gemini 3.1 Pro FAQ

How much does Gemini 3.1 Pro cost?

Gemini 3.1 Pro is priced at $2.00 (under 200K; $4.00 above) per 1M input tokens and $12.00 (under 200K; $18.00 above) per 1M output tokens (public API list price), with a 2M tokens context window. A typical call of about 10K input and 2K output tokens costs roughly $0.044.

What is Gemini 3.1 Pro best for?

Gemini 3.1 Pro by Google is best for long-document and multimodal workloads, RAG over huge corpora, and Google Cloud-native apps.

How does Gemini 3.1 Pro pricing compare to Grok 4.5?

Gemini 3.1 Pro input costs $2.00 (under 200K; $4.00 above) per 1M tokens versus $2.00 for Grok 4.5, roughly 1.0x more expensive on input. Output is $12.00 (under 200K; $18.00 above) vs $6.00.

Is Gemini 3.1 Pro multimodal?

Gemini 3.1 Pro supports text, image, audio, video.

Compare Gemini 3.1 Pro head-to-head

Tools that use Gemini 3.1 Pro

Other models

All models →
01Claude Fable 5Anthropic$10.00
02Claude Opus 5Anthropic$5.00
03GPT-5.6 SolOpenAI$5.00
04GPT-5.5OpenAI$5.00