Gemini 3.1 Pro
Google's flagship in 2026 with a 2M context window; tiered pricing above 200K tokens. Context caching can cut cached reads to ~$0.20-$0.40/1M.
Gemini 3.1 Pro strengths
- Largest mainstream context (2M)
- Strong multimodal reasoning
- Native Google ecosystem integration
- Competitive pricing
Pricing & context
| Context window | 2M tokens |
| Input price /1M | $2.00 (under 200K; $4.00 above) |
| Output price /1M | $12.00 (under 200K; $18.00 above) |
| Modalities | text, image, audio, video |
Cost guide: a typical call of about 10K input + 2K output tokens costs roughly $0.044 at list prices. Worth modelling against cheaper tiers before committing high-volume traffic.
When to choose Gemini 3.1 Pro
Gemini 3.1 Pro is best for long-document and multimodal workloads, RAG over huge corpora, and Google Cloud-native apps. If your workload is more cost-sensitive, weigh it against gpt-oss-120b (≈$0.03 input /1M) first.
Gemini 3.1 Pro FAQ
How much does Gemini 3.1 Pro cost?
Gemini 3.1 Pro is priced at $2.00 (under 200K; $4.00 above) per 1M input tokens and $12.00 (under 200K; $18.00 above) per 1M output tokens (public API list price), with a 2M tokens context window. A typical call of about 10K input and 2K output tokens costs roughly $0.044.
What is Gemini 3.1 Pro best for?
Gemini 3.1 Pro by Google is best for long-document and multimodal workloads, RAG over huge corpora, and Google Cloud-native apps.
How does Gemini 3.1 Pro pricing compare to Grok 4.5?
Gemini 3.1 Pro input costs $2.00 (under 200K; $4.00 above) per 1M tokens versus $2.00 for Grok 4.5, roughly 1.0x more expensive on input. Output is $12.00 (under 200K; $18.00 above) vs $6.00.
Is Gemini 3.1 Pro multimodal?
Gemini 3.1 Pro supports text, image, audio, video.
Compare Gemini 3.1 Pro head-to-head
Tools that use Gemini 3.1 Pro
Other models
All models →| 01 | Claude Fable 5 | Anthropic | $10.00 | → |
| 02 | Claude Opus 5 | Anthropic | $5.00 | → |
| 03 | GPT-5.6 Sol | OpenAI | $5.00 | → |
| 04 | GPT-5.5 | OpenAI | $5.00 | → |