Gemini 3.1 Pro is the value leader of the closed frontier: roughly $2 per 1M input tokens, a 1M-token context window, and reasoning that sits near the very top (GPQA Diamond ≈ 94.3%). It reads text, images, video and audio natively, which makes it the default for analysis tasks that mix media.
For high-volume work — summarizing documents, processing long transcripts, classifying images at scale — its price-to-capability ratio is hard to beat among proprietary models. The faster, cheaper Gemini 3.5 Flash ($1.50/$9) covers latency-sensitive jobs.
Where it gives ground is the hardest agentic coding (SWE-bench Pro), where Opus 4.8 and GPT-5.5 still lead. Output pricing is also reported inconsistently across sources, so confirm current rates before budgeting.
Best for
- High-volume reasoning at low cost
- Document, image and video analysis
- Cost-sensitive multimodal workloads
Pros & cons
Strengths
- Best value of the closed frontier (~$2/1M input, 1M context)
- Top-tier reasoning (GPQA Diamond ≈ 94.3%)
- Native text, image, video and audio understanding
Limitations
- Coding trails Opus 4.8 / GPT-5.5 on the hardest SWE-bench Pro tasks
- Output pricing reports vary by source