LIVE
Introducing Cosmos 3 Edge20/07/26 · Hugging Face|Kimi K3, Qwen 3.8, and Anthropic's (Potential) Unravelling20/07/26 · Anthropic|At SIGGRAPH, NVIDIA Advances Graphics and Simulation With Agentic and Physical AI20/07/26 · NVIDIA|China's open-weights AI strategy is winning20/07/26|Bristol Myers Squibb Building Life Science Industry’s Most Advanced AI Factory on NVIDIA Vera Rubin20/07/26 · NVIDIA|Safety and alignment in an era of long-horizon models20/07/26 · OpenAI|Claude Fable produced a counterexample to the Jacobian Conjecture20/07/26 · Anthropic|Apply for Anthropic’s AI for Science rare disease research grants20/07/26 · Anthropic|AI advice made people less accurate but more confident – sudy19/07/26|Vincentwei1021/video-shotcraft: AI video skill for Claude Code & Codex — cinematic product videos with Remotion: 106 shot recipe cards, 161 motion previews, a production-ready template19/07/26 · Anthropic|Claude Code uses Bun written in Rust now19/07/26 · Anthropic|OpenAI reduces Codex Model Context Size from 372k to 272k19/07/26 · OpenAI|Introducing Cosmos 3 Edge20/07/26 · Hugging Face|Kimi K3, Qwen 3.8, and Anthropic's (Potential) Unravelling20/07/26 · Anthropic|At SIGGRAPH, NVIDIA Advances Graphics and Simulation With Agentic and Physical AI20/07/26 · NVIDIA|China's open-weights AI strategy is winning20/07/26|Bristol Myers Squibb Building Life Science Industry’s Most Advanced AI Factory on NVIDIA Vera Rubin20/07/26 · NVIDIA|Safety and alignment in an era of long-horizon models20/07/26 · OpenAI|Claude Fable produced a counterexample to the Jacobian Conjecture20/07/26 · Anthropic|Apply for Anthropic’s AI for Science rare disease research grants20/07/26 · Anthropic|AI advice made people less accurate but more confident – sudy19/07/26|Vincentwei1021/video-shotcraft: AI video skill for Claude Code & Codex — cinematic product videos with Remotion: 106 shot recipe cards, 161 motion previews, a production-ready template19/07/26 · Anthropic|Claude Code uses Bun written in Rust now19/07/26 · Anthropic|OpenAI reduces Codex Model Context Size from 372k to 272k19/07/26 · OpenAI|
🇺🇸 United States🧠 GeneralistReleased February 2026

Gemini 3.1 Pro

Google
Free
Verdict

A versatile model, balanced across most use cases.

Overview

Gemini 3.1 Pro is Google DeepMind's flagship, built on a Mixture-of-Experts architecture. It leads most reasoning benchmarks (94.3% GPQA Diamond) while keeping Gemini 3 Pro's pricing ($2/$12 per million). Natively multimodal (text, image, audio, video) with a 1M-token context, it's one of the best cost/performance picks at the frontier.

Skill profile

FrenchReasoningSpeedCreativitySafety
Verified benchmarks
GPQA94.3%
SWE-Bench80.6%
MATH95.1%

What do these scores mean?

GPQA94.3%world-class

PhD-level science questions (physics, chemistry, biology), with no tool access.

SWE-Bench80.6%excellent

Resolving real GitHub issues under real conditions (SWE-bench Verified).

MATH95.1%world-class

Competition-level math problems.

Strengths

  • Top-tier reasoning
  • Excellent at code
  • Strong French quality
  • Very long context

Limitations

  • No native GDPR guarantee

Who is it for

A good fit if…
  • you need advanced reasoning or analysis
  • you build with a coding agent
  • you work in French
Skip it if…
  • you handle sensitive EU data

Ideal use cases

  • Complex reasoning and research
  • Multimodal analysis (image, audio, video)
  • Very long context (codebases, large corpora)

Access & availability

Paid API per tokenFree tier available

Key specifications

Context
1M
Input price
$2.00 $/M
Output price
$12.00 $/M
Speed
144 tok/s
Price verified Jun 1, 2026·Official sourceBenchmarks not audited

Estimate your monthly cost

Per-token API pricing
$145/ month with Gemini 3.1 Pro

For the same usage

Indicative estimate based on standard API rates (excluding caching, batch and volume discounts). Always check the official pricing before committing.

Privacy

Not GDPR-compliant
Advanced data · for experts
Architecture, modalities, detailed cost, full benchmarks
Architecture
Mixture-of-Experts
Size
Cutoff date
Dec 2024
Inputs
text, image, audio, video
Outputs
text
License
commercial
Hosting
Hallucination score
Estimated cost (API)
Typical exchange (~3k in / 1k out)
≈ $0.0180
1M in + 1M out
≈ $14.00
All benchmarks
Arena Elo
MMLU
GPQA94.3%
HumanEval
SWE-Bench80.6%
MATH95.1%

Similar models