LIVE
Which tools do Claude, Codex and Cursor choose? We measured 17k runs to find out03/09/26 · Anthropic|OpenAI's GPT-6 Astra on ARC-AGI-303/09/26 · OpenAI|GPT-6 Astra03/09/26 · OpenAI|OpenAI begins rolling out GPT-6 Astra03/09/26 · OpenAI|ESPO: Error-Structured Prompt Optimization via Diagnose, Diversify, and Stabilize03/09/26 · Anthropic|Sparks Fly: NVIDIA Accelerates Local AI at IFA 202603/09/26 · NVIDIA|Introducing WeatherNext 3, our most advanced and accurate global weather AI model03/09/26 · Google|Porting my 1993 Amiga game to Godot, with an LLM reading the 68000 assembly03/09/26 · Anthropic|Claude outage – Resolved03/09/26 · Anthropic|Daybreak for Frontline Defenders: $1B to protect essential services03/09/26 · OpenAI|NeoMME: an efficient Multimodal-native and Multilingual Encoder03/09/26 · Hugging Face|‘NBA 2K27’ With NVIDIA DLSS 5 Leads 28 New Games Coming to GeForce NOW03/09/26 · NVIDIA|Which tools do Claude, Codex and Cursor choose? We measured 17k runs to find out03/09/26 · Anthropic|OpenAI's GPT-6 Astra on ARC-AGI-303/09/26 · OpenAI|GPT-6 Astra03/09/26 · OpenAI|OpenAI begins rolling out GPT-6 Astra03/09/26 · OpenAI|ESPO: Error-Structured Prompt Optimization via Diagnose, Diversify, and Stabilize03/09/26 · Anthropic|Sparks Fly: NVIDIA Accelerates Local AI at IFA 202603/09/26 · NVIDIA|Introducing WeatherNext 3, our most advanced and accurate global weather AI model03/09/26 · Google|Porting my 1993 Amiga game to Godot, with an LLM reading the 68000 assembly03/09/26 · Anthropic|Claude outage – Resolved03/09/26 · Anthropic|Daybreak for Frontline Defenders: $1B to protect essential services03/09/26 · OpenAI|NeoMME: an efficient Multimodal-native and Multilingual Encoder03/09/26 · Hugging Face|‘NBA 2K27’ With NVIDIA DLSS 5 Leads 28 New Games Coming to GeForce NOW03/09/26 · NVIDIA|
🇺🇸 United States🔬 ReasoningReleased December 2025

Claude Opus 4.7

Anthropic
FreeRGPD
Verdict

The default choice when quality matters more than budget.

Overview

Claude Opus 4.7 excels particularly on tasks requiring deep reasoning, agentic coding, and complex document analysis. Its 500K-token context window allows it to ingest entire codebases or legal cases. The model is known for reliability (low hallucination rate), French quality, and intellectual honesty (it would rather say it doesn't know than fabricate).

Skill profile

FrenchReasoningSpeedCreativitySafety
Arena Elo1502
Verified benchmarks
GPQA71.8%
SWE-Bench72.4%
MATH88.6%
MMLU89.5%
HumanEval94.2%

What do these scores mean?

GPQA71.8%very good

PhD-level science questions (physics, chemistry, biology), with no tool access.

SWE-Bench72.4%very good

Resolving real GitHub issues under real conditions (SWE-bench Verified).

MATH88.6%excellent

Competition-level math problems.

MMLU89.5%excellent

General knowledge across dozens of academic subjects.

HumanEval94.2%world-class

Python code generation from specifications.

Strengths

  • Top-tier reasoning
  • Excellent at code
  • Strong French quality
  • World-class on Arena

Limitations

  • High output cost
  • Average speed

Who is it for

A good fit if…
  • you need advanced reasoning or analysis
  • you build with a coding agent
  • you work in French
  • you have GDPR constraints
Skip it if…
  • your budget is tight
  • you need very fast responses

Ideal use cases

  • Long document analysis
  • Agentic coding
  • Complex research
  • Premium writing
  • Strategic tasks

Access & availability

Paid API per tokenFree tier availableConsumer subscription (~20€/mo)

Key specifications

Context
500K
Input price
$5.00 $/M
Output price
$25.00 $/M
Speed
75 tok/s
Price verified Jun 1, 2026·Official sourceBenchmarks not audited

Estimate your monthly cost

Per-token API pricing
$327/ month with Claude Opus 4.7

For the same usage

Indicative estimate based on standard API rates (excluding caching, batch and volume discounts). Always check the official pricing before committing.

Privacy

GDPR-compliantHosting : US/EUAnonymizable data
Advanced data · for experts
Architecture, modalities, detailed cost, full benchmarks
Architecture
Dense
Size
undisclosed
Cutoff date
Apr 2025
Inputs
text, image, document
Outputs
text, code
License
commercial
Hosting
US/EU
Hallucination score
5/5
Estimated cost (API)
Typical exchange (~3k in / 1k out)
≈ $0.0400
1M in + 1M out
≈ $30.00
All benchmarks
Arena Elo1502
MMLU89.5%
GPQA71.8%
HumanEval94.2%
SWE-Bench72.4%
MATH88.6%

Similar models