LIVE
Introducing Cosmos 3 Edge20/07/26 · Hugging Face|Kimi K3, Qwen 3.8, and Anthropic's (Potential) Unravelling20/07/26 · Anthropic|At SIGGRAPH, NVIDIA Advances Graphics and Simulation With Agentic and Physical AI20/07/26 · NVIDIA|China's open-weights AI strategy is winning20/07/26|Bristol Myers Squibb Building Life Science Industry’s Most Advanced AI Factory on NVIDIA Vera Rubin20/07/26 · NVIDIA|Safety and alignment in an era of long-horizon models20/07/26 · OpenAI|Claude Fable produced a counterexample to the Jacobian Conjecture20/07/26 · Anthropic|Apply for Anthropic’s AI for Science rare disease research grants20/07/26 · Anthropic|AI advice made people less accurate but more confident – sudy19/07/26|Vincentwei1021/video-shotcraft: AI video skill for Claude Code & Codex — cinematic product videos with Remotion: 106 shot recipe cards, 161 motion previews, a production-ready template19/07/26 · Anthropic|Claude Code uses Bun written in Rust now19/07/26 · Anthropic|OpenAI reduces Codex Model Context Size from 372k to 272k19/07/26 · OpenAI|Introducing Cosmos 3 Edge20/07/26 · Hugging Face|Kimi K3, Qwen 3.8, and Anthropic's (Potential) Unravelling20/07/26 · Anthropic|At SIGGRAPH, NVIDIA Advances Graphics and Simulation With Agentic and Physical AI20/07/26 · NVIDIA|China's open-weights AI strategy is winning20/07/26|Bristol Myers Squibb Building Life Science Industry’s Most Advanced AI Factory on NVIDIA Vera Rubin20/07/26 · NVIDIA|Safety and alignment in an era of long-horizon models20/07/26 · OpenAI|Claude Fable produced a counterexample to the Jacobian Conjecture20/07/26 · Anthropic|Apply for Anthropic’s AI for Science rare disease research grants20/07/26 · Anthropic|AI advice made people less accurate but more confident – sudy19/07/26|Vincentwei1021/video-shotcraft: AI video skill for Claude Code & Codex — cinematic product videos with Remotion: 106 shot recipe cards, 161 motion previews, a production-ready template19/07/26 · Anthropic|Claude Code uses Bun written in Rust now19/07/26 · Anthropic|OpenAI reduces Codex Model Context Size from 372k to 272k19/07/26 · OpenAI|
🇺🇸 United States🧠 GeneralistReleased October 2025

Claude Sonnet 4.6

Anthropic
FreeRGPD
Verdict

A good performance/compliance balance for a European company.

Overview

Claude Sonnet 4.6 is the most-used model in the lineup: it delivers 90% of Opus capabilities at 5x lower price and higher speed. It's the default pick for most use cases: assistance, coding, content, support. French quality is excellent.

Skill profile

FrenchReasoningSpeedCreativitySafety
Arena Elo1352
Verified benchmarks
GPQA65.4%
SWE-Bench65%
MATH82.1%
MMLU86.2%
HumanEval91.5%

What do these scores mean?

GPQA65.4%very good

PhD-level science questions (physics, chemistry, biology), with no tool access.

SWE-Bench65%very good

Resolving real GitHub issues under real conditions (SWE-bench Verified).

MATH82.1%excellent

Competition-level math problems.

MMLU86.2%excellent

General knowledge across dozens of academic subjects.

HumanEval91.5%world-class

Python code generation from specifications.

Strengths

  • Excellent at code
  • Strong French quality
  • World-class on Arena
  • Has a free tier

Who is it for

A good fit if…
  • you build with a coding agent
  • you work in French
  • you have GDPR constraints
Skip it if…

    Ideal use cases

    • Daily assistant
    • Pair-programming
    • Summarization
    • Content creation
    • Support chatbots

    Access & availability

    Paid API per tokenFree tier availableConsumer subscription (~20€/mo)

    Key specifications

    Context
    200K
    Input price
    $3.00 $/M
    Output price
    $15.00 $/M
    Speed
    110 tok/s
    Price not auditedBenchmarks not audited

    Estimate your monthly cost

    Per-token API pricing
    $196/ month with Claude Sonnet 4.6

    For the same usage

    Indicative estimate based on standard API rates (excluding caching, batch and volume discounts). Always check the official pricing before committing.

    Privacy

    GDPR-compliantHosting : US/EUAnonymizable data
    Advanced data · for experts
    Architecture, modalities, detailed cost, full benchmarks
    Architecture
    Dense
    Size
    undisclosed
    Cutoff date
    Mar 2025
    Inputs
    text, image, document
    Outputs
    text, code
    License
    commercial
    Hosting
    US/EU
    Hallucination score
    5/5
    Estimated cost (API)
    Typical exchange (~3k in / 1k out)
    ≈ $0.0240
    1M in + 1M out
    ≈ $18.00
    All benchmarks
    Arena Elo1352
    MMLU86.2%
    GPQA65.4%
    HumanEval91.5%
    SWE-Bench65.0%
    MATH82.1%

    Similar models