LIVE
Which tools do Claude, Codex and Cursor choose? We measured 17k runs to find out03/09/26 · Anthropic|OpenAI's GPT-6 Astra on ARC-AGI-303/09/26 · OpenAI|GPT-6 Astra03/09/26 · OpenAI|OpenAI begins rolling out GPT-6 Astra03/09/26 · OpenAI|ESPO: Error-Structured Prompt Optimization via Diagnose, Diversify, and Stabilize03/09/26 · Anthropic|Sparks Fly: NVIDIA Accelerates Local AI at IFA 202603/09/26 · NVIDIA|Introducing WeatherNext 3, our most advanced and accurate global weather AI model03/09/26 · Google|Porting my 1993 Amiga game to Godot, with an LLM reading the 68000 assembly03/09/26 · Anthropic|Claude outage – Resolved03/09/26 · Anthropic|Daybreak for Frontline Defenders: $1B to protect essential services03/09/26 · OpenAI|NeoMME: an efficient Multimodal-native and Multilingual Encoder03/09/26 · Hugging Face|‘NBA 2K27’ With NVIDIA DLSS 5 Leads 28 New Games Coming to GeForce NOW03/09/26 · NVIDIA|Which tools do Claude, Codex and Cursor choose? We measured 17k runs to find out03/09/26 · Anthropic|OpenAI's GPT-6 Astra on ARC-AGI-303/09/26 · OpenAI|GPT-6 Astra03/09/26 · OpenAI|OpenAI begins rolling out GPT-6 Astra03/09/26 · OpenAI|ESPO: Error-Structured Prompt Optimization via Diagnose, Diversify, and Stabilize03/09/26 · Anthropic|Sparks Fly: NVIDIA Accelerates Local AI at IFA 202603/09/26 · NVIDIA|Introducing WeatherNext 3, our most advanced and accurate global weather AI model03/09/26 · Google|Porting my 1993 Amiga game to Godot, with an LLM reading the 68000 assembly03/09/26 · Anthropic|Claude outage – Resolved03/09/26 · Anthropic|Daybreak for Frontline Defenders: $1B to protect essential services03/09/26 · OpenAI|NeoMME: an efficient Multimodal-native and Multilingual Encoder03/09/26 · Hugging Face|‘NBA 2K27’ With NVIDIA DLSS 5 Leads 28 New Games Coming to GeForce NOW03/09/26 · NVIDIA|
🇺🇸 United States🧠 GeneralistReleased April 2026

GPT-5.5

OpenAI
Free
Verdict

The default choice when quality matters more than budget.

Overview

GPT-5.5 is OpenAI's most capable model at launch. It understands a task earlier, plans, uses tools and checks its own work until completion. It targets agentic coding, computer use and knowledge work. 1M-token context. High output pricing, best reserved for cases where quality matters more than budget.

Skill profile

FrenchReasoningSpeedCreativitySafety
Arena Elo1482
Verified benchmarks
SWE-Bench82.6%

What do these scores mean?

SWE-Bench82.6%excellent

Resolving real GitHub issues under real conditions (SWE-bench Verified).

Strengths

  • Top-tier reasoning
  • Excellent at code
  • Strong French quality
  • World-class on Arena

Limitations

  • High output cost
  • No native GDPR guarantee

Who is it for

A good fit if…
  • you need advanced reasoning or analysis
  • you build with a coding agent
  • you work in French
Skip it if…
  • your budget is tight
  • you handle sensitive EU data

Ideal use cases

  • Agentic coding and long-horizon tasks
  • Research and document analysis
  • Multi-tool automation

Access & availability

Paid API per tokenFree tier available

Key specifications

Context
1M
Input price
$5.00 $/M
Output price
$30.00 $/M
Speed
Price verified Jun 1, 2026·Official sourceBenchmarks not audited

Estimate your monthly cost

Per-token API pricing
$363/ month with GPT-5.5

For the same usage

Indicative estimate based on standard API rates (excluding caching, batch and volume discounts). Always check the official pricing before committing.

Privacy

Not GDPR-compliant
Advanced data · for experts
Architecture, modalities, detailed cost, full benchmarks
Architecture
Size
Cutoff date
Nov 2025
Inputs
text, image
Outputs
text
License
commercial
Hosting
Hallucination score
Estimated cost (API)
Typical exchange (~3k in / 1k out)
≈ $0.0450
1M in + 1M out
≈ $35.00
All benchmarks
Arena Elo1482
MMLU
GPQA
HumanEval
SWE-Bench82.6%
MATH

Similar models