LIVE
Corporate America increasingly turns to open-source AI04/09/26|Study finds Google's AI Mode surfaces pricier products than classic search04/09/26 · Google|Study measures how coding agents pick third-party tools03/09/26|OpenAI's GPT-6 Astra sets new state-of-the-art on ARC-AGI-303/09/26 · OpenAI|ESPO: A Prompt Optimization Method That Beats GEPA With Shorter, More Stable Prompts03/09/26|"Last Translation Benchmark": a large-scale collaborative effort for a definitive MT benchmark03/09/26|NVIDIA pushes local AI at IFA 2026 with RTX Spark PCs and a personal inference router03/09/26 · NVIDIA|Google DeepMind unveils WeatherNext 3, its most accurate weather AI model yet03/09/26 · Google DeepMind|OpenAI Launches $1B Daybreak Program for Frontline Cybersecurity Defenders03/09/26 · OpenAI|Hcompany releases NeoMME, a compact multimodal-native encoder for retrieval03/09/26 · Hcompany|Playco cuts manual fixes by 50% using GPT-6 Astra for prototyping03/09/26 · OpenAI|Legora reviews 41 financial documents in minutes with GPT-6 Astra03/09/26 · OpenAI|Corporate America increasingly turns to open-source AI04/09/26|Study finds Google's AI Mode surfaces pricier products than classic search04/09/26 · Google|Study measures how coding agents pick third-party tools03/09/26|OpenAI's GPT-6 Astra sets new state-of-the-art on ARC-AGI-303/09/26 · OpenAI|ESPO: A Prompt Optimization Method That Beats GEPA With Shorter, More Stable Prompts03/09/26|"Last Translation Benchmark": a large-scale collaborative effort for a definitive MT benchmark03/09/26|NVIDIA pushes local AI at IFA 2026 with RTX Spark PCs and a personal inference router03/09/26 · NVIDIA|Google DeepMind unveils WeatherNext 3, its most accurate weather AI model yet03/09/26 · Google DeepMind|OpenAI Launches $1B Daybreak Program for Frontline Cybersecurity Defenders03/09/26 · OpenAI|Hcompany releases NeoMME, a compact multimodal-native encoder for retrieval03/09/26 · Hcompany|Playco cuts manual fixes by 50% using GPT-6 Astra for prototyping03/09/26 · OpenAI|Legora reviews 41 financial documents in minutes with GPT-6 Astra03/09/26 · OpenAI|
🇨🇳 China💻 CodeReleased July 2025

Qwen3-Coder 480B

Alibaba
Open-sourceSelf-hostRGPD
Verdict

A solid pick for self-hosting and full data control.

Overview

Qwen3-Coder-480B-A35B is Alibaba's most ambitious coding model yet: a 480-billion-parameter MoE (35B active) built to reason over entire repositories, with a native 256K-token context extendable to 1M via Yarn. It targets agentic coding and tool use, with results the vendor claims rival Claude Sonnet — a comparison that still awaits independent verification. One limit: it ships without an explicit reasoning or thinking mode. Released under Apache 2.0, it can be self-hosted on European servers to keep data local, though its scale demands a serious multi-GPU cluster.

Skill profile

Arena Elo1387

Not disclosed

Strengths

  • World-class on Arena
  • Open-source and self-hostable

Limitations

  • API pricing not disclosed

Who is it for

A good fit if…
  • you build with a coding agent
  • you want to control cost or self-host
  • you have GDPR constraints
Skip it if…

    Ideal use cases

    • Autonomous coding agents
    • Whole-repo refactoring
    • Dev tool automation
    • Large-scale code review
    • Sovereign on-prem deployment

    Access & availability

    Self-hostable (open weights)

    Key specifications

    Context
    256K
    Input price
    Output price
    Speed
    Price not auditedBenchmarks not audited

    Privacy

    GDPR-compliant

    Run it locally

    Dedicated GPU server32k1k
    QuantizationDiskRAM / VRAMTypical hardware
    Q4 · recommended273.6 GB303 GBServer GPU infra (H100/A100…)
    Q8 · balanced513.6 GB567 GBServer GPU infra (H100/A100…)
    FP16 · max quality960 GB1058 GBServer GPU infra (H100/A100…)

    Estimates for a moderate context. Long contexts need more RAM (KV cache).

    🤗 Hugging Face
    ollama run qwen3-coder:480b

    Deploy

    Copy-ready commands generated from this card. Adjust context length and GPU count to your hardware.

    Estimated memory · FP16 1058 Go · Q4 303 Go

    Easiest way to try it on a workstation. Install Ollama, then:

    ollama run qwen3-coder:480b
    Qwen/Qwen3-Coder-480B-A35B-Instruct
    Advanced data · for experts
    Architecture, modalities, detailed cost, full benchmarks
    Architecture
    MoE
    Size
    480B (MoE, 35B actifs)
    Cutoff date
    Inputs
    text
    Outputs
    text
    License
    apache-2.0
    Hosting
    Hallucination score
    Estimated cost (API)
    Typical exchange (~3k in / 1k out)
    Not disclosed
    1M in + 1M out
    Not disclosed
    All benchmarks
    Arena Elo1387
    MMLU
    GPQA
    HumanEval
    SWE-Bench
    MATH

    Similar models