LIVE
Introducing Cosmos 3 Edge20/07/26 · Hugging Face|Kimi K3, Qwen 3.8, and Anthropic's (Potential) Unravelling20/07/26 · Anthropic|At SIGGRAPH, NVIDIA Advances Graphics and Simulation With Agentic and Physical AI20/07/26 · NVIDIA|China's open-weights AI strategy is winning20/07/26|Bristol Myers Squibb Building Life Science Industry’s Most Advanced AI Factory on NVIDIA Vera Rubin20/07/26 · NVIDIA|Safety and alignment in an era of long-horizon models20/07/26 · OpenAI|Claude Fable produced a counterexample to the Jacobian Conjecture20/07/26 · Anthropic|Apply for Anthropic’s AI for Science rare disease research grants20/07/26 · Anthropic|AI advice made people less accurate but more confident – sudy19/07/26|Vincentwei1021/video-shotcraft: AI video skill for Claude Code & Codex — cinematic product videos with Remotion: 106 shot recipe cards, 161 motion previews, a production-ready template19/07/26 · Anthropic|Claude Code uses Bun written in Rust now19/07/26 · Anthropic|OpenAI reduces Codex Model Context Size from 372k to 272k19/07/26 · OpenAI|Introducing Cosmos 3 Edge20/07/26 · Hugging Face|Kimi K3, Qwen 3.8, and Anthropic's (Potential) Unravelling20/07/26 · Anthropic|At SIGGRAPH, NVIDIA Advances Graphics and Simulation With Agentic and Physical AI20/07/26 · NVIDIA|China's open-weights AI strategy is winning20/07/26|Bristol Myers Squibb Building Life Science Industry’s Most Advanced AI Factory on NVIDIA Vera Rubin20/07/26 · NVIDIA|Safety and alignment in an era of long-horizon models20/07/26 · OpenAI|Claude Fable produced a counterexample to the Jacobian Conjecture20/07/26 · Anthropic|Apply for Anthropic’s AI for Science rare disease research grants20/07/26 · Anthropic|AI advice made people less accurate but more confident – sudy19/07/26|Vincentwei1021/video-shotcraft: AI video skill for Claude Code & Codex — cinematic product videos with Remotion: 106 shot recipe cards, 161 motion previews, a production-ready template19/07/26 · Anthropic|Claude Code uses Bun written in Rust now19/07/26 · Anthropic|OpenAI reduces Codex Model Context Size from 372k to 272k19/07/26 · OpenAI|
🇨🇳 China🧠 GeneralistReleased December 2024

DeepSeek V3

DeepSeek
Open-sourceFreeSelf-host
Verdict

A solid pick for self-hosting and full data control.

Overview

DeepSeek V3 is a revolution: 671B parameters in MoE, performance close to GPT-4o, but 30x lower price and MIT license (commercial use, self-hostable). Limits: French performance slightly behind, and for European enterprise, the official API is in China (prefer self-hosted deployment).

Skill profile

FrenchReasoningSpeedCreativitySafety
Arena Elo1340
Verified benchmarks
GPQA59%
SWE-Bench42%
MATH90.2%
MMLU87.5%
HumanEval89%

What do these scores mean?

GPQA59%solid

PhD-level science questions (physics, chemistry, biology), with no tool access.

SWE-Bench42%fair

Resolving real GitHub issues under real conditions (SWE-bench Verified).

MATH90.2%world-class

Competition-level math problems.

MMLU87.5%excellent

General knowledge across dozens of academic subjects.

HumanEval89%excellent

Python code generation from specifications.

Strengths

  • Excellent at code
  • Has a free tier
  • Open-source and self-hostable
  • Competitive input pricing

Limitations

  • Average speed
  • No native GDPR guarantee

Who is it for

A good fit if…
  • you want to control cost or self-host
Skip it if…
  • you handle sensitive EU data
  • you need very fast responses

Ideal use cases

  • Ultra-low-cost coding
  • Self-hosting
  • Academic research
  • Budget-tight apps

Access & availability

Paid API per tokenFree tier availableSelf-hostable (open weights)

Key specifications

Context
128K
Input price
$0.27 $/M
Output price
$1.10 $/M
Speed
60 tok/s
Price not auditedBenchmarks not audited

Estimate your monthly cost

Per-token API pricing
$16/ month with DeepSeek V3

For the same usage

Indicative estimate based on standard API rates (excluding caching, batch and volume discounts). Always check the official pricing before committing.

Privacy

Not GDPR-compliantHosting : CN/selfAnonymizable data

Run it locally

Dedicated GPU server926k3k
QuantizationDiskRAM / VRAMTypical hardware
Q4 · recommended390.4 GB431 GBServer GPU infra (H100/A100…)
Q8 · balanced733 GB808 GBServer GPU infra (H100/A100…)
FP16 · max quality1370 GB1509 GBServer GPU infra (H100/A100…)

Estimates for a moderate context. Long contexts need more RAM (KV cache).

🤗 Hugging Face
ollama run deepseek-v3
Advanced data · for experts
Architecture, modalities, detailed cost, full benchmarks
Architecture
MoE
Size
671B (37B actifs MoE)
Cutoff date
Jun 2024
Inputs
text
Outputs
text, code
License
mit
Hosting
CN/self
Hallucination score
3/5
Estimated cost (API)
Typical exchange (~3k in / 1k out)
≈ $0.0019
1M in + 1M out
≈ $1.37
All benchmarks
Arena Elo1340
MMLU87.5%
GPQA59.0%
HumanEval89.0%
SWE-Bench42.0%
MATH90.2%

Similar models