LIVE
Introducing Cosmos 3 Edge20/07/26 · Hugging Face|Kimi K3, Qwen 3.8, and Anthropic's (Potential) Unravelling20/07/26 · Anthropic|At SIGGRAPH, NVIDIA Advances Graphics and Simulation With Agentic and Physical AI20/07/26 · NVIDIA|China's open-weights AI strategy is winning20/07/26|Bristol Myers Squibb Building Life Science Industry’s Most Advanced AI Factory on NVIDIA Vera Rubin20/07/26 · NVIDIA|Safety and alignment in an era of long-horizon models20/07/26 · OpenAI|Claude Fable produced a counterexample to the Jacobian Conjecture20/07/26 · Anthropic|Apply for Anthropic’s AI for Science rare disease research grants20/07/26 · Anthropic|AI advice made people less accurate but more confident – sudy19/07/26|Vincentwei1021/video-shotcraft: AI video skill for Claude Code & Codex — cinematic product videos with Remotion: 106 shot recipe cards, 161 motion previews, a production-ready template19/07/26 · Anthropic|Claude Code uses Bun written in Rust now19/07/26 · Anthropic|OpenAI reduces Codex Model Context Size from 372k to 272k19/07/26 · OpenAI|Introducing Cosmos 3 Edge20/07/26 · Hugging Face|Kimi K3, Qwen 3.8, and Anthropic's (Potential) Unravelling20/07/26 · Anthropic|At SIGGRAPH, NVIDIA Advances Graphics and Simulation With Agentic and Physical AI20/07/26 · NVIDIA|China's open-weights AI strategy is winning20/07/26|Bristol Myers Squibb Building Life Science Industry’s Most Advanced AI Factory on NVIDIA Vera Rubin20/07/26 · NVIDIA|Safety and alignment in an era of long-horizon models20/07/26 · OpenAI|Claude Fable produced a counterexample to the Jacobian Conjecture20/07/26 · Anthropic|Apply for Anthropic’s AI for Science rare disease research grants20/07/26 · Anthropic|AI advice made people less accurate but more confident – sudy19/07/26|Vincentwei1021/video-shotcraft: AI video skill for Claude Code & Codex — cinematic product videos with Remotion: 106 shot recipe cards, 161 motion previews, a production-ready template19/07/26 · Anthropic|Claude Code uses Bun written in Rust now19/07/26 · Anthropic|OpenAI reduces Codex Model Context Size from 372k to 272k19/07/26 · OpenAI|
🇨🇳 China🔬 ReasoningReleased January 2025

DeepSeek-R1

DeepSeek
Open-sourceSelf-host
Verdict

A solid pick for self-hosting and full data control.

Skill profile

Not disclosed

Strengths

  • Open-source and self-hostable

Limitations

  • No native GDPR guarantee
  • API pricing not disclosed

Who is it for

A good fit if…
  • you want to control cost or self-host
Skip it if…
  • you handle sensitive EU data

Access & availability

Self-hostable (open weights)

Key specifications

Context
128K
Input price
Output price
Speed
Price not auditedBenchmarks not audited

Privacy

Not GDPR-compliant

Run it locally

Dedicated GPU server8.6M13k
QuantizationDiskRAM / VRAMTypical hardware
Q4 · recommended382.5 GB423 GBServer GPU infra (H100/A100…)
Q8 · balanced718 GB792 GBServer GPU infra (H100/A100…)
FP16 · max quality1342 GB1478 GBServer GPU infra (H100/A100…)

Estimates for a moderate context. Long contexts need more RAM (KV cache).

🤗 Hugging Face
ollama run deepseek-r1:671b
Advanced data · for experts
Architecture, modalities, detailed cost, full benchmarks
Architecture
Size
671B (MoE, 37B actifs)
Cutoff date
Inputs
text
Outputs
text
License
mit
Hosting
Hallucination score
Estimated cost (API)
Typical exchange (~3k in / 1k out)
Not disclosed
1M in + 1M out
Not disclosed
All benchmarks
Arena Elo
MMLU
GPQA
HumanEval
SWE-Bench
MATH

Similar models