EN DIRECT
Introducing Cosmos 3 Edge20/07/26 · Hugging Face|Kimi K3, Qwen 3.8, and Anthropic's (Potential) Unravelling20/07/26 · Anthropic|At SIGGRAPH, NVIDIA Advances Graphics and Simulation With Agentic and Physical AI20/07/26 · NVIDIA|China's open-weights AI strategy is winning20/07/26|Bristol Myers Squibb Building Life Science Industry’s Most Advanced AI Factory on NVIDIA Vera Rubin20/07/26 · NVIDIA|Safety and alignment in an era of long-horizon models20/07/26 · OpenAI|Claude Fable produced a counterexample to the Jacobian Conjecture20/07/26 · Anthropic|Apply for Anthropic’s AI for Science rare disease research grants20/07/26 · Anthropic|AI advice made people less accurate but more confident – sudy19/07/26|Vincentwei1021/video-shotcraft: AI video skill for Claude Code & Codex — cinematic product videos with Remotion: 106 shot recipe cards, 161 motion previews, a production-ready template19/07/26 · Anthropic|Claude Code uses Bun written in Rust now19/07/26 · Anthropic|OpenAI reduces Codex Model Context Size from 372k to 272k19/07/26 · OpenAI|Introducing Cosmos 3 Edge20/07/26 · Hugging Face|Kimi K3, Qwen 3.8, and Anthropic's (Potential) Unravelling20/07/26 · Anthropic|At SIGGRAPH, NVIDIA Advances Graphics and Simulation With Agentic and Physical AI20/07/26 · NVIDIA|China's open-weights AI strategy is winning20/07/26|Bristol Myers Squibb Building Life Science Industry’s Most Advanced AI Factory on NVIDIA Vera Rubin20/07/26 · NVIDIA|Safety and alignment in an era of long-horizon models20/07/26 · OpenAI|Claude Fable produced a counterexample to the Jacobian Conjecture20/07/26 · Anthropic|Apply for Anthropic’s AI for Science rare disease research grants20/07/26 · Anthropic|AI advice made people less accurate but more confident – sudy19/07/26|Vincentwei1021/video-shotcraft: AI video skill for Claude Code & Codex — cinematic product videos with Remotion: 106 shot recipe cards, 161 motion previews, a production-ready template19/07/26 · Anthropic|Claude Code uses Bun written in Rust now19/07/26 · Anthropic|OpenAI reduces Codex Model Context Size from 372k to 272k19/07/26 · OpenAI|
🇨🇳 Chine🔬 RaisonnementSortie janvier 2025

DeepSeek-R1

DeepSeek
Open-sourceSelf-host
Verdict

Une option solide pour l'auto-hébergement et le contrôle total des données.

Profil de compétences

Non communiqué

Forces

  • Open-source et auto-hébergeable

Limites

  • Pas de garantie RGPD native
  • Tarification API non publiée

À qui ça s'adresse

Pour vous si…
  • vous voulez maîtriser vos coûts ou auto-héberger
À éviter si…
  • vous manipulez des données sensibles européennes

Accès & disponibilité

Auto-hébergeable (poids ouverts)

Caractéristiques clés

Contexte
128K
Prix entrée
Prix sortie
Vitesse
Prix non auditéBenchmarks non audités

Confidentialité

Non conforme RGPD

Faire tourner en local

Serveur GPU dédié8.6M13k
QuantizationDisqueRAM / VRAMMatériel type
Q4 · recommandé382.5 Go423 GoInfra GPU serveur (H100/A100…)
Q8 · équilibré718 Go792 GoInfra GPU serveur (H100/A100…)
FP16 · qualité max1342 Go1478 GoInfra GPU serveur (H100/A100…)

Estimations pour un contexte modéré. Un long contexte augmente la RAM nécessaire (KV cache).

🤗 Hugging Face
ollama run deepseek-r1:671b
Données avancées · pour experts
Architecture, modalités, coût détaillé, benchmarks complets
Architecture
Taille
671B (MoE, 37B actifs)
Date de coupe
Entrées
text
Sorties
text
Licence
mit
Hébergement
Score d'hallucination
Coût estimé (API)
Échange type (~3k entrée / 1k sortie)
Non communiqué
1M entrée + 1M sortie
Non communiqué
Tous les benchmarks
Arena Elo
MMLU
GPQA
HumanEval
SWE-Bench
MATH

Modèles similaires