EN DIRECT
Introducing Cosmos 3 Edge20/07/26 · Hugging Face|Kimi K3, Qwen 3.8, and Anthropic's (Potential) Unravelling20/07/26 · Anthropic|At SIGGRAPH, NVIDIA Advances Graphics and Simulation With Agentic and Physical AI20/07/26 · NVIDIA|China's open-weights AI strategy is winning20/07/26|Bristol Myers Squibb Building Life Science Industry’s Most Advanced AI Factory on NVIDIA Vera Rubin20/07/26 · NVIDIA|Safety and alignment in an era of long-horizon models20/07/26 · OpenAI|Claude Fable produced a counterexample to the Jacobian Conjecture20/07/26 · Anthropic|Apply for Anthropic’s AI for Science rare disease research grants20/07/26 · Anthropic|AI advice made people less accurate but more confident – sudy19/07/26|Vincentwei1021/video-shotcraft: AI video skill for Claude Code & Codex — cinematic product videos with Remotion: 106 shot recipe cards, 161 motion previews, a production-ready template19/07/26 · Anthropic|Claude Code uses Bun written in Rust now19/07/26 · Anthropic|OpenAI reduces Codex Model Context Size from 372k to 272k19/07/26 · OpenAI|Introducing Cosmos 3 Edge20/07/26 · Hugging Face|Kimi K3, Qwen 3.8, and Anthropic's (Potential) Unravelling20/07/26 · Anthropic|At SIGGRAPH, NVIDIA Advances Graphics and Simulation With Agentic and Physical AI20/07/26 · NVIDIA|China's open-weights AI strategy is winning20/07/26|Bristol Myers Squibb Building Life Science Industry’s Most Advanced AI Factory on NVIDIA Vera Rubin20/07/26 · NVIDIA|Safety and alignment in an era of long-horizon models20/07/26 · OpenAI|Claude Fable produced a counterexample to the Jacobian Conjecture20/07/26 · Anthropic|Apply for Anthropic’s AI for Science rare disease research grants20/07/26 · Anthropic|AI advice made people less accurate but more confident – sudy19/07/26|Vincentwei1021/video-shotcraft: AI video skill for Claude Code & Codex — cinematic product videos with Remotion: 106 shot recipe cards, 161 motion previews, a production-ready template19/07/26 · Anthropic|Claude Code uses Bun written in Rust now19/07/26 · Anthropic|OpenAI reduces Codex Model Context Size from 372k to 272k19/07/26 · OpenAI|
🇺🇸 États-Unis🧠 GénéralisteSortie avril 2025

Llama 4 Behemoth

Meta
Open-sourceGratuitSelf-host
Verdict

Une option solide pour l'auto-hébergement et le contrôle total des données.

Présentation

Llama 4 Behemoth est le plus gros modèle de Meta, multimodal natif et 1M de contexte. La licence Llama Community permet l'usage commercial pour les organisations <700M MAU. Idéal pour les setups self-host de niveau enterprise. Beaucoup plus lent que les frontiers fermés.

Profil de compétences

FrançaisRaisonn.VitesseCréativitéSûreté
Arena Elo1330
Benchmarks vérifiés
GPQA66%
SWE-Bench45%
MATH86%
MMLU87%
HumanEval88.5%

Que signifient ces scores ?

GPQA66%très bon

Questions scientifiques de niveau doctorat (physique, chimie, biologie), sans accès à des outils.

SWE-Bench45%correct

Résolution de vrais tickets GitHub en conditions réelles (SWE-bench Verified).

MATH86%excellent

Problèmes mathématiques de niveau compétition.

MMLU87%excellent

Connaissances générales réparties sur des dizaines de matières académiques.

HumanEval88.5%excellent

Génération de code Python à partir de spécifications.

Forces

  • Excellent en code
  • Très bonne qualité en français
  • Très long contexte
  • Dispose d'un tier gratuit

Limites

  • Vitesse moyenne
  • Pas de garantie RGPD native

À qui ça s'adresse

Pour vous si…
  • vous voulez maîtriser vos coûts ou auto-héberger
  • vous travaillez en français
À éviter si…
  • vous manipulez des données sensibles européennes
  • vous avez besoin de réponses très rapides

Cas d'usage idéaux

  • Self-host pro à grande échelle
  • Recherche académique
  • Multimodal long contexte

Accès & disponibilité

API payante au tokenTier gratuit disponibleAuto-hébergeable (poids ouverts)

Caractéristiques clés

Contexte
1M
Prix entrée
$3.00 $/M
Prix sortie
$9.00 $/M
Vitesse
40 tok/s
Prix non auditéBenchmarks non audités

Estimez votre coût mensuel

Tarifs API au token
$154/ mois avec Llama 4 Behemoth

Pour le même usage

Estimation indicative sur les tarifs API standard (hors cache, batch et remises volume). Vérifiez toujours la grille officielle avant de vous engager.

Confidentialité

Non conforme RGPDHébergement : US/selfDonnées anonymisables
Données avancées · pour experts
Architecture, modalités, coût détaillé, benchmarks complets
Architecture
MoE
Taille
~2T (288B actifs MoE)
Date de coupe
nov. 2024
Entrées
text, image
Sorties
text, code
Licence
llama-community
Hébergement
US/self
Score d'hallucination
3/5
Coût estimé (API)
Échange type (~3k entrée / 1k sortie)
≈ $0.0180
1M entrée + 1M sortie
≈ $12.00
Tous les benchmarks
Arena Elo1330
MMLU87.0%
GPQA66.0%
HumanEval88.5%
SWE-Bench45.0%
MATH86.0%

Modèles similaires