EN DIRECT
Introducing Cosmos 3 Edge20/07/26 · Hugging Face|Kimi K3, Qwen 3.8, and Anthropic's (Potential) Unravelling20/07/26 · Anthropic|At SIGGRAPH, NVIDIA Advances Graphics and Simulation With Agentic and Physical AI20/07/26 · NVIDIA|China's open-weights AI strategy is winning20/07/26|Bristol Myers Squibb Building Life Science Industry’s Most Advanced AI Factory on NVIDIA Vera Rubin20/07/26 · NVIDIA|Safety and alignment in an era of long-horizon models20/07/26 · OpenAI|Claude Fable produced a counterexample to the Jacobian Conjecture20/07/26 · Anthropic|Apply for Anthropic’s AI for Science rare disease research grants20/07/26 · Anthropic|AI advice made people less accurate but more confident – sudy19/07/26|Vincentwei1021/video-shotcraft: AI video skill for Claude Code & Codex — cinematic product videos with Remotion: 106 shot recipe cards, 161 motion previews, a production-ready template19/07/26 · Anthropic|Claude Code uses Bun written in Rust now19/07/26 · Anthropic|OpenAI reduces Codex Model Context Size from 372k to 272k19/07/26 · OpenAI|Introducing Cosmos 3 Edge20/07/26 · Hugging Face|Kimi K3, Qwen 3.8, and Anthropic's (Potential) Unravelling20/07/26 · Anthropic|At SIGGRAPH, NVIDIA Advances Graphics and Simulation With Agentic and Physical AI20/07/26 · NVIDIA|China's open-weights AI strategy is winning20/07/26|Bristol Myers Squibb Building Life Science Industry’s Most Advanced AI Factory on NVIDIA Vera Rubin20/07/26 · NVIDIA|Safety and alignment in an era of long-horizon models20/07/26 · OpenAI|Claude Fable produced a counterexample to the Jacobian Conjecture20/07/26 · Anthropic|Apply for Anthropic’s AI for Science rare disease research grants20/07/26 · Anthropic|AI advice made people less accurate but more confident – sudy19/07/26|Vincentwei1021/video-shotcraft: AI video skill for Claude Code & Codex — cinematic product videos with Remotion: 106 shot recipe cards, 161 motion previews, a production-ready template19/07/26 · Anthropic|Claude Code uses Bun written in Rust now19/07/26 · Anthropic|OpenAI reduces Codex Model Context Size from 372k to 272k19/07/26 · OpenAI|
🇨🇳 Chine generalistSortie juillet 2025

GLM-4.5 Air

Z.ai
Open-sourceSelf-host
Verdict

Une option solide pour l'auto-hébergement et le contrôle total des données.

Profil de compétences

Non communiqué

Forces

  • Open-source et auto-hébergeable

Limites

  • Pas de garantie RGPD native
  • Tarification API non publiée

À qui ça s'adresse

Pour vous si…
  • vous voulez maîtriser vos coûts ou auto-héberger
À éviter si…
  • vous manipulez des données sensibles européennes

Accès & disponibilité

Auto-hébergeable (poids ouverts)

Caractéristiques clés

Contexte
128K
Prix entrée
Prix sortie
Vitesse
Prix non auditéBenchmarks non audités

Confidentialité

Non conforme RGPD

Faire tourner en local

Serveur GPU dédié418k614
QuantizationDisqueRAM / VRAMMatériel type
Q4 · recommandé60.4 Go68 GoMac Studio 96-128 Go / 4× GPU
Q8 · équilibré113.4 Go127 GoInfra GPU serveur (H100/A100…)
FP16 · qualité max212 Go235 GoInfra GPU serveur (H100/A100…)

Estimations pour un contexte modéré. Un long contexte augmente la RAM nécessaire (KV cache).

Données avancées · pour experts
Architecture, modalités, coût détaillé, benchmarks complets
Architecture
Taille
106B (MoE, 12B actifs)
Date de coupe
Entrées
text
Sorties
text
Licence
mit
Hébergement
Score d'hallucination
Coût estimé (API)
Échange type (~3k entrée / 1k sortie)
Non communiqué
1M entrée + 1M sortie
Non communiqué
Tous les benchmarks
Arena Elo
MMLU
GPQA
HumanEval
SWE-Bench
MATH

Modèles similaires