EN DIRECT
Introducing Cosmos 3 Edge20/07/26 · Hugging Face|Kimi K3, Qwen 3.8, and Anthropic's (Potential) Unravelling20/07/26 · Anthropic|At SIGGRAPH, NVIDIA Advances Graphics and Simulation With Agentic and Physical AI20/07/26 · NVIDIA|China's open-weights AI strategy is winning20/07/26|Bristol Myers Squibb Building Life Science Industry’s Most Advanced AI Factory on NVIDIA Vera Rubin20/07/26 · NVIDIA|Safety and alignment in an era of long-horizon models20/07/26 · OpenAI|Claude Fable produced a counterexample to the Jacobian Conjecture20/07/26 · Anthropic|Apply for Anthropic’s AI for Science rare disease research grants20/07/26 · Anthropic|AI advice made people less accurate but more confident – sudy19/07/26|Vincentwei1021/video-shotcraft: AI video skill for Claude Code & Codex — cinematic product videos with Remotion: 106 shot recipe cards, 161 motion previews, a production-ready template19/07/26 · Anthropic|Claude Code uses Bun written in Rust now19/07/26 · Anthropic|OpenAI reduces Codex Model Context Size from 372k to 272k19/07/26 · OpenAI|Introducing Cosmos 3 Edge20/07/26 · Hugging Face|Kimi K3, Qwen 3.8, and Anthropic's (Potential) Unravelling20/07/26 · Anthropic|At SIGGRAPH, NVIDIA Advances Graphics and Simulation With Agentic and Physical AI20/07/26 · NVIDIA|China's open-weights AI strategy is winning20/07/26|Bristol Myers Squibb Building Life Science Industry’s Most Advanced AI Factory on NVIDIA Vera Rubin20/07/26 · NVIDIA|Safety and alignment in an era of long-horizon models20/07/26 · OpenAI|Claude Fable produced a counterexample to the Jacobian Conjecture20/07/26 · Anthropic|Apply for Anthropic’s AI for Science rare disease research grants20/07/26 · Anthropic|AI advice made people less accurate but more confident – sudy19/07/26|Vincentwei1021/video-shotcraft: AI video skill for Claude Code & Codex — cinematic product videos with Remotion: 106 shot recipe cards, 161 motion previews, a production-ready template19/07/26 · Anthropic|Claude Code uses Bun written in Rust now19/07/26 · Anthropic|OpenAI reduces Codex Model Context Size from 372k to 272k19/07/26 · OpenAI|
🇨🇦 Canada multilingualSortie octobre 2024

Aya Expanse 32B

Cohere
Open-sourceSelf-host
Verdict

Une option solide pour l'auto-hébergement et le contrôle total des données.

Profil de compétences

Non communiqué

Forces

  • Open-source et auto-hébergeable

Limites

  • Pas de garantie RGPD native
  • Tarification API non publiée

À qui ça s'adresse

Pour vous si…
  • vous voulez maîtriser vos coûts ou auto-héberger
À éviter si…
  • vous manipulez des données sensibles européennes

Accès & disponibilité

Auto-hébergeable (poids ouverts)

Caractéristiques clés

Contexte
128K
Prix entrée
Prix sortie
Vitesse
Prix non auditéBenchmarks non audités

Confidentialité

Non conforme RGPD

Faire tourner en local

PC avec bon GPU / Mac 32 Go+4k298
QuantizationDisqueRAM / VRAMMatériel type
Q4 · recommandé18.4 Go22 GoMac 32 Go / RTX 3090-4090
Q8 · équilibré34.6 Go40 GoMac 64 Go / bi-GPU 24 Go
FP16 · qualité max64.6 Go73 GoMac Studio 96-128 Go / 4× GPU

Estimations pour un contexte modéré. Un long contexte augmente la RAM nécessaire (KV cache).

🤗 Hugging Face
ollama run aya-expanse:32b
Données avancées · pour experts
Architecture, modalités, coût détaillé, benchmarks complets
Architecture
Taille
32.3B
Date de coupe
Entrées
text
Sorties
text
Licence
cc-by-nc-4.0
Hébergement
Score d'hallucination
Coût estimé (API)
Échange type (~3k entrée / 1k sortie)
Non communiqué
1M entrée + 1M sortie
Non communiqué
Tous les benchmarks
Arena Elo
MMLU
GPQA
HumanEval
SWE-Bench
MATH

Modèles similaires