LIVE
Which tools do Claude, Codex and Cursor choose? We measured 17k runs to find out03/09/26 · Anthropic|OpenAI's GPT-6 Astra on ARC-AGI-303/09/26 · OpenAI|GPT-6 Astra03/09/26 · OpenAI|OpenAI begins rolling out GPT-6 Astra03/09/26 · OpenAI|ESPO: Error-Structured Prompt Optimization via Diagnose, Diversify, and Stabilize03/09/26 · Anthropic|Sparks Fly: NVIDIA Accelerates Local AI at IFA 202603/09/26 · NVIDIA|Introducing WeatherNext 3, our most advanced and accurate global weather AI model03/09/26 · Google|Porting my 1993 Amiga game to Godot, with an LLM reading the 68000 assembly03/09/26 · Anthropic|Claude outage – Resolved03/09/26 · Anthropic|Daybreak for Frontline Defenders: $1B to protect essential services03/09/26 · OpenAI|NeoMME: an efficient Multimodal-native and Multilingual Encoder03/09/26 · Hugging Face|‘NBA 2K27’ With NVIDIA DLSS 5 Leads 28 New Games Coming to GeForce NOW03/09/26 · NVIDIA|Which tools do Claude, Codex and Cursor choose? We measured 17k runs to find out03/09/26 · Anthropic|OpenAI's GPT-6 Astra on ARC-AGI-303/09/26 · OpenAI|GPT-6 Astra03/09/26 · OpenAI|OpenAI begins rolling out GPT-6 Astra03/09/26 · OpenAI|ESPO: Error-Structured Prompt Optimization via Diagnose, Diversify, and Stabilize03/09/26 · Anthropic|Sparks Fly: NVIDIA Accelerates Local AI at IFA 202603/09/26 · NVIDIA|Introducing WeatherNext 3, our most advanced and accurate global weather AI model03/09/26 · Google|Porting my 1993 Amiga game to Godot, with an LLM reading the 68000 assembly03/09/26 · Anthropic|Claude outage – Resolved03/09/26 · Anthropic|Daybreak for Frontline Defenders: $1B to protect essential services03/09/26 · OpenAI|NeoMME: an efficient Multimodal-native and Multilingual Encoder03/09/26 · Hugging Face|‘NBA 2K27’ With NVIDIA DLSS 5 Leads 28 New Games Coming to GeForce NOW03/09/26 · NVIDIA|
🇨🇳 China🧠 GeneralistReleased October 2024

Yi-Lightning

01.AI
Free
Verdict

Strong value for high-volume usage.

Overview

Yi-Lightning by 01.AI (Kai-Fu Lee's team) is designed for Asian consumer apps needing volume and low latency. Very good at Chinese/English, weak on French.

Skill profile

FrenchReasoningSpeedCreativitySafety
Arena Elo1328
Verified benchmarks
MMLU81.5%
HumanEval82%

What do these scores mean?

MMLU81.5%excellent

General knowledge across dozens of academic subjects.

HumanEval82%excellent

Python code generation from specifications.

Strengths

  • Has a free tier
  • Competitive input pricing

Limitations

  • No native GDPR guarantee

Who is it for

A good fit if…
  • you want to control cost or self-host
Skip it if…
  • you handle sensitive EU data

Ideal use cases

  • Volume + speed in Chinese/English
  • Low-cost Asia apps

Access & availability

Paid API per tokenFree tier available

Key specifications

Context
32K
Input price
$0.14 $/M
Output price
$0.14 $/M
Speed
180 tok/s
Price not auditedBenchmarks not audited

Estimate your monthly cost

Per-token API pricing
$5.20/ month with Yi-Lightning

For the same usage

Indicative estimate based on standard API rates (excluding caching, batch and volume discounts). Always check the official pricing before committing.

Privacy

Not GDPR-compliantHosting : CN
Advanced data · for experts
Architecture, modalities, detailed cost, full benchmarks
Architecture
MoE
Size
undisclosed
Cutoff date
Jul 2024
Inputs
text
Outputs
text, code
License
commercial
Hosting
CN
Hallucination score
3/5
Estimated cost (API)
Typical exchange (~3k in / 1k out)
≈ $0.0006
1M in + 1M out
≈ $0.28
All benchmarks
Arena Elo1328
MMLU81.5%
GPQA
HumanEval82.0%
SWE-Bench
MATH

Similar models