LIVE
Introducing Cosmos 3 Edge20/07/26 · Hugging Face|Kimi K3, Qwen 3.8, and Anthropic's (Potential) Unravelling20/07/26 · Anthropic|At SIGGRAPH, NVIDIA Advances Graphics and Simulation With Agentic and Physical AI20/07/26 · NVIDIA|China's open-weights AI strategy is winning20/07/26|Bristol Myers Squibb Building Life Science Industry’s Most Advanced AI Factory on NVIDIA Vera Rubin20/07/26 · NVIDIA|Safety and alignment in an era of long-horizon models20/07/26 · OpenAI|Claude Fable produced a counterexample to the Jacobian Conjecture20/07/26 · Anthropic|Apply for Anthropic’s AI for Science rare disease research grants20/07/26 · Anthropic|AI advice made people less accurate but more confident – sudy19/07/26|Vincentwei1021/video-shotcraft: AI video skill for Claude Code & Codex — cinematic product videos with Remotion: 106 shot recipe cards, 161 motion previews, a production-ready template19/07/26 · Anthropic|Claude Code uses Bun written in Rust now19/07/26 · Anthropic|OpenAI reduces Codex Model Context Size from 372k to 272k19/07/26 · OpenAI|Introducing Cosmos 3 Edge20/07/26 · Hugging Face|Kimi K3, Qwen 3.8, and Anthropic's (Potential) Unravelling20/07/26 · Anthropic|At SIGGRAPH, NVIDIA Advances Graphics and Simulation With Agentic and Physical AI20/07/26 · NVIDIA|China's open-weights AI strategy is winning20/07/26|Bristol Myers Squibb Building Life Science Industry’s Most Advanced AI Factory on NVIDIA Vera Rubin20/07/26 · NVIDIA|Safety and alignment in an era of long-horizon models20/07/26 · OpenAI|Claude Fable produced a counterexample to the Jacobian Conjecture20/07/26 · Anthropic|Apply for Anthropic’s AI for Science rare disease research grants20/07/26 · Anthropic|AI advice made people less accurate but more confident – sudy19/07/26|Vincentwei1021/video-shotcraft: AI video skill for Claude Code & Codex — cinematic product videos with Remotion: 106 shot recipe cards, 161 motion previews, a production-ready template19/07/26 · Anthropic|Claude Code uses Bun written in Rust now19/07/26 · Anthropic|OpenAI reduces Codex Model Context Size from 372k to 272k19/07/26 · OpenAI|
🇺🇸 United States🧠 GeneralistReleased November 2025

Gemini 3 Flash

Google
FreeRGPD
Verdict

Strong value for high-volume usage.

Overview

Gemini 3 Flash is extremely fast (250 tokens/sec) and one of the cheapest on the market while remaining multimodal. Excellent for consumer apps, high-traffic chatbots, indexing or large-scale video search.

Skill profile

FrenchReasoningSpeedCreativitySafety
Arena Elo1295
Verified benchmarks
GPQA60%
MATH78%
MMLU79.5%
HumanEval82%

What do these scores mean?

GPQA60%solid

PhD-level science questions (physics, chemistry, biology), with no tool access.

MATH78%very good

Competition-level math problems.

MMLU79.5%very good

General knowledge across dozens of academic subjects.

HumanEval82%excellent

Python code generation from specifications.

Strengths

  • Strong French quality
  • Very long context
  • Has a free tier
  • Robust safety filters

Who is it for

A good fit if…
  • you want to control cost or self-host
  • you work in French
  • you have GDPR constraints
Skip it if…

    Ideal use cases

    • High-volume multimodal tasks
    • Real-time apps
    • Indexing
    • Video search
    • Gemini free tier

    Access & availability

    Paid API per tokenFree tier available

    Key specifications

    Context
    1M
    Input price
    $0.50 $/M
    Output price
    $3.00 $/M
    Speed
    250 tok/s
    Price verified Jun 1, 2026·Official sourceBenchmarks not audited

    Estimate your monthly cost

    Per-token API pricing
    $36/ month with Gemini 3 Flash

    For the same usage

    Indicative estimate based on standard API rates (excluding caching, batch and volume discounts). Always check the official pricing before committing.

    Privacy

    GDPR-compliantHosting : US/EUAnonymizable data
    Advanced data · for experts
    Architecture, modalities, detailed cost, full benchmarks
    Architecture
    MoE
    Size
    undisclosed
    Cutoff date
    Jun 2025
    Inputs
    text, image, audio, video
    Outputs
    text, code
    License
    commercial
    Hosting
    US/EU
    Hallucination score
    3/5
    Estimated cost (API)
    Typical exchange (~3k in / 1k out)
    ≈ $0.0045
    1M in + 1M out
    ≈ $3.50
    All benchmarks
    Arena Elo1295
    MMLU79.5%
    GPQA60.0%
    HumanEval82.0%
    SWE-Bench
    MATH78.0%

    Similar models