LIVE
Corporate America increasingly turns to open-source AI04/09/26|Study finds Google's AI Mode surfaces pricier products than classic search04/09/26 · Google|Study measures how coding agents pick third-party tools03/09/26|OpenAI's GPT-6 Astra sets new state-of-the-art on ARC-AGI-303/09/26 · OpenAI|ESPO: A Prompt Optimization Method That Beats GEPA With Shorter, More Stable Prompts03/09/26|"Last Translation Benchmark": a large-scale collaborative effort for a definitive MT benchmark03/09/26|NVIDIA pushes local AI at IFA 2026 with RTX Spark PCs and a personal inference router03/09/26 · NVIDIA|Google DeepMind unveils WeatherNext 3, its most accurate weather AI model yet03/09/26 · Google DeepMind|OpenAI Launches $1B Daybreak Program for Frontline Cybersecurity Defenders03/09/26 · OpenAI|Hcompany releases NeoMME, a compact multimodal-native encoder for retrieval03/09/26 · Hcompany|Playco cuts manual fixes by 50% using GPT-6 Astra for prototyping03/09/26 · OpenAI|Legora reviews 41 financial documents in minutes with GPT-6 Astra03/09/26 · OpenAI|Corporate America increasingly turns to open-source AI04/09/26|Study finds Google's AI Mode surfaces pricier products than classic search04/09/26 · Google|Study measures how coding agents pick third-party tools03/09/26|OpenAI's GPT-6 Astra sets new state-of-the-art on ARC-AGI-303/09/26 · OpenAI|ESPO: A Prompt Optimization Method That Beats GEPA With Shorter, More Stable Prompts03/09/26|"Last Translation Benchmark": a large-scale collaborative effort for a definitive MT benchmark03/09/26|NVIDIA pushes local AI at IFA 2026 with RTX Spark PCs and a personal inference router03/09/26 · NVIDIA|Google DeepMind unveils WeatherNext 3, its most accurate weather AI model yet03/09/26 · Google DeepMind|OpenAI Launches $1B Daybreak Program for Frontline Cybersecurity Defenders03/09/26 · OpenAI|Hcompany releases NeoMME, a compact multimodal-native encoder for retrieval03/09/26 · Hcompany|Playco cuts manual fixes by 50% using GPT-6 Astra for prototyping03/09/26 · OpenAI|Legora reviews 41 financial documents in minutes with GPT-6 Astra03/09/26 · OpenAI|
🇺🇸 United States🧠 GeneralistReleased November 2025

Gemini 3 Flash

Google
FreeRGPD
Verdict

Strong value for high-volume usage.

Overview

Gemini 3 Flash is extremely fast (250 tokens/sec) and one of the cheapest on the market while remaining multimodal. Excellent for consumer apps, high-traffic chatbots, indexing or large-scale video search.

Skill profile

FrenchReasoningSpeedCreativitySafety
Arena Elo1474
Verified benchmarks
GPQA60%
MATH78%
MMLU79.5%
HumanEval82%

What do these scores mean?

GPQA60%solid

PhD-level science questions (physics, chemistry, biology), with no tool access.

MATH78%very good

Competition-level math problems.

MMLU79.5%very good

General knowledge across dozens of academic subjects.

HumanEval82%excellent

Python code generation from specifications.

Strengths

  • Strong French quality
  • World-class on Arena
  • Very long context
  • Has a free tier

Who is it for

A good fit if…
  • you want to control cost or self-host
  • you work in French
  • you have GDPR constraints
Skip it if…

    Ideal use cases

    • High-volume multimodal tasks
    • Real-time apps
    • Indexing
    • Video search
    • Gemini free tier

    Access & availability

    Paid API per tokenFree tier available

    Key specifications

    Context
    1M
    Input price
    $0.50/M
    Output price
    $3.00/M
    Speed
    250 tok/s
    Price verified Jun 1, 2026·Official sourceBenchmarks not audited

    Estimate your monthly cost

    Per-token API pricing
    $36/ month with Gemini 3 Flash

    For the same usage

    Indicative estimate based on standard API rates (excluding caching, batch and volume discounts). Always check the official pricing before committing.

    Privacy

    GDPR-compliantHosting : US/EUAnonymizable data
    Advanced data · for experts
    Architecture, modalities, detailed cost, full benchmarks
    Architecture
    MoE
    Size
    undisclosed
    Cutoff date
    Jun 2025
    Inputs
    text, image, audio, video
    Outputs
    text, code
    License
    proprietary
    Hosting
    US/EU
    Hallucination score
    3/5
    Estimated cost (API)
    Typical exchange (~3k in / 1k out)
    ≈ $0.0045
    1M in + 1M out
    ≈ $3.50
    All benchmarks
    Arena Elo1474
    MMLU79.5%
    GPQA60.0%
    HumanEval82.0%
    SWE-Bench
    MATH78.0%

    Similar models

    Articles mentioning it