LIVE
Formalizing Fermat's Last Theorem04/09/26 · Anthropic|OpenAI and Anthropic suffer simultaneous outages with no official explanation04/09/26|Show HN: Open-Source eInk Bike Computer04/09/26|Corporate America increasingly turns to open-source AI04/09/26|Study finds Google's AI Mode surfaces pricier products than classic search04/09/26 · Google|Study measures how coding agents pick third-party tools03/09/26|OpenAI's GPT-6 Astra sets new state-of-the-art on ARC-AGI-303/09/26 · OpenAI|ESPO: A Prompt Optimization Method That Beats GEPA With Shorter, More Stable Prompts03/09/26|"Last Translation Benchmark": a large-scale collaborative effort for a definitive MT benchmark03/09/26|NVIDIA pushes local AI at IFA 2026 with RTX Spark PCs and a personal inference router03/09/26 · NVIDIA|Google DeepMind unveils WeatherNext 3, its most accurate weather AI model yet03/09/26 · Google DeepMind|OpenAI Launches $1B Daybreak Program for Frontline Cybersecurity Defenders03/09/26 · OpenAI|Formalizing Fermat's Last Theorem04/09/26 · Anthropic|OpenAI and Anthropic suffer simultaneous outages with no official explanation04/09/26|Show HN: Open-Source eInk Bike Computer04/09/26|Corporate America increasingly turns to open-source AI04/09/26|Study finds Google's AI Mode surfaces pricier products than classic search04/09/26 · Google|Study measures how coding agents pick third-party tools03/09/26|OpenAI's GPT-6 Astra sets new state-of-the-art on ARC-AGI-303/09/26 · OpenAI|ESPO: A Prompt Optimization Method That Beats GEPA With Shorter, More Stable Prompts03/09/26|"Last Translation Benchmark": a large-scale collaborative effort for a definitive MT benchmark03/09/26|NVIDIA pushes local AI at IFA 2026 with RTX Spark PCs and a personal inference router03/09/26 · NVIDIA|Google DeepMind unveils WeatherNext 3, its most accurate weather AI model yet03/09/26 · Google DeepMind|OpenAI Launches $1B Daybreak Program for Frontline Cybersecurity Defenders03/09/26 · OpenAI|
🇨🇳 China🎨 Image generationReleased November 2025

Z-Image Turbo

Alibaba
Open-sourceSelf-hostRGPD
Verdict

A solid pick for self-hosting and full data control.

Overview

Z-Image Turbo is Alibaba's distilled, 8-step version of the Z-Image foundation model, a 6B single-stream DiT (S3-DiT) built for speed. It delivers photorealistic output and bilingual English/Chinese text rendering with sub-second latency, even on a 16GB consumer GPU. The trade-off: output diversity is lower than the base Z-Image model, and text rendering isn't tuned for French or other European languages. Apache-2.0 open weights mean fully local deployment — good for data residency, though the model itself originates from China.

Skill profile

Not disclosed

Strengths

  • Open-source and self-hostable

Limitations

  • API pricing not disclosed

Who is it for

A good fit if…
  • you want to control cost or self-host
  • you have GDPR constraints
Skip it if…

    Ideal use cases

    • Marketing image generation
    • Rapid creative prototyping
    • Fully local deployment
    • Bilingual EN/CN text rendering
    • Consumer-GPU app integration

    Access & availability

    Self-hostable (open weights)

    Key specifications

    Context
    0
    Input price
    Output price
    Speed
    Price not auditedBenchmarks not audited

    Privacy

    GDPR-compliantHosting : self

    Run it locally

    Runs on a laptop682k5k
    QuantizationDiskRAM / VRAMTypical hardware
    Q4 · recommended3.4 GB6 GBAny recent PC/Mac
    Q8 · balanced6.4 GB9 GB16GB PC / M1+ Mac / 8GB GPU
    FP16 · max quality12 GB15 GB32GB PC / 24GB Mac / RTX 4070 Ti+

    Estimates for a moderate context. Long contexts need more RAM (KV cache).

    Deploy

    Copy-ready commands generated from this card. Adjust context length and GPU count to your hardware.

    Estimated memory · FP16 15 Go · Q4 6 Go

    OpenAI-compatible server for production on NVIDIA GPUs.

    pip install vllm
    vllm serve Tongyi-MAI/Z-Image-Turbo \
      --max-model-len 32768 \
      --tensor-parallel-size 1 \
      --dtype auto
    Tongyi-MAI/Z-Image-Turbo
    Advanced data · for experts
    Architecture, modalities, detailed cost, full benchmarks
    Architecture
    dense
    Size
    6B
    Cutoff date
    Inputs
    text
    Outputs
    text
    License
    apache-2.0
    Hosting
    self
    Hallucination score
    Estimated cost (API)
    Typical exchange (~3k in / 1k out)
    Not disclosed
    1M in + 1M out
    Not disclosed
    All benchmarks
    Arena Elo
    MMLU
    GPQA
    HumanEval
    SWE-Bench
    MATH

    Similar models