LIVE
Formalizing Fermat's Last Theorem04/09/26 · Anthropic|OpenAI and Anthropic suffer simultaneous outages with no official explanation04/09/26|Show HN: Open-Source eInk Bike Computer04/09/26|Corporate America increasingly turns to open-source AI04/09/26|Study finds Google's AI Mode surfaces pricier products than classic search04/09/26 · Google|Study measures how coding agents pick third-party tools03/09/26|OpenAI's GPT-6 Astra sets new state-of-the-art on ARC-AGI-303/09/26 · OpenAI|ESPO: A Prompt Optimization Method That Beats GEPA With Shorter, More Stable Prompts03/09/26|"Last Translation Benchmark": a large-scale collaborative effort for a definitive MT benchmark03/09/26|NVIDIA pushes local AI at IFA 2026 with RTX Spark PCs and a personal inference router03/09/26 · NVIDIA|Google DeepMind unveils WeatherNext 3, its most accurate weather AI model yet03/09/26 · Google DeepMind|OpenAI Launches $1B Daybreak Program for Frontline Cybersecurity Defenders03/09/26 · OpenAI|Formalizing Fermat's Last Theorem04/09/26 · Anthropic|OpenAI and Anthropic suffer simultaneous outages with no official explanation04/09/26|Show HN: Open-Source eInk Bike Computer04/09/26|Corporate America increasingly turns to open-source AI04/09/26|Study finds Google's AI Mode surfaces pricier products than classic search04/09/26 · Google|Study measures how coding agents pick third-party tools03/09/26|OpenAI's GPT-6 Astra sets new state-of-the-art on ARC-AGI-303/09/26 · OpenAI|ESPO: A Prompt Optimization Method That Beats GEPA With Shorter, More Stable Prompts03/09/26|"Last Translation Benchmark": a large-scale collaborative effort for a definitive MT benchmark03/09/26|NVIDIA pushes local AI at IFA 2026 with RTX Spark PCs and a personal inference router03/09/26 · NVIDIA|Google DeepMind unveils WeatherNext 3, its most accurate weather AI model yet03/09/26 · Google DeepMind|OpenAI Launches $1B Daybreak Program for Frontline Cybersecurity Defenders03/09/26 · OpenAI|
🇺🇸 United States🎙️ Voice synthesisReleased April 2025

Dia 1.6B

Nari Labs
Open-sourceSelf-hostRGPD
Verdict

A solid pick for self-hosting and full data control.

Overview

Dia 1.6B is a text-to-speech model from Nari Labs that generates dialogue directly from a transcript, with tone control, voice cloning and nonverbal sounds like laughs or sighs. Released under Apache 2.0, it remains a research-stage project: it only generates English, needs roughly 10GB of VRAM on GPU (no CPU support or quantized build yet), and voice consistency across runs isn't guaranteed without an audio prompt or fixed seed. Self-hosted, it keeps all data on the user's own infrastructure regardless of the vendor's US base.

Skill profile

Not disclosed

Strengths

  • Open-source and self-hostable

Limitations

  • API pricing not disclosed

Who is it for

A good fit if…
  • you want to control cost or self-host
  • you have GDPR constraints
Skip it if…

    Ideal use cases

    • dialogue voice generation
    • podcast script prototyping
    • controlled voice cloning
    • nonverbal audio effects
    • on-prem GPU-based TTS

    Access & availability

    Self-hostable (open weights)

    Key specifications

    Context
    0
    Input price
    Output price
    Speed
    Price not auditedBenchmarks not audited

    Privacy

    GDPR-compliantHosting : self

    Run it locally

    Runs on a laptop30k3k
    QuantizationDiskRAM / VRAMTypical hardware
    Q4 · recommended0.9 GB3 GBAny recent PC/Mac
    Q8 · balanced1.7 GB4 GBAny recent PC/Mac
    FP16 · max quality3.2 GB6 GBAny recent PC/Mac

    Estimates for a moderate context. Long contexts need more RAM (KV cache).

    Deploy

    Copy-ready commands generated from this card. Adjust context length and GPU count to your hardware.

    Estimated memory · FP16 6 Go · Q4 3 Go

    OpenAI-compatible server for production on NVIDIA GPUs.

    pip install vllm
    vllm serve nari-labs/Dia-1.6B \
      --max-model-len 32768 \
      --tensor-parallel-size 1 \
      --dtype auto
    nari-labs/Dia-1.6B
    Advanced data · for experts
    Architecture, modalities, detailed cost, full benchmarks
    Architecture
    Size
    1.6B
    Cutoff date
    Inputs
    text
    Outputs
    text
    License
    apache-2.0
    Hosting
    self
    Hallucination score
    Estimated cost (API)
    Typical exchange (~3k in / 1k out)
    Not disclosed
    1M in + 1M out
    Not disclosed
    All benchmarks
    Arena Elo
    MMLU
    GPQA
    HumanEval
    SWE-Bench
    MATH

    Similar models