LIVE
Formalizing Fermat's Last Theorem04/09/26 · Anthropic|OpenAI and Anthropic suffer simultaneous outages with no official explanation04/09/26|Show HN: Open-Source eInk Bike Computer04/09/26|Corporate America increasingly turns to open-source AI04/09/26|Study finds Google's AI Mode surfaces pricier products than classic search04/09/26 · Google|Study measures how coding agents pick third-party tools03/09/26|OpenAI's GPT-6 Astra sets new state-of-the-art on ARC-AGI-303/09/26 · OpenAI|ESPO: A Prompt Optimization Method That Beats GEPA With Shorter, More Stable Prompts03/09/26|"Last Translation Benchmark": a large-scale collaborative effort for a definitive MT benchmark03/09/26|NVIDIA pushes local AI at IFA 2026 with RTX Spark PCs and a personal inference router03/09/26 · NVIDIA|Google DeepMind unveils WeatherNext 3, its most accurate weather AI model yet03/09/26 · Google DeepMind|OpenAI Launches $1B Daybreak Program for Frontline Cybersecurity Defenders03/09/26 · OpenAI|Formalizing Fermat's Last Theorem04/09/26 · Anthropic|OpenAI and Anthropic suffer simultaneous outages with no official explanation04/09/26|Show HN: Open-Source eInk Bike Computer04/09/26|Corporate America increasingly turns to open-source AI04/09/26|Study finds Google's AI Mode surfaces pricier products than classic search04/09/26 · Google|Study measures how coding agents pick third-party tools03/09/26|OpenAI's GPT-6 Astra sets new state-of-the-art on ARC-AGI-303/09/26 · OpenAI|ESPO: A Prompt Optimization Method That Beats GEPA With Shorter, More Stable Prompts03/09/26|"Last Translation Benchmark": a large-scale collaborative effort for a definitive MT benchmark03/09/26|NVIDIA pushes local AI at IFA 2026 with RTX Spark PCs and a personal inference router03/09/26 · NVIDIA|Google DeepMind unveils WeatherNext 3, its most accurate weather AI model yet03/09/26 · Google DeepMind|OpenAI Launches $1B Daybreak Program for Frontline Cybersecurity Defenders03/09/26 · OpenAI|
🇺🇸 United States🎧 Speech recognitionReleased August 2025

Parakeet TDT 0.6B v3

NVIDIA
Open-sourceSelf-hostRGPD
Verdict

A solid pick for self-hosting and full data control.

Overview

Parakeet TDT 0.6B v3 is NVIDIA's compact ASR model, small enough (0.6B parameters) to run on laptop-class hardware. It auto-detects and transcribes 25 European languages with punctuation and timestamps, making it a practical base for multilingual transcription or subtitling tools. Quality isn't uniform across languages, and NVIDIA flags a specific gap on Portuguese, trained on the European variant but often benchmarked against Brazilian Portuguese. Open weights under CC-BY-4.0 mean that, self-hosted, all audio and text stay on the user's own infrastructure regardless of the vendor's US origin.

Skill profile

Not disclosed

Strengths

  • Open-source and self-hostable

Limitations

  • API pricing not disclosed

Who is it for

A good fit if…
  • you want to control cost or self-host
  • you have GDPR constraints
Skip it if…

    Ideal use cases

    • Multilingual meeting transcription
    • Automated video subtitling
    • On-premise voice assistants
    • Confidential call analytics

    Access & availability

    Self-hostable (open weights)

    Key specifications

    Context
    0
    Input price
    Output price
    Speed
    Price not auditedBenchmarks not audited

    Privacy

    GDPR-compliantHosting : self

    Run it locally

    Runs on a laptop761k1k
    QuantizationDiskRAM / VRAMTypical hardware
    Q4 · recommended0.3 GB2 GBAny recent PC/Mac
    Q8 · balanced0.6 GB3 GBAny recent PC/Mac
    FP16 · max quality1.2 GB3 GBAny recent PC/Mac

    Estimates for a moderate context. Long contexts need more RAM (KV cache).

    Deploy

    Copy-ready commands generated from this card. Adjust context length and GPU count to your hardware.

    Estimated memory · FP16 3 Go · Q4 2 Go

    OpenAI-compatible server for production on NVIDIA GPUs.

    pip install vllm
    vllm serve nvidia/parakeet-tdt-0.6b-v3 \
      --max-model-len 32768 \
      --tensor-parallel-size 1 \
      --dtype auto
    nvidia/parakeet-tdt-0.6b-v3
    Advanced data · for experts
    Architecture, modalities, detailed cost, full benchmarks
    Architecture
    Size
    600M
    Cutoff date
    Inputs
    text
    Outputs
    text
    License
    cc-by-4.0
    Hosting
    self
    Hallucination score
    Estimated cost (API)
    Typical exchange (~3k in / 1k out)
    Not disclosed
    1M in + 1M out
    Not disclosed
    All benchmarks
    Arena Elo
    MMLU
    GPQA
    HumanEval
    SWE-Bench
    MATH

    Similar models