LOADING DATASHEET
LOADING DATASHEET
by Nvidia
RANKED #22 OF 32 AUDIO · OVERALL #391 OF 7,062 · VIBE SCORE 5.4 · 25 VOICES
parakeet-tdt-0.6b-v2 on Hugging Face (automatic speech recognition). 2,195,476 downloads. Open weights for local or hosted use.
8 mentions
1 mentions
26 weeks · 37 voices
AGING WELL
The crowd is warmer now than it was at the start.
Open weights. You can use a hosted service, or download it and run it yourself, free.
People are mostly frustrated with parakeet tdt 0.6B right now.
24 POSTS · 1 COMMENTS · GITHUB · REDDIT · X
0 thumbs up · 6 thumbs down
Other models the crowd has fully reviewed, starting with audio models like this one.
I built a 100% local, CPU-only voice loop for Ollama — talk to your models hands-free (Silero VAD + Parakeet STT + Supertonic TTS 3)
🎙️ Offline Speech-to-Text with NVIDIA Parakeet-TDT 0.6B v2
parakeet.wgsl – Fast, accurate ASR in the browser, via raw WebGPU & SIMD WASM
I built a CPU-only local voice stack for AI agents (Claude Code, OpenCode, Codex) - Silero VAD + Parakeet STT + Supertonic TTS, one-command install, macOS/Linux/Windows
I had Grok 4.5 run an autoresearch optimization ladder on Parakeet CPU ASR. Result: Upto 2.1x faster STT
One of the new Personal Brains - What do you guys think?? 1. Brain — llama.cpp b10069, tuned like vLLM/sglang, GPU1 only
Fixes the crashloop that #3105 shipped. Kokoro was unaffected and is running on the GPU already. What happened Parakeet downloaded correctly and then the pod crashlooped building the onnxruntime session: onnxruntime.capi.onnxruntime_pybind11_state.Fail:…
Add an optional Parakeet TDT v3 engine (single pass, much faster than Whisper)
Follows #3104, which built the image this uses. Ruling Tom, 2026-09-22 (#2960 / #3101): the existing RTX A2000 12 GB in talosm01 hosts vexa-whisper (3.9 GB, WHISPERTTL=-1, holds the node's only nvidia.com/gpu unit) plus Home Assistant's local STT and TTS. Not
Run transcription on Parakeet TDT 0.6b v3 (FluidAudio), retiring whisper-server (#204)
Default local ASR to Parakeet TDT 0.6B v2 with Whisper fallback
AI catalog: refresh the text shortlist to Qwen3.5/Olmo 3/Granite, and trim the rows that were never the right pick