LOADING DATASHEET
LOADING DATASHEET
by Pyannote
RANKED #13 OF 19 AUDIO · OVERALL #259 OF 6,520 · VIBE SCORE 5.6 · 33 VOICES
voice-activity-detection on Hugging Face (automatic speech recognition). 4,300,717 downloads. Open weights for local or hosted use.
1 mentions
1 mentions
26 weeks · 14 voices
TOO QUIET
Not enough weekly voices to call a trend yet.
Open weights. You can use a hosted service, or download it and run it yourself, free.
Not enough discussion about Voice Activity Detection yet to call it. We found 33 posts but people did not say much either way.
32 POSTS · 1 COMMENTS · HACKER NEWS · LEMMY · REDDIT · GITHUB · DEV.TO
Other models the crowd has fully reviewed, starting with audio models like this one.
TTS Audio Suite v4.15 - Step Audio EditX Engine & Universal Inline Edit Tags
I built a 100% local, CPU-only voice loop for Ollama — talk to your models hands-free (Silero VAD + Parakeet STT + Supertonic TTS 3)
Full speech pipeline in native Swift/MLX — ASR, TTS, diarization, speech-to-speech, all on-device
Real-Time Speech-to-Speech Chatbot: Whisper, Llama 3.1, Kokoro, and Silero VAD 🚀
Was looking for open source AI dictation app for typing long prompts, finally built one - OmniDictate
My dream project is finally live: An open-source AI voice agent framework.
Self-hosted voice for any agent/harness of your choice (open-source)
Show HN: Open-source turn detection model for voice AI
My dream project is finally live: An open-source AI voice agent framework.
Introducing Two Libraries for Voice-Based Applications
ChatGPT helped me create this Ai powered Haunted Mirror for Halloween
TTS Audio Suite v4.15 - Step Audio EditX Engine & Universal Inline Edit Tags