ASR nvidia/nemotron-speech-streaming-en-0.6b Automatic Speech Recognition • 0.6B • Updated Aug 5 • 414k • 629
nvidia/nemotron-speech-streaming-en-0.6b Automatic Speech Recognition • 0.6B • Updated Aug 5 • 414k • 629
INDIC TTS DATASETS my own collection of TTS Datasets for finetuning models on Indic languages. edwixx/Gujarati40h Updated Oct 18, 2024 • 4
Audio Models Collection of best text-to-audio models. stabilityai/stable-audio-open-1.0 Text-to-Audio • 1B • Updated Jun 19, 2025 • 20.7k • 1.65k Running on Zero Agents 326 TangoFlux 🚀 326 Text to Audio (Sound SFX) Generator openbmb/MiniCPM-o-2_6 Any-to-Any • 9B • Updated Aug 18 • 343k • 1.3k
TTS Collection of some of the TTS models i found cool SWivid/F5-TTS Text-to-Speech • Updated Mar 21, 2025 • 903k • 1.21k fishaudio/fish-speech-1.4 Text-to-Speech • Updated Nov 5, 2024 • 268 • 460 coqui/XTTS-v2 Text-to-Speech • Updated Dec 11, 2023 • 6.85M • 3.82k microsoft/speecht5_tts Text-to-Speech • Updated Nov 8, 2023 • 49.1k • 845
ASR nvidia/nemotron-speech-streaming-en-0.6b Automatic Speech Recognition • 0.6B • Updated Aug 5 • 414k • 629
nvidia/nemotron-speech-streaming-en-0.6b Automatic Speech Recognition • 0.6B • Updated Aug 5 • 414k • 629
Audio Models Collection of best text-to-audio models. stabilityai/stable-audio-open-1.0 Text-to-Audio • 1B • Updated Jun 19, 2025 • 20.7k • 1.65k Running on Zero Agents 326 TangoFlux 🚀 326 Text to Audio (Sound SFX) Generator openbmb/MiniCPM-o-2_6 Any-to-Any • 9B • Updated Aug 18 • 343k • 1.3k
INDIC TTS DATASETS my own collection of TTS Datasets for finetuning models on Indic languages. edwixx/Gujarati40h Updated Oct 18, 2024 • 4
TTS Collection of some of the TTS models i found cool SWivid/F5-TTS Text-to-Speech • Updated Mar 21, 2025 • 903k • 1.21k fishaudio/fish-speech-1.4 Text-to-Speech • Updated Nov 5, 2024 • 268 • 460 coqui/XTTS-v2 Text-to-Speech • Updated Dec 11, 2023 • 6.85M • 3.82k microsoft/speecht5_tts Text-to-Speech • Updated Nov 8, 2023 • 49.1k • 845