ASR nvidia/nemotron-speech-streaming-en-0.6b Automatic Speech Recognition • 0.6B • Updated 15 days ago • 139k • 611
nvidia/nemotron-speech-streaming-en-0.6b Automatic Speech Recognition • 0.6B • Updated 15 days ago • 139k • 611
INDIC TTS DATASETS my own collection of TTS Datasets for finetuning models on Indic languages. edwixx/Gujarati40h Updated Oct 18, 2024 • 6
Audio Models Collection of best text-to-audio models. stabilityai/stable-audio-open-1.0 Text-to-Audio • 1B • Updated Jun 19, 2025 • 21.5k • 1.56k Running on Zero Agents 326 TangoFlux 🚀 326 Text to Audio (Sound SFX) Generator openbmb/MiniCPM-o-2_6 Any-to-Any • 9B • Updated 2 days ago • 232k • 1.3k
TTS Collection of some of the TTS models i found cool SWivid/F5-TTS Text-to-Speech • Updated Mar 21, 2025 • 778k • 1.2k fishaudio/fish-speech-1.4 Text-to-Speech • Updated Nov 5, 2024 • 465 • 460 coqui/XTTS-v2 Text-to-Speech • Updated Dec 11, 2023 • 8.49M • 3.74k microsoft/speecht5_tts Text-to-Speech • Updated Nov 8, 2023 • 59.8k • 841
ASR nvidia/nemotron-speech-streaming-en-0.6b Automatic Speech Recognition • 0.6B • Updated 15 days ago • 139k • 611
nvidia/nemotron-speech-streaming-en-0.6b Automatic Speech Recognition • 0.6B • Updated 15 days ago • 139k • 611
Audio Models Collection of best text-to-audio models. stabilityai/stable-audio-open-1.0 Text-to-Audio • 1B • Updated Jun 19, 2025 • 21.5k • 1.56k Running on Zero Agents 326 TangoFlux 🚀 326 Text to Audio (Sound SFX) Generator openbmb/MiniCPM-o-2_6 Any-to-Any • 9B • Updated 2 days ago • 232k • 1.3k
INDIC TTS DATASETS my own collection of TTS Datasets for finetuning models on Indic languages. edwixx/Gujarati40h Updated Oct 18, 2024 • 6
TTS Collection of some of the TTS models i found cool SWivid/F5-TTS Text-to-Speech • Updated Mar 21, 2025 • 778k • 1.2k fishaudio/fish-speech-1.4 Text-to-Speech • Updated Nov 5, 2024 • 465 • 460 coqui/XTTS-v2 Text-to-Speech • Updated Dec 11, 2023 • 8.49M • 3.74k microsoft/speecht5_tts Text-to-Speech • Updated Nov 8, 2023 • 59.8k • 841