-
facebook/seamless-streaming
Text-to-Speech ⢠Updated ⢠285 -
pyannote/speaker-diarization-3.1
Automatic Speech Recognition ⢠Updated ⢠9.96M ⢠3.15k -
pyannote/segmentation
Voice Activity Detection ⢠Updated ⢠4.13M ⢠692 -
pyannote/segmentation-3.0
Voice Activity Detection ⢠Updated ⢠6.21M ⢠1.54k
Myem
Sergioso
AI & ML interests
None yet
Organizations
None yet
GOOD
- RunningAgents514
tts Text To Speech
š514Text-to-speech (TTS) with Next-gen Kaldi
- PausedAgents11
Text To Speech
š11Transcribe audio to text
- Build errorAgents18
UTMOS Demo
š¢18Evaluate audio quality with MOS score
- Runtime errorAgentsFeatured220
SpeechT5 Speech Synthesis Demo
š©220
SD_Comfy_IMG
SDModels
AudioVideo
TTSSS
- Runtime errorAgents21
Youtube Video Translator
šØ21Translate YouTube videos to different languages
- RunningAgents514
tts Text To Speech
š514Text-to-speech (TTS) with Next-gen Kaldi
- Running on ZeroAgents510
AICoverGen
š510Launch a web UI for interacting with the model
- Runtime errorAgents316
Tortoise Tts
š¢316ExpressivText-to-Speech
Subtitle
Sdiff
AISTS
-
facebook/seamless-streaming
Text-to-Speech ⢠Updated ⢠285 -
pyannote/speaker-diarization-3.1
Automatic Speech Recognition ⢠Updated ⢠9.96M ⢠3.15k -
pyannote/segmentation
Voice Activity Detection ⢠Updated ⢠4.13M ⢠692 -
pyannote/segmentation-3.0
Voice Activity Detection ⢠Updated ⢠6.21M ⢠1.54k
AudioVideo
GOOD
- RunningAgents514
tts Text To Speech
š514Text-to-speech (TTS) with Next-gen Kaldi
- PausedAgents11
Text To Speech
š11Transcribe audio to text
- Build errorAgents18
UTMOS Demo
š¢18Evaluate audio quality with MOS score
- Runtime errorAgentsFeatured220
SpeechT5 Speech Synthesis Demo
š©220
TTSSS
- Runtime errorAgents21
Youtube Video Translator
šØ21Translate YouTube videos to different languages
- RunningAgents514
tts Text To Speech
š514Text-to-speech (TTS) with Next-gen Kaldi
- Running on ZeroAgents510
AICoverGen
š510Launch a web UI for interacting with the model
- Runtime errorAgents316
Tortoise Tts
š¢316ExpressivText-to-Speech
SD_Comfy_IMG
Subtitle
SDModels
Sdiff