-
facebook/seamless-streaming
Text-to-Speech ⢠Updated ⢠287 -
pyannote/speaker-diarization-3.1
Automatic Speech Recognition ⢠Updated ⢠8.55M ⢠3.63k -
pyannote/segmentation
Voice Activity Detection ⢠Updated ⢠1.96M ⢠696 -
pyannote/segmentation-3.0
Voice Activity Detection ⢠Updated ⢠5.73M ⢠1.81k
Myem
Sergioso
AI & ML interests
None yet
Organizations
None yet
GOOD
- RunningAgents520
tts Text To Speech
š520Text-to-speech (TTS) with Next-gen Kaldi
- PausedAgents11
Text To Speech
š11Transcribe audio to text
- Build errorAgents18
UTMOS Demo
š¢18Evaluate audio quality with MOS score
- Runtime errorAgentsFeatured220
SpeechT5 Speech Synthesis Demo
š©220
SD_Comfy_IMG
SDModels
AudioVideo
TTSSS
- Runtime errorAgents21
Youtube Video Translator
šØ21Translate YouTube videos to different languages
- RunningAgents520
tts Text To Speech
š520Text-to-speech (TTS) with Next-gen Kaldi
- Running on ZeroAgents514
AICoverGen
š514Launch a web UI for interacting with the model
- Runtime errorAgents316
Tortoise Tts
š¢316ExpressivText-to-Speech
Subtitle
Sdiff
AISTS
-
facebook/seamless-streaming
Text-to-Speech ⢠Updated ⢠287 -
pyannote/speaker-diarization-3.1
Automatic Speech Recognition ⢠Updated ⢠8.55M ⢠3.63k -
pyannote/segmentation
Voice Activity Detection ⢠Updated ⢠1.96M ⢠696 -
pyannote/segmentation-3.0
Voice Activity Detection ⢠Updated ⢠5.73M ⢠1.81k
AudioVideo
GOOD
- RunningAgents520
tts Text To Speech
š520Text-to-speech (TTS) with Next-gen Kaldi
- PausedAgents11
Text To Speech
š11Transcribe audio to text
- Build errorAgents18
UTMOS Demo
š¢18Evaluate audio quality with MOS score
- Runtime errorAgentsFeatured220
SpeechT5 Speech Synthesis Demo
š©220
TTSSS
- Runtime errorAgents21
Youtube Video Translator
šØ21Translate YouTube videos to different languages
- RunningAgents520
tts Text To Speech
š520Text-to-speech (TTS) with Next-gen Kaldi
- Running on ZeroAgents514
AICoverGen
š514Launch a web UI for interacting with the model
- Runtime errorAgents316
Tortoise Tts
š¢316ExpressivText-to-Speech
SD_Comfy_IMG
Subtitle
SDModels
Sdiff