-
facebook/seamless-streaming
Text-to-Speech ⢠Updated ⢠289 -
pyannote/speaker-diarization-3.1
Automatic Speech Recognition ⢠Updated ⢠6.86M ⢠4.27k -
pyannote/segmentation
Voice Activity Detection ⢠Updated ⢠616k ⢠702 -
pyannote/segmentation-3.0
Voice Activity Detection ⢠Updated ⢠5.22M ⢠2.16k
Myem
Sergioso
AI & ML interests
None yet
Organizations
None yet
GOOD
- RunningAgents526
tts Text To Speech
š526Text-to-speech (TTS) with Next-gen Kaldi
- PausedAgents11
Text To Speech
š11Transcribe audio to text
- Build errorAgents18
UTMOS Demo
š¢18Evaluate audio quality with MOS score
- Runtime errorAgentsFeatured220
SpeechT5 Speech Synthesis Demo
š©220
SD_Comfy_IMG
SDModels
AudioVideo
TTSSS
- Runtime errorAgents21
Youtube Video Translator
šØ21Translate YouTube videos to different languages
- RunningAgents526
tts Text To Speech
š526Text-to-speech (TTS) with Next-gen Kaldi
- Running on ZeroAgents526
AICoverGen
š526Launch a web UI for interacting with the model
- Runtime errorAgents316
Tortoise Tts
š¢316ExpressivText-to-Speech
Subtitle
Sdiff
AISTS
-
facebook/seamless-streaming
Text-to-Speech ⢠Updated ⢠289 -
pyannote/speaker-diarization-3.1
Automatic Speech Recognition ⢠Updated ⢠6.86M ⢠4.27k -
pyannote/segmentation
Voice Activity Detection ⢠Updated ⢠616k ⢠702 -
pyannote/segmentation-3.0
Voice Activity Detection ⢠Updated ⢠5.22M ⢠2.16k
AudioVideo
GOOD
- RunningAgents526
tts Text To Speech
š526Text-to-speech (TTS) with Next-gen Kaldi
- PausedAgents11
Text To Speech
š11Transcribe audio to text
- Build errorAgents18
UTMOS Demo
š¢18Evaluate audio quality with MOS score
- Runtime errorAgentsFeatured220
SpeechT5 Speech Synthesis Demo
š©220
TTSSS
- Runtime errorAgents21
Youtube Video Translator
šØ21Translate YouTube videos to different languages
- RunningAgents526
tts Text To Speech
š526Text-to-speech (TTS) with Next-gen Kaldi
- Running on ZeroAgents526
AICoverGen
š526Launch a web UI for interacting with the model
- Runtime errorAgents316
Tortoise Tts
š¢316ExpressivText-to-Speech
SD_Comfy_IMG
Subtitle
SDModels
Sdiff