-
facebook/seamless-streaming
Text-to-Speech ⢠Updated ⢠287 -
pyannote/speaker-diarization-3.1
Automatic Speech Recognition ⢠Updated ⢠9.49M ⢠3.36k -
pyannote/segmentation
Voice Activity Detection ⢠Updated ⢠4.69M ⢠694 -
pyannote/segmentation-3.0
Voice Activity Detection ⢠Updated ⢠6M ⢠1.66k
Myem
Sergioso
AI & ML interests
None yet
Organizations
None yet
GOOD
- RunningAgents516
tts Text To Speech
š516Text-to-speech (TTS) with Next-gen Kaldi
- PausedAgents11
Text To Speech
š11Transcribe audio to text
- Build errorAgents18
UTMOS Demo
š¢18Evaluate audio quality with MOS score
- Runtime errorAgentsFeatured220
SpeechT5 Speech Synthesis Demo
š©220
SD_Comfy_IMG
SDModels
AudioVideo
TTSSS
- Runtime errorAgents21
Youtube Video Translator
šØ21Translate YouTube videos to different languages
- RunningAgents516
tts Text To Speech
š516Text-to-speech (TTS) with Next-gen Kaldi
- Running on ZeroAgents512
AICoverGen
š512Launch a web UI for interacting with the model
- Runtime errorAgents316
Tortoise Tts
š¢316ExpressivText-to-Speech
Subtitle
Sdiff
AISTS
-
facebook/seamless-streaming
Text-to-Speech ⢠Updated ⢠287 -
pyannote/speaker-diarization-3.1
Automatic Speech Recognition ⢠Updated ⢠9.49M ⢠3.36k -
pyannote/segmentation
Voice Activity Detection ⢠Updated ⢠4.69M ⢠694 -
pyannote/segmentation-3.0
Voice Activity Detection ⢠Updated ⢠6M ⢠1.66k
AudioVideo
GOOD
- RunningAgents516
tts Text To Speech
š516Text-to-speech (TTS) with Next-gen Kaldi
- PausedAgents11
Text To Speech
š11Transcribe audio to text
- Build errorAgents18
UTMOS Demo
š¢18Evaluate audio quality with MOS score
- Runtime errorAgentsFeatured220
SpeechT5 Speech Synthesis Demo
š©220
TTSSS
- Runtime errorAgents21
Youtube Video Translator
šØ21Translate YouTube videos to different languages
- RunningAgents516
tts Text To Speech
š516Text-to-speech (TTS) with Next-gen Kaldi
- Running on ZeroAgents512
AICoverGen
š512Launch a web UI for interacting with the model
- Runtime errorAgents316
Tortoise Tts
š¢316ExpressivText-to-Speech
SD_Comfy_IMG
Subtitle
SDModels
Sdiff