- alibaba
Pinned Loading
- modelscope/FunASR
modelscope/FunASR PublicOpen-source speech recognition toolkit for training, inference, streaming ASR, VAD, punctuation, speaker diarization pipelines, and OpenAI-compatible/MCP serving.
- QwenAudio/CosyVoice
QwenAudio/CosyVoice PublicMulti-lingual large voice generation model, providing inference, training and deployment full-stack ability.
- QwenAudio/SenseVoice
QwenAudio/SenseVoice PublicOpen-source SenseVoiceSmall model for Mandarin, Cantonese, English, Japanese, and Korean ASR, language ID, emotion recognition, and audio event detection.
- modelscope/FunClip
modelscope/FunClip PublicFunASR-powered video transcription, subtitle generation, and LLM-assisted clipping tool with a local Gradio UI.
- ddlBoJack/emotion2vec
ddlBoJack/emotion2vec Public[ACL 2024] Official PyTorch code for extracting features and training downstream models with emotion2vec: Self-Supervised Pre-Training for Speech Emotion Representation
- X-LANCE/SLAM-LLM
X-LANCE/SLAM-LLM PublicA Framework for Speech, Language, Audio, Music Processing with Large Language Model
If the problem persists, check the GitHub status page or contact support.
Uh oh!
There was an error while loading. Please reload this page.




