S
SenseVoice
QwenAudio/SenseVoice
Open-source SenseVoiceSmall model for Mandarin, Cantonese, English, Japanese, and Korean ASR, language ID, emotion recognition, and audio event detection.
★9.0kstars
C
MIT
Updated: Today
📋 Project at a Glance
Tap to expand
What's this?A , built with C open-source project in the Models category, core strengths: asr/audio-analysis
Who made it?Maintained by QwenAudio team, 9K⭐ on GitHub, #166 out of 3201 in Models
Why does it exist?The QwenAudio team recognized that existing asr tools in Models were hard to use. SenseVoice was designed to make audio-analysis more accessible.
What can it do?Key use cases: audio-event-detection, cantonese, cross-lingual
How to install with AI?Use an AI coding assistant to follow the README and automatically handle the install and environment setup.
🔗 github.com/QwenAudio/SenseVoice | 官网 https://huggingface.co/spaces/FunAudioLLM/SenseVoice
🔗 github.com/QwenAudio/SenseVoice | 官网 https://huggingface.co/spaces/FunAudioLLM/SenseVoice
Topics
asraudio-analysisaudio-event-detectioncantonesecross-lingualemotion-detectionfunasrlanguage-identificationllama-cppmultilingualmultilingual-asrpytorchsensevoicespeech-emotion-recognitionspeech-recognitionspeech-to-textspeech-understandingtranscriptionvoice-aiwhisper-alternative