M
MOSS-TTS
OpenMOSS/MOSS-TTS
MOSS‑TTS Family is an open‑source speech and sound generation model family from MOSI.AI and the OpenMOSS team. It is designed for high‑fidelity, high‑expressiveness, and complex real‑world scenarios, covering stable long‑form speech, multi‑speaker dialogue, voice/character design, environmental sound effects, and real‑time streaming TTS.
★4.0kstars
Python
Apache-2.0
Updated: Today
📋 Project at a Glance
Tap to expand
What's this?A audio/audio-tokenizer tool in the Models category, built with Python, open-source
Who made it?Maintained by OpenMOSS team, 4K⭐ on GitHub, #383 out of 3201 in Models
Why does it exist?The OpenMOSS team recognized that existing audio tools in Models were hard to use. MOSS-TTS was designed to make audio-tokenizer more accessible.
What can it do?Key use cases: llm, multimodal, text-to-speech
How to install with AI?Use an AI coding assistant (Claude Code, Cursor, Copilot) to automatically set up pip dependencies and virtual env. Follow the README — the AI handles the rest.
🔗 github.com/OpenMOSS/MOSS-TTS | 官网 https://mosi.cn/models/moss-tts
🔗 github.com/OpenMOSS/MOSS-TTS | 官网 https://mosi.cn/models/moss-tts
Topics
audioaudio-tokenizerllmmultimodaltext-to-speechvoice-cloning