-
Fudan University
- Shanghai
-
01:18
(UTC -12:00) - https://gyt1145028706.github.io/
Highlights
- Pro
Pinned Loading
-
XY-Tokenizer
XY-Tokenizer PublicThis is the code for paper: XY-Tokenizer: Mitigating the Semantic-Acoustic Conflict in Low-Bitrate Speech Codecs
-
OpenMOSS/MOSS-Audio-Tokenizer
OpenMOSS/MOSS-Audio-Tokenizer PublicA 1.6B causal Transformer audio tokenizer with streaming, variable bitrates, and semantic alignment across speech, sound, and music
-
OpenMOSS/MOSS-TTS
OpenMOSS/MOSS-TTS PublicAn open-source model family for long-form speech, dialogue synthesis, voice design, sound effects, and real-time streaming TTS
-
OpenMOSS/MOSS-TTSD
OpenMOSS/MOSS-TTSD PublicA multilingual model for long-form, multi-speaker dialogue synthesis with flexible speaker control and zero-shot voice cloning
-
SpeechGPT-2.0-preview
SpeechGPT-2.0-preview PublicForked from OpenMOSS/SpeechGPT-2.0-preview
GPT-4o-level, real-time spoken dialogue system.
Python
-
OpenMOSS/MOSS-TTS-Nano
OpenMOSS/MOSS-TTS-Nano PublicA 100M-parameter multilingual TTS model for real-time CPU inference, voice cloning, and 48 kHz stereo generation
If the problem persists, check the GitHub status page or contact support.

