Skip to content
View gyt1145028706's full-sized avatar

Highlights

  • Pro

Block or report gyt1145028706

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse

Pinned Loading

  1. XY-Tokenizer XY-Tokenizer Public

    This is the code for paper: XY-Tokenizer: Mitigating the Semantic-Acoustic Conflict in Low-Bitrate Speech Codecs

    Python 97 5

  2. OpenMOSS/MOSS-Audio-Tokenizer OpenMOSS/MOSS-Audio-Tokenizer Public

    A 1.6B causal Transformer audio tokenizer with streaming, variable bitrates, and semantic alignment across speech, sound, and music

    Python 255 18

  3. OpenMOSS/MOSS-TTS OpenMOSS/MOSS-TTS Public

    An open-source model family for long-form speech, dialogue synthesis, voice design, sound effects, and real-time streaming TTS

    Python 4.1k 368

  4. OpenMOSS/MOSS-TTSD OpenMOSS/MOSS-TTSD Public

    A multilingual model for long-form, multi-speaker dialogue synthesis with flexible speaker control and zero-shot voice cloning

    Python 1.4k 136

  5. SpeechGPT-2.0-preview SpeechGPT-2.0-preview Public

    Forked from OpenMOSS/SpeechGPT-2.0-preview

    GPT-4o-level, real-time spoken dialogue system.

    Python

  6. OpenMOSS/MOSS-TTS-Nano OpenMOSS/MOSS-TTS-Nano Public

    A 100M-parameter multilingual TTS model for real-time CPU inference, voice cloning, and 48 kHz stereo generation

    Python 4.3k 548