Trending repositories: audio-generation

4 tracked repositories tagged with audio-generation, ordered by stars. Use the topic filters below to narrow further.

Filter by topic

4 of 4 repositories

  • OpenBMB/VoxCPM

    VoxCPM2: Tokenizer-Free TTS for Multilingual Speech Generation, Creative Voice Design, and True-to-Life Cloning

    AI summary: A highly efficient, tokenizer-free Text-to-Speech (TTS) generation model.

    35,013ai-mlPythonApache-2.0
  • QwenLM/Qwen3-TTS

    Qwen3-TTS is an open-source series of TTS models developed by the Qwen team at Alibaba Cloud, supporting stable, expressive, and streaming speech generation, free-form voice design, and vivid voice cloning.

    AI summary: A highly natural, multilingual Text-to-Speech model from the Qwen team, capable of zero-shot voice cloning.

    12,836ai-mlPythonApache-2.0
  • MisoLabsAI/MisoTTS

    Miso TTS is an 8 billion, highly emotive text-to-speech model

    AI summary: A state-of-the-art 8B parameter text-to-speech RVQ Transformer model.

    3,192ai-mlPythonOther
  • samuel-vitorino/sopro

    A lightweight text-to-speech model with zero-shot voice cloning

    AI summary: A lightweight, streaming text-to-speech model offering zero-shot voice cloning at extreme speeds on CPU.

    877ai-mlPythonApache-2.0