VoxCPM2: Tokenizer-Free TTS for Multilingual Speech Generation, Creative Voice Design, and True-to-Life Cloning
-
Updated
Sep 30, 2026 - Python
VoxCPM2: Tokenizer-Free TTS for Multilingual Speech Generation, Creative Voice Design, and True-to-Life Cloning
AI-powered multi-voice audiobook generator — LLM script annotation, voice cloning, voice design, LoRA training, per-line style control, and export to MP3, chaptered M4B, or Audacity multi-track. Built on Qwen3-TTS.
A programmable version of Neil Thapen's Pink Trombone
LA Studio is a local-first AI audio platform for exploring, downloading, and testing speech-to-text, text-to-speech, and voice cloning models
Open-source AI audiobook studio. A free, private alternative to ElevenLabs. 3 voice modes, per-sentence voice & emotion control, LLM smart character analysis, mixed-voice generation. Runs 100% locally on your GPU with zero API costs.
VoxFlow 声流 — AI 声音到 AI 音乐,一套工作流做到自动上架
A local, multi-voice audiobook generator: an LLM casts every character in its own voice, kept consistent across a whole series — with voice design, per-line emotion, seven languages, and M4B / Audiobookshelf export. Runs on your own machine; you own the files. Kokoro · Qwen3-TTS.
Qwen3-TTS Audiobook Studio: Ultimate local multi-role AI audiobook generator. Built-in 3s Voice Clone & Design. Portable one-click launch for Mac/Win. 极致本地 AI 有声书制作工坊。
Native macOS text-to-speech app powered by Qwen3-TTS and Apple Silicon (MLX). Voice cloning, voice design, and custom voices — all running locally.
五引擎语音合成平台:VoxCPM2 / IndexTTS 2.5 / IndexTTS 2.0 / OpenVoice / Step-Audio-EditX,声音克隆、声音设计、多角色剧本配音、流式生成,中英日韩界面,Windows/Linux/Docker 一键部署 | Multi-engine TTS platform (5 engines) with voice cloning, voice design, multi-character dubbing & streaming, zh/en/ja/ko UI, one-click deploy
14 curated Spanish voices (clone + design) for Qwen3-TTS — Spain, Mexico, Argentina, Chile, Colombia. Runs locally on Apple Silicon via MLX.
🎙️ Qwen3-TTS-DubFlow: An open-source, human-in-the-loop AI dubbing workbench for novels, games, podcasts, and more. Features a "Design-then-Clone" workflow powered by Qwen3-TTS to achieve consistent identity and context-aware emotional performance.
Generate realistic speech and clone voices in ComfyUI with VoxCPM custom nodes, featuring token-free TTS, style guidance, and LoRA training support.
Local, offline text-to-speech and voice cloning on audio.cpp — multi-voice narration, voice design, sound/music generation, and a real-time conversational voice loop with a web UI. Runs on AMD (ROCm), NVIDIA (CUDA), Vulkan, or CPU.
VOX-1 Audiobook Maker is a local, GPU-accelerated studio for creating professional-quality audiobooks. Powered by OmniVoice TTS, it allows users to design voices from text descriptions or clone narrators from short audio samples.
CLI for Qwen3-TTS: voice clone, voice design & custom voice generation on local GPU. 10 languages, 9 speakers, batch mode.
An open-source Claude skill that builds a unique, consistent LinkedIn voice. Framework-driven persona interview (Jungian archetypes + NN/g tone model + anti-voice) → daily post generation on a 30/25/20/15/10 content mix. Works for one person or a whole team. Install with one npx command.
用喜歡的歌學語言:貼上 YouTube MV 網址,Gemini 自動轉錄歌詞,加上拼音、繁體中文翻譯與文法說明,再由 Gemini 3.8 Flash TTS 設計的老師逐句示範發音(正常速與慢速)。支援日文、韓文、英文。以 Next.js 與 Python 打造,部署在 Cloud Run 並以 IAP 保��。
VoiceFlow - Modern text-to-speech web application with real-time word highlighting, customizable voice settings, and content management. Built with React, TypeScript, and Web Speech API.
AI读书伴侣 — 把读书变成愉悦、享受、沉浸的体验。三种模式:基础TTS(8种内置音色)/ 声音设计(文字描述创造声线)/ 声音克隆(5秒录音)。支持 PDF/TXT/MD/EPUB。
To associate your repository with the voice-design topic, visit your repo's landing page and select "manage topics."