High-Quality Voice Cloning TTS for 600+ Languages
Categories & Topics
Repository Stats
Related Tools
GPA
[AutoArk] GPA (General Purpose Audio) can do ASR, TTS and voice conversion with one tiny model!
KittenTTS
KittenTTS enables users to generate lifelike speech from text, making it easier to create engaging audio content. With a focus on accessibility, it allows anyone to transform written information into spoken words effortlessly.
lingbot-world
Lingbot-world enables users to create and train chatbots that can communicate in multiple languages, making it easier for businesses and individuals to interact globally. By simplifying the process of developing language-specific bots, it enhances customer engagement and accessibility across different regions.
LuxTTS
LuxTTS is a voice cloning model that allows for high-quality text-to-speech conversion, achieving speeds up to 150 times faster than real-time. It offers an efficient solution for creating realistic voiceovers quickly, making it valuable for various applications in media and communication.
MOSS-TTS
MOSS‑TTS Family is an open‑source speech and sound generation model family from MOSI.AI and the OpenMOSS team. It is designed for high‑fidelity, high‑expressiveness, and complex real‑world scenarios, covering stable long‑form speech, multi‑speaker dialogue, voice/character design, environmental sound effects, and real‑time streaming TTS.
MOSS-TTS-Nano
MOSS-TTS-Nano is an open-source multilingual tiny speech generation model from MOSI.AI and the OpenMOSS team. With only 0.1B parameters, it is designed for realtime speech generation, can run directly on CPU without a GPU, and keeps the deployment stack simple enough for local demos, web serving, and lightweight product integration.