🎙️ Deepfake Audio – A neural voice cloning studio powered by SV2TTS technology.
-
Updated
Feb 25, 2026 - Python
🎙️ Deepfake Audio – A neural voice cloning studio powered by SV2TTS technology.
Open research toward end-to-end spoken dialogue systems. Ships SoviaMate-Codec — a neural audio codec for LLM integration with ASR-constrained encoding, enhancement training, and zero-shot speaker adaptation.
基于 VoxCPM2 的 HTTP TTS 服务,面向 Legado 与有声书场景,支持零样本音色克隆、多角色路由、流式语音合成、ASR 提示词自动生成及跨平台部署。
Python SDK for the TTS.ai text-to-speech API
JavaScript/Node.js SDK for TTS.ai API — text-to-speech, voice cloning, speech-to-text
Local multi-voice audiobook pipeline for the Shadow Slave web novel - LLM diarization + IndexTTS2 emotional voice cloning on a single 12 GB GPU
Add a description, image, and links to the zero-shot-voice-cloning topic page so that developers can more easily learn about it.
To associate your repository with the zero-shot-voice-cloning topic, visit your repo's landing page and select "manage topics."