GitHub 项目简介: Clone a voice in 5 seconds to generate arbitrary speech in real-time
README: SV2TTS is a deep learning framework in three stages. In the first stage, one creates a digital representation of a voice from a few seconds of audio. In the second and third stage…
README: this repo has quickly gotten old. Many SaaS apps (often paying) will give you a better audio quality than this repository will. If you wish for an open-source solution with a high…
README: Install [ffmpeg](https://ffmpeg.org/download.html#get-packages). This is necessary for reading audio files.