GitHub 项目简介: A TTS model capable of generating ultra-realistic dialogue in one pass.
README: Dia directly generates highly realistic dialogue from a transcript. You can condition the output on audio, enabling emotion and tone control. The model can also produce nonverbal…
README: To accelerate research, we are providing access to pretrained model checkpoints and inference code. The model weights are hosted on Hugging Face.
README: This project offers a high-fidelity speech generation model intended for research and educational use.
README: Docker support for ARM architecture and MacOS.