难度中等,需克隆仓库并预先配置 Python 3.12+、PyTorch 2.7+ 及 CUDA 环境,通过运行 'uv sync --extra all --extra cu13' 或 'pip install nemo-toolkit[asr,tts]' 进行安装。
核验依据
GitHub 项目简介: A scalable generative AI framework built for researchers and developers working on Large Language Models, Multimodal, and Speech AI (Automatic Speech Recognition and Text-to-Speec…
README: Project Status: Active -- The project has reached a stable, usable state and is being actively developed.
README: Checkout our HuggingFace🤗 collection for the latest open weight checkpoints and demos!
README: Nemotron-3.5-ASR-Streaming-0.6B has been released with 40 languages supported, controllable latency 80ms-1s, and 240-2400 1xH100 concurrent streams.
README: NVIDIA NeMo Speech is built for researchers and PyTorch developers working on Speech models including Automatic Speech Recognition (ASR), Text to Speech (TTS), and Speech LLMs. It…