Verified project record
NVIDIA-NeMo/Speech
A PyTorch-based generative AI framework for researchers and developers to build, customize, and deploy Automatic Speech Recognition, Text-to-Speech, and Speech Large Language Models. The project requires a pre-configured Python environment with CUDA-capable NVIDIA GPUs.