Speech, vision & media
NVIDIA NeMo Speech
The current speech-focused successor repository to NeMo, with reusable training and inference components.
Overview
Research summary
NVIDIA NeMo Speech provides training and inference infrastructure for automatic speech recognition, text-to-speech and speech language models. It combines reusable model implementations, pretrained checkpoints, examples and configuration-driven workflows for developers working in PyTorch. The repository formerly named NeMo now focuses on speech and audio after the upstream repository split, so this entry describes the current Speech project.
Source installation and published containers offer different ways to obtain its dependencies. Training targets NVIDIA GPU environments, while inference requirements depend on the selected model. Framework code uses Apache-2.0, with checkpoint-specific licensing recorded separately upstream.
Repository summary
- Stars
- 18,538
- Open issues
- Unavailable
- Last push
- 2026-10-01
- Commits, 90 days
- Unavailable
- Repository activity
- Not scored
- Version
- Unavailable
Recorded catalogue figures. View repository data and provenance →
Classification
Pricing & services
Paid services unknown
Whether the provider offers paid products or services has not been established.
Licence scope
Implementation
Recorded implementation details and interfaces for NVIDIA NeMo Speech.
Implementation details
Python/PyTorch speech framework with configuration-driven training and inference
- Languages
- Python
- Repository type
- source
Recorded interfaces and capabilities
Licence scope
Repository
Repository snapshots, release information and recorded maintenance signals.
Repository snapshot
- Stars
- 18,538
- Open issues
- Unavailable
- Last push
- 2026-10-01
- Commits, 90 days
- Unavailable
- Repository activity
- Not scored
- Archived
- Not recorded
Repository activity is a snapshot, not a quality or popularity ranking. It combines recent-push freshness (50%), 90-day commits (30%) and issue pressure (20%).
Maintenance and provenance
- Catalogue snapshot
- 2026-10-06
- Stars source
- Recorded fallback
- Last push source
- Recorded fallback
Documentation
Recorded references and research provenance for this entry.
Recorded sources 4
- https://docs.nvidia.com/nemo/speech/nightly/index.html Project page · Documentation
- https://github.com/NVIDIA-NeMo/Speech Linked repository · Research reference
- https://github.com/NVIDIA-NeMo/Speech/blob/main/README.md Research reference
- https://github.com/NVIDIA-NeMo/Speech/blob/main/LICENSE Research reference
Research metadata
- Research date
- 2026-10-02
- Catalogue snapshot
- 2026-10-06