Speech

by NVIDIA-NeMo · AI

A scalable generative AI framework built for researchers and developers working on Large Language Models, Multimodal, an

New0 ratings18,397 starsActive
AIDeveloper ToolPythonJupyter Notebook
Open project ↗↓ Download v3.0.0GitHub⚑ Report
Speech preview

About this project

NVIDIA NeMo Speech Checkout our HuggingFace🤗 collection for the latest open weight checkpoints and demos! Updates NeMo Speech 3.0 is now available as release v3.0.0 and in the 26.07.00 NeMo Speech NGC container. The final NeMo release before the repository split was v2.7.3, available in the 26.04 NeMo NGC container. — 2026-07: MagpieTTS v2607 has been released with support for 3 new languages (Ar, Ko, Pt) + 9 existing languages (En, Es, De, Fr, Vi, It, Zh, Hi, Ja). Try out the demo! — 2026-06: Nemotron-3.5-ASR-Streaming-0.6B has been released with 40 languages supported, controllable latency 80ms-1s, and 240-2400 1xH100 concurrent streams. Built on cache-aware Fastconformer…

Technologies

ShellPythonDockerfileJupyter Notebookgenerative-aiJinjadeeplearningasr

Project health

Actively maintained
Last update2 days ago
Contributors400
Latest releasev3.0.0
Open issues & PRs304
LicenseApache-2.0
On GitHubsince 2019

GitHub

18,397
stars
3,606
forks
400
contributors
304
open issues & PRs
Python
language
2 days ago
last commit
View on GitHub ↗All releases ↗

Reviews

out of 5 · 0 ratings
★★★★★
0%
★★★★
0%
★★★
0%
★★
0%
0%
Sign in to write a review
No reviews yet
Be the first to review Speech.

Built by

NVIDIA-NeMo
Imported from GitHub · not yet claimed on GitPalace
View developer pageSign in with GitHub to claim

Maintain NVIDIA-NeMo/Speech? Claiming verifies admin access through your GitHub account and gives you control of this listing.

You might also like