Skip to content
#

nvidia-nemo

Here are 63 public repositories matching this topic...

Automatic transcriber made with the Nvidia NeMo AI toolkit. Used to transcribe speech to text in real-time from any source. Requires CUDA capable GPU to run on the local machine, if setup using virtual audio cables can transcribe the audio that is being played in real-time without any other requirements.

  • Updated Oct 18, 2020
  • Python

Swift library for Speaker Embedding extraction and verification using NVIDIA NeMo TitaNet model converted to CoreML. Extract 192-dim speaker embeddings, verify speakers, and perform real-time speaker diarization on iOS/macOS.

  • Updated Feb 9, 2026
  • Swift

Improve this page

Add a description, image, and links to the nvidia-nemo topic page so that developers can more easily learn about it.

Curate this topic

Add this topic to your repo

To associate your repository with the nvidia-nemo topic, visit your repo's landing page and select "manage topics."

Learn more