Automatic Speech Recognition
NeMo
PyTorch
English
speech
audio
Transducer
TDT
FastConformer
Conformer
NeMo
hf-asr-leaderboard
Eval Results (legacy)
Eval Results
Instructions to use nvidia/parakeet-tdt-0.6b-v2 with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- NeMo
How to use nvidia/parakeet-tdt-0.6b-v2 with NeMo:
import nemo.collections.asr as nemo_asr asr_model = nemo_asr.models.ASRModel.from_pretrained("nvidia/parakeet-tdt-0.6b-v2") transcriptions = asr_model.transcribe(["file.wav"]) - Notebooks
- Google Colab
- Kaggle
Fine tuning examples
#67
by sakgoyal - opened
I have my own custom dataset with a bunch of audio files and their transcriptions in webvtt format.
How can I use this to fine tune the model? I am having a hard time finding anything on the docs.
I would prefer to have training in a python script and not use a bash script if possible.
I am curious about this as well. I created AI models from scratch in the past but this will be my first time fine tuning an AI model from nvidia.