🩺 Parakeet TDT 0.6B English Medical 🇬🇧 A fine tune of nvidia/parakeet tdt 0.6b v3 on the English subset of MultiMed mixed with Common Voice 17 English (train + validation) . The mix is the trick: it pushes the model toward medical vocabulary (TAVI, intervertebral disc herniation, drug names, dosing instructions) while keeping the everyday English it already knew. Outputs cased English text with punctuation. Drop in for the base Parakeet: same NeMo API, same long form support, same timestamps. 🔥 Quick start 📊 Results One model, one training mix (MultiMed en train + Common Voice 17 en train + validation, concatenated and shuffled per epoch — same .nemo for every row below). The two rows are the same checkpoint evaluated on two different held out test sets : one in domain (medical) and one out of domain (general English). Neither test set was seen during training. The zero shot column is the unmodified nvidia/parakeet tdt 0.6b v3 , measured on the same test set with the same evaluator. All numbers are normalized (lowercase + strip punctuation), the standard protocol used by the MultiMed paper and the Open ASR Leaderboard, so they are directly comparable to other published results…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy