NB Whisper Large Introducing the Norwegian NB Whisper Large model , proudly developed by the National Library of Norway. NB Whisper is a cutting edge series of models designed for automatic speech recognition (ASR) and speech translation. These models are based on the work of OpenAI's Whisper. Each model in the series has been trained for 250,000 steps, utilizing a diverse dataset of 8 million samples. These samples consist of aligned audio clips, each 30 seconds long, culminating in a staggering 66,000 hours of speech. For an in depth understanding of our training methodology and dataset composition, keep an eye out for our upcoming article. Model Size Parameters Model Tiny 39M NB Whisper Tiny Base 74M NB Whisper Base Small 244M NB Whisper Small Medium 769M NB Whisper Medium Large 1550M NB Whisper Large Verbatim Model While the main models are suitable for most transcription task, we demonstrate how easy it is to change the output of the main model. The following models are trained 250 additional steps from the main models above, and might be suitable for more targetted use cases: Verbatim version : This lower cased variant is more literal and suitable for tasks requiring detailed…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy