Model Card for Model ID This model is a Yiddish finetune (continued training) of the OpenAI Whisper Large v3 model. Model Details Model Description Developed by: ivrit ai Language(s) (NLP): Yiddish License: Apache 2.0 Finetuned from model openai/whisper large v3 Training Date Oct 2025 Bias, Risks, and Limitations Language detection capability of this model has been degraded during training it is intended for mostly hebrew audio transcription. Language token should be explicitly set to Yiddish Additionally, the translation task was not trained and also degraded. This model would not be able to translate in any reasonable capacity. How to Get Started with the Model Please follow the original model card for usage details replacing with this model name. You can also find other weight formats and quantizations on the ivrit ai HF page. We created some simple example scripts using this model and weights for other inference runtimes. Find those in the "examples" folder within the training GitHub repo. Training Details Training Data This model was trained on the following datasets: ivrit ai/crowd recital yi whisper training Crowd sourced recording of Wikipedia/Michlol article snippets. ~78h…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy