Kazakh Whisper Large v3 Turbo π Best open source speech recognition model for the Kazakh language. Kazakh Whisper Large v3 Turbo is a Whisper based speech to text model fine tuned specifically for Kazakh using a large scale mixture of public speech datasets. The model is built on OpenAI's Whisper Large v3 Turbo architecture and trained on over 841,000 speech transcript pairs representing approximately 1,500+ hours of Kazakh speech collected from multiple public sources. The resulting checkpoint is a merged full Transformers model that can be used directly without LoRA adapters or custom loading code. Key Features Optimized specifically for Kazakh ASR Trained on 841k+ speech transcript pairs 1,500+ hours of speech Based on Whisper Large v3 Turbo Ready to use Transformers checkpoint Evaluated on external and internal benchmarks Compatible with Hugging Face pipelines Suitable for production and research use Benchmark Summary FLEURS Kazakh Test Model WER β CER β : : Kazakh Whisper Large v3 Turbo 11.80% 4.98% Whisper Large v3 Turbo 19.75% 5.05% Wav2Vec2 XLSR Kazakh 21.75% 6.24% Whisper Large v3 31.10% 6.57% Whisper Medium 48.69% 10.92% Whisper Small 70.45% 21.23% The model achieves theβ¦
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy