Typhoon Whisper Turbo Typhoon Whisper Turbo is a state of the art Thai Automatic Speech Recognition (ASR) model fine tuned on the OpenAI Whisper Large v3 Turbo architecture. It is engineered to deliver a balance of high accuracy and low latency for offline transcription, significantly outperforming standard Whisper models in throughput while maintaining robust performance on Thai speech. The model was presented in the paper Typhoon ASR Real time: FastConformer Transducer for Thai Automatic Speech Recognition. By using this model, you agree to the OpenTyphoon Terms and Conditions and acknowledge the Privacy Notice: https://opentyphoon.ai/tac · https://opentyphoon.ai/privacy Project Page: opentyphoon.ai GitHub Repository: scb 10x/typhoon asr The model was trained on approximately 11,000 hours of Thai audio, curated and normalized using the SCB 10X Typhoon data pipeline to ensure consistent handling of Thai numbers, repetition markers, and context dependent ambiguities. Model Overview Architecture: Whisper Large v3 Turbo (4 decoder layers vs 32 in standard Large v3) Language: Thai Dataset: ~11,000 hours of normalized Thai speech (Gigaspeech2, CommonVoice, Internal Curated Public Media…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy