Pretrained on 10k hours WenetSpeech L subset. More details in TencentGameMate/chinese speech pretrain This model does not have a tokenizer as it was pretrained on audio alone. In order to use this model speech recognition, a tokenizer should be created and the model should be fine tuned on labeled text data. python package: transformers==4.16.2
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy