This is a re trained 3 layer RoBERTa wwm ext model. Chinese BERT with Whole Word Masking For further accelerating Chinese natural language processing, we provide Chinese pre trained BERT with Whole Word Masking . Pre Training with Whole Word Masking for Chinese BERT Yiming Cui, Wanxiang Che, Ting Liu, Bing Qin, Ziqing Yang, Shijin Wang, Guoping Hu This repository is developed based on:https://github.com/google research/bert You may also interested in, Chinese BERT series: https://github.com/ymcui/Chinese BERT wwm Chinese MacBERT: https://github.com/ymcui/MacBERT Chinese ELECTRA: https://github.com/ymcui/Chinese ELECTRA Chinese XLNet: https://github.com/ymcui/Chinese XLNet Knowledge Distillation Toolkit TextBrewer: https://github.com/airaria/TextBrewer More resources by HFL: https://github.com/ymcui/HFL Anthology Citation If you find the technical report or resource is useful, please cite the following technical report in your paper. Primary: https://arxiv.org/abs/2004.13922 Secondary: https://arxiv.org/abs/1906.08101
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy