datasets: UMLS [news] A cross lingual extension of SapBERT will appear in the main onference of ACL 2021 ! [news] SapBERT will appear in the conference proceedings of NAACL 2021 ! SapBERT PubMedBERT SapBERT by Liu et al. (2020). Trained with UMLS 2020AA (English only), using microsoft/BiomedNLP PubMedBERT base uncased abstract fulltext as the base model. Expected input and output The input should be a string of biomedical entity names, e.g., "covid infection" or "Hydroxychloroquine". The [CLS] embedding of the last layer is regarded as the output. Extracting embeddings from SapBERT The following script converts a list of strings (entity names) into embeddings. For more details about training and eval, see SapBERT github repo. Citation
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy