The aligned corpus constructed using the knowledge-anchored method is combined with a multi task training strategy to continue training XLM-R, thus obtaining KBioXLM. It is the first multilingual biomedical pre-trained language model we know that has cross-lingual understanding capabilities in medical domain. It was introduced in the paper
KBioXLM: A Knowledge-anchored Biomedical
Multilingual Pretrained Language Model and released in
this repository.
KBioXLM model can be fintuned on downstream tasks. The downstream tasks here refer to biomedical cross-lingual understanding tasks, such as biomedical entity recognition, biomedical relationship extraction and biomedical text classification.
1from transformers import RobertaModel
2model=RobertaModel.from_pretrained('ngwlh/KBioXLM')
Coming soon.