计算机科学
关系抽取
规范化(社会学)
人工智能
代表(政治)
图形
自然语言处理
帕金森病
机器学习
信息抽取
疾病
理论计算机科学
医学
病理
社会学
政治
人类学
政治学
法学
作者
Xiaoming Zhang,Can Yu,Rui Yan
标识
DOI:10.1016/j.jbi.2024.104624
摘要
The relational triple extraction of unstructured medical texts about Parkinson's disease is critical for the construction of a medical knowledge graph. However, the triple entities in Parkinson's disease are usually complicated and overlapped, which impedes the accuracy of triple extraction, especially in the case of rarely available corpus. Therefore, this study first builds a corpus about Parkinson's disease. Then, a tagging-based three-stage relational triple extraction model is proposed, named ParTRE. To enhance the contextual representation of sentences, the proposed model employs BiLSTM modules to capture fine-grained semantic information. Additionally, a conditional normalization layer is used so that entity pairs can be extracted accurately from two complementary directions. As for the imbalanced relationship categories, an adaptive loss function strategy based on focal loss is derived by assigning different weights to relationship categories and reducing the loss of easy-to-classify samples. The model performance is evaluated on the Parkinson's corpus and public datasets. The results indicate that the proposed model achieves an overall F1-score of 93.3 % on the Parkinson's corpus and comparable performance on public datasets compared with the state-of-the-art methods. Moreover, a satisfactory result is achieved by the proposed model on conquering the overlapped entities and imbalanced relationship categories. Owing to demonstrated availability and validity, the proposed method can be integrated with medical knowledge graphs and therefore benefits medical intelligence.
科研通智能强力驱动
Strongly Powered by AbleSci AI