注释
计算机科学
自然语言处理
医学术语
术语
情报检索
命名实体识别
人工智能
语言学
哲学
经济
管理
任务(项目管理)
作者
Zhe Wang,Lihong Liu,Keyu Yao,Junhui Wang,Yan Zhu
标识
DOI:10.1109/bibm55620.2022.9994940
摘要
Objective: To construct a natural language processing (NLP) system focused on named entity recognition (NER) and semantic relation extraction (RE) of ancient Chinese medical books, it supports annotated corpora management and semantic knowledge retrieval. Methods: We integrate the 47 ontologies and terminologies as the terminology database. After that, we trained a preprocessing NER model using spaCy and used a hybrid approach combining automated annotation and manual review to annotate corpora of ancient Chinese medical books. Results: The semantic annotation system of Chinese ancient texts named traditional Chinese medicine - semantic annotation system (TCM-SAS), was constructed based on ontologies and terminologies. Annotations and knowledge retrieval of TCM's ancient texts were realized. Conclusion: TCM-SAS is a user-friendly semantic annotation system for ancient Chinese medical books that includes a large-scale manual annotation of TCM literature and semantic knowledge of TCM. TCM-SAS could provide users with two modes of automatic and manual NER and RE for ancient Chinese texts, as well as annotated entity and corpora management. Support the discovery of new knowledge from ancient Chinese medical texts in the future.
科研通智能强力驱动
Strongly Powered by AbleSci AI