计算机科学
命名实体识别
人工智能
自然语言处理
特征(语言学)
信息抽取
序列标记
词(群论)
递归(计算机科学)
依赖关系(UML)
条件随机场
特征提取
算法
语言学
经济
管理
哲学
任务(项目管理)
标识
DOI:10.1038/s41598-024-56166-3
摘要
Chinese is characterized by high syntactic complexity, chaotic annotation granularity, and slow convergence. Joint learning models can effectively improve the accuracy of Chinese Named Entity Recognition (NER), but they focus too much on local feature information and reduce the ability of long sequence feature extraction. To address the limitations of long sequence feature extraction ability, we propose a Chinese NER model called Incorporating Recurrent Cell and Information State Recursion (IRCSR-NER). The model integrates recurrent cells and information state recursion to improve the recognition ability of long entity boundaries. To solve the problem that Chinese and English have different focuses in syntactic analysis. We use the syntactic dependency approach to add lexical relationship information to sentences represented at the word level. The IRCSR-NER is applied to sequence feature extraction to improve the model efficiency and long-text feature extraction ability. The model captures contextual long-distance dependent information while focusing on local feature information. We evaluated our proposed model using four public datasets and compared it with other mainstream models. Experimental results demonstrate that our model outperforms traditional and mainstream models.
科研通智能强力驱动
Strongly Powered by AbleSci AI