李丽双

个人信息Personal Information

教授

博士生导师

硕士生导师

性别:女

毕业院校:大连理工大学

学位:博士

所在单位:计算机科学与技术学院

学科:计算机应用技术. 计算机软件与理论

办公地点:创新大厦A930

电子邮箱:lils@dlut.edu.cn

扫描关注

论文成果

当前位置: 中文主页 >> 科学研究 >> 论文成果

The Protein-Protein Interaction Extraction Based on Full Texts

点击次数:

论文类型:会议论文

发表时间:2014-01-01

收录刊物:CPCI-S、Scopus

页面范围:493-496

关键字:Protein-Protien Interaction; full text; feature selection; syntactic pattern; tree kernel

摘要:Protein-Protein Interaction (PPI) extraction from literatures is becoming a more and more significant task in the biomedical information extraction. Though many methods for PPI extraction have achieved promising results, they all concentrated on the abstracts of literatures rather than full texts. In this paper, we append full-text features, namely Location and Co-occurrence to extract PPIs from full texts. Location describes where the protein pair appears in the article. Co-occurrence is the frequency of each protein pair occurring in the article. In addition, syntactic patterns are extracted as features, and then feature selection is applied to improve the performance and reduce the dimension of feature vectors in SVM. Finally, the selected features are combined with two-level DET tree kernel. Experimental results show that the presented approach can achieve an F-score of 74.46% and an AUC of 78.50%.