Relation Classification via Sequence Features and Bi-Directional LSTMs

来源 :Wuhan University Journal of Natural Sciences | 被引量 : 0次 | 上传用户:ncla02
下载到本地 , 更方便阅读
声明 : 本文档内容版权归属内容提供方 , 如果您对本文有版权争议 , 可与客服联系进行内容授权或下架
论文部分内容阅读
Structure features need complicated pre-processing, and are probably domain-dependent. To reduce time cost of pre-processing, we propose a novel neural network architecture which is a bi-directional long-short-term-memory recurrent-neural-network(Bi-LSTM-RNN) model based on low-cost sequence features such as words and part-of-speech(POS) tags, to classify the relation of two entities. First, this model performs bi-directional recurrent computation along the tokens of sentences. Then, the sequence is divided into five parts and standard pooling functions are applied over the token representations of each part. Finally, the token representations are concatenated and fed into a softmax layer for relation classification. We evaluate our model on two standard benchmark datasets in different domains, namely Sem Eval-2010 Task 8 and Bio NLP-ST 2016 Task BB3. In Sem Eval-2010 Task 8, the performance of our model matches those of the state-of-the-art models, achieving 83.0% in F1. In Bio NLP-ST 2016 Task BB3, our model obtains F_1 51.3% which is comparable with that of the best system. Moreover, we find that the context between two target entities plays an important role in relation classification and it can be a replacement of the shortest dependency path. To reduce the cost of pre-processing, we propose a novel neural network architecture which is a bi-directional long-short-term-memory recurrent-neural-network Bi-LSTM-RNN) model based on low-cost sequence features such as words and part-of-speech (POS) tags, to classify the relation of two entities. First, this model performs bi-directional recurrent computation along the tokens of Finally, the token representations are concatenated and fed into a softmax layer for relation classification. We evaluate our model on two standard benchmark datasets in different domains, namely Sem Eval-2010 Task 8 and Bio NLP-ST 2016 Task BB3. In Sem Eval-2010 Task 8, the performance of our model matches those of the state-of-the-art models, achieving 83.0% in F1. In Bio NLP-ST 2016 Task BB3, our model obtains F_1 51.3% which is comparable with that of the best system. Moreover, we find that the context between two target entities plays an important role in relation classification and it can be a replacement of the shortest dependency path.
其他文献
铜山早薹韭是江苏省徐州市铜山县早薹韭研究会会长、高级农艺师杨建民经多年繁育,从地方品种的变异株中精心选育的特早熟薹韭良种。该品种在我地试种几年来,效益可观,一般露
21世纪,世界进入到了一个信息社会,信息产业的蓬勃发展使世界联系得更加紧密。在世界经济国际化、一体化、全球化的发展趋势下,国际经济合作已成为各国发展进步的必由之路。在中
当前,我国进行的经济结构调整,对产业技术进行的升级已经进入了关键阶段。在这种形势下,如何提高社会的自主创新能力,发挥知识产权制度对创新的保障和激励作用,增强整个社会
该文从挂篮荷载计算、施工流程、支座及临时固结施工、挂篮安装及试验、合拢段施工、模板制作安装、钢筋安装、混凝土的浇筑及养生、测量监控等方面人手,介绍了S226海滨大桥
目的:探讨绝经后取宫内节育器(IUD)中米索前列醇片配伍盐酸丁卡因胶浆的应用价值。方法:对68例绝经后要求取宫内节育器妇女患者随机分成治疗组35例,术前2小时引导后穹窿放置
我国棉花良繁技术的理论基础在于保持原品种种性 ,而做法表现在保持和选择两个方面。本文对长期沿用前苏联的“三圃制”及类似技术进行了分析 ,并探讨了利用四级种子生产程序
浦江县森林公安在县委、县政府的正确领导下,在上级林业、公安部门的关心支持下,深入开展“两学一做”活动,通过抓队伍、强素质、提能力,持续开展保护林区生态、服务林业改革等各项工作,为林业生态建设、林区治安稳定翻开了崭新的一页。  精准突破 剿猎行动成效卓著  2017年4月2日,郑宅镇堂头村的骆奶奶在村后山采药时被野猪夹所夹;4月8日,檀溪镇村民陈大伯在山上采观音柴叶被野猪夹夹伤;4月11日,前吴乡
期刊
当代学校教育注重的是全面发展,不单单追求分数.体育做为学校教育的一门学科,受到各方的重视.一方面,体育对个人而言有重要意义,另一方面,体育对全民体育、终身体育等国家政
期刊
云南省昆明市嵩明县牛栏江镇(以下简称“牛栏江镇”)地处嵩明县东部,总面积228.8 km2,下辖16个村委会106个自然村,总人口5.5597万人。辖区内第一产业占56%,第二产业占31%,第
期刊