arrow
返回

Virtual prompt pre-training for prototype-based few-shot relation extraction

delete2023-03-01
delete34
delete
OA
AI
K
Kai He
Y
Yucheng Huang
R
Rui Mao
T
Tieliang Gong
C
Chen Li
E
Erik Cambria *
DOI:10.1016/j.eswa.2022.118927delete
delete原文链接
delete分享
delete收藏
查看原文
摘要

摘要

En 中文
Prompt tuning with pre-trained language models (PLM) has exhibited outstanding performance by reducing the gap between pre-training tasks and various downstream applications, which requires additional labor efforts in label word mappings and prompt template engineering. However, in a label intensive research domain, e.g., few-shot relation extraction (RE), manually defining label word mappings is particularly challenging, because the number of utilized relation label classes with complex relation names can be extremely large. Besides, the manual prompt development in natural language is subjective to individuals. To tackle these issues, we propose a virtual prompt pre-training method, projecting the virtual prompt to latent space, then fusing with PLM parameters. The pre-training is entity-relation-aware for RE, including the tasks of mask entity prediction, entity typing, distant supervised RE, and contrastive prompt pre-training. The proposed pre-training method can provide robust initialization for prompt encoding, while maintaining the interaction with the PLM. Furthermore, the virtual prompt can effectively avoid the labor efforts and the subjectivity issue in label word mapping and prompt template engineering. Our proposed prompt-based prototype network delivers a novel learning paradigm to model entities and relations via the probability distribution and Euclidean distance of the predictions of query instances and prototypes. The results indicate that our model yields an averaged accuracy gain of 4.21% on two few-shot datasets over strong RE baselines. Based on our proposed framework, our pre-trained model outperforms the strongest RE-related PLM by 6.52%.
Keyword:
Few-shot learning
Information extraction
Prompt tuning
Pre-trained Language Model
AI总结

AI总结

对已上传原文的论文进行重点信息的提取,主要内容包括:简要概述、研究摘要、背景介绍、关键亮点、图文解析、展望与总结。

期刊

Expert Systems with Applications 封面图
Expert Systems with Applications
IF:
7.5
论文数:
3.0W
被引数:
10.2W

机构

X
xi'an jiaotong university
学者数:
9.3W
论文数: 6.7W
被引数: 75
N
Nanyang Technological University
学者数:
4.9W
论文数: 4.8W
被引数: 8.1W
引用论文

引用论文

Enhanced prototypical network for few-shot relation extraction
err2021-07-01
err31
PREAI
errWen, Wen; Liu, Yongbin; Ouyang, Chunping; Lin, Qiang; Chung, Tonglee
err分享
err收藏
MetaPro: A computational metaphor processing model for text pre-processing
err2022-10-01
err53
PREAI
errMao, Rui; Li, Xiao; Ge, Mengshi; Cambria, Erik
err分享
err收藏
Database resources of the national center for biotechnology information国家生物技术信息中心数据库资源
err2007-12-23
err641
errOAAI
errWheeler, David L.; Barrett, Tanya; Benson, Dennis A.; Bryant, Stephen H.; Canese, Kathi; Chetvernin, Vyacheslav; Church, Deanna M.; DiCuccio, Michael; Edgar, Ron; Federhen, Scott; Feolo, Michael; Geer, Lewis Y.; Helmberg, Wolfgang; Kapustin, Yuri; Khovayko, Oleg; Landsman, David; Lipman, David J.; Madden, Thomas L.; Maglott, Donna R.; Miller, Vadim; Ostell, James; Pruitt, Kim D.; Schuler, Gregory D.; Shumway, Martin; Sequeira, Edwin; Sherry, Steven T.; Sirotkin, Karl; Souvorov, Alexandre; Starchenko, Grigory; Tatusov, Roman L.; Tatusova, Tatiana A.; Wagner, Lukas; Yaschenko, Eugene
err分享
err收藏
Improving Zero-Shot Learning Baselines with Commonsense Knowledge
err2022-07-14
err16
PREAI
errRoy, Abhinaba; Ghosal, Deepanway; Cambria, Erik; Majumder, Navonil; Mihalcea, Rada; Poria, Soujanya
err分享
err收藏
A Deep Look into neural ranking models for information retrieval
err2020-11-01
err180
errOAAI
errGuo, Jiafeng; Fan, Yixing; Pang, Liang; Yang, Liu; Ai, Qingyao; Zamani, Hamed; Wu, Chen; Croft, W. Bruce; Cheng, Xueqi
err分享
err收藏
没有更多内容