arrow
返回

Eliciting knowledge from language models with automatically generated continuous prompts

delete2024-04-01
delete1
PRE
AI
Y
Yadang Chen *
G
Gang Yang
D
Duolin Wang
D
Dichao Li
DOI:10.1016/j.eswa.2023.122327delete
delete原文链接
delete原文求助
delete分享
delete收藏
摘要

摘要

En 中文
Pre-trained Language Models (PLMs) have demonstrated remarkable performance in Natural Language Under-standing (NLU) tasks, with continuous prompt-based fine-tuning further enhancing their capabilities. However, current methods rely on hand-crafted discrete prompts to initialize continuous prompts, which are sensitive to subtle changes and inherently limited by the constraints of natural language. To address these limitations, this study introduces an innovative AutoPrompt-based Prompt Tuning (APT) approach. APT optimizes the initialization of continuous prompts by employing a gradient-guided automatic search to generate ideal discrete templates and identify trigger tokens. As the semantic features are already captured from the target task dataset, the continuous parameters initialized by trigger tokens are highly relevant, providing a superior starting point for prompt-tuning. APT searches for optimal prompts across various NLU tasks, enabling the PLM to learn task-related knowledge effectively. The APT method significantly improves PLM performance in both few-shot and fully supervised settings, eliminating the need for extensive prompt engineering. In the knowledge exploration (Language Model Analysis (LAMA)) benchmark, APT achieved a remarkable 58.6% (P@1) performance without additional text, representing a 3.6% improvement over the previous best result. Additionally, APT outperformed state-of-the-art methods in the SuperGLUE benchmark.
Keyword:
Prompt learning
Initialization
Trigger token
Continuous parameters

期刊

Expert Systems with Applications 封面图
Expert Systems with Applications
IF:
7.5
论文数:
3.0W
被引数:
10.2W

机构

暂无机构信息
引用论文

引用论文

Evidence of synergy between Thy-1 and CD3/TCR complex in signal delivery to murine thymocytes for cell death.
err1991-08-15
err0
errOAAI
errI Nakashima; Y H Zhang; S M Rahman; T Yoshida; K Isobe; L N Ding; T Iwamoto; M Hamaguchi; H Ikezawa; R Taguchi
err分享
err收藏
Self-Supervised Learning: Generative or Contrastive自我监督学习: 生成性或对比性
err2021-01-01
err522
errOAAI
errLiu, Xiao; Zhang, Fanjin; Hou, Zhenyu; Mian, Li; Wang, Zhaoyu; Zhang, Jing; Tang, Jie
err分享
err收藏
A portable fluorescence-based recombinase polymerase amplification assay for the detection of mal secco disease by Plenodomus tracheiphilus
err2024-10-01
err0
errOAAI
errErmes Ivan Rovetto; Matteo Garbelotto; Salvatore Moricca; Marcos Amato; Federico La Spada; Santa Olga Cacciola
err分享
err收藏
SpanBERT: Improving Pre-training by Representing and Predicting Spans
err2020-12-01
err1.1K
errOAAI
errJoshi, Mandar; Chen, Danqi; Liu, Yinhan; Weld, Daniel S.; Zettlemoyer, Luke; Levy, Omer
err分享
err收藏
没有更多内容