Input-Tuning: Adapting Unfamiliar Inputs to Frozen Pretrained Models
Input-Tuning: Adapting Unfamiliar Inputs to Frozen Pretrained Models
复制标题
DOI:
10.48550/arxiv.2203.03131
复制
发表时间:
2022-03
期刊:
影响因子:
--
通讯作者:
Shengnan An;Yifei Li;Zeqi Lin;Qian Liu;Bei Chen-;Qiang Fu;Weizhu Chen;Nanning Zheng;Jian-Guang Lou
中科院分区:
文献类型:
--
作者:
Shengnan An;Yifei Li;Zeqi Lin;Qian Liu;Bei Chen-;Qiang Fu;Weizhu Chen;Nanning Zheng;Jian-Guang Lou
Recently the prompt-tuning paradigm has attracted significant attention. By only tuning continuous prompts with a frozen pre-trained language model (PLM), prompt-tuning takes a step towards deploying a shared frozen PLM to serve numerous downstream tasks. Although prompt-tuning shows good performance on certain natural language understanding (NLU) tasks, its effectiveness on natural language generation (NLG) tasks is still under-explored. In this paper, we argue that one of the factors hindering the development of prompt-tuning on NLG tasks is the unfamiliar inputs (i.e., inputs are linguistically different from the pretraining corpus). For example, our preliminary exploration reveals a large performance gap between prompt-tuning and fine-tuning when unfamiliar inputs occur frequently in NLG tasks. This motivates us to propose input-tuning, which fine-tunes both the continuous prompts and the input representations, leading to a more effective way to adapt unfamiliar inputs to frozen PLMs. Our proposed input-tuning is conceptually simple and empirically powerful. Experimental results on seven NLG tasks demonstrate that input-tuning is significantly and consistently better than prompt-tuning. Furthermore, on three of these tasks, input-tuning can achieve a comparable or even better performance than fine-tuning.