Eliciting Knowledge from Pretrained Language Models for Prototypical Prompt Verbalizer

descriptionPublicationkeyboard_double_arrow_right Part of book or chapter of book , Article , Preprint 01 Jan 2022Embargo end date: 01 Jan 2022 English Publisher:Springer Nature Switzerland

Authors: Wei, Yinyi; Mo, Tong; Jiang, Yongtao; Li, Weiping; Zhao, Wen;

doi: 10.1007/978-3-031-15931-2_19 , 10.48550/arxiv.2201.05411

arXiv: 2201.05411

Eliciting Knowledge from Pretrained Language Models for Prototypical Prompt Verbalizer

- Summary
- Subjects
- Related research
  (11)
- Metrics

Abstract

Recent advances on prompt-tuning cast few-shot classification tasks as a masked language modeling problem. By wrapping input into a template and using a verbalizer which constructs a mapping between label space and label word space, prompt-tuning can achieve excellent results in zero-shot and few-shot scenarios. However, typical prompt-tuning needs a manually designed verbalizer which requires domain expertise and human efforts. And the insufficient label space may introduce considerable bias into the results. In this paper, we focus on eliciting knowledge from pretrained language models and propose a prototypical prompt verbalizer for prompt-tuning. Labels are represented by prototypical embeddings in the feature space rather than by discrete words. The distances between the embedding at the masked position of input and prototypical embeddings are used as classification criterion. For zero-shot settings, knowledge is elicited from pretrained language models by a manually designed template to form initial prototypical embeddings. For few-shot settings, models are tuned to learn meaningful and interpretable prototypical embeddings. Our method optimizes models by contrastive learning. Extensive experimental results on several many-class text classification datasets with low-resource settings demonstrate the effectiveness of our approach compared with other verbalizer construction methods. Our implementation is available at https://github.com/Ydongd/prototypical-prompt-verbalizer.

Related Organizations

Peking University
China (People's Republic of)
PEKING UNIVERSITY
China (People's Republic of)
Peking University
China (People's Republic of)
Peking University
China (People's Republic of)
PEKING UNIVERSITY
China (People's Republic of)

View all View all

Keywords

FOS: Computer and information sciences, Computer Science - Computation and Language, Computation and Language (cs.CL)

11 Research products, page 1 of 2

Enhancing Entity Representations with Prompt Learning for Biomedical Entity Linking
2022IsAmongTopNSimilarDocuments
Automating Method Naming with Context-Aware Prompt-Tuning
2023IsAmongTopNSimilarDocuments
Input-Tuning: Adapting Unfamiliar Inputs to Frozen Pretrained Models
2022IsAmongTopNSimilarDocuments
Continuous Detection, Rapidly React: Unseen Rumors Detection based on Continual Prompt-Tuning
2022IsAmongTopNSimilarDocuments
PromptEM
2022IsAmongTopNSimilarDocuments
Prompt-tuned Code Language Model as a Neural Knowledge Base for Type Inference in Statically-Typed Partial Code
2022IsAmongTopNSimilarDocuments
PanDa: Prompt Transfer Meets Knowledge Distillation for Efficient Model Adaptation
2024IsAmongTopNSimilarDocuments
Knowledgeable Prompt-tuning: Incorporating Knowledge into Prompt Verbalizer for Text Classification
2022IsAmongTopNSimilarDocuments
Long: Knowledgeable Prompt-tuning: Incorporating Knowledge into Prompt Verbalizer for Text Classification
2022IsAmongTopNSimilarDocuments
CCPrefix: Counterfactual Contrastive Prefix-Tuning for Many-Class Classification
2024IsAmongTopNSimilarDocuments

chevron_left
1
2
chevron_right

Impact byBIP!

	selected citations These citations are derived from selected sources. This is an alternative to the "Influence" indicator, which also reflects the overall/total impact of an article in the research community at large, based on the underlying citation network (diachronically).	16
	popularity This indicator reflects the "current" impact/attention (the "hype") of an article in the research community at large, based on the underlying citation network.	Top 10%
	influence This indicator reflects the overall/total impact of an article in the research community at large, based on the underlying citation network (diachronically).	Top 10%
	impulse This indicator reflects the initial momentum of an article directly after its publication, based on the underlying citation network.	Top 10%

Found an issue? Give us feedback

16

Top 10%

Green

Fields of Science (4) View all

natural sciences

Fields of Science

natural sciences

View all