Gen-Z: Generative Zero-Shot Text Classification with Contextualized Label Descriptions

Kumar, Sachin; Park, Chan_Young; Tsvetkov, Yulia

Citation Details

Language model (LM) prompting—a popular paradigm for solving NLP tasks—has been shown to be susceptible to miscalibration and brittleness to slight prompt variations, caused by its discriminative prompting approach, i.e., predicting the label given the input. To address these issues, we propose Gen-Z—a generative prompting framework for zero-shot text classification. GEN-Z is generative, as it measures the LM likelihood of input text, conditioned on natural language descriptions of labels. The framework is multivariate, as label descriptions allow us to seamlessly integrate additional contextual information about the labels to improve task performance. On various standard classification benchmarks, with six open-source LM families, we show that zero-shot classification with simple contextualization of the data source of the evaluation set consistently outperforms both zero-shot and few-shot baselines while improving robustness to prompt variations. Further, our approach enables personalizing classification in a zero-shot manner by incorporating author, subject, or reader information in the label descriptions. more »

Award ID(s):: 2142739 2203097 2125201

PAR ID:: 10520221

Author(s) / Creator(s):: Kumar, Sachin; Park, Chan_Young; Tsvetkov, Yulia

Publisher / Repository:: International Conference on Learning Representations

Date Published:: 2024-05-15

Format(s):: Medium: X

Sponsoring Org:: National Science Foundation

Free Publicly Accessible Full Text
Accepted Manuscript1.0
Conference Paper:
The DOI is not currently available.

More Like this