What Comes Next? Evaluating Uncertainty in Neural Text Generators Against Human Production Variability

Name: What Comes Next? Evaluating Uncertainty in Neural Text Generators Against Human Production Variability
Keywords: Decoding algorithm, FOS: Computer and information sciences, Computer Science - Machine Learning, Computer Science - Computation and Language, Syntactic variability, Computer Science - Artificial Intelligence, Model calibration, Aleatoric uncertainty, Human production variability, Semantic variability

Mario Giulianelli; Joris Baan; Wilker Aziz; Raquel Fernández; Barbara Plank

Found an issue? Give us feedback

arXiv.org e-Print Ar...arrow_drop_down

arXiv.org e-Print Archive

Preprint . 2023

Data sources: arXiv.org e-Print Archive

Universiteit van Amsterdam (UvA) Institutional Repository UvA-DARE

Conference object . 2023

License: CC BY

Data sources: Universiteit van Amsterdam (UvA) Institutional Repository UvA-DARE

Universiteit van Amsterdam: Digital Academic Repository (UvA DARE)

Article . 2023

Data sources: Bielefeld Academic Search Engine (BASE)

Research database - IT-University of Copenhagen

Article . 2023

Data sources: Bielefeld Academic Search Engine (BASE)

https://doi.org/10.18653/v1/20...

Article . 2023 . Peer-reviewed

Data sources: Crossref

https://dx.doi.org/10.48550/ar...

Article . 2023

License: arXiv Non-Exclusive Distribution

Data sources: Datacite

DBLP

Conference object

Data sources: DBLP

DBLP

Article

Data sources: DBLP

http://dx.doi.org/10.18653/v1/...

Conference object . 2023

Data sources: European Union Open Data Portal

http://dx.doi.org/10.18653/v1/...

Conference object

Data sources: Sygma

What Comes Next? Evaluating Uncertainty in Neural Text Generators Against Human Production Variability

descriptionPublicationkeyboard_double_arrow_right Article , Preprint , Conference object 01 Jan 2023Embargo end date: 01 Jan 2023 Netherlands, Denmark Publisher:Association for Computational Linguistics (ACL)Journal:Proceedings of the 2023 Conference on Empirical Methods in Natural Language ProcessingFunded by:EC | DIALECT, EC | UTTER, EC | DREAM

Authors: Mario Giulianelli; Joris Baan; Wilker Aziz; Raquel Fernández; Barbara Plank;

doi: 10.18653/v1/2023.emnlp-main.887 , 10.48550/arxiv.2305.11707

arXiv: 2305.11707

handle: 11245.1/3b31f778-3bcc-43a5-a48d-429399ef37a8

What Comes Next? Evaluating Uncertainty in Neural Text Generators Against Human Production Variability

- Summary
- Subjects
- Related research
  (1)
- Metrics

Abstract

In Natural Language Generation (NLG) tasks, for any input, multiple communicative goals are plausible, and any goal can be put into words, or produced, in multiple ways. We characterise the extent to which human production varies lexically, syntactically, and semantically across four NLG tasks, connecting human production variability to aleatoric or data uncertainty. We then inspect the space of output strings shaped by a generation system's predicted probability distribution and decoding algorithm to probe its uncertainty. For each test input, we measure the generator's calibration to human production variability. Following this instance-level approach, we analyse NLG models and decoding strategies, demonstrating that probing a generator with multiple samples and, when possible, multiple references, provides the level of detail necessary to gain understanding of a model's representation of uncertainty. Code available at https://github.com/dmg-illc/nlg-uncertainty-probes.

Camera ready version for EMNLP 2023

Countries

Netherlands, Denmark

Related Organizations

IT University of Copenhagen
Denmark
Ludwig-Maximilians-Universität München
Germany
University of Copenhagen
Denmark
University of Amsterdam
Netherlands

Keywords

Decoding algorithm, FOS: Computer and information sciences, Computer Science - Machine Learning, Computer Science - Computation and Language, Syntactic variability, Computer Science - Artificial Intelligence, Model calibration, Aleatoric uncertainty, Human production variability, Semantic variability, 004, 620, Machine Learning (cs.LG), Natural Language Generation, Artificial Intelligence (cs.AI), Uncertainty representation, Lexical variability, Communicative goals, Computation and Language (cs.CL)

1 Research products, page 1 of 1

nlg-uncertainty-probes software on GitHub
IsRelatedTo

Impact byBIP!

	selected citations These citations are derived from selected sources. This is an alternative to the "Influence" indicator, which also reflects the overall/total impact of an article in the research community at large, based on the underlying citation network (diachronically).	3
	popularity This indicator reflects the "current" impact/attention (the "hype") of an article in the research community at large, based on the underlying citation network.	Top 10%
	influence This indicator reflects the overall/total impact of an article in the research community at large, based on the underlying citation network (diachronically).	Average
	impulse This indicator reflects the initial momentum of an article directly after its publication, based on the underlying citation network.	Average

Found an issue? Give us feedback

3

Top 10%

Average

Green

Funded by

EC| DIALECT, EC| UTTER, EC| DREAM

Related to Research communities

Netherlands Research Portal

UArctic

What Comes Next? Evaluating Uncertainty in Neural Text Generators Against Human Production Variability

What Comes Next? Evaluating Uncertainty in Neural Text Generators Against Human Production Variability

1 Research products, page 1 of 1

nlg-uncertainty-probes software on GitHub