Abstract
Quantifying uncertainty in automatically generated text is important for letting humans check potential hallucinations and making systems more reliable. Conformal prediction is an attractive framework to provide predictions imbued with statistical guarantees, however, its application to text generation is challenging since any i.i.d. assumptions are not realistic. In this paper, we bridge this gap by leveraging recent results on *non-exchangeable* conformal prediction, which still ensures bounds on coverage. The result, *non-exchangeable conformal nucleus sampling*, is a novel extension of the conformal prediction framework to generation based on nearest neighbors. Our method can be used post-hoc for an arbitrary model without extra training and supplies token-level, calibrated prediction sets equipped with statistical guarantees. Experiments in machine translation and language modeling show encouraging results in generation quality. By also producing tighter prediction sets with good coverage, we thus give a more theoretically principled way to perform sampling with conformal guarantees.
| Original language | English |
|---|---|
| Title of host publication | Findings of the Association for Computational Linguistics: EACL 2024 |
| Editors | Yvette Graham, Matthew Purver |
| Number of pages | 20 |
| Volume | EACL |
| Publisher | Association for Computational Linguistics |
| Publication date | 17 Mar 2024 |
| Edition | 2024 |
| Pages | 1909-1929 |
| Publication status | Published - 17 Mar 2024 |
| Event | Conference of the European Chapter of the Association for Computational Linguistics - St. Julian's, Malta Duration: 17 Mar 2024 → 22 Mar 2024 Conference number: 18 https://dblp.org/db/conf/eacl/eacl2024-2.html https://aclanthology.org/volumes/2024.eacl-long/ https://dblp.org/db/conf/eacl/eacl2024f.html |
Conference
| Conference | Conference of the European Chapter of the Association for Computational Linguistics |
|---|---|
| Number | 18 |
| Country/Territory | Malta |
| City | St. Julian's |
| Period | 17/03/2024 → 22/03/2024 |
| Internet address |
Keywords
- Uncertainty Quantification
- Text Generation
- Conformal Prediction
- Non-Exchangeable Conformal Nucleus Sampling
- Statistical Guarantees
Fingerprint
Dive into the research topics of 'Non-Exchangeable Conformal Language Generation with Nearest Neighbors'. Together they form a unique fingerprint.Cite this
- APA
- Author
- BIBTEX
- Harvard
- Standard
- RIS
- Vancouver