REVIEW 3 cited by
Uncertainty in Natural Language Generation: From Theory to Applications
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Recent advances of powerful Language Models have allowed Natural Language Generation (NLG) to emerge as an important technology that can not only perform traditional tasks like summarisation or translation, but also serve as a natural language interface to a variety of applications. As such, it is crucial that NLG systems are trustworthy and reliable, for example by indicating when they are likely to be wrong; and supporting multiple views, backgrounds and writing styles -- reflecting diverse human sub-populations. In this paper, we argue that a principled treatment of uncertainty can assist in creating systems and evaluation protocols better aligned with these goals. We first present the fundamental theory, frameworks and vocabulary required to represent uncertainty. We then characterise the main sources of uncertainty in NLG from a linguistic perspective, and propose a two-dimensional taxonomy that is more informative and faithful than the popular aleatoric/epistemic dichotomy. Finally, we move from theory to applications and highlight exciting research directions that exploit uncertainty to power decoding, controllable generation, self-assessment, selective answering, active learning and more.
Forward citations
Cited by 3 Pith papers
-
Large Language Models Can Be a Viable Substitute for Expert Political Surveys When a Shock Disrupts Traditional Measurement Approaches
LLM-based pairwise comparisons can recover pre-shock perceptions of federal agencies, including a new knowledge-institution measure that predicts DOGE layoffs.
-
Efficient Hallucination Detection for LLMs Using Uncertainty-Aware Attention Heads
RAUQ detects hallucinated LLM output by selecting one attention head per layer (the one that most attends to the preceding token), propagating that attention together with token probabilities, and taking the maximum u...
-
Position: Uncertainty Quantification Needs Reassessment for Large-language Model Agents
A position paper arguing that aleatoric/epistemic uncertainty splits fail for LLM agents and proposing underspecification, interaction, and output-based uncertainty research.
Discussion (0). Continue with ORCID to comment.