Compression, Transduction, and Creation: A Unified Framework for Evaluating Natural Language Generation

Deng, Mingkai; Tan, Bowen; Liu, Zhengzhong; Xing, Eric P.; Hu, Zhiting

Computer Science > Computation and Language

arXiv:2109.06379v1 (cs)

[Submitted on 14 Sep 2021 (this version), latest version 21 Jan 2022 (v2)]

Title:Compression, Transduction, and Creation: A Unified Framework for Evaluating Natural Language Generation

Authors:Mingkai Deng, Bowen Tan, Zhengzhong Liu, Eric P. Xing, Zhiting Hu

View PDF

Abstract:Natural language generation (NLG) spans a broad range of tasks, each of which serves for specific objectives and desires different properties of generated text. The complexity makes automatic evaluation of NLG particularly challenging. Previous work has typically focused on a single task and developed individual evaluation metrics based on specific intuitions. In this paper, we propose a unifying perspective based on the nature of information change in NLG tasks, including compression (e.g., summarization), transduction (e.g., text rewriting), and creation (e.g., dialog). Information alignment between input, context, and output text plays a common central role in characterizing the generation. With automatic alignment prediction models, we develop a family of interpretable metrics that are suitable for evaluating key aspects of different NLG tasks, often without need of gold reference data. Experiments show the uniformly designed metrics achieve stronger or comparable correlations with human judgement compared to state-of-the-art metrics in each of diverse tasks, including text summarization, style transfer, and knowledge-grounded dialog.

Comments:	EMNLP 2021, Code available at this https URL
Subjects:	Computation and Language (cs.CL); Machine Learning (cs.LG)
Cite as:	arXiv:2109.06379 [cs.CL]
	(or arXiv:2109.06379v1 [cs.CL] for this version)
	https://doi.org/10.48550/arXiv.2109.06379

Submission history

From: Mingkai Deng [view email]
[v1] Tue, 14 Sep 2021 01:00:42 UTC (1,225 KB)
[v2] Fri, 21 Jan 2022 23:29:22 UTC (1,225 KB)

Computer Science > Computation and Language

Title:Compression, Transduction, and Creation: A Unified Framework for Evaluating Natural Language Generation

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computation and Language

Title:Compression, Transduction, and Creation: A Unified Framework for Evaluating Natural Language Generation

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators