UNION: An Unreferenced Metric for Evaluating Open-ended Story Generation

Guan, Jian; Huang, Minlie

Computer Science > Computation and Language

arXiv:2009.07602 (cs)

[Submitted on 16 Sep 2020]

Title:UNION: An Unreferenced Metric for Evaluating Open-ended Story Generation

Authors:Jian Guan, Minlie Huang

View PDF

Abstract:Despite the success of existing referenced metrics (e.g., BLEU and MoverScore), they correlate poorly with human judgments for open-ended text generation including story or dialog generation because of the notorious one-to-many issue: there are many plausible outputs for the same input, which may differ substantially in literal or semantics from the limited number of given references. To alleviate this issue, we propose UNION, a learnable unreferenced metric for evaluating open-ended story generation, which measures the quality of a generated story without any reference. Built on top of BERT, UNION is trained to distinguish human-written stories from negative samples and recover the perturbation in negative stories. We propose an approach of constructing negative samples by mimicking the errors commonly observed in existing NLG models, including repeated plots, conflicting logic, and long-range incoherence. Experiments on two story datasets demonstrate that UNION is a reliable measure for evaluating the quality of generated stories, which correlates better with human judgments and is more generalizable than existing state-of-the-art metrics.

Comments:	Long paper; Accepted by EMNLP2020
Subjects:	Computation and Language (cs.CL)
Cite as:	arXiv:2009.07602 [cs.CL]
	(or arXiv:2009.07602v1 [cs.CL] for this version)
	https://doi.org/10.48550/arXiv.2009.07602

Submission history

From: Jian Guan [view email]
[v1] Wed, 16 Sep 2020 11:01:46 UTC (794 KB)

Full-text links:

Access Paper:

view license

Current browse context:

cs.CL

< prev | next >

new | recent | 2020-09

Change to browse by:

References & Citations

DBLP - CS Bibliography

listing | bibtex

Jian Guan
Minlie Huang

export BibTeX citation

Computer Science > Computation and Language

Title:UNION: An Unreferenced Metric for Evaluating Open-ended Story Generation

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computation and Language

Title:UNION: An Unreferenced Metric for Evaluating Open-ended Story Generation

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators