PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
October 6, 20203 citationsOpen Access

GRUEN for Evaluating Linguistic Quality of Generated Text

View Full Paper
WZWanzheng ZhuSBSuma Bhat

Key Points

Key points are not available for this paper at this time.

Abstract

Automatic evaluation metrics are indispensable for evaluating generated text. To date, these metrics have focused almost exclusively on the content selection aspect of the system output, ignoring the linguistic quality aspect altogether. We bridge this gap by proposing GRUEN for evaluating Grammaticality, non-Redundancy, focUs, structure and coherENce of generated text. GRUEN utilizes a BERT-based model and a class of syntactic, semantic, and contextual features to examine the system output. Unlike most existing evaluation metrics which require human references as an input, GRUEN is reference-less and requires only the system output. Besides, it has the advantage of being unsupervised, deterministic, and adaptable to various tasks. Experiments on seven datasets over four language generation tasks show that the proposed metric correlates highly with human judgments.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Zhu et al. (2020) studied this question.

synapsesocial.com/papers/6a1bcc07c97d63156a5ef3e7https://doi.org/10.48550/arxiv.2010.02498
Ask AI
Helpful
Bookmark
Share
View Full Paper

Also Consider

Synapse has enriched 5 closely related papers on similar clinical questions. Consider them for comparative context:

  1. 1Abstractive Text Summarization using Sequence-to-sequence RNNs and Beyond2016 · 2,234 citations
  2. 2Sentence Mover’s Similarity: Automatic Evaluation for Multi-Sentence Texts2019 · 143 citations
  3. 3Meteor++ 2.0: Adopt Syntactic Level Paraphrase Knowledge into Machine Translation Evaluation2019 · 31 citations
  4. 4SUM-QE: a BERT-based Summary Quality Estimation Model2019 · 43 citations
  5. 5Blend: a Novel Combined MT Metric Based on Direct Assessment — CASICT-DCU submission to WMT17 Metrics Task2017 · 48 citations