PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
November 22, 2020Computer Speech & Language171 citationsOpen Access

Human evaluation of automatically generated text: Current trends and best practice guidelines

View Full Paper
CLChris van der LeeAGAlbert GattEMEmiel van Miltenburg

Key Points

Key points are not available for this paper at this time.

Abstract

Currently, there is little agreement as to how Natural Language Generation (NLG) systems should be evaluated, with a particularly high degree of variation in the way that human evaluation is carried out. This paper provides an overview of how (mostly intrinsic) human evaluation is currently conducted and presents a set of best practices, grounded in the literature. These best practices are also linked to the stages that researchers go through when conducting an evaluation research (planning stage; execution and release stage), and the specific steps in these stages. With this paper, we hope to contribute to the quality and consistency of human evaluations in NLG.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Lee et al. (2020) studied this question.

synapsesocial.com/papers/69d6bceff174babf6cab3553https://doi.org/10.1016/j.csl.2020.101151
Ask AI
Helpful
Bookmark
Share
View Full Paper