PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
January 1, 202117 citationsOpen Access

What happens if you treat ordinal ratings as interval data? Human evaluations in NLP are even more under-powered than you think

DHDavid M. HowcroftVRVerena Rieser

Key Points

Key points are not available for this paper at this time.

Abstract

Previous work has shown that human evaluations in NLP are notoriously under-powered.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Howcroft et al. (2021) studied this question.

synapsesocial.com/papers/69d6bceff174babf6cab3552https://doi.org/10.18653/v1/2021.emnlp-main.703
Ask AI
Helpful
Bookmark
Share
View Full Paper