PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
April 25, 20260 citationsOpen Access

The Blind Spots in Automated Feedback Generation for Academic Writing

TSToru SasakiRCRianne ConijnMWMartijn C. Willemsen

Key Points

Key points are not available for this paper at this time.

Abstract

Machine learning-based automated essay scoring (AES) and feedback generation (AFG) tools have been developed since the 1960s, with some commercially deployed. Such systems are expected to lighten the labor-intensive task of essay scoring and help provide timely feedback to students. However, these tools have not been used effectively enough in practice despite the long history of this research field. We aim to determine how accurately currently available models can make the necessary corrections and detect improvable segments of text. Latest attempts of AFG utilize state-of-the-art generative large language models (LLMs) for writing evaluation. To the best of our knowledge, however, none of the past studies have included generative LLM-based models in a fine-grained sentence-to-sentence comparison with human feedback or among AI tools. To fill these gaps, we conduct an experimental comparison of human feedback and three AI tools developed at different stages of technological advancement. Findings indicate that the overlap between human and AI feedback is predominantly limited to surface-level linguistic features and that generative AI-augmented tools demonstrate a markedly higher capability than a tool based on conventional rule-based AI.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Sasaki et al. (2026) studied this question.

synapsesocial.com/papers/6a08f91d1be1a34de49d0565https://doi.org/10.1145/3785022.3785120
Ask AI
Helpful
Bookmark
Share
View Full Paper

Also Consider

Synapse has enriched 5 closely related papers on similar clinical questions. Consider them for comparative context:

  1. 1Can Large Language Models Provide Feedback to Students? A Case Study on ChatGPT2023 · 264 citations
  2. 2The Effectiveness of Using Grammarly to Improve Students' Writing Skills2020 · 56 citations
  3. 3Large pre-trained language models contain human-like biases of what is right and wrong to do2022 · 295 citations
  4. 4A Neural Approach to Automated Essay Scoring2016 · 512 citations
  5. 5Did You Read the Instructions? Rethinking the Effectiveness of Task Definitions in Instruction Learning2023 · 8 citations