PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
April 25, 20260 citationsOpen Access

The Blind Spots in Automated Feedback Generation for Academic Writing

TSToru SasakiEindhoven University of TechnologyRCRianne ConijnEindhoven University of TechnologyMWMartijn C. WillemsenHuman Computer Interaction (Switzerland)

Key Points

Key points are not available for this paper at this time.

Abstract

Machine learning-based automated essay scoring (AES) and feedback generation (AFG) tools have been developed since the 1960s, with some commercially deployed. Such systems are expected to lighten the labor-intensive task of essay scoring and help provide timely feedback to students. However, these tools have not been used effectively enough in practice despite the long history of this research field. We aim to determine how accurately currently available models can make the necessary corrections and detect improvable segments of text. Latest attempts of AFG utilize state-of-the-art generative large language models (LLMs) for writing evaluation. To the best of our knowledge, however, none of the past studies have included generative LLM-based models in a fine-grained sentence-to-sentence comparison with human feedback or among AI tools. To fill these gaps, we conduct an experimental comparison of human feedback and three AI tools developed at different stages of technological advancement. Findings indicate that the overlap between human and AI feedback is predominantly limited to surface-level linguistic features and that generative AI-augmented tools demonstrate a markedly higher capability than a tool based on conventional rule-based AI.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Sasaki et al. (2026) studied this question.

synapsesocial.com/papers/6a08f91d1be1a34de49d0565https://doi.org/10.1145/3785022.3785120
Ask AI
Helpful
Bookmark
Share
View Full Paper