PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
February 22, 20247 citationsOpen Access

Using Large Language Models for Automated Grading of Student Writing about Science

View Full Paper
CIChris ImpeyUniversity of ArizonaMWMatthew WengerProvidence Health CareNGNikhil GarudaUniversity of Arizona

Key Points

Key points are not available for this paper at this time.

Abstract

Abstract A challenge in teaching large classes for formal or informal learners is assessing writing. As a result, most large classes, especially in science, use objective assessment tools like multiple choice quizzes. The rapid maturation of AI has created the possibility of using large language models (LLMs) to assess student writing. An experiment was carried out using GPT-3.5 and GPT-4 to see if machine learning methods based on LLMs can rival peer grading for reliability and automation in evaluating short writing assignments on topics in astronomy. The audience was lifelong learners in three massive open online courses (MOOCs) offered through Coursera. However, the results should also be applicable to non-science majors in university settings. The data was answers from 120 students on 12 questions across the three courses. The LLM was fed with total grades, model answers, and rubrics from an instructor for all three questions. In addition to seeing how reliably the LLMs reproduced instructor grades, the LLMs were asked to generate their own rubrics. Overall, the LLMs were more reliable than peer grading, both in the aggregate and by individual student, and they came much closer to the instructor grades for all three of the online courses. GPT-4 generally outperformed GPT-3.5. The implication is that LLMs can be used for automated, reliable, and scalable grading of student science writing.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Impey et al. (2024) studied this question.

synapsesocial.com/papers/68e781e8b6db6435876f4c0fhttps://doi.org/10.21203/rs.3.rs-3962175/v1
Ask AI
Helpful
Bookmark
Share
View Full Paper

Also Consider

Synapse has enriched 5 closely related papers on similar clinical questions. Consider them for comparative context:

  1. 1ChatGPT for Good? On Opportunities and Challenges of Large Language Models for Education2023 · 535 citations
  2. 2The social and ethical impacts of artificial intelligence in agriculture: mapping the agricultural AI literature2022 · 171 citations
  3. 3Academic dishonesty and trustworthy assessment in online learning: A systematic literature review2022 · 125 citations
  4. 4Online learner engagement: Conceptual definitions, research themes, and supportive practices2022 · 214 citations
  5. 5A Modified Claim, Evidence, Reasoning Organizer to Support Writing in the Science Classroom2023 · 4 citations