PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
March 29, 2026Scientific Reports0 citationsOpen Access

Enhancing IELTS writing automated scoring with M-LoRA fine-tuned LLAMA-3 and human feedback-driven PPO reinforcement learning

WXWenbo XuMKM. S. S. KassimRMRohana Mahmud

Key Points

  • The study aims to enhance automated essay scoring (AES) and feedback quality for IELTS writing, addressing current limitations of scoring models.
  • Utilized a private dataset of 5,088 IELTS essays with expert feedback for training.
  • Employed multi-task supervised fine-tuning of the LLaMA-3 model across four IELTS scoring dimensions.
  • Developed a reward model to enhance feedback generation quality.
  • Applied reinforcement learning based on fine-grained human feedback for further refining feedback.
  • Achieved significant improvements in essay scoring accuracy.
  • Enhanced the quality of personalized feedback for essays.
  • Showcased practical applications of the enhanced model in educational settings.

Abstract

This paper proposes an innovative automated essay scoring(AES) and feedback generation method based on the LLaMA-3 model and Multi-task LoRA (M-LoRA) fine-tune technology, aimed at improving the accuracy of IELTS essay scoring and the quality of personalized feedback generation. Our approach consists of three key stages: multi-task supervised fine-tuning of LLaMA-3 model, designing of the reward model, and reinforcement learning model training based on fine-grained human feedback. Before the experiment, we collected a private dataset of 5,088 IELTS essays with expert-annotated feedback and used this dataset to train and fine-tune the entire model. Firstly, through multi-task supervised fine-tuning, we successfully captured features efficiently across the four key dimensions of IELTS essay scoring: Task Response, Coherence and Cohesion, Lexical Resource, and Grammatical Range and Accuracy, effectively addressing the issue of catastrophic forgetting in scoring tasks. Secondly, we designed and trained a reward model to optimize the ability to generate feedback by scoring the quality of the feedback. Finally, we further fine-tuned the generated feedback using a reinforcement learning strategy model based on fine-grained human feedback, making the feedback more refined and personalized. Our findings demonstrate significant improvements in both essay scoring and feedback generation, showcasing practical applications in real-world educational settings. This research highlights the limitations of current large language models in grasping the complexities of essay scoring, emphasizing the need for more effective methods like ours to advance this field.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Xu et al. (2026) studied this question.

synapsesocial.com/papers/69c8c2d1de0f0f753b39d3a8https://doi.org/10.1038/s41598-026-43318-w
Ask AI
Helpful
Bookmark
Share
View Full Paper

Also Consider

Synapse has enriched 5 closely related papers on similar clinical questions. Consider them for comparative context:

  1. 1Natural Language Processing: Recent Development and Applications2023 · 8 citations
  2. 2Let Their Voices be Heard: IELTS Candidates’ Problems with the IELTS Academic Writing Test2023 · 6 citations
  3. 3An unsupervised approach to automated selection of good essays2011 · 9 citations
  4. 4N-Gram Based Approach for Automatic Prediction of Essay Rubric Marks2018 · 5 citations
  5. 5The State of the Art of Natural Language Processing—A Systematic Automated Review of NLP Literature Using NLP Techniques2023 · 47 citations