evidoria

← Back to browse

Good practice Imported

AIGC-Assisted Critical Writing Feedback — A Double-Blind Randomised Controlled Trial at Fudan University

China · Shanghai · See the China profile · See the Shanghai profile

Evidence: Randomised controlled trial Top 82% 33/100 · Ask Evidence Copilot about this practice

A double-blind, pre–post RCT with 259 Fudan University students found a fine-tuned Qwen-based AI feedback system significantly improved undergraduate critical-writing quality versus instructor-only feedback, with the largest gains in essay organisation (β=0.311, p<0.001).

259 undergraduates
Study sample size
125 students
Treatment group size
134 students
Control group size
β=0.149, p<0.001
Overall treatment effect on writing performance
β=0.311, p<0.001
Effect on essay organisation
β=0.191, p<0.001
Effect on content development
4 weeks
Intervention duration

Details

Maturity
Pilot
Promoter
Fudan University — Informatization Office
Period
2024–2025
Keywords
Higher education, writing instruction, AI feedback systems

Context

Researchers at Fudan University's Informatization Office ran a double-blind, pre–post randomised controlled trial with 259 eligible undergraduates (125 treatment, 134 control), screened for native Mandarin proficiency, intermediate English competency (CET-4 425–550) and no prior AI-writing-tool experience.

Objectives

The trial tested whether AI-generated feedback could improve undergraduate critical-writing quality compared with traditional instructor-only feedback.

Activities

The treatment group received feedback from a fine-tuned Qwen-7B-Chat model customised for academic-writing evaluation over a four-week intervention; the control group received traditional instructor-only feedback.

Results

Using difference-in-differences analysis and structural equation modelling, the study found a significant overall treatment-by-time effect on writing performance (β=0.149, p<0.001). Effects were strongest on essay organisation (β=0.311, p<0.001) and content development (β=0.191, p<0.001), with smaller but significant gains in language usage (β=0.070, p<0.05) and self-reported writing motivation (β=0.077, p<0.05). A technology-acceptance pathway analysis further linked perceived ease of use to perceived usefulness (β=0.326, p<0.001) and usefulness/attitude to actual system usage (β=0.431, p<0.001).

Conclusions

The trial was conducted at a single institution over one four-week cycle, so evidence of durability or transfer beyond this cohort and course is not yet available; the study itself does not report any classroom-scale rollout beyond the trial.

Implementation

Indicative cost
Low (< €50k)
Time to results
Short (< 1 year)
Staffing & skills
Fudan University Informatization Office (research team)

Conditions for success

  • Fine-tuned domain-specific LLM (Qwen-7B-Chat) rather than a generic model
  • Rigorous eligibility screening and randomisation protocol
  • Double-blind trial design to reduce bias

Do you run this practice? Claim it — verified implementers get a public contact pathway and can propose corrections.

Data sources

Where this practice's information was retrieved from, and when.

Similar practices you may find useful