← Back to Research News
A teacher prompts a diverse group of secondary students to reason with question cards while an unused answer-generating tablet rests at the edge of the table
Conference PaperConference paper202618 Aug 2026· 2 min

Students' essays supported a scaffold-don't-substitute principle, but not a causal harm estimate

Lucile Favero, Juan Antonio Pérez-Ortiz, Tanja Käser, Nuria Oliver

ACM AI Leadership Summit

500-word summary

A teacher prompts a diverse group of secondary students to reason with question cards while an unused answer-generating tablet rests at the edge of the table

Listen to the paper summary

Audio summary

0:00/0:00

Lucile Favero, Juan Antonio Pérez-Ortiz, Tanja Käser and Nuria Oliver argue that educational AI becomes harmful when it substitutes for the human capacities education is intended to develop. Their framework connects four dimensions: cognition, agency, emotional wellbeing, and ethics. Cognitive offloading can reduce effort; reduced effort can weaken a learner's sense of control; dependence and uncertainty can affect wellbeing; and these changes can intensify questions about responsibility and fairness. The authors describe this as a self-reinforcing cycle rather than four isolated risks.

The paper grounds the framework in an exploratory analysis of 49 International Baccalaureate argumentative essays about AI's impact. Eighty percent of the essays reported that reliance on AI reduces thinking. Students also described a preferred alternative: systems that withhold immediate answers, prompt recall, ask questions, and encourage reflection. The authors connect these requests with established learning-science principles and summarize the resulting design stance as scaffold, do not substitute.

Student voice is a strength of the study. Learners are often treated as recipients of AI policy rather than contributors to product design. Their essays reveal concerns and desired interactions in their own educational context. Still, the number 80 percent must be interpreted narrowly. The sample is small, comes from a particular programme, and consists of essays written for an argument task. It does not measure actual tool use, cognitive change, achievement, mental health, or the prevalence of an effect across students. Coding self-reports cannot establish that AI caused the harms described.

The proposed design principle is therefore best treated as a hypothesis and a requirement for testing. A scaffolded tutor might ask a learner to retrieve an idea before showing a hint, request a prediction before a simulation, or provide a partial step followed by an explanation prompt. A substitutive tool might produce the completed essay or solution immediately. Researchers can compare these designs on unaided retention, transfer, help-seeking, confidence calibration, frustration, agency, accessibility, and time. They should also examine whether withholding help disadvantages learners who need accommodations or foundational instruction.

Teachers can apply the principle without banning AI. They can specify phases in which learners first attempt, explain, or retrieve; allow graduated hints; require source checks; and assess a later task without assistance. Students should help define when support feels productive and when it feels controlling or evasive. Product teams should expose hint policies and let educators align them with age, subject, and learner needs.

For AIEDHK, the paper supplies a memorable orientation while modelling evidential restraint. The student essays justify taking substitution risks seriously, but they do not quantify causal harm. The next step is participatory, prospective evaluation of designs that preserve effort without denying appropriate support. Educational AI should be judged by the capacities learners retain and can exercise independently, not by the amount of work the system completes on their behalf. Designers should report when scaffolds frustrate, exclude, or delay learners as carefully as they report successful reflection, because support quality depends on responsive adjustment.

Related papers

A university student explains a geometry construction to a lecturer while a classmate follows and a laptop displays a related digital diagram
Industry7 Sept 2026
Industry 112

Commentary: Astra's AGI claim puts evidence of human learning at the centre of education

AIED.HK Editorial

AI Product News Commentary

OpenAI launched GPT-6 Astra on 3 September 2026 amid claims about the arrival of AGI. This commentary treats that label as a claim, not an established consensus. For education, the immediate challenge is to distinguish what an AI can produce from what a learner can explain, question and transfer independently—and to use stronger agents to support that learning.

product newscommentaryGPT-6 Astra
Read 500-word summary →
Three education and software colleagues review illustrated lesson cards, an annotated chart and a digital prototype in a bright university design studio
Industry7 Sept 2026
Industry 113

Commentary: Fable 5.1 brings longer AI workflows to AIED—and makes educational validation more important

AIED.HK Editorial

AI Product News Commentary

Anthropic released Claude Fable 5.1 on 1 September 2026 with stronger long-running coding and knowledge-work capabilities and cheaper cache reads. For AIED, the opportunity is a faster cycle from teaching idea to reviewable prototype and research analysis. The test is whether teams can turn that speed into better pedagogy and credible evidence, while accounting for total cost, data conditions and human review.

product newscommentaryClaude Fable 5.1
Read 500-word summary →
A university student compares an AI explanation with handwritten concept notes while an instructor and peers work in a seminar room
Journal Paper2026
Journal Paper 50

Experimental evidence on the learning impact of generative AI: gains persisted when students used it for explanation rather than automation

Zara Contractor, Germán Reyes

arXiv working paper

A randomized, proctored experiment reported that undergraduate access to off-the-shelf generative AI raised immediate factual and conceptual test performance by 0.27 standard deviations and that the gains persisted one week later. The working paper also finds a consequential usage pattern: students who used AI to explain concepts showed stronger delayed gains than students who used it to automate drafting.

generative AIrandomized experimenthigher education
Read 500-word summary →