Back to Research News
A university student compares an AI explanation with handwritten concept notes while an instructor and peers work in a seminar room
Journal PaperPeer-reviewed study20267 Aug 2026· 2 min

Experimental evidence on the learning impact of generative AI: gains persisted when students used it for explanation rather than automation

500-word summary

A university student compares an AI explanation with handwritten concept notes while an instructor and peers work in a seminar room

Listen to the paper summary

Audio summary

0:00/0:00

Contractor and Reyes examine a question that is easy to obscure in product demonstrations: when students have access to a general-purpose generative-AI system while learning a new topic, do they learn more once the system is removed? Their July 2026 working paper reports a randomized experiment conducted in proctored, in-person undergraduate sessions. Participants studied an unfamiliar subject and wrote an analytical essay either with or without access to an off-the-shelf generative-AI tool. They then completed unaided assessments immediately and one week later. The study measures both knowledge tests and open-ended writing, so it separates short-term task performance from later independent learning more clearly than a satisfaction survey can.

The authors report that AI access increased immediate factual and conceptual test scores by 0.27 standard deviations. They also report that the advantage persisted at the one-week assessment. That result matters because a common concern is that AI may improve the visible product while shifting effort away from understanding. In this setting, the reported knowledge gains did not disappear when students worked without the tool. The paper also reports little change in essay quality while AI was available, but better style and relevance in unaided writing one week later.

The most useful finding is not a general claim that AI access is beneficial. The researchers distinguish augmentation-oriented use from automation-oriented use. Students who used the system to obtain explanations of concepts had stronger delayed gains than students who used it primarily to generate text. The paper links the result to reported changes in effort: AI users shifted time away from drafting and toward reading and searching for information, while also reporting greater learning enjoyment. These measures identify plausible mechanisms, but they do not prove every learner followed the same path or that every tool configuration will produce the same effect.

The evidence needs careful interpretation. This is an arXiv working paper rather than a peer-reviewed journal article, and the results should not be generalized without replication. The task involved an unfamiliar topic, proctored in-person sessions and a specific assessment schedule; different courses, age groups, prompting supports, incentives or unrestricted home use may lead to different behavior. The paper's own results suggest that usage quality is central. Giving students a tool without a learning design can encourage either explanation, inquiry and revision or fast drafting with little durable understanding.

For higher education and Hong Kong classrooms, a defensible pilot would make the augmentation route explicit. Teachers can ask students to request explanations, compare them with course sources, annotate what changed in their understanding and complete an independent follow-up task. Rubrics can reward source evaluation, reasoning and revision rather than polished first drafts. Process logs should support reflection rather than surveillance, and assessments should include moments when learners demonstrate what they can do unaided. The study offers a promising but provisional message: AI may support learning when it redirects effort toward sense-making, not when it quietly replaces it.

Related papers

Four diverse adults analyze a business problem with a laptop, charts and an unassisted written follow-up in a workforce-learning laboratory
Journal Paper2026
Journal Paper 54

Generative AI closed three quarters of an education-based performance gap during assisted work, but effort shaped what carried forward

Guillermo Cruces, Diego Fernández Meijide, Sebastian Galiani, Ramiro H. Gálvez, María Lombardi

arXiv working paper

In a preregistered randomized online experiment with 1,174 Argentine adults, GPT-4.1 assistance raised workplace-style problem-solving performance for both education groups and reduced the baseline gap from 0.548 to 0.139 standard deviations. Lower-education participants retained a modest gain after AI was removed, but stronger follow-up performance appeared when intensive assistance was paired with sustained human effort.

generative AIrandomized experimenteducation inequality
Read 500-word summary
Four diverse university students practise prompting and source checking with an instructor at a library learning table
Journal Paper2026
Journal Paper 52

A 90-minute GenAI literacy course improved knowledge, prompting, source checking and self-efficacy across 65 university sections

Allison E. Connell Pensky, Lydia E. Eckstein, Michael C. Melville, Laura O. Pottmeyer, Zach Mineroff, Avi Chawla, Judy Brooks, Chad Hershock, Marsha C. Lovett

Computers & Education

In a large experiment involving 1,368 undergraduate and graduate students across 65 university course sections, a 90-minute asynchronous GenAI learning module improved knowledge of how the technology works, prompt-engineering performance, fact- and source-checking, and self-efficacy. It did not improve critical evaluation of bias, showing that short foundational training needs deeper practice for responsible judgment.

generative AI literacyrandomized experimenthigher education
Read 500-word summary
University students discuss transparent and responsible generative-AI use during a collaborative assignment while a teacher facilitates peer reflection
Journal Paper2026
Journal Paper 84

Perceived classmate GenAI use was associated with lower trust, while perceived AI literacy attenuated the direct link

Zhen Zhang, Jiaying Geng, Chunhui Qi

Behavioral Sciences

A cross-sectional survey of 406 students at two institutions found that perceiving a classmate as using more GenAI was associated with lower perceived warmth, competence and interpersonal trust. Perceived target AI literacy weakened only the direct association, while the design cannot establish that AI use caused distrust.

generative AIinterpersonal trustAI literacy
Read 500-word summary