Back to Research News
Editorial cover for a systematic review of technology-enhanced CLIL
ReviewEvidence synthesis202515 Jul 2026· 3 min

Technology-enhanced content and language integrated learning: A systematic review of empirical studies

Dongpin Hu, Juanjuan Chen, Yan Li, Minhong Wang

Educational Research Review

500-word summary

Editorial cover for a systematic review of technology-enhanced CLIL

Listen to the 500-word paper summary

Audio summary

0:00/0:00

Hu, Chen, Li, and Wang review how technology has been used in content and language integrated learning, or CLIL, where students learn a subject through a non-native language. The paper addresses an important AIED question: does adding digital technology help learners manage the dual demands of understanding subject content and developing language? Rather than assuming that technology is beneficial, the authors map what has been studied, which tools have been used, what outcomes have been measured, and which designs offer useful examples for practice.

The review searched Web of Science for English-language SSCI journal articles published from 2014 to 2023. From 390 records, the authors retained 20 empirical studies after title, abstract, and full-text screening. Two researchers screened and coded the studies independently, resolving disagreements through discussion; coding agreement was high, with a Kappa value of 0.91. The review classified education level, learning mode, subject domain, learners' first language, language of instruction, language skills, technology type, and cognitive or affective outcomes. For comparative evidence, it also synthesized nine experiments from seven controlled studies using a random-effects model.

The evidence base is concentrated. Fourteen of the 20 studies were conducted in Asia, five in Europe, and one in South America. Three quarters involved tertiary education, and 90 percent used English as the language of instruction. Productive language skills such as speaking and writing received more attention than reading and listening, although many papers did not specify a language-skill focus. The authors identify seven technology categories: telecollaboration tools, online learning platforms, multimedia, virtual learning environments, social-network applications, language tools, and intelligent conversation tools. Telecollaboration was most common, while intelligent chatbots appeared in only one included study.

The outcome evidence is more cautious than the technology range might suggest. All 20 studies reported cognitive outcomes, commonly language knowledge or subject-content knowledge, while eight examined affective outcomes. Only one addressed collaboration and communication, and none assessed higher-order thinking as an outcome. Across the controlled studies, the estimated effect on content knowledge was small at 0.29, with a 95 percent confidence interval from -0.07 to 0.66. The estimated effect on language achievement was 0.16, with a confidence interval from -0.03 to 0.35. Because both intervals include zero and the number of experiments is small, neither estimate provides firm evidence of a positive overall effect.

Three exemplary cases make the design implications more concrete: cooperative digital storytelling in a Taiwanese primary school, an educational chatbot in a Greek secondary-school history unit, and a cloud learning environment in a mainland Chinese university course. Across these examples, technology is most useful when it supports rich input, scaffolding, interaction, purposeful learner output, higher-order activity, and balanced development of both content and language. The chatbot case also showed that novelty or student preference did not automatically produce better achievement.

For Hong Kong educators, the review argues for design discipline. Technology should serve explicit content and language goals, with scaffolds that manage the cognitive load of learning subject matter through another language. Evaluation should include motivation, self-efficacy, interaction, and durable language development, not only immediate knowledge tests. The review is limited to English-language SSCI articles in one database, includes only 20 studies, and groups some EMI work with CLIL, so its estimates should not be treated as final. Its strongest contribution is a practical warning: effective TECLIL depends on pedagogy and evidence, not the presence of a digital tool alone.

Related papers

A Black female lecturer and two diverse university students review an audio transcript, curriculum binder and organized learning cards in a media studio
Industry9 Aug 2026
Industry 55

Product news: GPT Transcribe, Claude memory and Gemini Classroom make learning context persistent

OpenAI, Anthropic, Google for Education

AI Product and Learning Report

Product news: OpenAI released GPT Transcribe and GPT Live Transcribe for file and streaming speech, Anthropic changed Claude memory into categorized entries that update across conversations, and Google is connecting Gemini learning activities to teacher-selected Classroom materials. Together, the products make consent, correction and purposeful forgetting central to educational AI design.

product newsGPT TranscribeClaude memory
Read 500-word summary
A racially diverse university language class evaluates AI-supported writing, speaking, translation and feedback activities with an instructor
Review2026
Review 78

A review of 101 higher-education English-language studies maps five roles for AI and recurring evidence gaps

Yanping Wang, Zuwati Hasim, Ling Wu, Yujia Fang

Cogent Education

A systematic review retained 101 empirical studies of AI in higher-education ESL and EFL from 1,536 records. It identified five broad roles for AI and generally positive reports, while finding regional and skill imbalances, uneven methods, overreliance risks and limited evidence about culture, equity and durable outcomes.

AI in educationEnglish language teachingsystematic review
Read 500-word summary
A diverse mathematics class works with an AI tutoring system while a teacher compares student reasoning and assessment evidence
Review2026
Review 74

Review of 12 studies finds promising mathematics-tutoring gains while generative-AI achievement evidence remains thin

Neo Molemane, Moeketsi Mosia, Felix O. Egara

Discover Education

A PRISMA systematic review screened 1,749 records and included 12 studies of AI tutoring and assessment in mathematics. Results were generally promising, especially for some lower-performing learners, but heterogeneous methods and limited direct achievement evidence for generative AI prevent a universal efficacy claim.

AI tutoringmathematics educationautomated assessment
Read 500-word summary