← Voltar às notícias de pesquisa
Editorial cover for a review of generative AI in programming education
RevisãoEvidence synthesis202525/06/2026· 2 min

Literature Review on the Integration of Generative AI in Programming Education

Jemimah Nathaniel, Solomon Sunday Oyelere, Jarkko Suhonen, Matti Tedre

International Journal of Artificial Intelligence in Education

Resumo de 500 palavras

Editorial cover for a review of generative AI in programming education

Nathaniel, Oyelere, Suhonen, and Tedre review a question that is now central to computer science education: how can generative AI tools be integrated into programming education without weakening students' foundational logic, problem solving, and higher-order thinking skills? The paper is useful for AIEDHK because it moves beyond generic enthusiasm for ChatGPT or Copilot. It asks whether the tools are embedded in teaching methods, assessment routines, and learning processes that still require students to understand code rather than only generate it.

The review synthesizes 40 empirical studies using PRISMA 2020 and Kitchenham-style review methods. Its focus is not simply whether GenAI can solve programming tasks. Instead, it examines how studies connect GenAI tools with programming curricula, teaching methods, assessment designs, integration processes, and student skill development. That framing is important because programming education has a long history of intelligent tutoring systems, automated feedback, Parsons problems, code explanation tools, and step-based support. GenAI adds flexibility and natural-language interaction, but it also increases the risk that learners accept generated code without understanding algorithms, syntax, data structures, or debugging logic.

The paper's findings are deliberately implementation-focused. The authors argue that successful integration depends on intentional teaching strategies, thoughtfully designed assessments, and structured integration processes. They also identify barriers: limited accessibility support, insufficient bias mitigation, weak curriculum alignment, and tool selection driven by availability rather than educational fit. These are practical concerns for any school or university considering AI-assisted coding. A tool that improves productivity for experienced developers can still be harmful for novice learners if it bypasses the struggle needed to build mental models.

The review also proposes a GenAI-Ped framework that combines self-regulated learning, universal design principles, prompt-engineering support, and iterative feedback. For AIEDHK, this is the most actionable contribution. It suggests that GenAI in coding courses should be framed as a guided learning partner, not an answer machine. Students need orientation on when to ask for help, how to inspect generated code, how to explain a solution, and how to reflect on what they have learned. Teachers need assessment formats that reveal reasoning, not only final code output. Product teams need interfaces that encourage explanation, comparison, revision, and metacognitive checks.

The paper is especially relevant for Hong Kong because programming education is multilingual, high-stakes, and often linked to future workforce claims. GenAI coding support can make programming more accessible, but only if it is aligned with local curricula, language needs, teacher capacity, and assessment expectations. AIEDHK can use this review to evaluate AI coding tutors, coding assistants, and student copilots through a clear test: does the system help learners develop durable programming understanding, or does it mainly make correct-looking code easier to obtain?

Artigos relacionados

A university student explains a geometry construction to a lecturer while a classmate follows and a laptop displays a related digital diagram
Política / ética7/09/2026
Política / ética 112

Commentary: Astra's AGI claim puts evidence of human learning at the centre of education

AIED.HK Editorial

AI Product News Commentary

OpenAI launched GPT-6 Astra on 3 September 2026 amid claims about the arrival of AGI. This commentary treats that label as a claim, not an established consensus. For education, the immediate challenge is to distinguish what an AI can produce from what a learner can explain, question and transfer independently—and to use stronger agents to support that learning.

product newscommentaryGPT-6 Astra
Ler resumo de 500 palavras →
A programming lecturer and two diverse university students inspect compiled code, an inheritance diagram, and a grading rubric in a computer laboratory
Artigo de revista2026
Artigo de revista 102

Five AI systems outscored the average OOP cohort but still failed compilation and advanced concepts

Marina Lepp, Joosep Kaimre

arXiv preprint

Lepp and Kaimre evaluated ChatGPT-5.2, DeepSeek-V3, Gemini 2.5 Flash, Claude Sonnet 4.5, and Microsoft 365 Copilot on authentic introductory OOP tests and examinations using student grading criteria. Systems exceeded the historical average and often solved long tasks, yet some code did not compile and interfaces, abstract classes, inheritance, and image-based questions remained difficult. The results challenge take-home assessment validity without proving student learning.

programming assessmentobject-oriented programminggenerative AI
Ler resumo de 500 palavras →
A Black university student debugs from her own notes while an instructor supports her beside a graduated cyan help ladder whose final solution rung is locked
Ferramenta / conjunto de dados2026
Ferramenta / conjunto de dados 92

A guarded LLM tutor reached its withholding targets in scripted tests, but student learning remains unmeasured

Yusuf Pisan

arXiv preprint

Pisan reports a deployed programming tutor that places an eight-rung help ceiling outside the generating LLM, strips solution code deterministically and judges risky replies against a per-turn contract. Across roughly two dozen scripted turns per calibration run, earnest-reply revisions fell from 43% to 0% and audited ceiling compliance rose from about 77% to 100% after measurement and policy defects were repaired. The evaluation used synthetic personas, not students, so it establishes contract compliance rather than durable learning.

Socratic tutoringanswer withholdingprogramming education
Ler resumo de 500 palavras →