AIEDHKAI in Education Hub of Knowledge
HomeMissionNewsAcademyAbout
HomeMissionNewsAcademyAbout
← Back to Research News

#programming assessment

programming assessment

1 papers

A programming lecturer and two diverse university students inspect compiled code, an inheritance diagram, and a grading rubric in a computer laboratory
Journal Paper2026
Journal Paper 102

Five AI systems outscored the average OOP cohort but still failed compilation and advanced concepts

Marina Lepp, Joosep Kaimre

arXiv preprint

Lepp and Kaimre evaluated ChatGPT-5.2, DeepSeek-V3, Gemini 2.5 Flash, Claude Sonnet 4.5, and Microsoft 365 Copilot on authentic introductory OOP tests and examinations using student grading criteria. Systems exceeded the historical average and often solved long tasks, yet some code did not compile and interfaces, abstract classes, inheritance, and image-based questions remained difficult. The results challenge take-home assessment validity without proving student learning.

programming assessmentobject-oriented programminggenerative AI
Read 500-word summary →
AIEDHKAI in Education Hub of Knowledge

AIEDHK is a multilingual knowledge hub for AI in Education research, development, and responsible learning innovation.

Navigation

HomeMissionNewsAcademyAbout

Ecosystem links

Dr. Peter Hu DongpinPedaNova TechnologyMAISCAIS
© 2026 AIEDHK. All rights reserved.