Retour aux actualités recherche
Four diverse university students practise prompting and source checking with an instructor at a library learning table
Article de revuePeer-reviewed study20268 août 2026· 2 min

A 90-minute GenAI literacy course improved knowledge, prompting, source checking and self-efficacy across 65 university sections

Allison E. Connell Pensky, Lydia E. Eckstein, Michael C. Melville, Laura O. Pottmeyer, Zach Mineroff, Avi Chawla, Judy Brooks, Chad Hershock, Marsha C. Lovett

Computers & Education

Résumé de 500 mots

Four diverse university students practise prompting and source checking with an instructor at a library learning table

Connell Pensky and colleagues test a practical question for universities: can a short, reusable online course improve students' capacity to work with generative AI across disciplines? Their 2026 article in Computers & Education evaluates four asynchronous modules designed at Carnegie Mellon University. The approximately 90-minute sequence combined explanations, examples, practice and immediate feedback. It addressed how generative-AI systems work, why appropriate use depends on the person and task, ethical concerns, strategies for responsible educational use and students' confidence in applying those strategies.

The study involved 1,368 undergraduate and graduate students in 65 sections of 53 courses taught by 46 instructors. Course sections, rather than individual students, were assigned to complete the modules between a pre-test and post-test or to a wait-list control condition. That design reduced the chance that students in one section would receive different versions of the intervention, while the analysis accounted for students being nested within courses. Measures included knowledge questions, self-efficacy ratings and authentic tasks for prompt engineering and output evaluation. A randomly selected subset of 174 students had authentic-task responses scored independently without raters knowing the condition or testing time.

Students who completed the modules improved more than control students in knowledge of how large language models work, prompt-engineering skill and self-efficacy. The published abstract also reports gains in fact- and source-checking. Effects were not uniform across every outcome: the modules did not improve the skill of critically evaluating potential bias in generated output, and the university's study report found no added improvement in responsible-use knowledge or overall output analysis. This contrast is important. A concise module can build a useful foundation, but knowing model basics and writing a better prompt are not the same as judging fairness, omissions, evidence or downstream harm.

The authors report that improvements were equitable across the examined categories of birth sex, race and ethnicity, student level, first-generation status and academic discipline. That finding means the analysis did not detect differential benefits across those groups in this setting; it does not prove that one module will remove every access or participation gap. The study was conducted at one university with volunteer instructors, and the post-test followed soon after the intervention. It therefore does not establish long-term retention, transfer into real coursework, changes in academic performance or responsible behavior when tools and incentives differ. Generative-AI products also change quickly, so examples and assessments require maintenance.

For Hong Kong universities, the intervention offers a feasible starting architecture: a shared foundational module, embedded practice, immediate feedback and authentic tasks that require students to create prompts and inspect outputs. Institutions should add discipline-specific cases, multilingual examples and repeated source, bias and uncertainty checks across a semester. Evaluation should include delayed assessments, course artifacts and observed application, not completion rates alone. The result is encouraging precisely because it is bounded: short training can improve several competencies at scale, while critical judgment needs sustained, contextual practice.

Articles liés

Four diverse adults analyze a business problem with a laptop, charts and an unassisted written follow-up in a workforce-learning laboratory
Article de revue2026
Article de revue 54

Generative AI closed three quarters of an education-based performance gap during assisted work, but effort shaped what carried forward

Guillermo Cruces, Diego Fernández Meijide, Sebastian Galiani, Ramiro H. Gálvez, María Lombardi

arXiv working paper

In a preregistered randomized online experiment with 1,174 Argentine adults, GPT-4.1 assistance raised workplace-style problem-solving performance for both education groups and reduced the baseline gap from 0.548 to 0.139 standard deviations. Lower-education participants retained a modest gain after AI was removed, but stronger follow-up performance appeared when intensive assistance was paired with sustained human effort.

generative AIrandomized experimenteducation inequality
Lire le résumé de 500 mots
A university student compares an AI explanation with handwritten concept notes while an instructor and peers work in a seminar room
Article de revue2026
Article de revue 50

Experimental evidence on the learning impact of generative AI: gains persisted when students used it for explanation rather than automation

Zara Contractor, Germán Reyes

arXiv working paper

A randomized, proctored experiment reported that undergraduate access to off-the-shelf generative AI raised immediate factual and conceptual test performance by 0.27 standard deviations and that the gains persisted one week later. The working paper also finds a consequential usage pattern: students who used AI to explain concepts showed stronger delayed gains than students who used it to automate drafting.

generative AIrandomized experimenthigher education
Lire le résumé de 500 mots
Editorial cover of undergraduate learners and a lecturer examining a course-grounded RAG chatbot alongside flat learning and motivation outcome traces
Article de revue2026
Article de revue 36

AI chatbots in higher education: Comparing expectations to evidence

Andrew Thoeni, Luke K. Fryer

Computers in Human Behavior Reports

A semester-long randomized field experiment with 454 undergraduates found that access to a course-grounded RAG chatbot did not significantly improve interest, self-efficacy, engagement, or test performance, despite students reporting that they liked the tool.

RAG chatbotrandomized field experimenthigher education
Lire le résumé de 500 mots