Skip to content
Review

Evaluating the Effects of Automated Test Results, Model Solutions, and Checklists on Peer Code Review: A Quantitative Study

Unknown authors
· 0 citations · 37 references

TL;DR

Qu quantitative analysis of 622 peer reviews on review time, length, and correctness reveals that students are better at assessing actual correct submissions as correct than incorrect ones as incorrect, with test results and checklists slightly increasing their performance in identifying incorrect submissions.

View source

Similar papers

Open access Aug 2026

Comparative Evaluation of Large Language Models in Computer Programming Education

A comparative analysis of six LLMs for generating formative feedback on introductory Java programs containing predefined defects under controlled conditions reveals substantial cross-model variation, particularly in multi-defect scenarios.

Melina Najimi, Saba Yazdani, Marzieh Ahmadzadeh · 0 citations
2026

From Repositories to Practice: LLM-Based Personalization in Programming Education

The design and implementation of a web application that connects to a university version control system, analyzes student-selected repositories, and generates programming challenges targeted at weaknesses identified in the submitted source code is presented.

M. Horváth, Michaela Durkovicová, Lenka Bubenková et al. · 0 citations
Review Open access Aug 2026

Human-in-the-Loop LLM Assessment for Programming Education: Design and Empirical Validation

The system attained a System Usability Scale (SUS) score of 88.5 and cut grading time by 87.5 percent, facilitating focused instructor review in LLM-supported programming assessments, facilitating focused instructor review in LLM-supported programming assessments.

A. Ibrahim, Runal Rezkiawan · 0 citations
Book Open access Aug 2026

Scaffolding Autocomplete: Improving Guidance for Learners using Generative Code Suggestions

A scaffolded programming exercise designed to support student differentiation between good and bad GenAI code suggestions based on negative expertise–that identifying why an answer is wrong is part of developing conceptual knowledge.

J. Prather, Stephen MacNeil, Andrew Luxton-Reilly et al. · 0 citations
Review

Prompt engineering applied to code generation: a preliminary systematic review

This Systematic Literature Review examines prompt engineering in automatic code generation using large language models (LLMs) and shows that prompt engineering has been established as a key discipline for optimizing interaction with LLMs and improve the accuracy, robustness, and applicability of the generated code.

E. Camacho, Y. Gutierrez, César Pardo · 0 citations
Review Aug 2026

CodeStylist: Supporting Early Undergraduate Programmers with Course-Aware Code Style Feedback

Findings are interpreted as evidence that course-aware style feedback is promising as a pre-submission revision aid, but that future versions should combine deterministic rule checks with LLM-generated explanations, rule citations, and stronger verification support.

Ethan Dickey, L. Vento, Peter Kurto et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.