Skip to content

From clustering to classification: mining code schemas for accurate identification of student solution strategies

Jul 2026 · International Journal of Data Science and Analysis · Vol 22 · 0 citations · 52 references

TL;DR

SchemaNet is introduced, a high-accuracy supervised model that detects student solution strategies via multi-view features, especially code schemas extracted at sub-problem granularity by the proposed SchemaMiner framework, and validated on a newly constructed, manually annotated dataset of 1,612 Python solutions.

View source

Similar papers

Open access

Knowledge component-constrained diagnostic prompting for automated knowledge gap detection

Introductory programming courses face challenges to give scalable feedback on students’ understanding of core programming concepts. Automated grading shows whether code passes its test cases, but not the concept gaps behind the errors. Knowledge Tracing models target those concepts, but they demand extensive historical...

Pranay Sureshrao Ghuge · 0 citations
Open access 2026

Leveraging large language models for scalable analysis of the end-of-the-course student feedback

An LLM-based feedback analytics pipeline designed to transform students’ open-ended feedback into structured, actionable insights is proposed, ultimately supporting data-informed improvements in teaching and course management.

J. Jovanović, Irena Vodenska, V. Devedžić · 0 citations
Sep 2026

Automated LLM-based Classification of Software Requirements

The adoption of large language models (LLMs) in software engineering has enabled the potential to automate complex activities such as requirements analysis. This paper presents an empirical performance analysis of four modern LLMs: GPT-4o, Aya, Gemma and Phi-4 on the task of automated classification of atomic software...

Nourchène Elleuch Ben Ayed, Jaber Jemai, Keletso J. Letsholo et al. · 0 citations
Open access

Evaluation and Distillation of Source Code Generation Tasks by Large Language Models

Two novel contributions are introduced: CodeEval and CodeQual, an open-source execution framework that provides researchers with a ready-to-use evaluation pipeline for evaluating and improving LLMs in software engineering contexts, encompassing both functional correctness assessment and subjective code quality evaluati...

Danny Brahman · 0 citations
Preprint Aug 2026

RepoProbe: Benchmarking Architecture-Aware Repository Comprehension with Checklists

This work introduces RepoProbe, a novel benchmark for evaluating repository-level code understanding through open-ended Q&A using GitHub Discussions, which focuses on open-ended architectural inquiries rather than defect reporting and proposes a Checklist-Based Verification Protocol that decomposes answers into atomic,...

Yue Yang, Alyssa Wu, Ji Luo et al. · 1 citation
Aug 2026

PAMI-GPT: knowledge-grounded LLMs for reliable pattern mining workflow generation

Experimental evaluation on 300 realistic pattern mining tasks demonstrates consistent improvements in algorithm configuration accuracy, parameter compliance, and dataset specification correctness across zero-shot, one-shot, and few-shot settings, highlighting the effectiveness of inference-time domain grounding for ena...

Madhavi Palla, Uday Kiran Rage, Arjun Chakravarthi Pogaku · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.