Oct 2026· Proceedings of the 14th Nordic Conference on Human-Computer Interaction· 0 citations· 58 references
TL;DR
A design case study of a JupyterLab addon that delivers Socratic hints instead of direct answers, and six design hypotheses for developers of constrained AI programming assistants, addressing hint escalation, selective dialogue, context granularity, vocabulary calibration, onboarding transparency, and difficulty-aware scaffolding.
Abstract
Generative AI tools such as ChatGPT Codex and Claude Code can produce complete solutions to programming exercises, raising concerns about over-reliance and reduced learning among novice programmers. Constraining AI output is a promising but under-explored design strategy. We present a design case study of a JupyterLab addon that delivers Socratic hints instead of direct answers. The system enforces four deliberate constraints: (1) no free-form chat input, (2) no code generation, (3) automatic first hints triggered by cell execution, and (4) Socratic questioning as the sole response format. We deployed the system in a usability evaluation with 12 first-year undergraduates from a non-CS bachelor program working primarily on Scala exercises (one used Python). Drawing on open-ended survey responses, think-aloud transcriptions, and interaction log episodes, we identify five design tensions that emerged from student interactions with the constrained interface: the helpfulness–guardedness trade-off, the one-way interaction dilemma, the hint progression gap, the adaptivity ceiling, and the language complexity barrier. We derive six design hypotheses for developers of constrained AI programming assistants, addressing hint escalation, selective dialogue, context granularity, vocabulary calibration, onboarding transparency, and difficulty-aware scaffolding. Our findings inform the design space of constrained AI tools for programming education.
This work designed a tool called dBlocks with the following features: blocks to scope content, a context manager to edit context, and inline execution to verify code, and shows how human-centered design can guide the development of LLM-integrated tools.
Undergraduate students can now obtain complete programming answers from generative artificial intelligence systems in seconds, but many of those answers are accepted with little checking. This paper examines whether explanation-first prompting makes such answers easier to review rather than whether it improves student...
Orest Raiter· Automation, Control, and Inf...· 0 citations
The rapid adoption of artificial intelligence (AI) programming assistants has raised questions about the actual benefits they provide in professional software development. This exploratory longitudinal case study compares configurations of three AI programming assistants (GitHub Copilot, ChatGPT, and Cursor)—specific c...
Goran Đambić, Anton Maurovic, Ivana Ogrizek Biškupić et al.· Information· 0 citations
Generative artificial intelligence is reshaping programming education, yet its effects on skill development depend partly on how learners interact with artificial intelligence-supported systems. This study introduces the Artificial Intelligence-Scaffolding Interaction Framework, which conceptualizes constrained, questi...
M. Yenidogan, Zeynep Cömert, Duygu Çakır· IEEE Access· 0 citations
A scaffolded programming exercise designed to support student differentiation between good and bad GenAI code suggestions based on negative expertise–that identifying why an answer is wrong is part of developing conceptual knowledge.
J. Prather, Stephen MacNeil, Andrew Luxton-Reilly et al.· International Computing Educ...· 0 citations
Related blog posts
MIT News · Artificial Intelligence· news.mit.eduOct 8, 2026
Exploring how generative AI could make machine vision more accessible to businesses. The post GenEye in a Box: Making Machine Vision Something You Can Just Ask For appeared first on GPT-Lab.
MIT News · Artificial Intelligence· news.mit.eduOct 8, 2026
Training AI agents with reinforcement learning can be challenging because their tools, context, and decision-making are managed by complex frameworks. Agent Lightning connects existing agents to RL training, making it easier to improve them without rebuilding them. The post Agent Lightning v1.0: A 3,500-Line Lightweight Agentic RL Framework for Training Agents with Real Harnesses appeared first on Microsoft Research.
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.