While agents aid initial task completion, they harm users' code comprehension and thus do not prepare users to extend their code, and low-effort agent interaction types, like copy+paste prompts and auto-accepted edits, are linked with lower comprehension.
Abstract
Coding agents (e.g., Cursor) improve developer productivity by optimizing task completion, but shifting users from writing code to prompting and reviewing may harm their understanding, impeding oversight, learning, and communication. To probe this, we have 54 students create a website with one of two AI systems: an agent that edits user code; or a chatbot where users write code alone or adapt generic code snippets. We test understanding via comprehension questions and a task where users extend their code without agents, showing: (1) While agents aid initial task completion, they harm users'code comprehension and thus do not prepare users to extend their code; (2) Low-effort agent interaction types, like copy+paste prompts and auto-accepted edits, are linked with lower comprehension; and (3) Despite self-reported weaker understanding, users still prefer coding agents because they are quick and easy to use. While users stay in the loop for coding workflows, understanding should not be forgotten. Towards this goal, we distill our analyses into future research directions for coding agent developers: dissuading low-effort prompting, creating readable code, and promoting active engagement.
SWE-Touch is introduced, a framework that stress-tests this setting through validated Counter-Edits: plausible edits to task-relevant code that conflict with task completion, and point to detecting workspace changes, reconciling conflicting edits with the task, and verifying the affected behavior as key capabilities fo...
Yuqiao Tan, Jinxiang Meng, Fangyu Lei et al.· 1 citation
VibeJam, a browser-based user study platform for users to collaborate with AI agents to develop websites, and open-source VibeJam to spur extensions and support studies on how coding agents can help users.
Nishant Balepur, Connor Baumler, Valerie Chen et al.· 0 citations
Large language model coding agents have recently become useful for software tasks, but weaker or open-weight agents still struggle to reliably interpret user intent and execute complex multi-step workflows. This gap is especially visible in long-horizon settings, where an agent must repeatedly inspect prior outcomes, d...
AI agents are becoming a fundamental part of modern software creation, helping developers in generating code, debugging, designing systems, etc. But there is a clear difference between how beginners and experienced software engineers get benefits from these tools. Newbies usually depend on agents for one-time prompts a...
Madhurima Kommuru, Srujana Pulipaka· International Journal of Mod...· 0 citations
It is argued that AI coding assistants should not only be evaluated by the code they generate, but also by how they mediate the transfer of that code into software artifacts, and soft barriers are proposed as one class of handoff-aware mechanisms.
I. E. Olatunji, Albérick Euraste Djiré, Jacques Klein et al.· 1 citation
This work proposes a framework for extracting reusable developer preferences from interaction traces, generates personalized skills through rule-based bootstrapping and evidence-grounded refinement, and evaluates them using a reproducible replay framework with an interactive, trajectory-conditioned LLM-based human deve...
Shuyan Huang, Kai Du, Andrew S. Lan· 3 citations· ⚡2
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.