By pairing controllable live agents with behavioral and self-report measures, HARP enables systematic testing of how AI design choices affect users.
Abstract
Large language models (LLMs) have shifted human--computer interaction from `traditional''interface journeys toward more conversational exchanges. Researchers studying HCI and UI use moderated usability sessions, interviews, surveys, transcript analysis, and static prototypes. However, static prototypes provide limited opportunities to study interaction with live AI systems or systematically control how an LLM behaves across participants and scenarios. Conversation transcripts reveal little about how users formulate, revise, and hesitate over prompts before submission. We designed the Human--AI Research Platform (HARP) for researchers, designers, and anyone who has ever wondered, `What if AI did this?'HARP places participants in controlled mock scenarios with live, configurable AI agents. Researchers can control agent prompts, model parameters, response characteristics, and experimental conditions; trigger surveys at predefined moments; and record prompt composition time, response latency, deletions, and keystroke pauses. Planned capabilities include voice, facial expression, gesture, and, where legally and ethically appropriate, emotion analysis. We illustrate HARP through a study examining how technical specificity and response length affect retention of LLM output. By pairing controllable live agents with behavioral and self-report measures, HARP enables systematic testing of how AI design choices affect users.
It is found that companionship behaviors reduced likability, humanlikeness, and trust in AI chatbots, and women and older participants saw companionship chatbots as less likable, humanlike, and trustworthy.
J. R. Anthis, Mark Díaz, Renee Shelby· 0 citations
This paper applies speech act and politeness theory to a corpus-pragmatic analysis of 2,000 English-language prompts drawn from publicly shared ChatGPT conversations, showing a consistent movement toward indirect, implicit, and fragmentary realizations of directive force, accompanied by a decline in politeness marking.
Conversational Task Assistants (CTAs) are multimodal dialogue systems that support users in complex real-world tasks such as cooking and DIY through voice, text, image, and video interactions. Prior user studies have focused on controlled settings, leaving limited understanding of real-world CTA usage at scale. In this...
Rafael Ferreira, Diogo Tavares, Diogo Glória-Silva et al.· 0 citations
This study proposes a transparent rule-based framework for estimating persistence, delegation-related lexical patterns, and refinement-related lexical markers without inferring psychological dependence in conversational generative models.
Aracely Mera-Navarrete, Solange Revelo, Jefferson Beltrán-Morales et al.· Frontiers in Artificial Inte...· 0 citations
ChatGPT for Intelligent Human–AI Interaction: Opportunities and Limitations provides a comprehensive review of the technological foundations, capabilities, applications, and constraints associated with ChatGPT, emphasizing its role in enabling intelligent human–AI collaboration.
Antoine Morel, Camille Laurent· International Bulletin of Ap...· 0 citations
It is suggested that chatbot communication style influences users’ perceptions of conversational agents and may improve performance relative to less supportive chatbot designs, but the overall value of chatbot interaction depends on the task context.
Erik Derner, Dalibor Kučera, Aditya Gulati et al.· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.