Skip to content
Review

HARP: The Human-AI Research Platform

Jul 2026 · arXiv.org · Vol abs/2607.20773 · 0 citations · 7 references
Computer Science

TL;DR

By pairing controllable live agents with behavioral and self-report measures, HARP enables systematic testing of how AI design choices affect users.

Abstract

Large language models (LLMs) have shifted human--computer interaction from `traditional''interface journeys toward more conversational exchanges. Researchers studying HCI and UI use moderated usability sessions, interviews, surveys, transcript analysis, and static prototypes. However, static prototypes provide limited opportunities to study interaction with live AI systems or systematically control how an LLM behaves across participants and scenarios. Conversation transcripts reveal little about how users formulate, revise, and hesitate over prompts before submission. We designed the Human--AI Research Platform (HARP) for researchers, designers, and anyone who has ever wondered, `What if AI did this?'HARP places participants in controlled mock scenarios with live, configurable AI agents. Researchers can control agent prompts, model parameters, response characteristics, and experimental conditions; trigger surveys at predefined moments; and record prompt composition time, response latency, deletions, and keystroke pauses. Planned capabilities include voice, facial expression, gesture, and, where legally and ethically appropriate, emotion analysis. We illustrate HARP through a study examining how technical specificity and response length affect retention of LLM output. By pairing controllable live agents with behavioral and self-report measures, HARP enables systematic testing of how AI design choices affect users.

View source

Similar papers

#natural language process... Preprint Aug 2026

How To Do Things With Prompts

This paper applies speech act and politeness theory to a corpus-pragmatic analysis of 2,000 English-language prompts drawn from publicly shared ChatGPT conversations, showing a consistent movement toward indirect, implicit, and fragmentary realizations of directive force, accompanied by a decline in politeness marking.

Kristina Šekrst, Virna Karlić · 0 citations
#artificial intelligence Preprint Sep 2026

Large-Scale User Behavior Analysis in Multimodal AI-Assisted Manual Task Execution

Conversational Task Assistants (CTAs) are multimodal dialogue systems that support users in complex real-world tasks such as cooking and DIY through voice, text, image, and video interactions. Prior user studies have focused on controlled settings, leaving limited understanding of real-world CTA usage at scale. In this...

Rafael Ferreira, Diogo Tavares, Diogo Glória-Silva et al. · 0 citations
#large language models Open access Sep 2026

Corpus-level behavioral intensity in human–AI conversations

This study proposes a transparent rule-based framework for estimating persistence, delegation-related lexical patterns, and refinement-related lexical markers without inferring psychological dependence in conversational generative models.

Aracely Mera-Navarrete, Solange Revelo, Jefferson Beltrán-Morales et al. · 0 citations
Review Open access Aug 2026

ChatGPT for Intelligent Human–AI Interaction: Opportunities and Limitations

ChatGPT for Intelligent Human–AI Interaction: Opportunities and Limitations provides a comprehensive review of the technological foundations, capabilities, applications, and constraints associated with ChatGPT, emphasizing its role in enabling intelligent human–AI collaboration.

Antoine Morel, Camille Laurent · 0 citations

Computers in Human Behavior: Artificial Humans

It is suggested that chatbot communication style influences users’ perceptions of conversational agents and may improve performance relative to less supportive chatbot designs, but the overall value of chatbot interaction depends on the task context.

Erik Derner, Dalibor Kučera, Aditya Gulati et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.