Skip to content
Review Open access

Conversational Grounding in Large Language Models: Evaluation Methods, Challenges and Future Directions

2026 · SIGDIAL Conferences · pp. 151-163 · 0 citations · 66 references
Computer Science

TL;DR

This paper surveys how conversational grounding is evaluated in task-oriented dialogue in the current era of LLMs and focuses on how conversational grounding is modelled explicitly—using dialogue acts and by modelling the participant mental state.

Abstract

Conversational grounding is the collaborative process through which speakers establish and maintain mutual understanding. It is essential for the success of a dialogue. While it is inherent in human conversations, it remains a challenge for instruction-following Large Language Models (LLM). This paper surveys how conversational grounding is evaluated in task-oriented dialogue in the current era of LLMs. First, we focus on how conversational grounding is modelled explicitly—using dialogue acts and by modelling the participant mental state. Then, we review collaborative tasks that enable the evaluation of conversational grounding implicitly at the global level based on outcomes. Finally, we highlight the limitations of evaluation, notably the current methodology and metrics used, and outline research directions in conversational grounding and its evaluation.

Read PDF

Similar papers

Book Open access Sep 2026

Conversational Style in Open Domain Dialogue Systems: What Makes a Response Sound Natural

Conversational style in this setting is best understood as a pragmatic orientation toward the user rather than toward the system’s own content, and that the system must be mixed-initiative to manifest a conversational style.

Vrindavan Harrison, M. Walker · 0 citations
Book Open access Aug 2026

Post-Thinking in NPC Dialogue: A Paradigm for Reflective Character Models

Immersive and believable NPC dialogue requires characters that feel intentional. They remember specific information about themselves, follow through on their goals, and stay true to their personalities across long conversations. We introduce Post-Thinking, a technique that maintains a rolling reflection trace across ch...

Keegan Carey, Hexi Wang · 0 citations
#natural language process... Preprint Aug 2026

You Know What I Mean: A Benchmark for Agentic Conversational Reference Grounding

The results show that CoRG remains challenging for current agents, even the best agent reaches only 67.0% success rate, leaving one third of references unresolved, and position CoRG as a concrete benchmark for studying how agents search, inspect, and verify information in realistic multi-tool environments.

Karen Fuchs, Uri Katz, Yoav Goldberg · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.