Intent-Driven Situation States (IDSS) is proposed, a training-free framework that maintains an explicit situation state alongside the dialogue that allows agents to avoid infeasible actions, advance dependent goals, and reuse relevant information without repeatedly searching raw history.
Abstract
User-centric multi-turn agents must act on an evolving task situation shaped by changing user intents, accumulated tool-grounded facts, missing information, and execution constraints. Existing context-management methods improve the use of past interaction history, but rarely maintain an explicit situation state that separates grounded facts from task-state judgments. As a result, agents often need to infer fine-grained attributes, task dependencies, and constraint satisfaction implicitly from dialogue traces. We propose Intent-Driven Situation States (IDSS), a training-free framework that maintains an explicit situation state alongside the dialogue. IDSS parses tool returns into provenance-aware entities and attributes, tracks user intents, required variables, constraints, and execution status, and propagates new facts to task constraints to update action executability. This allows agents to avoid infeasible actions, advance dependent goals, and reuse relevant information without repeatedly searching raw history. Experiments on three interactive benchmarks across eight LLMs show that IDSS improves task completion, preference elicitation, and interaction efficiency, with clear gains on tasks involving multi-entity coordination, evolving user constraints, and constraint-aware replanning. Ablations and error analyses show that these improvements come from the interaction between fact persistence, intent-centered state tracking, and constraint modeling. These results suggest that explicit situation tracking offers an effective alternative to history-centric context management for reliable user-centric multi-turn agents.
GSC-QA (Goal-based Sequential Conversation QA), a framework that integrates three complementary components into a unified enterprise dialogue architecture that combines retrieval, instruction enforcement, and goal persistence in a single coordinated loop built on LangGraph, is introduced.
The Procedural Graph is introduced: just as a knowledge graph organizes factual knowledge into (entity, relation, entity) triplets for what-is questions, a Procedural Graph organizes procedural knowledge into (procedure, relation, procedure) triplets for what-to-do questions.
Yu-Xing Lu, Yi-Cheng Chen, Shan-Chan Wu et al.· 0 citations
SAGE (State-Grounded Abstention-Aware Evaluation), which compiles a workflow specification and per-turn state diff into atomic, schema-grounded criteria and routes each through a cascade of symbolic and encoder/NLI verifiers that abstain rather than guess, aggregating criterion verdicts into a turn-level decision with...
Rayan Khoury, Shih-Yao Lin, P. Mishra· 0 citations
SAIN is presented, a zero-shot framework that turns active dialogue into persistent navigation state and supports dialogue-to-state conversion as an effective zero-shot mechanism for long-horizon interactive instance navigation.
Evaluating language-guided mobile agents has recently shifted from rule-based to model-based approaches to achieve scalable and automated assessments. However, existing holistic evaluation paradigms process entire trajectories at once, leading to substantial context overload. Moreover, they primarily focus on task comp...
Peng-Jian Yang, Zijing Gao, Xue Yu et al.· 1 citation
OODA-Tool, a typed closed-loop policy designed to mitigate state preservation from action realization, consistently improves task success across model sizes, with larger gains on smaller models and on tasks whose actions depend strongly on information accumulated across turns and prior tool results.
Rong-Feng Guo, Yin-Xuan Huang, Yusen Wu et al.· 1 citation
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.