This work presents TACT (Taxonomy-Aligned Conversational Tutor), a human-grounded framework for post-training and evaluating pedagogically adaptive ESL tutors and develops two complementary taxonomies: the Tutor-Strategy Taxonomy with 13 tutor response strategies and the Student-Move Taxonomy characterizing learner behavior by move type and status.
Abstract
Large language models (LLMs) are increasingly used to provide conversational practice for English-as-a-second-language (ESL) learners. Effective ESL tutoring, however, requires more than fluent response generation: a tutor must select an appropriate pedagogical action based on learner behavior and dialogue context. Human-tutoring research offers principles for adaptive support, but they are often task-specific and remain insufficiently integrated into LLM-based ESL tutor training and evaluation. We present TACT (Taxonomy-Aligned Conversational Tutor), a human-grounded framework for post-training and evaluating pedagogically adaptive ESL tutors. Drawing on established literature, we develop two complementary taxonomies: the Tutor-Strategy Taxonomy with 13 tutor response strategies and the Student-Move Taxonomy characterizing learner behavior by move type and status. Using these taxonomies, we construct TACTCorpus, which enriches 260 authentic teacher-student conversations with 32,379 annotations and quality-controlled augmented training data. We then post-train Qwen3.5-4B through supervised fine-tuning followed by taxonomy-aligned Group Relative Policy Optimization, producing TACTutor and optimizing it for scaffolding quality rather than reference imitation alone. On TACTBench, a strategy-balanced diagnostic benchmark comprising 78 authentic tutoring contexts, TACTutor improves over its backbone by 20.30% and outperforms all evaluated proprietary baselines under the same protocol, while maintaining backbone performance on established external educational benchmarks; in a blinded study with 50 learners, it also receives the highest overall mean rating among the evaluated tutors. We release the data, benchmark, and model weights, providing an open foundation for developing pedagogically adaptive ESL tutors.
It is indicated that an integrated learner-state-aware RAG architecture can receive stronger expert ratings for pedagogical text quality on a fixed task set while also introducing trace leakage and evidence-boundary risks that must be controlled before deployment.
Jude A. Adenuga· Frontiers in Education· 1 citation
Current Large Language Models (LLMs) excel at solving complex mathematical problems, yet this proficiency does not inherently translate into effective tutoring. While advanced LLM tutors may leverage multi-agent frameworks or fine-tuning, most still lack a mechanism to systematically accumulate and reuse pedagogical ex...
Jian-Heng Zhou, Chao-Li Zhang, Xing-Jun Wei et al.· 0 citations
An Intelligent Writing Tutor that integrates corpus-informed error analysis, natural language processing, and rule-based reasoning to generate individualized and explainable writing feedback for ESL learners is proposed and demonstrates the feasibility of integrating explainable NLP techniques and second language acqui...
R. L. Ladwingon· International Journal of Adv...· 0 citations
The use of Large Language Models to automatically code indicators of complex psychological constructs from student chat logs collected through a conversation-based assessment for middle school mathematics highlights the promise of human-LLM collaboration to enhance qualitative coding efficiency and validity in educatio...
Teresa M. Ober, Shan Zhang, Diego Zapata-Rivera et al.· 2 citations
The results suggest that learner- and curriculum-aware alignment may matter more for effective tutoring than model category alone, and that such alignment is both measurable and improvable.
Benjamin Barlog, Hudson Craig, Ze-Dong Peng· IEEE International Conferenc...· 0 citations
We use cookies to run the site and, with your consent, for analytics and to show ads.
See our Cookie Policy.