Skip to content
Conference

SCLAT: An LLM-Interpretable User Story Quality Evaluation Framework

Jul 2026 · Annual International Computer Software and Applications Conference · pp. 406-415 · 0 citations · 28 references
Computer Science

Abstract

User stories serve as the core requirement carriers in agile development, and their quality directly determines the development team's understanding of requirements, as well as software delivery efficiency and quality. With the expansion of software specifications, user story quality evaluation is constrained by issues such as subjectivity and low efficiency in manual assessment, as well as lack of comprehensiveness and insufficient accuracy in the evaluation dimensions of automated methods, such as Natural Language Processing(NLP), Machine Learning(ML), and Large Language Models (LLM). This paper proposes an LLM-interpretable user story quality evaluation framework. Constructed by integrating the 3C principle, IN-VEST criteria, and IEEE 830 standards, the framework divides 40 core quality attributes into four dimensions: Structural Specification, Content Specification, Logical Soundness, and Actionability & Traceability (collectively SCLAT framework). To complement the framework, an LLM-interpretable application paradigm is proposed (known as K-CoT). This paradigm provides detailed, LLM-tailored designs for core attributes, introduces a negative example guidance mechanism, and a Chain-of-Thought (CoT) prompting strategy. Experimental validation on the NFDI4Cat dataset shows that the SCLAT framework enables LLMs to achieve an average F1-score of 89.9%, representing a significant improvement over traditional frameworks. The experimental results indicate that the interpretability design for LLM can comprehensively improve the dimensions, efficiency, and quality of user story defect detection.

View source

Similar papers

Preprint Aug 2026

Large Language Models for Requirements Engineering: A Cross-Task Empirical Evaluation

This work presents the first cross-task empirical evaluation of LLMs spanning five RE-related activities, as well as replication materials supporting reproducibility, and a broader understanding of the capabilities, limitations, and practical readiness of current LLMs for RE.

Jacek Dabrowski, Manjeshwar Aniruddh Mallya, Alessio Ferrari et al. · 1 citation
Preprint Aug 2026

How Well Do LLMs Generate Taxonomies in the SE Domain? A Multi-perspective Evaluation Framework

It is suggested that TnT-LLM and CLIMB can be used in practical situations in the SE domain, while researchers should first assess the complexity of the generated taxonomies and their cost using a subset of the target data to decide whether to use automated methods or human experts.

Sota Nakashima, Yuta Ishimoto, Masanari Kondo et al. · 0 citations
Conference Aug 2026

Human-LLM Collaboration for Context-Dependent Requirements Engineering Tasks

Many requirements engineering (RE) tasks, such as requirements elicitation, documentation, and quality assessment, are inherently context-sensitive: what counts as a missing, wellwritten, or defective requirement varies by stakeholder-intent, domain, process, as well as countless other potential factors. Existing autom...

Max Unterbusch · 0 citations
Book Open access Oct 2026

From Features to Value: A Metric Planner for Linking UX and Business Outcomes

The Feature Metric Planner is a web-based interactive tool designed to operationalize feature-level metric planning and is perceived as useful for structuring metric planning early in the development process, improving alignment between feature design and UX-related metrics, and reducing the effort required to define m...

Gessé Evangelista, Tatiana Alencar, M. Sami et al. · 0 citations
Open access Mar 2025

LLMs’ reshaping of people, processes, products, and society in software development: a qualitative exploration with early adopters

Interviews with sixteen early-adopter software professionals who integrated LLM-based tools into their day-to-day work in early to mid-2023 offer actionable implications for developers, organizations, educators, and tool designers seeking to integrate LLMs responsibly into professional software practice.

Benyamin T. Tabarsi, Heidi Reichert, Sam Gilson et al. · 22 citations · ⚡1
Open access 2026

From EventStorming Artifacts to User Stories: A Semi-Automated Requirements Extraction Approach

Requirements engineering is a critical phase of software development that directly affects project scope, cost, and quality. In software development companies, a requirements list is typically prepared before creating a commercial proposal and signing a contract. EventStorming workshops are widely used for requirements...

Mantas Jurgelaitis, Antanas Ramanauskas, Gvidas Ambrozaitis et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.