Skip to content
Review Open access

A human-centered cognitive support framework for manual industrial assembly: integrating perception, agentic reasoning, and augmented reality guidance

Jul 2026 · The International Journal of Advanced Manufacturing Technology · Vol 145, pp. 6679 - 6709 · 0 citations · 127 references

TL;DR

A human-centered cognitive support framework for manual industrial assembly that integrates perception, agentic reasoning, knowledge grounding, and augmented reality guidance is proposed and translated into a layered implementation architecture comprising physical, perception, cognitive, guidance, knowledge, and application layers.

Abstract

Manual industrial assembly remains essential in high-variety and customized production, but increasing product and process complexity places substantial cognitive demands on operators. Existing assistance systems, particularly augmented reality-based solutions, improve instruction visualization and task guidance, yet they often remain weakly connected to the real assembly state and limited in reasoning, adaptation, and decision support. This paper proposes a human-centered cognitive support framework for manual industrial assembly that integrates perception, agentic reasoning, knowledge grounding, and augmented reality guidance. The study is informed by a systematic search, bibliometric overview, and literature analysis of 129 Scopus-indexed documents. The analysis shows that augmented reality dominates current cognitive support approaches, while AI-based methods are increasingly used for object recognition, contextual interpretation, adaptive information delivery, and error detection. However, perception, reasoning, guidance, and knowledge grounding are still commonly treated as isolated functions. Based on these findings, 11 review-derived design requirements are formulated and used to develop a conceptual perception-cognition-guidance framework that represents cognitive support as a closed human-centered loop grounded in procedures, constraints, rules, and memory. The framework is then translated into a layered implementation architecture comprising physical, perception, cognitive, guidance, knowledge, and application layers. Structured data contracts clarify how sensory and interaction data can be transformed into perception evidence, structured assembly states, cognitive support decisions, and device-specific guidance commands. A toy-train assembly demonstrator illustrates how procedural state modeling, multi-camera perception, YOLO11-based object detection, projector-based guidance, and contract-based data exchange can connect physical assembly events with structured reasoning and operator-facing feedback.

Read PDF

Similar papers

Open access Sep 2026

AR and Intelligent Assistance for Technicians: Design, Delivery, and Real-World Impact

AR and Intelligent Assistance for Technicians: Design, Delivery, and Real-World Impact Table of Contents CHAPTER 1 IntroductionCHAPTER 2 The Evolution of Technical Support and AssistanceCHAPTER 3 Foundational Principles of Augmented RealityCHAPTER 4 Cognitive Ergonomics and Human-Computer InteractionCHAPTER 5 Hardwar...

Unknown authors · 0 citations
Open access Sep 2026

Toward an Integrated Cognitive-Ergonomic Architecture for Human-Machine Interaction: Combining Cognitive Models with Human Factors Ergonomics

This paper presents an integrated approach to modeling human competencies by combining the theoretical foundations of cognitive architectures with principles from Human Factors Ergonomics (HFE). Through a comparative analysis of established cognitive models-SOAR, ACT-R, LIDA, and COCOM-we synthesize a tailored architec...

Antoine Lénat, Olivier Cheminat, Damien Chablat et al. · 0 citations
Oct 2026

Fostering Computational Thinking for an Adaptive Construction Workforce: Experimental Study of VR-Based Human–Robot Interaction

Construction work is becoming increasingly technology-intensive, requiring transferable skills that enable effective human–robot interaction (HRI). We designed and evaluated a virtual-reality (VR) training environment that pedagogically integrates computational thinking (CT) elements (i.e., decomposition, pattern rec...

Hameedreza Gucci, J. Morse, Amirhosein Jafari et al. · 0 citations
Open access Aug 2026

CARTA: Context-Aware Dual Retrieval and Chain-of-Thought Task Allocation for Environment-Grounded Elderly Care

Enabling elderly individuals to age independently at home requires intelligent assistive systems that can understand complex care needs and coordinate appropriate responses. While large language models (LLMs) show promise for adaptive assistance, current eldercare systems suffer from critical limitations: they generate...

Thanh Son Le, Huu-Sy Le, Le Minh Toan Truong et al. · 0 citations
Review Sep 2026

From Code to Collaboration: A Cognitive Agent Framework for Large Language Model (LLM)-Based Human-Vehicle Teaming

This study develops a human-centered cognitive-agent framework for understanding how large language models (LLMs) can support human-vehicle teaming in automated driving. Following PRISMA guidelines, we reviewed 1,126 records published between 2021 and 2025 and included 52 studies after screening and full-text assessmen...

Jing-Jie Wang, Brandon J. Pitts · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.