Skip to content
#explainable ai Open access

Intelligent Martial Arts Coaching Framework Using Artificial Intelligence for Real-Time Action Detection and Performance Feedback

Sep 2026 · International Journal of e-collaboration · 0 citations · 16 references
Human Pose and Action Recognition

TL;DR

MMAF-Net is proposed, a multi-branch deep learning architecture that integrates visual, pose-estimation and inertial sensor streams to extract complementary motion features, fused via a temporal attention module to classify 23 karate action types accurately.

Abstract

Karate training relies on subjective manual evaluation of intricate motions, hindering scalable, consistent coaching. This paper proposes MMAF-Net, a multimodal AI framework for real-time karate action recognition and automated performance feedback. Its three-branch deep learning architecture integrates visual, pose-estimation and inertial sensor streams to extract complementary motion features, fused via a temporal attention module to classify 23 karate action types accurately. Trained and validated on MS-KARD (2.8 million video frames and 5.6 million sensor readings from dual cameras and three IMUs), the model generates explainable coaching tips through rule-based modules referencing prediction confidence, posture bias and motion stability. Tests yield 96.3% accuracy and 95.1% F1-score, outperforming benchmarks like KarateNet. With only 24 ms inference latency, this real-time system suits interactive martial arts training scenarios.

Read PDF

Similar papers

Open access Aug 2026

Martial Arts Routine Action Recognition and Scoring Using YOLOv8 and OpenPose

A joint framework combining YOLOv8 and time-optimized OpenPose to mitigate pose estimation jitter and detection inaccuracies caused by rapid motion and occlusion in human motion analysis and offers technical reference for multimodal perception and dynamic scene understanding in advanced electromagnetic sensing applicat...

D. Zhao, Y.-Q. Ma · 0 citations
Conference Sep 2026

Deep pose estimation-based action recognition and performance analysis for intelligent motion understanding

Human motion understanding requires not only accurate action recognition but also interpretable performance evaluation capable of reflecting motion quality. This paper proposes Pose-ARPA, a unified framework that combines deep pose estimation with spatiotemporal representation learning for comprehensive action recognit...

Qi Yang, Long-Hui Wen · 0 citations
Review Sep 2026

An Integrated Video-AI Platform for Action-Level Microanastomosis Training and Performance Feedback

Developing microanastomosis skill requires repeated practice with timely, action-specific feedback, yet expert review of lengthy microscope videos does not scale to frequent or distributed training. We present an integrated video-AI platform that turns a complete simulated procedure into inspectable, interactive feedba...

Yan Meng, Daniel A. Donoho · 0 citations
Open access Aug 2026

Design of University Instrumental Music Performance Movement Recognition and Accurate Feedback System Integrating OpenPose and LSTM

Accurate recognition of fine-grained human movements with low-latency feedback is fundamental to intelligent perception systems and real-time human–machine interaction in modern engineering applications. This study presents an end-to-end motion recognition and feedback framework for university instrumental music perfor...

X. Zhou, W. He · 0 citations
Preprint Aug 2026

Autonomous Telerehabilitation via Skeletal Motion Prediction and Joint-Level Performance Assessment

Autonomous rehabilitation systems must not only recognize human motion but also provide structured feedback to support users without continuous therapist supervision. This paper presents a telerehabilitation pipeline that integrates skeleton-based exercise quality assessment and short-term motion prediction into a two-...

Lara Pereira, J. Paulo, P. Santos et al. · 0 citations
Open access Aug 2026

Research on Rehabilitation Training Movement Recognition and Real-time Feedback Model Based on Computer Vision

This study provides a replicable technical path for the validation of rehabilitation evaluation algorithms without clinical data collection through the adaptive fusion mechanism to dynamically integrate the confidence of the deep network and the matching score of dynamic time warping template.

Mingxiang Yang · 1 citation

Related blog posts

Microsoft Research Blog Oct 7, 2026

Agent Lightning v1.0: A 3,500-Line Lightweight Agentic RL Framework for Training Agents with Real Harnesses

Training AI agents with reinforcement learning can be challenging because their tools, context, and decision-making are managed by complex frameworks. Agent Lightning connects existing agents to RL training, making it easier to improve them without rebuilding them. The post Agent Lightning v1.0: A 3,500-Line Lightweight Agentic RL Framework for Training Agents with Real Harnesses appeared first on Microsoft Research.

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.