Skip to content
Preprint

STEAM: A Spatio-TEmporal Alignment Mixture-of-Experts Model with Hierarchical Pre-training for EEG Decoding

Aug 2026 · 0 citations · 47 references
Computer Science

TL;DR

STEAM is presented, a hierarchical transfer framework that reconciles general-purpose representation learning with paradigm-specific specialization in EEG foundation models and attains the best average rank among the compared methods at a competitive inference cost measured in FLOPs.

Abstract

Brain-computer interfaces (BCIs) have been widely used in motor rehabilitation, disease diagnosis, and other neural engineering scenarios. However, conventional neural signal decoding algorithms often suffer from limited generalizability and high adaptation costs, motivating recent interest in BCI foundation models. Existing approaches still struggle to jointly achieve general transferability, accurate decoding, and efficient downstream adaptation. We present STEAM, a hierarchical transfer framework that reconciles general-purpose representation learning with paradigm-specific specialization in EEG foundation models. The framework is instantiated as a dual-branch spatio-temporal encoder in which a shared soft mixture-of-experts (SSMoE) module aligns the spatial and temporal branches, allowing complementary representations to exchange information through a compact set of soft slots. Across seven downstream datasets and fourteen evaluation settings, STEAM attains the best average rank among the compared methods at a competitive inference cost measured in FLOPs. Building upon the Stage-I general initialization, the hierarchical pre-training strategy further specializes the model to a target paradigm without retraining from scratch, yielding consistent gains in paradigm-specific decoding accuracy.

View source

Similar papers

Aug 2026

HCFT: A Hierarchical Convolutional Fusion Transformer for Cross-Task EEG Decoding.

Electroencephalography (EEG) decoding remains challenging due to the non-stationary nature of neural signals and the limited generalization of existing models across tasks and subjects. To address this challenge, we propose a lightweight and generalizable decoding framework named Hierarchical Convolutional Fusion Trans...

Haodong Zhang, Jiapeng Zhu, Yitong Chen et al. · 0 citations
Jul 2026

EEGForceFusion: Joint Tokenised-Continuous Representation Learning for Subject-Independent Grasp Force Decoding

Brain-machine interfaces provide a link between neural activity and external devices, enabling restoration of motor function and advancing human-machine interaction using non-invasive electroencephalography (EEG). However, continuous grasp force decoding remains challenging due to complex temporal dynamics, high inter-...

Sankalp Sunil Turankar, Y. Meena · 0 citations
Review Open access Sep 2026

Cross-variability decoding for motor imagery EEG signals: a comprehensive review

A comprehensive taxonomy of MI EEG cross-variability decoding studies from 2020 to 2025 is presented, systematically organizing advances in deep learning and transfer learning and critically evaluate core algorithmic approaches, including Convolutional Neural Networks, transformers, feature alignment, domain adaptation...

Li-Jun Wang, Yue-Ying Zhou, Peng-Pai Wang et al. · 0 citations
Jul 2026

MSBraM: A Multi-scale Self-supervised Brain Foundation Model for Hierarchical EEG Dynamics Learning

Self-supervised foundation models have recently shown strong potential for electroencephalogram (EEG)-based analysis. However, existing approaches struggle to capture the inherently multi-scale temporal structure of EEG signals, where local neural patterns and long-range dependencies jointly encode task-relevant inform...

Tao Zhou, Jing Han, Lingyu Shu et al. · 0 citations
Open access Sep 2026

SEDAT: a hybrid tokenizer for large EEG models

Objective. The fidelity of neural representations learned by large EEG foundation models depends on how raw brain signals are tokenized. Existing methods suffer from arbitrary temporal boundaries misaligned with neural state transitions, neglecting inter-channel spatial information, and fixed segmentation criteria that...

Muhammad Zulkifal Aziz, Yue Zhuo, Bin-Wen Huang et al. · 0 citations
Open access Aug 2026

FoME: A foundation model for EEG using adaptive temporal-lateral attention scaling.

Electroencephalography (EEG) is a vital tool to measure and record brain activity in neuroscience and clinical applications, yet its potential is constrained by signal heterogeneity, low signal-to-noise ratios, and limited labeled datasets. In this paper, we propose FoME (Foundation Model for EEG), a novel approach usi...

Enze Shi, Kui Zhao, Qilong Yuan et al. · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.