Pith. sign in

REVIEW 22 cited by

Transformer-based Spatial-Temporal Feature Learning for EEG Decoding

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2106.11170 v1 pith:VLWJ42TM submitted 2021-06-11 eess.SP cs.AIcs.LG

Transformer-based Spatial-Temporal Feature Learning for EEG Decoding

classification eess.SP cs.AIcs.LG
keywords attentiondatadecodingtimetransformingcnnsdimensionfeatures
verification ladder T0 review T1 audit T2 compute T3 formal T4 reserved
0 comments
Share X Bluesky LinkedIn Reddit HN
read the original abstract

At present, people usually use some methods based on convolutional neural networks (CNNs) for Electroencephalograph (EEG) decoding. However, CNNs have limitations in perceiving global dependencies, which is not adequate for common EEG paradigms with a strong overall relationship. Regarding this issue, we propose a novel EEG decoding method that mainly relies on the attention mechanism. The EEG data is firstly preprocessed and spatially filtered. And then, we apply attention transforming on the feature-channel dimension so that the model can enhance more relevant spatial features. The most crucial step is to slice the data in the time dimension for attention transforming, and finally obtain a highly distinguishable representation. At this time, global averaging pooling and a simple fully-connected layer are used to classify different categories of EEG data. Experiments on two public datasets indicate that the strategy of attention transforming effectively utilizes spatial and temporal features. And we have reached the level of the state-of-the-art in multi-classification of EEG, with fewer parameters. As far as we know, it is the first time that a detailed and complete method based on the transformer idea has been proposed in this field. It has good potential to promote the practicality of brain-computer interface (BCI). The source code can be found at: \textit{https://github.com/anranknight/EEG-Transformer}.

discussion (0)

Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.

Forward citations

Cited by 22 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

  1. B[FM]$^2$: Brain Foundation Model via Flow Matching with SplitUNet

    cs.LG 2026-06 unverdicted novelty 7.0

    B[FM]^2 pretrains an EEG foundation model on raw signals with flow matching and SplitUNet, reaching SOTA on 7 of 9 tasks using ~30x less data and generating neurologist-indistinguishable synthetic EEG.

  2. EvoBrain: Continual Learning of EEG Foundation Models Across Heterogeneous BCI Tasks

    cs.AI 2026-06 unverdicted novelty 7.0

    EvoBrain introduces a continual learning method with Neuro-Spectral Task Normalization and Response-Affinity Distillation to enable unified EEG decoding across heterogeneous BCI tasks.

  3. CaMBRAIN: Real-time, Continuous EEG Inference with Causal State Space Models

    cs.AI 2026-05 unverdicted novelty 7.0

    CaMBRAIN introduces a causal Mamba-based SSM with a multi-stage self-supervised training pipeline that achieves SOTA results on three EEG datasets while enabling linear-time long-range inference.

  4. DARE-EEG: A Foundation Model for Mining Dual-Aligned Representation of EEG

    cs.AI 2026-05 unverdicted novelty 7.0

    DARE-EEG is a self-supervised EEG foundation model that enforces mask-invariance via contrastive mask alignment and momentum anchor alignment, plus conv-linear-probing for heterogeneous setups, achieving SOTA accuracy...

  5. Visualizing the Invisible: Generative Visual Grounding Empowers Universal EEG Understanding in MLLMs

    cs.AI 2026-05 unverdicted novelty 6.0

    Generative Visual Grounding creates visual proxy images from EEG to enhance MLLM understanding of brain signals beyond text-only alignment.

  6. Visualizing the Invisible: Generative Visual Grounding Empowers Universal EEG Understanding in MLLMs

    cs.AI 2026-05 unverdicted novelty 6.0

    Generative Visual Grounding creates instance-specific visual proxy images from EEG signals to enhance MLLM understanding of brain activity beyond text-only alignment.

  7. PRiSE-EEG: A Prior-Guided Foundation Model with Depth-Stratified Experts for Cross-Paradigm EEG Representation Learning

    eess.SP 2026-05 unverdicted novelty 6.0

    PRiSE-EEG is a prior-guided EEG foundation model that allocates shared and specialized experts across depth using CKA-derived sigmoid mappings and reports strong cross-paradigm results on 12 benchmarks.

  8. Brain-OF: An Omnifunctional Foundation Model for fMRI, EEG and MEG

    cs.LG 2026-02 unverdicted novelty 6.0

    Brain-OF is a multimodal foundation model for fMRI, EEG and MEG using any-resolution sampling, DINT attention with sparse MoE, and masked temporal-frequency pretraining on ~40 datasets to achieve superior downstream p...

  9. UniMind: Unleashing the Power of LLMs for Unified Multi-Task Brain Decoding

    cs.HC 2025-06 unverdicted novelty 6.0

    UniMind unifies multi-task brain decoding from EEG by bridging signals to LLMs via a Neuro-Language Connector and dynamic task queries, outperforming prior models by 12% on average across ten datasets.

  10. CodeBrain: Bridging Decoupled Tokenizer and Multi-Scale Architecture for EEG Foundation Model

    cs.LG 2025-06 unverdicted novelty 6.0

    CodeBrain introduces a decoupled TFDual-Tokenizer and multi-scale EEGSSM architecture for an EEG foundation model pretrained on a large corpus, claiming strong generalization across eight downstream tasks and ten datasets.

  11. Tokenizing Single-Channel EEG with Time-Frequency Motif Learning

    cs.LG 2025-02 unverdicted novelty 6.0

    TFM-Tokenizer learns a vocabulary of time-frequency motifs from single-channel EEG via a dual-path masked architecture and encodes signals into discrete tokens, reporting up to 11% Cohen's Kappa gains on benchmarks an...

  12. DiffEEG: A Self-Supervised Denoising Diffusion Model for Learning EEG Generic Representations

    cs.LG 2026-07 conditional novelty 5.5

    A diffusion U-Net pretrained on unlabeled TUHSZ EEG plus an F1-maximizing RL decision layer yields clinically usable patient-wise seizure detection and subtyping under severe class imbalance.

  13. Interpretable EEG biomarkers with bag-of-waves: Spatial and temporal waveform dictionaries for low-data regimes

    cs.LG 2026-07 conditional novelty 5.0

    A small unsupervised EEG waveform dictionary with transition counts matches deep-model performance on three tasks while staying interpretable.

  14. MSBraM: A Multi-scale Self-supervised Brain Foundation Model for Hierarchical EEG Dynamics Learning

    cs.AI 2026-07 conditional novelty 5.0

    A multi-scale, VQ-VAE-based EEG foundation model with curriculum masking beats prior EEG foundation models on 11 of 12 benchmarks, but two test sets overlap with its pretraining corpus.

  15. Foundation Models for Epileptogenic Zone Identification in Drug-Resistant Epilepsy

    cs.LG 2026-06 unverdicted novelty 5.0

    A signal foundation model trained on over 100,000 minutes of sEEG plus a language model achieves 0.978 contact-level PPV for epileptogenic zone identification under leave-one-patient-out evaluation.

  16. BEAM: Brainwave Empathy Assessment Model for Early Childhood

    cs.LG 2025-09 conditional novelty 5.0

    BEAM, a multi-view EEG deep learning model, predicts high vs low empathy in 4-6 year olds with 64.7% accuracy and 0.008 standard deviation on 57 children.

  17. Dynamic Survival Prediction using Longitudinal Images based on Transformer

    eess.IV 2025-08 conditional novelty 5.0

    A Transformer combining vision and sequence encoders with a Cox survival head for dynamic survival prediction from longitudinal MRI, evaluated on Alzheimer's disease data.

  18. MSCGC-KAN: Multi-scale Causal Graph Convolution and Kolmogorov-Arnold Feature Mapping for EEG Emotion Recognition

    cs.CV 2026-05 unverdicted novelty 4.0

    MSCGC-KAN adds multi-scale causal graph convolution and Kolmogorov-Arnold feature mapping as a structured task head on a pre-trained CBraMod backbone, reporting balanced accuracy gains of 5.91 and 2.03 points on FACED...

  19. Towards Unified Multi-task EEG Analysis with Low-Rank Adaptation

    cs.LG 2026-04 unverdicted novelty 4.0

    MTEEG uses task-specific LoRA modules to jointly adapt a pre-trained EEG model across multiple tasks, outperforming single-task baselines on most metrics in evaluations on six downstream tasks.

  20. Task-guided Spatiotemporal Network with Diffusion Augmentation for EEG-based Dementia Diagnosis and MMSE Prediction

    cs.LG 2026-04 unverdicted novelty 4.0

    TGSN reports 97.78% accuracy on AD/FTD classification and RMSE of 1.93 for MMSE prediction on the XY02 EEG dataset, outperforming baselines by large margins.

  21. NeuroWeaver: An Autonomous Evolutionary Agent for Exploring the Programmatic Space of EEG Analysis Pipelines

    cs.AI 2026-02 unverdicted novelty 4.0

    NeuroWeaver reformulates EEG pipeline design as constrained evolutionary optimization with domain-informed initialization, yielding lightweight pipelines that outperform task-specific methods and match foundation mode...

  22. Transformer Based Model for Spatiotemporal Feature Learning in EEG Emotion Recognition

    cs.LG 2026-06 unverdicted novelty 3.0

    EEG-TransNet combines wavelet denoising, ResNet feature extraction, local self-attention, and a fuzzy-attention synchronous transformer to outperform prior methods on BETA, SEED, and DepEEG datasets for EEG emotion re...