REVIEW 4 cited by
Spatial Temporal Graph Convolutional Networks for Skeleton-Based Action Recognition
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Spatial Temporal Graph Convolutional Networks for Skeleton-Based Action Recognition
read the original abstract
Dynamics of human body skeletons convey significant information for human action recognition. Conventional approaches for modeling skeletons usually rely on hand-crafted parts or traversal rules, thus resulting in limited expressive power and difficulties of generalization. In this work, we propose a novel model of dynamic skeletons called Spatial-Temporal Graph Convolutional Networks (ST-GCN), which moves beyond the limitations of previous methods by automatically learning both the spatial and temporal patterns from data. This formulation not only leads to greater expressive power but also stronger generalization capability. On two large datasets, Kinetics and NTU-RGBD, it achieves substantial improvements over mainstream methods.
Forward citations
Cited by 4 Pith papers
-
Physical Self-Supervised Learning: IMU Sensing without Manual Labels
A label-free IMU sensing framework whose neural encoder feeds a learnable physics decoder achieves tracking and mocap accuracy that beats supervised baselines on public benchmarks.
-
Physical Self-Supervised Learning: IMU Sensing without Manual Labels
A physics-based self-supervised autoencoder achieves label-free IMU tracking and motion capture that outperforms supervised baselines in generalization tests.
-
StormNet: Improving storm surge predictions with a GNN-based spatio-temporal offset forecasting model
A spatio-temporal GNN model reduces storm surge water-level forecast RMSE by more than 70% for 48-hour horizons and over 50% for 72-hour horizons on U.S. Gulf Coast hurricane data.
-
CascadeFormer: A Family of Two-stage Cascading Transformers for Skeleton-based Human Action Recognition
A masked-pretrained skeleton transformer with a second fine-tuning transformer and cross-attention fusion reaches 94.66% on Penn Action, 91.16% on N-UCLA, and 81.01%/88.17% on NTU RGB+D 60 cross-subject/cross-view.
discussion (0)
Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.