REVIEW 4 cited by
HYperbolic Self-Paced Learning for Self-Supervised Skeleton-based Action Representations
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Self-paced learning has been beneficial for tasks where some initial knowledge is available, such as weakly supervised learning and domain adaptation, to select and order the training sample sequence, from easy to complex. However its applicability remains unexplored in unsupervised learning, whereby the knowledge of the task matures during training. We propose a novel HYperbolic Self-Paced model (HYSP) for learning skeleton-based action representations. HYSP adopts self-supervision: it uses data augmentations to generate two views of the same sample, and it learns by matching one (named online) to the other (the target). We propose to use hyperbolic uncertainty to determine the algorithmic learning pace, under the assumption that less uncertain samples should be more strongly driving the training, with a larger weight and pace. Hyperbolic uncertainty is a by-product of the adopted hyperbolic neural networks, it matures during training and it comes with no extra cost, compared to the established Euclidean SSL framework counterparts. When tested on three established skeleton-based action recognition datasets, HYSP outperforms the state-of-the-art on PKU-MMD I, as well as on 2 out of 3 downstream tasks on NTU-60 and NTU-120. Additionally, HYSP only uses positive pairs and bypasses therefore the complex and computationally-demanding mining procedures required for the negatives in contrastive techniques. Code is available at https://github.com/paolomandica/HYSP.
Forward citations
Cited by 4 Pith papers
-
Less is More: Compact-Token Masked Feature Prediction for Skeleton Representation Learning
A decoder-free teacher-student masked feature predictor with semantic tube masking and skeleton-aware augmentations reaches state-of-the-art accuracy on NTU-60/120 and PKU-MMD II using a compact 8x25 token grid.
-
Towards Efficient General Feature Prediction in Masked Skeleton Modeling
GFP predicts hierarchical high-level features from a jointly trained lightweight target network instead of reconstructing joint coordinates, speeding up masked skeleton pretraining 6.2x while improving downstream accuracy.
-
Learning Along the Arrow of Time: Hyperbolic Geometry for Backward-Compatible Representation Learning
HBCT lifts embeddings into Lorentz hyperbolic space, uses entailment cones to keep new embeddings inside old ones' cones, and weights contrastive alignment by an uncertainty estimate, improving backward-compatible ret...
-
3D Skeleton-Based Action Recognition: A Review
A task-oriented review of skeleton-based action recognition that reorganizes known methods along a data processing pipeline and contains no new experimental result.
Discussion (0). Continue with ORCID to comment.