Pith. sign in

REVIEW 2 cited by

CroSSL: Cross-modal Self-Supervised Learning for Time-series through Latent Masking

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2307.16847 v3 pith:JLFLJPXZ submitted 2023-07-31 cs.LG

classification cs.LG
keywords datalearningcross-modalcrosslmaskingmissinghandlinglabeled
verification ladder T0 review T1 audit T2 compute T3 formal

Signed reviews

No signed human review yet.

0 comments
read the original abstract

Limited availability of labeled data for machine learning on multimodal time-series extensively hampers progress in the field. Self-supervised learning (SSL) is a promising approach to learning data representations without relying on labels. However, existing SSL methods require expensive computations of negative pairs and are typically designed for single modalities, which limits their versatility. We introduce CroSSL (Cross-modal SSL), which puts forward two novel concepts: masking intermediate embeddings produced by modality-specific encoders, and their aggregation into a global embedding through a cross-modal aggregator that can be fed to down-stream classifiers. CroSSL allows for handling missing modalities and end-to-end cross-modal learning without requiring prior data preprocessing for handling missing inputs or negative-pair sampling for contrastive learning. We evaluate our method on a wide range of data, including motion sensors such as accelerometers or gyroscopes and biosignals (heart rate, electroencephalograms, electromyograms, electrooculograms, and electrodermal) to investigate the impact of masking ratios and masking strategies for various data types and the robustness of the learned representations to missing data. Overall, CroSSL outperforms previous SSL and supervised benchmarks using minimal labeled data, and also sheds light on how latent masking can improve cross-modal learning. Our code is open-sourced at https://github.com/dr-bell/CroSSL.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Wearable Accelerometer Foundation Models for Health via Knowledge Distillation

    cs.LG 2024-12 conditional novelty 6.0 of 10

    A knowledge-distilled accelerometry encoder, taught by an unsupervised PPG teacher on 20 million minutes of paired wearable data, predicts heart rate, heart-rate variability, demographics, and 46 health conditions fro...

  2. CiTrus: Squeezing Extra Performance out of Low-data Bio-signal Transfer Learning

    cs.LG 2024-12 conditional novelty 5.0 of 10

    A CNN-transformer hybrid with frequency-based masked autoencoding and source-frequency resampling improves low-data bio-signal transfer learning on several benchmark tasks.

Pith tools