Pith. sign in

REVIEW 2 cited by

Masked Event Modeling: Self-Supervised Pretraining for Event Cameras

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2212.10368 v3 pith:IKPSLKIN submitted 2022-12-20 cs.CV

classification cs.CV
keywords eventaccuracydatataskcamerasclassificationeventshigh
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Event cameras asynchronously capture brightness changes with low latency, high temporal resolution, and high dynamic range. However, annotation of event data is a costly and laborious process, which limits the use of deep learning methods for classification and other semantic tasks with the event modality. To reduce the dependency on labeled event data, we introduce Masked Event Modeling (MEM), a self-supervised framework for events. Our method pretrains a neural network on unlabeled events, which can originate from any event camera recording. Subsequently, the pretrained model is finetuned on a downstream task, leading to a consistent improvement of the task accuracy. For example, our method reaches state-of-the-art classification accuracy across three datasets, N-ImageNet, N-Cars, and N-Caltech101, increasing the top-1 accuracy of previous work by significant margins. When tested on real-world event data, MEM is even superior to supervised RGB-based pretraining. The models pretrained with MEM are also label-efficient and generalize well to the dense task of semantic image segmentation.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. SpikeWorld: Fast-State Adaptation for Frozen Spiking World Models

    cs.CV 2026-08 conditional novelty 6.0 of 10

    A frozen 1.45M-parameter spiking world model with a small external fast-state module raises frozen-policy reward by 7.90 (CI [2.48, 14.06]) and improves held-out prediction under shear and attenuation while inherited ...

  2. Making Every Event Count: Balancing Data Efficiency and Accuracy in Event Camera Subsampling

    cs.CV 2025-05 conditional novelty 6.0 of 10

    A causal, density-based event subsampling method preserves classification accuracy better than random, spatial, temporal, event-count, and corner-based baselines in sparse regimes, except when event counts vary widely...

Pith tools