REVIEW 2 cited by
BrainVis: Exploring the Bridge between Brain and Visual Signals via Image Reconstruction
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Analyzing and reconstructing visual stimuli from brain signals effectively advances the understanding of human visual system. However, the EEG signals are complex and contain significant noise. This leads to substantial limitations in existing works of visual stimuli reconstruction from EEG, such as difficulties in aligning EEG embeddings with the fine-grained semantic information and a heavy reliance on additional large self-collected dataset for training. To address these challenges, we propose a novel approach called BrainVis. Firstly, we divide the EEG signals into various units and apply a self-supervised approach on them to obtain EEG time-domain features, in an attempt to ease the training difficulty. Additionally, we also propose to utilize the frequency-domain features to enhance the EEG representations. Then, we simultaneously align EEG time-frequency embeddings with the interpolation of the coarse and fine-grained semantics in the CLIP space, to highlight the primary visual components and reduce the cross-modal alignment difficulty. Finally, we adopt the cascaded diffusion models to reconstruct images. Using only 10\% training data of the previous work, our proposed BrainVis outperforms state of the arts in both semantic fidelity reconstruction and generation quality. The code is available at https://github.com/RomGai/BrainVis.
Forward citations
Cited by 2 Pith papers
-
Uncovering the EEG Temporal Representation of Low-dimensional Object Properties
Using a pre-trained EEG decoder and temporal masking, the authors find concept-specific activation windows and prototypical temporal clusters in THINGS-EEG data.
-
CATVis: Context-Aware Thought Visualization
CATVis combines a Conformer EEG classifier, CLIP-based caption retrieval and re-ranking, and Stable Diffusion to generate images from EEG, reporting large gains over prior work.
Discussion (0). Sign in to comment.