Pith. sign in

REVIEW 3 cited by

MinD-3D++: Advancing fMRI-Based 3D Reconstruction with High-Quality Textured Mesh Generation and a Comprehensive Dataset

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2409.11315 v3 pith:LQKYOFDF submitted 2024-09-17 cs.CV

classification cs.CV
keywords datasetfmridatamind-3dobjectstexturedfmri-objaversefmri-shape
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Reconstructing 3D visuals from functional Magnetic Resonance Imaging (fMRI) data, introduced as Recon3DMind, is of significant interest to both cognitive neuroscience and computer vision. To advance this task, we present the fMRI-3D dataset, which includes data from 15 participants and showcases a total of 4,768 3D objects. The dataset consists of two components: fMRI-Shape, previously introduced and available at https://huggingface.co/datasets/Fudan-fMRI/fMRI-Shape, and fMRI-Objaverse, proposed in this paper and available at https://huggingface.co/datasets/Fudan-fMRI/fMRI-Objaverse. fMRI-Objaverse includes data from 5 subjects, 4 of whom are also part of the core set in fMRI-Shape. Each subject views 3,142 3D objects across 117 categories, all accompanied by text captions. This significantly enhances the diversity and potential applications of the dataset. Moreover, we propose MinD-3D++, a novel framework for decoding textured 3D visual information from fMRI signals. The framework evaluates the feasibility of not only reconstructing 3D objects from the human mind but also generating, for the first time, 3D textured meshes with detailed textures from fMRI data. We establish new benchmarks by designing metrics at the semantic, structural, and textured levels to evaluate model performance. Furthermore, we assess the model's effectiveness in out-of-distribution settings and analyze the attribution of the proposed 3D pari fMRI dataset in visual regions of interest (ROIs) in fMRI signals. Our experiments demonstrate that MinD-3D++ not only reconstructs 3D objects with high semantic and spatial accuracy but also provides deeper insights into how the human brain processes 3D visual information. Project page: https://jianxgao.github.io/MinD-3D.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 3 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. MindAligner: Explicit Brain Functional Alignment for Cross-Subject Visual Decoding from Limited fMRI Data

    cs.CV 2025-02 conditional novelty 5.0 of 10

    MindAligner aligns a new subject's fMRI to a known subject's brain space with a low-rank transfer matrix and cross-stimulus losses, improving cross-subject visual decoding from one hour of data.

  2. Making Your Dreams A Reality: Decoding the Dreams into a Coherent Video Story from fMRI Signals

    cs.CV 2025-01 reject novelty 5.0 of 10

    A zero-shot pipeline claims to generate dream videos from sleep fMRI by transferring a model trained on awake visual perception, with only three dream segments validated.

  3. Multimodal Brain-Computer Interfaces: AI-powered Decoding Methodologies

    cs.HC 2025-02 conditional novelty 3.0 of 10

    This review organizes multimodal brain-computer interface decoding into three algorithmic task types and surveys AI methods for visual, speech, and affective decoding.

Pith tools