REVIEW 7 cited by
Spatial Functa: Scaling Functa to ImageNet Classification and Generation
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Neural fields, also known as implicit neural representations, have emerged as a powerful means to represent complex signals of various modalities. Based on this Dupont et al. (2022) introduce a framework that views neural fields as data, termed *functa*, and proposes to do deep learning directly on this dataset of neural fields. In this work, we show that the proposed framework faces limitations when scaling up to even moderately complex datasets such as CIFAR-10. We then propose *spatial functa*, which overcome these limitations by using spatially arranged latent representations of neural fields, thereby allowing us to scale up the approach to ImageNet-1k at 256x256 resolution. We demonstrate competitive performance to Vision Transformers (Steiner et al., 2022) on classification and Latent Diffusion (Rombach et al., 2022) on image generation respectively.
Forward citations
Cited by 7 Pith papers
-
NISF++: Geometrically-grounded implicit representations of 3D+time cardiac function from 2D short- and long-axis MR views
NISF++ reconstructs a continuous 3D+time cardiac intensity and segmentation field from mixed short- and long-axis 2D MRI views, with learned rigid slice motion correction.
-
Picasso: Holistic Scene Reconstruction with Physics-Constrained Sampling
Picasso produces multi-object scene reconstructions that are both geometrically accurate and physically plausible by using physics-constrained rejection sampling over an inferred contact graph, outperforming prior met...
-
VidFuncta: Towards Generalizable Neural Representations for Ultrasound Videos
VidFuncta encodes ultrasound videos into static and time-varying latent vectors, improving reconstruction over 2D and 3D baselines while enabling efficient downstream analysis.
-
Equivariant Eikonal Neural Networks: Grid-Free, Scalable Travel-Time Prediction on Homogeneous Spaces
E-NES uses Lie-group point-cloud conditioning and equivariant neural fields to make grid-free eikonal travel-time prediction steerable under rotations and translations, with complete invariant features and competitive...
-
PDEfuncta: Spectrally-Aware Neural Representation for PDE Solution Modeling
A Fourier-based weight modulation for shared INR networks improves reconstruction of high-frequency PDE fields and enables bidirectional inference between paired solution spaces.
-
Fourier-Modulated Implicit Neural Representation for Multispectral Satellite Image Compression
ImpliSat compresses multispectral satellite images using an implicit neural network with hypernetwork-generated Fourier modulations per band, reporting higher PSNR than shift and scale modulation baselines.
-
CINeMA: Conditional Implicit Neural Multi-Modal Atlas for a Spatio-Temporal Representation of the Perinatal Brain
A conditional implicit neural atlas trains on MRI and labels to generate high-resolution fetal and neonatal brain atlases in minutes, with conditioning on age, birth age, ventricle volume, and corpus callosum presence.
Discussion (0). Sign in to comment.