REVIEW 13 cited by
Spatial Functa: Scaling Functa to ImageNet Classification and Generation
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
Neural fields, also known as implicit neural representations, have emerged as a powerful means to represent complex signals of various modalities. Based on this Dupont et al. (2022) introduce a framework that views neural fields as data, termed *functa*, and proposes to do deep learning directly on this dataset of neural fields. In this work, we show that the proposed framework faces limitations when scaling up to even moderately complex datasets such as CIFAR-10. We then propose *spatial functa*, which overcome these limitations by using spatially arranged latent representations of neural fields, thereby allowing us to scale up the approach to ImageNet-1k at 256x256 resolution. We demonstrate competitive performance to Vision Transformers (Steiner et al., 2022) on classification and Latent Diffusion (Rombach et al., 2022) on image generation respectively.
Forward citations
Cited by 13 Pith papers
-
NISF++: Geometrically-grounded implicit representations of 3D+time cardiac function from 2D short- and long-axis MR views
NISF++ reconstructs a continuous 3D+time cardiac intensity and segmentation field from mixed short- and long-axis 2D MRI views, with learned rigid slice motion correction.
-
Weight-Space Mixture-of-Experts for Implicit Neural Representation Classification
A hierarchical mixture-of-experts transformer that classifies images from the weights of implicit neural representations achieves state-of-the-art accuracy in weight-space learning.
-
Picasso: Holistic Scene Reconstruction with Physics-Constrained Sampling
Picasso produces multi-object scene reconstructions that are both geometrically accurate and physically plausible by using physics-constrained rejection sampling over an inferred contact graph, outperforming prior met...
-
Seeing the Many: Exploring Parameter Distributions Conditioned on Features in Surrogates
A Bayesian system that samples and visualizes the distribution of simulation parameters consistent with a user-specified output feature, using a density prior and Hamiltonian Monte Carlo.
-
VidFuncta: Towards Generalizable Neural Representations for Ultrasound Videos
VidFuncta encodes ultrasound videos into static and time-varying latent vectors, improving reconstruction over 2D and 3D baselines while enabling efficient downstream analysis.
-
Equivariant Eikonal Neural Networks: Grid-Free, Scalable Travel-Time Prediction on Homogeneous Spaces
E-NES uses Lie-group point-cloud conditioning and equivariant neural fields to make grid-free eikonal travel-time prediction steerable under rotations and translations, with complete invariant features and competitive...
-
Efficient Neural Video Representation with Temporally Coherent Modulation
NVTM uses flow-guided coordinate alignment and shared modulation codes from 2D grids to speed up and slim down implicit neural video representation.
-
SCENT: Robust Spatiotemporal Learning for Continuous Scientific Data via Scalable Conditioned Neural Fields
SCENT, a single-stage transformer-based conditioned neural field, jointly handles reconstruction, interpolation, and forecasting on sparse and noisy scientific data, reporting state-of-the-art results on Navier-Stokes...
-
Level-Set Parameters: Novel Representation for 3D Shape Analysis
Aligned SDF network weights, conditioned on pose by a hypernetwork, act as a continuous 3D shape representation that performs well on classification, retrieval, and pose estimation.
-
INRFlow: Flow Matching for INRs in Ambient Space
A domain-agnostic PerceiverIO-style transformer, trained with a point-wise flow-matching loss in ambient space, generates competitive images, point clouds, and protein structures without a separate data compressor.
-
PDEfuncta: Spectrally-Aware Neural Representation for PDE Solution Modeling
A Fourier-based weight modulation for shared INR networks improves reconstruction of high-frequency PDE fields and enables bidirectional inference between paired solution spaces.
-
Fourier-Modulated Implicit Neural Representation for Multispectral Satellite Image Compression
ImpliSat compresses multispectral satellite images using an implicit neural network with hypernetwork-generated Fourier modulations per band, reporting higher PSNR than shift and scale modulation baselines.
-
CINeMA: Conditional Implicit Neural Multi-Modal Atlas for a Spatio-Temporal Representation of the Perinatal Brain
A conditional implicit neural atlas trains on MRI and labels to generate high-resolution fetal and neonatal brain atlases in minutes, with conditioning on age, birth age, ventricle volume, and corpus callosum presence.
Discussion (0). Continue with ORCID to comment.