REVIEW 3 cited by
DISCO: accurate Discrete Scale Convolutions
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
Scale is often seen as a given, disturbing factor in many vision tasks. When doing so it is one of the factors why we need more data during learning. In recent work scale equivariance was added to convolutional neural networks. It was shown to be effective for a range of tasks. We aim for accurate scale-equivariant convolutional neural networks (SE-CNNs) applicable for problems where high granularity of scale and small kernel sizes are required. Current SE-CNNs rely on weight sharing and kernel rescaling, the latter of which is accurate for integer scales only. To reach accurate scale equivariance, we derive general constraints under which scale-convolution remains equivariant to discrete rescaling. We find the exact solution for all cases where it exists, and compute the approximation for the rest. The discrete scale-convolution pays off, as demonstrated in a new state-of-the-art classification on MNIST-scale and on STL-10 in the supervised learning setting. With the same SE scheme, we also improve the computational effort of a scale-equivariant Siamese tracker on OTB-13.
Forward citations
Cited by 3 Pith papers
-
EquiFusion: Kinematics-Agnostic Human Motion Prediction via Equivariant Latent Diffusion
A permutation-equivariant latent diffusion model treats skeleton connectivity as input, enabling the first kinematics-agnostic stochastic human motion predictor that generalizes zero-shot to unseen and partial skeletons.
-
MACRO: Training-free Multi-plane Attention for Closeup Render Optimization
Training-free multi-plane attention with image-space scale-matched reference crops restores correct close-up detail from 3DGS without retraining the enhancer.
-
Neural Operators for Forward and Inverse Potential-Density Mappings in Classical Density Functional Theory
In 1D hard-rod cDFT, Fourier neural operators learn the density-to-direct-correlation-function map more accurately than DeepONet variants and dense networks, with squared ReLU giving the best extrapolation.
Discussion (0). Continue with ORCID to comment.