super hub Mixed citations

ShapeNet: An Information-Rich 3D Model Repository

Angel X. Chang, Leonidas Guibas, Pat Hanrahan, Qixing Huang, Thomas Funkhouser, Zimo Li · 2015 · cs.GR · arXiv 1512.03012

Mixed citation behavior. Most common role is background (57%).

134 Pith papers citing it

Background 57% of classified citations

open full Pith review browse 134 citing papers more from Angel X. Chang arXiv PDF

abstract

We present ShapeNet: a richly-annotated, large-scale repository of shapes represented by 3D CAD models of objects. ShapeNet contains 3D models from a multitude of semantic categories and organizes them under the WordNet taxonomy. It is a collection of datasets providing many semantic annotations for each 3D model such as consistent rigid alignments, parts and bilateral symmetry planes, physical sizes, keywords, as well as other planned annotations. Annotations are made available through a public web-based interface to enable data visualization of object attributes, promote data-driven geometric analysis, and provide a large-scale quantitative benchmark for research in computer graphics and vision. At the time of this technical report, ShapeNet has indexed more than 3,000,000 models, 220,000 models out of which are classified into 3,135 categories (WordNet synsets). In this report we describe the ShapeNet effort as a whole, provide details for all currently available datasets, and summarize future plans.

hub tools

JSON dossier citing papers JSON arXiv source

citation-role summary

dataset 12 background 8 method 1

citation-polarity summary

background 12 use dataset 7 unclear 1 use method 1

claims ledger

abstract We present ShapeNet: a richly-annotated, large-scale repository of shapes represented by 3D CAD models of objects. ShapeNet contains 3D models from a multitude of semantic categories and organizes them under the WordNet taxonomy. It is a collection of datasets providing many semantic annotations for each 3D model such as consistent rigid alignments, parts and bilateral symmetry planes, physical sizes, keywords, as well as other planned annotations. Annotations are made available through a public web-based interface to enable data visualization of object attributes, promote data-driven geometri

authors

Angel X. Chang Leonidas Guibas Pat Hanrahan Qixing Huang Thomas Funkhouser Zimo Li

co-cited works

representative citing papers

Towards Realistic 3D Emission Materials: Dataset, Baseline, and Evaluation for Emission Texture Generation

cs.CV · 2026-04-13 · unverdicted · novelty 8.0

The work creates the first dataset and baseline for generating emission textures on 3D objects to reproduce glowing materials from input images.

ARKitScenes: A Diverse Real-World Dataset For 3D Indoor Scene Understanding Using Mobile RGB-D Data

cs.CV · 2021-11-17 · accept · novelty 8.0

ARKitScenes is the largest real-world indoor RGB-D dataset captured with mobile LiDAR, including high-resolution depth maps and 3D furniture bounding box annotations for advancing object detection and depth upsampling.

WarpHammer: Densifying Scene Warps with 3D Object Priors for Extreme View Synthesis

cs.CV · 2026-06-30 · unverdicted · novelty 7.0

WarpHammer densifies scene warps with 3D object priors from generative models and fuses pose-unknown auxiliary views via multi-view geometry to enable stable extreme novel view synthesis.

3D-CoS: A New 3D Reconstruction Paradigm Based on VLM Code Synthesis

cs.CV · 2026-06-09 · unverdicted · novelty 7.0

3D-CoS represents 3D objects as Blender code generated by VLMs, with workflows for planning, RAG, and agents, showing better edit fidelity than point-cloud baselines.

Rethinking 3D Shape Generation: Diffusion over Superquadrics

cs.CV · 2026-06-08 · unverdicted · novelty 7.0

Diffusion for 3D shapes is moved from dense geometry to compact superquadric parameter sets, cutting state size to roughly 7 KB per shape and enabling faster generation plus new editing capabilities.

Why Far Looks Up: Probing Spatial Representation in Vision-Language Models

cs.CV · 2026-05-28 · conditional · novelty 7.0

VLMs exhibit consistent vertical-distance entanglement in embeddings from perspective bias in natural images, producing accuracy gaps that a new synthetic benchmark SpatialTunnel exposes as model-intrinsic.

Category-Level 3D Correspondence in Camera Space via Morphable Object Priors

cs.CV · 2026-05-27 · unverdicted · novelty 7.0

Morpheus learns morphable category-level shape priors to produce implicit 3D correspondences in camera space without explicit supervision and releases the HouseCorr3D benchmark with amodal and symmetry annotations.

Metric--Phase Fields: Decoupling Distance and Sign for Thin-Structure Reconstruction from Unoriented Point Clouds

cs.CV · 2026-05-25 · unverdicted · novelty 7.0

Metric-Phase Fields decouple unsigned metric proximity from a smooth phase field with learnable sharpness to enable faithful reconstruction of thin and open structures from point clouds.

ArtSplat: Feed-Forward Articulated 3D Gaussian Splatting from Sparse Multi-State Uncalibrated Views

cs.CV · 2026-05-23 · unverdicted · novelty 7.0

ArtSplat is the first feed-forward framework for articulated 3D Gaussian Splatting that reconstructs geometry and joints from sparse multi-state uncalibrated views in one pass.

MAPS: A Synthetic Dataset for Probing Vision Models in a Controlled 3D Scene Space

cs.CV · 2026-05-19 · unverdicted · novelty 7.0

MAPS provides 2618 validated 3D meshes and a controllable rendering pipeline to attribute vision model recognition failures to specific scene parameters, finding camera distance and elevation as the dominant failure factors across 20 tested models.

OffsetAxis: UDF Mesh Reconstruction via Offset-Volume Medial Axis Extraction

cs.GR · 2026-05-14 · unverdicted · novelty 7.0

OffsetAxis reconstructs meshes from unsigned distance fields by extracting the medial axis of the alpha-offset volume using ray casting and variational medial ball optimization.

Min Generalized Sliced Gromov Wasserstein: A Scalable Path to Gromov Wasserstein

cs.LG · 2026-05-13 · unverdicted · novelty 7.0

min-GSGW learns coupled nonlinear slicers to produce a rigid-motion-invariant, scalable approximation to the Gromov-Wasserstein distance and its transport plans.

Img2CADSeq: Image-to-CAD Generation via Sequence-Based Diffusion

cs.CV · 2026-05-13 · unverdicted · novelty 7.0

Img2CADSeq generates standard CAD sequences from images via a multi-stage pipeline with three-level hierarchical codebook encoding, importance-guided compression, and contrastive point-cloud conditioning of a VQ-Diffusion model, outperforming prior methods on new CAD-220K and PrintCAD datasets.

Count Anything at Any Granularity

cs.CV · 2026-05-11 · unverdicted · novelty 7.0

Multi-grained counting is introduced with five granularity levels, supported by the new KubriCount dataset generated via 3D synthesis and editing, and HieraCount model that combines text and visual exemplars for improved accuracy.

The Wittgensteinian Representation Hypothesis: Is Language the Attractor of Multimodal Convergence?

cs.AI · 2026-05-10 · unverdicted · novelty 7.0

Language representations serve as the asymptotic attractor for convergence in independently trained multimodal neural networks due to feature density asymmetry.

MeshFIM: Local Low-Poly Mesh Editing via Fill-in-the-Middle Autoregressive Generation

cs.GR · 2026-05-09 · unverdicted · novelty 7.0

MeshFIM enables local low-poly mesh editing by autoregressively filling target regions conditioned on context, using boundary markers, positional embeddings, and a gated geometry encoder to enforce attachment, topology, and region limits.

Rollback-Free Stable Brick Structures Generation

cs.LG · 2026-05-07 · unverdicted · novelty 7.0

Reinforcement learning internalizes physical stability rules for brick structures, enabling the first rollback-free generation with orders-of-magnitude faster inference.

Two Steps Are All You Need: Efficient 3D Point Cloud Anomaly Detection with Consistency Models

cs.CV · 2026-05-06 · unverdicted · novelty 7.0

Consistency learning reformulates 3D point cloud anomaly detection to predict clean geometry directly in one or two steps, yielding up to 80 times faster inference while matching state-of-the-art accuracy.

ADS: Random Sampling of Occupancy Functions using Adaptive Delaunay Scaffolding

cs.GR · 2026-05-05 · unverdicted · novelty 7.0

ADS adaptively refines a Delaunay scaffold to produce unbiased random samples on occupancy function surfaces together with a connecting mesh, using far fewer evaluations than existing approaches.

Generative Modeling with Orbit-Space Particle Flow Matching

cs.GR · 2026-05-04 · unverdicted · novelty 7.0

OGPP is a particle flow-matching method using orbit-space canonicalization and geometric paths that achieves lower error and fewer steps than prior approaches on 3D benchmarks.

AirZoo: A Unified Large-Scale Dataset for Grounding Aerial Geometric 3D Vision

cs.CV · 2026-04-29 · unverdicted · novelty 7.0 · 2 refs

AirZoo is a new dataset covering 378 regions across 22 countries with pixel-level metric depth and 6-DoF poses, shown via benchmarks to improve SoTA models on aerial image retrieval, cross-view matching, and multi-view 3D reconstruction.

Topo-ADV: Generating Topology-Driven Imperceptible Adversarial Point Clouds

cs.CV · 2026-04-10 · unverdicted · novelty 7.0

Topo-ADV uses differentiable persistent homology to create topology-altering perturbations that achieve up to 100% attack success on point cloud classifiers like PointNet while remaining geometrically imperceptible.

Training-free Spatially Grounded Geometric Shape Encoding (Technical Report)

cs.CV · 2026-04-08 · unverdicted · novelty 7.0 · 2 refs

XShapeEnc encodes arbitrary 2D spatially grounded shapes into compact invertible representations by decomposing them into unit-disk geometry and harmonic pose fields then applying Zernike bases with frequency propagation.

3D-Fixer: Coarse-to-Fine In-place Completion for 3D Scenes from a Single Image

cs.CV · 2026-04-06 · unverdicted · novelty 7.0

3D-Fixer performs in-place 3D asset completion from single-view partial point clouds via coarse-to-fine generation with ORFA conditioning, plus a new ARSG-110K dataset, to achieve higher geometric accuracy than MIDI and Gen3DSR while keeping diffusion efficiency.

citing papers explorer

Showing 50 of 134 citing papers.

Towards Realistic 3D Emission Materials: Dataset, Baseline, and Evaluation for Emission Texture Generation cs.CV · 2026-04-13 · unverdicted · none · ref 3 · internal anchor
The work creates the first dataset and baseline for generating emission textures on 3D objects to reproduce glowing materials from input images.
ARKitScenes: A Diverse Real-World Dataset For 3D Indoor Scene Understanding Using Mobile RGB-D Data cs.CV · 2021-11-17 · accept · none · ref 8 · internal anchor
ARKitScenes is the largest real-world indoor RGB-D dataset captured with mobile LiDAR, including high-resolution depth maps and 3D furniture bounding box annotations for advancing object detection and depth upsampling.
WarpHammer: Densifying Scene Warps with 3D Object Priors for Extreme View Synthesis cs.CV · 2026-06-30 · unverdicted · none · ref 9 · internal anchor
WarpHammer densifies scene warps with 3D object priors from generative models and fuses pose-unknown auxiliary views via multi-view geometry to enable stable extreme novel view synthesis.
3D-CoS: A New 3D Reconstruction Paradigm Based on VLM Code Synthesis cs.CV · 2026-06-09 · unverdicted · none · ref 7 · internal anchor
3D-CoS represents 3D objects as Blender code generated by VLMs, with workflows for planning, RAG, and agents, showing better edit fidelity than point-cloud baselines.
Rethinking 3D Shape Generation: Diffusion over Superquadrics cs.CV · 2026-06-08 · unverdicted · none · ref 16 · internal anchor
Diffusion for 3D shapes is moved from dense geometry to compact superquadric parameter sets, cutting state size to roughly 7 KB per shape and enabling faster generation plus new editing capabilities.
Why Far Looks Up: Probing Spatial Representation in Vision-Language Models cs.CV · 2026-05-28 · conditional · none · ref 8 · internal anchor
VLMs exhibit consistent vertical-distance entanglement in embeddings from perspective bias in natural images, producing accuracy gaps that a new synthetic benchmark SpatialTunnel exposes as model-intrinsic.
Category-Level 3D Correspondence in Camera Space via Morphable Object Priors cs.CV · 2026-05-27 · unverdicted · none · ref 4 · internal anchor
Morpheus learns morphable category-level shape priors to produce implicit 3D correspondences in camera space without explicit supervision and releases the HouseCorr3D benchmark with amodal and symmetry annotations.
Metric--Phase Fields: Decoupling Distance and Sign for Thin-Structure Reconstruction from Unoriented Point Clouds cs.CV · 2026-05-25 · unverdicted · none · ref 6 · internal anchor
Metric-Phase Fields decouple unsigned metric proximity from a smooth phase field with learnable sharpness to enable faithful reconstruction of thin and open structures from point clouds.
ArtSplat: Feed-Forward Articulated 3D Gaussian Splatting from Sparse Multi-State Uncalibrated Views cs.CV · 2026-05-23 · unverdicted · none · ref 2 · internal anchor
ArtSplat is the first feed-forward framework for articulated 3D Gaussian Splatting that reconstructs geometry and joints from sparse multi-state uncalibrated views in one pass.
MAPS: A Synthetic Dataset for Probing Vision Models in a Controlled 3D Scene Space cs.CV · 2026-05-19 · unverdicted · none · ref 13 · internal anchor
MAPS provides 2618 validated 3D meshes and a controllable rendering pipeline to attribute vision model recognition failures to specific scene parameters, finding camera distance and elevation as the dominant failure factors across 20 tested models.
OffsetAxis: UDF Mesh Reconstruction via Offset-Volume Medial Axis Extraction cs.GR · 2026-05-14 · unverdicted · none · ref 1 · internal anchor
OffsetAxis reconstructs meshes from unsigned distance fields by extracting the medial axis of the alpha-offset volume using ray casting and variational medial ball optimization.
Min Generalized Sliced Gromov Wasserstein: A Scalable Path to Gromov Wasserstein cs.LG · 2026-05-13 · unverdicted · none · ref 2 · internal anchor
min-GSGW learns coupled nonlinear slicers to produce a rigid-motion-invariant, scalable approximation to the Gromov-Wasserstein distance and its transport plans.
Img2CADSeq: Image-to-CAD Generation via Sequence-Based Diffusion cs.CV · 2026-05-13 · unverdicted · none · ref 27 · internal anchor
Img2CADSeq generates standard CAD sequences from images via a multi-stage pipeline with three-level hierarchical codebook encoding, importance-guided compression, and contrastive point-cloud conditioning of a VQ-Diffusion model, outperforming prior methods on new CAD-220K and PrintCAD datasets.
Count Anything at Any Granularity cs.CV · 2026-05-11 · unverdicted · none · ref 14 · internal anchor
Multi-grained counting is introduced with five granularity levels, supported by the new KubriCount dataset generated via 3D synthesis and editing, and HieraCount model that combines text and visual exemplars for improved accuracy.
The Wittgensteinian Representation Hypothesis: Is Language the Attractor of Multimodal Convergence? cs.AI · 2026-05-10 · unverdicted · none · ref 45 · internal anchor
Language representations serve as the asymptotic attractor for convergence in independently trained multimodal neural networks due to feature density asymmetry.
MeshFIM: Local Low-Poly Mesh Editing via Fill-in-the-Middle Autoregressive Generation cs.GR · 2026-05-09 · unverdicted · none · ref 30 · internal anchor
MeshFIM enables local low-poly mesh editing by autoregressively filling target regions conditioned on context, using boundary markers, positional embeddings, and a gated geometry encoder to enforce attachment, topology, and region limits.
Rollback-Free Stable Brick Structures Generation cs.LG · 2026-05-07 · unverdicted · none · ref 2 · internal anchor
Reinforcement learning internalizes physical stability rules for brick structures, enabling the first rollback-free generation with orders-of-magnitude faster inference.
Two Steps Are All You Need: Efficient 3D Point Cloud Anomaly Detection with Consistency Models cs.CV · 2026-05-06 · unverdicted · none · ref 4 · internal anchor
Consistency learning reformulates 3D point cloud anomaly detection to predict clean geometry directly in one or two steps, yielding up to 80 times faster inference while matching state-of-the-art accuracy.
ADS: Random Sampling of Occupancy Functions using Adaptive Delaunay Scaffolding cs.GR · 2026-05-05 · unverdicted · none · ref 7 · internal anchor
ADS adaptively refines a Delaunay scaffold to produce unbiased random samples on occupancy function surfaces together with a connecting mesh, using far fewer evaluations than existing approaches.
Generative Modeling with Orbit-Space Particle Flow Matching cs.GR · 2026-05-04 · unverdicted · none · ref 18 · internal anchor
OGPP is a particle flow-matching method using orbit-space canonicalization and geometric paths that achieves lower error and fewer steps than prior approaches on 3D benchmarks.
AirZoo: A Unified Large-Scale Dataset for Grounding Aerial Geometric 3D Vision cs.CV · 2026-04-29 · unverdicted · none · ref 6 · 2 links · internal anchor
AirZoo is a new dataset covering 378 regions across 22 countries with pixel-level metric depth and 6-DoF poses, shown via benchmarks to improve SoTA models on aerial image retrieval, cross-view matching, and multi-view 3D reconstruction.
Topo-ADV: Generating Topology-Driven Imperceptible Adversarial Point Clouds cs.CV · 2026-04-10 · unverdicted · none · ref 6 · internal anchor
Topo-ADV uses differentiable persistent homology to create topology-altering perturbations that achieve up to 100% attack success on point cloud classifiers like PointNet while remaining geometrically imperceptible.
Training-free Spatially Grounded Geometric Shape Encoding (Technical Report) cs.CV · 2026-04-08 · unverdicted · none · ref 4 · 2 links · internal anchor
XShapeEnc encodes arbitrary 2D spatially grounded shapes into compact invertible representations by decomposing them into unit-disk geometry and harmonic pose fields then applying Zernike bases with frequency propagation.
3D-Fixer: Coarse-to-Fine In-place Completion for 3D Scenes from a Single Image cs.CV · 2026-04-06 · unverdicted · none · ref 4 · internal anchor
3D-Fixer performs in-place 3D asset completion from single-view partial point clouds via coarse-to-fine generation with ORFA conditioning, plus a new ARSG-110K dataset, to achieve higher geometric accuracy than MIDI and Gen3DSR while keeping diffusion efficiency.
Deformation-based In-Context Learning for Point Cloud Understanding cs.CV · 2026-04-03 · unverdicted · none · ref 5 · internal anchor
DeformPIC deforms query point clouds under prompt guidance for in-context learning, outperforming prior methods with lower Chamfer Distance on reconstruction, denoising, and registration tasks.
Align then Adapt: Rethinking Parameter-Efficient Transfer Learning in 4D Perception cs.CV · 2026-02-26 · unverdicted · none · ref 7 · internal anchor
PointATA is a parameter-efficient transfer learning method that aligns 3D-4D modality gaps via optimal transport before adapting a frozen 3D model with video-specific modules to achieve strong 4D perception results.
CLIPoint3D: Language-Grounded Few-Shot Unsupervised 3D Point Cloud Domain Adaptation cs.CV · 2026-02-23 · unverdicted · none · ref 6 · internal anchor
CLIPoint3D is the first CLIP-based framework for few-shot unsupervised 3D point cloud domain adaptation that reports 3-16% accuracy gains on PointDA-10 and GraspNetPC-10.
Physically Guided Visual Mass Estimation from a Single RGB Image cs.CV · 2026-01-28 · unverdicted · none · ref 4 · internal anchor
A method estimates mass from single RGB images by fusing depth-based volume cues with vision-language model density semantics via adaptive gating and separate regression heads trained on mass labels only.
Streaming Sliced Optimal Transport cs.LG · 2025-05-11 · unverdicted · none · ref 11 · internal anchor
A low-memory streaming estimator for sliced Wasserstein distance using quantile approximations on random projections with theoretical error guarantees.
Hard-Label Black-Box Attacks on 3D Point Clouds cs.CV · 2024-11-30 · unverdicted · none · ref 78 · internal anchor
A spectrum-aware decision boundary algorithm enables effective hard-label black-box adversarial attacks on 3D point cloud models by fusing spectral information across classes and performing curvature-aware iterative optimization.
LRM: Large Reconstruction Model for Single Image to 3D cs.CV · 2023-11-08 · conditional · none · ref 4 · internal anchor
LRM is a large transformer that predicts a NeRF directly from a single image after training on a million-object multi-view dataset.
Objaverse-XL: A Universe of 10M+ 3D Objects cs.CV · 2023-07-11 · accept · none · ref 9 · internal anchor
Objaverse-XL supplies over 10 million diverse 3D objects that, when used to render 100 million views, improve zero-shot novel-view synthesis in models such as Zero123.
Fast Graph Representation Learning with PyTorch Geometric cs.LG · 2019-03-06 · accept · none · ref 9 · internal anchor
PyTorch Geometric is a PyTorch library that delivers fast graph neural network training through sparse GPU kernels and variable-size mini-batching.
SuperFlex: Deformable Superquadrics for Point Cloud Decomposition cs.CV · 2026-07-01 · unverdicted · none · ref 6 · internal anchor
SuperFlex extends superquadrics with deformations and a new loss for higher-accuracy point cloud decomposition and trains a model robust to partial real-world data.
GenSP: Consistent Spherical Parameterization via Learning Shape Generative Models cs.CV · 2026-07-01 · unverdicted · none · ref 10 · internal anchor
GenSP learns a continuous neural deformation model from sphere coordinates and latent codes to produce consistent spherical parameterizations for genus-0 shapes.
From Grasps to Dexterity: Large-Scale Grasp Pretraining for Dexterous Manipulation cs.RO · 2026-06-29 · unverdicted · none · ref 48 · internal anchor
Grasp pretraining on 355k trajectories improves full-task success on six articulated tool-use tasks by 33.3 pp over DP3 in real-world experiments.
Emergence of a Shared Canonical Object Frame from In-the-Wild Videos cs.CV · 2026-06-29 · unverdicted · none · ref 8 · internal anchor
A coarse canonical mesh bottleneck plus multi-view consistency lets a shared object frame emerge from self-supervised training on in-the-wild videos without canonical labels or category conditioning.
GeoEdit: Geometry-Aware Object Editing via Dual-Branch Denoising cs.CV · 2026-06-29 · unverdicted · none · ref 7 · internal anchor
GeoEdit introduces a Lift-Manipulate-Render-Denoise pipeline with dual-branch denoising and variance-homogeneous injection for 3D-consistent object editing in single photos.
Anomaly Factory 3D: A Modular Framework for Diverse Pseudo-Anomaly Synthesis in Unsupervised 3D Anomaly Detection cs.CV · 2026-06-28 · unverdicted · none · ref 4 · internal anchor
AF3AD is a modular synthesis framework using center-conditioned parametric deformations in local PCA frames to create diverse pseudo-anomalies, improving unsupervised 3D anomaly detection on AnomalyShapeNet and Real3D-AD.
DeepJEB++: Foundation Model-Driven Large-Scale 3D Engineering Dataset via 2D Latent Space Augmentation cs.LG · 2026-06-11 · unverdicted · none · ref 10 · internal anchor
DeepJEB++ expands a small seed set of jet engine brackets into 15,360 labeled 3D designs via 2D latent diffusion augmentation, VLM filtering, generative 3D lifting, and automated finite-element labeling.
MAMVI: 3D Test-Time Adaptation via Masked Multi-View Point Clouds cs.CV · 2026-06-11 · unverdicted · none · ref 4 · internal anchor
MAMVI performs unified single-step TTA on masked multi-view point clouds with hybrid masking and confidence-adaptive learning rates, reporting SOTA on ShapeNet-C and ScanObjectNN-C plus 4.9-8.9x speedup.
Tac-DINO: Learning Vision-Tactile Features with Patch Alignment cs.CV · 2026-06-10 · unverdicted · none · ref 26 · internal anchor
Tac-DINO constructs a large tactile dataset and Vis-Tac Holographic Matching Benchmark, then proposes Vision-Tactile Patch Alignment (VTPA) methods that outperform non-aligned baselines on local-to-global feature matching.
SynthICL: Scalable In-context Imitation Learning with Synthetic Data cs.RO · 2026-06-06 · unverdicted · none · ref 29 · internal anchor
SynthICL trains flow-matching transformer policies for in-context imitation learning entirely from synthetic RGB data and reports 79% average success on 16 unseen real manipulation tasks with one test-time demonstration.
PAR3D: A Unified 3D-MLLM with Part-Aware Representation for Scene Understanding cs.CV · 2026-06-04 · unverdicted · none · ref 6 · internal anchor
PAR3D is a part-aware 3D-MLLM framework with ScenePart dataset, Part-Aware 3D Representation Learning, and Hierarchical Segmentation Query Generation to improve part-level 3D scene understanding.
EqGINO: Equivariant Geometry-Informed Fourier Neural Operators for 3D PDEs cs.LG · 2026-06-02 · unverdicted · none · ref 14 · internal anchor
EqGINO adds a spectral isotropy prior to FNOs to guarantee discrete equivariance and enable generalization to continuous SE(3) transformations on 3D PDEs with limited training data.
From Extrinsic to Intrinsic: Geodesic-Guided Representation Learning for 3D Geometric Data cs.CV · 2026-06-01 · unverdicted · none · ref 2 · internal anchor
PRISM is a pre-training method that learns isometric latent embeddings by explicitly recovering surface geodesic distances with a topology-enforcing loss and a two-stage training schedule.
HOLA: Holistic Multi-Modal Alignment for Open-Set 3D Recognition cs.CV · 2026-05-31 · unverdicted · none · ref 2 · internal anchor
HOLA introduces multi-view multi-text alignment and a decoupled contrastive loss for state-of-the-art open-vocabulary 3D recognition on long-tail benchmarks.
FoundObj: Self-supervised Foundation Models as Rewards for Label-free 3D Object Segmentation cs.CV · 2026-05-26 · unverdicted · none · ref 9 · internal anchor
FoundObj uses foundation-model priors as RL rewards to discover multi-class 3D objects from point clouds without scene-level labels.
DinoComplete: 3D Shape Completion with Distilled Semantic Priors and State Space Models cs.CV · 2026-05-26 · unverdicted · none · ref 7 · internal anchor
DinoComplete augments geometric 3D shape completion with voxel-aligned DINO semantic priors and multi-scale voxel Mamba modeling to improve results on unseen categories with lower compute.
BrickAnything: Geometry-Conditioned Buildable Brick Generation with Structure-Aware Tokenization cs.AI · 2026-05-25 · unverdicted · none · ref 50 · internal anchor
BrickAnything generates buildable brick structures from 3D point clouds via geometry-conditioned autoregressive prediction with structure-aware tree tokenization and post-training for stability.

ShapeNet: An Information-Rich 3D Model Repository

hub tools

citation-role summary

citation-polarity summary

claims ledger

authors

co-cited works

fields

years

verdicts

roles

polarities

representative citing papers

citing papers explorer