Pith. sign in

REVIEW 15 cited by

Rethinking Network Design and Local Geometry in Point Cloud: A Simple Residual MLP Framework

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2202.07123 v2 pith:VRNGVMNP submitted 2022-02-15 cs.CV cs.AI

Rethinking Network Design and Local Geometry in Point Cloud: A Simple Residual MLP Framework

classification cs.CV cs.AI
keywords pointmlpcloudlocalpointanalysissophisticatedextractorsfaster
verification ladder T0 review T1 audit T2 compute T3 formal T4 reserved
0 comments
read the original abstract

Point cloud analysis is challenging due to irregularity and unordered data structure. To capture the 3D geometries, prior works mainly rely on exploring sophisticated local geometric extractors using convolution, graph, or attention mechanisms. These methods, however, incur unfavorable latency during inference, and the performance saturates over the past few years. In this paper, we present a novel perspective on this task. We notice that detailed local geometrical information probably is not the key to point cloud analysis -- we introduce a pure residual MLP network, called PointMLP, which integrates no sophisticated local geometrical extractors but still performs very competitively. Equipped with a proposed lightweight geometric affine module, PointMLP delivers the new state-of-the-art on multiple datasets. On the real-world ScanObjectNN dataset, our method even surpasses the prior best method by 3.3% accuracy. We emphasize that PointMLP achieves this strong performance without any sophisticated operations, hence leading to a superior inference speed. Compared to most recent CurveNet, PointMLP trains 2x faster, tests 7x faster, and is more accurate on ModelNet40 benchmark. We hope our PointMLP may help the community towards a better understanding of point cloud analysis. The code is available at https://github.com/ma-xu/pointMLP-pytorch.

discussion (0)

Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.

Forward citations

Cited by 15 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

  1. On the Generation and Mitigation of Harmful Geometry in Image-to-3D Models

    cs.CR 2026-05 conditional novelty 8.0

    Image-to-3D models successfully generate harmful geometries in most cases with under 0.3% caught by commercial filters; existing safeguards are weak but a stacked defense cuts harmful outputs to under 1% at 11% false-...

  2. RelFlexformer: Efficient Attention 3D-Transformers for Integrable Relative Positional Encodings

    cs.LG 2026-05 unverdicted novelty 7.0

    RelFlexformers enable flexible integrable 3D RPE in attention via NU-FFT, generalizing prior methods to heterogeneous token positions with O(L log L) complexity.

  3. DAP: Doppler-aware Point Network for Heterogeneous mmWave Action Recognition

    cs.CV 2026-05 unverdicted novelty 7.0

    Introduces the first heterogeneous multi-source mmWave point cloud HAR dataset and DAP-Net architecture with Doppler reparameterization and text alignment for cross-source robustness.

  4. Detecting Dental Landmarks from Intraoral 3D Scans: the 3DTeethLand challenge

    cs.CV 2025-12 accept novelty 7.0

    A public benchmark dataset and competition results for 3D dental landmark detection from intraoral scans, with the top team reaching 0.91 rank score using a stratified transformer and DBSCAN.

  5. LLaMA-Adapter: Efficient Fine-tuning of Language Models with Zero-init Attention

    cs.CV 2023-03 conditional novelty 7.0

    LLaMA-Adapter turns frozen LLaMA 7B into a capable instruction follower using only 1.2M new parameters and zero-init attention, matching Alpaca while extending to image-conditioned reasoning on ScienceQA and COCO.

  6. DAP: Doppler-aware Point Network for Heterogeneous mmWave Action Recognition

    cs.CV 2026-05 unverdicted novelty 6.0

    Introduces the first heterogeneous multi-source mmWave point cloud HAR dataset and DAP-Net, which uses Doppler patterns for source-invariant action recognition and outperforms prior methods.

  7. Beyond Defenses: Manifold-Aligned Regularization for Intrinsic 3D Point Cloud Robustness

    cs.CV 2026-05 unverdicted novelty 6.0

    MAPR aligns latent and intrinsic geometries in 3D point cloud models via regularization on curvature and diffusion features plus consistency loss, yielding +20% average robustness gains on ModelNet40 without adversari...

  8. Beyond Defenses: Manifold-Aligned Regularization for Intrinsic 3D Point Cloud Robustness

    cs.CV 2026-05 unverdicted novelty 6.0

    MAPR improves adversarial robustness in 3D point cloud networks by aligning latent predictions with intrinsic manifold geometry via curvature/diffusion features and a consistency loss.

  9. Delaunay Canopy: Building Wireframe Reconstruction from Airborne LiDAR Point Clouds via Delaunay Graph

    cs.CV 2026-04 unverdicted novelty 6.0

    Delaunay Canopy uses Delaunay graphs as a geometric prior with region-wise curvature scoring to reconstruct accurate building wireframes from sparse and noisy airborne LiDAR point clouds.

  10. CLAMP: Contrastive Learning for 3D Multi-View Action-Conditioned Robotic Manipulation Pretraining

    cs.RO 2026-01 unverdicted novelty 6.0

    CLAMP pretrains 3D multi-view encoders with contrastive learning on point clouds and actions, then initializes diffusion policies for more sample-efficient fine-tuning on robotic tasks.

  11. $\text{VG}^2$GT: Voxel-Gaussian Splatting Visual Geometry Grounded Transformer

    cs.CV 2026-06 unverdicted novelty 5.0

    VG²GT regresses Gaussian primitive parameters from multi-scale voxel features of a frozen VFM and uses stochastic solid volume rendering for depth supervision to produce geometrically accurate reconstructions that out...

  12. Full-field prediction for engineering-scale three-dimensional aircraft with multigrid-hierarchical learning

    physics.flu-dyn 2026-05 unverdicted novelty 5.0

    MHLF combines multigrid geometry representation with hierarchical learning to predict full flow fields for engineering-scale 3D aircraft, accelerating CFD convergence 3-8x across subsonic to supersonic regimes without...

  13. A Camera-Cooperative ISAC Framework for Multimodal Non-Cooperative UAVs Sensing

    cs.AI 2026-05 unverdicted novelty 5.0

    The CC-ISAC framework aligns camera visuals with radio echoes via cross-attention and fuses multimodal data to reduce beam steering overhead by 71% and tracking overhead by 1.69-11.15% on the DeepSense 6G dataset whil...

  14. Heterogeneous and Adept Snapshot Distillation for 3D Semantic Segmentation

    cs.CV 2026-06 unverdicted novelty 4.0

    HAS-KD combines information-oriented heterogeneous distillation from multi-modal models with adept snapshot distillation from training checkpoints to reach SOTA 3D semantic segmentation on ScanNetV2 and S3DIS without ...

  15. Scene Reconstruction as Mapping Priors for 3D Detection

    cs.CV 2026-05 unverdicted novelty 4.0

    Automatically constructed mapping priors from sensor aggregation are integrated via the MPA3D framework to achieve state-of-the-art 3D detection results on the Waymo Open Dataset.