Pith. sign in

REVIEW 4 cited by

GS-Pose: Generalizable Segmentation-based 6D Object Pose Estimation with 3D Gaussian Splatting

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2403.10683 v2 pith:3CC3O24V submitted 2024-03-15 cs.CV

classification cs.CV
keywords gs-poseobjectposedatabaseestimatinggaussiannovelobjects
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

This paper introduces GS-Pose, a unified framework for localizing and estimating the 6D pose of novel objects. GS-Pose begins with a set of posed RGB images of a previously unseen object and builds three distinct representations stored in a database. At inference, GS-Pose operates sequentially by locating the object in the input image, estimating its initial 6D pose using a retrieval approach, and refining the pose with a render-and-compare method. The key insight is the application of the appropriate object representation at each stage of the process. In particular, for the refinement step, we leverage 3D Gaussian splatting, a novel differentiable rendering technique that offers high rendering speed and relatively low optimization time. Off-the-shelf toolchains and commodity hardware, such as mobile phones, can be used to capture new objects to be added to the database. Extensive evaluations on the LINEMOD and OnePose-LowTexture datasets demonstrate excellent performance, establishing the new state-of-the-art. Project page: https://dingdingcai.github.io/gs-pose.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 4 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. T-Rex: Task-Adaptive Spatial Representation Extraction for Robotic Manipulation with Vision-Language Models

    cs.RO 2025-06 conditional novelty 6.0 of 10

    A zero-training framework that adaptively selects spatial representation extractors per object and per task stage improves real-world robot manipulation success and efficiency over fixed-representation baselines.

  2. GCE-Pose: Global Context Enhancement for Category-level Object Pose Estimation

    cs.CV 2025-02 conditional novelty 6.0 of 10

    Adding a category-level semantic shape prior, reconstructed from partial RGB-D input, improves 6D pose and size estimation for unseen objects on HouseCat6D and NOCS-REAL275.

  3. GSGTrack: Gaussian Splatting-Guided Object Pose Tracking from RGB Videos

    cs.CV 2024-12 conditional novelty 5.0 of 10

    GSGTrack jointly optimizes Gaussian Splatting geometry and object pose to track unknown objects in RGB video, reporting large accuracy gains over SLAM baselines.

  4. Advancing Extended Reality with 3D Gaussian Splatting: Innovations and Prospects

    cs.CV 2024-12 conditional novelty 4.0 of 10

    3D Gaussian Splatting research relevant to Extended Reality is organized into a five-part taxonomy with suggested future directions.

Pith tools