Pith. sign in

REVIEW 1 cited by

KeypointNet: A Large-scale 3D Keypoint Dataset Aggregated from Numerous Human Annotations

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2002.12687 v6 pith:VJ5MOEHN submitted 2020-02-28 cs.CV

classification cs.CV
keywords keypointsannotationsdatasetkeypointkeypointnetdatasetshumanlarge-scale
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Detecting 3D objects keypoints is of great interest to the areas of both graphics and computer vision. There have been several 2D and 3D keypoint datasets aiming to address this problem in a data-driven way. These datasets, however, either lack scalability or bring ambiguity to the definition of keypoints. Therefore, we present KeypointNet: the first large-scale and diverse 3D keypoint dataset that contains 103,450 keypoints and 8,234 3D models from 16 object categories, by leveraging numerous human annotations. To handle the inconsistency between annotations from different people, we propose a novel method to aggregate these keypoints automatically, through minimization of a fidelity loss. Finally, ten state-of-the-art methods are benchmarked on our proposed dataset. Our code and data are available on https://github.com/qq456cvb/KeypointNet.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models

    cs.CV 2024-12 conditional novelty 6.0 of 10

    ZeroKey detects 3D keypoints on unseen object categories by prompting the Molmo vision-language model on multiple rendered views and aggregating the back-projected points, with no 3D annotations required.

Pith tools