Pith. sign in

REVIEW 1 cited by

Unsupervised Multi-Person 3D Human Pose Estimation From 2D Poses Alone

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2309.14865 v3 pith:E6YTHEGZ submitted 2023-09-26 cs.CV

classification cs.CV
keywords posesposeunsupervisedd-3destimationhumanmulti-personalone
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Current unsupervised 2D-3D human pose estimation (HPE) methods do not work in multi-person scenarios due to perspective ambiguity in monocular images. Therefore, we present one of the first studies investigating the feasibility of unsupervised multi-person 2D-3D HPE from just 2D poses alone, focusing on reconstructing human interactions. To address the issue of perspective ambiguity, we expand upon prior work by predicting the cameras' elevation angle relative to the subjects' pelvis. This allows us to rotate the predicted poses to be level with the ground plane, while obtaining an estimate for the vertical offset in 3D between individuals. Our method involves independently lifting each subject's 2D pose to 3D, before combining them in a shared 3D coordinate system. The poses are then rotated and offset by the predicted elevation angle before being scaled. This by itself enables us to retrieve an accurate 3D reconstruction of their poses. We present our results on the CHI3D dataset, introducing its use for unsupervised 2D-3D pose estimation with three new quantitative metrics, and establishing a benchmark for future research.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Reconstructing People, Places, and Cameras

    cs.CV 2024-12 conditional novelty 6.0 of 10

    HSfM jointly optimizes human meshes, dense scene pointmaps, and camera poses in a metric world frame, reducing world-frame human joint error from 3.5m to 1.0m on EgoHumans.

Pith tools