REVIEW 2 cited by
R-CNNs for Pose Estimation and Action Detection
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
We present convolutional neural networks for the tasks of keypoint (pose) prediction and action classification of people in unconstrained images. Our approach involves training an R-CNN detector with loss functions depending on the task being tackled. We evaluate our method on the challenging PASCAL VOC dataset and compare it to previous leading approaches. Our method gives state-of-the-art results for keypoint and action prediction. Additionally, we introduce a new dataset for action detection, the task of simultaneously localizing people and classifying their actions, and present results using our approach.
Forward citations
Cited by 2 Pith papers
-
HOIverse: A Synthetic Scene Graph Dataset With Human Object Interactions
HOIverse is a synthetic indoor scene graph dataset with dense human-object interaction and parametric relation annotations, benchmarked with scene graph generation models.
-
MultiTaskVIF: Segmentation-oriented visible and infrared image fusion via multi-task learning
A training framework with a dual-branch decoder lets visible-infrared fusion networks learn semantic segmentation as an auxiliary task, improving fused-image segmentation without a separate cascade model.
Discussion (0). Continue with ORCID to comment.