Pith. sign in

REVIEW 1 cited by

Joint Learning of Social Groups, Individuals Action and Sub-group Activities in Videos

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2007.02632 v2 pith:DJDHSEDJ submitted 2020-07-06 cs.CV

classification cs.CV
keywords socialactivitygrouptaskindividualssceneactionscall
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

The state-of-the art solutions for human activity understanding from a video stream formulate the task as a spatio-temporal problem which requires joint localization of all individuals in the scene and classification of their actions or group activity over time. Who is interacting with whom, e.g. not everyone in a queue is interacting with each other, is often not predicted. There are scenarios where people are best to be split into sub-groups, which we call social groups, and each social group may be engaged in a different social activity. In this paper, we solve the problem of simultaneously grouping people by their social interactions, predicting their individual actions and the social activity of each social group, which we call the social task. Our main contributions are: i) we propose an end-to-end trainable framework for the social task; ii) our proposed method also sets the state-of-the-art results on two widely adopted benchmarks for the traditional group activity recognition task (assuming individuals of the scene form a single group and predicting a single group activity label for the scene); iii) we introduce new annotations on an existing group activity dataset, re-purposing it for the social task.

Discussion (0). Sign in to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

  1. A Pairwise Human-Human Interaction Detection and Recognition Framework for Mobile Service Robots

    cs.RO 2026-02 conditional novelty 4.0 of 10

    On JRDB, geometry plus optical flow alone reaches 84.3% accuracy for classifying walking/standing/sitting pairs, outperforming the same pipeline augmented with a frozen visual backbone, and transfers zero-shot to a la...

Pith tools