Pith. sign in

REVIEW 2 cited by

Beyond Confusion: A Fine-grained Dialectical Examination of Human Activity Recognition Benchmark Datasets

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2412.09037 v1 pith:AHAHGUMC submitted 2024-12-12 cs.LG

classification cs.LG
keywords datasetsbenchmarkdataactivityfine-grainedhumanmetricsproblems
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

The research of machine learning (ML) algorithms for human activity recognition (HAR) has made significant progress with publicly available datasets. However, most research prioritizes statistical metrics over examining negative sample details. While recent models like transformers have been applied to HAR datasets with limited success from the benchmark metrics, their counterparts have effectively solved problems on similar levels with near 100% accuracy. This raises questions about the limitations of current approaches. This paper aims to address these open questions by conducting a fine-grained inspection of six popular HAR benchmark datasets. We identified for some parts of the data, none of the six chosen state-of-the-art ML methods can correctly classify, denoted as the intersect of false classifications (IFC). Analysis of the IFC reveals several underlying problems, including ambiguous annotations, irregularities during recording execution, and misaligned transition periods. We contribute to the field by quantifying and characterizing annotated data ambiguities, providing a trinary categorization mask for dataset patching, and stressing potential improvements for future data collections.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. TinierHAR: Towards Ultra-Lightweight Deep Learning Models for Efficient Human Activity Recognition on Edge Devices

    cs.CV 2025-07 conditional novelty 5.0 of 10

    TinierHAR is an ultra-lightweight HAR model that matches TinyHAR's F1 score with 2.7x fewer parameters and 6.4x fewer MACs across 14 datasets.

  2. FedFitTech: A Baseline in Federated Learning for Fitness Tracking

    cs.LG 2025-06 conditional novelty 4.0 of 10

    An open-source Flower-based federated learning baseline for fitness tracking, plus a case study showing client-side early stopping cuts communication 13% with a 1% F1 drop.

Pith tools