Pith. sign in

REVIEW 2 cited by

NurViD: A Large Expert-Level Video Database for Nursing Procedure Activity Understanding

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2310.13347 v1 pith:2LLMFWZW submitted 2023-10-20 cs.CV cs.AI

classification cs.CVcs.AI
keywords nursingactivitydatasetsprocedureactionnurvidproceduresunderstanding
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

The application of deep learning to nursing procedure activity understanding has the potential to greatly enhance the quality and safety of nurse-patient interactions. By utilizing the technique, we can facilitate training and education, improve quality control, and enable operational compliance monitoring. However, the development of automatic recognition systems in this field is currently hindered by the scarcity of appropriately labeled datasets. The existing video datasets pose several limitations: 1) these datasets are small-scale in size to support comprehensive investigations of nursing activity; 2) they primarily focus on single procedures, lacking expert-level annotations for various nursing procedures and action steps; and 3) they lack temporally localized annotations, which prevents the effective localization of targeted actions within longer video sequences. To mitigate these limitations, we propose NurViD, a large video dataset with expert-level annotation for nursing procedure activity understanding. NurViD consists of over 1.5k videos totaling 144 hours, making it approximately four times longer than the existing largest nursing activity datasets. Notably, it encompasses 51 distinct nursing procedures and 177 action steps, providing a much more comprehensive coverage compared to existing datasets that primarily focus on limited procedures. To evaluate the efficacy of current deep learning methods on nursing activity understanding, we establish three benchmarks on NurViD: procedure recognition on untrimmed videos, procedure and action recognition on trimmed videos, and action detection. Our benchmark and code will be available at \url{https://github.com/minghu0830/NurViD-benchmark}.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. M$^3$-Med: A Benchmark for Multi-lingual, Multi-modal, and Multi-hop Reasoning in Medical Instructional Video Understanding

    cs.CV 2025-07 conditional novelty 7.0 of 10

    M3-Med is a new bilingual benchmark that tests video-language AI models on multi-hop temporal reasoning in medical instructional videos, and it shows a large gap between current models and human performance.

  2. Meta-SurDiff: Classification Diffusion Model Optimized by Meta Learning is Reliable for Online Surgical Phase Recognition

    cs.CV 2025-06 reject novelty 4.0 of 10

    Meta-SurDiff combines a classification diffusion model with meta-learned sample weighting and reports state-of-the-art results on five surgical video datasets, but the derivation of the reverse process contains a nume...

Pith tools