A new benchmark dataset of 4,000 lecture video frames, 1,000 manually annotated with four visual object categories and 3,000 auto-labeled via a fine-tuned YOLOv11 model.
Visual content detection in educational videos with transfer learn- ing and dataset enrichment,
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
citation-role summary
background 1
citation-polarity summary
fields
cs.CV 1years
2025 1verdicts
CONDITIONAL 1roles
background 1polarities
unclear 1representative citing papers
citing papers explorer
-
Lecture Video Visual Objects (LVVO) Dataset: A Benchmark for Visual Object Detection in Educational Videos
A new benchmark dataset of 4,000 lecture video frames, 1,000 manually annotated with four visual object categories and 3,000 auto-labeled via a fine-tuned YOLOv11 model.