REVIEW 2 cited by
The VIA Annotation Software for Images, Audio and Video
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
In this paper, we introduce a simple and standalone manual annotation tool for images, audio and video: the VGG Image Annotator (VIA). This is a light weight, standalone and offline software package that does not require any installation or setup and runs solely in a web browser. The VIA software allows human annotators to define and describe spatial regions in images or video frames, and temporal segments in audio or video. These manual annotations can be exported to plain text data formats such as JSON and CSV and therefore are amenable to further processing by other software tools. VIA also supports collaborative annotation of a large dataset by a group of human annotators. The BSD open source license of this software allows it to be used in any academic project or commercial application.
Forward citations
Cited by 2 Pith papers
-
Toward quantitative fractography using convolutional neural networks
A U-net semantic segmentation model trained on MgAl2O4 fracture surfaces quantifies intergranular and transgranular modes in SEM images, with reported mean IoU of 91.1% on the training material and 94% on untrained Al2O3.
-
Advancing Utility Pole and Sign Detection Through Deep Learning
A DETR-based detector with a segmentation head detects utility poles and signs and estimates lean angle with about 1 degree mean error on a new UK Street View dataset, though YOLOv8 matches or beats it on several metrics.
Discussion (0). Continue with ORCID to comment.