Pith. sign in

REVIEW 2 cited by

The VIA Annotation Software for Images, Audio and Video

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1904.10699 v3 pith:V4N3RXUT submitted 2019-04-24 cs.CV

classification cs.CV
keywords softwarevideoannotationaudioimagesallowsannotatorshuman
verification ladder T0 review T1 audit T2 compute T3 formal

Signed reviews

No signed human review yet.

0 comments
read the original abstract

In this paper, we introduce a simple and standalone manual annotation tool for images, audio and video: the VGG Image Annotator (VIA). This is a light weight, standalone and offline software package that does not require any installation or setup and runs solely in a web browser. The VIA software allows human annotators to define and describe spatial regions in images or video frames, and temporal segments in audio or video. These manual annotations can be exported to plain text data formats such as JSON and CSV and therefore are amenable to further processing by other software tools. VIA also supports collaborative annotation of a large dataset by a group of human annotators. The BSD open source license of this software allows it to be used in any academic project or commercial application.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Toward quantitative fractography using convolutional neural networks

    eess.IV 2019-08 conditional novelty 6.0 of 10

    A U-net semantic segmentation model trained on MgAl2O4 fracture surfaces quantifies intergranular and transgranular modes in SEM images, with reported mean IoU of 91.1% on the training material and 94% on untrained Al2O3.

  2. Advancing Utility Pole and Sign Detection Through Deep Learning

    cs.CV 2026-08 conditional novelty 5.0 of 10

    A DETR-based detector with a segmentation head detects utility poles and signs and estimates lean angle with about 1 degree mean error on a new UK Street View dataset, though YOLOv8 matches or beats it on several metrics.

Pith tools