Pith. sign in

REVIEW 1 cited by

audino: A Modern Annotation Tool for Audio and Speech

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2006.05236 v2 pith:MVYPUDDQ submitted 2020-06-09 cs.SD cs.CLeess.AS

classification cs.SDcs.CLeess.AS
keywords toolannotationspeechallowsaudioadminannotationsaudino
verification ladder T0 review T1 audit T2 compute T3 formal

Signed reviews

No signed human review yet.

0 comments
read the original abstract

In this paper, we introduce a collaborative and modern annotation tool for audio and speech: audino. The tool allows annotators to define and describe temporal segmentation in audios. These segments can be labelled and transcribed easily using a dynamically generated form. An admin can centrally control user roles and project assignment through the admin dashboard. The dashboard also enables describing labels and their values. The annotations can easily be exported in JSON format for further analysis. The tool allows audio data and their corresponding annotations to be uploaded and assigned to a user through a key-based API. The flexibility available in the annotation tool enables annotation for Speech Scoring, Voice Activity Detection (VAD), Speaker Diarisation, Speaker Identification, Speech Recognition, Emotion Recognition tasks and more. The MIT open source license allows it to be used for academic and commercial projects.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. CAFE A Novel Code switching Dataset for Algerian Dialect French and English

    cs.SD 2024-11 conditional novelty 6.0 of 10

    CAFE is a new spontaneous speech corpus for Algerian dialect, French, and English code-switching, with 2.6 hours manually annotated and a Whisper benchmark reaching MER 0.310.

Pith tools