Pith. sign in

REVIEW 1 cited by

MeetEval: A Toolkit for Computation of Word Error Rates for Meeting Transcription Systems

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2307.11394 v3 pith:MDZLU2TS submitted 2023-07-21 cs.CL eess.AS

classification cs.CLeess.AS
keywords timecomputationleadsmatchingtranscriptionword-levelannotationsconstraint
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

MeetEval is an open-source toolkit to evaluate all kinds of meeting transcription systems. It provides a unified interface for the computation of commonly used Word Error Rates (WERs), specifically cpWER, ORC-WER and MIMO-WER along other WER definitions. We extend the cpWER computation by a temporal constraint to ensure that only words are identified as correct when the temporal alignment is plausible. This leads to a better quality of the matching of the hypothesis string to the reference string that more closely resembles the actual transcription quality, and a system is penalized if it provides poor time annotations. Since word-level timing information is often not available, we present a way to approximate exact word-level timings from segment-level timings (e.g., a sentence) and show that the approximation leads to a similar WER as a matching with exact word-level annotations. At the same time, the time constraint leads to a speedup of the matching algorithm, which outweighs the additional overhead caused by processing the time stamps.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. DNCASR: End-to-End Training for Speaker-Attributed ASR

    eess.AS 2025-06 conditional novelty 5.0 of 10

    DNCASR links speaker clustering and ASR decoders with cross-attention, achieving a 9.0% relative cpWER reduction on AMI-MDM Eval over a parallel (unlinked) system.

Pith tools