Pith. sign in

Paper Citation Record · LEDGER

WeNet: Production oriented Streaming and Non-streaming End-to-End Speech Recognition Toolkit

As of 21 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2102.01547.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2102.01547 v5

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-12T04:59:13.266712Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-23T17:45:46.256504Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation b619631c-44b6-4e38-a4f8-7a3fad3813f3 · inbound

DGSNA: Dynamic Generative Scene-based Noise Addition method cites this paper.

DGSNA: Dynamic Generative Scene-based Noise Addition method WeNet: Production oriented Streaming and Non-streaming End-to-End Speech Recognition Toolkit

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-05-23T17:45:46.259659Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-23T17:43:47.086524Z digest=sha256:ae7d286452aa23d49a51b0c86d2c963c5c8bb1cb489c65ea8164e25d115bd72e

Observation 6260f095-42ec-4c31-9990-41765f8c7ea3 · inbound

Complexity boosted adaptive training for better low resource ASR performance cites this paper.

Complexity boosted adaptive training for better low resource ASR performance WeNet: Production oriented Streaming and Non-streaming End-to-End Speech Recognition Toolkit

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-12T04:59:13.266712Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:59:13.266712Z digest=sha256:60ffb5bbbab5242b34b027f3d6318519d45c1fb29cf5eb81d9823b0f5dadd60c

Observation 369bca63-dcd9-4716-865b-18eba5fa604d · inbound

YingSound: Video-Guided Sound Effects Generation with Multi-modal Chain-of-Thought Controls cites this paper.

YingSound: Video-Guided Sound Effects Generation with Multi-modal Chain-of-Thought Controls WeNet: Production oriented Streaming and Non-streaming End-to-End Speech Recognition Toolkit

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-11T17:17:49.362817Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:17:49.362817Z digest=sha256:2dcae3ad611e9fc29e4691fe678709ff8fd2c4ca794b802c3b9f4b03e36d9d59

Observation 84d1e35b-2251-46cb-8d05-d5ecb225bc9e · inbound

TouchASP: Elastic Automatic Speech Perception that Everyone Can Touch cites this paper.

TouchASP: Elastic Automatic Speech Perception that Everyone Can Touch WeNet: Production oriented Streaming and Non-streaming End-to-End Speech Recognition Toolkit

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-11T11:18:34.568621Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T11:18:34.568621Z digest=sha256:d886e70d6cccee249d52e423035fb3cd419c8ea8d38b3a9a2e6eddfe2f452174

Observation d4cffacc-9597-448f-8c66-fdefb9c92978 · inbound

When End-to-End is Overkill: Rethinking Cascaded Speech-to-Text Translation cites this paper.

When End-to-End is Overkill: Rethinking Cascaded Speech-to-Text Translation WeNet: Production oriented Streaming and Non-streaming End-to-End Speech Recognition Toolkit

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-09T19:16:37.847173Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T19:16:37.847173Z digest=sha256:cb73a34dd1a48e76a8c150b2ba7371406e95781baa96ec5608c37ef3244788fc

Observation 65eb1298-afbb-42b7-a8fc-b97ef1131cea · inbound

A Self-Training Approach for Whisper to Enhance Long Dysarthric Speech Recognition cites this paper.

A Self-Training Approach for Whisper to Enhance Long Dysarthric Speech Recognition WeNet: Production oriented Streaming and Non-streaming End-to-End Speech Recognition Toolkit

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T22:02:27.799231Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:02:27.799231Z digest=sha256:c22c65ea760415bbbf0e7930d51b03fd54266c5572f96349b78d540c5c05f882