Pith. sign in

Paper Citation Record · LEDGER

WenetSpeech: A 10000+ Hours Multi-domain Mandarin Corpus for Speech Recognition

As of 19 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 7 inbound Pith citation observations for arXiv:2110.03370.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2110.03370 v5

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 7 of 7 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 7 of 7 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:30:37.358858Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

32
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 20ba5e4b-8a6c-45a4-b627-74fb13feaee1 · inbound

Pseudo Labels-based Neural Speech Enhancement for the AVSR Task in the MISP-Meeting Challenge cites this paper.

Pseudo Labels-based Neural Speech Enhancement for the AVSR Task in the MISP-Meeting Challenge WenetSpeech: A 10000+ Hours Multi-domain Mandarin Corpus for Speech Recognition

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-07T12:30:37.358858Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:30:37.358858Z digest=sha256:9f24d1c384f5147fb2154cb786e62587236bf37b33b0c965a694f7e09db60c21

Observation 19cd9c47-a240-4442-9e8b-0599fd5ce481 · inbound

Ming-Omni: A Unified Multimodal Model for Perception and Generation cites this paper.

Ming-Omni: A Unified Multimodal Model for Perception and Generation WenetSpeech: A 10000+ Hours Multi-domain Mandarin Corpus for Speech Recognition

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T04:58:10.654286Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:58:10.654286Z digest=sha256:76fcda01d5b61349be31438135b2ae122d365019ffb30a36d4b73851c9365018

Observation 8b54d6d9-9aaf-4136-a244-4645ab435966 · inbound

Stream-Omni: Simultaneous Multimodal Interactions with Large Language-Vision-Speech Model cites this paper.

Stream-Omni: Simultaneous Multimodal Interactions with Large Language-Vision-Speech Model WenetSpeech: A 10000+ Hours Multi-domain Mandarin Corpus for Speech Recognition

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T00:33:19.462794Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:33:19.462794Z digest=sha256:bf1aea60f659bf913d55a378af258681ddf3689855f1ff32e2715887cc6cc3ee

Observation 17c8c44b-5f15-4b71-a916-c3bd436ea1a7 · inbound

Breaking the Transcription Bottleneck: Fine-tuning ASR Models for Extremely Low-Resource Fieldwork Languages cites this paper.

Breaking the Transcription Bottleneck: Fine-tuning ASR Models for Extremely Low-Resource Fieldwork Languages WenetSpeech: A 10000+ Hours Multi-domain Mandarin Corpus for Speech Recognition

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-06T23:35:32.604685Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:35:32.604685Z digest=sha256:bb6f8bd80ae363f546542544b7480fc7c4e705167cb964d224187cff299b6852

Observation 0fdc7dad-ae15-41b2-b68d-4684f58df13e · inbound

Step-Audio 2 Technical Report cites this paper.

Step-Audio 2 Technical Report WenetSpeech: A 10000+ Hours Multi-domain Mandarin Corpus for Speech Recognition

Reference 80

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T05:59:51.107662Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-16T05:59:50.900436Z digest=sha256:602b9a8fa95ef07e9433ac8a91c0ea7b39d331630fd4b545af9627bfad64b931

Observation b71c0222-688f-4a27-9fe8-e363ad42f8be · inbound

Non-Intrusive Automatic Speech Recognition Refinement: A Survey cites this paper.

Non-Intrusive Automatic Speech Recognition Refinement: A Survey WenetSpeech: A 10000+ Hours Multi-domain Mandarin Corpus for Speech Recognition

Reference 136

Resolution
verified exact
arxiv_id, observed 2026-05-21T23:35:46.418333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-21T23:34:52.249143Z digest=sha256:01d39ef48388c9b80d1de00d42547673b4f493019b60035f409585775dd7a519

Observation aed6896a-4747-4327-8c52-3fd2774666c3 · inbound

Data-Efficient On-Policy Distillation for Automatic Speech Recognition cites this paper.

Data-Efficient On-Policy Distillation for Automatic Speech Recognition WenetSpeech: A 10000+ Hours Multi-domain Mandarin Corpus for Speech Recognition

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-06-29T12:03:23.187584Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-29T12:02:35.332633Z digest=sha256:ea90d3c8a11adf357d73b9c3e68148728330652fcb5d0f5c1009dbd5d7183c25