Pith. sign in

Paper Citation Record · LEDGER

Unified Streaming and Non-streaming Two-pass End-to-end Model for Speech Recognition

As of 23 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 8 inbound Pith citation observations for arXiv:2012.05481.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2012.05481 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 8 of 8 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 8 of 8 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T19:58:55.037683Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-05T13:21:06.219289Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation e798313b-24dc-4d09-bfe5-56d522e7283b · inbound

Adapting Whisper for Streaming Speech Recognition via Two-Pass Decoding cites this paper.

Adapting Whisper for Streaming Speech Recognition via Two-Pass Decoding Unified Streaming and Non-streaming Two-pass End-to-end Model for Speech Recognition

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T01:07:24.172626Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T01:07:24.172626Z digest=sha256:20ac107aaf6351efe9b5be3f17bde4bc186bc1e661091427edd4161c00f56b08

Observation d7c5c7a2-23aa-4bf3-92e7-0d4095e5b152 · inbound

Unifying Streaming and Non-streaming Zipformer-based ASR cites this paper.

Unifying Streaming and Non-streaming Zipformer-based ASR Unified Streaming and Non-streaming Two-pass End-to-end Model for Speech Recognition

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-15T19:58:55.037683Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T19:58:55.037683Z digest=sha256:f8f42bf435b70d04605458466cd78416df4a7429bd20a6c23ac96471ab92d46f

Observation c15b48dd-76b3-4ddd-be3e-19c2916165fc · inbound

A Study of Data Selection Strategies for Pre-training Self-Supervised Speech Models cites this paper.

A Study of Data Selection Strategies for Pre-training Self-Supervised Speech Models Unified Streaming and Non-streaming Two-pass End-to-end Model for Speech Recognition

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-16T10:17:43.805945Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-16T10:17:15.726726Z digest=sha256:b681c9706c284089a641eeb20fdd9eec9ae227660b0b28e2dd363494791b5a40

Observation ac20fda9-1cab-4441-8df6-468ebed687c7 · inbound

NIM4-ASR: Towards Efficient, Robust, and Customizable Real-Time LLM-Based ASR cites this paper.

NIM4-ASR: Towards Efficient, Robust, and Customizable Real-Time LLM-Based ASR Unified Streaming and Non-streaming Two-pass End-to-end Model for Speech Recognition

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-11T12:21:05.877936Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-10T03:48:14.211240Z digest=sha256:3b04bfc88706c1a938afb7f4272843bbb9cf6398ccc50c9f34cfb33cab1e6293

Observation d5abe31f-6cd1-4942-b71c-8d865372d548 · inbound

NIM4-ASR: Towards Efficient, Robust, and Customizable Real-Time LLM-Based ASR cites this paper.

NIM4-ASR: Towards Efficient, Robust, and Customizable Real-Time LLM-Based ASR Unified Streaming and Non-streaming Two-pass End-to-end Model for Speech Recognition

Reference 31

Resolution
verified exact
local_arxiv, observed 2026-07-05T13:21:06.222299Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-07-05T13:15:50.794969Z digest=sha256:bb9fcd9089a0b30151c9a6bf92c874b668bf7783a79d1944e3d7c7076b868c66

Observation 18b90027-1df9-4311-aadd-992dba7bb252 · inbound

TRADE: Transducer-Augmented Decoder for Speech LLM cites this paper.

TRADE: Transducer-Augmented Decoder for Speech LLM Unified Streaming and Non-streaming Two-pass End-to-end Model for Speech Recognition

Reference 49

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T22:47:25.701123Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-27T18:40:19.688550Z digest=sha256:92476652826444d5b3ea8ec384dc68dffaddb2ef167f4fad93967d98c95341ea

Observation 2cb1b489-1675-4243-9293-65dc665d5e53 · inbound

Online Predictive Coding for Dual-Mode Self-Supervised Speech Model cites this paper.

Online Predictive Coding for Dual-Mode Self-Supervised Speech Model Unified Streaming and Non-streaming Two-pass End-to-end Model for Speech Recognition

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-07-04T07:39:38.781197Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-26T13:15:04.922107Z digest=sha256:aebb84c5d135ad88eac6a964f95e16c96939faf88c65c729c7fbfd7b9630435f

Observation 77d40865-6b3f-4ab8-9f6e-45ff83bf3fcc · inbound

JoyAI-Talker: Full-Duplex Speech Interactive Large Model Built for Empathetic Voice Agents cites this paper.

JoyAI-Talker: Full-Duplex Speech Interactive Large Model Built for Empathetic Voice Agents Unified Streaming and Non-streaming Two-pass End-to-end Model for Speech Recognition

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T00:30:12.153813Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:30:12.153813Z digest=sha256:b13593753a63c42e7fb29517b8c954da4c49d9c818f7b519f565db26a4537f0a