Pith. sign in

Paper Citation Record · LEDGER

Unified Streaming and Non-streaming Two-pass End-to-end Model for Speech Recognition

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2012.05481.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2012.05481 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T00:30:12.153813Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-05T13:21:06.219289Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation c15b48dd-76b3-4ddd-be3e-19c2916165fc · inbound

A Study of Data Selection Strategies for Pre-training Self-Supervised Speech Models cites this paper.

A Study of Data Selection Strategies for Pre-training Self-Supervised Speech Models Unified Streaming and Non-streaming Two-pass End-to-end Model for Speech Recognition

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-16T10:17:43.805945Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-16T10:17:15.726726Z digest=sha256:a3e9da42205de741e893294541d124b8008f6b0e5e2d613b72a7521ba6611053

Observation ac20fda9-1cab-4441-8df6-468ebed687c7 · inbound

NIM4-ASR: Towards Efficient, Robust, and Customizable Real-Time LLM-Based ASR cites this paper.

NIM4-ASR: Towards Efficient, Robust, and Customizable Real-Time LLM-Based ASR Unified Streaming and Non-streaming Two-pass End-to-end Model for Speech Recognition

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-11T12:21:05.877936Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T03:48:14.211240Z digest=sha256:74e59b0ebcf6d1da4884846449679b9d7d52ff86cbcd00626bab7f4e261d86d3

Observation d5abe31f-6cd1-4942-b71c-8d865372d548 · inbound

NIM4-ASR: Towards Efficient, Robust, and Customizable Real-Time LLM-Based ASR cites this paper.

NIM4-ASR: Towards Efficient, Robust, and Customizable Real-Time LLM-Based ASR Unified Streaming and Non-streaming Two-pass End-to-end Model for Speech Recognition

Reference 31

Resolution
verified exact
local_arxiv, observed 2026-07-05T13:21:06.222299Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-05T13:15:50.794969Z digest=sha256:29d6f37087f4ab4686582e2fc7d7aa9416cae4c61e521818b5cb4c0d42f733d9

Observation 18b90027-1df9-4311-aadd-992dba7bb252 · inbound

TRADE: Transducer-Augmented Decoder for Speech LLM cites this paper.

TRADE: Transducer-Augmented Decoder for Speech LLM Unified Streaming and Non-streaming Two-pass End-to-end Model for Speech Recognition

Reference 49

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T22:47:25.701123Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-06-27T18:40:19.688550Z digest=sha256:199a8cd91201df36c00a3102a9760fc66d08b876b0bc4bccc83421e31d79206f

Observation 2cb1b489-1675-4243-9293-65dc665d5e53 · inbound

Online Predictive Coding for Dual-Mode Self-Supervised Speech Model cites this paper.

Online Predictive Coding for Dual-Mode Self-Supervised Speech Model Unified Streaming and Non-streaming Two-pass End-to-end Model for Speech Recognition

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-07-04T07:39:38.781197Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-26T13:15:04.922107Z digest=sha256:c26a835b6cdcc2fcce8e640309c7f87f408423b5e9bb97c06ec293e6646e0794

Observation 77d40865-6b3f-4ab8-9f6e-45ff83bf3fcc · inbound

JoyAI-Talker: Full-Duplex Speech Interactive Large Model Built for Empathetic Voice Agents cites this paper.

JoyAI-Talker: Full-Duplex Speech Interactive Large Model Built for Empathetic Voice Agents Unified Streaming and Non-streaming Two-pass End-to-end Model for Speech Recognition

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T00:30:12.153813Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:30:12.153813Z digest=sha256:304d3324fd1e2da0732cd3b3564236074521f2bf7b59617fda814247b9417a2c