Pith. sign in

Paper Citation Record · LEDGER

Decoder-only Architecture for Streaming End-to-end Speech Recognition

As of 20 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2406.16107.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2406.16107 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 5 of 5 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T11:30:22.001244Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T11:02:26.631012Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation a884769e-bc77-4026-b8ef-54b81ba2eca0 · inbound

Transducer-Llama: Integrating LLMs into Streamable Transducer-based Speech Recognition cites this paper.

Transducer-Llama: Integrating LLMs into Streamable Transducer-based Speech Recognition Decoder-only Architecture for Streaming End-to-end Speech Recognition

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-11T10:37:52.407658Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T10:37:52.407658Z digest=sha256:6f0241c5fdcc3f2be9b5ed3a65572d2abf8f51506d448b01689fc80a645e66dd

Observation 968d45f1-92ec-4447-9021-36170cbc3b77 · inbound

Prepending or Cross-Attention for Speech-to-Text? An Empirical Comparison cites this paper.

Prepending or Cross-Attention for Speech-to-Text? An Empirical Comparison Decoder-only Architecture for Streaming End-to-end Speech Recognition

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-10T22:18:01.804040Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T22:18:01.804040Z digest=sha256:1a5b174549abc7ee13a1a55c4876c3a8d0858c6b5ae0e0810d9cd625d80f22d2

Observation 821afaf1-fa38-4854-bf93-c572627d5498 · inbound

SimulS2S-LLM: Unlocking Simultaneous Inference of Speech LLMs for Speech-to-Speech Translation cites this paper.

SimulS2S-LLM: Unlocking Simultaneous Inference of Speech LLMs for Speech-to-Speech Translation Decoder-only Architecture for Streaming End-to-end Speech Recognition

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-16T11:30:22.001244Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:30:22.001244Z digest=sha256:47dc659e2a1fe9816f75b22c8f1e2814ada76ac0c6bab6283573e5c694edd9a5

Observation a85ce056-a4a0-4df2-adae-bceb707f7f9b · inbound

MFLA: Monotonic Finite Look-ahead Attention for Streaming Speech Recognition cites this paper.

MFLA: Monotonic Finite Look-ahead Attention for Streaming Speech Recognition Decoder-only Architecture for Streaming End-to-end Speech Recognition

Reference 36

Resolution
verified exact
local_arxiv, observed 2026-08-07T11:02:26.674736Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-07T11:02:26.521767Z digest=sha256:20a6057d1e6b4904eed09904137a5e1210e61855988b6f8671e2bdafefe15904

Observation 3c586111-c9b3-401a-b74a-38bc08e448e1 · inbound

SimulS2ST-Omni: Data-Efficient Streaming Speech-to-Speech Translation via Explicit Trajectory Supervision cites this paper.

SimulS2ST-Omni: Data-Efficient Streaming Speech-to-Speech Translation via Explicit Trajectory Supervision Decoder-only Architecture for Streaming End-to-end Speech Recognition

Reference 164

Resolution
unresolved
no resolver link, observed 2026-08-01T11:43:06.492839Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T11:43:06.492839Z digest=sha256:3a3e5ec9ce1a71c81402b8c82042074d6f5d5a2cb2991790219ccc2d4e1d09d2