Pith. sign in

Paper Citation Record · LEDGER

ParaASR: Multi-Token Prediction for Fast and Long-Context LLM-Based Speech Recognition

As of 10 August 2026, this Paper Citation Record lists 41 of 41 outbound references and 0 inbound Pith citation observations for arXiv:2607.29279.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.29279 v1

Coverage vector

measured 41 of 41 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-03T10:13:57.954776Z

measured 41 of 41 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

41 of 41 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved41
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 786023c1-079a-4bf8-82c1-d41f5b28afd8 · outbound

This paper cites arXiv preprint arXiv:2509.12508 (2025).

ParaASR: Multi-Token Prediction for Fast and Long-Context LLM-Based Speech Recognition arXiv preprint arXiv:2509.12508 (2025)

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-03T10:13:53.900211Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:13:53.900211Z digest=sha256:1b8e77cfd5d3a48de4593432956b50c86eef618566a10af131c3ce15c2ea223c

Observation 98d73dab-bdee-4925-b285-723f707ab2c4 · outbound

This paper cites In: Proceedings of the Twelfth Language Resources and Evaluation Conference.

ParaASR: Multi-Token Prediction for Fast and Long-Context LLM-Based Speech Recognition In: Proceedings of the Twelfth Language Resources and Evaluation Conference

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-03T10:13:53.962467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:13:53.962467Z digest=sha256:01f690a962a8d25d4c99df2c3bf59502bfec1e9097ca73d08082799b38c2cd82

Observation 59897fb3-9159-4e10-b6ff-cd66f47b02d4 · outbound

This paper cites an unresolved cited work.

ParaASR: Multi-Token Prediction for Fast and Long-Context LLM-Based Speech Recognition Unresolved cited work

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-03T10:13:54.044695Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:13:54.044695Z digest=sha256:89dd78ae41e172795543ddb1643f444c044f04507b49bf1fde0f5fbf2db793b5

Observation f105ed3b-93ca-4826-b8d7-e8988352e9c9 · outbound

This paper cites an unresolved cited work.

ParaASR: Multi-Token Prediction for Fast and Long-Context LLM-Based Speech Recognition Unresolved cited work

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-03T10:13:54.189749Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:13:54.189749Z digest=sha256:6d54f79d3099d6c59349b2afee0fc8a04add7a196b13c4e6ca4daf220ca341e1

Observation 7d9304c1-3dee-415d-92b3-b42649a97cf5 · outbound

This paper cites Seed-ASR: Understanding Diverse Speech and Contexts with LLM-based Speech Recognition.

ParaASR: Multi-Token Prediction for Fast and Long-Context LLM-Based Speech Recognition Seed-ASR: Understanding Diverse Speech and Contexts with LLM-based Speech Recognition

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-03T10:13:54.253318Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:13:54.253318Z digest=sha256:388afeb1bf8f15a2c280976e52a95a905ce3bd00162423b7646bf5c0e2965125

Observation 60aa5429-51d5-4dd8-a379-f1250ea58b3f · outbound

This paper cites WhisperX: Time-Accurate Speech Transcription of Long-Form Audio.

ParaASR: Multi-Token Prediction for Fast and Long-Context LLM-Based Speech Recognition WhisperX: Time-Accurate Speech Transcription of Long-Form Audio

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-03T10:13:54.327892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:13:54.327892Z digest=sha256:d0eda01fa585f153f4617a2aca37c58412b28d27e58c0112b5475682efa0b542

Observation 6b018eb8-dd4f-450c-8ef0-256b36962a39 · outbound

This paper cites In: 2017 20th Conference of the Oriental Chapter of the International Coordinating Committee on Speech Databases and Speech I/O Systems and Assessment.

ParaASR: Multi-Token Prediction for Fast and Long-Context LLM-Based Speech Recognition In: 2017 20th Conference of the Oriental Chapter of the International Coordinating Committee on Speech Databases and Speech I/O Systems and Assessment

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-03T10:13:54.405095Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:13:54.405095Z digest=sha256:5c659e6da22c0a3351fd71ec0e52b9a6444eb2fe14f07f1506ef5417ed0c4b28

Observation a873db63-7423-4755-a66c-75739bc9399d · outbound

This paper cites Medusa: Simple LLM Inference Acceleration Framework with Multiple Decoding Heads.

ParaASR: Multi-Token Prediction for Fast and Long-Context LLM-Based Speech Recognition Medusa: Simple LLM Inference Acceleration Framework with Multiple Decoding Heads

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-03T10:13:54.567049Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:13:54.567049Z digest=sha256:99a84a4c867b47faa05981477c2494ff452180fe8c5faaf01c0c74c4ef93d4ba

Observation c4d5d563-3613-49ad-aad0-996cf83555a0 · outbound

This paper cites Listen, Attend and Spell.

ParaASR: Multi-Token Prediction for Fast and Long-Context LLM-Based Speech Recognition Listen, Attend and Spell

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-03T10:13:54.677464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:13:54.677464Z digest=sha256:940c142e819d3ace550a7282c2f8b1e540aba439f46eb9df2dbfee4ad8bfc0f6

Observation 8c35b0d7-c2cf-44c1-93b4-09c538a9986a · outbound

This paper cites Accelerating Large Language Model Decoding with Speculative Sampling.

ParaASR: Multi-Token Prediction for Fast and Long-Context LLM-Based Speech Recognition Accelerating Large Language Model Decoding with Speculative Sampling

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-03T10:13:54.841372Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:13:54.841372Z digest=sha256:1136e76e4503d08fcf60314a73ec72f31b704e80cfe706ffa085a87049b72a18

Observation 0a00891e-d3f5-4e73-b27d-8578017b231d · outbound

This paper cites FLEURS: Few-shot Learning Evaluation of Universal Representations of Speech.

ParaASR: Multi-Token Prediction for Fast and Long-Context LLM-Based Speech Recognition FLEURS: Few-shot Learning Evaluation of Universal Representations of Speech

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-03T10:13:54.959703Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:13:54.959703Z digest=sha256:b07f37e73f1635cd3093501cd81edf175a449aca64a399bce7949871994d92f9

Observation b914f1c1-138d-43ea-9bc0-1c88ea8f8d0b · outbound

This paper cites arXiv preprint arXiv:2511.00850 (2025).

ParaASR: Multi-Token Prediction for Fast and Long-Context LLM-Based Speech Recognition arXiv preprint arXiv:2511.00850 (2025)

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-03T10:13:55.143015Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:13:55.143015Z digest=sha256:168683f873e045c778203c58fa44eee64f24bc333a4b739cfd6eec8828a00df5

Observation 84590f14-97e8-4627-96c8-0976a52559ce · outbound

This paper cites AISHELL-2: Transforming Mandarin ASR Research Into Industrial Scale.

ParaASR: Multi-Token Prediction for Fast and Long-Context LLM-Based Speech Recognition AISHELL-2: Transforming Mandarin ASR Research Into Industrial Scale

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-03T10:13:55.256063Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:13:55.256063Z digest=sha256:869c6b4e9dc12c24df0668e33ff01d57db5895894860320b04d105d862ef3313

Observation 8a79d292-1f8a-4733-bdfd-72157ec3dde8 · outbound

This paper cites In: 1997 IEEE Workshop on Automatic Speech Recognition and Understanding Proceedings.

ParaASR: Multi-Token Prediction for Fast and Long-Context LLM-Based Speech Recognition In: 1997 IEEE Workshop on Automatic Speech Recognition and Understanding Proceedings

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-03T10:13:55.344559Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:13:55.344559Z digest=sha256:abc532ec74d57873503d7c6fcb940858340d18687900c1ed20e666e28cc15e24

Observation a699f2ed-2560-42c0-be04-d8e298569471 · outbound

This paper cites Better & Faster Large Language Models via Multi-token Prediction.

ParaASR: Multi-Token Prediction for Fast and Long-Context LLM-Based Speech Recognition Better & Faster Large Language Models via Multi-token Prediction

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-03T10:13:55.501150Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:13:55.501150Z digest=sha256:09b8e205d99181a9c54fbcc7ed99af0a6f87477610ea3b2dbd9df9941fd08bbf

Observation 701c2852-13e5-40de-b600-afa692b17e6a · outbound

This paper cites In: Supervised sequence labelling with recurrent neural networks, pp.

ParaASR: Multi-Token Prediction for Fast and Long-Context LLM-Based Speech Recognition In: Supervised sequence labelling with recurrent neural networks, pp

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-03T10:13:55.684968Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:13:55.684968Z digest=sha256:18af306aa25c432c2b701a17556373fe7ee6960ce55d97ccc90ec9b30a77fbec

Observation e98b20d4-e12e-451d-af7a-6cfb656d79e5 · outbound

This paper cites Sequence Transduction with Recurrent Neural Networks.

ParaASR: Multi-Token Prediction for Fast and Long-Context LLM-Based Speech Recognition Sequence Transduction with Recurrent Neural Networks

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-03T10:13:55.806367Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:13:55.806367Z digest=sha256:86953cd19f7124999df52f8a651216dadc0c4811a896943ea024e67b7aac64cd

Observation 2b3adb49-ff95-4ec7-a6e3-4ee3ce5b227e · outbound

This paper cites arXiv preprint arXiv:2602.10604 (2026).

ParaASR: Multi-Token Prediction for Fast and Long-Context LLM-Based Speech Recognition arXiv preprint arXiv:2602.10604 (2026)

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-03T10:13:55.896842Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:13:55.896842Z digest=sha256:e2b3ad6b9e7d30cd9a3d17910c7b2255294a660ce56ee26a1ac664327238a0e0

Observation 7fc80d42-bd14-4677-96a1-0555c731bc76 · outbound

This paper cites In: ICASSP 2024-2024 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP).

ParaASR: Multi-Token Prediction for Fast and Long-Context LLM-Based Speech Recognition In: ICASSP 2024-2024 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP)

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-03T10:13:55.951331Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:13:55.951331Z digest=sha256:fbc01ae1c5f75a2c5f66af6885008926c5607b0b2584115b7416b4b54a843623

Observation 16ed9af0-bddf-4445-bedd-e80a5bb5dabd · outbound

This paper cites PMLR (2023).

ParaASR: Multi-Token Prediction for Fast and Long-Context LLM-Based Speech Recognition PMLR (2023)

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-03T10:13:56.028066Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:13:56.028066Z digest=sha256:26bfbeb841807f3e468c39feb5b3f66823c97a128ce90518f1635e301d40cef5

Observation a05dba3a-3f4d-45fe-ad13-fa4c033e29f2 · outbound

This paper cites StepAudio 2.5 Technical Report.

ParaASR: Multi-Token Prediction for Fast and Long-Context LLM-Based Speech Recognition StepAudio 2.5 Technical Report

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-03T10:13:56.110991Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:13:56.110991Z digest=sha256:682b0cb92719fb5720372acea7a819f95609509892dfcde34c9b523904716350

Observation fec2c2ba-554f-41c4-8d0b-b1181f731ba6 · outbound

This paper cites Boosting Omni-Modal Language Models: Staged Post-Training with Visually Debiased Evaluation.

ParaASR: Multi-Token Prediction for Fast and Long-Context LLM-Based Speech Recognition Boosting Omni-Modal Language Models: Staged Post-Training with Visually Debiased Evaluation

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-03T10:13:56.212641Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:13:56.212641Z digest=sha256:f779a334b8441f426563405c79554aed12d3df419013211adb5732d3a7bd7678

Observation e6ea5e0b-160c-4d36-be00-ea350751c1dd · outbound

This paper cites arXiv preprint arXiv:2509.24310 (2025).

ParaASR: Multi-Token Prediction for Fast and Long-Context LLM-Based Speech Recognition arXiv preprint arXiv:2509.24310 (2025)

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-03T10:13:56.287599Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:13:56.287599Z digest=sha256:2631f70d0e04ba18ef7b783a3ae039f2e018b4edab060d5dc3d451e17883e636

Observation e9eafcfa-2543-4c00-8956-5556515a8dc7 · outbound

This paper cites Multimedia Tools and Applications80(6), 9411–9457 (2021).

ParaASR: Multi-Token Prediction for Fast and Long-Context LLM-Based Speech Recognition Multimedia Tools and Applications80(6), 9411–9457 (2021)

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-03T10:13:56.371783Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:13:56.371783Z digest=sha256:55342e93aa670b4322e9e020621c71222d1c3c93ecbc40d2a4122ce8cf03aa2d

Observation 4294a865-6f85-425d-8e02-7c88ce2fa08b · outbound

This paper cites In: 2015 IEEE International Conference on Acoustics, Speech and Signal Processing.

ParaASR: Multi-Token Prediction for Fast and Long-Context LLM-Based Speech Recognition In: 2015 IEEE International Conference on Acoustics, Speech and Signal Processing

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-03T10:13:56.439971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:13:56.439971Z digest=sha256:29a56a625b29756d012c793fb9eea91587362876ae903f79bba1a9f771088bc4

Observation 670a9bb4-64f1-483b-8ea2-8f39f00f5179 · outbound

This paper cites In: Interspeech 2019.

ParaASR: Multi-Token Prediction for Fast and Long-Context LLM-Based Speech Recognition In: Interspeech 2019

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-03T10:13:56.559132Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:13:56.559132Z digest=sha256:ec82dc789142c4ad6054faa34a12518053783c8185e0de194fba62a7cabdc838

Observation 85426615-17cb-45b6-863e-7499fc858af1 · outbound

This paper cites arXiv preprint arXiv:2601.18184 (2026).

ParaASR: Multi-Token Prediction for Fast and Long-Context LLM-Based Speech Recognition arXiv preprint arXiv:2601.18184 (2026)

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-03T10:13:56.661673Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:13:56.661673Z digest=sha256:f301c2e392b67047a2e3acc88d9308b94d307c0a0799c4c9eb3874c403e19407

Observation 34b7b98c-a437-43b3-9e85-a862db828b91 · outbound

This paper cites In: International conference on machine learning.

ParaASR: Multi-Token Prediction for Fast and Long-Context LLM-Based Speech Recognition In: International conference on machine learning

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-03T10:13:56.747587Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:13:56.747587Z digest=sha256:a13d3dacac148180c549bc749ce4413571c941a56bcd59129d80bfd9f24955d8

Observation 49398783-16d2-4f38-8552-36d00e3cdcfd · outbound

This paper cites Qwen3-ASR Technical Report.

ParaASR: Multi-Token Prediction for Fast and Long-Context LLM-Based Speech Recognition Qwen3-ASR Technical Report

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-03T10:13:56.837223Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:13:56.837223Z digest=sha256:0f5296b59f21a45020c26b4bdd5986aca9ae1cc977e182d2968b1bf57c25af31

Observation 96b7d5c6-40bd-4372-b052-9a4518c7b4ba · outbound

This paper cites arXiv preprint arXiv:2511.15848 (2025).

ParaASR: Multi-Token Prediction for Fast and Long-Context LLM-Based Speech Recognition arXiv preprint arXiv:2511.15848 (2025)

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-03T10:13:56.948109Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:13:56.948109Z digest=sha256:12d1ae3f2bb9450b0a3beca412492e74db4ca830088d04bbc50007ca6e81741e

Observation fb3eaa16-e7e1-43db-a0ca-c7e1b0191f4c · outbound

This paper cites Step-Audio 2 Technical Report.

ParaASR: Multi-Token Prediction for Fast and Long-Context LLM-Based Speech Recognition Step-Audio 2 Technical Report

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-03T10:13:57.074975Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:13:57.074975Z digest=sha256:22b95d80e4167157bd29107625147473ffe820da4498cfefecd191494356d3e2

Observation f65dfae5-4eca-46bc-98c1-0ede1ccce1c9 · outbound

This paper cites Mind-Paced Speaking: A Dual-Brain Approach to Real-Time Reasoning in Spoken Language Models.

ParaASR: Multi-Token Prediction for Fast and Long-Context LLM-Based Speech Recognition Mind-Paced Speaking: A Dual-Brain Approach to Real-Time Reasoning in Spoken Language Models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-03T10:13:57.190594Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:13:57.190594Z digest=sha256:a0fefc12ae9632c3108a88470028ee953e5c42a4a5755c06d4220288aa0f8054

Observation e66aebf7-654d-4632-9ae6-b30d514d25ac · outbound

This paper cites Qwen3-Omni Technical Report.

ParaASR: Multi-Token Prediction for Fast and Long-Context LLM-Based Speech Recognition Qwen3-Omni Technical Report

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-03T10:13:57.308232Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:13:57.308232Z digest=sha256:db361b164e32ff07664080b2b56dca82172825b85b5049998ec3ff7aea7c131d

Observation 41173a3e-492f-4ecf-8dd3-bf59f3429dba · outbound

This paper cites In: ICASSP 2023-2023 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP).

ParaASR: Multi-Token Prediction for Fast and Long-Context LLM-Based Speech Recognition In: ICASSP 2023-2023 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP)

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-03T10:13:57.421541Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:13:57.421541Z digest=sha256:3c845a5fda91228d4231a4b856bf22aa57a08ef6e2779d7db63dabb2d8ed31a4

Observation c37169ab-571b-462d-9532-9e62f73f0700 · outbound

This paper cites In: Proc.

ParaASR: Multi-Token Prediction for Fast and Long-Context LLM-Based Speech Recognition In: Proc

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-03T10:13:57.509325Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:13:57.509325Z digest=sha256:0d7376c372cceaaa612051292928173bb491fff1e5538cd0974caeababc0be82

Observation 68e690a5-c5a6-4dc7-9bad-08c34e9a21f5 · outbound

This paper cites In: ICASSP 2022 - 2022 IEEE International Conference on Acoustics, Speech and Signal Processing.

ParaASR: Multi-Token Prediction for Fast and Long-Context LLM-Based Speech Recognition In: ICASSP 2022 - 2022 IEEE International Conference on Acoustics, Speech and Signal Processing

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-03T10:13:57.569911Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:13:57.569911Z digest=sha256:84539bed0e9c7e834d280cc478ca34225207d9f76daec81d2848b670a6000eda

Observation e5a96d3b-4301-4639-8848-e9d03eedb21f · outbound

This paper cites DuplexSLA: A Full-Duplex Spoken Language Model with Synchronized Speech, Language, and Action.

ParaASR: Multi-Token Prediction for Fast and Long-Context LLM-Based Speech Recognition DuplexSLA: A Full-Duplex Spoken Language Model with Synchronized Speech, Language, and Action

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-03T10:13:57.633052Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:13:57.633052Z digest=sha256:6f29d826cb9233db06f070fbe7c3b7710c89dfbcee60031cf58433215e511c34

Observation d6e5ce31-033b-4fde-86cd-ac72e05ad0a5 · outbound

This paper cites Impact of Frame Rates on Speech Tokenizer: A Case Study on Mandarin and English.

ParaASR: Multi-Token Prediction for Fast and Long-Context LLM-Based Speech Recognition Impact of Frame Rates on Speech Tokenizer: A Case Study on Mandarin and English

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-03T10:13:57.724078Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:13:57.724078Z digest=sha256:a1634ab720bacfa8eb74602b62f921dfbc8a0be1cf51921c3387e29afa34ffb0

Observation 92418081-62ad-4080-9924-bd5911203b0b · outbound

This paper cites The WER Trap: Shattering the Illusion of Unified Tokens in Speech Language Models.

ParaASR: Multi-Token Prediction for Fast and Long-Context LLM-Based Speech Recognition The WER Trap: Shattering the Illusion of Unified Tokens in Speech Language Models

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-03T10:13:57.787394Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:13:57.787394Z digest=sha256:ea86fd4f513a7977152c551c8a93a5bf401d82b8968d671e4375a9458fc09127

Observation 3b207db3-9e4f-48f1-b3d8-f17750f315d2 · outbound

This paper cites IEEE Transactions on Audio, Speech and Language Processing (2025).

ParaASR: Multi-Token Prediction for Fast and Long-Context LLM-Based Speech Recognition IEEE Transactions on Audio, Speech and Language Processing (2025)

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-03T10:13:57.856488Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:13:57.856488Z digest=sha256:a61cdf7b48a57f6b02d9ee9a69eeb85c84f33fa88e7bae98ac5c24f53481820b

Observation cad70a16-691f-4939-ab2e-c04847e2c8b7 · outbound

This paper cites Step-Audio-R1.5 Technical Report.

ParaASR: Multi-Token Prediction for Fast and Long-Context LLM-Based Speech Recognition Step-Audio-R1.5 Technical Report

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-03T10:13:57.954776Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:13:57.954776Z digest=sha256:3227c19bd2da6e294939c9135df1634a8d746101264e8eaca4da027a3f88e094

Pith citing papers

No inbound Pith citation observations are available.