Pith. sign in

Paper Citation Record · LEDGER

Seed-ASR: Understanding Diverse Speech and Contexts with LLM-based Speech Recognition

As of 11 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 40 inbound Pith citation observations for arXiv:2407.04675.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2407.04675 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 40 of 40 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 40 of 40 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-11T18:08:22.102985Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-05T13:21:06.342183Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 647a6ebd-3673-4ad8-b35c-6f63a7df5f33 · inbound

TouchTTS: An Embarrassingly Simple TTS Framework that Everyone Can Touch cites this paper.

TouchTTS: An Embarrassingly Simple TTS Framework that Everyone Can Touch Seed-ASR: Understanding Diverse Speech and Contexts with LLM-based Speech Recognition

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-11T18:08:22.102985Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T18:08:22.102985Z digest=sha256:c85672223992a7537df08aaec7d23a557b518a944744a3ab127ff67ca46e8ec6

Observation cb05912a-b141-4d17-b5f9-69cd95bac048 · inbound

MERaLiON-SpeechEncoder: Towards a Speech Foundation Model for Singapore and Beyond cites this paper.

MERaLiON-SpeechEncoder: Towards a Speech Foundation Model for Singapore and Beyond Seed-ASR: Understanding Diverse Speech and Contexts with LLM-based Speech Recognition

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-11T14:54:03.571131Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T14:54:03.571131Z digest=sha256:dcfc5dc942eb9bbfb3c87df9f0a6263c84bb8461ea798f76c07284a04cd4d525

Observation 7da03c5d-ffa7-4dad-86fa-2e57f10cecaa · inbound

TouchASP: Elastic Automatic Speech Perception that Everyone Can Touch cites this paper.

TouchASP: Elastic Automatic Speech Perception that Everyone Can Touch Seed-ASR: Understanding Diverse Speech and Contexts with LLM-based Speech Recognition

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-11T11:18:33.773017Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T11:18:33.773017Z digest=sha256:7c191749f38b2d508085f60db7b7cbadc3a0904741acef0ac0b95983d1f4593a

Observation a85879f1-7564-47f2-8945-97aa42eeac38 · inbound

Interleaved Speech-Text Language Models for Simple Streaming Text-to-Speech Synthesis cites this paper.

Interleaved Speech-Text Language Models for Simple Streaming Text-to-Speech Synthesis Seed-ASR: Understanding Diverse Speech and Contexts with LLM-based Speech Recognition

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-11T10:50:48.808748Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T10:50:48.808748Z digest=sha256:6b1857d86923eb6ce2d52c990a8ef117f5feadf58f6a1acddfed61ec752b76ff

Observation b6c0bdbc-a086-47c4-b397-24b63caebdb7 · inbound

MuQ: Self-Supervised Music Representation Learning with Mel Residual Vector Quantization cites this paper.

MuQ: Self-Supervised Music Representation Learning with Mel Residual Vector Quantization Seed-ASR: Understanding Diverse Speech and Contexts with LLM-based Speech Recognition

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-10T22:40:21.995444Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T22:40:21.995444Z digest=sha256:d83771259cb23ea7827e15750193969a53838407535a1cf2faa088007c1d7818

Observation 22b1f541-060f-41ed-8f1c-ad0228b1741b · inbound

Audio-Language Models for Audio-Centric Tasks: A Systematic Survey cites this paper.

Audio-Language Models for Audio-Centric Tasks: A Systematic Survey Seed-ASR: Understanding Diverse Speech and Contexts with LLM-based Speech Recognition

Reference 100

Resolution
unresolved
no resolver link, observed 2026-08-10T14:36:19.597943Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:36:19.597943Z digest=sha256:b97f1959236cf5bba12fddb67c2508c70fc1c51bfc44ffa71e33eca317c7b668

Observation bac61da4-d3b8-46ea-b07f-19f951584281 · inbound

Qwen2.5-Omni Technical Report cites this paper.

Qwen2.5-Omni Technical Report Seed-ASR: Understanding Diverse Speech and Contexts with LLM-based Speech Recognition

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-10T17:54:03.268853Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-10T17:54:03.225439Z digest=sha256:d4126aeccb8adc2cc2636766037baef1bb4a1cb4f729bd0aae6d1238f31c92d7

Observation 888acafe-876e-438d-99b9-60c6a4d3cabd · inbound

From Tens of Hours to Tens of Thousands: Scaling Back-Translation for Speech Recognition cites this paper.

From Tens of Hours to Tens of Thousands: Scaling Back-Translation for Speech Recognition Seed-ASR: Understanding Diverse Speech and Contexts with LLM-based Speech Recognition

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T14:55:54.427250Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:55:54.427250Z digest=sha256:80feb92648c36ce83609c7f62b984282d41dc3bd3c797fdb77663e0cafa77e4a

Observation b4e8642d-3e79-4e28-b2cc-48be220962f8 · inbound

Weakly Supervised Data Refinement and Flexible Sequence Compression for Efficient Thai LLM-based ASR cites this paper.

Weakly Supervised Data Refinement and Flexible Sequence Compression for Efficient Thai LLM-based ASR Seed-ASR: Understanding Diverse Speech and Contexts with LLM-based Speech Recognition

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T13:21:01.949770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:21:01.949770Z digest=sha256:511a7c29892950bcf1b33dd5ae200b5293319d79db0a6a4ce95671e1cf275e66

Observation e5ce7bd1-b68d-4f4e-aaa7-1dca9bbc100b · inbound

Leveraging Large Language Models in Visual Speech Recognition: Model Scaling, Context-Aware Decoding, and Iterative Polishing cites this paper.

Leveraging Large Language Models in Visual Speech Recognition: Model Scaling, Context-Aware Decoding, and Iterative Polishing Seed-ASR: Understanding Diverse Speech and Contexts with LLM-based Speech Recognition

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T13:26:50.575343Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:26:50.575343Z digest=sha256:955497a6392fac5e753619bd92104d0b67294edb263d081bdab3aa93ed304c34

Observation 7a4edad8-3398-4a25-ac9f-c211a6ad67ee · inbound

Large Language models for Time Series Analysis: Techniques, Applications, and Challenges cites this paper.

Large Language models for Time Series Analysis: Techniques, Applications, and Challenges Seed-ASR: Understanding Diverse Speech and Contexts with LLM-based Speech Recognition

Reference 113

Resolution
unresolved
no resolver link, observed 2026-08-07T15:27:00.351719Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:27:00.351719Z digest=sha256:b77f65dbc2f723fd4480cf705faabd1e5c71dc47f236d3881c84fbebbb551908

Observation ba4234d8-d03c-420c-8bc8-b5ebc8443421 · inbound

PMF-CEC: Phoneme-augmented Multimodal Fusion for Context-aware ASR Error Correction with Error-specific Selective Decoding cites this paper.

PMF-CEC: Phoneme-augmented Multimodal Fusion for Context-aware ASR Error Correction with Error-specific Selective Decoding Seed-ASR: Understanding Diverse Speech and Contexts with LLM-based Speech Recognition

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T12:09:55.809973Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:09:55.809973Z digest=sha256:82fa5e2621f25ff2ebab8998e39445dbdac7c968edb13fd3dadc9a56b927aa0e

Observation a1f20979-b550-4282-b092-91f0057acbfd · inbound

CMT-LLM: Contextual Multi-Talker ASR Utilizing Large Language Models cites this paper.

CMT-LLM: Contextual Multi-Talker ASR Utilizing Large Language Models Seed-ASR: Understanding Diverse Speech and Contexts with LLM-based Speech Recognition

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T12:09:29.130086Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:09:29.130086Z digest=sha256:6ef2ebd10a2f927ea392e4edfbb482e35356fd12bbd877a29bf5570793bb1b99

Observation e5bf7f3a-5aa8-4e1e-9ae7-b0fc88ca08ec · inbound

Improving Contextual ASR via Multi-grained Fusion with Large Language Models cites this paper.

Improving Contextual ASR via Multi-grained Fusion with Large Language Models Seed-ASR: Understanding Diverse Speech and Contexts with LLM-based Speech Recognition

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T16:56:51.734646Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:56:51.734646Z digest=sha256:bdd6e949396fd03a20d594c0353848a296854db3972e35422f0ce82323f17e82

Observation eb5c602e-02af-4a2b-a973-a8bddc81995d · inbound

Step-Audio 2 Technical Report cites this paper.

Step-Audio 2 Technical Report Seed-ASR: Understanding Diverse Speech and Contexts with LLM-based Speech Recognition

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-16T05:59:51.053374Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-16T05:59:50.900436Z digest=sha256:b17ec141f9eac769e98a50b44ab307f5971d68587b02635e9679b52b3b278560

Observation dddc8553-5bbc-4af2-9566-eaf72d02e81b · inbound

Seed LiveInterpret 2.0: End-to-end Simultaneous Speech-to-speech Translation with Your Voice cites this paper.

Seed LiveInterpret 2.0: End-to-end Simultaneous Speech-to-speech Translation with Your Voice Seed-ASR: Understanding Diverse Speech and Contexts with LLM-based Speech Recognition

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T14:52:42.951815Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:52:42.951815Z digest=sha256:274a0c303da57915605b8d6deecbbe7e2d35e1822de400d9c30854c4b6390b4d

Observation 28c0212d-65f8-4799-9caa-4ca13afb055e · inbound

The TEA-ASLP System for Multilingual Conversational Speech Recognition and Speech Diarization in MLC-SLM 2025 Challenge cites this paper.

The TEA-ASLP System for Multilingual Conversational Speech Recognition and Speech Diarization in MLC-SLM 2025 Challenge Seed-ASR: Understanding Diverse Speech and Contexts with LLM-based Speech Recognition

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T14:45:25.066761Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:45:25.066761Z digest=sha256:9f8925acf80d7b914d328a5f334a661ffef6fe3ab75b69c1d4ccc7f5557bcf51

Observation b80dc020-9196-44e7-8b93-c1503b356253 · inbound

SpecASR: Accelerating LLM-based Automatic Speech Recognition via Speculative Decoding cites this paper.

SpecASR: Accelerating LLM-based Automatic Speech Recognition via Speculative Decoding Seed-ASR: Understanding Diverse Speech and Contexts with LLM-based Speech Recognition

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T14:42:56.037314Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:42:56.037314Z digest=sha256:8a26dddcd18b4c2c86ddd18ec98b4bbe4f720df6461f7c0bf719796a239fb2e9

Observation 49e04d82-0591-4a2b-bcfa-165cee1248a3 · inbound

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems cites this paper.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Seed-ASR: Understanding Diverse Speech and Contexts with LLM-based Speech Recognition

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:42.749315Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:42.749315Z digest=sha256:6ff1ad908ff7028d5e050030b0079e76d5ebc11c2cf4b6aca4b3ea76f4cdd26b

Observation 5442749b-5a01-481a-9369-9a5ec5530831 · inbound

Cross-Learning Fine-Tuning Strategy for Dysarthric Speech Recognition Via CDSD database cites this paper.

Cross-Learning Fine-Tuning Strategy for Dysarthric Speech Recognition Via CDSD database Seed-ASR: Understanding Diverse Speech and Contexts with LLM-based Speech Recognition

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-05T16:18:50.808762Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T16:18:50.808762Z digest=sha256:1f6c4ef219c4c7fa21d508c4aee155d7f1c30a216fb7e118a9371671106ac053

Observation 09b12923-4b21-4f31-a03c-f2275b0f650f · inbound

Denoising GER: A Noise-Robust Generative Error Correction with LLM for Speech Recognition cites this paper.

Denoising GER: A Noise-Robust Generative Error Correction with LLM for Speech Recognition Seed-ASR: Understanding Diverse Speech and Contexts with LLM-based Speech Recognition

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-05T10:16:04.880310Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:16:04.880310Z digest=sha256:612ad93d817d94cdd5b09822e3e8edd2dba6348b70e727e02f038011bb13c8ee

Observation 326f4580-d5aa-41b9-b654-8a60e83ef953 · inbound

UniVoice: Unifying Autoregressive ASR and Flow-Matching based TTS with Large Language Models cites this paper.

UniVoice: Unifying Autoregressive ASR and Flow-Matching based TTS with Large Language Models Seed-ASR: Understanding Diverse Speech and Contexts with LLM-based Speech Recognition

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-04T11:29:30.053883Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T11:29:30.053883Z digest=sha256:efeeb0f611373f58e54df24cfb5f2c0bc08f3282860b44b6d8cb3c728205dde6

Observation 102af8f7-1d4f-4709-9e34-64fc416c88ca · inbound

LLMs and Speech: Integration vs. Combination cites this paper.

LLMs and Speech: Integration vs. Combination Seed-ASR: Understanding Diverse Speech and Contexts with LLM-based Speech Recognition

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-15T10:45:28.319473Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-15T10:41:22.138517Z digest=sha256:e9385119e3d7818e3001300b52c8e5a242de6930b80199110dd5f627356aa21f

Observation da4427f9-e0cb-4de6-9412-be9608b96132 · inbound

LLMs and Speech: Integration vs. Combination cites this paper.

LLMs and Speech: Integration vs. Combination Seed-ASR: Understanding Diverse Speech and Contexts with LLM-based Speech Recognition

Reference 17

Resolution
unresolved
no resolver link, observed 2026-07-14T20:46:17.286576Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T20:46:17.286576Z digest=sha256:c1572905de2a273dcbc0ee0da3e62e1b558fd7eb82d05f5520a83367e25c5c6a

Observation 3370149f-6561-4fa2-8df7-87d529d972e4 · inbound

Rethinking Entropy Allocation in LLM-based ASR: Understanding the Dynamics between Speech Encoders and LLMs cites this paper.

Rethinking Entropy Allocation in LLM-based ASR: Understanding the Dynamics between Speech Encoders and LLMs Seed-ASR: Understanding Diverse Speech and Contexts with LLM-based Speech Recognition

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-11T05:30:57.139265Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-10T18:06:50.408402Z digest=sha256:26c069ea2b88811f8d4a8f9f2b099f33ab433003f285c97ad9a38e45e5f26a3b

Observation 0bae1a1b-aacb-4176-a926-4b627897d09d · inbound

Interactive ASR: Towards Human-Like Interaction and Semantic Coherence Evaluation for Agentic Speech Recognition cites this paper.

Interactive ASR: Towards Human-Like Interaction and Semantic Coherence Evaluation for Agentic Speech Recognition Seed-ASR: Understanding Diverse Speech and Contexts with LLM-based Speech Recognition

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-11T00:41:07.223921Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-10T18:22:08.670559Z digest=sha256:98d554342378dfce9c1c287b59e9712bdc59a2015435f720eb151cf317ce686c

Observation f82fd620-648e-48fa-83ec-afae6d06c116 · inbound

NIM4-ASR: Towards Efficient, Robust, and Customizable Real-Time LLM-Based ASR cites this paper.

NIM4-ASR: Towards Efficient, Robust, and Customizable Real-Time LLM-Based ASR Seed-ASR: Understanding Diverse Speech and Contexts with LLM-based Speech Recognition

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-11T12:21:05.820918Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-10T03:48:14.211240Z digest=sha256:085275e020761265f1872520ee4b61ea3a06c2d7f78d35a9015d32558634b2b4

Observation 7da1bb08-7c34-4060-b9c9-caab82427ab1 · inbound

NIM4-ASR: Towards Efficient, Robust, and Customizable Real-Time LLM-Based ASR cites this paper.

NIM4-ASR: Towards Efficient, Robust, and Customizable Real-Time LLM-Based ASR Seed-ASR: Understanding Diverse Speech and Contexts with LLM-based Speech Recognition

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-07-05T13:21:06.346052Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-07-05T13:15:50.794969Z digest=sha256:0f0d9397ab4e10e4c407937ed76910ed4d126c3aecfc0f94eb92351ab39fb280

Observation 2e54763f-9c1a-474c-b355-1c5d1fa33262 · inbound

From Synthesis to Clinical Assistance: A Strategy-Aware Agent Framework for Autism Intervention based on Real Clinical Dataset cites this paper.

From Synthesis to Clinical Assistance: A Strategy-Aware Agent Framework for Autism Intervention based on Real Clinical Dataset Seed-ASR: Understanding Diverse Speech and Contexts with LLM-based Speech Recognition

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T06:30:59.129914Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-10T17:36:58.858864Z digest=sha256:89194596a5318a2094a6fda7067c23da7af04b265952703955a6f3dabf9a5847

Observation bcd6b14d-8c89-4547-97c8-52ce9ab99065 · inbound

VocalParse: Towards Unified and Scalable Singing Voice Transcription with Large Audio Language Models cites this paper.

VocalParse: Towards Unified and Scalable Singing Voice Transcription with Large Audio Language Models Seed-ASR: Understanding Diverse Speech and Contexts with LLM-based Speech Recognition

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-11T17:51:09.717541Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-08T16:59:25.973854Z digest=sha256:912fb693445068383cc853f8d8a90e48dd9133f6602b7929c87acd73646229c4

Observation d9ededf9-0e59-4227-8d4e-c17da6fba2ce · inbound

JSPG: Dynamic Dictionary Filtering via Joint Semantic-Pinyin-Glyph Retrieval for Chinese Contextual ASR cites this paper.

JSPG: Dynamic Dictionary Filtering via Joint Semantic-Pinyin-Glyph Retrieval for Chinese Contextual ASR Seed-ASR: Understanding Diverse Speech and Contexts with LLM-based Speech Recognition

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-19T20:57:46.907545Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-19T20:56:04.778077Z digest=sha256:62a30105fc13551ae4fdfd4f1918be0f3e46a8ffdf3c5e9c8acf64213f31eb39

Observation 91001d88-a7f2-4930-8541-dc76486a95f5 · inbound

StepAudio 2.5 Technical Report cites this paper.

StepAudio 2.5 Technical Report Seed-ASR: Understanding Diverse Speech and Contexts with LLM-based Speech Recognition

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-25T02:55:16.515698Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-25T02:52:22.610397Z digest=sha256:0fc98a0176d4fb391715dd77d0410383659702b1289529e753d2994793b7fe33

Observation 42e76763-52ed-49d5-94cd-e443283301dd · inbound

Audio Interaction Model cites this paper.

Audio Interaction Model Seed-ASR: Understanding Diverse Speech and Contexts with LLM-based Speech Recognition

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-07-02T10:46:52.389064Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-06-28T04:57:05.062465Z digest=sha256:14f0f78ddee84b16f9d7963dc17f7ff53919aad0c6de0ce898be6364783b0ed3

Observation 1044f1c1-3172-47e7-a2df-c706a6340555 · inbound

TRADE: Transducer-Augmented Decoder for Speech LLM cites this paper.

TRADE: Transducer-Augmented Decoder for Speech LLM Seed-ASR: Understanding Diverse Speech and Contexts with LLM-based Speech Recognition

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-07-02T22:47:25.686348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-06-27T18:40:19.688550Z digest=sha256:b504054b0da4424416673a5421e0002d113cf38cefeb0b514fb497ee46b6ddd5

Observation e18ee8f2-5b9b-449a-9983-987079b5e7ee · inbound

Rethinking Speech-LLM Integration for ASR: Effective Joint Speech-Text Training by Interleaving cites this paper.

Rethinking Speech-LLM Integration for ASR: Effective Joint Speech-Text Training by Interleaving Seed-ASR: Understanding Diverse Speech and Contexts with LLM-based Speech Recognition

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-07-03T15:18:32.637916Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-07-03T15:17:52.966144Z digest=sha256:abef2d8c5ab61e78ba6093ed74052bb18382c6197f1f12326bb6a0d8b3855e62

Observation 4641fd45-9eeb-4ad6-84e6-8c6a34f9414e · inbound

Context-Aware ASR for Mandarin Technical Lectures cites this paper.

Context-Aware ASR for Mandarin Technical Lectures Seed-ASR: Understanding Diverse Speech and Contexts with LLM-based Speech Recognition

Reference 17

Resolution
unresolved
no resolver link, observed 2026-07-11T09:26:54.286790Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T09:26:54.286790Z digest=sha256:38f4fec5da96cef2ffcf613e17dc63d786aab8679431bbfcb5fb183f044dcbe7

Observation 7d9304c1-3dee-415d-92b3-b42649a97cf5 · inbound

ParaASR: Multi-Token Prediction for Fast and Long-Context LLM-Based Speech Recognition cites this paper.

ParaASR: Multi-Token Prediction for Fast and Long-Context LLM-Based Speech Recognition Seed-ASR: Understanding Diverse Speech and Contexts with LLM-based Speech Recognition

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-03T10:13:54.253318Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:13:54.253318Z digest=sha256:388afeb1bf8f15a2c280976e52a95a905ce3bd00162423b7646bf5c0e2965125

Observation 85433cbc-b637-4948-8f02-ace2f144e042 · inbound

InteracVid: Building a Real Interactive Audio-Visual Response Dataset from Live-Chat Videos cites this paper.

InteracVid: Building a Real Interactive Audio-Visual Response Dataset from Live-Chat Videos Seed-ASR: Understanding Diverse Speech and Contexts with LLM-based Speech Recognition

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T00:31:25.062446Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:31:25.062446Z digest=sha256:71ac45e3094a53a194bf9c5bbc31c8d7585c30b3f3905f4bda7406ef41e52ba0

Observation 071cbf86-bfd9-4d4e-9c88-3ad455cd2ed5 · inbound

Language-Specialized Multi-Teacher On-Policy Distillation for Multilingual LLM-Based ASR cites this paper.

Language-Specialized Multi-Teacher On-Policy Distillation for Multilingual LLM-Based ASR Seed-ASR: Understanding Diverse Speech and Contexts with LLM-based Speech Recognition

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-05T16:07:18.513493Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T16:07:18.513493Z digest=sha256:c37733eceb22fc98753f53f85c58e7e869cda5716cc0c821404cfb43c460d939

Observation d5cd0330-cb34-4039-8265-c9ab44e1570d · inbound

Language-Specialized Multi-Teacher On-Policy Distillation for Multilingual LLM-Based ASR cites this paper.

Language-Specialized Multi-Teacher On-Policy Distillation for Multilingual LLM-Based ASR Seed-ASR: Understanding Diverse Speech and Contexts with LLM-based Speech Recognition

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-11T04:22:00.929409Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T04:22:00.929409Z digest=sha256:49032783347623d679596433e3e82efd57f55c8d41fca0eadad69309a553142b