Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 6 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 32 inbound Pith citation observations for arXiv:2402.08846.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-06T04:17:35.289242Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-10T20:37:34.440026Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation d2ab4cb9-f143-45b2-9f74-ba441f2be27a · inbound
Dual Information Speech Language Models for Emotional Conversations An Embarrassingly Simple Approach for LLM with Strong ASR Capacity
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8db43db0-dba1-4fca-8ce9-bd411e41521b · inbound
RAG-Boost: Retrieval-Augmented Generation Enhanced LLM-based Speech Recognition An Embarrassingly Simple Approach for LLM with Strong ASR Capacity
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b8005c93-cc00-44fa-bb0f-4ec3663f192a · inbound
Transsion Multilingual Speech Recognition System for MLC-SLM 2025 Challenge An Embarrassingly Simple Approach for LLM with Strong ASR Capacity
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ca7456a1-cfec-42dc-8ed9-fd7c83588c02 · inbound
Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs An Embarrassingly Simple Approach for LLM with Strong ASR Capacity
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 38f2f1b3-3766-4a55-a261-a2edd83ae08c · inbound
TokenVerse++: Towards Flexible Multitask Learning with Dynamic Task Activation An Embarrassingly Simple Approach for LLM with Strong ASR Capacity
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dabc5971-19b3-45e2-8aa3-f0fbed91b423 · inbound
Group Relative Policy Optimization for Speech Recognition An Embarrassingly Simple Approach for LLM with Strong ASR Capacity
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 16abd9eb-df37-4f74-9542-77af2a53c1b2 · inbound
Denoising GER: A Noise-Robust Generative Error Correction with LLM for Speech Recognition An Embarrassingly Simple Approach for LLM with Strong ASR Capacity
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0a2508d2-110c-4be6-afcc-a424821966ce · inbound
SpeechLLM: Unified Speech and Language Model for Enhanced Multi-Task Understanding in Low Resource Settings An Embarrassingly Simple Approach for LLM with Strong ASR Capacity
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f0f964f5-1a8a-4393-b402-e7f00e95610b · inbound
UniVoice: Unifying Autoregressive ASR and Flow-Matching based TTS with Large Language Models An Embarrassingly Simple Approach for LLM with Strong ASR Capacity
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a0b59a83-9cb8-49fa-a2a4-4b2c8699e090 · inbound
Reducing Prompt Sensitivity in LLM-based Speech Recognition Through Learnable Projection An Embarrassingly Simple Approach for LLM with Strong ASR Capacity
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation ffd01fc1-cdd3-4e8c-b2a2-fd18b9a8ff10 · inbound
LLMs and Speech: Integration vs. Combination An Embarrassingly Simple Approach for LLM with Strong ASR Capacity
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 21e856e1-d7e0-4d88-bdde-df5b4978a8f5 · inbound
LLMs and Speech: Integration vs. Combination An Embarrassingly Simple Approach for LLM with Strong ASR Capacity
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8fd1d5bc-f9d8-4245-ac25-5860d4d827e1 · inbound
Closing the Speech-Text Gap with Limited Audio for Effective Domain Adaptation in LLM-Based ASR An Embarrassingly Simple Approach for LLM with Strong ASR Capacity
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 63ac30c2-c9f2-4acb-97ca-3450bc6c544e · inbound
Interactive ASR: Towards Human-Like Interaction and Semantic Coherence Evaluation for Agentic Speech Recognition An Embarrassingly Simple Approach for LLM with Strong ASR Capacity
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation a6796af1-5374-4c57-8f2f-e507b4237054 · inbound
Phonemes vs. Projectors: An Investigation of Speech-Language Interfaces for LLM-based ASR An Embarrassingly Simple Approach for LLM with Strong ASR Capacity
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 96265c51-0a44-4540-9c46-f7f1196ab634 · inbound
Speaker Attributed Automatic Speech Recognition Using Speech Aware LLMS An Embarrassingly Simple Approach for LLM with Strong ASR Capacity
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation d5c21b94-693d-4df0-ae81-e480800009b4 · inbound
Contextual Biasing for ASR in Speech LLM with Common Word Cues and Bias Word Position Prediction An Embarrassingly Simple Approach for LLM with Strong ASR Capacity
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 4f734acb-6cfd-4bdc-a430-ccea693940fc · inbound
In-Sync: Adaptation of Speech Aware Large Language Models for ASR with Word Level Timestamp Predictions An Embarrassingly Simple Approach for LLM with Strong ASR Capacity
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 1cc8376e-1167-47d1-84d9-45cd2b614353 · inbound
Refining Pseudo-Audio Prompts with Speech-Text Alignment for Text-Only Domain Adaptation in LLM-Based ASR An Embarrassingly Simple Approach for LLM with Strong ASR Capacity
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation c0f63a5a-61d1-4360-9057-207586259bc6 · inbound
Refining Pseudo-Audio Prompts with Speech-Text Alignment for Text-Only Domain Adaptation in LLM-Based ASR An Embarrassingly Simple Approach for LLM with Strong ASR Capacity
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation bdd076fe-6dd0-445d-b3f1-f57e0a6feeb6 · inbound
SoulX-Transcriber: A Robust End-to-End Framework for Multi-Speaker Speech Transcription An Embarrassingly Simple Approach for LLM with Strong ASR Capacity
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation d4277d40-1e2b-4807-87b2-982b3b0785da · inbound
TRADE: Transducer-Augmented Decoder for Speech LLM An Embarrassingly Simple Approach for LLM with Strong ASR Capacity
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 632df31e-02d9-4c56-b09d-cab846e7c658 · inbound
RespiraMFM: A Multimodal Foundation Model with Contrastive Audio-Language Alignment for Respiratory Disease Identification An Embarrassingly Simple Approach for LLM with Strong ASR Capacity
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation ea183f7d-918b-4e36-970f-7b983777c7f2 · inbound
Speech Meets ELF: Audio Conditional Continuous-Target Diffusion for Speech Recognition and Translation An Embarrassingly Simple Approach for LLM with Strong ASR Capacity
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation d7aacf95-318f-4cbf-965b-9d006b84d17b · inbound
Enhancing Multilingual LLM-based ASR with Mixture of Experts and Dynamic Downsampling An Embarrassingly Simple Approach for LLM with Strong ASR Capacity
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation cfb87c54-0c80-4817-b9a6-c43cd55448eb · inbound
Entropy-Aware Domain-Routed Mixture-of-Experts Speech-LLM Framework: A Case Study of Multi-Domain Child-Adult ASR An Embarrassingly Simple Approach for LLM with Strong ASR Capacity
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 94579e9f-e5a5-4763-b6ad-d074958d75a7 · inbound
Aligning MusicLLM with Emotion using Instruction Tuning and Feedback-Driven Alignment An Embarrassingly Simple Approach for LLM with Strong ASR Capacity
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation bdccff6e-7a42-4bff-9dbd-f250118945e4 · inbound
Does Translation-Enhanced Speech Encoder Pre-training Affect Speech LLMs? An Embarrassingly Simple Approach for LLM with Strong ASR Capacity
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 0867094c-0434-40cb-9141-7303a8ed75e6 · inbound
LuxSQA: Ask Me in Luxembourgish with TTS-Augmented Spoken Question Answering An Embarrassingly Simple Approach for LLM with Strong ASR Capacity
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3f7e0a52-5d58-43fa-926d-e60fa96128a2 · inbound
Compress the Cache, Not the Speech Embedding: KV Compression for Efficient Speech LLMs An Embarrassingly Simple Approach for LLM with Strong ASR Capacity
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation c72269b7-e68e-40fd-9e73-4195f71fe3d5 · inbound
When Synthetic Speech Is All You Have: Better Call GRPO An Embarrassingly Simple Approach for LLM with Strong ASR Capacity
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 809360d3-6ffb-4409-a6c6-00ac8e76d07b · inbound
MEUSLI: a Multilingual Projector for LLM-based ASR and Beyond An Embarrassingly Simple Approach for LLM with Strong ASR Capacity
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.