Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T11:26:23.320594Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 39 of 39 outbound references and 2 inbound Pith citation observations for arXiv:2506.02457.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T11:26:23.320594Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T11:26:22.389818Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-21T07:39:48.680240Z
39 of 39 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 24a79a7e-e7f3-46b6-8e29-f2eabedc2326 · outbound
SOVA-Bench: Benchmarking the Speech Conversation Ability for LLM-based Voice Assistant SOVA-Bench: Benchmarking the Speech Conversation Ability for LLM-based Voice Assistant
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6935127e-fa43-4f0d-b77e-26f4d1a3a331 · outbound
SOVA-Bench: Benchmarking the Speech Conversation Ability for LLM-based Voice Assistant Speech LLM Speech LLM extends the understanding capability to speech flow, performing modality alignment between speech and text via an encoder with adaptors
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 335ffbc3-ee3e-4578-a789-24643a2af936 · outbound
SOVA-Bench: Benchmarking the Speech Conversation Ability for LLM-based Voice Assistant Unresolved cited work
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3007e468-ce9c-436f-8f31-d8c14bf5e103 · outbound
SOVA-Bench: Benchmarking the Speech Conversation Ability for LLM-based Voice Assistant Unresolved cited work
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 741100b8-5df6-4d9e-8de8-ec39610b01af · outbound
SOVA-Bench: Benchmarking the Speech Conversation Ability for LLM-based Voice Assistant Unresolved cited work
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation eed943fb-c74a-42e8-86bb-cac7be6255bb · outbound
SOVA-Bench: Benchmarking the Speech Conversation Ability for LLM-based Voice Assistant Unresolved cited work
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 04f99b43-0c47-40cd-b373-185506acfab5 · outbound
SOVA-Bench: Benchmarking the Speech Conversation Ability for LLM-based Voice Assistant GPT-4o System Card
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b1ef39e5-7134-495b-aee8-c300409e6154 · outbound
SOVA-Bench: Benchmarking the Speech Conversation Ability for LLM-based Voice Assistant Mini-Omni: Language Models Can Hear, Talk While Thinking in Streaming
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3a38c437-6768-4ac8-b6b1-defd92aa8332 · outbound
SOVA-Bench: Benchmarking the Speech Conversation Ability for LLM-based Voice Assistant LLaMA-Omni: Seamless Speech Interaction with Large Language Models
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f821b026-4326-4f47-b148-bb854874ed7f · outbound
SOVA-Bench: Benchmarking the Speech Conversation Ability for LLM-based Voice Assistant Moshi: a speech-text foundation model for real-time dialogue
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 657ac16c-0126-4684-af44-6e56729f7170 · outbound
SOVA-Bench: Benchmarking the Speech Conversation Ability for LLM-based Voice Assistant Dynamic- SUPERB: Towards a dynamic, collaborative, and comprehensive instruction-tuning benchmark for speech,
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5aad3ced-54b1-4a76-8309-61beb49207e2 · outbound
SOVA-Bench: Benchmarking the Speech Conversation Ability for LLM-based Voice Assistant Dynamic-SUPERB Phase-2: A Collaboratively Expanding Benchmark for Measuring the Capabilities of Spoken Language Models with 180 Tasks
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0db694d5-841a-41cf-99e1-1c9efe1a8c85 · outbound
SOVA-Bench: Benchmarking the Speech Conversation Ability for LLM-based Voice Assistant AudioBench: A Universal Benchmark for Audio Large Language Models
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1dae2d60-7300-4edf-ad27-3acb7069dff2 · outbound
SOVA-Bench: Benchmarking the Speech Conversation Ability for LLM-based Voice Assistant AIR-Bench: Benchmarking Large Audio-Language Models via Generative Comprehension
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 277382e1-99a2-4fc9-ab1b-e2432272bb66 · outbound
SOVA-Bench: Benchmarking the Speech Conversation Ability for LLM-based Voice Assistant VoiceBench: Benchmarking LLM-Based Voice Assistants
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3208b10f-2a13-4b05-85bf-ef46656c5d91 · outbound
SOVA-Bench: Benchmarking the Speech Conversation Ability for LLM-based Voice Assistant SALMONN:Towards generic hearing abilities for large language models,
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c243da26-5a10-41da-9c47-b0fc12b019fb · outbound
SOVA-Bench: Benchmarking the Speech Conversation Ability for LLM-based Voice Assistant SpeechGPT: Empowering large language models with intrinsic cross-modal conversational abilities,
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2f99f1ab-e977-4b39-b9f2-dc13e6840bf7 · outbound
SOVA-Bench: Benchmarking the Speech Conversation Ability for LLM-based Voice Assistant Qwen2-Audio Technical Report
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d72b60f0-b1c0-4ad4-86ed-f944745a5b04 · outbound
SOVA-Bench: Benchmarking the Speech Conversation Ability for LLM-based Voice Assistant SNAC: Multi- scale neural audio codec,
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 78bf5a99-e0a4-4aac-ba3e-64c54fd2a8ab · outbound
SOVA-Bench: Benchmarking the Speech Conversation Ability for LLM-based Voice Assistant Hubert: Self-supervised speech represen- tation learning by masked prediction of hidden units,
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 25d6d93a-a35b-43e3-a76c-698b211e4008 · outbound
SOVA-Bench: Benchmarking the Speech Conversation Ability for LLM-based Voice Assistant HiFi-GAN: Generative adversar- ial networks for efficient and high fidelity speech synthesis,
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 81babd28-831b-43ef-be8d-cef09381c961 · outbound
SOVA-Bench: Benchmarking the Speech Conversation Ability for LLM-based Voice Assistant Speech resynthesis from discrete disentangled self-supervised representations,
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 37d13974-9f50-4609-bf99-3a907d1da2c4 · outbound
SOVA-Bench: Benchmarking the Speech Conversation Ability for LLM-based Voice Assistant Westlake-Omni,
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation fd4d79c4-b7f4-466a-bb6d-a80d9ec59e96 · outbound
SOVA-Bench: Benchmarking the Speech Conversation Ability for LLM-based Voice Assistant Freeze-Omni: A Smart and Low Latency Speech-to-speech Dialogue Model with Frozen LLM
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f725328a-7a76-49dc-a579-cb24b7ecc36c · outbound
SOVA-Bench: Benchmarking the Speech Conversation Ability for LLM-based Voice Assistant Be- yond turn-based interfaces: Synchronous llms as full-duplex dia- logue agents,
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 101c2dd5-5392-419c-820a-30b3c3d19b08 · outbound
SOVA-Bench: Benchmarking the Speech Conversation Ability for LLM-based Voice Assistant OmniFlatten: An End-to-end GPT Model for Seamless Voice Conversation
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b577a389-a753-4102-adbf-4c98624ca753 · outbound
SOVA-Bench: Benchmarking the Speech Conversation Ability for LLM-based Voice Assistant Baichuan-Omni-1.5 technical report,
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f9735e47-eb7d-479c-a8ac-03dc3d0f8cf8 · outbound
SOVA-Bench: Benchmarking the Speech Conversation Ability for LLM-based Voice Assistant TriviaQA: A large scale distantly supervised challenge dataset for reading comprehension,
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2733f0ee-bbe7-44dd-b717-497f1ce9d094 · outbound
SOVA-Bench: Benchmarking the Speech Conversation Ability for LLM-based Voice Assistant Lib- rispeech: an asr corpus based on public domain audio books,
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ffa586ec-7787-4748-bc32-a3b4aa3794d3 · outbound
SOVA-Bench: Benchmarking the Speech Conversation Ability for LLM-based Voice Assistant LibriSQA: A novel dataset and framework for spoken question answering with large language models,
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 4ad395e0-aac3-43e9-820e-feafba8e888d · outbound
SOVA-Bench: Benchmarking the Speech Conversation Ability for LLM-based Voice Assistant Spoken SQuAD: A study of mitigating the impact of speech recognition errors on listening comprehension,
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6d061c8f-99cc-4c34-97b1-977e6766fafb · outbound
SOVA-Bench: Benchmarking the Speech Conversation Ability for LLM-based Voice Assistant IEMOCAP: Interactive emotional dyadic motion capture database,
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3f915456-202c-420d-b2bd-e2b094f68824 · outbound
SOVA-Bench: Benchmarking the Speech Conversation Ability for LLM-based Voice Assistant Common V oice: A massively-multilingual speech corpus,
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 16d4b01d-6eed-45ce-aa65-308b97d668b2 · outbound
SOVA-Bench: Benchmarking the Speech Conversation Ability for LLM-based Voice Assistant Stanford alpaca: an instruction- following llama model (2023),
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3ca2bfd7-9187-4d2a-945d-c5e9210fa9d0 · outbound
SOVA-Bench: Benchmarking the Speech Conversation Ability for LLM-based Voice Assistant The T05 System for The VoiceMOS Challenge 2024: Transfer Learning from Deep Image Classifier to Naturalness MOS Prediction of High-Quality Synthetic Speech
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f82a96c8-3bec-43ff-899d-1ad5b1167973 · outbound
SOVA-Bench: Benchmarking the Speech Conversation Ability for LLM-based Voice Assistant CosyVoice: A Scalable Multilingual Zero-shot Text-to-speech Synthesizer based on Supervised Semantic Tokens
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3141e860-23e4-4cdc-b2df-98ae5a89919b · outbound
SOVA-Bench: Benchmarking the Speech Conversation Ability for LLM-based Voice Assistant Robust speech recognition via large-scale weak supervision,
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4b7c8061-2bde-426b-8c21-ff56e53f4e71 · outbound
SOVA-Bench: Benchmarking the Speech Conversation Ability for LLM-based Voice Assistant Mini-Omni2: Towards Open-source GPT-4o with Vision, Speech and Duplex Capabilities
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8b15efbd-c7c1-496d-b92e-9470a0a2afa2 · outbound
SOVA-Bench: Benchmarking the Speech Conversation Ability for LLM-based Voice Assistant GLM-4-Voice: Towards Intelligent and Human-Like End-to-End Spoken Chatbot
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 24a79a7e-e7f3-46b6-8e29-f2eabedc2326 · inbound
SOVA-Bench: Benchmarking the Speech Conversation Ability for LLM-based Voice Assistant SOVA-Bench: Benchmarking the Speech Conversation Ability for LLM-based Voice Assistant
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 77213481-1a76-4b50-ab3d-5c2d1777249c · inbound
A Survey of Large Audio Language Models: Generalization, Trustworthiness, and Outlook SOVA-Bench: Benchmarking the Speech Conversation Ability for LLM-based Voice Assistant
Reference 190
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.