Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-12T18:57:28.666194Z
Paper Citation Record · LEDGER
As of 4 August 2026, this Paper Citation Record lists 37 of 37 outbound references and 100 inbound Pith citation observations for arXiv:2311.07919.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-12T18:57:28.666194Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-03T06:30:56.289259+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-04T03:03:32.342985Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-10T20:27:36.579792Z
37 of 37 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation f82b18c7-0b96-499e-bba2-fcb3970209e0 · outbound
Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models Spice: Semantic propositional image caption evaluation
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 0b34eb66-5d5e-4197-b53c-291d3fa9631a · outbound
Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models PaLM 2 Technical Report
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 1a9a4096-7be8-487e-bd11-85d4f50bebe4 · outbound
Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models SpeechT5: Unified-Modal Encoder-Decoder Pre-Training for Spoken Language Processing
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 686a88c1-2cb5-4cc3-9743-96b07e7caff0 · outbound
Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models Qwen Technical Report
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 978cba6e-4754-4698-89bd-496af42f7de5 · outbound
Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models AISHELL-1: an open-source mandarin speech corpus and a speech recognition baseline
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 40075b4d-b472-4db9-9b1e-be1236df4437 · outbound
Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models Shikra: Unleashing Multimodal LLM's Referential Dialogue Magic
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation c2e4f870-f674-4337-aaa5-83566ad92f9f · outbound
Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models SpeechNet: A Universal Modularized Model for Speech Processing Tasks
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 0389ca6e-f17d-4aec-b5e2-8381638cda49 · outbound
Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models PaLM: Scaling Language Modeling with Pathways
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 513cee5e-03c7-4c91-98bd-c697f2efda83 · outbound
Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models High Fidelity Neural Audio Compression
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 527b4f02-96cd-4cc8-b586-ec04dde22436 · outbound
Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models Clotho: an audio captioning dataset
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 68ff7646-e1bf-41d7-bcfa-203ecfe6a788 · outbound
Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models AISHELL-2: Transforming Mandarin ASR Research Into Industrial Scale
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 6f5b030d-787a-4213-8656-d5a8ac082472 · outbound
Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models CLAP: Learning Audio Concepts From Natural Language Supervision
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 8797769d-2f3c-4613-9ead-011e07268843 · outbound
Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models Engel, Cinjon Resnick, Adam Roberts, Sander Dieleman, Mohammad Norouzi, Douglas Eck, and Karen Simonyan
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 3ad2844e-c9f5-4714-bff0-1a4b7959d5fd · outbound
Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models FunASR: A Fundamental End-to-End Speech Recognition Toolkit
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 36888e76-61f8-40ff-8187-3bcc31335b98 · outbound
Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models Vocalsound: Adatasetforimprovinghumanvocalsoundsrecognition
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation f7b85189-ba7f-4e93-8573-4e11a561165d · outbound
Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models author Zhou, A
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 2af46fc4-917a-4c63-8f0e-57d6da47ab1e · outbound
Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models LoRA: Low-Rank Adaptation of Large Language Models
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation fe563662-f13d-4f5b-860d-d8d85889f490 · outbound
Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models AudioGPT: Understanding and Generating Speech, Music, Sound, and Talking Head
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 06221470-b5d6-497b-bdc2-9c210a4d8566 · outbound
Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models CochlScene: Acquisition of acoustic scene data using crowdsourcing
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 59b741bb-9023-4ffb-849d-3efec1ae8228 · outbound
Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models Unresolved cited work
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 85bb758f-2015-4066-8751-0af96520986f · outbound
Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models Clotho-aqa: A crowdsourceddatasetforaudioquestionanswering
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 0e5f935e-91f7-42dd-90d9-b091d4afbb92 · outbound
Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models Macaw-LLM: Multi-Modal Language Modeling with Image, Audio, Video, and Text Integration
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 8f2a679e-c6bd-4afd-850c-f503f410178f · outbound
Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models Voxtlm: unified decoder-only models for consolidating speech recognition/synthesis and speech/text continuation tasks
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 86c9db3c-e837-44ff-a9fc-598806601471 · outbound
Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models Montreal forced aligner: Trainable text-speech alignment using kaldi
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 9138ff39-1341-42d4-9330-8c7a9faac660 · outbound
Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models DCASE2017 challenge setup: Tasks, datasets and baseline system
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 7d1d4f61-5064-42a0-8e1d-3aa78034ab6a · outbound
Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models Librispeech: AnASRcorpusbasedon public domain audio books
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 24f53383-474b-4b6d-b5df-d15bf63ea882 · outbound
Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models Park, William Chan, Yu Zhang, Chung-Cheng Chiu, Barret Zoph, Ekin D
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 4578af68-d155-46d8-8b0e-6e1618754984 · outbound
Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models Kosmos-2: Grounding Multimodal Large Language Models to the World
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 1b0ec48a-8858-49fd-8f34-0c73b749ae2a · outbound
Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models MELD: A multimodal multi-party dataset for emotion recognition in conversations
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 5fe6b423-28aa-4c97-9865-085fe94460da · outbound
Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models Alec Radford, Jong Wook Kim, Tao Xu, Greg Brockman, Christine McLeavey, and Ilya Sutskever
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation fb10e4d5-8f2d-43af-8bde-5bdd7fc47cd6 · outbound
Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models HuggingGPT: Solving AI Tasks with ChatGPT and its Friends in Hugging Face
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 2bcc33b7-1fb3-4b57-ad7c-82fe00b902f9 · outbound
Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models LLaSM: Large Language and Speech Model
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation d1b73c67-50be-45a3-9491-94e911eddeed · outbound
Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models Emu: Generative Pretraining in Multimodality
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 0877f768-621a-4a4f-a3aa-e5cdd7c7ac10 · outbound
Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models LLaMA: Open and Efficient Foundation Language Models
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 492184fe-5f82-4b06-a1a1-3b6b33fe2ea8 · outbound
Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models CoVoST 2 and Massively Multilingual Speech-to-Text Translation
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation e09f75fc-0890-4b05-b290-52f56213f9f7 · outbound
Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models SpeechGPT: Empowering Large Language Models with Intrinsic Cross-Modal Conversational Abilities
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation fe3be170-510f-402f-8ce1-f13f025a6ccb · outbound
Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models Whisper-large-v2 Qwen-audio 1st-stage LLM init
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 1fce556b-0194-4d5f-9565-dc42f0d4f564 · inbound
Qwen2-Audio Technical Report Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 23e023fc-61b4-45fd-a4c8-b6b1a7d5011d · inbound
VoiceBench: Benchmarking LLM-Based Voice Assistants Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation af400f69-26f3-42aa-8581-4de09686fd5e · inbound
GLM-4-Voice: Towards Intelligent and Human-Like End-to-End Spoken Chatbot Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 5a11210e-e87e-4a6c-b8c7-d7f1f4be233d · inbound
WorldSense: Evaluating Real-world Omnimodal Understanding for Multimodal LLMs Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation fee4a322-b101-4a6d-8353-d3e47ddd677e · inbound
Multimodal Chain-of-Thought Reasoning: A Comprehensive Survey Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 199
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation f7f4e2d5-eef8-4afe-ac6e-be2886e2d488 · inbound
On The Landscape of Spoken Language Models: A Comprehensive Survey Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 069f554c-bd79-4cfd-a636-58f85cf251ce · inbound
Kimi-Audio Technical Report Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation a68178db-90bf-4a1b-a8a9-5ee404ebe3b1 · inbound
Sat2Sound: A Unified Framework for Zero-Shot Soundscape Mapping Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 8b25ced2-3437-4776-af5b-64709d4dc2ab · inbound
LLaDA-V: Large Language Diffusion Models with Visual Instruction Tuning Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 2f947a66-9d89-4f63-8508-05c68e9a5cdb · inbound
VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation bb4ccade-bec5-4864-aff7-30c54bb5815c · inbound
Orak: A Foundational Benchmark for Training and Evaluating LLM Agents on Diverse Video Games Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 93
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation a85c5273-a4ad-4bb3-83da-cfd9fe8cb236 · inbound
Step-Audio 2 Technical Report Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 5ba525c0-e879-4e1c-ac0c-dfc514a71efe · inbound
MECAT: A Multi-Experts Constructed Benchmark for Fine-Grained Audio Understanding Tasks Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 9459bfa7-63e8-4767-b177-81c377a0642e · inbound
Enhancing Speech Large Language Models through Reinforced Behavior Alignment Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation dd84d51e-7157-4d6f-a6f6-17f81ad08e06 · inbound
Direct Simultaneous Translation Activation for Large Audio-Language Models Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation a15dbc66-93b9-4f32-825e-c029cb124bea · inbound
GraphMend: Code Transformations for Fixing Graph Breaks in PyTorch 2 Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation f87db4fa-bc05-4d2f-8cab-d25c11610b5e · inbound
Qwen3-Omni Technical Report Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation e78fa995-c65f-4887-8e7c-cdeffa3ff2e6 · inbound
Investigating Modality Contribution in Audio LLMs for Music Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation a0e35937-98ce-43e6-bdbc-85b965b995a4 · inbound
End-to-end Contrastive Language-Speech Pretraining Model For Long-form Spoken Question Answering Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 5a4f7985-4e58-499b-94de-6d9b360f85cf · inbound
MCAT: Scaling Many-to-Many Speech-to-Text Translation with MLLMs to 70 Languages Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation b2fca029-712f-4cfe-9d69-cc6c5f6b7842 · inbound
Dynamic Content Moderation in Livestreams: Combining Supervised Classification with MLLM-Boosted Similarity Matching Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 938d5d20-137b-438b-92bb-7c6fc392fde4 · inbound
MOSS Transcribe Diarize Technical Report Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 82d211a8-aba7-478b-aa7b-03ef17f646a1 · inbound
FastSLM: Hierarchical Temporal Abstraction for Efficient Long-Form Speech Adaptation Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b4e00556-f75b-4533-98b7-ea35e4ccfbd6 · inbound
AQUA-Bench: Beyond Finding Answers to Knowing When There Are None in Audio Question Answering Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 915656ab-cbcd-4b88-b8c6-f189a0119684 · inbound
The Alignment Curse: Modality Alignment Supercharges Audio Attacks via Text Transfer Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 736bbc75-f1be-4855-a73d-8c2fb4eef428 · inbound
RA-QA: A Benchmarking System for Respiratory Audio Question Answering Under Real-World Heterogeneity Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4eaee3f0-f2e7-4f2a-89ec-0f84c995de76 · inbound
TW-Sound580K: A Regional Audio-Text Dataset with Verification-Guided Curation for Localized Audio-Language Modeling Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation dec17be9-5a17-4f83-abe7-6e3b2f56bacb · inbound
Language-Aware Distillation for Multilingual Instruction-Following Speech LLMs with ASR-Only Supervision Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c879398c-5410-45c8-b011-e0edb2ffcbaa · inbound
FoleyDesigner: Immersive Stereo Foley Generation with Precise Spatio-Temporal Alignment for Film Clips Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 290cb93c-7d3d-4bcb-bfbd-7e3b0930fc6b · inbound
Interactive ASR: Towards Human-Like Interaction and Semantic Coherence Evaluation for Agentic Speech Recognition Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation e9a965b2-0d0a-4ef0-9396-243c375e80d4 · inbound
Whisper-AuT: Domain-Adapted Audio Encoder for Efficient Audio-LLM Training Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation bd9ea5f8-f59b-491e-ad1c-ff72371925c7 · inbound
Speaker Attributed Automatic Speech Recognition Using Speech Aware LLMS Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation c749619a-078b-4728-8ab2-c3c25b7369e6 · inbound
HumDial-EIBench: A Human-Recorded Multi-Turn Emotional Intelligence Benchmark for Audio Language Models Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation c3034d23-7de3-4507-864f-79a5b5ad6212 · inbound
SpotSound: Enhancing Large Audio-Language Models with Fine-Grained Temporal Grounding Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation b52494af-c2e3-4ba0-853e-437c93d78287 · inbound
Towards Fine-grained Temporal Perception: Post-Training Large Audio-Language Models with Audio-Side Time Prompt Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 8379af5e-6e66-4f3f-ad8f-1aa9fea82c16 · inbound
Character Beyond Speech: Leveraging Role-Playing Evaluation in Audio Large Language Models via Reinforcement Learning Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 7547d7e1-6daa-43f2-8e63-084d53cb4443 · inbound
A Synonymous Variational Perspective on the Rate-Distortion-Perception Tradeoff Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5b6f4c32-1675-40bd-987f-f73297d3e4fa · inbound
Hijacking Large Audio-Language Models via Context-Agnostic and Imperceptible Auditory Prompt Injection Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 70ea96b3-229c-4098-bf72-f96e84c1f355 · inbound
Qwen3.5-Omni Technical Report Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation d0052591-51b9-4b3f-9b14-905b6495f83f · inbound
TinyMU: A Compact Audio-Language Model for Music Understanding Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 95e3d43a-fc71-4bc5-87b6-3b5faa34f5a9 · inbound
Omni-Embed-Audio: Leveraging Multimodal LLMs for Robust Audio-Text Retrieval Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 2edfa591-0bad-47f3-b02a-3ae6c8c974dc · inbound
HalluAudio: A Comprehensive Benchmark for Hallucination Detection in Large Audio-Language Models Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation e5ad04c3-579a-43ea-aa50-76833e525b9c · inbound
Indic-CodecFake meets SATYAM: Towards Detecting Neural Audio Codec Synthesized Speech Deepfakes in Indic Languages Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 193
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation fda59e30-64c1-4c59-bba0-294afcb332f5 · inbound
ONOTE: Benchmarking Omnimodal Notation Processing for Expert-level Music Intelligence Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 2c5cbd56-3b87-4777-9b2b-6e83b530d87e · inbound
In-Sync: Adaptation of Speech Aware Large Language Models for ASR with Word Level Timestamp Predictions Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation abfbf2bf-15ea-4e16-bd20-bb0be9f09530 · inbound
When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation bdaf2b62-393c-44b1-885c-7505b775a31c · inbound
AffectGPT-RL: Revealing Roles of Reinforcement Learning in Open-Vocabulary Emotion Recognition Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 9e68c7f7-ce71-4e29-b9ae-f44abb081e5c · inbound
VITA-QinYu: Expressive Spoken Language Model for Role-Playing and Singing Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 5729abf1-bc37-4611-9f1e-e94c51ac38bf · inbound
MIST: Multimodal Interactive Speech-based Tool-calling Conversational Assistants for Smart Homes Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation f5cc76f9-0d30-4fda-a649-90c29cfca4d9 · inbound
Polyphonia: Zero-Shot Timbre Transfer in Polyphonic Music with Acoustic-Informed Attention Calibration Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 79
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 620ffce7-c856-4ca7-9bb0-3c8f10fadcb3 · inbound
NAACA: Training-Free NeuroAuditory Attentive Cognitive Architecture with Oscillatory Working Memory for Salience-Driven Attention Gating Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation cdc99437-c98d-415e-b494-73f7e9b21483 · inbound
SpeakerLLM: A Speaker-Specialized Audio-LLM for Speaker Understanding and Verification Reasoning Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation df4a1cc0-e0b4-4445-bf82-31cf959c3e95 · inbound
Safety Geometry Collapse in Multimodal LLMs and Adaptive Drift Correction Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 1e53ada0-156c-4bd4-97c7-dbbb9a207388 · inbound
Beyond the Cartesian Illusion: Testing Two-Stage Multi-Modal Theory of Mind under Perceptual Bottlenecks Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 9e8fa346-5ab2-42b9-a605-1bb28993f172 · inbound
Heterogeneity-Aware Dataset Scheduling for Efficient Audio Large Language Model Training Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 4f00b568-ccd7-42fa-9273-d5de3b870bea · inbound
AffectVerse: Emotional World Models for Multimodal Affective Computing Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation eb3712d9-5bb2-4d85-8236-08592f036432 · inbound
A Survey of Large Audio Language Models: Generalization, Trustworthiness, and Outlook Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation d0601d4d-da6c-4519-8b2d-1c2ceed0618d · inbound
Codec-Robust Attacks on Audio LLMs Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 479df41f-2545-4208-89ef-c1e4ca2aa2a3 · inbound
Codec-Robust Attacks on Audio LLMs Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 7f4ccaf0-1710-4abd-8708-534d46d21743 · inbound
Academic Text-to-Music Grand Challenge: Datasets, Baselines, and Evaluation Methods Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 5c25e27e-88b1-45a4-872a-19bd9b4dd46c · inbound
Academic Text-to-Music Grand Challenge: Datasets, Baselines, and Evaluation Methods Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 740e0a2d-bb52-478b-b45b-c7f7d49cb0e8 · inbound
Toward Native Multimodal Modeling: A Roadmap Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation da3ac5f8-e845-4de2-bfd5-97485eec1690 · inbound
Learning When to Think While Listening in Large Audio-Language Models Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 0bc2763b-f789-42ea-9194-ebae95b4c9f7 · inbound
Bandwidth-Efficient and Privacy-Preserving Edge-Cloud Many-to-Many Speech Translation Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 9d928bee-1a71-4611-b74a-b293c28cb711 · inbound
Decoding Strategies for Diffusion-Based ASR: A Systematic Evaluation of Confidence-Based Thresholding Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation aefe8a3d-a213-4bab-a83d-2b41b781e83d · inbound
UNISON: A Unified Sound Generation and Editing Framework via Deep LLM Fusion Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 179085e3-992f-4d2a-ae0a-f132a28e8a89 · inbound
SALSA: Speech Aware LLM Adaptation via Learned Steering Activation Vectors Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation f4ca3108-7a9c-4749-b2ad-8a07fd214955 · inbound
MOSS-Audio Technical Report Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 8849e7a8-7d9b-4497-bb3f-0ad7da2baa51 · inbound
SpeakerCard-1M: An Evidence-Grounded Corpus for In-the-Wild Speaker Verification Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 723b62a9-8d87-4d49-b458-0387ac1675a5 · inbound
SpeakerCard-1M: An Evidence-Grounded Corpus for In-the-Wild Speaker Verification Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation e2bb7787-a114-40f1-8466-be63fe23eef8 · inbound
Entity Binding Failures in Speech LLM Reasoning: Diagnosis and Chain-of-Thought Intervention Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation eb536c04-8d0a-4304-9e23-2ec0aa406ed2 · inbound
UAT: Unified Audio-Text Diffusion for Audio Generation, Editing, and Captioning Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation a5855a8f-e10c-4cb7-bb7c-fab514158adb · inbound
Audio Interaction Model Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 932b0ab3-6afd-45de-b862-4d56d7a095e1 · inbound
Beyond Semantic Dominance: Cognitive Affective Reasoning and Empathetic Response Alignment in Audio Language Models Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 3176a1fb-2ea5-4502-bd1c-55e08d7d8913 · inbound
Making the Most of Limited Data: Score-Aware Training for Text-to-Music Generation Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 2eae3c81-fd6b-4d33-881c-12a91f9888d2 · inbound
Titans-as-a-Layer: Test-Time Memory for Conversational Speech Emotion Recognition Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 4ba04366-f80d-42f3-b828-93d0309bccac · inbound
Is Text All You Need? Text as a Universal Information Bottleneck for Speech LLMs Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation c241abff-92c0-4a8e-9fa7-a78e8cf487c2 · inbound
A Finetuned SpeechLLM for Joint Multi-Granular L2 Assessment and Natural-Language Rationales Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation e4925622-4525-4f07-9bf9-2d1dc972c69e · inbound
Speech Encoder Fusion for LLM-based Automatic Speech Recognition Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation b6ca04b8-b71d-4992-bd17-b1f28879ad73 · inbound
Afrispeech Semantics: Evaluating Audio Semantic Reasoning in Spoken Language Models Across Domains and Accents Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 87
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 9452a331-52ee-4638-9edb-3d048f2d355a · inbound
RAIL: Rethinking Auditory Intelligence in Large Audio-Language Models with a CHC-Grounded Benchmark Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation f1b06a76-9818-43c9-900a-f37b5ca039e7 · inbound
DeceptionX: From Multimodal Evidence to Explainable Deception Detection Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 9b7135e4-56ed-4edb-95a9-bb132db28181 · inbound
DeceptionX: From Multimodal Evidence to Explainable Deception Detection Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ac7ff34f-1db1-4410-b224-25bed8213b59 · inbound
Continuous Audio Thinking for Large Audio Language Models Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 619b0362-c950-408b-be0b-f8183fec4b63 · inbound
ESPnet3: Infrastructure for Scalable Speech and Audio Research in the Foundation Model Era Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 0b2d104a-d72a-4026-8a0a-2d0483a7cfd6 · inbound
Uncertainty-based Debiasing and Unlearning for Decontamination Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 5ea7e87f-1ab4-41c0-9425-0c9ebd3b1236 · inbound
Omni-Perception Policy Optimization for Multimodal Emotion Reasoning Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 141
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 4fb3ee76-5ee9-4a24-9bed-c477ba7f6356 · inbound
Evaluating Japanese Dialect Robustness Across Speech and Text-based Large Language Models Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 4d7d6bd8-907e-4ffd-8ad1-d78f8f3a39b2 · inbound
Does Translation-Enhanced Speech Encoder Pre-training Affect Speech LLMs? Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 9df4c430-c924-4e36-9570-2ed681fabf5e · inbound
wav2tok 2.0: Scalable Audio Tokenization Maintaining Explicit Pairwise Token Alignment for Efficient Audio Retrieval Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 83072bef-a268-4225-a828-3e1b98b061ed · inbound
MER-R1: Multimodal Emotion Reasoning via Slow-Fast Thinking Synergy Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation a1f64d49-bfc4-47c2-8d5c-48b413637a29 · inbound
How to Leverage Synthetic Speech for LLM-Based ASR Systems? Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation 89de6386-c342-448a-951d-515551432d5f · inbound
How to Leverage Synthetic Speech for LLM-Based ASR Systems? Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0891cb16-4399-4f0e-b866-501bbb99ede7 · inbound
Preference-ASR: A Preference-Aware Test Set for Benchmarking ASR in the Era of Speech LLMs Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation cc259b6a-0406-45d8-868c-c62268c37d7d · inbound
Beyond Binary Instrument QA: Probing Instrument Grounding in Music Audio-Language Models Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.
Observation b8215101-0e5e-467c-8a2a-491459e6a730 · inbound
CaReCoS: A Spectrogram based Visual Benchmark for Cardiac, Respiratory and Cough Sounds Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2f60d2a8-ff7e-4bd3-a57e-289062cf1888 · inbound
Auto-AEG: Scalable Data Construction for Open-Vocabulary Audio Event Grounding Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ad3f1393-442d-4c5a-b34a-89707874ac6f · inbound
Auto-AEG: Scalable Data Construction for Open-Vocabulary Audio Event Grounding Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2c0c3017-8736-467a-85ae-40de1a7a8f12 · inbound
Context-Aware ASR for Mandarin Technical Lectures Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 71b2450c-44af-4992-b2b3-d8adbe4b8090 · inbound
Unified Audio Intelligence Without Regressing on Text Intelligence Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 136
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.