Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-02T05:20:03.868153Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 70 of 70 outbound references and 1 inbound Pith citation observation for arXiv:2607.13408.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-02T05:20:03.868153Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-02T05:19:55.932658Z
A source-named dated measurement, never combined with another source.
Source: cited_works
70 of 70 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 897b7fd3-7c2c-496e-abb6-b3bc3fcc7be8 · outbound
Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 93faeea3-4eac-431b-a8db-3976a2457343 · outbound
Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models Unresolved cited work
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c552ab95-18f0-40c9-bd4d-1a1f569710d2 · outbound
Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models Rain is falling con- tinuously
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7bed7c70-e320-44be-81fc-a962e42fcab3 · outbound
Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models It starts with the sound ofe 1, shifts toe 2, and ends withe 3
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 261244ee-e32a-471f-851a-887ae4373772 · outbound
Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models Verifying ALLMs as Judges Table 1 reports ALLM performance on audio understanding benchmarks with ground-truth annotations
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8d3e11d5-03f1-4d28-bbe9-0b3d50843993 · outbound
Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models While recent systems achieve strong percep- tual realism, our study shows that they often overlook fine- grained requirements such as sound event completeness and temporal ordering
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 73364c19-88fc-4a25-ab64-3abf510cc4e2 · outbound
Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models Unresolved cited work
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 83e62017-cf58-492c-8725-d771dc9ac622 · outbound
Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models AudioGen: Textually Guided Audio Generation
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3b8dd948-6f65-466b-a3cc-46a277751328 · outbound
Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models MusicLM: Generating Music From Text
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cd851abf-2a24-4ff8-b554-cb7618658e91 · outbound
Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models AudioLDM: Text-to-audio generation with latent diffusion models,
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cd8dac01-cc08-4c03-b8de-40e5cddd78c1 · outbound
Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models Text-to-Audio Generation using Instruction-Tuned LLM and Latent Diffusion Model
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 348e50f9-7e8a-4b55-b7c6-6fe9059b08ca · outbound
Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models Diffsound: Discrete diffusion model for text-to-sound genera- tion,
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5c55add6-096b-489f-90d4-2bc60330662f · outbound
Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models Towards general-purpose text-instruction-guided voice conversion,
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7c05fc79-d7d6-4cb5-999f-5c2659632ce1 · outbound
Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models Simple and controllable music generation,
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7ad275b0-2ef7-460e-b96a-95909d3f937f · outbound
Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models Audiolm: a language modeling approach to audio gener- ation,
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c3acb87d-c10d-4f9a-8ab0-17fa3de23416 · outbound
Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models SoundStorm: Efficient Parallel Audio Generation
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 99729a8c-4d66-420f-867e-6f3bdbb85b88 · outbound
Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models Make-An-Audio 2: Temporal-Enhanced Text-to-Audio Generation
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f1542adc-c504-493e-b5cc-0d8a9df23b56 · outbound
Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models Tango 2: Aligning diffusion-based text-to-audio generations through direct preference optimization,
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation af68f29d-04c2-4789-a505-a562ef0ef7b0 · outbound
Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models Tangoflux: Super fast and faithful text to audio generation with flow matching and clap- ranked preference optimization,
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c7753fe2-0ae1-45b0-a65f-29f274af05eb · outbound
Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models EzAudio: Enhancing Text-to-Audio Generation with Efficient Diffusion Transformer
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 38c0a192-db20-4807-8219-d48127efaa21 · outbound
Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models Audioldm 2: Learn- ing holistic audio generation with self-supervised pretraining,
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c69f1c51-9f5b-4103-901d-165ab4231797 · outbound
Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models Stable audio open,
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bf2e3ebf-2a9f-4891-995d-e67f26ee13e9 · outbound
Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models Etta: Elucidating the design space of text-to-audio models,
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 137f8c6c-2bdb-45d0-a2e6-47e2e1cae64a · outbound
Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models Impact: Iter- ative mask-based parallel decoding for text-to-audio generation with diffusion modeling,
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bf6b0368-4ac5-4de3-9de7-09775fd8cde0 · outbound
Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models Gen- erative audio language modeling with continuous-valued tokens and masked next-token prediction,
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation faf29cef-d559-44e2-a290-673d1bece393 · outbound
Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models Fr ´echet audio distance: A reference-free metric for evaluating music enhancement algorithms,
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f28e2bb0-5c2d-47e1-800d-a3f9a2f950b1 · outbound
Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models Improved techniques for training gans,
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 68fafd2c-9de4-478b-8e25-965cce8f216a · outbound
Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models Clap learning audio concepts from natural language supervision,
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 823b4a65-1815-4089-bea7-f33063c84884 · outbound
Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models Large-scale contrastive language-audio pretraining with feature fusion and keyword-to-caption augmentation,
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cebd6c1b-4a32-480c-92df-1ae59215ff7e · outbound
Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models Ritta: Model- ing event relations in text-to-audio generation,
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e1032269-da52-42aa-bf86-def53b563435 · outbound
Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models Aurelius: Relation aware text-to-audio generation at scale,
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 902de066-9176-4336-8df2-c8a8c476b0ca · outbound
Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models Compa: Address- ing the gap in compositional reasoning in audio-language mod- els,
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f9f1773e-935e-4788-a6ef-ea24d36f11a5 · outbound
Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models T2a-feedback: Improving basic capabilities of text-to-audio generation via fine-grained ai feed- back,
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 90a1d0a4-25db-4132-a3a1-62995b35a5b4 · outbound
Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models Aqascore: Evaluating semantic alignment in text-to-audio generation via audio question answering,
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c044ba76-7e09-45c9-997c-dfaf3b24b02c · outbound
Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models The Sound of Absence: Audio-Language Embedding Models Struggle with Negation
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ea01226f-cdac-4df2-a716-7a4c648b1874 · outbound
Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models GPT-4 Technical Report
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 858fae87-adb7-410e-9d45-36588d562764 · outbound
Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models Joint audio and speech understanding,
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0f35bbe4-350f-4beb-bfd6-a448f097bb72 · outbound
Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models BLSP: Bootstrapping Language-Speech Pre-training via Behavior Alignment of Continuation Writing
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eeb694ec-2354-4fac-a910-630a9669a28b · outbound
Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models Audiochatllama: Towards general-purpose speech abilities for llms,
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 404f282e-2fdc-4d66-a562-0a1f6a605364 · outbound
Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models Speech-copilot: Leveraging large language models for speech processing via task decomposition, modularization, and program generation,
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b6d54ccc-0e7d-4188-864e-5c65a1e75323 · outbound
Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models Speechprompt: Prompting speech language models for speech processing tasks,
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2b6a02bc-8860-4fbe-b2d4-40658e22f1b0 · outbound
Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models Blsp- emo: Towards empathetic large speech-language models,
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 63214a6e-6535-42e2-aec1-579dd8c11ab5 · outbound
Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models Audio flamingo 3: Advancing audio intelligence with fully open large audio language models,
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 608bbfa2-db29-48f0-b7a6-465c09ffa510 · outbound
Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models Voxtral
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 24cfd24c-19b1-49d1-98a5-cd9b15b8f084 · outbound
Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models Qwen2.5-Omni Technical Report
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 643cea81-7953-4c92-b12c-91f3407746b5 · outbound
Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models Teaching audio-aware large language models what does not hear: Mitigating hallucinations through synthesized negative samples,
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4db4dd76-378d-41cc-9a06-1be24eb3645d · outbound
Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models From alignment to advancement: Bootstrapping audio- language alignment with synthetic data,
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c0c038e6-b470-4114-998a-f34b9a784af4 · outbound
Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models Phi-4-Mini Technical Report: Compact yet Powerful Multimodal Language Models via Mixture-of-LoRAs
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 51859531-fef3-4eb6-a497-416d668ceead · outbound
Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 82966f7a-01b6-4c22-b376-bebc7cb45090 · outbound
Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models On the landscape of spoken language models: A comprehensive survey,
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1c2430b7-ebd3-4959-89b8-d6be49899dd5 · outbound
Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models Audio flamingo 2: An audio- language model with long-audio understanding and expert reason- ing abilities,
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 437b6dca-2413-4b2e-a766-918f3ce965c9 · outbound
Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models Understanding sounds, missing the questions: The challenge of object hallucination in large audio-language models,
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d21e50ec-7b39-44ad-bc3d-c16394ff7da3 · outbound
Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models Can large audio-language models truly hear? tackling hallucinations with multi-task assessment and stepwise audio reasoning,
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 22d14206-0d01-4eab-9f4a-3efedc4a53f9 · outbound
Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models Dynamic-superb: Towards a dynamic, col- laborative, and comprehensive instruction-tuning benchmark for speech,
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8f6b06dc-021c-424d-b487-6393d494879e · outbound
Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models Dynamic-superb phase-2: A collaboratively expanding benchmark for measuring the capabilities of spoken language models with 180 tasks,
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bc7f16a0-4878-48a4-95f3-1e4acbb3d1cb · outbound
Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models Mmau: A mas- sive multi-task audio understanding and reasoning benchmark,
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6725413c-ccf2-44a3-92cf-440a8c763c73 · outbound
Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models MMAU-Pro: A Challenging and Comprehensive Benchmark for Holistic Evaluation of Audio General Intelligence
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 845f6050-4f64-412b-ab73-8ac9afaa3b9a · outbound
Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models MMAR: A Challenging Benchmark for Deep Reasoning in Speech, Audio, Music, and Their Mix
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bbe15c22-57ad-49e3-922d-0172e2fe2e30 · outbound
Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models Game-time: Evaluating temporal dynamics in spoken language models,
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 96f2e11a-f2dc-4ef6-a2ca-3096fb59c02c · outbound
Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models Aqua-bench: Beyond finding answers to knowing when there are none in audio question answering,
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4a70f1b7-c7c8-4ec6-ad7f-eb66b453401b · outbound
Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models Baton: aligning text-to- audio model using human preference feedback,
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 11e27ef6-1daa-4dc8-89f9-9050d62389a4 · outbound
Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models Walking Through Uncertainty: An Empirical Study of Uncertainty Estimation for Audio-Aware Large Language Models
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bf460991-dbb7-491d-aefe-0c7c596b4ab9 · outbound
Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models Audio large language models can be descrip- tive speech quality evaluators,
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b988b51f-4411-49b7-b030-50c369e351cb · outbound
Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models AudioJudge: Understanding What Works in Large Audio Model Based Speech Evaluation
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fcfb353f-9297-47b6-a57a-556b13f36469 · outbound
Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models Audio-aware large language models as judges for speaking styles,
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2cc58875-3fa4-4603-9419-b2cabe24bced · outbound
Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models InstructTTSEval: Benchmarking Complex Natural-Language Instruction Following in Text-to-Speech Systems
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ddab289b-5042-4030-a525-02d0dc13abbd · outbound
Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models Audioeval: Automatic dual-perspective and multi-dimensional evaluation of text-to-audio-generation,
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0a5ecb1b-92f2-4fb1-8830-1452c7d0a9d3 · outbound
Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models Audiocaps: Generat- ing captions for audios in the wild,
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1c71ab8d-13c3-487b-afad-0214601ac569 · outbound
Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models Audiotime: A temporally- aligned audio-text benchmark dataset,
Reference 69
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fc05b91d-98ce-4a16-b5c4-cc609a2bd884 · outbound
Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models ESC: Dataset for Environmental Sound Classifi- cation,
Reference 70
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 897b7fd3-7c2c-496e-abb6-b3bc3fcc7be8 · inbound
Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models Improving Text-to-Audio Instruction Following via Fine-Grained Feedback from Audio-Aware Large Language Models
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.