Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-14T15:40:22.428261Z
Paper Citation Record · LEDGER
As of 17 August 2026, this Paper Citation Record lists 37 of 37 outbound references and 0 inbound Pith citation observations for arXiv:2608.13239.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-14T15:40:22.428261Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
37 of 37 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation fcb3e689-f5cc-4edf-b386-61120ec4cc92 · outbound
Reasoning for Social Audio-Visual Question Answering: Where Do We Stand? In: Proceedings of the 33rd ACM International Confer- ence on Multimedia
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation ebf1d7cf-2b22-48f8-97c5-989a2cedf960 · outbound
Reasoning for Social Audio-Visual Question Answering: Where Do We Stand? In: Belgrave, D., Zhang, C., Lin, H., Pascanu, R., Koniusz, P., Ghassemi, M., Chen, N
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation b8ad8321-5d4f-47d3-a567-5808432c5690 · outbound
Reasoning for Social Audio-Visual Question Answering: Where Do We Stand? Gemma 2: Improving Open Language Models at a Practical Size
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c167fe86-a646-46ec-8322-762a47c04df8 · outbound
Reasoning for Social Audio-Visual Question Answering: Where Do We Stand? Nature645(8081), 633–638 (2025).https://doi.org/10.1038/s41586-025-09422-z,http://dx
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 71ab9d94-9509-4867-b6f8-2b12e5127677 · outbound
Reasoning for Social Audio-Visual Question Answering: Where Do We Stand? pre- ferring shorter thinking chains for improved llm reasoning
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dce71d75-8e6a-4879-9609-24d2d8e8ea70 · outbound
Reasoning for Social Audio-Visual Question Answering: Where Do We Stand? In: The Fourteenth Inter- national Conference on Learning Representations (2026),https://openreview
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 1431cff5-861d-4542-97df-30a83cda322a · outbound
Reasoning for Social Audio-Visual Question Answering: Where Do We Stand? Unresolved cited work
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation b26da899-3e01-432e-bcb1-63c107e0ea75 · outbound
Reasoning for Social Audio-Visual Question Answering: Where Do We Stand? In: IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) (2026)
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 94effb2d-3ea3-45a5-bdb6-51465b73bae9 · outbound
Reasoning for Social Audio-Visual Question Answering: Where Do We Stand? In: Proceedings of the AAAI Conference on Artificial Intelligence
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 1bac5ce9-d9ce-4f25-8c19-c0abeef21e1d · outbound
Reasoning for Social Audio-Visual Question Answering: Where Do We Stand? Transactions on Machine Learning Re- search (TMLR) (2026) 16 K
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 94d9d133-0221-4577-b26d-72cd84f0b53a · outbound
Reasoning for Social Audio-Visual Question Answering: Where Do We Stand? In: The Thirty-ninth Annual Conference on Neural Information Processing Systems Datasets and Benchmarks Track (2026), https://openreview.net/forum?id=SSF4qgsNYE
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 331b1625-dbc1-4104-b370-44f3383ab879 · outbound
Reasoning for Social Audio-Visual Question Answering: Where Do We Stand? Unresolved cited work
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b76175d4-b330-4aa5-b3ed-76396b6e712d · outbound
Reasoning for Social Audio-Visual Question Answering: Where Do We Stand? Explainable Multimodal Emotion Recognition
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d2008d3f-f1e2-4300-8549-b3ef134667e3 · outbound
Reasoning for Social Audio-Visual Question Answering: Where Do We Stand? Advances in Neural Information Processing Systems36, 34892–34916 (2023)
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 0c1bdeb2-c3ea-4cfd-bde9-562187734be3 · outbound
Reasoning for Social Audio-Visual Question Answering: Where Do We Stand? The Llama 3 Herd of Models
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cc2c9d48-c62e-4373-9422-43e9d060b4b9 · outbound
Reasoning for Social Audio-Visual Question Answering: Where Do We Stand? International Conference on Learning Representations (ICLR) (2026)
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 48de3eb2-7ea5-465d-ad4e-fd162ee927a7 · outbound
Reasoning for Social Audio-Visual Question Answering: Where Do We Stand? MODF-SIR: A Multi-agent Omni-modal Distilled Framework for Social Intelligence Reasoning
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 5bf02b1c-97f5-4650-bfbd-5f8c224913fc · outbound
Reasoning for Social Audio-Visual Question Answering: Where Do We Stand? In: Proceedings of the 2025 Conference on Empirical Methods in Natural Language Processing
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 05fc6fc0-db95-4e07-b562-4ee30c328158 · outbound
Reasoning for Social Audio-Visual Question Answering: Where Do We Stand? Unresolved cited work
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation c46bf565-a18f-4c7b-8d69-717186081ca5 · outbound
Reasoning for Social Audio-Visual Question Answering: Where Do We Stand? Unresolved cited work
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 6ead7a3d-e8ed-448d-a535-028650533434 · outbound
Reasoning for Social Audio-Visual Question Answering: Where Do We Stand? In: Proceedings of the AAAI Conference on Artificial Intelligence
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation a5661164-765a-40f0-a528-c1ce20f39c39 · outbound
Reasoning for Social Audio-Visual Question Answering: Where Do We Stand? Qwen3.5-Omni Technical Report
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b2fdcf59-6add-4887-8a55-43043127a5c1 · outbound
Reasoning for Social Audio-Visual Question Answering: Where Do We Stand? Proceedings of the AAAI Conference on Arti- ficial Intelligence40(3), 2029–2037 (Mar 2026).https://doi.org/10.1609/aaai
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4de6f7d6-9041-438e-aa85-4b7b2131ab37 · outbound
Reasoning for Social Audio-Visual Question Answering: Where Do We Stand? DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ef2d15aa-22c3-406c-a6a3-99f58c4eed8f · outbound
Reasoning for Social Audio-Visual Question Answering: Where Do We Stand? IEEE Transactions on Affective Computing pp
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 864397c2-61c0-4e69-88c3-bd32fbf521c0 · outbound
Reasoning for Social Audio-Visual Question Answering: Where Do We Stand? NIPS ’22, Curran Associates Inc., Red Hook, NY, USA (2022) Reasoning for Social AV-QA: Where Do We Stand? 17
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 99d44657-f5ad-439d-8113-0272b25107f4 · outbound
Reasoning for Social Audio-Visual Question Answering: Where Do We Stand? Unresolved cited work
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 9a5be136-94c5-48f3-9271-805bdd1e91cc · outbound
Reasoning for Social Audio-Visual Question Answering: Where Do We Stand? In: The Fourteenth International Conference on Learning Representations (2026),https://openreview.net/forum?id=KttCXdjj4w
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 9dde63e2-6531-4509-8203-9b8cbd09f804 · outbound
Reasoning for Social Audio-Visual Question Answering: Where Do We Stand? Qwen2.5-Omni Technical Report
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 311bef0b-dcd4-4b30-829b-aeffeb57682e · outbound
Reasoning for Social Audio-Visual Question Answering: Where Do We Stand? HumanOmniV2: From Understanding to Omni-Modal Reasoning with Context
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 146a99af-8849-4886-a608-74b881604adc · outbound
Reasoning for Social Audio-Visual Question Answering: Where Do We Stand? In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation c4488e5e-0328-4d8b-b5fb-e61f1fcafda1 · outbound
Reasoning for Social Audio-Visual Question Answering: Where Do We Stand? In: The Fourteenth International Confer- ence on Learning Representations (2026),https://openreview.net/forum?id= xindJJLSr1
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 7595718e-6042-458e-b0b3-44aa283e9948 · outbound
Reasoning for Social Audio-Visual Question Answering: Where Do We Stand? In: Proceedings of the 2023 Conference on Empirical Methods in Natural Language Processing: System Demonstrations
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 1351b028-c51b-4d92-a936-9235dfbebfbb · outbound
Reasoning for Social Audio-Visual Question Answering: Where Do We Stand? R1-Omni: Explainable Omni-Multimodal Emotion Recognition with Reinforcement Learning
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3aecfed7-7764-4b78-9b86-e7a58187b973 · outbound
Reasoning for Social Audio-Visual Question Answering: Where Do We Stand? arXiv preprint arXiv:2512.09616 (2025)
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b47a2d16-4a8f-482e-ac4d-67e3d64bc694 · outbound
Reasoning for Social Audio-Visual Question Answering: Where Do We Stand? Reinforced MLLM: A Survey on RL-Based Reasoning in Multimodal Large Language Models
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d9427a58-448b-418a-82a9-a38044e6c743 · outbound
Reasoning for Social Audio-Visual Question Answering: Where Do We Stand? arXiv preprint arXiv:2505.17862 (2025) 18 K
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.