Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-06-28T14:36:53.295540Z
Paper Citation Record · LEDGER
As of 6 August 2026, this Paper Citation Record lists 85 of 85 outbound references and 0 inbound Pith citation observations for arXiv:2606.07643.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-06-28T14:36:53.295540Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
85 of 85 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 7eab8998-10dd-4032-bfff-a4ef379e2b8b · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs 2024 , eprint=
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 77902575-2ec8-45be-a020-5b7d0116048f · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs GPT-4o System Card
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 6eca5690-5363-4edb-b11f-58962318387d · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 512b2825-40ee-47e2-8b50-3a8e3177d431 · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs 2023 , eprint=
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6ddadbb4-f963-48a1-88e7-609000220dfe · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs Llama 2: Open Foundation and Fine-Tuned Chat Models
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 259a1f8d-3c64-4e9b-ab5d-6a18e128e250 · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs The Llama 3 Herd of Models
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 0373cec7-b233-4249-afd1-d8bb5fb0423b · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs 2023 , eprint=
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5ff05621-1995-4090-a9ae-a8381507d4c7 · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs Qwen2.5 Technical Report
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 124a9a04-d6ff-41b8-a25c-4938003809d6 · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 1c9f38ba-f619-4fe0-b813-dcd4f1b0e1ab · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 03cd71e0-4698-4c2e-a126-47c12f83ff40 · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs Advances in neural information processing systems , volume=
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d9a7c800-734b-4fd5-9482-44ef9b258c56 · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs 2023 , url =
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6b132ff5-61e0-406d-9c0f-cdad268a7692 · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs Advances in Neural Information Processing Systems , volume=
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f73eee58-a95f-4b3b-8712-3da237ae403b · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs Kimi-Audio Technical Report
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 69f2883e-3ec2-4f81-9b62-78aa3388b00d · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs Qwen2-Audio Technical Report
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 9fed51af-1648-4b02-b58c-48c1fbe4630b · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs SpeechGPT: Empowering Large Language Models with Intrinsic Cross-Modal Conversational Abilities
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation f760dea0-d8a1-4938-afc0-fb811b8ab46b · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs PandaGPT: One Model To Instruction-Follow Them All
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation ca55d9fa-6838-47ad-8f95-b35888bceb2a · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs Advances in Neural Information Processing Systems , volume=
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3933720a-cdc7-4355-9816-4cfd579d92fd · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs Forty-first International Conference on Machine Learning , year=
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 01482320-8175-4f71-ae28-60370b3353d9 · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs AnyGPT: Unified Multimodal LLM with Discrete Sequence Modeling
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation ada64e49-89e5-4159-b6be-efc9be3dc9cb · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs Video-LLaMA: An Instruction-tuned Audio-Visual Language Model for Video Understanding
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 03689645-6d08-4faa-b24a-2f705339135f · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs 2025 , eprint=
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 42dd175d-8b13-42fd-a150-27dbe2c49a27 · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs AVHBench: A Cross-Modal Hallucination Benchmark for Audio-Visual Large Language Models
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation c8ba0510-7d73-4c6c-9fb1-34ac20f19080 · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs 2025 , eprint=
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3c5484df-8921-4bd1-b9eb-c1a35d82b311 · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs 2023 , eprint=
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 22da7afc-03dd-4598-a03b-a8c4a95d83f2 · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs 2024 , eprint=
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2b0317bc-baa6-490e-815f-2d0073361aed · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs HumanOmni: A Large Vision-Speech Language Model for Human-Centric Video Understanding
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 8b9b6cb0-7057-4580-8937-d4b939526d4d · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs Baichuan-omni technical report
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 43a93783-f378-4293-8374-2d54b854bef6 · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs Qwen2.5-Omni Technical Report
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation d71c8850-c65d-4c82-a6fe-d097db8bfe65 · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs 2024 , eprint=
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2cff3767-114a-4b3a-81b2-a27fb5e2465c · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs 2025 , eprint=
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6b7387cf-e11a-4262-87a7-3310372b1b71 · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs OmnixR: Evaluating Omni-modality Language Models on Reasoning across Modalities
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation f45cb49a-d6f0-4171-92d7-bd7672cb6225 · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs In: Proceedings of the 30th ACM International Conference on Multimedia
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 3efb50be-37aa-4cdf-bf77-c48eb6fc0e70 · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs IEEE Conference on Computer Vision and Pattern Recognition (CVPR) , year =
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6f764bd1-c41f-4b79-b7b9-0cde6e78e852 · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs WorldQA: Multimodal World Knowledge in Videos through Long-Chain Reasoning
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 73b812e0-94b1-48af-aa80-06f250226bb2 · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs Multimodal Chain-of-Thought Reasoning: A Comprehensive Survey
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation fbc68ea5-6d45-4c17-b63c-ed485cd01f64 · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs Journal of Artificial General Intelligence , volume=
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1c55f1c5-60de-4a54-8364-479fe1057f1d · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs Nature Communications , volume=
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e5fbfe49-4ac7-4f20-aca4-0629f563089f · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs 2023 , publisher=
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f11e1a63-ec2f-4de3-8af6-49b295106a5a · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs 2007 , publisher=
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f248a72a-a0e0-43d1-8920-a65a8acbbbc7 · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs Advances in neural information processing systems , volume=
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1684874b-3080-44a1-b9e1-4cefdb1db4c3 · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs Forty-first International Conference on Machine Learning , year=
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 405639c0-0048-44e9-a116-5cdd7a8613b7 · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs Training Verifiers to Solve Math Word Problems
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation b5116ad0-73cf-4d52-80f8-dd8d3cbf9026 · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs Evaluating Large Language Models Trained on Code
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 711dfbc5-a0cb-4863-a6cf-f4bc91566c03 · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs Advances in neural information processing systems , volume=
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8a77d57d-3757-46c9-a35a-d6a7eebef130 · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs Proceedings of the European conference on computer vision (ECCV) , pages=
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 358d3fbc-86f2-4537-ba00-a055eff396f1 · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 523bea54-7155-469c-9270-28c7a33d5377 · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs European Conference on Computer Vision , pages=
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 285e210d-0a59-48f6-983a-e86204f4382d · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs European Conference on Computer Vision , pages=
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8ec5c04d-9b69-4ecf-b772-9de5fe61245c · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs Nature reviews neuroscience , volume=
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8421f70d-ce34-4a31-a770-557c3de079e5 · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs Unresolved cited work
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 95cedbdb-05c3-4a06-ac54-ae88dc719f89 · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs Unresolved cited work
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 306d2ccf-2767-4b63-b476-2b7cb3111e3a · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs Annual review of vision science , volume=
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6bc095df-d393-4fff-bd6d-09a67dc7b56d · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs Current Biology , volume=
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d1430315-2e72-409f-aaf1-7cc05310b29b · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs Neuropsychopharmacology , volume=
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c8b9f8af-727b-4968-881a-cdfe97acfe78 · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs IEEE Open Journal of Signal Processing , year=
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ead9513f-0f4b-4682-8143-5ea973647017 · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs IEEE Transactions on Pattern Analysis and Machine Intelligence , year=
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 65eca983-78dc-43d8-af42-ee4c9ea5a895 · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs Ola: Pushing the Frontiers of Omni-Modal Language Model
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 8a4f4a3c-7a4c-44aa-ab98-d71ba68d7101 · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs Baichuan-omni-1.5 technical report
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 5553b942-08b1-4b22-a187-1ccbe864fa2a · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs R1-Omni: Explainable Omni-Multimodal Emotion Recognition with Reinforcement Learning
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation af222d8e-548a-4726-bb9f-ee3dffeb813a · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs Phi-4-Mini Technical Report: Compact yet Powerful Multimodal Language Models via Mixture-of-LoRAs
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation f137f6d5-c43a-4677-9922-2d3efa2a795b · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs Perception, Reason, Think, and Plan: A Survey on Large Multimodal Reasoning Models
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 43432d11-d636-49d1-ba2b-56cb6e0df3fb · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs On Path to Multimodal Generalist: General-Level and General-Bench
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation e4c6475b-f81d-4af8-a2e0-5d8304dea673 · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , pages=
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7c1f08db-ba8b-464e-bc9a-54c54fd8d491 · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs SEED-Bench: Benchmarking Multimodal LLMs with Generative Comprehension
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 28df9847-0e7c-41ca-b84d-889692bb0d2e · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs MMAU: A Massive Multi-Task Audio Understanding and Reasoning Benchmark
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 36603af1-061d-4692-b661-d9cc357fc349 · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs European Conference on Computer Vision , pages=
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 73b3c104-6f09-47f0-87dc-56088d024a10 · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs European Conference on Computer Vision , pages=
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7437bf96-f31c-41bf-a731-7a019c1d5e2e · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs IEEE Access , volume=
Reference 69
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 86ea63b5-35e0-4eeb-a983-593355edac8b · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs Advances in neural information processing systems , volume=
Reference 70
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 130d376f-fb46-4d91-be11-cd66874094dc · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs International journal of Remote sensing , volume=
Reference 71
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bb3c1aa5-fba1-47b9-ade8-14da3e146d58 · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , pages=
Reference 72
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6f732fe6-8efe-4257-a26e-1e82583ca98a · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs Proceedings of the IEEE/CVF International Conference on Computer Vision , pages=
Reference 73
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d968893a-e41c-4071-98a4-1e07c77b1b77 · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs ACM Transactions on Multimedia Computing, Communications and Applications , volume=
Reference 74
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cc820081-9fd9-4594-9cd2-9174f3a85055 · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs Proceedings of the IEEE/CVF International Conference on Computer Vision , pages=
Reference 75
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0f8aee72-c88f-4de0-b55b-5b06dbcc371f · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs 2024 , isbn =
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 49e5515e-8e9d-4fa0-8b03-4f9ab08d8811 · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs European Conference on Computer Vision , pages=
Reference 77
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d3b95ef4-a88e-4e68-9463-6be83c61a2a2 · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs ChatGLM: A Family of Large Language Models from GLM-130B to GLM-4 All Tools
Reference 78
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation a7d7f17a-5f35-478f-9cd9-388ff19e134c · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs Do Audio-Visual Segmentation Models Truly Segment Sounding Objects?
Reference 79
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation f129b174-1b38-463b-a386-c40262bf78de · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs Tarsier: Recipes for Training and Evaluating Large Video Description Models
Reference 80
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 51799bc7-832b-4828-a5f4-fb065bbcfb93 · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs Proceedings of the Computer Vision and Pattern Recognition Conference , pages=
Reference 81
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 95db0db5-b3ce-4464-b784-cd1cbc29a304 · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs WorldSense: Evaluating Real-world Omnimodal Understanding for Multimodal LLMs
Reference 82
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation 477e8220-4d10-41cd-8307-17383bc18592 · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs Mlvu: Benchmarking multi-task long video understanding
Reference 83
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation c7fc47f5-d11c-4520-8194-9a3234e73e52 · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs Video-Holmes: Can MLLM Think Like Holmes for Complex Video Reasoning?
Reference 84
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.
Observation c8fbf330-f926-438c-89ce-90c093d3762d · outbound
AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs Unresolved cited work
Reference 85
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.