Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T21:18:59.867228Z
Paper Citation Record · LEDGER
As of 23 August 2026, this Paper Citation Record lists 100 of 113 outbound references and 45 inbound Pith citation observations for arXiv:2501.05444.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T21:18:59.867228Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-16T11:33:47.134304Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T13:09:50.233935Z
100 of 113 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 46b02530-cd70-4fc8-9b73-00e9ce62e98e · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark Greg Lan- druwu2024plot2codem, 8(31.10):5281, 2013
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 14fa4d6e-2547-4c55-a821-15c2d7a109f8 · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark Khan academy
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3d093895-4ab0-4d83-a0e0-bde7659ab4e2 · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark GPT-4 Technical Report
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3211d22e-a5e8-4f81-b9e5-b6479ac530e1 · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark VISREAS: Complex Visual Reasoning with Unanswerable Questions
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 076f321a-ad4a-425a-9f09-a46c3a3634a1 · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark Claude 3.5 sonnet
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 82627291-00dd-40e2-b146-0a320bfa6a21 · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark Vqa: Visual question answering
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f3cb3538-a059-4694-859f-d1f529e922b2 · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark Qwen Technical Report
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d8f4ba97-6a51-4967-a794-a984b2a9ef25 · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark Viseval: A benchmark for data visualization in the era of large language models
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 70951517-0338-4286-9c48-afd52e19ba1d · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark Uniter: Universal image-text representation learning
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e14bc7da-57d3-437f-930e-d95c1e01be9e · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ee5f873c-6dee-4731-8241-a898d9499189 · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark CoMT: A Novel Benchmark for Chain of Multi-modal Thought on Large Vision-Language Models
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eba7d77c-7753-4477-ac84-c127ba18b377 · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark On the Measure of Intelligence
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1a1c06c1-4593-4752-bc9f-0dbf6cca4292 · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark Training Verifiers to Solve Math Word Problems
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ba0b953c-972e-4b30-9c11-fa618e9f2640 · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark Codeforces
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c7e84cf4-f6d5-4005-b2e4-6fd89c07ae6e · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark EXAMS-V: A Multi-Discipline Multilingual Multimodal Exam Benchmark for Evaluating Vision Language Models
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6cd4b9f5-91e0-4d7b-90c5-96dd9a1cae71 · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark Gemini 2.0 flash thinking mode
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 90b3e9d7-14ba-4eae-a20e-36798e4c72fe · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark Introducing gemini 2.0: our new ai model for the agentic era
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aa282e4b-3422-47e8-bf01-f17ed1e52da6 · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark Deepseek-r1-lite-preview is now live: un- leashing supercharged reasoning power! https://api- docs.deepseek.com/news/news1120, 2024
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2565ee91-18a7-4971-9f5d-691e5bb64fa7 · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark The Llama 3 Herd of Models
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation babb6af9-8727-4a53-b713-e0b26a5d3279 · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark IsoBench: Benchmarking Multimodal Foundation Models on Isomorphic Representations
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 70a7997a-fe80-416a-bdb6-13e5673290e2 · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark The epistemology of visual thinking in mathematics, Feb 2020
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 37cda82d-31bc-42e2-942a-d063efc0c57c · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark Making the v in vqa matter: Elevating the role of image understanding in visual question answer- ing
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ddbd4355-7759-44ce-bd94-73a657a3a50a · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark ChartLlama: A Multimodal LLM for Chart Understanding and Generation
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2036de14-bdd4-4dff-ab7c-0b17df18037c · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark OlympiadBench: A Challenging Benchmark for Promoting AGI with Olympiad-Level Bilingual Multimodal Scientific Problems
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b3716284-a980-40f7-9cfd-662dd62fa70e · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark Measuring Massive Multitask Language Understanding
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8d5336ba-2e17-4a82-a8ec-3bc5f20d878a · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark Measuring Mathematical Problem Solving With the MATH Dataset
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 904be90d-30bf-4e36-91d0-1081315e14d6 · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark Novachart: A large- scale dataset towards chart understanding and generation of multimodal large language models
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c3be78fa-f225-494e-99aa-c25eaadca2f7 · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark Towards Reasoning in Large Language Models: A Survey
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a0b885fd-fe6b-4bfa-b9fe-70a49c2421cc · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark Large Language Models Cannot Self-Correct Reasoning Yet
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 52f2157c-e6eb-4379-9ffe-b163cafbc6fe · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark Gqa: A new dataset for real-world visual reasoning and compositional question answering
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dfe477f8-c437-4249-9e75-feb26edff9df · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark Cladder: Assessing causal reasoning in language models
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ea7cef7c-dff0-40a5-8ba3-40574c4ac517 · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark Gonzalez, Hao Zhang, and Ion Stoica
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e4616ed0-a19d-4ef4-9e42-5903514ce8c5 · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark SMiCRM: A Benchmark Dataset of Mechanistic Molecular Images
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 6b3e57b4-b5e3-4591-917b-d272259502de · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark LLaVA-OneVision: Easy Visual Task Transfer
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fcd78bbc-f9b5-431e-89a9-a88d5dd2fcbd · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark Name reactions
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2172043a-f3c7-4aa5-b2bf-0baf33665701 · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c46d5a91-088d-4f2a-b0a3-079f476777ea · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark MMCode: Benchmarking Multimodal Large Language Models for Code Generation with Visually Rich Programming Problems
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 81907be8-e07d-4f0f-8991-5cd04cc98e82 · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark Oscar: Object-semantics aligned pre-training for vision-language tasks
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b91bf0dd-22be-4ac7-a1a6-2f1a1caed467 · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark Mmsci: A multimodal multi-discipline dataset for phd-level scientific comprehen- sion
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 418e932c-63b7-4325-bd55-0a3affc2ae66 · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark Let's Verify Step by Step
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 32a37897-1d1f-465c-90f6-3b860a3363cc · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark DePlot: One-shot visual language reasoning by plot-to-table translation
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation daa5824b-5ab5-4d97-878e-4b0af8525f67 · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark Improved baselines with visual instruction tuning
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 04b9763d-f4c4-478f-ac6e-c34945a1b47d · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark Visual instruction tuning
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ce440d86-6ee9-4f15-97b2-1810d10d93d5 · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark DeepSeek-VL: Towards Real-World Vision-Language Understanding
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e97a3218-0411-4795-aad3-91cb949c7922 · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark Vilbert: Pretraining task-agnostic visiolinguistic representations for vision-and-language tasks
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dc3a9eec-b736-468a-85df-6558e857ae3c · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark Mathvista: Evaluating mathemat- ical reasoning of foundation models in visual contexts
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4d223824-231b-4af4-bcb2-f15754cbcf7a · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark Learn to explain: Multimodal reasoning via thought chains for science question answering
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9efa7098-8324-4f19-8bd5-3dee3d45a696 · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark Ok-vqa: A visual question answering benchmark requiring external knowledge
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ec75f540-35fc-460b-8cd3-8424254bb626 · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark ChartQA: A Benchmark for Question Answering about Charts with Visual and Logical Reasoning
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bb9c855c-498c-4feb-81e7-5ce5e028ed8f · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark Plotqa: Reasoning over scientific plots
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 20217427-0147-4b64-8e23-c4a60a556ade · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark Hello gpt-4o
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5cea753d-68a1-4577-9f8e-c96e07628b5f · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark Introducing chatgpt pro
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 406c6e8e-e7bb-4e91-b9d8-cc0ef78359ca · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark Learning to reason with llms
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 260d2619-6031-4d65-a215-074bc253c68f · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark Learn ap physics
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9f10e2cc-25ac-4f8e-82d7-59cb9e5225fd · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark Qwq: Reflect deeply on the boundaries of the unknown
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 643d4f34-7607-4387-9370-3be0e6c30dc3 · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark Learning transferable visual models from natural language supervi- sion
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 903bcecf-2a35-461b-9a52-c0224d9c6cbc · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark Does Spatial Cognition Emerge in Frontier Models?
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eb48029e-d953-4e6d-b4b0-47d3b7d202d0 · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark GPQA: A Graduate-Level Google-Proof Q&A Benchmark
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9eb40ce7-f312-43a6-8f5a-56d6f48d77ac · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark Visual CoT: Advancing Multi-Modal Language Models with a Comprehensive Dataset and Benchmark for Chain-of-Thought Reasoning
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b926b98a-5a2b-4ad1-98e8-d7d6a8e3d96d · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark ChartMimic: Evaluating LMM's Cross-Modal Reasoning Capability via Chart-to-Code Generation
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 04b26edc-cc0b-4a97-b13d-af27afa29223 · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6b047635-622b-480c-85a1-ddcbde21afd7 · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark Varco arena: A tourna- ment approach to reference-free benchmarking large lan- guage models
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 94cc29a8-17ae-4fe3-aa24-a73ec130fc32 · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2c19a6f9-f116-45e4-8275-9764fc426a7e · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark Challenging BIG-Bench Tasks and Whether Chain-of-Thought Can Solve Them
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b4744dfb-bc4e-42a2-b3d7-be7e0870efb7 · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark LXMERT: Learning Cross-Modality Encoder Representations from Transformers
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 51e65c02-b20e-49d0-9a19-59aa7d024572 · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark Gemini 1.5: Unlocking multimodal under- standing across millions of tokens of context (2024)
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 0742aa80-6b3c-48fa-9ec1-b196452cf0a0 · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark Examples
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 7b9733a9-9b99-4fa3-a9df-25512b5dd4e2 · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark Internvl2: Better than the best— expanding performance boundaries of open-source multi- modal models with the progressive scaling strategy.https: / / internvl
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 9f07bf62-9d4d-4e98-af49-6325d5bfaeca · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark Measuring multimodal mathemat- ical reasoning with math-vision dataset, 2024
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation c83163f6-0433-432a-9d37-7562fef9ad19 · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution
Reference 71
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 98b439a5-bac1-4166-a0b1-2a79c2f0b81f · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark CharXiv: Charting Gaps in Realistic Chart Understanding in Multimodal LLMs
Reference 72
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 39202520-efa4-4ef1-9b55-5da3841ac2b5 · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark Chain-of-thought prompting elicits reasoning in large lan- guage models
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation b025c019-e937-47db-9e0d-03731ffd5513 · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark Unresolved cited work
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation a7ce70ec-02bc-4ff6-ac9a-b52d084494e3 · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark Plot2Code: A Comprehensive Benchmark for Evaluating Multi-modal Large Language Models in Code Generation from Scientific Plots
Reference 75
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fe070716-1d2f-428c-ab00-7e1f46e5e61c · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark ChartX & ChartVLM: A Versatile Benchmark and Foundation Model for Complicated Chart Reasoning
Reference 76
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 65e9244b-fc31-4104-852f-30695a3c8420 · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark Qwen2 Technical Report
Reference 77
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8a458036-67c4-4727-8596-ce5c897211f8 · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark Qwen2.5-Math Technical Report: Toward Mathematical Expert Model via Self-Improvement
Reference 78
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d8ed0517-b6c4-4619-9233-68357788667e · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark Jimenez, Alex L
Reference 79
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 94ee971f-2195-4d7b-ab91-d332cb320c83 · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark MMT-Bench: A Comprehensive Multimodal Benchmark for Evaluating Large Vision-Language Models Towards Multitask AGI
Reference 80
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b32eabd4-57fd-4a1c-b012-c4c8e58938bc · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark CoCa: Contrastive Captioners are Image-Text Foundation Models
Reference 81
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 345be9ac-78c1-4c1e-bf27-b86c6d76d8bb · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark MM-Vet: Evaluating Large Multimodal Models for Integrated Capabilities
Reference 82
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 21804f5c-f753-458e-b07a-f607a8c784c3 · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark Mmmu: A massive multi-discipline multimodal understand- ing and reasoning benchmark for expert agi
Reference 83
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 5d3f8ba1-f2c9-474f-8a80-b20c75bdf66b · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark MMMU-Pro: A More Robust Multi-discipline Multimodal Understanding Benchmark
Reference 84
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 050afd63-1e9e-4872-8d22-36f4c4c66b20 · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark Raven: A dataset for relational and analogical vi- sual reasoning
Reference 85
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation d12b1bc9-7eb0-4ade-8d4d-52a7cb3fb672 · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark Vinvl: Revisiting visual representations in vision-language models
Reference 86
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aaae7146-96f1-443b-828f-a01d8c631281 · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark MathVerse: Does Your Multi-modal LLM Truly See the Diagrams in Visual Math Problems?
Reference 87
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d463a017-94cf-4691-95b5-a2027b7c85b5 · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark Is gpt-4v (ision) all you need for automating academic data vi- sualization? exploring vision-language models’ capability in reproducing academic charts
Reference 88
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 2711860d-6858-42ec-ad0a-3ce0d9c7f0ce · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark Multimodal Chain-of-Thought Reasoning in Language Models
Reference 89
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a51f4df2-e185-40cc-9410-f8be7f4d02ca · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark Marco-o1: Towards Open Reasoning Models for Open-Ended Solutions
Reference 90
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f90a52ac-5fe9-401d-a65e-16d176c36b0c · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark Image-of-Thought Prompting for Visual Reasoning Refinement in Multimodal Large Language Models
Reference 91
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 97664272-fe20-429a-9cd6-eca10c4c8298 · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models
Reference 92
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 65fa8981-6ffa-48d0-b0ff-1855e2056abc · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark Unresolved cited work
Reference 94
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 3fc875dd-3c52-4489-a6b8-215e1f8df131 · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark The vertical distance is 2 units up and 2 units down for each peak
Reference 95
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 29f8b3f7-ddbd-4776-9343-2d4e6f03fac6 · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark The vertical distance is 1.5 units up and 1.5 units down for each peak
Reference 96
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation d7180cf0-e327-42c9-8baa-85222805dbd3 · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark Therefore, the total distance between the two points is the same for all routes
Reference 97
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 5a8737b9-7df3-4bb5-9c18-f3a409915ee2 · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark The second cube shows three different faces: a green triangle, a blue circle, a brown arrow
Reference 98
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 522282fa-f8fe-4e66-b9a3-bfb8ed61b6e4 · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark On the first cube, the green triangle is adjacent to the red square and the yellow star
Reference 99
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 4fc386dc-f6ed-41b3-9499-db7bcc435702 · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark Therefore, the face opposite the green triangle must be the kangaroo (which is not shown in <image2> but is given in <image1>)
Reference 100
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 2b062b73-8c22-44b2-bf54-ef27eeab1d6e · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark The face opposite the kangaroo must be the green triangle
Reference 101
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation e7ea0d9e-c8ac-4e9f-9377-c2cdd81f6daf · outbound
Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark Final Answer: \boxed{B} Figure 17
Reference 102
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 59217165-9cd2-4838-ace1-1f645cccb2b2 · inbound
Position: Multimodal Large Language Models Can Significantly Advance Scientific Reasoning Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation bea1f8ac-2ae8-4271-97b5-1c28b2ddafb2 · inbound
Towards Reasoning Era: A Survey of Long Chain-of-Thought for Reasoning Large Language Models Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark
Reference 241
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation fc20a570-a8c1-499f-a529-97c46d9fcd85 · inbound
Multimodal Chain-of-Thought Reasoning: A Comprehensive Survey Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark
Reference 167
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation a431e228-57cb-4df6-a848-fcc25fc8e8f5 · inbound
OpenVLThinker: Complex Vision-Language Reasoning via Iterative SFT-RL Cycles Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 6de50937-63e5-4a52-8c9a-237f56fc1afa · inbound
Learning to Reason under Off-Policy Guidance Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 4e8e5026-c49a-4518-9216-5cdfe46d6385 · inbound
VisuLogic: A Benchmark for Evaluating Visual Reasoning in Multi-modal Large Language Models Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8cf6e07d-d06f-4b29-88ea-c79f4e77bcec · inbound
Reinforced MLLM: A Survey on RL-Based Reasoning in Multimodal Large Language Models Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation df00b970-8667-45ef-b1e6-a5e22a608ddf · inbound
SurveillanceVQA-589K: A Benchmark for Comprehensive Surveillance Video-Language Understanding with Large Models Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ca4d68d7-8604-484e-b562-19be0e03b0cf · inbound
Towards Spoken Mathematical Reasoning: Benchmarking Speech-based Models over Multi-faceted Math Problems Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8cafd74e-8f58-4fde-a158-02ea21ec09c9 · inbound
lmgame-Bench: How Good are LLMs at Playing Games? Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3269c397-9cf3-456c-a0c7-69b46502952b · inbound
PhyX: Does Your Model Have the "Wits" for Physical Reasoning? Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 13c6dd49-fa38-4685-86aa-7bee9f02578a · inbound
FullFront: Benchmarking MLLMs Across the Full Front-End Engineering Workflow Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e892892b-c10d-45a6-8065-9432d3feab2f · inbound
Reinforcement Fine-Tuning Powers Reasoning Capability of Multimodal Large Language Models Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 82ab3d3d-13f9-4382-8cf0-d5cfdb58dcb1 · inbound
Unveiling the Compositional Ability Gap in Vision-Language Reasoning Model Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0af8138d-b50b-4120-96ae-298ddf751c9a · inbound
Point-RFT: Improving Multimodal Reasoning with Visually Grounded Reinforcement Finetuning Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 060b87ff-030b-457b-bbd4-e859554de971 · inbound
MME-Reasoning: A Comprehensive Benchmark for Logical Reasoning in MLLMs Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation df31dd7b-4de3-468f-8f15-7c3087c86cca · inbound
CSVQA: A Chinese Multimodal Benchmark for Evaluating STEM Reasoning Capabilities of VLMs Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a5964cce-ce42-4bca-9940-f8b194a075a7 · inbound
Seeing is Not Reasoning: MVPBench for Graph-based Evaluation of Multi-path Visual Physical CoT Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 58541b15-d921-4064-b89d-6dee79b1e6b8 · inbound
Mimicking or Reasoning: Rethinking Multi-Modal In-Context Learning in Vision-Language Models Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c27708dd-cd6c-4f19-9f5c-898587f63c4b · inbound
Wait, We Don't Need to "Wait"! Removing Thinking Tokens Improves Reasoning Efficiency Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9024f7a6-09d5-44b8-a2c3-b8d1e33ac7af · inbound
ViCrit: A Verifiable Reinforcement Learning Proxy Task for Visual Perception in VLMs Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 55df135a-9ca6-4fe7-ad75-4bb19f440d8c · inbound
MARBLE: A Hard Benchmark for Multimodal Spatial Reasoning and Planning Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6d4fe260-a268-459c-bf38-30fcf3dd1d80 · inbound
CaughtCheating: Is Your MLLM a Good Cheating Detective? Exploring the Boundary of Visual Perception and Reasoning Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d1c9e39c-da3b-451e-b3c2-7c06538b096d · inbound
Skywork-R1V3 Technical Report Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b7ae7c65-1c24-4eef-8281-58c46bfdc0d1 · inbound
VL-Cogito: Progressive Curriculum Reinforcement Learning for Advanced Multimodal Reasoning Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9d5b0801-966a-45f1-9e27-1ecb18d0c745 · inbound
SEAM: Semantically Equivalent Across Modalities Benchmark for Vision-Language Models Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f86f6a0d-31ae-4ba1-9bb2-869119a11ede · inbound
KRETA: A Benchmark for Korean Reading and Reasoning in Text-Rich VQA Attuned to Diverse Visual Contexts Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aa570e30-2544-4fe1-8f67-24262a221426 · inbound
LLaVA-Critic-R1: Your Critic Model is Secretly a Strong Policy Model Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c66eeefd-99ec-410e-8435-a40e3634a2e5 · inbound
MiniCPM-V 4.5: Cooking Efficient MLLMs via Architecture, Data, and Training Recipe Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation ddc2eb4c-96b9-4766-992b-70958a3ee1c7 · inbound
SoM-1K: A Thousand-Problem Benchmark Dataset for Strength of Materials Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b5e52865-c1e2-4a43-b5f1-2576227b5ab7 · inbound
AgroCoT: A Chain-of-Thought Benchmark for Evaluating Reasoning in Vision-Language Models for Agriculture Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation aa028a27-b5f1-47b4-99c1-6f0219a0a6b1 · inbound
MentisOculi: Revealing the Limits of Reasoning with Mental Imagery Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b8dcbd05-29aa-456e-a0fa-05bb7fc6ef92 · inbound
SPM-Bench: Benchmarking Large Language Models for Scanning Probe Microscopy Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 314fcb5f-e6d6-4dd8-9e5f-f82c27d3b6bf · inbound
Seed1.8 Model Card: Towards Generalized Real-World Agency Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation fabe2c58-eaca-4c41-9005-42d8daed444e · inbound
Dual-Cluster Memory Agent: Resolving Multi-Paradigm Ambiguity in Optimization Problem Solving Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark
Reference 84
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 3325c29f-6bc7-4f06-b15b-841b61ccdeb9 · inbound
OptiVerse: A Comprehensive Benchmark towards Optimization Problem Solving Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark
Reference 100
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 67764f88-0862-4b1c-9f4a-002e83632f6b · inbound
Reflection Anchors for Propagation-Aware Visual Retention in Long-Chain Multimodal Reasoning Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation edaac7b9-5988-4235-82c7-2e1c8513f50e · inbound
Bad Seeing or Bad Thinking? Rewarding Perception for Multimodal Reasoning Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark
Reference 94
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation eb529714-10e1-4e6b-88ff-28e8a71ad0a3 · inbound
Reversing the Flow: Generation-to-Understanding Synergy in Large Multimodal Models Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 7e1c70c0-7481-4bd7-af97-e107c43eae73 · inbound
AnE: Pushing the Reasoning Frontier of Multimodal LLMs via Anchor Evolution Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 41f675df-f331-4c6b-9d3f-3dd540b15384 · inbound
Causal Scaffolding for Physical Reasoning: A Benchmark for Causally-Informed Physical World Understanding in VLMs Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 01469653-fd64-47f6-bc32-e6f8c0ffde20 · inbound
From Hallucination to Grounding: Diagnosing Visual Spatial Intelligence via CRISP Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation dcc576b0-35a1-4d70-8ba8-bdb99db859d1 · inbound
Seed2.0 Model Card: Towards Intelligence Frontier for Real-World Complexity Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 1ad6a5de-5851-4a93-8ca7-e9b1d05f83b4 · inbound
ProLaViT: Learning Progressive Latent Visual Thoughts in Structured Latent Space Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e6992855-794e-4037-b7d3-00ab9c601dcb · inbound
PerceptionBench: Evaluating Atomic Visual Perception in Multimodal Large Language Models Can MLLMs Reason in Multimodality? EMMA: An Enhanced MultiModal ReAsoning Benchmark
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.