Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-23T20:10:59.264484Z
Paper Citation Record · LEDGER
As of 5 August 2026, this Paper Citation Record lists 92 of 92 outbound references and 1 inbound Pith citation observation for arXiv:2410.04509.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-23T20:10:59.264484Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-05-23T04:30:38.804702Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-05-23T04:32:32.712337Z
92 of 92 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 391c2e3d-8ab0-4bb8-a7e8-0f1680031792 · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection Complexity in declarative process models: Metrics and multi-modal assessment of cognitive load
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 055423cd-2fb1-4a5f-a3ed-bdbab039e4d1 · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 081df14e-80ae-4bb1-ab08-30c62b6666a9 · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection Scaling laws for generative mixed-modal language models
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 13237bbd-396e-47ce-8b20-4238944258aa · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection Large Language Models for Mathematical Reasoning: Progresses and Challenges
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation ccb84f09-1457-4698-b7fa-8999f8a525ca · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection Claude 3, 2024 a
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 5e6cf1ad-5f67-473c-bd17-555a775c03ce · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection Claude 3.5, 2024 b
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation df8fda6c-3f85-440e-8684-a2f46e46d525 · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 8d8b027e-93cc-4830-a1cc-f388a5eda008 · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection Turning large language models into cognitive models
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 4ca4f385-91cd-4b24-95d3-a76458269871 · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection Theoremqa: A theorem-driven question answering dataset
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 3be8ded9-a037-42a7-97eb-24e30023ecc7 · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection InternVL: Scaling up Vision Foundation Models and Aligning for Generic Visual-Linguistic Tasks
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation ce162e8b-2af0-4b96-8ef1-f74d6e85f1d2 · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection Training Verifiers to Solve Math Word Problems
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 9dc97bf6-0f72-4722-8dcd-ec1e2a400ba1 · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection A survey on multimodal large language models for autonomous driving
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 16026aa9-6449-4604-a689-020543105388 · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection Advancing mathematics by guiding human intuition with ai
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 21518bc5-3b88-484c-822b-15c1805cd1d6 · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection Visual representations in the human brain are aligned with large language models
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 52922b22-c1d3-45ca-bf92-467dfa39452a · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection Muffin or chihuahua? challenging multimodal large language models with multipanel vqa
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation c17fe3b3-ce1f-488e-ab6a-d6b5706b8563 · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection Trends in Integration of Knowledge and Large Language Models: A Survey and Taxonomy of Methods, Benchmarks, and Applications
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation e53d8485-5e44-4821-8c36-a047c854cde0 · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection IsoBench: Benchmarking Multimodal Foundation Models on Isomorphic Representations
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 33cedaae-dab5-4308-868d-eabed4c4ea68 · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection ChatGLM: A Family of Large Language Models from GLM-130B to GLM-4 All Tools
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 61b1e7d4-020b-45a2-a222-07c3f5485eaf · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection UrbanVLP: Multi-Granularity Vision-Language Pretraining for Urban Socioeconomic Indicator Prediction
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 868a8b7a-3584-45a6-9cac-c12faeedf4c6 · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection PeFoMed: Parameter Efficient Fine-tuning of Multimodal Large Language Models for Medical Imaging
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation afc6ffdf-2f90-4c99-9de8-680fd9e1e13c · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection CMMU: A Benchmark for Chinese Multi-modal Multi-type Question Understanding and Reasoning
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 61c2cfe1-22c8-449d-b678-ebbc4c2dbc27 · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection Measuring Mathematical Problem Solving With the MATH Dataset
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 8900b923-fec1-4134-bf1e-df2e0167bfa0 · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection Visual Sketchpad: Sketching as a Visual Chain of Thought for Multimodal Language Models
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation ec757e1c-078f-417a-9525-dc588045b480 · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection MMNeuron: Discovering Neuron-Level Domain-Specific Interpretation in Multimodal Large Language Model
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation ccc2ac53-320e-476b-a810-444d8c4da6bf · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection Describe-then-Reason: Improving Multimodal Mathematical Reasoning through Visual Comprehension Training
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 1817a52e-bc03-4163-845e-e23b32334315 · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection New generation deep learning for video object detection: A survey
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 1e079590-4b40-45db-8bf9-78b8a447502e · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection Learning instance-level representation for large-scale multi-modal pretraining in e-commerce
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation fd89f66e-3a2a-42cd-9720-86d0ad964e32 · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection Large language models struggle to learn long-tail knowledge
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 10420b75-8e73-4209-bad3-e386591f6652 · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection Scaling Laws for Neural Language Models
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 73b4a142-6d39-40c3-9099-df106ddfdfb6 · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection Cognitive load theory: An applied reintroduction for special and general educators
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation d3ceb4ea-2a12-403f-92e2-70ad657c6bfc · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection Large language models are zero-shot reasoners
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation bc925469-d196-4f80-a92e-8f719e36071c · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection Solving quantitative reasoning problems with language models
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation ca615388-f73a-4a0b-9df3-4df1a512c422 · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection Bringing Generative AI to Adaptive Learning in Education
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 14d40029-233f-43c0-99e7-30d1188c3bdf · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection Evaluating Mathematical Reasoning of Large Language Models: A Focus on Error Identification and Correction
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation e4a372f3-2979-437e-8640-101eb7831d61 · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection CMMaTH: A Chinese Multi-modal Math Skill Evaluation Benchmark for Foundation Models
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 981eb948-3025-48ee-83e2-0d6677026d28 · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection Llava-next: Improved reasoning, ocr, and world knowledge, January 2024 a
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation a4fb8dfe-28f1-432b-9c51-6b8ead269c3b · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection MathBench: Evaluating the Theory and Application Proficiency of LLMs with a Hierarchical Mathematics Benchmark
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 213f5ab0-6c1f-454d-809c-372f739c275f · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection Are LLMs Capable of Data-based Statistical and Causal Reasoning? Benchmarking Advanced Quantitative Reasoning with Data
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation bb2f1cd8-6e24-43f5-8ac0-7e500414bbac · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection DeepSeek-VL: Towards Real-World Vision-Language Understanding
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 4add98e9-d240-4094-9622-d47fb5d22cda · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection A Survey of Deep Learning for Mathematical Reasoning
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 740ff5f3-23ef-4f57-877e-1a072ddbc86d · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection MathVista: Evaluating Mathematical Reasoning of Foundation Models in Visual Contexts
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 5ae4255b-12a5-483b-a923-2ccbaea78250 · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection Chameleon: Plug-and-play compositional reasoning with large language models
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation a4d2500f-7082-4485-a0e7-69bbf93779de · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection Large Language Models: A Survey
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 1d180948-b687-4322-b825-7e8fb872c1e7 · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection Scaling data-constrained language models
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 91aa87ec-9310-4dcc-a70c-71666846f229 · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection GPT-4 Technical Report
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation bc700717-d009-44ea-9e7b-272d59800955 · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection GPT-4V(ision) system card, 2024 a
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation bfbc8825-449e-4ddc-9c43-9c4e997b51e8 · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection Gpt-4o mini: advancing cost-efficient intelligence, 2024 b
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 5d29731a-20f6-4213-8362-89fdc2eb83b0 · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection Cognitive load theory and instructional design: Recent developments
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation a1e38f9c-de1a-42cd-9fb2-0d325458eb1a · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection Gemini Goes to Med School: Exploring the Capabilities of Multimodal Large Language Models on Medical Challenge Problems & Hallucinations
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 0f08144f-7f53-4bce-967b-e99f9d347ae5 · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection MultiMath: Bridging Visual and Mathematical Reasoning for Large Language Models
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation b92f1d45-4b97-46a8-b689-3c3579cf5d2b · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection We-Math: Does Your Large Multimodal Model Achieve Human-like Mathematical Reasoning?
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 6109d41c-3620-4bfd-bf38-7ec138d14b65 · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection Elementary math learning through piaget's cognitive development stages
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation c5183aab-f5a4-4d10-ac62-c2ca24b4d04e · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 1cdd9210-316a-4bb2-a8c4-de08309256d0 · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection Detecting Pretraining Data from Large Language Models
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation dd81da68-afc6-487a-b30d-16383dfef233 · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection Math-LLaVA: Bootstrapping Mathematical Reasoning for Multimodal Large Language Models
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 89ecd939-1a04-4291-8044-b49b18e9cc38 · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection How to Bridge the Gap between Modalities: Survey on Multimodal Large Language Model
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 81b3a9e7-e85a-4bea-b552-7cf98bfc58c5 · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection Scieval: A multi-level large language model evaluation benchmark for scientific research
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 6696d416-a815-4cb5-bea7-38f0a3071df7 · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection Aligning Large Multimodal Models with Factually Augmented RLHF
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation d0ae4849-dc13-4b84-8d16-27965b2d6f13 · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection Memorization without overfitting: Analyzing the training dynamics of large language models
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation d7ddc412-30ba-4a2d-acbc-506790c88a42 · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection Measuring Multimodal Mathematical Reasoning with MATH-Vision Dataset
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 6d1ba5fa-f238-4a5e-aa60-3ad5e5606653 · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection Large Language Models for Education: A Survey and Outlook
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation dedad807-91d7-4d41-ad1f-0185068474a2 · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection CogVLM: Visual Expert for Pretrained Language Models
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 2841365e-1eb1-47e2-818b-53502139cdcc · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection Large-scale multi-modal pre-trained models: A comprehensive survey
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation e1bbde25-40b2-4a34-b348-a6b4424acfbe · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection SciBench: Evaluating College-Level Scientific Problem-Solving Abilities of Large Language Models
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 0b2da07f-e806-4b7e-a552-f13f425ffe47 · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection Exploring the Reasoning Abilities of Multimodal Large Language Models (MLLMs): A Comprehensive Survey on Emerging Trends in Multimodal Reasoning
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 2e8d19e2-7d15-4fb1-98f4-a5faad70e5bf · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection Are deep neural networks adequate behavioral models of human visual perception? Annual Review of Vision Science, 9 0 (1): 0 501--524
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation b7bf2295-098e-429e-bd4b-71e3631f542d · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection A Comprehensive Survey of Large Language Models and Multimodal Large Language Models in Medicine
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation ad01e102-61e5-4ae0-be2e-9cc1545c302f · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection MIND: Multimodal Shopping Intention Distillation from Large Vision-language Models for E-commerce Purchase Understanding
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 752b09f9-3757-4f31-99a5-b587d6533ca1 · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection SuperCLUE-Math6: Graded Multi-Step Math Reasoning Benchmark for LLMs in Chinese
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 6ed7585e-c8b0-45b2-b324-b47400a2c4e1 · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection Raise a Child in Large Language Model: Towards Effective and Generalizable Fine-tuning
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation eeb102f7-ccbe-4efe-9ff6-0a01b33774c4 · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection Emerging Synergies Between Large Language Models and Machine Learning in Ecommerce Recommendations
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation c1a9acb3-51a9-4a54-8acc-a199d0c57ce1 · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection GeoReasoner: Reasoning On Geospatially Grounded Context For Natural Language Understanding
Reference 72
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation ae8143ff-b337-457f-ac72-5a7d537990fe · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection Urbanclip: Learning text-enhanced urban region profiling with contrastive language-image pretraining from the web
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 0372a539-9a1d-41d3-9673-a0093cb903f5 · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection Exploring diverse in-context configurations for image captioning
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 17c4ad5f-dc70-4c37-ad5e-3797dcbb1393 · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection MiniCPM-V: A GPT-4V Level MLLM on Your Phone
Reference 75
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 2bab7c3f-75c0-4163-9282-e0ca44687f23 · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection Yi: Open Foundation Models by 01.AI
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 7dd62501-1f2d-42ad-929b-c34c289c1fb3 · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection Large language model as attributed training data generator: A tale of diversity and bias
Reference 77
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation fdf4f964-e45d-4c93-ad69-829831adaa9a · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection MR-Ben: A Meta-Reasoning Benchmark for Evaluating System-2 Thinking in LLMs
Reference 78
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 1fd7739d-d3c1-447c-9fb2-ef9a29d4363c · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection MathVerse: Does Your Multi-modal LLM Truly See the Diagrams in Visual Math Problems?
Reference 79
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 60de8a06-b8b0-4a34-a59e-75de3c30507a · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection A Survey of Large Language Models
Reference 80
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 26e69ea4-d823-4e60-88c0-5da7588b24d4 · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection Reefknot: A Comprehensive Benchmark for Relation Hallucination Evaluation, Analysis and Mitigation in Multimodal Large Language Models
Reference 81
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation d7ccc97d-03ac-4148-a9a0-67fb8b57f831 · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection UrbanCross: Enhancing Satellite Image-Text Retrieval with Cross-Domain Adaptation
Reference 82
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 6a4c28c4-ad13-4c28-8dfe-f0bb8de28cef · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection Mathscape: Evaluating mllms in multimodal math scenarios through a hierarchical benchmark
Reference 83
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 774c5a7e-bd97-4aff-a820-5a3e8933d4cc · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection Large Language Model for Participatory Urban Planning
Reference 84
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 6be2948c-6684-4b74-9e22-fbe581ffa754 · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection Large language models for information retrieval: A survey
Reference 85
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 882d7d89-01da-45cb-8219-31bb3569df79 · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection Math-PUMA: Progressive Upward Multimodal Alignment to Enhance Mathematical Reasoning
Reference 86
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation eb14ca03-8ff0-4d6f-988c-96131b7302d4 · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection Deep learning for cross-domain data fusion in urban computing: Taxonomy, advances, and outlook
Reference 87
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 0ce46bd0-3a45-4866-b6a3-476d24ef4787 · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection Object detection in 20 years: A survey
Reference 88
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 6892dd8d-ae52-453f-bd0a-45fa91b59f44 · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection write newline
Reference 89
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 2229d2cb-fbcf-4124-8669-856d25a312ec · outbound
Reference 90
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation b4822c65-1ae0-4888-937b-1572a7a59bd0 · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection Unresolved cited work
Reference 91
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 0393e8f1-3007-41c8-a9e9-5de496b1dab9 · outbound
ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection Unresolved cited work
Reference 92
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 0eabe05b-56f9-4448-8a95-946e31910d21 · inbound
Position: Multimodal Large Language Models Can Significantly Advance Scientific Reasoning ErrorRadar: Benchmarking Complex Mathematical Reasoning of Multimodal Large Language Models Via Error Detection
Reference 224
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.