Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-17T00:24:45.343430Z
Paper Citation Record · LEDGER
As of 6 August 2026, this Paper Citation Record lists 22 of 22 outbound references and 57 inbound Pith citation observations for arXiv:2502.09560.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-17T00:24:45.343430Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-06T00:13:24.699600Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-05T02:28:24.338817Z
22 of 22 outbound references displayed
External citation measurements
1
pith, observed 2026-08-05T02:28:24.338817Z
Observation 077f13c7-f4f6-4281-9a93-74407eed553e · outbound
EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents Put washed lettuce in the refrigerator
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 83542d9a-a771-4696-8036-4bc678362fac · outbound
EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents Unresolved cited work
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 5e3f3d06-7814-40bf-8c03-e2845c01765d · outbound
EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents Avoid performing actions that do not meet the defined validity criteria
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 35028448-e93f-4a45-8109-acf312d1bafe · outbound
EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents You can explore these instances if you do not find the desired object in the current receptacle
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 797cd098-8497-4ada-862e-56bd9744abb0 · outbound
EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents If the last action is invalid, reflect on the reason, such as not adhering to action rules or missing preliminary actions, and adjust your plan accordingly
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 71e67403-b852-4a13-97c4-542ec1991b58 · outbound
EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents Each plan should include no more than 20 actions
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 977209b6-ef16-472f-8278-f2a71daa2f90 · outbound
EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents Unresolved cited work
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation dee88191-3283-4676-a5c7-88ac72b6fe26 · outbound
EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents Avoid performing actions that do not meet the defined validity criteria
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 36546a12-1133-4870-ac32-3cb138ea2feb · outbound
EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents Try to modify the action sequence because previous actions do not lead to success
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 299b9357-228e-485f-b20b-91522f8445d7 · outbound
EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents You can explore these instances if you do not find the desired object in the current receptacle
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 403ea813-25c6-4af7-9e4d-837180734fce · outbound
EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents If the last action is invalid, reflect on the reason, such as not adhering to action rules or missing preliminary actions, and adjust your plan accordingly
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 72382bf5-739c-46d4-a07b-a06d9db221e9 · outbound
EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents try to be as close as possible
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 7350b6a4-db4a-4193-9f47-bf12a4df4cdd · outbound
EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents on the front left side, a few steps from the current standing point)
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation b5693947-0715-44f1-8fc4-047add080fa7 · outbound
EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents When planning for movement, reason based on target object’s location and obstacles around you
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation aa04e17b-1ded-4bf1-9cd1-2f1961248a69 · outbound
EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents In other words, do not overly focus on correcting invalid actions when direct movement toward the target object can still bring you closer
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 72944b41-f8aa-43b6-9ec0-a8cb8d3cadc9 · outbound
EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents If so, plan nothing but ONE ROTATION at a step until that object appears in your view
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 40c206ad-1f5e-46a5-9c25-0ff9c6852f30 · outbound
EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents red", "maroon
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 8097f701-ec5a-42d1-9348-8ca586228f76 · outbound
EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents There are two copper-colored pots visible on the stovetop
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation a08ff14e-16e8-4dae-b047-36d8e03a1dc4 · outbound
EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents pick up the spoon3
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation acb841d4-5da2-4479-b6ed-3428036d8655 · outbound
EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents pick up the sponge9
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 9d54dd95-e210-44a7-81ba-c1d0bcfb4efb · outbound
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 7cc72fbe-ec6d-4038-8692-4a40cfdd318f · outbound
EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents Rotate to the left by 90 degrees18
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation cd0faa7a-c349-44bf-baa1-866bfa5a84ca · inbound
From LLM Reasoning to Autonomous AI Agents: A Comprehensive Review EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents
Reference 124
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 1c01cf7b-0cb0-43b0-a351-1d7b11dc2069 · inbound
VS-Bench: Evaluating VLMs for Strategic Abilities in Multi-Agent Environments EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents
Reference 83
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 2fbb1055-bac0-4de1-afab-f3b463cd6f0b · inbound
A Survey of Self-Evolving Agents: What, When, How, and Where to Evolve on the Path to Artificial Super Intelligence EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents
Reference 113
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 2ce6b19d-f7b0-40cc-bf15-783a3356c17d · inbound
RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0c3c7d1d-2359-447e-acd6-b3946edd022e · inbound
World Simulation with Video Foundation Models for Physical AI EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents
Reference 90
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 001bf018-0459-48d5-a0b6-fc31a33ace18 · inbound
SWITCH: Benchmarking Modeling and Handling of Tangible Interfaces in Long-horizon Embodied Scenarios EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e46753b1-8384-4fb1-8788-a17ba1a5348e · inbound
Beyond Description: Cognitively Benchmarking Fine-Grained Action for Embodied Agents EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6c686cc6-2972-49c6-a179-b04d943a1882 · inbound
Vision-Language Memory for Spatial Reasoning EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9c4ac8e5-6d71-49fb-a73a-d9d4a7249104 · inbound
Multimodal Reinforcement Learning with Adaptive Verifier for AI Agents EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation f552b55e-3a86-4cb2-9843-7e8bfaf3efeb · inbound
Multimodal Reinforcement Learning with Adaptive Verifier for AI Agents EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9718d69a-9d11-4238-a967-d9ec96a96563 · inbound
Unified Embodied VLM Reasoning with Robotic Action via Autoregressive Discretized Pre-training EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c503a699-ab9c-40d3-8e94-86609930ed8f · inbound
TCAP: Tri-Component Attention Profiling for Unsupervised Backdoor Detection in MLLM Fine-Tuning EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 96c42348-e4f9-4d2a-9dd5-373bf61df92a · inbound
PLanAR: Planning-Language-Grounded Agentic Reasoning for Robot Manipulation EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 134909e8-416c-4e2d-bab2-370b3069cce1 · inbound
ST-BiBench: Benchmarking Multi-Stream Multimodal Coordination in Bimanual Embodied Tasks for MLLMs EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents
Reference 121
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 0275c72d-a767-486b-becb-64c8b9991453 · inbound
LLM-WikiRace Benchmark: How Far Can LLMs Plan over Real-World Knowledge Graphs? EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b1c5c87f-0953-4960-91d0-503c13e898c5 · inbound
MapTab: A Diagnostic Benchmark for Long-Horizon Multi-Criteria Multimodal Reasoning on Heterogeneous Topological Graphs EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents
Reference 91
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation cb1a712b-eeda-4722-a10e-094923d1432e · inbound
MapTab: A Diagnostic Benchmark for Long-Horizon Multi-Criteria Multimodal Reasoning on Heterogeneous Topological Graphs EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents
Reference 91
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation e8579d2d-49ba-4eb7-b5a9-329f5edd818a · inbound
MapTab: A Diagnostic Benchmark for Long-Horizon Multi-Criteria Multimodal Reasoning on Heterogeneous Topological Graphs EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents
Reference 90
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b79e4a07-6e16-4773-b3f8-617128355912 · inbound
Evaluation as Evolution: Transforming Adversarial Diffusion into Closed-Loop Curricula for Autonomous Vehicles EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 941cf336-ccfb-4be9-8e3c-6ccd52b2dcf8 · inbound
Online Reasoning Video Object Segmentation EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 0c4c9b45-3d9f-4660-81e4-5f8065c3b317 · inbound
ESCAPE: Episodic Spatial Memory and Adaptive Execution Policy for Long-Horizon Mobile Manipulation EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation e4e02e69-4502-4cd0-a939-dcc4063492cc · inbound
MirrorBench: Evaluating Self-centric Intelligence in MLLMs by Introducing a Mirror EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 9f71d0a8-205d-4d73-b998-9c77e60e71ee · inbound
BrainMem: Brain-Inspired Evolving Memory for Embodied Agent Task Planning EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation a29bcb38-f12a-4216-bd2c-f3aeee16f432 · inbound
Chain Of Interaction Benchmark (COIN): When Reasoning meets Embodied Interaction EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation eac2dfab-431f-4838-84c8-48b4b5a2c81f · inbound
GaLa: Hypergraph-Guided Visual Language Models for Procedural Planning EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 55884cfe-fbf7-41af-a01c-d9a6181c63f4 · inbound
Environmental Understanding Vision-Language Model for Embodied Agent EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 3389aba7-1e28-4e1d-aac3-d84a8eb2fd1f · inbound
Safactory: A Scalable Agentic Infrastructure for Training Trustworthy Autonomous Intelligence EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents
Reference 92
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 2437dcb2-8887-46af-a3ac-3b02cdb6aefd · inbound
Safactory: A Scalable Agentic Infrastructure for Training Trustworthy Autonomous Intelligence EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents
Reference 92
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 5e155ef8-cdd4-4b86-8fb9-022f5a54d294 · inbound
MemCompiler: Compile, Don't Inject -- State-Conditioned Memory for Embodied Agents EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation b37b9033-653c-4d26-b790-b11e552c08be · inbound
MemCompiler: Compile, Don't Inject -- State-Conditioned Memory for Embodied Agents EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 0246b88a-f056-425e-9c3f-0ac5c4bef107 · inbound
Done, But Not Sure: Disentangling World Completion from Self-Termination in Embodied Agents EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 3bfa3d12-6d43-4469-bee3-b4bd9c966058 · inbound
Done, But Not Sure: Disentangling World Completion from Self-Termination in Embodied Agents EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 6ae12fb6-355e-4744-98df-f7c72b39cc31 · inbound
Done, But Not Sure: Disentangling World Completion from Self-Termination in Embodied Agents EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 33fd9655-e68a-4a65-b101-b72f71d2bb8b · inbound
SceneFunRI: Reasoning the Invisible for Task-Driven Functional Object Localization EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 3bfad6e8-56bb-4bb2-8fa6-a580031e4416 · inbound
CosFly-Track: A Large-Scale Multi-Modal Dataset for UAV Visual Tracking via Multi-Constraint Trajectory Optimization EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 0617a6c1-888f-49ab-9de0-57633cbab29f · inbound
CosFly-Track: A Large-Scale Multi-Modal Dataset for UAV Visual Tracking via Multi-Constraint Trajectory Optimization EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation ea6908c6-e8a1-42ca-b176-5751481be93c · inbound
AtlasVA: Self-Evolving Visual Skill Memory for Teacher-Free VLM Agents EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 2bed4e9a-b408-4a18-bcde-da92fcdf242c · inbound
DexHoldem: Playing Texas Hold'em with Dexterous Embodied System EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 572d38ee-4d2e-4dab-97d9-c762949e15f2 · inbound
WorldMemArena: Evaluating Multimodal Agent Memory Through Action-World Interaction EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation e4c219b1-1b7e-43f8-8449-93a2d4a6f311 · inbound
OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 45dda8e8-3acf-491b-b450-a7a1226d1c4c · inbound
Sample-Efficient Post-Training for LEGO Spatial-Physics Reasoning EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 5541e0bd-d6fb-49d9-9bbf-de7e114fb2f1 · inbound
ChronoPhyBench: Do MLLMs Truly Understand the World or Merely Exploit Language Priors? EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation e4bdfe65-5a21-4455-9f69-007194d073a2 · inbound
SpatialWorld: Benchmarking Interactive Spatial Reasoning of Multimodal Agents in Real-World Tasks EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents
Reference 82
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 5fba39a7-eb8e-43a2-9743-d7a4f974a161 · inbound
Embodied-BenchClaw: An Autonomous Multi-Agent System for Embodied Spatial Intelligence Benchmark Construction EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 77b3e495-0d87-411f-a1c9-1adb18b4b53b · inbound
InternVideo3: Agentify Foundation Models with Multimodal Contextual Reasoning EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents
Reference 103
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation d6be2457-669e-4684-8a08-77c73f53f3e0 · inbound
Intelligent Automation for Embodied Benchmark Construction: Pipelines, Embodiments, Simulators, and Trends EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation e5c0c01f-3b83-4292-8ff4-68cc6d9a4205 · inbound
GroundControl: Anticipating Navigation Failures in Vision-Language Agents via Trajectory-Consistent Uncertainty Estimates EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation f773a281-2f9c-4f7e-a320-495ce06fbbb2 · inbound
MultiUAV-Plat: An LLM-Oriented Platform, Benchmark and Framework for Multi-UAV Collaborative Task Planning EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation dbb92018-0b8a-4cb0-8eed-9f322a2f62f9 · inbound
Multi-scale Mixture of World Models for Embodied Agents in Evolving Environments EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation ed3391f0-0627-4f70-ab1b-fd1ab6c758c0 · inbound
LLM Agents for Deliberative Collaboration: A Study on Joint Decision Making Under Partial Observability EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents
Reference 174
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 5b528198-cced-4cf0-9480-ec9bea17795a · inbound
Who&When Pro: Can LLMs Really Attribute Failures in AI Agents? EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f0de9766-8888-461b-9763-29a1d24acac7 · inbound
ABot-AgentOS: A General Robotic Agent OS with Lifelong Multi-modal Memory EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents
Reference 105
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8e922246-0a77-497f-8128-b5740718dd89 · inbound
ABot-AgentOS: A General Robotic Agent OS with Lifelong Multi-modal Memory EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents
Reference 105
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 44386acd-0f27-472c-9eb1-2b18717fd673 · inbound
Leveraging Trajectory Graphs for Pre-Execution Error Diagnosis in Agentic LLM Systems EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 437c2389-511d-4f3b-ada4-e114328adeaa · inbound
SafeNexus: Discovering and Steering Modality-Universal Safety Neurons in MLLMs EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ac8bf5ec-4b22-432b-bf15-3e24b956d6e0 · inbound
Long-Horizon Embodied Decision-Making via Multimodal Memory Compression EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 69fb0e21-b44b-4472-ac1b-fd15c76213df · inbound
Is Inter-Seed Cross-Play Enough? Evaluating the Robustness of Zero-Shot Coordination Algorithms to Implementation Details EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.