Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-08T11:25:49.254278Z
Paper Citation Record · LEDGER
As of 19 August 2026, this Paper Citation Record lists 49 of 49 outbound references and 2 inbound Pith citation observations for arXiv:2502.07949.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-08T11:25:49.254278Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-11T20:19:16.013755Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-07T05:04:05.836097Z
49 of 49 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation db6d4381-44e3-4be6-89e6-1a7aff1885ea · outbound
Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning Relative Entropy Regularized Policy Iteration
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ad348645-6fc5-4b15-a8d7-52f35171101e · outbound
Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning Maximum a Posteriori Policy Optimisation
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation afb04b57-749f-4a82-bae2-30382afbd04b · outbound
Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning Andrychowicz, F
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 940735a0-7fbc-4adb-a395-994715bfa4f0 · outbound
Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning DigiRL: Training In-The-Wild Device-Control Agents with Autonomous Reinforcement Learning
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e46950f5-deb2-461b-b25e-8f166c7fb301 · outbound
Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ab3c5455-2785-453a-a08c-15817c87c355 · outbound
Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning Dota 2 with Large Scale Deep Reinforcement Learning
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f50414d4-631b-4de0-babf-65f21a162716 · outbound
Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning Chane-Sane, C
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 9a0861c7-54df-4708-b225-9e2f16cd2e1e · outbound
Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning Unresolved cited work
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 01e8418b-e7c6-4b6d-9364-cc4c414ab025 · outbound
Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning Chevalier-Boisvert, D
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 6ca99384-01e5-4238-8fdf-f7fe09267918 · outbound
Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning Dayan and G
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 170d37da-5366-418e-a2e8-e0df0d179100 · outbound
Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning MetaGPT: Meta Programming for A Multi-Agent Collaborative Framework
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation df641b50-b2f7-4fde-8402-03936d23a97d · outbound
Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning Unresolved cited work
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2183db8f-1589-4447-86ce-66fdffe95b51 · outbound
Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning Jiang and A
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation a53ce53f-bd3a-46cb-93a2-7073fe2b8ebf · outbound
Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning Jurgenson, O
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation f278e4ea-5612-4a4e-a7df-e544b1fb432a · outbound
Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning Reinforcement Learning and Control as Probabilistic Inference: Tutorial and Review
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 106e6232-3df4-4967-9c6b-b83b751e6ca6 · outbound
Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning Goal-Conditioned Reinforcement Learning: Problems and Solutions
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation acf8ca8a-b7fe-472b-84aa-d2d0445cd6a6 · outbound
Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning VisualAgentBench: Towards Large Multimodal Models as Visual Foundation Agents
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cfdbcdd3-fcaf-4b3f-b471-1f435bcbb150 · outbound
Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning Unresolved cited work
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 4b6e6fa3-0b7b-4c50-9300-2a45812dfec0 · outbound
Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning Hierarchical Foresight: Self-Supervised Learning of Long-Horizon Tasks via Visual Subgoal Generation
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 57ea3c29-b182-46d1-b602-baddc31d951c · outbound
Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning Unresolved cited work
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 210d1af5-1431-4db5-82c3-8d2e472ba75a · outbound
Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning Gpt-4v(ision) technical work and authors
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 9791b529-d7f0-4d83-bff0-7cc2ab1f5821 · outbound
Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning Unresolved cited work
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation edc6f4e7-e33c-4748-b23e-10fe327137a3 · outbound
Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning Unresolved cited work
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 1a828d38-91fb-433c-8a94-c1d945a87768 · outbound
Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning Divide-and-Conquer Monte Carlo Tree Search For Goal-Directed Planning
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 7a761859-ce8f-4eb8-9b89-becb5342b309 · outbound
Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning Advantage-Weighted Regression: Simple and Scalable Off-Policy Reinforcement Learning
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3d05642c-2028-4729-99d8-165576046229 · outbound
Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning WebRL: Training LLM Web Agents via Self-Evolving Online Curriculum Reinforcement Learning
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 854c3f89-17f3-49bf-89f8-8ad27d167861 · outbound
Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning Rawles, A
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation c0df10cc-6fa3-4edf-bb43-6e49a0db3b95 · outbound
Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning Vision-Language Models are Zero-Shot Reward Models for Reinforcement Learning
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9a24d763-76fb-4fad-b423-68a61695c180 · outbound
Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning Proximal Policy Optimization Algorithms
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b8cc4d12-7acf-43ea-a751-20b619dc1bdf · outbound
Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning Unresolved cited work
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 7a76d994-278f-4b76-9196-b54776bbd237 · outbound
Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning Silver, J
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 582c930b-ebba-4c21-9bd8-a90a116ef71d · outbound
Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning Unresolved cited work
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 202465b4-947a-4c1e-bd6d-cac05b52a381 · outbound
Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bb2b12ed-f448-4500-8c4c-cf6cb2a2a9db · outbound
Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning AndroidEnv: A Reinforcement Learning Platform for Android
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 567c1408-798b-4872-ba04-3232fc222b34 · outbound
Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning DistRL: An Asynchronous Distributed Reinforcement Learning Framework for On-Device Control Agents
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b42ea7ae-1131-4da9-b935-0010228faa95 · outbound
Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning Variational Delayed Policy Optimization
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 7515b4c1-1019-4a5f-986d-22fcb949a0dc · outbound
Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning Auto-GPT for Online Decision Making: Benchmarks and Additional Opinions
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bc0803b3-03aa-4096-ac07-a215800f9cf2 · outbound
Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning Set-of-Mark Prompting Unleashes Extraordinary Visual Grounding in GPT-4V
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 61a16baa-4b4f-4967-91f0-ffe0c112cd86 · outbound
Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning Guiding Long-Horizon Task and Motion Planning with Vision Language Models
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d2ed289d-5155-40c0-b929-36367458568b · outbound
Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning AppAgent: Multimodal Agents as Smartphone Users
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a17b7127-cfaf-4dc4-94d9-9d3551404b2c · outbound
Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning You Only Look at Screens: Multimodal Chain-of-Action Agents
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 807ca801-9339-4c20-8842-6922488bfbfc · outbound
Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning EPO: Hierarchical LLM Agents with Environment Preference Optimization
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 92fe8298-9bb5-4372-8eb4-ec63f69a5133 · outbound
Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning GPT-4V(ision) is a Generalist Web Agent, if Grounded
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 311eadef-630c-46c7-897d-10395bb54b1d · outbound
Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning WebArena: A Realistic Web Environment for Building Autonomous Agents
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 87c60618-e859-47a4-b8eb-cff15b0e3c29 · outbound
Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning Unresolved cited work
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation f2008d08-eb33-4c26-be9b-e403c122ee27 · outbound
Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning Unresolved cited work
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation b6931b3b-4451-40ec-a80b-30339a29d45c · outbound
Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning Unresolved cited work
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation ec9c52ae-6f6e-4974-8959-0018819e9ce2 · outbound
Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning Unresolved cited work
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation abd9c80c-484e-48d0-a7e8-696dae34f0c3 · outbound
Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning The generator decomposes the goal of navigating the maze into subgoals like opening specific doors sequentially
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 8dc86540-026a-4449-8965-31049f5667c0 · inbound
Atomic-to-Compositional Generalization for Mobile Agents with A New Benchmark and Scheduling System Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 35fbfc91-c98e-4d06-ba1e-6c53e24663ec · inbound
Software Engineering for and with GUI Agent Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning
Reference 247
Source-reported events for the cited work
Unavailable: canonical work link unavailable.