Pith. sign in

Paper Citation Record · LEDGER

Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning

As of 19 August 2026, this Paper Citation Record lists 49 of 49 outbound references and 2 inbound Pith citation observations for arXiv:2502.07949.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.07949 v2

Coverage vector

measured 49 of 49 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-08T11:25:49.254278Z

measured 51 of 51 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-11T20:19:16.013755Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T05:04:05.836097Z

Reference resolution

49 of 49 outbound references displayed

  • verified exact2
  • verified fuzzy10
  • unresolved35
  • parse uncertain2
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation db6d4381-44e3-4be6-89e6-1a7aff1885ea · outbound

This paper cites Relative Entropy Regularized Policy Iteration.

Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning Relative Entropy Regularized Policy Iteration

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-08T11:25:49.077848Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:25:49.077848Z digest=sha256:0ad016b997a8e41ebef370a8d5dd056cf988c21f2d290542e600cf19fc5048c4

Observation ad348645-6fc5-4b15-a8d7-52f35171101e · outbound

This paper cites Maximum a Posteriori Policy Optimisation.

Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning Maximum a Posteriori Policy Optimisation

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-08T11:25:49.082805Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:25:49.082805Z digest=sha256:2c02014400962ca1a81c2910e35d2192dd0b6ab37c7d7751d220a0d2a1487ccb

Observation afb04b57-749f-4a82-bae2-30382afbd04b · outbound

This paper cites Andrychowicz, F.

Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning Andrychowicz, F

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T11:25:49.747985Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-08T11:25:49.086838Z digest=sha256:d081788986a608da52aa41257c6c5c058b6a110467664bd1e50c7f456cc93af5

Observation 940735a0-7fbc-4adb-a395-994715bfa4f0 · outbound

This paper cites DigiRL: Training In-The-Wild Device-Control Agents with Autonomous Reinforcement Learning.

Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning DigiRL: Training In-The-Wild Device-Control Agents with Autonomous Reinforcement Learning

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-08T11:25:49.090555Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:25:49.090555Z digest=sha256:f0658c9d34c801a5c40191c3b511085e4966cf38a3449dc2d7f66bd978fa186c

Observation e46950f5-deb2-461b-b25e-8f166c7fb301 · outbound

This paper cites Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond.

Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-08T11:25:49.094639Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:25:49.094639Z digest=sha256:97228fe9b3f050d4aaf55f4688e2795a2d2c632a125cbecac53c2c977b9e2dc4

Observation ab3c5455-2785-453a-a08c-15817c87c355 · outbound

This paper cites Dota 2 with Large Scale Deep Reinforcement Learning.

Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning Dota 2 with Large Scale Deep Reinforcement Learning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-08T11:25:49.098504Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:25:49.098504Z digest=sha256:7b97ce2b42952427d0679f9c8c3b82d94550a50b239e8cc1e70b1e6f766006a6

Observation f50414d4-631b-4de0-babf-65f21a162716 · outbound

This paper cites Chane-Sane, C.

Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning Chane-Sane, C

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T11:25:49.738098Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-08T11:25:49.102552Z digest=sha256:41a77879fcaee69d6faab07cb9f1c69c6f4c101e07ef829320e645dc4a38c440

Observation 9a0861c7-54df-4708-b225-9e2f16cd2e1e · outbound

This paper cites an unresolved cited work.

Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-08T11:25:49.728264Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-08T11:25:49.106133Z digest=sha256:6ddbfa05f12f6fe2723b83b969c2a93322fa56c1ebf61484272ab422bb8308c1

Observation 01e8418b-e7c6-4b6d-9364-cc4c414ab025 · outbound

This paper cites Chevalier-Boisvert, D.

Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning Chevalier-Boisvert, D

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T11:25:49.718546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-08T11:25:49.109437Z digest=sha256:60ae1bfafe260ef3d77e5c30a451c8110301f6c75206e3c52dfa1fd791e26fc8

Observation 6ca99384-01e5-4238-8fdf-f7fe09267918 · outbound

This paper cites Dayan and G.

Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning Dayan and G

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T11:25:49.708639Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-08T11:25:49.112208Z digest=sha256:7df01c650cea232828fb64946775add6b3b9eb351170ff844194f1b3b17a884d

Observation 170d37da-5366-418e-a2e8-e0df0d179100 · outbound

This paper cites MetaGPT: Meta Programming for A Multi-Agent Collaborative Framework.

Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning MetaGPT: Meta Programming for A Multi-Agent Collaborative Framework

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-08T11:25:49.114831Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:25:49.114831Z digest=sha256:cc19f52e79cd339c98e110bb8932c1ad96e3bcc1cb0485b3089465db6aaeb37b

Observation df641b50-b2f7-4fde-8402-03936d23a97d · outbound

This paper cites an unresolved cited work.

Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning Unresolved cited work

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-08T11:25:49.117896Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:25:49.117896Z digest=sha256:d6c6d30b3aa3c3bdf87b98778cd6226f7633182ab0b4fe9252d842fc31877831

Observation 2183db8f-1589-4447-86ce-66fdffe95b51 · outbound

This paper cites Jiang and A.

Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning Jiang and A

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T11:25:49.692268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-08T11:25:49.120677Z digest=sha256:09de4d97886d094c627a7be1f3916dee8b8606381f731822637fad68902f447d

Observation a53ce53f-bd3a-46cb-93a2-7073fe2b8ebf · outbound

This paper cites Jurgenson, O.

Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning Jurgenson, O

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T11:25:49.682754Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-08T11:25:49.123476Z digest=sha256:6c518a3360e827408d0541bd0d8ddd0cf70d75add3b948a8a879e20d42539c86

Observation f278e4ea-5612-4a4e-a7df-e544b1fb432a · outbound

This paper cites Reinforcement Learning and Control as Probabilistic Inference: Tutorial and Review.

Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning Reinforcement Learning and Control as Probabilistic Inference: Tutorial and Review

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-08T11:25:49.126140Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:25:49.126140Z digest=sha256:be8e3afe312913807a478d9e8330288a7eb4be50c210c3a1fb8134a8d7319725

Observation 106e6232-3df4-4967-9c6b-b83b751e6ca6 · outbound

This paper cites Goal-Conditioned Reinforcement Learning: Problems and Solutions.

Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning Goal-Conditioned Reinforcement Learning: Problems and Solutions

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-08T11:25:49.129256Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:25:49.129256Z digest=sha256:d6441ad035aba3ffec23edcec1b54dbb21b86d263ccc56ebf0fe4a10e2a4937c

Observation acf8ca8a-b7fe-472b-84aa-d2d0445cd6a6 · outbound

This paper cites VisualAgentBench: Towards Large Multimodal Models as Visual Foundation Agents.

Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning VisualAgentBench: Towards Large Multimodal Models as Visual Foundation Agents

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-08T11:25:49.133307Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:25:49.133307Z digest=sha256:d28feebbbbbdcd2bf4435b35f8c51799cdabd325974ee03241854a2779f250b2

Observation cfdbcdd3-fcaf-4b3f-b471-1f435bcbb150 · outbound

This paper cites an unresolved cited work.

Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-08T11:25:49.673067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-08T11:25:49.137231Z digest=sha256:9f6b54135ef3887f3ddec55eaa0b2ad90c48815c603ce37c89fce66572747d5c

Observation 4b6e6fa3-0b7b-4c50-9300-2a45812dfec0 · outbound

This paper cites Hierarchical Foresight: Self-Supervised Learning of Long-Horizon Tasks via Visual Subgoal Generation.

Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning Hierarchical Foresight: Self-Supervised Learning of Long-Horizon Tasks via Visual Subgoal Generation

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-08T11:25:49.140608Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:25:49.140608Z digest=sha256:47fc6497d95a8b64988318d23a69a266e91c2d615577dcc0068d59a754e2c5fd

Observation 57ea3c29-b182-46d1-b602-baddc31d951c · outbound

This paper cites an unresolved cited work.

Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-08T11:25:49.663276Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-08T11:25:49.144215Z digest=sha256:22ac06e6e6a803e2ac0252f24514aaf6fbf11304ee95039692d8239b0c367619

Observation 210d1af5-1431-4db5-82c3-8d2e472ba75a · outbound

This paper cites Gpt-4v(ision) technical work and authors.

Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning Gpt-4v(ision) technical work and authors

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T11:25:49.653686Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-08T11:25:49.147554Z digest=sha256:6c38b0df39910e90dea955597ea20d6a32f3e3ad4ff74c26d1f0106e06150869

Observation 9791b529-d7f0-4d83-bff0-7cc2ab1f5821 · outbound

This paper cites an unresolved cited work.

Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning Unresolved cited work

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-08T11:25:49.150969Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:25:49.150969Z digest=sha256:15b4e85a733e272882e5914567dd7f15dc2c4e868803172aa8ac34a605f1bbb7

Observation edc6f4e7-e33c-4748-b23e-10fe327137a3 · outbound

This paper cites an unresolved cited work.

Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-08T11:25:49.637791Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-08T11:25:49.154436Z digest=sha256:b11f38817efc86bd88616b89d1df77b6ab6a16c5da4c90ac84d226b5b51fa76f

Observation 1a828d38-91fb-433c-8a94-c1d945a87768 · outbound

This paper cites Divide-and-Conquer Monte Carlo Tree Search For Goal-Directed Planning.

Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning Divide-and-Conquer Monte Carlo Tree Search For Goal-Directed Planning

Reference 24

Resolution
verified exact
local_arxiv, observed 2026-08-08T11:25:49.440516Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-08T11:25:49.157998Z digest=sha256:e13febdaf71c1365ecb4fc7bf3c786faaa6a162dfb1a5de6462fe69bf8b8f8f1

Observation 7a761859-ce8f-4eb8-9b89-becb5342b309 · outbound

This paper cites Advantage-Weighted Regression: Simple and Scalable Off-Policy Reinforcement Learning.

Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning Advantage-Weighted Regression: Simple and Scalable Off-Policy Reinforcement Learning

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-08T11:25:49.161620Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:25:49.161620Z digest=sha256:35f60b470629dab2d62dd1915c759f808b36206d686f9ed3c268e6ec38d040e1

Observation 3d05642c-2028-4729-99d8-165576046229 · outbound

This paper cites WebRL: Training LLM Web Agents via Self-Evolving Online Curriculum Reinforcement Learning.

Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning WebRL: Training LLM Web Agents via Self-Evolving Online Curriculum Reinforcement Learning

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-08T11:25:49.165415Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:25:49.165415Z digest=sha256:dfe50ed47b7d99bb242d1e885db2ad1a3b40f375ee572468cc9732e03ef03715

Observation 854c3f89-17f3-49bf-89f8-8ad27d167861 · outbound

This paper cites Rawles, A.

Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning Rawles, A

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T11:25:49.628464Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-08T11:25:49.168973Z digest=sha256:89d5c208978f2d3d6750574a3cdc8864e82870e53343def44d46427e4cc9565a

Observation c0df10cc-6fa3-4edf-bb43-6e49a0db3b95 · outbound

This paper cites Vision-Language Models are Zero-Shot Reward Models for Reinforcement Learning.

Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning Vision-Language Models are Zero-Shot Reward Models for Reinforcement Learning

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-08T11:25:49.172358Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:25:49.172358Z digest=sha256:9ca200721c55ae0b983550bb14bd9c5f56178bebe29872051c5b374a9508df7c

Observation 9a24d763-76fb-4fad-b423-68a61695c180 · outbound

This paper cites Proximal Policy Optimization Algorithms.

Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning Proximal Policy Optimization Algorithms

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-08T11:25:49.176014Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:25:49.176014Z digest=sha256:c17953d6fc873fc5b2438f4c8d00253aaca9a55f09beeeba5f7f44579d5e5771

Observation b8cc4d12-7acf-43ea-a751-20b619dc1bdf · outbound

This paper cites an unresolved cited work.

Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-08T11:25:49.618557Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-08T11:25:49.179540Z digest=sha256:febf974d891120f3ac100ba2cfd9104a3b21b43d7d6025b6f553beb97b0e3d09

Observation 7a76d994-278f-4b76-9196-b54776bbd237 · outbound

This paper cites Silver, J.

Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning Silver, J

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T11:25:49.609706Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-08T11:25:49.182692Z digest=sha256:d4cf8a74669bf378b3951973f3bccf4f7618e144309c3bed3b149cd4655d8293

Observation 582c930b-ebba-4c21-9bd8-a90a116ef71d · outbound

This paper cites an unresolved cited work.

Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-08T11:25:49.600811Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-08T11:25:49.186042Z digest=sha256:0cdfef593623afb5aea0cb9c2d7cb696884e527871724e48a8ff23f3d05c27fe

Observation 202465b4-947a-4c1e-bd6d-cac05b52a381 · outbound

This paper cites Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context.

Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-08T11:25:49.189339Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:25:49.189339Z digest=sha256:10569787e7b14df0f7624eb3dec2de5507e2ac70d851a1d483d416e3d534625c

Observation bb2b12ed-f448-4500-8c4c-cf6cb2a2a9db · outbound

This paper cites AndroidEnv: A Reinforcement Learning Platform for Android.

Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning AndroidEnv: A Reinforcement Learning Platform for Android

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-08T11:25:49.192822Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:25:49.192822Z digest=sha256:39f87e912dfa6fc178805871d76e4b631e5e8a37fcfcb4113b33bd560e75eb29

Observation 567c1408-798b-4872-ba04-3232fc222b34 · outbound

This paper cites DistRL: An Asynchronous Distributed Reinforcement Learning Framework for On-Device Control Agents.

Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning DistRL: An Asynchronous Distributed Reinforcement Learning Framework for On-Device Control Agents

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-08T11:25:49.196216Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:25:49.196216Z digest=sha256:a6d4de0e186aaf5eb1ba3dd01ade4ece1c4c215305664f07cee1467331d7bb3d

Observation b42ea7ae-1131-4da9-b935-0010228faa95 · outbound

This paper cites Variational Delayed Policy Optimization.

Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning Variational Delayed Policy Optimization

Reference 36

Resolution
verified exact
local_arxiv, observed 2026-08-08T11:25:49.360722Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-08T11:25:49.199644Z digest=sha256:66b894086b4f3af5177a5084f9ee23982dea6cf97defa97d179a6ffb9eccc2d8

Observation 7515b4c1-1019-4a5f-986d-22fcb949a0dc · outbound

This paper cites Auto-GPT for Online Decision Making: Benchmarks and Additional Opinions.

Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning Auto-GPT for Online Decision Making: Benchmarks and Additional Opinions

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-08T11:25:49.203048Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:25:49.203048Z digest=sha256:e38e9bedeb66a22bc7d3eeca8fc0e35622a7bbd144d828c1f7a61f5da6c968d6

Observation bc0803b3-03aa-4096-ac07-a215800f9cf2 · outbound

This paper cites Set-of-Mark Prompting Unleashes Extraordinary Visual Grounding in GPT-4V.

Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning Set-of-Mark Prompting Unleashes Extraordinary Visual Grounding in GPT-4V

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-08T11:25:49.206333Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:25:49.206333Z digest=sha256:7f7bca0d76639072cc44d752f71e53b6a2eeb3e75bfc45ff321fe2be1baf93d1

Observation 61a16baa-4b4f-4967-91f0-ffe0c112cd86 · outbound

This paper cites Guiding Long-Horizon Task and Motion Planning with Vision Language Models.

Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning Guiding Long-Horizon Task and Motion Planning with Vision Language Models

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-08T11:25:49.209337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:25:49.209337Z digest=sha256:aef3a781ba30b483fe1048eea0dde033746f4bb27cc81cede54bbb719de2048d

Observation d2ed289d-5155-40c0-b929-36367458568b · outbound

This paper cites AppAgent: Multimodal Agents as Smartphone Users.

Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning AppAgent: Multimodal Agents as Smartphone Users

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-08T11:25:49.215278Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:25:49.215278Z digest=sha256:802eaa736a609796677da3119d802387cfa92a12b0e62d24e0df05ba076943a7

Observation a17b7127-cfaf-4dc4-94d9-9d3551404b2c · outbound

This paper cites You Only Look at Screens: Multimodal Chain-of-Action Agents.

Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning You Only Look at Screens: Multimodal Chain-of-Action Agents

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-08T11:25:49.218075Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:25:49.218075Z digest=sha256:181154cc0c66955b22b019814a3d622ca85e749a3a82441221241fe942f8ad77

Observation 807ca801-9339-4c20-8842-6922488bfbfc · outbound

This paper cites EPO: Hierarchical LLM Agents with Environment Preference Optimization.

Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning EPO: Hierarchical LLM Agents with Environment Preference Optimization

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-08T11:25:49.220959Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:25:49.220959Z digest=sha256:a309fad2cca7a854e61c7cb20bfeeb7cc137336627ffe00ded6fda6e64e2a2a5

Observation 92fe8298-9bb5-4372-8eb4-ec63f69a5133 · outbound

This paper cites GPT-4V(ision) is a Generalist Web Agent, if Grounded.

Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning GPT-4V(ision) is a Generalist Web Agent, if Grounded

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-08T11:25:49.224379Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:25:49.224379Z digest=sha256:5782537a0cab8de72f01f27dc404f943a34eb60ee1b41a1c6cb08e900b36ded7

Observation 311eadef-630c-46c7-897d-10395bb54b1d · outbound

This paper cites WebArena: A Realistic Web Environment for Building Autonomous Agents.

Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning WebArena: A Realistic Web Environment for Building Autonomous Agents

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-08T11:25:49.227821Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:25:49.227821Z digest=sha256:efd375366b27545085a425a6596c20dd872aed83af8239e25f1765131461c9eb

Observation 87c60618-e859-47a4-b8eb-cff15b0e3c29 · outbound

This paper cites an unresolved cited work.

Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning Unresolved cited work

Reference 46

Resolution
unresolved
raw_fallback, observed 2026-08-08T11:25:49.590539Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-08T11:25:49.231350Z digest=sha256:7e8f83e0d97322ad4897420befb6e8b6c14a07edddaa2cbc168d41d7857d194b

Observation f2008d08-eb33-4c26-be9b-e403c122ee27 · outbound

This paper cites an unresolved cited work.

Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning Unresolved cited work

Reference 50

Resolution
unresolved
raw_fallback, observed 2026-08-08T11:25:49.559547Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-08T11:25:49.244728Z digest=sha256:4aa8f74120077d92d3013ac9cbe1a7817f7563916247a830fa70af4cbc9099f5

Observation b6931b3b-4451-40ec-a80b-30339a29d45c · outbound

This paper cites an unresolved cited work.

Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning Unresolved cited work

Reference 51

Resolution
parse uncertain
raw_fallback, observed 2026-08-08T11:25:49.570519Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-08T11:25:49.248056Z digest=sha256:01a0990d0944040a1b586d745b7d0620fa46eb7707c50f38808a356e2cd8a5ac

Observation ec9c52ae-6f6e-4974-8959-0018819e9ce2 · outbound

This paper cites an unresolved cited work.

Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning Unresolved cited work

Reference 52

Resolution
parse uncertain
raw_fallback, observed 2026-08-08T11:25:49.581345Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-08T11:25:49.251212Z digest=sha256:a6f8523a464e00276acf782eb2699037aec5495b5dea72fa4b915834e6076218

Observation abd9c80c-484e-48d0-a7e8-696dae34f0c3 · outbound

This paper cites The generator decomposes the goal of navigating the maze into subgoals like opening specific doors sequentially.

Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning The generator decomposes the goal of navigating the maze into subgoals like opening specific doors sequentially

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T11:25:49.548407Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-08T11:25:49.254278Z digest=sha256:ef6ef781cefd369c53eb58660f8dd480b69381eaf9fd7a62f72e45efe4732695

Pith citing papers

Observation 8dc86540-026a-4449-8965-31049f5667c0 · inbound

Atomic-to-Compositional Generalization for Mobile Agents with A New Benchmark and Scheduling System cites this paper.

Atomic-to-Compositional Generalization for Mobile Agents with A New Benchmark and Scheduling System Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning

Reference 59

Resolution
verified exact
local_arxiv, observed 2026-08-07T05:04:05.841970Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T05:04:05.548532Z digest=sha256:bc22d3ae3c47beb82e1220f5a3b88d864d5cfaeda4e3fd8047cd333654ca8c95

Observation 35fbfc91-c98e-4d06-ba1e-6c53e24663ec · inbound

Software Engineering for and with GUI Agent cites this paper.

Software Engineering for and with GUI Agent Advancing Autonomous VLM Agents via Variational Subgoal-Conditioned Reinforcement Learning

Reference 247

Resolution
unresolved
no resolver link, observed 2026-08-11T20:19:16.013755Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:19:16.013755Z digest=sha256:0be3021deeda1560e46253d809b9cdc42864fed4c806c513e54e76a46d9002b6