Pith. sign in

Paper Citation Record · LEDGER

SafeDojo: Safe Reinforcement Learning for VLA via Interactive World Model

As of 25 July 2026, this Paper Citation Record lists 46 of 46 outbound references and 0 inbound Pith citation observations for arXiv:2606.20698.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2606.20698 v1

Coverage vector

measured 46 of 46 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-06-27T04:13:22.598591Z

measured 46 of 46 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-07-25T06:30:59.84592+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

46 of 46 outbound references displayed

  • verified exact24
  • verified fuzzy0
  • unresolved18
  • parse uncertain0
  • malformed identifier2
  • metadata mismatch2

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 9431746a-41ec-4c11-bb62-00c893d0a1c3 · outbound

This paper cites Brunke, M.

SafeDojo: Safe Reinforcement Learning for VLA via Interactive World Model Brunke, M

Reference 1

Resolution
unresolved
no resolver link, observed 2026-06-27T04:13:22.598591Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T04:13:22.598591Z digest=sha256:3db47c6fb5d02a46bf5e4804a6bbdf00b3425ef96924971deebba2e3b6c76897

Observation e7d69a8d-1ceb-4808-b041-9641ede98884 · outbound

This paper cites an unresolved cited work.

SafeDojo: Safe Reinforcement Learning for VLA via Interactive World Model Unresolved cited work

Reference 2

Resolution
unresolved
no resolver link, observed 2026-06-27T04:13:22.598591Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T04:13:22.598591Z digest=sha256:b274b1ba75f5394816a7cd4648652a98882173f82268fc6a2632ebe0616e8ef5

Observation 081049cf-0812-40d0-beb8-9c4ced69abb5 · outbound

This paper cites Real-time obstacle avoidance for manipulators and mobile robots.

SafeDojo: Safe Reinforcement Learning for VLA via Interactive World Model Real-time obstacle avoidance for manipulators and mobile robots

Reference 3

Resolution
verified exact
doi, observed 2026-06-27T04:20:31.946980Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=pdf_text observed=2026-06-27T04:13:22.598591Z digest=sha256:927eb7137385360c7edb26b1717a808baa3bb8bff60bb918e70224860f72a45f

Observation 4feaa642-b149-4447-9e9f-278899e9ec8c · outbound

This paper cites Li, 3d fully convolutional network for vehicle de- tection in point cloud, in: 2017 IEEE/RSJ Interna- tional Conference on Intelligent Robots and Systems (IROS), 2017, pp.

SafeDojo: Safe Reinforcement Learning for VLA via Interactive World Model Li, 3d fully convolutional network for vehicle de- tection in point cloud, in: 2017 IEEE/RSJ Interna- tional Conference on Intelligent Robots and Systems (IROS), 2017, pp

Reference 4

Resolution
malformed identifier
doi_truncated, observed 2026-06-27T04:20:31.944810Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=pdf_text observed=2026-06-27T04:13:22.598591Z digest=sha256:874b32dfe948d56859b8aa18ada150eb8bd81a3112b06f722e35bfecb92c2258

Observation a7107b95-7ac7-4407-a392-fe203eef57b7 · outbound

This paper cites Haddadin, A.

SafeDojo: Safe Reinforcement Learning for VLA via Interactive World Model Haddadin, A

Reference 5

Resolution
verified exact
doi, observed 2026-06-27T04:20:31.964209Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=pdf_text observed=2026-06-27T04:13:22.598591Z digest=sha256:b02964b771bd29a4ec059aff08e2507a2d6e273347b445449c00044ae0ce1026

Observation d8156a4b-56fd-49b8-92f0-4e8fc73e2787 · outbound

This paper cites Haddadin, A.

SafeDojo: Safe Reinforcement Learning for VLA via Interactive World Model Haddadin, A

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-06-27T04:20:31.939211Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=pdf_text observed=2026-06-27T04:13:22.598591Z digest=sha256:91a0bb870841ba53e6e83e7fa0e312b4015551f6f4a0c3a9983d98759af0a6e1

Observation fbd61de5-8d9e-4c7e-8766-f9512872b78e · outbound

This paper cites an unresolved cited work.

SafeDojo: Safe Reinforcement Learning for VLA via Interactive World Model Unresolved cited work

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-06-27T04:20:31.961105Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=pdf_text observed=2026-06-27T04:13:22.598591Z digest=sha256:1535e9cbef18846f8d19218cc74c8a511ad2131da795d3de5a8f5b7e41828d5f

Observation d96feb02-4fbf-4e2a-a794-8f2c3ced5ba8 · outbound

This paper cites an unresolved cited work.

SafeDojo: Safe Reinforcement Learning for VLA via Interactive World Model Unresolved cited work

Reference 8

Resolution
verified exact
doi, observed 2026-06-27T04:20:31.961942Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=pdf_text observed=2026-06-27T04:13:22.598591Z digest=sha256:6c0c2c202f1dfac8e3a93b1750e997412a1886ed601764aa67cb345ee73073cb

Observation 690dd619-0aaf-4e12-8548-6ae840d1e846 · outbound

This paper cites Safety-Critical Optimal Control for Robotic Manipulators in A Cluttered Environment.

SafeDojo: Safe Reinforcement Learning for VLA via Interactive World Model Safety-Critical Optimal Control for Robotic Manipulators in A Cluttered Environment

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-06-27T04:20:31.945990Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=pdf_text observed=2026-06-27T04:13:22.598591Z digest=sha256:439953d3d606ddbb0d3a37f0ecf364c93ac3e3ce05e5c3d690740691189c3d06

Observation 4cf2563e-1524-4238-b49e-67b6fee37f82 · outbound

This paper cites Brohan, N.

SafeDojo: Safe Reinforcement Learning for VLA via Interactive World Model Brohan, N

Reference 10

Resolution
unresolved
no resolver link, observed 2026-06-27T04:13:22.598591Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T04:13:22.598591Z digest=sha256:ea25549ce503d5e603bc8a699292f26b3c713369503094797c645f88996cec24

Observation c7aa011a-3189-45be-a222-b61f1a3fa284 · outbound

This paper cites Open X-Embodiment: Robotic Learning Datasets and RT-X Models.

SafeDojo: Safe Reinforcement Learning for VLA via Interactive World Model Open X-Embodiment: Robotic Learning Datasets and RT-X Models

Reference 11

Resolution
verified exact
local_arxiv, observed 2026-07-03T17:28:44.187967Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=pdf_text observed=2026-06-27T04:13:22.598591Z digest=sha256:f7874ae233e5af9005844dd9afdc0457561b479e79df7f1e221c340ff22ab533

Observation f190b701-ae5c-4a67-b909-cc7fdddd5661 · outbound

This paper cites OpenVLA: An Open-Source Vision-Language-Action Model.

SafeDojo: Safe Reinforcement Learning for VLA via Interactive World Model OpenVLA: An Open-Source Vision-Language-Action Model

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-07-03T17:18:44.352219Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=pdf_text observed=2026-06-27T04:13:22.598591Z digest=sha256:e8ee8efc3496edaf82e7ca4bc41ea65b7815a3af11b3b21da08fdc87e3a88c48

Observation 5b671430-935d-48e4-be65-c56bdb085dd7 · outbound

This paper cites Octo: An Open-Source Generalist Robot Policy.

SafeDojo: Safe Reinforcement Learning for VLA via Interactive World Model Octo: An Open-Source Generalist Robot Policy

Reference 13

Resolution
verified exact
local_arxiv, observed 2026-07-03T17:28:44.190678Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=pdf_text observed=2026-06-27T04:13:22.598591Z digest=sha256:45786243bc691ea35011d6bdccd3ed919e6f0706712cba195a19f0301e3417b2

Observation f93f6a6b-4df2-4ee6-b78c-690ad3ecce7d · outbound

This paper cites $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control.

SafeDojo: Safe Reinforcement Learning for VLA via Interactive World Model $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-07-03T17:18:44.362058Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=pdf_text observed=2026-06-27T04:13:22.598591Z digest=sha256:c145d4491c0b844ca49b7871133dafa010b12af66f38f9119060dcad19a58113

Observation 0c0fdb36-0512-451a-83c7-9addeafba095 · outbound

This paper cites Control barrier functions: Theory and ap- plications.

SafeDojo: Safe Reinforcement Learning for VLA via Interactive World Model Control barrier functions: Theory and ap- plications

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-06-27T04:20:31.956729Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=pdf_text observed=2026-06-27T04:13:22.598591Z digest=sha256:35716bce76994a3904cbcf5ab6d59a65082e2008c4ecff4b20eb8eb6d0bf91df

Observation a615724e-ec73-492d-923c-578fa19abfe0 · outbound

This paper cites Huang, J.

SafeDojo: Safe Reinforcement Learning for VLA via Interactive World Model Huang, J

Reference 16

Resolution
unresolved
no resolver link, observed 2026-06-27T04:13:22.598591Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T04:13:22.598591Z digest=sha256:e1acaf4a669743c460f6540468008cbacb2cbd9f43dd23bb8e40741d599458df

Observation 89d69aa1-a097-46d6-b76e-e7f23ae362a4 · outbound

This paper cites Zhang, Y.

SafeDojo: Safe Reinforcement Learning for VLA via Interactive World Model Zhang, Y

Reference 17

Resolution
unresolved
no resolver link, observed 2026-06-27T04:13:22.598591Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T04:13:22.598591Z digest=sha256:6459a0da158c6fd8fd601b150ca37a5b25052a878cabf7a6481dd3d95517fdaf

Observation ae2ae3bc-8d02-4adf-9f62-6c90b0cb4920 · outbound

This paper cites VLSA: Vision-Language-Action Models with Plug-and-Play Safety Constraint Layer.

SafeDojo: Safe Reinforcement Learning for VLA via Interactive World Model VLSA: Vision-Language-Action Models with Plug-and-Play Safety Constraint Layer

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-07-03T17:18:44.348813Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=pdf_text observed=2026-06-27T04:13:22.598591Z digest=sha256:a7f9688da1d4e107120f2a8a531edae06e1752fa4e0cbeef074d3f3bc2b6ebe7

Observation 100c7a0a-2906-4bc3-801b-0321524c0e02 · outbound

This paper cites an unresolved cited work.

SafeDojo: Safe Reinforcement Learning for VLA via Interactive World Model Unresolved cited work

Reference 19

Resolution
unresolved
no resolver link, observed 2026-06-27T04:13:22.598591Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T04:13:22.598591Z digest=sha256:1167dde14992ca66bd0452b9d1334b6308fc9819c196159f4a8c26ae78301d40

Observation b1c9a51c-77bd-4dc7-83ee-5893cc069275 · outbound

This paper cites Safe-Night VLA: Seeing the Unseen via Thermal-Perceptive Vision-Language-Action Models for Safety-Critical Manipulation.

SafeDojo: Safe Reinforcement Learning for VLA via Interactive World Model Safe-Night VLA: Seeing the Unseen via Thermal-Perceptive Vision-Language-Action Models for Safety-Critical Manipulation

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-07-17T02:21:32.193530Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=pdf_text observed=2026-06-27T04:13:22.598591Z digest=sha256:eca2a904b39ace6eb2090f21f1a26a7c201f7cd1da9abde6700ffda439043d84

Observation b05cc9f2-d1d9-45ed-9aab-8e4de26304e1 · outbound

This paper cites Son, D.-K.

SafeDojo: Safe Reinforcement Learning for VLA via Interactive World Model Son, D.-K

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-06-27T04:20:31.966668Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=pdf_text observed=2026-06-27T04:13:22.598591Z digest=sha256:a7cbabde4e8690277de05258a02846d079035db3a64f30f70da2d8ea4fa4006a

Observation 903bb4bb-4e7b-4474-8b61-31feec2dfd03 · outbound

This paper cites Cofreevla: Collision-free dual-arm manipulation via vision-language-action model and risk estimation.

SafeDojo: Safe Reinforcement Learning for VLA via Interactive World Model Cofreevla: Collision-free dual-arm manipulation via vision-language-action model and risk estimation

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-07-03T17:18:44.360795Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=pdf_text observed=2026-06-27T04:13:22.598591Z digest=sha256:a063557644fdb1d2f927400787aa5e5be956b74c6f912c8abb577f3952aa2635

Observation 8b3b97fc-7017-4821-9172-561aa7638066 · outbound

This paper cites Pranav Guruprasad, Harshvardhan Sikka, Jaewoo Song, Yangyue Wang, and Paul Pu Liang.

SafeDojo: Safe Reinforcement Learning for VLA via Interactive World Model Pranav Guruprasad, Harshvardhan Sikka, Jaewoo Song, Yangyue Wang, and Paul Pu Liang

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-06-27T04:20:31.953710Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=pdf_text observed=2026-06-27T04:13:22.598591Z digest=sha256:0721c1080f94af8901c0662a88c30a125781771a9de150634ed3d34eea5f50e4

Observation 62cfe2bf-a65b-4d10-9fe0-c2e626669ab7 · outbound

This paper cites GRAPE: Generalizing Robot Policy via Preference Alignment.

SafeDojo: Safe Reinforcement Learning for VLA via Interactive World Model GRAPE: Generalizing Robot Policy via Preference Alignment

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-07-03T17:18:44.353030Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=pdf_text observed=2026-06-27T04:13:22.598591Z digest=sha256:66d40a9a10922a9fc564f1b3baa06768306ea0264fc472bef8abfa0a7e01401f

Observation dd067f52-1652-4806-8e97-4d554a3ac16d · outbound

This paper cites GRAPE: Generalizing Robot Policy via Preference Alignment.

SafeDojo: Safe Reinforcement Learning for VLA via Interactive World Model GRAPE: Generalizing Robot Policy via Preference Alignment

Reference 25

Resolution
metadata mismatch
arxiv_id, observed 2026-06-27T04:20:31.958581Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=pdf_text observed=2026-06-27T04:13:22.598591Z digest=sha256:1e32584c4164aadd8694b9ec1ea5169e79d541e8c0d3c3f563940dc38d5c723c

Observation 9893b730-fd55-465c-a4cf-af9022f63095 · outbound

This paper cites Altman.Constrained Markov Decision Processes.

SafeDojo: Safe Reinforcement Learning for VLA via Interactive World Model Altman.Constrained Markov Decision Processes

Reference 26

Resolution
unresolved
no resolver link, observed 2026-06-27T04:13:22.598591Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T04:13:22.598591Z digest=sha256:754b15625667c24fd5604f43c897b76c46a09f715aaff2a6b87f14c82ac99c89

Observation c41f090a-4f76-4b43-9b4e-c2b2e775c617 · outbound

This paper cites Achiam, D.

SafeDojo: Safe Reinforcement Learning for VLA via Interactive World Model Achiam, D

Reference 27

Resolution
unresolved
no resolver link, observed 2026-06-27T04:13:22.598591Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T04:13:22.598591Z digest=sha256:76b20ff8420af5df9e9e2648447460b32eb2d1157a1586514116b470e1dcf0e4

Observation 15830582-5561-479a-aa0c-904b2363af54 · outbound

This paper cites an unresolved cited work.

SafeDojo: Safe Reinforcement Learning for VLA via Interactive World Model Unresolved cited work

Reference 28

Resolution
unresolved
no resolver link, observed 2026-06-27T04:13:22.598591Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T04:13:22.598591Z digest=sha256:5ff54ae25a53c0fb3ed6bc62d68db7fe382059440442d1546445a0996323a1ae

Observation fd56d5c4-11b4-4dd7-9715-a14a2e4d7e1e · outbound

This paper cites Tessler, D.

SafeDojo: Safe Reinforcement Learning for VLA via Interactive World Model Tessler, D

Reference 29

Resolution
unresolved
no resolver link, observed 2026-06-27T04:13:22.598591Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T04:13:22.598591Z digest=sha256:07bf8f3fe485ea5e6dd89886122c3317615389e5b134ef643f8873913e236259

Observation 317fc33f-8ea0-4ba2-a4dd-5c3f78f439ba · outbound

This paper cites Zhang, Q.

SafeDojo: Safe Reinforcement Learning for VLA via Interactive World Model Zhang, Q

Reference 30

Resolution
unresolved
no resolver link, observed 2026-06-27T04:13:22.598591Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T04:13:22.598591Z digest=sha256:4949a42d33ee2bbacac74202d0c0ea2f28fa5048041aca9b071b3e2a87199dfc

Observation 38c8e800-49ad-4410-b39c-08fa4f35e556 · outbound

This paper cites Stooke, J.

SafeDojo: Safe Reinforcement Learning for VLA via Interactive World Model Stooke, J

Reference 31

Resolution
unresolved
no resolver link, observed 2026-06-27T04:13:22.598591Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T04:13:22.598591Z digest=sha256:68f5229a5989f231a1320ef9ee3ae7276c6670bcf5f91bd0661cab42d310642e

Observation a3352358-5c02-4ed6-b8ba-2a3f725fab1a · outbound

This paper cites Thomas, Y.

SafeDojo: Safe Reinforcement Learning for VLA via Interactive World Model Thomas, Y

Reference 32

Resolution
unresolved
no resolver link, observed 2026-06-27T04:13:22.598591Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T04:13:22.598591Z digest=sha256:76b36358367e1146ffca59092cc634f230e68c2125cd70a783529780c64b724e

Observation 2091cfbc-367d-4a83-9374-70a5c38c6217 · outbound

This paper cites Hogewind, T.

SafeDojo: Safe Reinforcement Learning for VLA via Interactive World Model Hogewind, T

Reference 33

Resolution
unresolved
no resolver link, observed 2026-06-27T04:13:22.598591Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T04:13:22.598591Z digest=sha256:bfe03ef9fc59b2fb334e341c4846cd944e3f0524a10f621517323db92e51fe0e

Observation ea4024e8-8ef0-47d5-a226-475ad721e258 · outbound

This paper cites Generalizing safety beyond collision-avoidance via latent- space reachability analysis.

SafeDojo: Safe Reinforcement Learning for VLA via Interactive World Model Generalizing safety beyond collision-avoidance via latent- space reachability analysis

Reference 34

Resolution
verified exact
doi, observed 2026-06-27T04:20:31.950974Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=pdf_text observed=2026-06-27T04:13:22.598591Z digest=sha256:5604bd8e9565e9366f2c12e7a89875670227a240b9b0a5a26a049dca506da430

Observation a82fa073-f547-4f52-829e-63c6c9e0dbd5 · outbound

This paper cites an unresolved cited work.

SafeDojo: Safe Reinforcement Learning for VLA via Interactive World Model Unresolved cited work

Reference 35

Resolution
unresolved
no resolver link, observed 2026-06-27T04:13:22.598591Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T04:13:22.598591Z digest=sha256:569027d3cf2282ab0dc665cd16b57cf0b700ece4eb85aea5d276b971e7c4565f

Observation 8a49b8e3-a146-4f4e-82ef-62fda9c106bc · outbound

This paper cites an unresolved cited work.

SafeDojo: Safe Reinforcement Learning for VLA via Interactive World Model Unresolved cited work

Reference 36

Resolution
unresolved
no resolver link, observed 2026-06-27T04:13:22.598591Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T04:13:22.598591Z digest=sha256:c0c436c2817044aa7b3ee9758204771319d19cb0737a8fc251c0b08a4d462500

Observation ed0564d4-9380-49df-ad19-fdd27c9db8dd · outbound

This paper cites WoVR: World Models as Reliable Simulators for Post-Training VLA Policies with RL.

SafeDojo: Safe Reinforcement Learning for VLA via Interactive World Model WoVR: World Models as Reliable Simulators for Post-Training VLA Policies with RL

Reference 37

Resolution
verified exact
local_arxiv, observed 2026-07-03T17:18:44.349846Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=pdf_text observed=2026-06-27T04:13:22.598591Z digest=sha256:b1139f102f4bad3d8d347dc99d5b7bcf984145ec79d30333426a16d6282eb4b1

Observation 37a32824-084f-4ae0-9251-8509f1759774 · outbound

This paper cites World-Env: Leveraging World Model as a Virtual Environment for VLA Post-Training.

SafeDojo: Safe Reinforcement Learning for VLA via Interactive World Model World-Env: Leveraging World Model as a Virtual Environment for VLA Post-Training

Reference 38

Resolution
verified exact
local_arxiv, observed 2026-07-03T17:18:44.356532Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=pdf_text observed=2026-06-27T04:13:22.598591Z digest=sha256:ba6674316c1b1cc4c412353dad16092d2bfa35eac44c7553056bfc559fcb4423

Observation a14210ec-0b4b-4428-9e74-1c03ea1775c2 · outbound

This paper cites Vla-rft: Vision- language-action reinforcement fine-tuning with veri- fied rewards in world simulators.

SafeDojo: Safe Reinforcement Learning for VLA via Interactive World Model Vla-rft: Vision- language-action reinforcement fine-tuning with veri- fied rewards in world simulators

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-06-27T04:20:31.969644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=pdf_text observed=2026-06-27T04:13:22.598591Z digest=sha256:5ce7e1a9423fbb7b634a447101d476d2406b3784a4c15845b065e38d331e9e83

Observation 42a2284b-a6ff-432c-b48a-29e11db0b450 · outbound

This paper cites World-VLA-Loop: Closed-Loop Learning of Video World Model and VLA Policy.

SafeDojo: Safe Reinforcement Learning for VLA via Interactive World Model World-VLA-Loop: Closed-Loop Learning of Video World Model and VLA Policy

Reference 40

Resolution
malformed identifier
local_arxiv, observed 2026-07-03T17:18:44.363907Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=pdf_text observed=2026-06-27T04:13:22.598591Z digest=sha256:d6db54f8c1cb381daed34b69ce2369c1d448b5148807e55563b8a75f27b053e5

Observation a0a980f1-3237-45c4-8161-26e3128bfc6b · outbound

This paper cites Wan: Open and Advanced Large-Scale Video Generative Models.

SafeDojo: Safe Reinforcement Learning for VLA via Interactive World Model Wan: Open and Advanced Large-Scale Video Generative Models

Reference 41

Resolution
verified exact
local_arxiv, observed 2026-07-03T17:18:44.345493Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=pdf_text observed=2026-06-27T04:13:22.598591Z digest=sha256:8b20047bef8a2bf1e8b90b822ae9065163790f694228fc5462592090ef4c5a48

Observation 147aa7a5-6fa3-4eb3-8450-7dcb4ca15299 · outbound

This paper cites Girgis, R.

SafeDojo: Safe Reinforcement Learning for VLA via Interactive World Model Girgis, R

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-07-03T17:18:44.346116Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=pdf_text observed=2026-06-27T04:13:22.598591Z digest=sha256:5d5fccf4b5325691764aee5ae63c83db110d7024b83a15ef93c0f59e40ea14b4

Observation db3bbbba-6749-460d-8547-ad2e0bd4c1cb · outbound

This paper cites an unresolved cited work.

SafeDojo: Safe Reinforcement Learning for VLA via Interactive World Model Unresolved cited work

Reference 43

Resolution
unresolved
no resolver link, observed 2026-06-27T04:13:22.598591Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T04:13:22.598591Z digest=sha256:e3ab06a7835d44d2056aca56db789e2e36bfdcc70aabe4f205647a5bb8e12b6a

Observation 141e5c11-0b4c-48f0-a59a-2ff7902f225b · outbound

This paper cites robosuite: A Modular Simulation Framework and Benchmark for Robot Learning.

SafeDojo: Safe Reinforcement Learning for VLA via Interactive World Model robosuite: A Modular Simulation Framework and Benchmark for Robot Learning

Reference 44

Resolution
verified exact
local_arxiv, observed 2026-07-03T17:28:44.193134Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=pdf_text observed=2026-06-27T04:13:22.598591Z digest=sha256:f93a4187e52a09a41ae31d4f6b6ac1f3ea83c82bdba5b80ffb1f7e61989491f6

Observation 82e32704-df5a-4e60-b6aa-0260083646b3 · outbound

This paper cites an unresolved cited work.

SafeDojo: Safe Reinforcement Learning for VLA via Interactive World Model Unresolved cited work

Reference 45

Resolution
unresolved
no resolver link, observed 2026-06-27T04:13:22.598591Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T04:13:22.598591Z digest=sha256:8bd5cfa663d9d90a12df8ed8d19a340020514f324b79f992ef1f200f3dd90433

Observation 28bf2df1-fab4-4452-be46-8a37b45d13e1 · outbound

This paper cites Fine-Tuning Vision-Language-Action Models: Optimizing Speed and Success.

SafeDojo: Safe Reinforcement Learning for VLA via Interactive World Model Fine-Tuning Vision-Language-Action Models: Optimizing Speed and Success

Reference 46

Resolution
metadata mismatch
doi, observed 2026-06-27T04:20:31.963475Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-25T06:30:59.84592+00:00.

source=pdf_text observed=2026-06-27T04:13:22.598591Z digest=sha256:c7e6e562f2de4608c4bb4f729ddea71019f64de2672b6b3e766fe3dbc42bdff0

Pith citing papers

No inbound Pith citation observations are available.