Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-03T03:34:07.067380Z
Paper Citation Record Β· LEDGER
As of 5 August 2026, this Paper Citation Record lists 66 of 66 outbound references and 0 inbound Pith citation observations for arXiv:2607.29613.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-03T03:34:07.067380Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
66 of 66 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 7fa90ab6-857f-4b8d-a7d3-59b97102d7d5 Β· outbound
WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f0f35e92-5023-4295-bcfa-3bf76c62f8a2 Β· outbound
WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning π0.5: Avision-language-actionmodelwithopen-worldgeneralization
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 83ab3ffb-07f9-4034-bfe5-34ba90136c3c Β· outbound
WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning ${\pi}_{0.7}$: a Steerable Generalist Robotic Foundation Model with Emergent Capabilities
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bd1eba51-00ff-4a53-aff5-652d4dd5198c Β· outbound
WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning Fine-Tuning Vision-Language-Action Models: Optimizing Speed and Success
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 343dbda9-a696-4e7f-93d9-645eeaf20350 Β· outbound
WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning OpenVLA: An Open-Source Vision-Language-Action Model
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1a4001d8-33fd-4efa-985f-ccd29084a647 Β· outbound
WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning What can rl bring to vla generalization? an empirical study.arXiv preprint arXiv:2505.19789, 2025
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d08c0812-ab4b-4173-aef3-12c0d56b7712 Β· outbound
WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning Srpo: Self-referential policy optimization for vision-language-action models.arXiv preprint arXiv:2511.15605, 2025
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6eca5fab-20cf-4c76-a10f-91fa5b50df44 Β· outbound
WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning SimpleVLA-RL: Scaling VLA Training via Reinforcement Learning
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 73cdb8fc-c7af-487f-8777-ccc53fcbe341 Β· outbound
WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning Rlinf-vla: A unified and efficient framework for vla+ rl training.arXiv preprint arXiv:2510.06710, 2025
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b647665a-7334-4173-b611-d5ef02bce27f Β· outbound
WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning $\pi^{*}_{0.6}$: a VLA That Learns From Experience
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 242eec4c-a43e-4b75-a5b1-bfd812172088 Β· outbound
WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning Unresolved cited work
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b83d68a0-1171-4bd1-86d9-1894af254b05 Β· outbound
WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning Rlinf-user: A unified and extensible system for real-world online policy learning in embodied ai
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation df350158-23c2-4575-b6b6-4ea84dd51391 Β· outbound
WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning Predictive representations of state.Advances in neural information processing systems, 14, 2001
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 53cfd8ac-add9-43aa-8fa0-aa4a4fe6a9e1 Β· outbound
WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning Learning predictive state representations
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bdb30716-95e7-4ae3-b169-de0b687edc1d Β· outbound
WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning Predictive State Representations: A New Theory for Modeling Dynamical Systems
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1ebfbe89-9d2b-40b7-acaa-08e4cdbe4917 Β· outbound
WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning When is partially observable reinforcement learning not scary? InConference on Learning Theory, pages 5175β5220
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6106fe5c-1350-4222-b9dd-df70d277639f Β· outbound
WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning Approximate information state for approximateplanningandreinforcementlearninginpartiallyobservedsystems.JournalofMachineLearningResearch, 23(12):1β83, 2022
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 123f041f-05e5-42a0-9235-954ec290d52b Β· outbound
WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8a9ee29c-22c2-413d-90fd-6ce154bb1a8b Β· outbound
WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning Reinforcement learning with latent flow.Advances in Neural Information Processing Systems, 34:22171β22183, 2021
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8c54494c-aa1f-4e5f-bf04-52f172b5b90d Β· outbound
WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning Provable reinforcement learning with a short-term memory
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 19972567-022c-4080-8240-65516e2fcbbc Β· outbound
WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning Improving sample efficiencyinmodel-freereinforcementlearningfromimages
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4e14a391-0959-47d7-9eab-7f60f8093f18 Β· outbound
WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning Weakly supervised representation learning with sparse perturbations.Advances in Neural Information Processing Systems, 35:15516β15528, 2022
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f2ecac8e-9a05-4392-a829-95c0f40aa3da Β· outbound
WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning GPT-4 Technical Report
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 02b0f225-4330-464f-90c6-73748e5f0158 Β· outbound
WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning Gemini: A Family of Highly Capable Multimodal Models
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 87d8efb6-e8b5-418f-9a6b-46290389fc05 Β· outbound
WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning Data-efficient reinforcement learning with self-predictive representations.International Conference on Learning Representations, 2020
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7003de70-2833-4330-8149-940cfaadb01c Β· outbound
WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning PaLM-E: An Embodied Multimodal Language Model
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 88171681-e18f-4375-9572-f08ab7c8fa8d Β· outbound
WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning Rt-2: Vision-language-action models transfer web knowledge to robotic control
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 708ea319-aba6-410d-8201-79efc96041d7 Β· outbound
WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning FAST: Efficient Action Tokenization for Vision-Language-Action Models
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 592d2881-d740-4607-8194-2e85d338ba83 Β· outbound
WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 822de1b3-429a-4177-ad8f-3c28a816cdcb Β· outbound
WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning Interactive Post-Training for Vision-Language-Action Models
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 853ca24b-1724-4fcd-b5b7-e97e81336a3a Β· outbound
WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning Proximal Policy Optimization Algorithms
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1b2310d6-92ae-40b1-b520-4c407fb0f638 Β· outbound
WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning DeepSeek-V3 Technical Report
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 44dfb27f-1f49-4687-b600-a0f442de68c8 Β· outbound
WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning Soft Actor-Critic Algorithms and Applications
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 57438874-8622-461f-8331-e7470c642835 Β· outbound
WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning Offline Reinforcement Learning with Implicit Q-Learning
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8aab59ed-a513-42f3-bd7a-1a6b6c3595f0 Β· outbound
WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning Advantage-Weighted Regression: Simple and Scalable Off-Policy Reinforcement Learning
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a11aba10-313d-4f63-905d-e8f4b6034c33 Β· outbound
WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning Reinflow: Fine-tuning flow matching policy with online reinforcement learning.arXiv preprint arXiv:2505.22094, 2025
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 68c8c9e2-c66e-4544-ac17-6f0fce99ed93 Β· outbound
WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning Flow-GRPO: Training Flow Matching Models via Online RL
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9f13788d-107f-4628-ba61-208e54ae897a Β· outbound
WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning Diffusion Policy Policy Optimization
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f698fb0e-85e1-4c77-9a51-3c1495ac8224 Β· outbound
WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning Steering Your Diffusion Policy with Latent Space Reinforcement Learning
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 73ca62b2-2d62-4f7c-a049-4a8f05474fff Β· outbound
WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning Precise and dexterous robotic manipulation via human-in- the-loop reinforcement learning.Science Robotics, 10(105):eads5033, 2025
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cfaad6db-cb59-4d64-aaf1-de56b7e983b3 Β· outbound
WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning Gigabrain-0.5 m*: a vla that learns from world model-based reinforcement learning.arXiv preprint arXiv:2602.12099, 2026
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e929bafd-2a44-4e76-a65b-9c98b9284e70 Β· outbound
WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning Optimal control of markov decision processes with incomplete state estimation.J
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 84aed31f-fe45-4273-9cab-e6d23b77c2b1 Β· outbound
WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning The optimal control of partially observable markov processes over a finite horizon.Operations research, 21(5):1071β1088, 1973
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 98d935dc-c828-43c3-aff9-317e556c7864 Β· outbound
WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning Reinforcement learning with augmented data.Advances in neural information processing systems, 33:19884β19895, 2020
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 78cd7873-4be1-4ea0-82f1-b37fe1bcb0c7 Β· outbound
WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning Contextual decision processes with low bellman rank are pac-learnable
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4e701487-ded9-4a2a-8fde-615fa793c6e4 Β· outbound
WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning Human-level control through deep reinforcement learning.nature, 518(7540):529β533, 2015
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0dadea8c-6c53-4ffa-a2f7-8b220aa2ae01 Β· outbound
WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning Deep recurrent q-learning for partially observable mdps
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3dca8b78-4ee0-49d2-a690-58ae25b7be3c Β· outbound
WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning Robust Reinforcement Learning in POMDPs with Incomplete and Noisy Observations
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 79d01f92-8e09-4a1d-a044-a73562782aeb Β· outbound
WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning Robust Finite-Memory Policy Gradients for Hidden-Model POMDPs
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5a65ab53-f0d4-4fc7-8650-380116ace64a Β· outbound
WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning MemoryVLA: Perceptual-Cognitive Memory in Vision-Language-Action Models for Robotic Manipulation
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6b3b49e7-5eb6-40f0-80b9-2abc4f9ac89e Β· outbound
WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning HAMLET: Switch your Vision-Language-Action Model into a History-Aware Policy
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5a1ba548-5787-4be7-8c7e-704f6aa7dac7 Β· outbound
WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning Cronusvla: Transferring latent motion across time for multi-frame prediction in manipulation.arXiv e-prints, pages arXivβ2506, 2025
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6a9c4f3a-696b-4556-816b-9316f18d4e5b Β· outbound
WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning Decision transformer: Reinforcement learning via sequence modeling.Advances in neural information processing systems, 34:15084β15097, 2021
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6ad8aa0f-39e6-4432-abc4-c809ee450e54 Β· outbound
WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning World Action Models are Zero-shot Policies
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e4186373-6e16-4fe1-8f3b-b2c0d989aa43 Β· outbound
WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning Cosmos Policy: Fine-Tuning Video Models for Visuomotor Control and Planning
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 252b9028-a7b9-48d0-a9bb-9921dc501027 Β· outbound
WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning Motus: A Unified Latent Action World Model
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6ab9c4a7-d9ce-4ee2-9fab-67980c921be9 Β· outbound
WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning Causal World Modeling for Robot Control
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7024ebcd-9741-4e84-95a0-7e1585df37cb Β· outbound
WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning Fast-WAM: Do World Action Models Need Test-time Future Imagination?
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 439b7114-9c0f-4163-9f03-e08b3947a626 Β· outbound
WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning LeJEPA: Provable and Scalable Self-Supervised Learning Without the Heuristics
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a3693436-16f0-403b-b9e0-2b1fc073cd44 Β· outbound
WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning LeWorldModel: Stable End-to-End Joint-Embedding Predictive Architecture from Pixels
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1e559c64-4cae-4de1-9a9a-0406d32435bc Β· outbound
WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning Learning transferable visual models from natural language supervision
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 484980e3-2d16-41dc-96b1-c0e3964ba11c Β· outbound
WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning Film: Visual reasoning with a general conditioning layer
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 99773263-4905-41ea-a982-3fde8f212d8e Β· outbound
WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning ManiSkill: Generalizable Manipulation Skill Benchmark with Large-Scale Demonstrations
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 59d15a0a-12a6-4016-81f4-17d904b8a8dc Β· outbound
WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning Meta- world: A benchmark and evaluation for multi-task and meta reinforcement learning
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 079bb149-e56d-498b-a474-c4f90377d6ef Β· outbound
WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning Unresolved cited work
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d3d21f88-ce07-42eb-963b-2567fafcbccf Β· outbound
WCM: A World Critic Model for Vision-Language-Action Reinforcement Learning LIBERO-Plus: In-depth Robustness Analysis of Vision-Language-Action Models
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.