Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-06-28T05:38:11.089753Z
Paper Citation Record · LEDGER
As of 6 August 2026, this Paper Citation Record lists 45 of 45 outbound references and 1 inbound Pith citation observation for arXiv:2606.05468.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-06-28T05:38:11.089753Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-02T11:31:34.620402Z
A source-named dated measurement, never combined with another source.
Source: cited_works
45 of 45 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 0e4a2a7b-9acb-4ea5-9ba6-216eb7efa12a · outbound
FlowPRO: Reward-Free Reinforced Fine-Tuning of Flow-Matching VLAs via Proximalized Preference Optimization $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation a84d4502-3a0b-483e-a35b-b828431ec7ec · outbound
FlowPRO: Reward-Free Reinforced Fine-Tuning of Flow-Matching VLAs via Proximalized Preference Optimization Zitkovich, T
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9d4b9225-f0bd-4788-b614-29b90f28b27f · outbound
FlowPRO: Reward-Free Reinforced Fine-Tuning of Flow-Matching VLAs via Proximalized Preference Optimization OpenVLA: An Open-Source Vision-Language-Action Model
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation da506971-fa08-47eb-94c5-e908cac25bdb · outbound
FlowPRO: Reward-Free Reinforced Fine-Tuning of Flow-Matching VLAs via Proximalized Preference Optimization Octo: An Open-Source Generalist Robot Policy
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 8015cbd1-3dec-4918-8d9e-bc5ffa58d7fd · outbound
FlowPRO: Reward-Free Reinforced Fine-Tuning of Flow-Matching VLAs via Proximalized Preference Optimization $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation bcd60db1-2b4b-409d-bda8-91382679e21b · outbound
FlowPRO: Reward-Free Reinforced Fine-Tuning of Flow-Matching VLAs via Proximalized Preference Optimization Unresolved cited work
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 84f5b215-e300-4979-817b-fb53cbe1ee15 · outbound
FlowPRO: Reward-Free Reinforced Fine-Tuning of Flow-Matching VLAs via Proximalized Preference Optimization Flow Matching for Generative Modeling
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 88195603-b028-4393-a0a0-5c3de5f9c7ca · outbound
FlowPRO: Reward-Free Reinforced Fine-Tuning of Flow-Matching VLAs via Proximalized Preference Optimization Flow Straight and Fast: Learning to Generate and Transfer Data with Rectified Flow
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 9aedf7db-f547-48ec-a8f4-d04a5e68147d · outbound
FlowPRO: Reward-Free Reinforced Fine-Tuning of Flow-Matching VLAs via Proximalized Preference Optimization What Matters in Learning from Offline Human Demonstrations for Robot Manipulation
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation bf6dabe3-9446-46be-9a50-3740297c0a36 · outbound
FlowPRO: Reward-Free Reinforced Fine-Tuning of Flow-Matching VLAs via Proximalized Preference Optimization Unresolved cited work
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 366f44bc-a214-4235-9c59-d04498b368cf · outbound
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1d71d986-d49f-4fc2-bb8a-0621822811f8 · outbound
FlowPRO: Reward-Free Reinforced Fine-Tuning of Flow-Matching VLAs via Proximalized Preference Optimization Unresolved cited work
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f5f1f064-0950-4c8a-807e-fa9479fcca7f · outbound
FlowPRO: Reward-Free Reinforced Fine-Tuning of Flow-Matching VLAs via Proximalized Preference Optimization Schulman, F
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 66141af6-a326-4bb8-94bc-a52f2edd3006 · outbound
FlowPRO: Reward-Free Reinforced Fine-Tuning of Flow-Matching VLAs via Proximalized Preference Optimization VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 62ec178a-1942-4964-aefe-d46acd212ed5 · outbound
FlowPRO: Reward-Free Reinforced Fine-Tuning of Flow-Matching VLAs via Proximalized Preference Optimization Unresolved cited work
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b1351cff-46a7-4dc9-a200-303dd7a210c5 · outbound
FlowPRO: Reward-Free Reinforced Fine-Tuning of Flow-Matching VLAs via Proximalized Preference Optimization Unresolved cited work
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1b030e9c-c707-4ca2-9d50-7611f9179f14 · outbound
FlowPRO: Reward-Free Reinforced Fine-Tuning of Flow-Matching VLAs via Proximalized Preference Optimization $\pi^{*}_{0.6}$: a VLA That Learns From Experience
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 50bdbd37-f54a-4d62-9397-1792e9d6d04f · outbound
FlowPRO: Reward-Free Reinforced Fine-Tuning of Flow-Matching VLAs via Proximalized Preference Optimization Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 5b939387-b20c-4efc-b60f-0a503c1c955f · outbound
FlowPRO: Reward-Free Reinforced Fine-Tuning of Flow-Matching VLAs via Proximalized Preference Optimization Offline Reinforcement Learning: Tutorial, Review, and Perspectives on Open Problems
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 02a91f91-65be-4ab1-a6f7-ca0aaa2d4ece · outbound
FlowPRO: Reward-Free Reinforced Fine-Tuning of Flow-Matching VLAs via Proximalized Preference Optimization GRAPE: Generalizing Robot Policy via Preference Alignment
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation c801d4a2-490e-4708-84d3-ae93e50b8d97 · outbound
FlowPRO: Reward-Free Reinforced Fine-Tuning of Flow-Matching VLAs via Proximalized Preference Optimization Rafailov, A
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bbb28a0e-da5e-4591-8c1a-15b065f0726b · outbound
FlowPRO: Reward-Free Reinforced Fine-Tuning of Flow-Matching VLAs via Proximalized Preference Optimization Unresolved cited work
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6ef050e2-f988-407d-9b35-1216cdef32f2 · outbound
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 98a38f4f-e9cb-4295-9fe7-61d3107ab9cb · outbound
FlowPRO: Reward-Free Reinforced Fine-Tuning of Flow-Matching VLAs via Proximalized Preference Optimization Kostrikov, A
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3a2c723f-5941-4219-b31b-ef33615b6131 · outbound
FlowPRO: Reward-Free Reinforced Fine-Tuning of Flow-Matching VLAs via Proximalized Preference Optimization Nakamoto, S
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a51271d3-4c59-4733-b050-58bffa008673 · outbound
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6d466824-7d3e-46fd-a6c4-3d9b9c9e6d93 · outbound
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 83df395c-3f9b-475d-a3e2-6049c4da235c · outbound
FlowPRO: Reward-Free Reinforced Fine-Tuning of Flow-Matching VLAs via Proximalized Preference Optimization KTO: Model Alignment as Prospect Theoretic Optimization
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation 75c04942-f7a4-49e3-81b1-6bda7cfbccb3 · outbound
FlowPRO: Reward-Free Reinforced Fine-Tuning of Flow-Matching VLAs via Proximalized Preference Optimization A General Theoretical Paradigm to Understand Learning from Human Preferences
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation b3e48ca0-1ce2-4742-bb8f-0fdabe1d4422 · outbound
FlowPRO: Reward-Free Reinforced Fine-Tuning of Flow-Matching VLAs via Proximalized Preference Optimization Unresolved cited work
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 34de7e37-25f0-4224-92fc-8cf88aed9166 · outbound
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 50dd785a-3d7a-4927-b24e-109c417bd9e0 · outbound
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 52c89370-f519-4b01-8e73-5cf3b319c8f9 · outbound
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 23c65083-dab4-4214-b78c-3511fa8d7f6c · outbound
FlowPRO: Reward-Free Reinforced Fine-Tuning of Flow-Matching VLAs via Proximalized Preference Optimization Unresolved cited work
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3706f97f-1fbf-4b2c-a064-d5da69ac0fc3 · outbound
FlowPRO: Reward-Free Reinforced Fine-Tuning of Flow-Matching VLAs via Proximalized Preference Optimization Improving Video Generation with Human Feedback
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.
Observation c5c31fad-de0e-4ae0-8776-5f4865e16516 · outbound
FlowPRO: Reward-Free Reinforced Fine-Tuning of Flow-Matching VLAs via Proximalized Preference Optimization Unresolved cited work
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d36f2619-0a99-40c7-8f4a-a36d8a92d57d · outbound
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3eadd73a-5309-4ada-bdc8-70d20e3c1ada · outbound
FlowPRO: Reward-Free Reinforced Fine-Tuning of Flow-Matching VLAs via Proximalized Preference Optimization Unresolved cited work
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 85320e61-e839-4b03-be3e-649330e1be89 · outbound
FlowPRO: Reward-Free Reinforced Fine-Tuning of Flow-Matching VLAs via Proximalized Preference Optimization Mantel and W
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7d353e49-49c0-4a2d-be97-c822563a6b65 · outbound
FlowPRO: Reward-Free Reinforced Fine-Tuning of Flow-Matching VLAs via Proximalized Preference Optimization Unresolved cited work
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 072df478-282c-42ec-864e-7c7561d2172e · outbound
FlowPRO: Reward-Free Reinforced Fine-Tuning of Flow-Matching VLAs via Proximalized Preference Optimization Unresolved cited work
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4695fdcd-8dc7-4f30-80fc-93934b4ed4af · outbound
FlowPRO: Reward-Free Reinforced Fine-Tuning of Flow-Matching VLAs via Proximalized Preference Optimization Unresolved cited work
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bfe98040-db98-4073-8a96-fab6e0ca3119 · outbound
FlowPRO: Reward-Free Reinforced Fine-Tuning of Flow-Matching VLAs via Proximalized Preference Optimization Unresolved cited work
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6e4a9f5b-cf17-4a89-9562-193885a123da · outbound
FlowPRO: Reward-Free Reinforced Fine-Tuning of Flow-Matching VLAs via Proximalized Preference Optimization Unresolved cited work
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 88880598-d42e-4036-88d5-ef4840487b44 · outbound
FlowPRO: Reward-Free Reinforced Fine-Tuning of Flow-Matching VLAs via Proximalized Preference Optimization correct-then-retrain
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3329dd95-3ae2-4c66-811f-3187e50d2f02 · inbound
Hy-Embodied-0.5-VLA: From Vision-Language-Action Models to a Real-World Robot Learning Stack FlowPRO: Reward-Free Reinforced Fine-Tuning of Flow-Matching VLAs via Proximalized Preference Optimization
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.