Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T12:44:08.215410Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 59 of 59 outbound references and 0 inbound Pith citation observations for arXiv:2505.24156.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T12:44:08.215410Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
59 of 59 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 2fac1c56-cc2f-4620-a7b0-064742dec1e4 · outbound
Towards a Generalizable Bimanual Foundation Policy via Flow-based Video Prediction Peract2: Benchmarking and learning for robotic bimanual manipulation tasks
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bcaf170e-95de-4a17-86b4-7433c74a3e05 · outbound
Towards a Generalizable Bimanual Foundation Policy via Flow-based Video Prediction RoboTwin: Dual-Arm Robot Benchmark with Generative Digital Twins (early version)
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a31546b0-097f-4998-86f7-64d31e6aeabb · outbound
Towards a Generalizable Bimanual Foundation Policy via Flow-based Video Prediction Zhao, Vikash Kumar, Sergey Levine, and Chelsea Finn
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 54af4f5a-da69-45e5-b09f-8b5933de8208 · outbound
Towards a Generalizable Bimanual Foundation Policy via Flow-based Video Prediction Sukhatme
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 56c91c17-8459-4094-8113-bb94cca24cfb · outbound
Towards a Generalizable Bimanual Foundation Policy via Flow-based Video Prediction Stabilize to act: Learning to coordinate for bimanual manipulation
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0430fe08-44b4-4835-aae1-d6576339ab25 · outbound
Towards a Generalizable Bimanual Foundation Policy via Flow-based Video Prediction Taco: Benchmarking generalizable bimanual tool-action-object understanding
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e137aa09-a528-4cbf-9254-a739d9d3dd01 · outbound
Towards a Generalizable Bimanual Foundation Policy via Flow-based Video Prediction Towards human-level bimanual dexterous manipulation with reinforcement learning.Advances in Neural Information Processing Systems, 35:5150–5163, 2022
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1e851552-939d-4968-a1d5-99057804227b · outbound
Towards a Generalizable Bimanual Foundation Policy via Flow-based Video Prediction Bi-touch: Bimanual tactile manipulation with sim-to-real deep reinforcement learning.IEEE Robotics and Automation Letters, 8(9):5472–5479, 2023
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b00183a8-b25b-4217-a162-3b8d970974bf · outbound
Towards a Generalizable Bimanual Foundation Policy via Flow-based Video Prediction OpenVLA: An Open-Source Vision-Language-Action Model
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d6a663d2-4eb8-4385-a608-6c619c808dff · outbound
Towards a Generalizable Bimanual Foundation Policy via Flow-based Video Prediction $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0408ea97-8c40-43cb-9232-0a03ef86a20e · outbound
Towards a Generalizable Bimanual Foundation Policy via Flow-based Video Prediction RDT-1B: a Diffusion Foundation Model for Bimanual Manipulation
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6a5143d7-b632-43db-b835-ab8acb693f40 · outbound
Towards a Generalizable Bimanual Foundation Policy via Flow-based Video Prediction GR00T N1: An Open Foundation Model for Generalist Humanoid Robots
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7d75ce22-ca01-4944-ac0b-53e1aca7ee89 · outbound
Towards a Generalizable Bimanual Foundation Policy via Flow-based Video Prediction AgiBot World Colosseo: A Large-scale Manipulation Platform for Scalable and Intelligent Embodied Systems
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c4286d83-0b49-482e-8780-22cbafe551a9 · outbound
Towards a Generalizable Bimanual Foundation Policy via Flow-based Video Prediction Cogvideox: Text-to-video diffusion models with an expert transformer
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6073b6e8-ad1b-42a6-a8d4-eb0e1cf6aa34 · outbound
Towards a Generalizable Bimanual Foundation Policy via Flow-based Video Prediction Open-television: Teleoperation with immersive active visual feedback
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0bf92015-6fa3-4b81-869b-0b24459e896e · outbound
Towards a Generalizable Bimanual Foundation Policy via Flow-based Video Prediction RoboMIND: Benchmark on Multi-embodiment Intelligence Normative Data for Robot Manipulation
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eb0a3954-2c33-4201-87a8-179a45056f8c · outbound
Towards a Generalizable Bimanual Foundation Policy via Flow-based Video Prediction Motion-i2v: Consistent and controllable image-to-video generation with explicit motion modeling
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 61677f15-766e-43c4-ad4d-81e8968df44c · outbound
Towards a Generalizable Bimanual Foundation Policy via Flow-based Video Prediction Raft: Recurrent all-pairs field transforms for optical flow
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cd00a0c8-fdf3-41fc-990d-0c67755ed727 · outbound
Towards a Generalizable Bimanual Foundation Policy via Flow-based Video Prediction Diffusion policy: Visuomotor policy learning via action diffusion.The International Journal of Robotics Research, page 02783649241273668, 2023
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2eccbf07-e555-4e86-97bf-6d047b2b7b58 · outbound
Towards a Generalizable Bimanual Foundation Policy via Flow-based Video Prediction Learning manipulation by sequencing motor primitives with a two-armed robot
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f3228e6d-c88d-4ca3-8bda-a61f047d91c5 · outbound
Towards a Generalizable Bimanual Foundation Policy via Flow-based Video Prediction A system for imitation learning of contact-rich bimanual manipulation policies
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5d8336de-a3bf-4548-87df-43ea6ebf4ae6 · outbound
Towards a Generalizable Bimanual Foundation Policy via Flow-based Video Prediction Deep imitation learning for bimanual robotic manipulation.Advances in neural information processing systems, 33:2327–2337, 2020
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 15afa95b-f4d9-4901-969a-88adc9cd4342 · outbound
Towards a Generalizable Bimanual Foundation Policy via Flow-based Video Prediction HumanoidBench: Simulated Humanoid Benchmark for Whole-Body Locomotion and Manipulation
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a5e7fb93-c5d5-4fa0-b5da-c3a0372b47c6 · outbound
Towards a Generalizable Bimanual Foundation Policy via Flow-based Video Prediction Bigym: A demo-driven mobile bi-manual manipulation benchmark
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1c11c2fe-234f-4b6c-8226-0a098b1da6b1 · outbound
Towards a Generalizable Bimanual Foundation Policy via Flow-based Video Prediction Zhao, and Chelsea Finn
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 167d60bd-69cf-4f17-82e8-9cd07240c600 · outbound
Towards a Generalizable Bimanual Foundation Policy via Flow-based Video Prediction Mimicgen: A data generation system for scalable robot learning using human demonstrations
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d272dfd6-bcb4-4c07-a694-0a90644aaa95 · outbound
Towards a Generalizable Bimanual Foundation Policy via Flow-based Video Prediction Dexmimicgen: Automated data generation for bimanual dexterous manipulation via imitation learning
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1eb90710-51fb-4167-9347-2ce984bc88d0 · outbound
Towards a Generalizable Bimanual Foundation Policy via Flow-based Video Prediction Bi-kvil: Keypoints-based visual imitation learning of bimanual manipulation tasks
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1dd43b64-256b-4be9-9f06-9c35b259422b · outbound
Towards a Generalizable Bimanual Foundation Policy via Flow-based Video Prediction Interactive imitation learning of bimanual movement primitives.IEEE/ASME Transactions on Mechatronics, 2023
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 60c6a14d-e412-4b26-9edf-55b1e1b1aa56 · outbound
Towards a Generalizable Bimanual Foundation Policy via Flow-based Video Prediction InterACT: Inter-dependency aware action chunking with hierarchical attention transformers for bimanual manipulation
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c28b8204-994e-4025-a49b-910d8f7ab99d · outbound
Towards a Generalizable Bimanual Foundation Policy via Flow-based Video Prediction Octo: An Open-Source Generalist Robot Policy
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5b935397-c482-4cc6-9d71-910456882a75 · outbound
Towards a Generalizable Bimanual Foundation Policy via Flow-based Video Prediction SpatialVLA: Exploring Spatial Representations for Visual-Language-Action Model
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 231118b3-8f9b-4594-a547-cde708953bfe · outbound
Towards a Generalizable Bimanual Foundation Policy via Flow-based Video Prediction CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3e83782f-60aa-4ab7-a5c5-87baef0780f6 · outbound
Towards a Generalizable Bimanual Foundation Policy via Flow-based Video Prediction Dexgraspvla: A vision-language-action framework towards general dexterous grasping.arXiv preprint arXiv:2502.20900, 2025
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6f11c451-0ad3-4ec4-acac-2e7d5d3fa5d5 · outbound
Towards a Generalizable Bimanual Foundation Policy via Flow-based Video Prediction MoLe-VLA: Dynamic Layer-skipping Vision Language Action Model via Mixture-of-Layers for Efficient Robot Manipulation
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 18a086b1-d1e2-4298-95f8-2b4e561103c0 · outbound
Towards a Generalizable Bimanual Foundation Policy via Flow-based Video Prediction Learning an actionable dis- crete diffusion policy via large-scale actionless video pre-training
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bee84a65-eff2-407f-86dd-eab3fd439225 · outbound
Towards a Generalizable Bimanual Foundation Policy via Flow-based Video Prediction Video diffusion models.Advances in Neural Information Processing Systems, 35:8633–8646, 2022
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5bcbadb6-88c3-414d-8bb9-ca315084392a · outbound
Towards a Generalizable Bimanual Foundation Policy via Flow-based Video Prediction CogVideo: Large-scale Pretraining for Text-to-Video Generation via Transformers
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation da57f813-7684-47be-ac5b-b5790e4c2d0f · outbound
Towards a Generalizable Bimanual Foundation Policy via Flow-based Video Prediction Learning universal policies via text-guided video generation.Advances in neural information processing systems, 36:9156–9172, 2023
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5c371ed9-8661-4b74-bd73-2598e2f92313 · outbound
Towards a Generalizable Bimanual Foundation Policy via Flow-based Video Prediction Tenenbaum
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d116d889-5c4c-4831-b474-abd130333a4f · outbound
Towards a Generalizable Bimanual Foundation Policy via Flow-based Video Prediction Grounding video models to actions through goal conditioned exploration
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 34c42fec-0338-40a2-9cf2-cdec63d56755 · outbound
Towards a Generalizable Bimanual Foundation Policy via Flow-based Video Prediction Robodreamer: Learning compositional world models for robot imagination
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3cefc23a-d791-41c4-a1b5-45c13c0bb7ff · outbound
Towards a Generalizable Bimanual Foundation Policy via Flow-based Video Prediction Vidman: Exploiting implicit dynamics from video diffusion model for effective robot manipulation.Advances in Neural Information Processing Systems, 37:41051–41075, 2024
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3d7bb084-2812-4bb6-9c60-6fdca5151224 · outbound
Towards a Generalizable Bimanual Foundation Policy via Flow-based Video Prediction Open-Sora: Democratizing Efficient Video Production for All
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2768bf18-0f9c-4e5f-8f63-771cb17a482c · outbound
Towards a Generalizable Bimanual Foundation Policy via Flow-based Video Prediction Open x-embodiment: Robotic learning datasets and rt-x models: Open x-embodiment collaboration 0
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d80fbe2f-7cd1-40ee-9807-32de246ad89f · outbound
Towards a Generalizable Bimanual Foundation Policy via Flow-based Video Prediction IRASim: A Fine-Grained World Model for Robot Manipulation
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 34e5f4aa-5e0c-4837-8273-c8eea0c9a849 · outbound
Towards a Generalizable Bimanual Foundation Policy via Flow-based Video Prediction AVID: Adapting Video Diffusion Models to World Models
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e1c12c8f-e986-43f3-8d92-73890d0503f6 · outbound
Towards a Generalizable Bimanual Foundation Policy via Flow-based Video Prediction AdaWM: Adaptive world model based planning for autonomous driving
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8d20f82c-8c48-4edd-a587-1ff4ce3d668d · outbound
Towards a Generalizable Bimanual Foundation Policy via Flow-based Video Prediction Flowformer++: Masked cost volume autoencoding for pretraining optical flow estimation
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 650addb1-45ad-418b-9f65-ca3a1138b024 · outbound
Towards a Generalizable Bimanual Foundation Policy via Flow-based Video Prediction 3D Diffusion Policy: Generalizable Visuomotor Policy Learning via Simple 3D Representations
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3a264020-d2e0-4d8d-a3a7-a8611b32084e · outbound
Towards a Generalizable Bimanual Foundation Policy via Flow-based Video Prediction Image quality metrics: Psnr vs
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a590c559-8eca-4f32-a6a0-249a978f4b86 · outbound
Towards a Generalizable Bimanual Foundation Policy via Flow-based Video Prediction Image quality assessment: from error visibility to structural similarity.IEEE transactions on image processing, 13(4):600–612, 2004
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 031ffeed-379c-4cf0-b6be-1a3a8a29b77e · outbound
Towards a Generalizable Bimanual Foundation Policy via Flow-based Video Prediction The unreasonable effectiveness of deep features as a perceptual metric
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7927f42b-564d-48ce-9124-096462d42db5 · outbound
Towards a Generalizable Bimanual Foundation Policy via Flow-based Video Prediction Towards Accurate Generative Models of Video: A New Metric & Challenges
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 46889b61-55be-4ffa-b472-269bd1302727 · outbound
Towards a Generalizable Bimanual Foundation Policy via Flow-based Video Prediction Bunny-VisionPro: Real-Time Bimanual Dexterous Teleoperation for Imitation Learning
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1cd04788-5a2c-48be-ba50-222e4c3e00a0 · outbound
Towards a Generalizable Bimanual Foundation Policy via Flow-based Video Prediction OmniH2O: Universal and Dexterous Human-to-Humanoid Whole-Body Teleoperation and Learning
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 179cb5ee-4512-475f-b9e9-989c0d8982f7 · outbound
Towards a Generalizable Bimanual Foundation Policy via Flow-based Video Prediction Using apple vision pro to train and control robots, 2024
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ac0e4fe9-bc07-4976-9e0c-bfabf194b869 · outbound
Towards a Generalizable Bimanual Foundation Policy via Flow-based Video Prediction Grasp the rope on the box and pull together to bring the box closer
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2b3325ee-2e6d-4d9a-9643-4f982f124578 · outbound
Towards a Generalizable Bimanual Foundation Policy via Flow-based Video Prediction Unresolved cited work
Reference 2024
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
No inbound Pith citation observations are available.