Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T13:07:09.738381Z
Paper Citation Record · LEDGER
As of 11 August 2026, this Paper Citation Record lists 84 of 84 outbound references and 37 inbound Pith citation observations for arXiv:2507.21053.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T13:07:09.738381Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-06T20:50:11.728877Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T06:19:37.772072Z
84 of 84 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 7d48a4f4-865f-4fbc-80ec-061b0d8d4127 · outbound
Flow Matching Policy Gradients Hierarchical Text-Conditional Image Generation with CLIP Latents
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1cf0b7e5-1090-468e-92ec-5c429fda7fe1 · outbound
Flow Matching Policy Gradients Photorealistic text-to-image diffusion models with deep language understanding
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 17e532ea-88af-460f-b415-c19b93531e5b · outbound
Flow Matching Policy Gradients Video generation models as world simulators
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 73329125-5b0e-4757-a629-9c0b07f1bb92 · outbound
Flow Matching Policy Gradients Movie Gen: A Cast of Media Foundation Models
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 080afe53-3d90-4311-80d9-af33d5eb7f3a · outbound
Flow Matching Policy Gradients Unresolved cited work
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 511d80a0-aa90-46af-93dc-954e0950af30 · outbound
Flow Matching Policy Gradients AudioLDM: Text-to-Audio Generation with Latent Diffusion Models
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9bb4e091-57f8-4c23-8170-c196d80b0b7d · outbound
Flow Matching Policy Gradients DiffWave: A Versatile Diffusion Model for Audio Synthesis
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1b9295c3-0ca0-42bb-bc66-3eeb1c3d8dbc · outbound
Flow Matching Policy Gradients Action-Minimization Meets Generative Modeling: Efficient Transition Path Sampling with the Onsager-Machlup Functional
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 87d264a7-87e9-4ca4-9188-a11c629ce240 · outbound
Flow Matching Policy Gradients SFT Memorizes, RL Generalizes: A Comparative Study of Foundation Model Post-training
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2e2ad645-3fe4-42bc-8729-bdd904b89985 · outbound
Flow Matching Policy Gradients DeepSeek-V2: A Strong, Economical, and Efficient Mixture-of-Experts Language Model
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 859eba96-90f8-4d58-a110-993b68a50282 · outbound
Flow Matching Policy Gradients Flow Matching for Generative Modeling
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 91c46414-c837-4d29-8744-b3125bb80fa6 · outbound
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 21c31ee7-1365-45fa-aa20-5e285033cb70 · outbound
Flow Matching Policy Gradients Sutton, David McAllester, Satinder P
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 48973bbe-0726-4fad-833d-1556a987a731 · outbound
Flow Matching Policy Gradients Simple statistical gradient-following algorithms for connectionist reinforce- ment learning
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 0c314530-806c-4eb5-9392-de4e099117ec · outbound
Flow Matching Policy Gradients Unresolved cited work
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 495be691-4082-452c-aa65-6fa5c7ccfafb · outbound
Flow Matching Policy Gradients Natural actor–critic
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation c942879b-aa84-40dc-8ebe-48948c74ef51 · outbound
Flow Matching Policy Gradients Trust region policy optimization
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 186eca33-80ff-48ef-bf90-69f4bacada81 · outbound
Flow Matching Policy Gradients Proximal Policy Optimization Algorithms
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3ba1e354-460b-4978-8393-5be06c232d93 · outbound
Flow Matching Policy Gradients Asynchronous methods for deep reinforcement learning
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation ed193931-a53c-4b6b-a351-c52365a42aea · outbound
Flow Matching Policy Gradients Sample efficient actor–critic with experience replay
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation c982a532-ee28-4f22-a50f-5a19931f6c5e · outbound
Flow Matching Policy Gradients DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 779f8cd8-dee5-4ecc-a2e7-4e9df2984d21 · outbound
Flow Matching Policy Gradients Benchmarking deep reinforcement learning for continuous control
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 86d4becd-bdb6-4f38-a85a-ec1a24722036 · outbound
Flow Matching Policy Gradients Open RL Benchmark: Comprehensive Tracked Experiments for Reinforcement Learning
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7708fd5f-4d27-4671-9520-997b8c14c45d · outbound
Flow Matching Policy Gradients Learning to walk in minutes using massively parallel deep reinforcement learning
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation e1edfbab-7162-44f1-97f3-925e696e1a8f · outbound
Flow Matching Policy Gradients Curiosity-driven learning of joint locomotion and manipulation tasks
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation c475e2b2-e639-485c-9083-183f583662c2 · outbound
Flow Matching Policy Gradients Sym- metry considerations for learning task symmetric robot policies
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3e435712-f0cc-4ace-a2c4-196104c75814 · outbound
Flow Matching Policy Gradients Visual Imitation Enables Contextual Humanoid Control
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4d984d50-21bc-48ce-b02e-5150c11eec4a · outbound
Flow Matching Policy Gradients Solving Rubik's Cube with a Robot Hand
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cb9b6c92-0a5f-437c-94b8-78b23ddef145 · outbound
Flow Matching Policy Gradients A system for general in-hand object re-orientation
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 8dd4c96f-4605-421d-b162-cc0972a884a1 · outbound
Flow Matching Policy Gradients General in-hand object rotation with vision and touch
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 19f2bce2-c743-464c-a611-e38be38635cb · outbound
Flow Matching Policy Gradients From Simple to Complex Skills: The Case of In-Hand Object Reorientation
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3a61accb-3afc-4990-b998-80ffce8087cf · outbound
Flow Matching Policy Gradients Training language models to follow instructions with human feedback
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation b146a6b7-6b72-4676-aab7-f1e09d02b197 · outbound
Flow Matching Policy Gradients Deep reinforcement learning from human preferences
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cfcbdff0-fef5-443f-91e4-606972a019ad · outbound
Flow Matching Policy Gradients DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b779c2bb-1bff-4a9e-b9c9-72297ed03027 · outbound
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b06b7f54-55a3-4150-b4f2-20e4844256e6 · outbound
Flow Matching Policy Gradients Denoising diffusion probabilistic models
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 01d1919f-4302-4b1d-a44d-c47cee24c712 · outbound
Flow Matching Policy Gradients Denoising Diffusion Implicit Models
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0a6b6828-98b8-461b-9a8a-6c3a8c510549 · outbound
Flow Matching Policy Gradients High-Resolution Image Synthesis with Latent Diffusion Models
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7fb39a58-50c2-4bed-a7a2-d05067b99914 · outbound
Flow Matching Policy Gradients Generative Modeling by Estimating Gradients of the Data Distribution
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c3cf78de-b476-405f-97bd-ea912929706c · outbound
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7e4e6e22-3ebc-4df0-9ec2-109cfb748b9a · outbound
Flow Matching Policy Gradients Make-A-Video: Text-to-Video Generation without Text-Video Data
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 39b0b94b-6eab-4796-b043-c23cc67f78b6 · outbound
Flow Matching Policy Gradients Imagen Video: High Definition Video Generation with Diffusion Models
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 90ec9298-fd08-447f-b4cc-4535480966b8 · outbound
Flow Matching Policy Gradients Grad-TTS: A Diffusion Probabilistic Model for Text-to-Speech
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 29b9bb6b-7f80-42a5-9a3c-dfe1a2e48497 · outbound
Flow Matching Policy Gradients WaveGrad 2: Iterative Refinement for Text-to-Speech Synthesis
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 0ce2233f-4dad-46fa-b9aa-660cf115acc1 · outbound
Flow Matching Policy Gradients $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2c11213d-6d19-4863-9189-f602f75545ad · outbound
Flow Matching Policy Gradients GR00T N1: An Open Foundation Model for Generalist Humanoid Robots
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f0cf4abc-8f27-431c-87b8-b8669653bb02 · outbound
Flow Matching Policy Gradients The Superposition of Diffusion Models Using the It\^o Density Estimator
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c8264315-8621-46be-883f-c88f7f5f9ec6 · outbound
Flow Matching Policy Gradients Diffusion policy: Visuomotor policy learning via action diffusion
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7b275053-1472-4143-81b3-889ac7bbb763 · outbound
Flow Matching Policy Gradients Tenenbaum, Tommi S
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 609b7561-f021-4622-8c15-a979737e5a3b · outbound
Flow Matching Policy Gradients Planning with Diffusion for Flexible Behavior Synthesis
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 72de1d12-188f-43e4-89f4-c072db53ad94 · outbound
Flow Matching Policy Gradients Aligning Text-to-Image Models using Human Feedback
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 608e9c56-f9b2-4354-8679-a2ae634a744a · outbound
Flow Matching Policy Gradients Training Diffusion Models with Reinforcement Learning
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1f94d5fb-f315-478a-91bd-91b0b2591b2b · outbound
Flow Matching Policy Gradients Flow-GRPO: Training Flow Matching Models via Online RL
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 307a410a-5ad7-4481-83ff-e91af1ade30c · outbound
Flow Matching Policy Gradients Learning a Diffusion Model Policy from Rewards via Q-Score Matching
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fbdc4bdc-a63f-468e-84ca-2fba47c89ba3 · outbound
Flow Matching Policy Gradients FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ca47697d-98b3-4c1b-bc21-a3ad36d553dc · outbound
Flow Matching Policy Gradients Addressing Function Approximation Error in Actor-Critic Methods
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1e583592-7a90-4ece-af9d-ac27cb444553 · outbound
Flow Matching Policy Gradients Diffusion Policy Policy Optimization
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c4a1389d-374d-4b3d-9fda-68b9f397afd1 · outbound
Flow Matching Policy Gradients High-Dimensional Continuous Control Using Generalized Advantage Estimation
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d3c463ad-121d-4b9b-a546-b3314b3a03b5 · outbound
Flow Matching Policy Gradients Elucidating the design space of diffusion-based generative models
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 4125040b-071b-4473-86bb-994b4fd49c8d · outbound
Flow Matching Policy Gradients Kingma, Tim Salimans, Ben Poole, and Jonathan Ho
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f87efa6e-56a4-4dcd-9361-095d50f425f2 · outbound
Flow Matching Policy Gradients Understanding Diffusion Objectives as the ELBO with Simple Data Augmentation
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ea5d47b3-29c3-43b5-aa7a-34dd658b3c87 · outbound
Flow Matching Policy Gradients Openai gym, 2016
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 781d152e-9e66-45d5-ae39-c03650a189a4 · outbound
Flow Matching Policy Gradients Gymnasium: A Standard Interface for Reinforcement Learning Environments
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 50e52d53-fe43-4d92-8b60-f8f8f4a83b6c · outbound
Flow Matching Policy Gradients Mujoco: A physics engine for model-based control
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 4a7887dd-6be2-40f4-9483-172d05172d79 · outbound
Flow Matching Policy Gradients Isaac Gym: High Performance GPU-Based Physics Simulation For Robot Learning
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 24209886-9031-4947-9111-e6fc18d5ca08 · outbound
Flow Matching Policy Gradients Ppo-for-beginners: A simple, well-styled ppo implementation in pytorch
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation f4537ca1-3755-46b8-9780-389e03224349 · outbound
Reference 69
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ecec4bba-2bd5-4fe4-87e7-277aaef169b2 · outbound
Flow Matching Policy Gradients dm_control: Software and tasks for continuous control
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation f16deb1d-05d3-4830-add6-e7f2cfc7833f · outbound
Flow Matching Policy Gradients Brax -- A Differentiable Physics Engine for Large Scale Rigid Body Simulation
Reference 71
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eb7e53cf-ead2-4328-83ba-89932559dc55 · outbound
Flow Matching Policy Gradients Adam: A Method for Stochastic Optimization
Reference 72
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2d1d3ee5-2ff8-4e33-b5b2-6831349903e6 · outbound
Flow Matching Policy Gradients Perpetual humanoid control for real-time simulated avatars
Reference 73
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 71d4ce65-67a4-48df-a9cd-22f56efb8c39 · outbound
Flow Matching Policy Gradients Deepmimic: Example- guided deep reinforcement learning of physics-based character skills
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 5ec79a79-0638-48c3-9eef-cc92b07fefdc · outbound
Flow Matching Policy Gradients Amass: Archive of motion capture as surface shapes
Reference 75
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3f6620e2-326b-4dba-8f1b-6df9e88abedc · outbound
Flow Matching Policy Gradients Maskedmimic: Unified physics-based character control through masked motion inpainting
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation da5426c2-df28-4a30-8564-0ee7fa8871f7 · outbound
Flow Matching Policy Gradients CLONE: Closed-Loop Whole-Body Humanoid Teleoperation for Long-Horizon Tasks
Reference 77
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fb531a57-4844-48df-be70-a7f9efc57875 · outbound
Flow Matching Policy Gradients Universal Humanoid Motion Representations for Physics-Based Control
Reference 78
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f15fcd56-9845-4b78-875e-14fbbc5ef57a · outbound
Flow Matching Policy Gradients Omnigrasp: Grasping diverse objects with simulated humanoids
Reference 79
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation fb6389f7-a526-425c-9733-ba6a960702ba · outbound
Flow Matching Policy Gradients Ai models collapse when trained on recursively generated data
Reference 80
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 5fcc6f94-7a74-4d85-9d2b-613b187b927d · outbound
Flow Matching Policy Gradients The Curse of Recursion: Training on Generated Data Makes Models Forget
Reference 81
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9609017e-463f-4320-a755-009de32ace2f · outbound
Flow Matching Policy Gradients Self-consuming generative models go mad
Reference 82
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 982b49df-d100-47d8-b962-11d9eee2036e · outbound
Flow Matching Policy Gradients Progressive Distillation for Fast Sampling of Diffusion Models
Reference 83
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6d70b079-8e59-4e96-bb42-e693d40ef204 · outbound
Flow Matching Policy Gradients Classifier-Free Diffusion Guidance
Reference 84
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cbd8537c-7cdd-4b41-b843-fb363b4b1b1e · outbound
Flow Matching Policy Gradients Low CFG scales tend to encourage bluriness while high CFG scales encourage saturation and sharp geometric artifacts
Reference 86
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation ff534587-8822-4c30-9401-e11a3b7bb46c · outbound
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 05e84f68-da6f-4416-a7fe-69d441ee2b57 · inbound
FM-IRL: Flow-Matching for Reward Modeling and Policy Regularization in Reinforcement Learning Flow Matching Policy Gradients
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 42306e54-3ee8-4118-a18e-aabd58d51bbb · inbound
Reinforcement Fine-Tuning of Flow-Matching Policies for Vision-Language-Action Models Flow Matching Policy Gradients
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 813d2a66-f479-4ab2-9210-41d0a7f218b0 · inbound
Training Diffusion Policies via Prior-Mapping Co-Evolution Flow Matching Policy Gradients
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 02cc83af-386b-48db-be1b-38a7ae9bbf7a · inbound
SERNF: Sample-Efficient Real-World Dexterous Policy Fine-Tuning via Action-Chunked Critics and Normalizing Flows Flow Matching Policy Gradients
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6110491d-3f52-421b-9f2f-7a52c6b18727 · inbound
FAIL: Flow Matching Adversarial Imitation Learning for Image Generation Flow Matching Policy Gradients
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e72ba1c7-d96b-452b-88d1-1d6dfddec020 · inbound
From Prior to Pro: Efficient Skill Mastery via Distribution Contractive RL Finetuning Flow Matching Policy Gradients
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4dbeea93-a4ba-4039-aa29-189b66e09596 · inbound
Genuine pair density wave order on the kagome lattice Flow Matching Policy Gradients
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 81cc340c-653a-4b55-88ec-34322423e827 · inbound
FP4 Explore, BF16 Train: Diffusion Reinforcement Learning via Efficient Rollout Scaling Flow Matching Policy Gradients
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation e789f086-ded3-4079-b850-39e5af030598 · inbound
Positive-Only Drifting Policy Optimization Flow Matching Policy Gradients
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 4cd40595-a073-464a-baad-2aefc12f74a9 · inbound
V-GRPO: Online Reinforcement Learning for Denoising Generative Models Is Easier than You Think Flow Matching Policy Gradients
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 99dd1daf-654c-4afd-9632-0e9e5074a3c9 · inbound
OGPO: Sample Efficient Full-Finetuning of Generative Control Policies Flow Matching Policy Gradients
Reference 161
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation d1f51504-4b4f-4ca0-9bee-40f1946ef096 · inbound
OGPO: Sample Efficient Full-Finetuning of Generative Control Policies Flow Matching Policy Gradients
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 7b1ba904-5f14-4f41-a195-a8a8dd17e79b · inbound
Generative Actor-Critic with Soft Bridge Policies Flow Matching Policy Gradients
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 9c0034ed-4c0b-4de1-abd3-a6314507a6c2 · inbound
Preserving Foundational Capabilities in Flow-Matching VLAs through Conservative SFT Flow Matching Policy Gradients
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 525ed746-329f-4c0b-8e25-a99392d0698d · inbound
Preserving Foundational Capabilities in Flow-Matching VLAs through Conservative SFT Flow Matching Policy Gradients
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 9925d1a7-0a26-4370-9f8d-a360209d5528 · inbound
UniSteer: Unified Noise Steering for Efficient Human-Guided VLA Adaptation Flow Matching Policy Gradients
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 26ebde54-747a-4775-b632-bda0ab5f1f03 · inbound
UniSteer: Unified Noise Steering for Efficient Human-Guided VLA Adaptation Flow Matching Policy Gradients
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 011018f0-67a2-453d-8fcf-545ddeb72532 · inbound
Discrete Flow Matching for Offline-to-Online Reinforcement Learning Flow Matching Policy Gradients
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 9f39d2f9-dc00-49a1-b84e-4f54eec339d3 · inbound
Driving Intents Amplify Planning-Oriented Reinforcement Learning Flow Matching Policy Gradients
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 49869b54-f9c1-4f88-acf4-a3fed8f6ad0f · inbound
Driving Intents Amplify Planning-Oriented Reinforcement Learning Flow Matching Policy Gradients
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation bb8ed032-0dba-404a-86d1-d4871142448b · inbound
Video Models Can Reason with Verifiable Rewards Flow Matching Policy Gradients
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 6a075063-c23c-48ff-97fe-63a7d6b107a0 · inbound
DISA: Offline Importance Sampling for Distribution-Matching LLM-RL Flow Matching Policy Gradients
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 23dc9362-165e-4fa4-94e4-ee0bd0bbce84 · inbound
DEFLECT: Delay-Robust Execution via Flow-matching Likelihood-Estimated Counterfactual Tuning for VLA Policies Flow Matching Policy Gradients
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation c6dda188-44b6-40bd-b964-c8790ddc3744 · inbound
Adversarial Dual On-Policy Distillation from Expressive Teacher Flow Matching Policy Gradients
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation ba7332b4-aa86-4a3a-93fe-e8b92efd1677 · inbound
Explicit Critic Guidance for Aligning Diffusion Models Flow Matching Policy Gradients
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 6f5d792f-788b-4217-9a3c-f27b0434a31c · inbound
GDSD: Reinforcement Learning as Guided Denoiser Self-Distillation for Diffusion Language Models Flow Matching Policy Gradients
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation dfe8937c-42c2-4882-a823-a84fba206635 · inbound
Sample-Efficient Diffusion-based Reinforcement Learning with Critic Guidance Flow Matching Policy Gradients
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 0fefc763-7a47-4966-a122-eddd48ef2301 · inbound
GenPO++: Generative Policy Optimization with Jacobian-free Likelihood Ratios Flow Matching Policy Gradients
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 54dff922-423a-4c63-9fd8-c2dc04662816 · inbound
Test-Time Gradient Guidance of Flow Policies in Reinforcement Learning Flow Matching Policy Gradients
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 644ab85c-d687-4f57-9b97-ce1248b4bbbe · inbound
DiPOD: Diffusion Policy Optimization without Drifting Apart Flow Matching Policy Gradients
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 3084d04a-393b-4e9c-802a-4c0bc3245ad7 · inbound
Transferring Contact, Not Just Motion: Compliant Grasping Across Dexterous Hands Flow Matching Policy Gradients
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 7ec13d94-630e-4ff1-9b03-e26a0bba8df9 · inbound
ReFPO: Reflow Regularization for Flow Matching Policy Gradients Flow Matching Policy Gradients
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation ac6772b6-e3e0-4d8e-909f-058b20597375 · inbound
Support-Constrained RL Enables Real-World Policy Improvement without Real-World Experience Flow Matching Policy Gradients
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 43635d58-6a0a-4d67-88f2-cebaaff81699 · inbound
NavCMPO: Critic-Guided MeanFlow Policy Optimization for Adaptive Navigation Flow Matching Policy Gradients
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fe9cea05-9176-457a-a327-ee2baa00f1af · inbound
PAVXploreRL: Physical-Action-Visual World Model Reinforcement Learning with Action Exploration Flow Matching Policy Gradients
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dda17a4b-e2c0-45c6-b9ca-5f0cc6d45e4f · inbound
RLMM-Flow: A Flow-based Mobile Manipulation Framework with Latent-Space Reinforcement Learning Flow Matching Policy Gradients
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5869d352-9738-41ae-8acc-8e07dbfb20cf · inbound
GASP: GPU-Accelerated Safe Planner for Real-Time Collision-Aware Motion Generation with Latent Trajectory Sampling Flow Matching Policy Gradients
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.