Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 23 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 38 inbound Pith citation observations for arXiv:2402.03681.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-16T05:31:55.956455Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
6
pith, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation cc57d9c5-0e98-4c8f-95e1-cf4e4564c3f0 · inbound
ViSTa Dataset: Do vision-language models understand sequential tasks? RL-VLM-F: Reinforcement Learning from Vision Language Foundation Model Feedback
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7face66c-8094-4861-8bb5-e0511b7b31ba · inbound
LLM-Based Offline Learning for Embodied Agents via Consistency-Guided Reward Ensemble RL-VLM-F: Reinforcement Learning from Vision Language Foundation Model Feedback
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e923c4c4-e55c-4e44-b176-e626888b8f78 · inbound
MALMM: Multi-Agent Large Language Models for Zero-Shot Robotics Manipulation RL-VLM-F: Reinforcement Learning from Vision Language Foundation Model Feedback
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 386c3901-3ef0-45c6-8a8e-f23fb05974a4 · inbound
SILMM: Self-Improving Large Multimodal Models for Compositional Text-to-Image Generation RL-VLM-F: Reinforcement Learning from Vision Language Foundation Model Feedback
Reference 69
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 08752a8d-6ae4-4138-9b49-da9046bd152a · inbound
YingSound: Video-Guided Sound Effects Generation with Multi-modal Chain-of-Thought Controls RL-VLM-F: Reinforcement Learning from Vision Language Foundation Model Feedback
Reference 100
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 00b8eccd-443e-422a-a52b-e4e2b94300da · inbound
VLM-RL: A Unified Vision Language Models and Reinforcement Learning Framework for Safe Autonomous Driving RL-VLM-F: Reinforcement Learning from Vision Language Foundation Model Feedback
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 72d9408e-e95e-407a-bcba-4cf965a72b72 · inbound
CLIP-RLDrive: Human-Aligned Autonomous Driving via CLIP-Based Reward Shaping in Reinforcement Learning RL-VLM-F: Reinforcement Learning from Vision Language Foundation Model Feedback
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2e8e269f-ae0d-4245-a649-6aad57ad9ed3 · inbound
Contrastive Learning from Exploratory Actions: Leveraging Natural Interactions for Preference Elicitation RL-VLM-F: Reinforcement Learning from Vision Language Foundation Model Feedback
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7463f495-a244-4402-95be-4bc5fd03023d · inbound
FDPP: Fine-tune Diffusion Policy with Human Preference RL-VLM-F: Reinforcement Learning from Vision Language Foundation Model Feedback
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1e78bded-9810-45f3-9f60-d1d02b92c022 · inbound
PRISM: Projection-based Reward Integration for Scene-Aware Real-to-Sim-to-Real Transfer with Few Demonstrations RL-VLM-F: Reinforcement Learning from Vision Language Foundation Model Feedback
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cb0a0399-3f18-42a0-bdf4-76f5bfcbbff7 · inbound
TREND: Tri-teaching for Robust Preference-based Reinforcement Learning with Demonstrations RL-VLM-F: Reinforcement Learning from Vision Language Foundation Model Feedback
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d6e0d8eb-dcc5-43a1-8cd8-7c1850be8884 · inbound
Sample Efficient Reinforcement Learning via Large Vision Language Model Distillation RL-VLM-F: Reinforcement Learning from Vision Language Foundation Model Feedback
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 507ccd29-01c6-4931-9078-8b810bcaa289 · inbound
ManiTaskGen: A Comprehensive Task Generator for Benchmarking and Improving Vision-Language Agents on Embodied Decision-Making RL-VLM-F: Reinforcement Learning from Vision Language Foundation Model Feedback
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aad52432-1358-4d8d-8388-110825fd8d6a · inbound
ROAD: Responsibility-Oriented Reward Design for Reinforcement Learning in Autonomous Driving RL-VLM-F: Reinforcement Learning from Vision Language Foundation Model Feedback
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1debe993-d244-4a08-84cc-9bc6ccd864dc · inbound
HALO: Human Preference Aligned Offline Reward Learning for Robot Navigation RL-VLM-F: Reinforcement Learning from Vision Language Foundation Model Feedback
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c5684732-36b4-4776-8e20-104819cb6eb6 · inbound
UAV-VL-R1: Generalizing Vision-Language Models via Supervised Fine-Tuning and Multi-Stage GRPO for UAV Visual Reasoning RL-VLM-F: Reinforcement Learning from Vision Language Foundation Model Feedback
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 060d5b69-72d5-4948-9dd6-85d684d9fe3c · inbound
Reflection-Based Task Adaptation for Self-Improving VLA RL-VLM-F: Reinforcement Learning from Vision Language Foundation Model Feedback
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 8145122e-b207-4fe9-ad62-9219e51fd3a4 · inbound
Meta-Learning and Meta-Reinforcement Learning -- Tracing the Path towards DeepMind's Adaptive Agent RL-VLM-F: Reinforcement Learning from Vision Language Foundation Model Feedback
Reference 182
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation b0f84f00-831b-4459-8c57-1980f1a76a1b · inbound
SOLE-R1: Video-Language Reasoning as the Sole Reward for On-Robot Reinforcement Learning RL-VLM-F: Reinforcement Learning from Vision Language Foundation Model Feedback
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation be41da0a-9061-4663-b6dd-701b2e2ac0e9 · inbound
AtlasVA: Self-Evolving Visual Skill Memory for Teacher-Free VLM Agents RL-VLM-F: Reinforcement Learning from Vision Language Foundation Model Feedback
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 595aa6c1-82b7-4519-8978-ab500322f516 · inbound
Reasoning Portability: Guiding Continual Learning for MLLMs in the RLVR Era RL-VLM-F: Reinforcement Learning from Vision Language Foundation Model Feedback
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 1cc2b38f-b9e8-463c-b334-a8bb8672514d · inbound
Beyond Pixels: Learning Invariant Rewards for Real-World Robotics From a Few Demonstrations RL-VLM-F: Reinforcement Learning from Vision Language Foundation Model Feedback
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 2e8cf5b5-b2c5-4f1d-8830-fa282443f118 · inbound
CV-Arena: An Open Benchmark for Instructional Computer Vision Problem Solving with Human-AI Collaborative Preferences RL-VLM-F: Reinforcement Learning from Vision Language Foundation Model Feedback
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 58dec1c0-6e7c-4150-8538-f0174bd95a96 · inbound
World Model Self-Distillation: Training World Models to Solve General Tasks RL-VLM-F: Reinforcement Learning from Vision Language Foundation Model Feedback
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 1549fe9b-7479-4347-8a31-cc63ad680120 · inbound
MAGNIFIED: RL Fine-tuning of Multimodal Large Language Models for Motion Planning RL-VLM-F: Reinforcement Learning from Vision Language Foundation Model Feedback
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation bed6605d-e35a-41f4-81bb-bc60db7357fe · inbound
Learning Process Rewards via Success Visitation Matching for Efficient RL RL-VLM-F: Reinforcement Learning from Vision Language Foundation Model Feedback
Reference 89
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 3c476429-1c35-4df5-8adf-9aa9847da207 · inbound
MAPL: Multi-Objective Preference Learning for Robot Locomotion RL-VLM-F: Reinforcement Learning from Vision Language Foundation Model Feedback
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 597a613c-efae-4c62-b916-b96e68cab2fb · inbound
Digitizing Coaching Intelligence: An Agentic Framework for Holistic Athlete Profiling using VLM and RAG RL-VLM-F: Reinforcement Learning from Vision Language Foundation Model Feedback
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 7942573f-050a-4597-90c3-7c9f599cb04a · inbound
Freeform Preference Learning for Robotic Manipulation RL-VLM-F: Reinforcement Learning from Vision Language Foundation Model Feedback
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 0caae5b4-41cc-49c5-a5a7-dcc07bdc7d52 · inbound
Freeform Preference Learning for Robotic Manipulation RL-VLM-F: Reinforcement Learning from Vision Language Foundation Model Feedback
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e14f6cc5-e031-4ecd-8ea1-afff73aff912 · inbound
Freeform Preference Learning for Robotic Manipulation RL-VLM-F: Reinforcement Learning from Vision Language Foundation Model Feedback
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f7fe787f-399f-467c-971e-38c5db5a2352 · inbound
QVal: Cheaply Evaluating Dense Supervision Signals for Long-Horizon LLM Agents RL-VLM-F: Reinforcement Learning from Vision Language Foundation Model Feedback
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 03ba039f-d16d-4838-a07f-6ea9688dbcf6 · inbound
LLM-as-a-Verifier: A General-Purpose Verification Framework RL-VLM-F: Reinforcement Learning from Vision Language Foundation Model Feedback
Reference 80
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.
Observation 2159aea4-c9d6-488c-bf68-f1ac3e07fd84 · inbound
LLM-as-a-Verifier: A General-Purpose Verification Framework RL-VLM-F: Reinforcement Learning from Vision Language Foundation Model Feedback
Reference 80
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cb1fefb5-d847-43f4-98f8-0233a4d79a31 · inbound
Prompt-Driven Exploration RL-VLM-F: Reinforcement Learning from Vision Language Foundation Model Feedback
Reference 115
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aaa22542-473a-4093-a51c-0125399e5ac2 · inbound
Learning More from Less: Reinforcement Learning from Hindsight RL-VLM-F: Reinforcement Learning from Vision Language Foundation Model Feedback
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 886dba74-5a71-4507-9b08-7e7097e3363a · inbound
MARS-RA: Rank Aggregation for Credit Assignment via Multimodal Comparisons in Embodied Multi-Agent Cooperation RL-VLM-F: Reinforcement Learning from Vision Language Foundation Model Feedback
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0513bdd3-5d69-46bb-8356-dde3addce4f4 · inbound
Weights or Skills? A Survey of Robot-Learning Techniques: from Action-Predicting Weights to Robots that Write their Own Skills RL-VLM-F: Reinforcement Learning from Vision Language Foundation Model Feedback
Reference 260
Source-reported events for the cited work
Unavailable: canonical work link unavailable.