Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:10:14.710517Z
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 47 of 47 outbound references and 0 inbound Pith citation observations for arXiv:2505.19767.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:10:14.710517Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
47 of 47 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 84d7d996-2113-49df-b533-066e031b2686 · outbound
RFTF: Reinforcement Fine-tuning for Embodied Agents with Temporal Feedback GPT-4 Technical Report
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0cefab2a-604e-4025-b595-1d9023d64e15 · outbound
RFTF: Reinforcement Fine-tuning for Embodied Agents with Temporal Feedback The claude 3 model family: Opus, sonnet, haiku
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 25d58bb7-f8a8-4b61-b200-6d99f679d5d5 · outbound
RFTF: Reinforcement Fine-tuning for Embodied Agents with Temporal Feedback PaliGemma: A versatile 3B VLM for transfer
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 26cd55c1-25b9-4c35-89ce-385b705f08d3 · outbound
RFTF: Reinforcement Fine-tuning for Embodied Agents with Temporal Feedback $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 10297ea3-b1a0-45fb-8768-1687b41cc52f · outbound
RFTF: Reinforcement Fine-tuning for Embodied Agents with Temporal Feedback Do as i can, not as i say: Grounding language in robotic affordances
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ca446ca0-4380-40a2-b952-0367c84dc081 · outbound
RFTF: Reinforcement Fine-tuning for Embodied Agents with Temporal Feedback AgiBot World Colosseo: A Large-scale Manipulation Platform for Scalable and Intelligent Embodied Systems
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2f5a51c2-fe4b-4098-a172-7fc68a980344 · outbound
RFTF: Reinforcement Fine-tuning for Embodied Agents with Temporal Feedback Closed-Loop Visuomotor Control with Generative Expectation for Robotic Manipulation
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 32791648-9a9f-474d-a85c-0e52b1a44714 · outbound
RFTF: Reinforcement Fine-tuning for Embodied Agents with Temporal Feedback PaLI-X: On Scaling up a Multilingual Vision and Language Model
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d2a96252-7877-47d8-aade-4f0386510bf9 · outbound
RFTF: Reinforcement Fine-tuning for Embodied Agents with Temporal Feedback ConRFT: A Reinforced Fine-tuning Method for VLA Models via Consistency Policy
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 132f271d-e618-43d2-8cfd-7213a697e4ad · outbound
RFTF: Reinforcement Fine-tuning for Embodied Agents with Temporal Feedback Internvl: Scaling up vision foundation models and aligning for generic visual-linguistic tasks
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation deb26e8c-db32-4229-b4a8-e404067fa895 · outbound
RFTF: Reinforcement Fine-tuning for Embodied Agents with Temporal Feedback DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a27cf8f7-cea0-43b4-b7c8-1ec0799ea99a · outbound
RFTF: Reinforcement Fine-tuning for Embodied Agents with Temporal Feedback Improving Vision-Language-Action Model with Online Reinforcement Learning
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9601e9e1-6c59-4e3b-982d-2d31622ec4e8 · outbound
RFTF: Reinforcement Fine-tuning for Embodied Agents with Temporal Feedback Diffusion Transformer Policy
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 595a48b4-583f-4508-961f-4d32ac9e6a73 · outbound
RFTF: Reinforcement Fine-tuning for Embodied Agents with Temporal Feedback FLaRe: Achieving Masterful and Adaptive Robot Policies with Large-Scale Reinforcement Learning Fine-Tuning
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f5ed7ac3-7079-4303-a82c-676d6c11c2f8 · outbound
RFTF: Reinforcement Fine-tuning for Embodied Agents with Temporal Feedback Look Before You Leap: Unveiling the Power of GPT-4V in Robotic Vision-Language Planning
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 48acf3a8-a8c5-495d-bb47-c09580192076 · outbound
RFTF: Reinforcement Fine-tuning for Embodied Agents with Temporal Feedback Inner Monologue: Embodied Reasoning through Planning with Language Models
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f5da6f77-aa0a-4f7a-aff8-73cba822e6c9 · outbound
RFTF: Reinforcement Fine-tuning for Embodied Agents with Temporal Feedback RoboBrain: A Unified Brain Model for Robotic Manipulation from Abstract to Concrete
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f2593529-d885-4a0c-88e6-041bbb5838f6 · outbound
RFTF: Reinforcement Fine-tuning for Embodied Agents with Temporal Feedback Prismatic vlms: Investigating the design space of visually-conditioned language models
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation fba6cf7e-a5ab-4ff0-bad7-344f4cb4c0e4 · outbound
RFTF: Reinforcement Fine-tuning for Embodied Agents with Temporal Feedback 3D Diffuser Actor: Policy Diffusion with 3D Scene Representations
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 170d07f2-10c9-4a20-b293-700f65812c2f · outbound
RFTF: Reinforcement Fine-tuning for Embodied Agents with Temporal Feedback OpenVLA: An Open-Source Vision-Language-Action Model
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 662bd5c9-a28e-4bb5-9d50-0f00bdfe42f0 · outbound
RFTF: Reinforcement Fine-tuning for Embodied Agents with Temporal Feedback Gr-mg: Leveraging partially-annotated data via multi-modal goal-conditioned policy
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 829083fe-0ec1-4bfc-8589-afd904549977 · outbound
RFTF: Reinforcement Fine-tuning for Embodied Agents with Temporal Feedback Manipllm: Embodied multimodal large language model for object-centric robotic manipulation
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4cb70d16-4877-422c-80bd-9af4a09c4f5d · outbound
RFTF: Reinforcement Fine-tuning for Embodied Agents with Temporal Feedback Visual instruction tuning
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a091da66-7b2d-410c-bd73-9b4b9b49537e · outbound
RFTF: Reinforcement Fine-tuning for Embodied Agents with Temporal Feedback RDT-1B: a Diffusion Foundation Model for Bimanual Manipulation
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8b0aca1d-5856-4d18-8f72-b1d6733a5134 · outbound
RFTF: Reinforcement Fine-tuning for Embodied Agents with Temporal Feedback Aligning Cyber Space with Physical World: A Comprehensive Survey on Embodied AI
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 04b7eaec-cacb-4cbb-85b3-b00ad11efe56 · outbound
RFTF: Reinforcement Fine-tuning for Embodied Agents with Temporal Feedback Exploring the Limit of Outcome Reward for Learning Mathematical Reasoning
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 718535f4-56c7-42e8-9b0a-1fe475137173 · outbound
RFTF: Reinforcement Fine-tuning for Embodied Agents with Temporal Feedback A Survey on Vision-Language-Action Models for Embodied AI
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dbe37bec-c507-46d9-954d-854dce1def88 · outbound
RFTF: Reinforcement Fine-tuning for Embodied Agents with Temporal Feedback RoboMatrix: A Skill-centric Hierarchical Framework for Scalable Robot Task Planning and Execution in Open-World
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fd506478-a35b-4b75-8882-5beed27e6043 · outbound
RFTF: Reinforcement Fine-tuning for Embodied Agents with Temporal Feedback Genrl: Multimodal-foundation world models for generalization in embodied agents.Neural Information Processing Systems (NeurIPS), 2024
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a926ba9d-a3ac-48e9-a991-a4c0054f2623 · outbound
RFTF: Reinforcement Fine-tuning for Embodied Agents with Temporal Feedback Calvin: A benchmark for language-conditioned policy learning for long-horizon robot manipulation tasks
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation c1af770b-cf65-43ff-bb77-4e464b3dd087 · outbound
RFTF: Reinforcement Fine-tuning for Embodied Agents with Temporal Feedback Policy invariance under reward transfor- mations: Theory and application to reward shaping
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 789a1e12-6441-4c4f-8265-60bea975c1c6 · outbound
RFTF: Reinforcement Fine-tuning for Embodied Agents with Temporal Feedback Training language models to follow instructions with human feedback
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 44169c1d-9b73-4100-9a24-cad41620917f · outbound
RFTF: Reinforcement Fine-tuning for Embodied Agents with Temporal Feedback Open x- embodiment: Robotic learning datasets and rt-x models: Open x-embodiment collaboration 0
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation cfce2e45-5d38-48af-a08b-d652d6b928e6 · outbound
RFTF: Reinforcement Fine-tuning for Embodied Agents with Temporal Feedback Open x- embodiment: Robotic learning datasets and rt-x models: Open x-embodiment collaboration 0
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a51e8918-7449-4f2a-9cc9-b6d474515a9d · outbound
RFTF: Reinforcement Fine-tuning for Embodied Agents with Temporal Feedback Diffusion Policy Policy Optimization
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 82e608ab-dec4-43f3-b186-d49f41d56a23 · outbound
RFTF: Reinforcement Fine-tuning for Embodied Agents with Temporal Feedback High-Dimensional Continuous Control Using Generalized Advantage Estimation
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 25990dc4-a7e6-43a9-a577-72cf3780da04 · outbound
RFTF: Reinforcement Fine-tuning for Embodied Agents with Temporal Feedback Proximal Policy Optimization Algorithms
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 278423af-23a4-4733-9b88-b9827ebb80bc · outbound
RFTF: Reinforcement Fine-tuning for Embodied Agents with Temporal Feedback DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9fc9c671-1819-4976-bc23-e2e8caef6cd0 · outbound
RFTF: Reinforcement Fine-tuning for Embodied Agents with Temporal Feedback Large language models as general- izable policies for embodied tasks
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 26c447e8-a739-4411-a434-4c72a1249eed · outbound
RFTF: Reinforcement Fine-tuning for Embodied Agents with Temporal Feedback Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f2b7222c-2b51-4f65-ae09-0e8342420b22 · outbound
RFTF: Reinforcement Fine-tuning for Embodied Agents with Temporal Feedback Predictive Inverse Dynamics Models are Scalable Learners for Robotic Manipulation
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 398732cb-7f57-4aaa-af4b-a73d90fe4d7a · outbound
RFTF: Reinforcement Fine-tuning for Embodied Agents with Temporal Feedback LLaMA: Open and Efficient Foundation Language Models
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f1bad90d-60d1-4953-ba4d-71d0bc6ef7e4 · outbound
RFTF: Reinforcement Fine-tuning for Embodied Agents with Temporal Feedback Reft: Rea- soning with reinforced fine-tuning
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1dc21c5f-2644-4847-a832-8a11beb6bd49 · outbound
RFTF: Reinforcement Fine-tuning for Embodied Agents with Temporal Feedback Embodied Task Planning with Large Language Models
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0f1717da-f9b7-49bc-a6e7-40611aea8a24 · outbound
RFTF: Reinforcement Fine-tuning for Embodied Agents with Temporal Feedback Qwen2.5 Technical Report
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 19ca6b3e-7f11-4af9-84f2-e563aad0994b · outbound
RFTF: Reinforcement Fine-tuning for Embodied Agents with Temporal Feedback Fine-tuning large vision-language models as decision-making agents via reinforcement learning
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 0f7a4f8f-8980-45f4-91f8-fbca89ebc6f3 · outbound
RFTF: Reinforcement Fine-tuning for Embodied Agents with Temporal Feedback Rt-2: Vision-language-action models transfer web knowledge to robotic control
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
No inbound Pith citation observations are available.