Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T20:59:41.281514Z
Paper Citation Record · LEDGER
As of 16 August 2026, this Paper Citation Record lists 50 of 50 outbound references and 0 inbound Pith citation observations for arXiv:2505.11221.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T20:59:41.281514Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
50 of 50 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 12b45feb-cf55-4525-be8a-32e6644f08ae · outbound
Sample Efficient Reinforcement Learning via Large Vision Language Model Distillation Mastering the game of go with deep neural networks and tree search,
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 374f6784-dfc4-457b-873e-63da11c83dc0 · outbound
Sample Efficient Reinforcement Learning via Large Vision Language Model Distillation Advanced planning for autonomous vehicles using reinforcement learning and deep inverse reinforcement learning,
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 2ada002b-50ab-4b48-999c-61eff2ab90db · outbound
Sample Efficient Reinforcement Learning via Large Vision Language Model Distillation Qt-opt: Scalable deep reinforcement learning for vision-based robotic manipu- lation,
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 88899db3-57af-4c87-973e-c91341cb9213 · outbound
Sample Efficient Reinforcement Learning via Large Vision Language Model Distillation Mastering atari, go, chess and shogi by planning with a learned model,
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation e761d655-2eb4-46a3-b0a8-a6924722d97d · outbound
Sample Efficient Reinforcement Learning via Large Vision Language Model Distillation Hindsight goal ranking on replay buffer for sparse reward environment,
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation cacf23d0-8b6d-43f3-afdb-3bcaa311684c · outbound
Sample Efficient Reinforcement Learning via Large Vision Language Model Distillation Utilizing skipped frames in action repeats for improving sample efficiency in reinforcement learning,
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 23b1ea9c-e676-46a9-9cc5-63fada2bb3e8 · outbound
Sample Efficient Reinforcement Learning via Large Vision Language Model Distillation Predictive coding for decision transformer,
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 6a11b1b3-ac27-41a0-9bfa-2f64ad4911f9 · outbound
Sample Efficient Reinforcement Learning via Large Vision Language Model Distillation Dota 2 with Large Scale Deep Reinforcement Learning
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e12c152e-d9b1-43e1-abd5-c26168028409 · outbound
Sample Efficient Reinforcement Learning via Large Vision Language Model Distillation A comprehensive survey on safe reinforce- ment learning,
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 21454ae8-9264-4d55-b328-08ade17e5816 · outbound
Sample Efficient Reinforcement Learning via Large Vision Language Model Distillation No falls, no resets: Reliable humanoid behavior in the darpa robotics challenge,
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 2819d1b1-c94e-4d10-aebe-a8f57790659b · outbound
Sample Efficient Reinforcement Learning via Large Vision Language Model Distillation Deep reinforcement learning: A brief survey,
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 004f6a71-28d9-429f-8bff-8b3771b7e502 · outbound
Sample Efficient Reinforcement Learning via Large Vision Language Model Distillation Language models are unsupervised multitask learners,
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1a12b399-4e20-4909-9477-33ca8df11950 · outbound
Sample Efficient Reinforcement Learning via Large Vision Language Model Distillation Language Models are Few-Shot Learners
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3b9b039e-4a3d-4c01-98fe-c08c0e227e66 · outbound
Sample Efficient Reinforcement Learning via Large Vision Language Model Distillation On the Opportunities and Risks of Foundation Models
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 309e9cdc-7ec7-45f7-8d12-f298c6c72fd6 · outbound
Sample Efficient Reinforcement Learning via Large Vision Language Model Distillation Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e853ad67-0be1-4149-8690-34eeb5e2efdc · outbound
Sample Efficient Reinforcement Learning via Large Vision Language Model Distillation Openai. gpt-4v,
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 73121656-f8a6-48f4-94f6-cee085346f26 · outbound
Sample Efficient Reinforcement Learning via Large Vision Language Model Distillation Do As I Can, Not As I Say: Grounding Language in Robotic Affordances
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0887c92a-18f5-4763-9aa5-0fb8b905530d · outbound
Sample Efficient Reinforcement Learning via Large Vision Language Model Distillation Language models as zero-shot planners: Extracting actionable knowledge for embodied agents,
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation fcbea04e-e131-4401-9fec-3787d0fbcbee · outbound
Sample Efficient Reinforcement Learning via Large Vision Language Model Distillation Large language models as generalizable policies for embodied tasks,
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation e1b7d85e-be24-4f27-acb9-2c7dc0e7d20f · outbound
Sample Efficient Reinforcement Learning via Large Vision Language Model Distillation Fine-Tuning Large Vision-Language Models as Decision-Making Agents via Reinforcement Learning
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 399ad55a-72b2-4784-9348-0ab362b99e56 · outbound
Sample Efficient Reinforcement Learning via Large Vision Language Model Distillation LLaMA: Open and Efficient Foundation Language Models
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 72b92636-9d51-4181-8bd3-1892a3bd2cb7 · outbound
Sample Efficient Reinforcement Learning via Large Vision Language Model Distillation Palm: Scaling language modeling with pathways,
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 98407cff-95d8-42a0-8cb6-470690c23c8a · outbound
Sample Efficient Reinforcement Learning via Large Vision Language Model Distillation Distilling the Knowledge in a Neural Network
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 304c07cf-6812-4672-98ba-78a28767842e · outbound
Sample Efficient Reinforcement Learning via Large Vision Language Model Distillation Proximal Policy Optimization Algorithms
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9375b1ad-ebf4-4abd-8857-850b3cbece2d · outbound
Sample Efficient Reinforcement Learning via Large Vision Language Model Distillation Asynchronous methods for deep reinforcement learning,
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 7ae865fb-960e-4774-9e16-aa28089e8908 · outbound
Sample Efficient Reinforcement Learning via Large Vision Language Model Distillation Guiding pretraining in reinforcement learning with large language models,
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation cb90abaf-fed4-450d-a731-507c58d1852f · outbound
Sample Efficient Reinforcement Learning via Large Vision Language Model Distillation Progprompt: Generating situated robot task plans using large language models,
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 921fad93-7cc5-4c14-8671-f21cf653fcb2 · outbound
Sample Efficient Reinforcement Learning via Large Vision Language Model Distillation Inner Monologue: Embodied Reasoning through Planning with Language Models
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ba6aa2d7-a84e-4c0f-88e2-76eb8fa93d00 · outbound
Sample Efficient Reinforcement Learning via Large Vision Language Model Distillation Enabling intelligent interactions between an agent and an llm: A reinforcement learning approach,
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 90e9c8ea-ab1c-4d35-9c1c-09e51f68285c · outbound
Sample Efficient Reinforcement Learning via Large Vision Language Model Distillation Grounding large language models in interactive environments with online reinforcement learning,
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 627dc17b-390e-4f3b-9360-b7d22f48d547 · outbound
Sample Efficient Reinforcement Learning via Large Vision Language Model Distillation Introducing gemini: our largest and most capable ai model,
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation dfcedf34-5a29-42cc-844b-085a800a7dfd · outbound
Sample Efficient Reinforcement Learning via Large Vision Language Model Distillation Qwen2.5 Technical Report
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8db894f0-eb4a-4bec-93d9-24ca9e2e60c6 · outbound
Sample Efficient Reinforcement Learning via Large Vision Language Model Distillation The claude 3 model family: Opus, sonnet, haiku
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c5d046f5-9430-464c-a32b-38894c703de5 · outbound
Sample Efficient Reinforcement Learning via Large Vision Language Model Distillation Plan-Seq-Learn: Language Model Guided RL for Solving Long Horizon Robotics Tasks
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f9347ac1-3a04-4b37-87f8-a2a89600d53f · outbound
Sample Efficient Reinforcement Learning via Large Vision Language Model Distillation Reward Design with Language Models
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4541a172-4582-4d2f-91c1-6f0dc5d04da7 · outbound
Sample Efficient Reinforcement Learning via Large Vision Language Model Distillation Vision-Language Models are Zero-Shot Reward Models for Reinforcement Learning
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d6e0d8eb-dcc5-43a1-8cd8-7c1850be8884 · outbound
Sample Efficient Reinforcement Learning via Large Vision Language Model Distillation RL-VLM-F: Reinforcement Learning from Vision Language Foundation Model Feedback
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 79cb93c5-bbab-4fc9-a2e6-0b13b27de1a8 · outbound
Sample Efficient Reinforcement Learning via Large Vision Language Model Distillation Voyager: An Open-Ended Embodied Agent with Large Language Models
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aa1c2f58-d289-44cb-b677-69551e889f3e · outbound
Sample Efficient Reinforcement Learning via Large Vision Language Model Distillation Bootstrap Your Own Skills: Learning to Solve New Tasks with Large Language Model Guidance
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8bdc0382-62f0-44ac-9437-0d12274b4263 · outbound
Sample Efficient Reinforcement Learning via Large Vision Language Model Distillation Embodiedgpt: Vision-language pre-training via embodied chain of thought,
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 0c388fba-6c6c-49d7-862e-aad30cc1e6e5 · outbound
Sample Efficient Reinforcement Learning via Large Vision Language Model Distillation Vision-Language Models Provide Promptable Representations for Reinforcement Learning
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 86f2b24e-5c2f-49cb-a5e0-0fa78e362dbe · outbound
Sample Efficient Reinforcement Learning via Large Vision Language Model Distillation Policy distillation,
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation ff176453-e3da-4a72-8bc3-f890816a73d3 · outbound
Sample Efficient Reinforcement Learning via Large Vision Language Model Distillation Actor-Mimic: Deep Multitask and Transfer Reinforcement Learning
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ae662865-6a6f-42fe-b1ad-6f57d153b6d1 · outbound
Sample Efficient Reinforcement Learning via Large Vision Language Model Distillation Guided policy search,
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 44c0d836-a331-474c-9338-6de3c43ae6bb · outbound
Sample Efficient Reinforcement Learning via Large Vision Language Model Distillation Neural network dynamics for model-based deep reinforcement learning with model-free fine-tuning,
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 497f7ed1-13c2-4fea-a19c-dea6c2b02efc · outbound
Sample Efficient Reinforcement Learning via Large Vision Language Model Distillation Agents teaching agents: a survey on inter-agent transfer learning,
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 9424baa2-0bc3-4fa1-b8f1-3a8f7965de6d · outbound
Sample Efficient Reinforcement Learning via Large Vision Language Model Distillation Reincarnating reinforcement learning: Reusing prior computation to accelerate progress,
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 7e66c295-55df-46b2-a47f-c04b0a8c8b1d · outbound
Sample Efficient Reinforcement Learning via Large Vision Language Model Distillation Kickstarting Deep Reinforcement Learning
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7ee6f8c5-bf57-4e97-bcb9-0211e6784704 · outbound
Sample Efficient Reinforcement Learning via Large Vision Language Model Distillation When does label smoothing help?
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 1a679df0-1267-4705-a418-1b30313f31c6 · outbound
Sample Efficient Reinforcement Learning via Large Vision Language Model Distillation Minigrid & miniworld: Modular & customizable reinforcement learning environments for goal- oriented tasks,
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
No inbound Pith citation observations are available.