Pith. sign in

Paper Citation Record · LEDGER

QT-Opt: Scalable Deep Reinforcement Learning for Vision-Based Robotic Manipulation

As of 4 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 24 inbound Pith citation observations for arXiv:1806.10293.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
1806.10293 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 24 of 24 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00

measured 24 of 24 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-02T23:48:58.328749Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-10T06:15:00.866473Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 840759de-2a17-442a-beb3-3abaa9eb7c1b · inbound

Open X-Embodiment: Robotic Learning Datasets and RT-X Models cites this paper.

Open X-Embodiment: Robotic Learning Datasets and RT-X Models QT-Opt: Scalable Deep Reinforcement Learning for Vision-Based Robotic Manipulation

Reference 66

Resolution
verified exact
arxiv_id, observed 2026-05-11T17:23:24.391566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-11T17:23:24.255829Z digest=sha256:7744551ab3289a37e4f031c4e4e15eebf624a4c4c2a0a3b9f03cd031b312ca74

Observation 4de130fc-355b-45c0-9e22-a33edfbbccb8 · inbound

CROP: Conservative Reward for Model-based Offline Policy Optimization cites this paper.

CROP: Conservative Reward for Model-based Offline Policy Optimization QT-Opt: Scalable Deep Reinforcement Learning for Vision-Based Robotic Manipulation

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-05-24T06:39:01.267629Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-24T06:37:57.104233Z digest=sha256:b65adc83a278dbcd105d3ca3601caf22802cae40018464306268f70a6ad3c214

Observation f13e0c82-ae6c-44f8-8061-c76ef624d109 · inbound

MimicGen: A Data Generation System for Scalable Robot Learning using Human Demonstrations cites this paper.

MimicGen: A Data Generation System for Scalable Robot Learning using Human Demonstrations QT-Opt: Scalable Deep Reinforcement Learning for Vision-Based Robotic Manipulation

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-05-17T09:47:55.088607Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-17T09:47:54.977716Z digest=sha256:4e2073dab988c4fec56bfee5ae3d2592220ace3d8239791ba43e4c90ad253d77

Observation 82047816-2996-47ac-8b03-cfbf02993db5 · inbound

DROID: A Large-Scale In-The-Wild Robot Manipulation Dataset cites this paper.

DROID: A Large-Scale In-The-Wild Robot Manipulation Dataset QT-Opt: Scalable Deep Reinforcement Learning for Vision-Based Robotic Manipulation

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-11T05:51:18.929401Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-11T05:51:18.508352Z digest=sha256:4521ea602ec42998307914c14b6392853ad82a250535f7b85b178c6840194830

Observation 4f95175f-f48e-492f-8da6-e2311f598eea · inbound

Octo: An Open-Source Generalist Robot Policy cites this paper.

Octo: An Open-Source Generalist Robot Policy QT-Opt: Scalable Deep Reinforcement Learning for Vision-Based Robotic Manipulation

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-11T00:26:15.513470Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-11T00:26:15.163358Z digest=sha256:69c2ea737648d87b3a27afb64861b9a5ef9c218b90c649d7ff7281f1f40703ad

Observation b98b57a8-6ea6-44e0-9ab1-3df0bf808f0d · inbound

RoboCasa: Large-Scale Simulation of Everyday Tasks for Generalist Robots cites this paper.

RoboCasa: Large-Scale Simulation of Everyday Tasks for Generalist Robots QT-Opt: Scalable Deep Reinforcement Learning for Vision-Based Robotic Manipulation

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-12T23:46:30.265942Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-12T23:46:30.198761Z digest=sha256:eed4b01258e0e6472272ddea97624001fdca7b54d6cceff60aeba576be53cdaa

Observation 55ba4592-6f4b-411b-bd2c-22ac4e4d1f65 · inbound

OpenVLA: An Open-Source Vision-Language-Action Model cites this paper.

OpenVLA: An Open-Source Vision-Language-Action Model QT-Opt: Scalable Deep Reinforcement Learning for Vision-Based Robotic Manipulation

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-10T14:46:36.382354Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T14:46:35.942338Z digest=sha256:bbc9f01d96b9dd4b6f5d3a7a10a8258ff87fbf1b5723b2439b44cef0bc0601b5

Observation 7efdcb50-c867-4e03-94a2-6a2acbba6ce9 · inbound

CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation cites this paper.

CogACT: A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation QT-Opt: Scalable Deep Reinforcement Learning for Vision-Based Robotic Manipulation

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-12T07:33:25.574455Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-12T07:33:25.188358Z digest=sha256:af42b40f25955c2d3d7e226d9341212094569f1375a48b5ba869782c240a492b

Observation d783e6fb-cf91-49ec-8ef1-fc865b6e8db4 · inbound

Deep Computerized Adaptive Testing cites this paper.

Deep Computerized Adaptive Testing QT-Opt: Scalable Deep Reinforcement Learning for Vision-Based Robotic Manipulation

Reference 27

Resolution
verified exact
local_arxiv, observed 2026-05-23T02:12:24.027626Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-23T02:11:42.178712Z digest=sha256:a388d846db43c3fabd009bf3ddc81eaa9fd422a0c2f8103379da631d9a6fc5e8

Observation b8235e85-3797-4d73-8757-7677093a5f32 · inbound

HybridVLA: Collaborative Diffusion and Autoregression in a Unified Vision-Language-Action Model cites this paper.

HybridVLA: Collaborative Diffusion and Autoregression in a Unified Vision-Language-Action Model QT-Opt: Scalable Deep Reinforcement Learning for Vision-Based Robotic Manipulation

Reference 84

Resolution
verified exact
local_arxiv, observed 2026-05-15T22:00:48.865695Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-15T22:00:48.667428Z digest=sha256:81b9371f778a8e1774e792fc69c9e4180fa865ef7c9baa78cfc090a69649fc4a

Observation 462a9486-5aaa-48b7-a90e-4a6af5fd45e5 · inbound

GraspVLA: a Grasping Foundation Model Pre-trained on Billion-scale Synthetic Action Data cites this paper.

GraspVLA: a Grasping Foundation Model Pre-trained on Billion-scale Synthetic Action Data QT-Opt: Scalable Deep Reinforcement Learning for Vision-Based Robotic Manipulation

Reference 55

Resolution
verified exact
local_arxiv, observed 2026-05-17T20:55:52.240827Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-17T20:55:52.109166Z digest=sha256:6615fc5a7ba2d67c32f1563e064864afab0dd91ebd8f023e447cd150791ba81a

Observation 897fa41e-adc0-437e-9c31-68dfd8d45e93 · inbound

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning cites this paper.

VLA-RL: Towards Masterful and General Robotic Manipulation with Scalable Reinforcement Learning QT-Opt: Scalable Deep Reinforcement Learning for Vision-Based Robotic Manipulation

Reference 35

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T12:55:40.423602Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T12:55:40.245908Z digest=sha256:5056c9c85727dfceb07f6f9bfc7b8f2aa9d3647b44f08e93b35944b076420b07

Observation 0a795517-5ea7-4ca0-922b-cf519cd4f3a5 · inbound

$\pi^{*}_{0.6}$: a VLA That Learns From Experience cites this paper.

$\pi^{*}_{0.6}$: a VLA That Learns From Experience QT-Opt: Scalable Deep Reinforcement Learning for Vision-Based Robotic Manipulation

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-12T10:34:59.265492Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-12T10:34:59.134604Z digest=sha256:a4740c5c9c45eab0e62877fe8110a16c1d13973105e7ffe8e66861f3c31274d0

Observation 0bf2d53a-8726-4b2a-a61d-2a2d6bd29683 · inbound

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models cites this paper.

Beyond Imitation: Reinforcement Learning-Based Sim-Real Co-Training for VLA Models QT-Opt: Scalable Deep Reinforcement Learning for Vision-Based Robotic Manipulation

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-02T23:48:58.328749Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:48:58.328749Z digest=sha256:e06feb53df49af7a8fba2d345f4bc8dc2f07c83e8e4eff1d9e9636f0afb0ca49

Observation 08e23531-bf65-4e6d-9080-7a517cb3414b · inbound

ALOE: Action-Level Off-Policy Evaluation for Vision-Language-Action Model Post-Training cites this paper.

ALOE: Action-Level Off-Policy Evaluation for Vision-Language-Action Model Post-Training QT-Opt: Scalable Deep Reinforcement Learning for Vision-Based Robotic Manipulation

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-02T23:47:55.298132Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:47:55.298132Z digest=sha256:862f1ded4f44311be4f67943f3b8bef769554824ac20f2b48a8a8ab58e3ac7e8

Observation 6b700035-24c1-42c0-8a75-0d1ccd1ef06a · inbound

Unmasking the Illusion of Embodied Reasoning in Vision-Language-Action Models cites this paper.

Unmasking the Illusion of Embodied Reasoning in Vision-Language-Action Models QT-Opt: Scalable Deep Reinforcement Learning for Vision-Based Robotic Manipulation

Reference 60

Resolution
verified exact
arxiv_id, observed 2026-05-11T11:56:29.320668Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T04:27:18.284698Z digest=sha256:dd5c4f8de9f1a4eb8d33a374978628549163adcd987bc20eb400c3a12be8671e

Observation 7d4d2e6f-2e01-4ba4-957f-353f646f32d7 · inbound

LaST-R1: Reinforcing Robotic Manipulation via Adaptive Physical Latent Reasoning cites this paper.

LaST-R1: Reinforcing Robotic Manipulation via Adaptive Physical Latent Reasoning QT-Opt: Scalable Deep Reinforcement Learning for Vision-Based Robotic Manipulation

Reference 62

Resolution
verified exact
arxiv_id, observed 2026-05-12T10:31:30.076897Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-07T05:47:17.494531Z digest=sha256:0423e3f58efacc23d64ba9233be033d38498cec653eca26b11c6dc78a1a480e6

Observation a2e8c942-994e-4981-ba89-82ad0bf793e5 · inbound

LaST-R1: Reinforcing Robotic Manipulation via Adaptive Physical Latent Reasoning cites this paper.

LaST-R1: Reinforcing Robotic Manipulation via Adaptive Physical Latent Reasoning QT-Opt: Scalable Deep Reinforcement Learning for Vision-Based Robotic Manipulation

Reference 62

Resolution
verified exact
arxiv_id, observed 2026-05-11T22:17:03.486360Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T03:00:26.352130Z digest=sha256:afaf982f6e0fb49ba4218f6b3806f062053613d8b70fd8eb291bdd1b3dcd8132

Observation 1ceee35e-da16-454e-9b98-ac75447cb6a1 · inbound

World Action Models: The Next Frontier in Embodied AI cites this paper.

World Action Models: The Next Frontier in Embodied AI QT-Opt: Scalable Deep Reinforcement Learning for Vision-Based Robotic Manipulation

Reference 118

Resolution
verified exact
arxiv_id, observed 2026-05-13T05:07:18.197943Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T05:01:16.802019Z digest=sha256:fb71afd5264c6436884f3872c341685d7d9de99feee6bc58093a257e74f83507

Observation bb9e507e-cf42-4f5c-b0f5-634804e9c5ab · inbound

Coarse-to-Control: Action-Token Planning for Vision-Language-Action Models cites this paper.

Coarse-to-Control: Action-Token Planning for Vision-Language-Action Models QT-Opt: Scalable Deep Reinforcement Learning for Vision-Based Robotic Manipulation

Reference 41

Resolution
verified exact
local_arxiv, observed 2026-06-27T22:01:20.745662Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-27T21:57:19.195413Z digest=sha256:59d7e54c014c7ca9156d5f41e7f791e960f8a377c232b454a7574a9135b80a44

Observation ba3c2609-49e7-4c3c-a6e0-b31cef026d2b · inbound

LaST-HD: Learning Latent Physical Reasoning from Scalable Human Data for Robot Manipulation cites this paper.

LaST-HD: Learning Latent Physical Reasoning from Scalable Human Data for Robot Manipulation QT-Opt: Scalable Deep Reinforcement Learning for Vision-Based Robotic Manipulation

Reference 73

Resolution
verified exact
local_arxiv, observed 2026-07-04T11:29:50.820061Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-26T07:58:17.225491Z digest=sha256:c1d70156ab8713628ef1320c9cd65b71a30e549d8d7b974c2309a88a698a4730

Observation a84e08f0-7034-40b6-a6fb-ba8efa6d0740 · inbound

InSight: Self-Guided Skill Acquisition via Steerable VLAs cites this paper.

InSight: Self-Guided Skill Acquisition via Steerable VLAs QT-Opt: Scalable Deep Reinforcement Learning for Vision-Based Robotic Manipulation

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-07-04T16:49:58.292475Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-26T00:10:51.721485Z digest=sha256:b86fb6e3392965426d0a3c3fada96fda33db579e8e7cadef13fc26947eb887d3

Observation 362013dd-6d24-45fc-84a3-32dbffafc886 · inbound

Predicting Grasping Compliance in Robotic Hands through Analytical-Model-Informed Neural Networks cites this paper.

Predicting Grasping Compliance in Robotic Hands through Analytical-Model-Informed Neural Networks QT-Opt: Scalable Deep Reinforcement Learning for Vision-Based Robotic Manipulation

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-01T17:46:31.844616Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T17:46:31.844616Z digest=sha256:a5d36458a2945ea427b0cf674a76e96689b68786eedc0104d7e3332b7a13db96

Observation a2c7b3eb-8ab6-4d06-b6fd-a41bd7f2c470 · inbound

Data Pyramid for Embodied Manipulation cites this paper.

Data Pyramid for Embodied Manipulation QT-Opt: Scalable Deep Reinforcement Learning for Vision-Based Robotic Manipulation

Reference 186

Resolution
unresolved
no resolver link, observed 2026-07-31T06:18:55.667721Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T06:18:55.667721Z digest=sha256:d06099f52d1c94ab2584f18397b4d111385a16aa861ed8d74c8b3b3c47c1166d