Pith. sign in

Paper Citation Record · LEDGER

Sample-efficient Deep Reinforcement Learning with Imaginary Rollouts for Human-Robot Interaction

As of 16 August 2026, this Paper Citation Record lists 17 of 17 outbound references and 0 inbound Pith citation observations for arXiv:1908.05546.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
1908.05546 v1

Coverage vector

measured 17 of 17 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-14T13:14:15.607688Z

measured 17 of 17 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

17 of 17 outbound references displayed

  • verified exact0
  • verified fuzzy10
  • unresolved7
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 5e9dcbd2-dce0-4de4-8bd8-d504a2db65ef · outbound

This paper cites Playing Atari with Deep Reinforcement Learning.

Sample-efficient Deep Reinforcement Learning with Imaginary Rollouts for Human-Robot Interaction Playing Atari with Deep Reinforcement Learning

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-14T13:14:15.532755Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T13:14:15.532755Z digest=sha256:fb554885fa4039c5cce678cf66d3da822101d920b883292efaf50819a62befb6

Observation f2947633-9cbf-4623-ba80-1cf07c2e8c16 · outbound

This paper cites End-to-end training of deep visuomotor policies,.

Sample-efficient Deep Reinforcement Learning with Imaginary Rollouts for Human-Robot Interaction End-to-end training of deep visuomotor policies,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-14T13:14:15.537872Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T13:14:15.537872Z digest=sha256:3ef198402cc8c97a77fbe13e7f7d7ec81bd441b1d50b55da5da801fee7f2ecff

Observation d33a40c5-11a7-4dfb-90c0-3f7846ccb8a6 · outbound

This paper cites Imagination- augmented agents for deep reinforcement learning,.

Sample-efficient Deep Reinforcement Learning with Imaginary Rollouts for Human-Robot Interaction Imagination- augmented agents for deep reinforcement learning,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:14:15.847267Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T13:14:15.542274Z digest=sha256:1e620a08843a9b9e9132e70be7f585bda83177610ad0f1b9f7eb73df9cd83b47

Observation 05b7f8e6-a5d5-40be-b52f-3902e919b3f7 · outbound

This paper cites Uncertainty-driven imagination for continuous deep reinforcement learning,.

Sample-efficient Deep Reinforcement Learning with Imaginary Rollouts for Human-Robot Interaction Uncertainty-driven imagination for continuous deep reinforcement learning,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:14:15.831849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T13:14:15.547555Z digest=sha256:266c6f15e00a5b3d4ffec2f6e8dd94e76557cab42b0f19faa9d336ce3b17c0f6

Observation 81b0dbc8-ac12-4382-867d-00a2f69fe217 · outbound

This paper cites Recurrent world models facilitate policy evolution,.

Sample-efficient Deep Reinforcement Learning with Imaginary Rollouts for Human-Robot Interaction Recurrent world models facilitate policy evolution,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:14:15.817553Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T13:14:15.552183Z digest=sha256:e1adc3ee3db3c10441e0b1a5a3ff3d3c0130ec5edbe46d9f33860f2c7ee5865c

Observation 87855e7f-005a-4083-8ef8-0cda25222758 · outbound

This paper cites Model-Based Value Estimation for Efficient Model-Free Reinforcement Learning.

Sample-efficient Deep Reinforcement Learning with Imaginary Rollouts for Human-Robot Interaction Model-Based Value Estimation for Efficient Model-Free Reinforcement Learning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-14T13:14:15.556569Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T13:14:15.556569Z digest=sha256:5b95b24d5588d0883e7baabcfb9f1bcdc717e1c5e9e50c367693c2687884e40e

Observation 8f59a860-c47c-441d-b4cd-75efb55d14b0 · outbound

This paper cites Sample- efficient reinforcement learning with stochastic ensemble value expan- sion,.

Sample-efficient Deep Reinforcement Learning with Imaginary Rollouts for Human-Robot Interaction Sample- efficient reinforcement learning with stochastic ensemble value expan- sion,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:14:15.801412Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T13:14:15.562684Z digest=sha256:1b9cd3ada0af83193f5ee5117dea93401ab92bf68a4f13b3205123448777749e

Observation a53fb595-8682-4328-92e3-dc6ca8059d29 · outbound

This paper cites Continuous control with deep reinforcement learning.

Sample-efficient Deep Reinforcement Learning with Imaginary Rollouts for Human-Robot Interaction Continuous control with deep reinforcement learning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-14T13:14:15.566942Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T13:14:15.566942Z digest=sha256:07bec3e8539fe24d02d094f6eb93c536093418edfc977481d7252c12984c0a49

Observation 5ecd990d-b298-4993-8644-c9815705d966 · outbound

This paper cites Deep reinforcement learn- ing for robotic manipulation with asynchronous off-policy updates,.

Sample-efficient Deep Reinforcement Learning with Imaginary Rollouts for Human-Robot Interaction Deep reinforcement learn- ing for robotic manipulation with asynchronous off-policy updates,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:14:15.786193Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T13:14:15.571524Z digest=sha256:5ce503a580d10e009a2d7b53a76f263a27af08251d25ce0005244664f543ae96

Observation e06276e6-48fe-406f-aa7f-8efade96daa0 · outbound

This paper cites Data-efficient Deep Reinforcement Learning for Dexterous Manipulation.

Sample-efficient Deep Reinforcement Learning with Imaginary Rollouts for Human-Robot Interaction Data-efficient Deep Reinforcement Learning for Dexterous Manipulation

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-14T13:14:15.576035Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T13:14:15.576035Z digest=sha256:f7d8d42105cfaad400028b23042bde2ef61bb67fc661e8b8b3c49e730eba5ab1

Observation 628c781b-e2cc-482b-a32c-fa868d1bf338 · outbound

This paper cites Robot gains social intelligence through multimodal deep reinforcement learn- ing,.

Sample-efficient Deep Reinforcement Learning with Imaginary Rollouts for Human-Robot Interaction Robot gains social intelligence through multimodal deep reinforcement learn- ing,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:14:15.773139Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T13:14:15.580749Z digest=sha256:157ee83f1aea133f9ea47a8462475a58e8262c1cead69609c51467c971bc67ab

Observation 3748524c-5c9e-485d-9c5f-23b07c53779c · outbound

This paper cites an unresolved cited work.

Sample-efficient Deep Reinforcement Learning with Imaginary Rollouts for Human-Robot Interaction Unresolved cited work

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-14T13:14:15.586294Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T13:14:15.586294Z digest=sha256:de4d27bc9dbaa796bcfe31149d6d205b972b952876baff4ceeb2a03bc335c981

Observation 7e0cc391-2626-49b4-8b5d-dd0c637f247f · outbound

This paper cites Auto-encoding variational bayes,.

Sample-efficient Deep Reinforcement Learning with Imaginary Rollouts for Human-Robot Interaction Auto-encoding variational bayes,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-14T13:14:15.590630Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T13:14:15.590630Z digest=sha256:304f07254cff1fcf55194aa2e7024b489453361d8c88e9b4dcd24323603c76a4

Observation 75118bde-7f9b-4275-9245-5e36d4237165 · outbound

This paper cites Mixture density networks,.

Sample-efficient Deep Reinforcement Learning with Imaginary Rollouts for Human-Robot Interaction Mixture density networks,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:14:15.745446Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T13:14:15.595219Z digest=sha256:c67e5ebfbab87756661109d945e722b23ed20f29b275ca0f8416f0fec9ddaf24

Observation 99c464b9-f910-439b-a859-7ee7bbfbe6b7 · outbound

This paper cites A syntactic approach to robot imitation learning using probabilistic activity grammars,.

Sample-efficient Deep Reinforcement Learning with Imaginary Rollouts for Human-Robot Interaction A syntactic approach to robot imitation learning using probabilistic activity grammars,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:14:15.731434Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T13:14:15.599198Z digest=sha256:724f019a0d183ecaf121a07075fc95f0e9cac7a827caa4373800df316ace9b76

Observation e52de069-0158-41ff-8d49-516cad033521 · outbound

This paper cites beta-vae: Learning basic visual concepts with a constrained variational framework,.

Sample-efficient Deep Reinforcement Learning with Imaginary Rollouts for Human-Robot Interaction beta-vae: Learning basic visual concepts with a constrained variational framework,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:14:15.717880Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T13:14:15.603707Z digest=sha256:1ffc78c2fa268c24c72c939f24bbf9224b2d45fdc4e375b2c41dfaab16d8a16b

Observation e6818271-abc2-4edd-8b4a-8f3fa271f1a6 · outbound

This paper cites Robot program- ming by demonstration,.

Sample-efficient Deep Reinforcement Learning with Imaginary Rollouts for Human-Robot Interaction Robot program- ming by demonstration,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T13:14:15.703740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-14T13:14:15.607688Z digest=sha256:b1034406e174ac49c9d6022dc1fed4b6b4e9d141e6921a32a29978f5868dd558

Pith citing papers

No inbound Pith citation observations are available.