Pith. sign in

Paper Citation Record · LEDGER

PixL2R: Guiding Reinforcement Learning Using Natural Language by Mapping Pixels to Rewards

As of 11 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 3 inbound Pith citation observations for arXiv:2007.15543.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2007.15543 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 3 of 3 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-30T22:11:35.277901Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

13
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 50e5cce8-3d6a-4c20-8e17-f9045a9a68e8 · inbound

Do As I Can, Not As I Say: Grounding Language in Robotic Affordances cites this paper.

Do As I Can, Not As I Say: Grounding Language in Robotic Affordances PixL2R: Guiding Reinforcement Learning Using Natural Language by Mapping Pixels to Rewards

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:24:06.368010Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-10T22:24:05.999350Z digest=sha256:3a508557b49f36575d7a916b4ae46e0cd3793014db6351a3d6a6abdf18e86ae0

Observation ba69724d-da44-41c1-9eb7-bc6dd39f1a6b · inbound

Code as Policies: Language Model Programs for Embodied Control cites this paper.

Code as Policies: Language Model Programs for Embodied Control PixL2R: Guiding Reinforcement Learning Using Natural Language by Mapping Pixels to Rewards

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-15T00:38:02.832095Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-15T00:38:02.723424Z digest=sha256:620279fadc8f626f3c72a92be8c5391f6698120851a9e6a482acd0b533ef9191

Observation ddb9b24c-f232-4937-b512-63c970e0b103 · inbound

Robust Instruction Compliance in Cooperative Multi-Agent Reinforcement Learning cites this paper.

Robust Instruction Compliance in Cooperative Multi-Agent Reinforcement Learning PixL2R: Guiding Reinforcement Learning Using Natural Language by Mapping Pixels to Rewards

Reference 23

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T22:15:06.044098Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-06-30T22:11:35.277901Z digest=sha256:7f610cfac08cccb6f0e0e3dfef737a5aa93a00baa43bc6a4ad586e7082bf3229