Pith. sign in

Paper Citation Record · LEDGER

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation

As of 21 August 2026, this Paper Citation Record lists 28 of 28 outbound references and 0 inbound Pith citation observations for arXiv:2608.03753.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.03753 v1

Coverage vector

measured 28 of 28 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T13:28:59.018485Z

measured 28 of 28 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

28 of 28 outbound references displayed

  • verified exact0
  • verified fuzzy23
  • unresolved5
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 43746d9c-b4ea-492b-9e5f-7bad0a2816e8 · outbound

This paper cites Policy invariance under reward transformations: Theory and application to reward shaping,.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation Policy invariance under reward transformations: Theory and application to reward shaping,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:29:03.025051Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T13:28:55.766491Z digest=sha256:148ce4a445627d77b87b1c3ffa6f8254745dc541c6df7e7c16eedf63090dea4a

Observation 0d79435e-88ae-42e6-ba54-e09b0614cd4d · outbound

This paper cites Recent advances in robot learning from demonstration,.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation Recent advances in robot learning from demonstration,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-05T13:28:55.878454Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T13:28:55.878454Z digest=sha256:d4136db09e7c956da10cc2a10a8f9292c4cc3d1d8c5004c605c085dcd5f0dc15

Observation 61db427f-5a54-472b-8026-6bb91a80dfc2 · outbound

This paper cites Learning by watching: A review of video-based learning approaches for robot manipulation,.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation Learning by watching: A review of video-based learning approaches for robot manipulation,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:29:02.876201Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T13:28:56.021102Z digest=sha256:60a30958e636d39119b5638ab19dd9d768d5ec8aabf78d0957cf09960861f871

Observation 4058901c-69a2-4466-b37b-253bad6ed30c · outbound

This paper cites Xirl: Cross-embodiment inverse reinforcement learning,.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation Xirl: Cross-embodiment inverse reinforcement learning,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:29:02.752237Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T13:28:56.146719Z digest=sha256:c61b3542ff8efe48e84eff64f3e207dd8202716be9b3937c16b7db6501cad5d3

Observation 36637702-3da0-4118-84ae-902f6365dc40 · outbound

This paper cites Hierarchical rein- forcement learning: A comprehensive survey,.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation Hierarchical rein- forcement learning: A comprehensive survey,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:29:02.607907Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T13:28:56.275251Z digest=sha256:5630ffcf03c65c839c1ba86c71dc1f5b55368cb9b3c02a5dc4d13ed107695adc

Observation bfb40785-5915-4fcd-9ed9-6c9800c7e64d · outbound

This paper cites Graph inverse reinforcement learning from diverse videos,.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation Graph inverse reinforcement learning from diverse videos,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:29:02.499103Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T13:28:56.413900Z digest=sha256:d7c250b641aab5ad7d6a7cca0a917ac699725b4da82b6f24e4db1a88681b450a

Observation 4b828829-bcf5-4293-961a-c5e838fc0d28 · outbound

This paper cites Maniskill3: Gpu parallelized robotics simulation and rendering for generalizable embodied ai,.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation Maniskill3: Gpu parallelized robotics simulation and rendering for generalizable embodied ai,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:29:02.286484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T13:28:56.520643Z digest=sha256:45b1c5a25b1115c9c9e69adcb291ce883483890d389b8f73626eb7b4539b8d1a

Observation e1838f07-ba6f-4aa1-8184-b547bdaca0b3 · outbound

This paper cites Time-contrastive networks: Self-supervised learning from video,.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation Time-contrastive networks: Self-supervised learning from video,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:29:02.144740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T13:28:56.619559Z digest=sha256:2cba7f279bf703a63622d4b2b5d9b4b30ffc52465d3f3b01a125f36fa84da75a

Observation 71387b7c-2690-4ca4-a581-18f8582a0e9c · outbound

This paper cites Temporal cycle-consistency learning,.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation Temporal cycle-consistency learning,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:29:01.965500Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T13:28:56.717175Z digest=sha256:117c4a3288504288b6e01fc7c6ca7d783fd5a821c1c9403e4052bcba2d14d52d

Observation 1312b98a-5ea3-4e95-8700-12a9588bfa7b · outbound

This paper cites Learning reward functions for robotic manipulation by observing humans,.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation Learning reward functions for robotic manipulation by observing humans,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:29:01.826215Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T13:28:56.815765Z digest=sha256:0576da16d40aa9412ec8c7e07c11f3bdf64cd7ec22dc21544e3bbc052c95a5a3

Observation 5d2b0d4e-38b0-4c6f-8b3d-d57a6c510fa1 · outbound

This paper cites VIP: Towards Universal Visual Reward and Representation via Value-Implicit Pre-Training.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation VIP: Towards Universal Visual Reward and Representation via Value-Implicit Pre-Training

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-05T13:28:56.958309Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T13:28:56.958309Z digest=sha256:e40a5bd9b57f841c013e7a6d86417af9c05613370b3cde3ee33424d188e539c0

Observation 77192bce-9238-4171-b0b1-2f38aa98215c · outbound

This paper cites Liv: Language-image representations and rewards for robotic control,.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation Liv: Language-image representations and rewards for robotic control,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:29:01.707603Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T13:28:57.039984Z digest=sha256:589e3bd0ec0874cac3c037208571fe454e6975ad75c9e28a7dee5209d8af22e7

Observation 51be2870-4581-4d7f-acfb-f102bb68b4fe · outbound

This paper cites Shadow: Leveraging segmentation masks for cross-embodiment policy transfer,.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation Shadow: Leveraging segmentation masks for cross-embodiment policy transfer,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:29:01.574806Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T13:28:57.119425Z digest=sha256:c28a1f71c959817527a13daa4f7ce2db8f936344d668bfaee94ce4edc1ddf703

Observation 8a6b1dc3-27a9-4627-a6ed-ce6a08289def · outbound

This paper cites Augmented reality for robots (arro): Pointing visuomotor policies towards visual robustness,.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation Augmented reality for robots (arro): Pointing visuomotor policies towards visual robustness,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:29:01.439488Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T13:28:57.214084Z digest=sha256:6901f4aa47ebfda8b9932cc9da40e4167ec3993114cbc0ec352322ad9ea49387

Observation 9f5eee01-10e3-492a-8dad-5da09a919572 · outbound

This paper cites Relay pol- icy learning: Solving long-horizon tasks via imitation and reinforcement learning,.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation Relay pol- icy learning: Solving long-horizon tasks via imitation and reinforcement learning,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:29:01.193167Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T13:28:57.299692Z digest=sha256:c2323765b6c9ddf93dac3d46306bdfb3b6546c8be695823cd92b8399e545185c

Observation 440ee1b4-e931-4f33-89d4-53f14afa9c31 · outbound

This paper cites Taco: Learning task decomposition via temporal alignment for control,.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation Taco: Learning task decomposition via temporal alignment for control,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:29:00.977857Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T13:28:57.453169Z digest=sha256:76b7e8ba8e789e05607024b24f6c4698eec6e25469d2eaac9ba2b868c360420e

Observation b1a97fdf-3329-44f5-80d3-cc1a5514466e · outbound

This paper cites Sequential dexterity: Chain- ing dexterous policies for long-horizon manipulation,.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation Sequential dexterity: Chain- ing dexterous policies for long-horizon manipulation,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:29:00.839721Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T13:28:57.561313Z digest=sha256:e59f7262d7ec3d106dea681f1a62ced069369935dc1cf34474cfd5b0a8c55393

Observation 0dadbb69-86a7-43d5-9c64-9da354f0a724 · outbound

This paper cites Universal visual decomposer: Long-horizon manipu- lation made easy,.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation Universal visual decomposer: Long-horizon manipu- lation made easy,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:29:00.677189Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T13:28:57.694668Z digest=sha256:de055f336192bb40d7c6e43f601389c02b55b8ee29160bd96d47bbca4ab67707

Observation 9fccc9fc-38d3-432e-9a31-df76b144a621 · outbound

This paper cites Do As I Can, Not As I Say: Grounding Language in Robotic Affordances.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation Do As I Can, Not As I Say: Grounding Language in Robotic Affordances

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-05T13:28:57.814979Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T13:28:57.814979Z digest=sha256:f61b82a5f8ecf44733929a454cce49e2dab2aff1c4d4f1cd792a50bc150009b1

Observation ab241ba7-957a-45f7-8383-457fd954c7e5 · outbound

This paper cites RoboGen: Towards unleashing infinite data for auto- mated robot learning via generative simulation,.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation RoboGen: Towards unleashing infinite data for auto- mated robot learning via generative simulation,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:29:00.528056Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T13:28:57.938137Z digest=sha256:17879eb9d556baad7ff019961f22e25b9a5e75e1a8aa958e4f2add407925073a

Observation 95221821-878f-4cb2-8a39-0cf7709fe922 · outbound

This paper cites RoboHorizon: An LLM-Assisted Multi-View World Model for Long-Horizon Robotic Manipulation.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation RoboHorizon: An LLM-Assisted Multi-View World Model for Long-Horizon Robotic Manipulation

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-05T13:28:58.085261Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T13:28:58.085261Z digest=sha256:0d9e6ac09f928debbc86267801b37e150e79180f388070231a910c634c60619a

Observation b1adfc05-3f3a-4804-af3c-4226bff78d06 · outbound

This paper cites Deco: Task decomposition and skill composition for zero-shot generalization in long-horizon 3d manipulation,.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation Deco: Task decomposition and skill composition for zero-shot generalization in long-horizon 3d manipulation,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:29:00.415566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T13:28:58.212697Z digest=sha256:38aea11f8ed3318b796a7b039d258513facb98980a36901d9e5b19e50cce4856

Observation ac624e5f-4399-4b46-a4a2-3a3436fffb19 · outbound

This paper cites Subtask-aware visual reward learning from segmented demonstrations,.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation Subtask-aware visual reward learning from segmented demonstrations,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:29:00.186836Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T13:28:58.311872Z digest=sha256:e8132db0d7de852c4d176ecf607278359eb99fd875222c9bc6ece0ec214b96c7

Observation 24548f0f-d283-47cc-ae15-d4b4f3c497e8 · outbound

This paper cites Egtr: Extracting graph from transformer for scene graph generation,.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation Egtr: Extracting graph from transformer for scene graph generation,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:28:59.911520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T13:28:58.404488Z digest=sha256:0d7d485c58a066b8c4cf3dba4611b5659c27fd1d90eb9c9108103ae980e8812a

Observation cb654b4e-b095-4044-9723-eda3874091b6 · outbound

This paper cites The magical benchmark for robust imitation,.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation The magical benchmark for robust imitation,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:28:59.644349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T13:28:58.576140Z digest=sha256:f7e00bcb2dd6d4a0acabf8e2edfee4188383ed2959c9dbd980abf8b79c5cca62

Observation 62fc6244-3c7e-4d47-b090-c0ce1fad3bc6 · outbound

This paper cites Rlbench: The robot learning benchmark & learning environment,.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation Rlbench: The robot learning benchmark & learning environment,

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-05T13:28:58.710531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T13:28:58.710531Z digest=sha256:ba417e81f6ce8fc95d0e296f875fedd933c06dcf276e54653f985aa0c81b43dc

Observation 810ff14c-eb9c-439e-bb70-95964814fb44 · outbound

This paper cites Soft actor-critic: Off- policy maximum entropy deep reinforcement learning with a stochastic actor,.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation Soft actor-critic: Off- policy maximum entropy deep reinforcement learning with a stochastic actor,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:28:59.512614Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T13:28:58.796384Z digest=sha256:69433b158ae56adf317c16e141f58e455c9d74484285b36f882bdb742488a0eb

Observation c4d12e7b-167a-43f1-bd1e-950c97a1e81e · outbound

This paper cites Graph transformer networks,.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation Graph transformer networks,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:28:59.277875Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-05T13:28:59.018485Z digest=sha256:1a92b4b895785d499f79e308349e4f25e35f90e9ccffbef2796fcb6c069e0e5e

Pith citing papers

No inbound Pith citation observations are available.