Pith. sign in

Paper Citation Record · LEDGER

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation

As of 20 August 2026, this Paper Citation Record lists 28 of 28 outbound references and 0 inbound Pith citation observations for arXiv:2608.03753.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.03753 v1

Coverage vector

measured 28 of 28 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T13:28:59.018485Z

measured 28 of 28 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

28 of 28 outbound references displayed

  • verified exact0
  • verified fuzzy23
  • unresolved5
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 43746d9c-b4ea-492b-9e5f-7bad0a2816e8 · outbound

This paper cites Policy invariance under reward transformations: Theory and application to reward shaping,.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation Policy invariance under reward transformations: Theory and application to reward shaping,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:29:03.025051Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-05T13:28:55.766491Z digest=sha256:b2e2aa11f353323fd2dba725fe4845579ab60b75ffabffb71119949826d77233

Observation 0d79435e-88ae-42e6-ba54-e09b0614cd4d · outbound

This paper cites Recent advances in robot learning from demonstration,.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation Recent advances in robot learning from demonstration,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-05T13:28:55.878454Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T13:28:55.878454Z digest=sha256:f4206bba6fb71a0bc6775bf0cdf73fcb3e6ceecf5f741af9c45f2cc23d9c65a0

Observation 61db427f-5a54-472b-8026-6bb91a80dfc2 · outbound

This paper cites Learning by watching: A review of video-based learning approaches for robot manipulation,.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation Learning by watching: A review of video-based learning approaches for robot manipulation,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:29:02.876201Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-05T13:28:56.021102Z digest=sha256:1564c7be2dfe3a84667659cdc2f224a2c711baa250658905a1fb272f53a16301

Observation 4058901c-69a2-4466-b37b-253bad6ed30c · outbound

This paper cites Xirl: Cross-embodiment inverse reinforcement learning,.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation Xirl: Cross-embodiment inverse reinforcement learning,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:29:02.752237Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-05T13:28:56.146719Z digest=sha256:c31231b2521dd21369008505105390529de6a80502f504e73bb8b645994d567d

Observation 36637702-3da0-4118-84ae-902f6365dc40 · outbound

This paper cites Hierarchical rein- forcement learning: A comprehensive survey,.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation Hierarchical rein- forcement learning: A comprehensive survey,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:29:02.607907Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-05T13:28:56.275251Z digest=sha256:d8b9f27227e0f20b569519401e6b0d0a91f8dd9b2b40f703a72784a5054bf7d6

Observation bfb40785-5915-4fcd-9ed9-6c9800c7e64d · outbound

This paper cites Graph inverse reinforcement learning from diverse videos,.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation Graph inverse reinforcement learning from diverse videos,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:29:02.499103Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-05T13:28:56.413900Z digest=sha256:04c4c1e23766e06b10acfc30b21fa205a18f679c31f288afde9f88126769203a

Observation 4b828829-bcf5-4293-961a-c5e838fc0d28 · outbound

This paper cites Maniskill3: Gpu parallelized robotics simulation and rendering for generalizable embodied ai,.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation Maniskill3: Gpu parallelized robotics simulation and rendering for generalizable embodied ai,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:29:02.286484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-05T13:28:56.520643Z digest=sha256:fb51e9ef1d25a9c5480a81e820320477736e86e86d99ed02d47d34f468238590

Observation e1838f07-ba6f-4aa1-8184-b547bdaca0b3 · outbound

This paper cites Time-contrastive networks: Self-supervised learning from video,.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation Time-contrastive networks: Self-supervised learning from video,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:29:02.144740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-05T13:28:56.619559Z digest=sha256:f9d5a46ca1e8d27a6b9c1f4ca881b0702c1b16e9d3db85efaeece7e4ad68232b

Observation 71387b7c-2690-4ca4-a581-18f8582a0e9c · outbound

This paper cites Temporal cycle-consistency learning,.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation Temporal cycle-consistency learning,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:29:01.965500Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-05T13:28:56.717175Z digest=sha256:86520b1cad7841a4ac1611f4e40cb77500ebf3a6a04ab770f1bcb0a2577dc87c

Observation 1312b98a-5ea3-4e95-8700-12a9588bfa7b · outbound

This paper cites Learning reward functions for robotic manipulation by observing humans,.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation Learning reward functions for robotic manipulation by observing humans,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:29:01.826215Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-05T13:28:56.815765Z digest=sha256:44b5ca4ededdfb170b6acb382d900080e4f6069e86a75c44f7227e83a7d663df

Observation 5d2b0d4e-38b0-4c6f-8b3d-d57a6c510fa1 · outbound

This paper cites VIP: Towards Universal Visual Reward and Representation via Value-Implicit Pre-Training.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation VIP: Towards Universal Visual Reward and Representation via Value-Implicit Pre-Training

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-05T13:28:56.958309Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T13:28:56.958309Z digest=sha256:aae173bf918e644f72351b6900b199006877d9481d330dd829c283ce9058b48f

Observation 77192bce-9238-4171-b0b1-2f38aa98215c · outbound

This paper cites Liv: Language-image representations and rewards for robotic control,.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation Liv: Language-image representations and rewards for robotic control,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:29:01.707603Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-05T13:28:57.039984Z digest=sha256:10cd7848b4442f863cb3a063335bf24b50d2d0611c4b1b39629f31043ba36f85

Observation 51be2870-4581-4d7f-acfb-f102bb68b4fe · outbound

This paper cites Shadow: Leveraging segmentation masks for cross-embodiment policy transfer,.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation Shadow: Leveraging segmentation masks for cross-embodiment policy transfer,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:29:01.574806Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-05T13:28:57.119425Z digest=sha256:890069f6ecd781b5f40beb62ead95ca9f2fdd223e6c161bc6e98d0b8d8bfb68d

Observation 8a6b1dc3-27a9-4627-a6ed-ce6a08289def · outbound

This paper cites Augmented reality for robots (arro): Pointing visuomotor policies towards visual robustness,.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation Augmented reality for robots (arro): Pointing visuomotor policies towards visual robustness,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:29:01.439488Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-05T13:28:57.214084Z digest=sha256:b04d2460b3542bcbeb36887aaa55945de2753c57b9f2e63020628dab87c02094

Observation 9f5eee01-10e3-492a-8dad-5da09a919572 · outbound

This paper cites Relay pol- icy learning: Solving long-horizon tasks via imitation and reinforcement learning,.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation Relay pol- icy learning: Solving long-horizon tasks via imitation and reinforcement learning,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:29:01.193167Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-05T13:28:57.299692Z digest=sha256:ac6c03a953735000a20c933101182926651b46fce0c00e1444f1fe8eec273ed9

Observation 440ee1b4-e931-4f33-89d4-53f14afa9c31 · outbound

This paper cites Taco: Learning task decomposition via temporal alignment for control,.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation Taco: Learning task decomposition via temporal alignment for control,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:29:00.977857Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-05T13:28:57.453169Z digest=sha256:14d813035911e87264764709aba06949784f38830b953e10c7e1e6673a3d09e3

Observation b1a97fdf-3329-44f5-80d3-cc1a5514466e · outbound

This paper cites Sequential dexterity: Chain- ing dexterous policies for long-horizon manipulation,.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation Sequential dexterity: Chain- ing dexterous policies for long-horizon manipulation,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:29:00.839721Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-05T13:28:57.561313Z digest=sha256:9917981f99bb61b446d73381a23e0d98c6036708b2e437269592098c2ee1a93e

Observation 0dadbb69-86a7-43d5-9c64-9da354f0a724 · outbound

This paper cites Universal visual decomposer: Long-horizon manipu- lation made easy,.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation Universal visual decomposer: Long-horizon manipu- lation made easy,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:29:00.677189Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-05T13:28:57.694668Z digest=sha256:cfd96fdd2ce18d096b7beaec48eae83c5c19558d5723ccde055819e43f2ae51b

Observation 9fccc9fc-38d3-432e-9a31-df76b144a621 · outbound

This paper cites Do As I Can, Not As I Say: Grounding Language in Robotic Affordances.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation Do As I Can, Not As I Say: Grounding Language in Robotic Affordances

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-05T13:28:57.814979Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T13:28:57.814979Z digest=sha256:b3a7c0f7441f46b65af71a8b7e0d8058affc5168cbdca1a3fc7c615aa1bbb8a9

Observation ab241ba7-957a-45f7-8383-457fd954c7e5 · outbound

This paper cites RoboGen: Towards unleashing infinite data for auto- mated robot learning via generative simulation,.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation RoboGen: Towards unleashing infinite data for auto- mated robot learning via generative simulation,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:29:00.528056Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-05T13:28:57.938137Z digest=sha256:c2802729ef46f1d39d041ecd710b7b9350106a17e1cf0bc32ab45e88c8db5aac

Observation 95221821-878f-4cb2-8a39-0cf7709fe922 · outbound

This paper cites RoboHorizon: An LLM-Assisted Multi-View World Model for Long-Horizon Robotic Manipulation.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation RoboHorizon: An LLM-Assisted Multi-View World Model for Long-Horizon Robotic Manipulation

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-05T13:28:58.085261Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T13:28:58.085261Z digest=sha256:7716ffe5a159f50fa04147f48b2064931eed9e19019bfc3446215bf3245ea241

Observation b1adfc05-3f3a-4804-af3c-4226bff78d06 · outbound

This paper cites Deco: Task decomposition and skill composition for zero-shot generalization in long-horizon 3d manipulation,.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation Deco: Task decomposition and skill composition for zero-shot generalization in long-horizon 3d manipulation,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:29:00.415566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-05T13:28:58.212697Z digest=sha256:5909b8c3fb73f5f8cd1bb5258feb3da3ef42ed84bfd8465ebf23b7bc2d708e9e

Observation ac624e5f-4399-4b46-a4a2-3a3436fffb19 · outbound

This paper cites Subtask-aware visual reward learning from segmented demonstrations,.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation Subtask-aware visual reward learning from segmented demonstrations,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:29:00.186836Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-05T13:28:58.311872Z digest=sha256:6d72343ac4e0c931d116e8093ca84993ed2785048004dc00a843cafee2b2f2b2

Observation 24548f0f-d283-47cc-ae15-d4b4f3c497e8 · outbound

This paper cites Egtr: Extracting graph from transformer for scene graph generation,.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation Egtr: Extracting graph from transformer for scene graph generation,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:28:59.911520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-05T13:28:58.404488Z digest=sha256:e1357d44ef166b3f058feb81f007cd845ee2a052c3db8eba9ef82cd657d08bc8

Observation cb654b4e-b095-4044-9723-eda3874091b6 · outbound

This paper cites The magical benchmark for robust imitation,.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation The magical benchmark for robust imitation,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:28:59.644349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-05T13:28:58.576140Z digest=sha256:e7f6441b5801dc7258c56b729a34d451bd9b816de429669636f174007ef61a9e

Observation 62fc6244-3c7e-4d47-b090-c0ce1fad3bc6 · outbound

This paper cites Rlbench: The robot learning benchmark & learning environment,.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation Rlbench: The robot learning benchmark & learning environment,

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-05T13:28:58.710531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T13:28:58.710531Z digest=sha256:ac9d55ec32da4d6c911cba0c023fd5872f21584e24ce24ffc96d902bdbaeef4f

Observation 810ff14c-eb9c-439e-bb70-95964814fb44 · outbound

This paper cites Soft actor-critic: Off- policy maximum entropy deep reinforcement learning with a stochastic actor,.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation Soft actor-critic: Off- policy maximum entropy deep reinforcement learning with a stochastic actor,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:28:59.512614Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-05T13:28:58.796384Z digest=sha256:fa79b1db5b572dbb9e9e2a1c4369a234b1c67cc54d5e11cfea334c15c53e61b9

Observation c4d12e7b-167a-43f1-bd1e-950c97a1e81e · outbound

This paper cites Graph transformer networks,.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation Graph transformer networks,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:28:59.277875Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-05T13:28:59.018485Z digest=sha256:8a2d83bb6c9e9f4369cd6957d25e5c60795b614acb697ea9f8ea83c0f0e557cc

Pith citing papers

No inbound Pith citation observations are available.