Pith. sign in

Paper Citation Record · LEDGER

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation

As of 20 August 2026, this Paper Citation Record lists 28 of 28 outbound references and 0 inbound Pith citation observations for arXiv:2608.03753.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.03753 v1

Coverage vector

measured 28 of 28 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T13:28:59.018485Z

measured 28 of 28 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

28 of 28 outbound references displayed

  • verified exact0
  • verified fuzzy23
  • unresolved5
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 43746d9c-b4ea-492b-9e5f-7bad0a2816e8 · outbound

This paper cites Policy invariance under reward transformations: Theory and application to reward shaping,.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation Policy invariance under reward transformations: Theory and application to reward shaping,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:29:03.025051Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T13:28:55.766491Z digest=sha256:13a14036e8d8a49ab87570a3685296d1cca530740495a5100c2fcef8d688b84a

Observation 0d79435e-88ae-42e6-ba54-e09b0614cd4d · outbound

This paper cites Recent advances in robot learning from demonstration,.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation Recent advances in robot learning from demonstration,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-05T13:28:55.878454Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T13:28:55.878454Z digest=sha256:f4206bba6fb71a0bc6775bf0cdf73fcb3e6ceecf5f741af9c45f2cc23d9c65a0

Observation 61db427f-5a54-472b-8026-6bb91a80dfc2 · outbound

This paper cites Learning by watching: A review of video-based learning approaches for robot manipulation,.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation Learning by watching: A review of video-based learning approaches for robot manipulation,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:29:02.876201Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T13:28:56.021102Z digest=sha256:860e2e9b98bc037ee0b2c9175e108d154b150ff11a8ed563d9f952f77f580784

Observation 4058901c-69a2-4466-b37b-253bad6ed30c · outbound

This paper cites Xirl: Cross-embodiment inverse reinforcement learning,.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation Xirl: Cross-embodiment inverse reinforcement learning,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:29:02.752237Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T13:28:56.146719Z digest=sha256:2e0523b69cea061bd2d9b1d2091503c7e69b2acdd503b600f60b7f9a0433713c

Observation 36637702-3da0-4118-84ae-902f6365dc40 · outbound

This paper cites Hierarchical rein- forcement learning: A comprehensive survey,.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation Hierarchical rein- forcement learning: A comprehensive survey,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:29:02.607907Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T13:28:56.275251Z digest=sha256:3d8b14675e2cc5727bdf80a04c6175687d76529b24050432ec27eb1eaf1b5ac2

Observation bfb40785-5915-4fcd-9ed9-6c9800c7e64d · outbound

This paper cites Graph inverse reinforcement learning from diverse videos,.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation Graph inverse reinforcement learning from diverse videos,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:29:02.499103Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T13:28:56.413900Z digest=sha256:1c08975973b68d853848c960104d380d55f18140aaf6f3ab1a0ba4b174bbc08d

Observation 4b828829-bcf5-4293-961a-c5e838fc0d28 · outbound

This paper cites Maniskill3: Gpu parallelized robotics simulation and rendering for generalizable embodied ai,.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation Maniskill3: Gpu parallelized robotics simulation and rendering for generalizable embodied ai,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:29:02.286484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T13:28:56.520643Z digest=sha256:5054b44c2bdcc2dc9c1ba9e37e76b26670fa9d4172be41f105e24b57c782d2b8

Observation e1838f07-ba6f-4aa1-8184-b547bdaca0b3 · outbound

This paper cites Time-contrastive networks: Self-supervised learning from video,.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation Time-contrastive networks: Self-supervised learning from video,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:29:02.144740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T13:28:56.619559Z digest=sha256:bf433b340a60b0a48fa161baca8b121d5d15711efc1a38a23784ed71897dd859

Observation 71387b7c-2690-4ca4-a581-18f8582a0e9c · outbound

This paper cites Temporal cycle-consistency learning,.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation Temporal cycle-consistency learning,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:29:01.965500Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T13:28:56.717175Z digest=sha256:bd7f2a8a5657d79ce97632f97b1b9472cde50b7672ee0c3b82f5481f2c607ba4

Observation 1312b98a-5ea3-4e95-8700-12a9588bfa7b · outbound

This paper cites Learning reward functions for robotic manipulation by observing humans,.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation Learning reward functions for robotic manipulation by observing humans,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:29:01.826215Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T13:28:56.815765Z digest=sha256:fee5e0ca7c834d6f593d5f74ddc26a40e5aa5d30330aa6a7c9d82ada0d4331f9

Observation 5d2b0d4e-38b0-4c6f-8b3d-d57a6c510fa1 · outbound

This paper cites VIP: Towards Universal Visual Reward and Representation via Value-Implicit Pre-Training.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation VIP: Towards Universal Visual Reward and Representation via Value-Implicit Pre-Training

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-05T13:28:56.958309Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T13:28:56.958309Z digest=sha256:aae173bf918e644f72351b6900b199006877d9481d330dd829c283ce9058b48f

Observation 77192bce-9238-4171-b0b1-2f38aa98215c · outbound

This paper cites Liv: Language-image representations and rewards for robotic control,.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation Liv: Language-image representations and rewards for robotic control,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:29:01.707603Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T13:28:57.039984Z digest=sha256:c66701473b676125dc9405425e4f649db25586e661a0777a128a69980210954f

Observation 51be2870-4581-4d7f-acfb-f102bb68b4fe · outbound

This paper cites Shadow: Leveraging segmentation masks for cross-embodiment policy transfer,.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation Shadow: Leveraging segmentation masks for cross-embodiment policy transfer,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:29:01.574806Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T13:28:57.119425Z digest=sha256:fc6183a12714307b81287dc55066e7ce957d602cc094b25b8b44c7afe809899c

Observation 8a6b1dc3-27a9-4627-a6ed-ce6a08289def · outbound

This paper cites Augmented reality for robots (arro): Pointing visuomotor policies towards visual robustness,.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation Augmented reality for robots (arro): Pointing visuomotor policies towards visual robustness,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:29:01.439488Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T13:28:57.214084Z digest=sha256:091b5806246d867ccbc5d1f0743588bfa65661a3ab7edde8d04f2a695b365bdc

Observation 9f5eee01-10e3-492a-8dad-5da09a919572 · outbound

This paper cites Relay pol- icy learning: Solving long-horizon tasks via imitation and reinforcement learning,.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation Relay pol- icy learning: Solving long-horizon tasks via imitation and reinforcement learning,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:29:01.193167Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T13:28:57.299692Z digest=sha256:f44d209963a450a89dda6c4320ec00f013f238a6c15961dc34290231da2a9035

Observation 440ee1b4-e931-4f33-89d4-53f14afa9c31 · outbound

This paper cites Taco: Learning task decomposition via temporal alignment for control,.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation Taco: Learning task decomposition via temporal alignment for control,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:29:00.977857Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T13:28:57.453169Z digest=sha256:3ba44764ba549eae595306bcacd26bcfef07c2dd22bec2f265da1c419b14fe74

Observation b1a97fdf-3329-44f5-80d3-cc1a5514466e · outbound

This paper cites Sequential dexterity: Chain- ing dexterous policies for long-horizon manipulation,.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation Sequential dexterity: Chain- ing dexterous policies for long-horizon manipulation,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:29:00.839721Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T13:28:57.561313Z digest=sha256:6e47932e0ef29d4afe94ee22261a866e5f733a9fcbd65aca3ab3e78a7c44b1a7

Observation 0dadbb69-86a7-43d5-9c64-9da354f0a724 · outbound

This paper cites Universal visual decomposer: Long-horizon manipu- lation made easy,.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation Universal visual decomposer: Long-horizon manipu- lation made easy,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:29:00.677189Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T13:28:57.694668Z digest=sha256:afb66bbc9d6427c65e5287f56daac18af25b23a28341d36872e6abd43e03e550

Observation 9fccc9fc-38d3-432e-9a31-df76b144a621 · outbound

This paper cites Do As I Can, Not As I Say: Grounding Language in Robotic Affordances.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation Do As I Can, Not As I Say: Grounding Language in Robotic Affordances

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-05T13:28:57.814979Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T13:28:57.814979Z digest=sha256:b3a7c0f7441f46b65af71a8b7e0d8058affc5168cbdca1a3fc7c615aa1bbb8a9

Observation ab241ba7-957a-45f7-8383-457fd954c7e5 · outbound

This paper cites RoboGen: Towards unleashing infinite data for auto- mated robot learning via generative simulation,.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation RoboGen: Towards unleashing infinite data for auto- mated robot learning via generative simulation,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:29:00.528056Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T13:28:57.938137Z digest=sha256:12f92f826871c51c6b9eb668a21d08ccd2a9e0131c330d817b34274b439eaafa

Observation 95221821-878f-4cb2-8a39-0cf7709fe922 · outbound

This paper cites RoboHorizon: An LLM-Assisted Multi-View World Model for Long-Horizon Robotic Manipulation.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation RoboHorizon: An LLM-Assisted Multi-View World Model for Long-Horizon Robotic Manipulation

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-05T13:28:58.085261Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T13:28:58.085261Z digest=sha256:7716ffe5a159f50fa04147f48b2064931eed9e19019bfc3446215bf3245ea241

Observation b1adfc05-3f3a-4804-af3c-4226bff78d06 · outbound

This paper cites Deco: Task decomposition and skill composition for zero-shot generalization in long-horizon 3d manipulation,.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation Deco: Task decomposition and skill composition for zero-shot generalization in long-horizon 3d manipulation,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:29:00.415566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T13:28:58.212697Z digest=sha256:2e1ca5c8fc882c278d5f046a0f6203665117bebd93cc0c5005d31826eddde14d

Observation ac624e5f-4399-4b46-a4a2-3a3436fffb19 · outbound

This paper cites Subtask-aware visual reward learning from segmented demonstrations,.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation Subtask-aware visual reward learning from segmented demonstrations,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:29:00.186836Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T13:28:58.311872Z digest=sha256:d6acb3e86c910f448766e64c5ff7b5a40a48aa3e7e906d2f7500f545c2ae2d05

Observation 24548f0f-d283-47cc-ae15-d4b4f3c497e8 · outbound

This paper cites Egtr: Extracting graph from transformer for scene graph generation,.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation Egtr: Extracting graph from transformer for scene graph generation,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:28:59.911520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T13:28:58.404488Z digest=sha256:140b5066357b473545e9df21690009bd73a6b8c86b3fcf48da6468adba1fd746

Observation cb654b4e-b095-4044-9723-eda3874091b6 · outbound

This paper cites The magical benchmark for robust imitation,.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation The magical benchmark for robust imitation,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:28:59.644349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T13:28:58.576140Z digest=sha256:4e9805677d72a7b152df77526bfae8811747853078e3485bcce4fb37f246f705

Observation 62fc6244-3c7e-4d47-b090-c0ce1fad3bc6 · outbound

This paper cites Rlbench: The robot learning benchmark & learning environment,.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation Rlbench: The robot learning benchmark & learning environment,

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-05T13:28:58.710531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T13:28:58.710531Z digest=sha256:ac9d55ec32da4d6c911cba0c023fd5872f21584e24ce24ffc96d902bdbaeef4f

Observation 810ff14c-eb9c-439e-bb70-95964814fb44 · outbound

This paper cites Soft actor-critic: Off- policy maximum entropy deep reinforcement learning with a stochastic actor,.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation Soft actor-critic: Off- policy maximum entropy deep reinforcement learning with a stochastic actor,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:28:59.512614Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T13:28:58.796384Z digest=sha256:d6a3aac9d5bff48ed7ac3f765470b64579e94840370dbac341881f18b37f27c9

Observation c4d12e7b-167a-43f1-bd1e-950c97a1e81e · outbound

This paper cites Graph transformer networks,.

GORDON: Graph-based Object-centric Rewards for Decomposition of Long-Horizon Manipulation Graph transformer networks,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T13:28:59.277875Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T13:28:59.018485Z digest=sha256:602358f588bd864ee8f1263ac962e0543923e85db752ed805646247e33b85487

Pith citing papers

No inbound Pith citation observations are available.