Pith. sign in

Paper Citation Record · LEDGER

RT-Affordance: Affordances are Versatile Intermediate Representations for Robot Manipulation

As of 22 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 21 inbound Pith citation observations for arXiv:2411.02704.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.02704 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 21 of 21 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 21 of 21 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T21:35:26.119165Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-30T13:24:40.645193Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 9d90b9b4-fcfb-46dc-ab87-eedb199260e0 · inbound

Motion Before Action: Diffusing Object Motion as Manipulation Condition cites this paper.

Motion Before Action: Diffusing Object Motion as Manipulation Condition RT-Affordance: Affordances are Versatile Intermediate Representations for Robot Manipulation

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-12T20:28:49.552296Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T20:28:49.552296Z digest=sha256:0af9e2f77e2c5c3390792af12196cd324fde17b061d54ef8a612de7c24494f88

Observation 786a3219-5367-4820-84ac-2271d019952d · inbound

Tra-MoE: Learning Trajectory Prediction Model from Multiple Domains for Adaptive Policy Conditioning cites this paper.

Tra-MoE: Learning Trajectory Prediction Model from Multiple Domains for Adaptive Policy Conditioning RT-Affordance: Affordances are Versatile Intermediate Representations for Robot Manipulation

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-12T15:24:14.332055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:24:14.332055Z digest=sha256:998f9973071014e79650ed75e3d1ca5870bfcf40f72e64a30c26de9bd7b30f02

Observation d5376ed9-0cf7-4a83-a290-622e4646770e · inbound

SeqAfford: Sequential 3D Affordance Reasoning via Multimodal Large Language Model cites this paper.

SeqAfford: Sequential 3D Affordance Reasoning via Multimodal Large Language Model RT-Affordance: Affordances are Versatile Intermediate Representations for Robot Manipulation

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-12T04:22:03.966607Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:22:03.966607Z digest=sha256:7c2c59b3f7a2a8cb930bd706198ee6b780fa93994a6f9e3bb244108c5d81c052

Observation 4af336ec-415b-4c43-bc6f-5b30944c0091 · inbound

You Only Teach Once: Learn One-Shot Bimanual Robotic Manipulation from Video Demonstrations cites this paper.

You Only Teach Once: Learn One-Shot Bimanual Robotic Manipulation from Video Demonstrations RT-Affordance: Affordances are Versatile Intermediate Representations for Robot Manipulation

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-10T15:18:52.426488Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:18:52.426488Z digest=sha256:30e24daba3071e84edb68bff74c59f46df7d80cdd52d024810f4463fd8898b11

Observation 2f8e15a0-b352-4636-9cff-c8c059d0d455 · inbound

Learning Long-Context Diffusion Policies via Past-Token Prediction cites this paper.

Learning Long-Context Diffusion Policies via Past-Token Prediction RT-Affordance: Affordances are Versatile Intermediate Representations for Robot Manipulation

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-15T21:35:26.119165Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:35:26.119165Z digest=sha256:c54edc9ee2e9c0e1591b3f15fc9e6d2388247fd430fbbb69cfa60381373383fe

Observation f254d4c4-3048-4429-9036-184ab0ef71d0 · inbound

SwitchVLA: Execution-Aware Task Switching for Vision-Language-Action Models cites this paper.

SwitchVLA: Execution-Aware Task Switching for Vision-Language-Action Models RT-Affordance: Affordances are Versatile Intermediate Representations for Robot Manipulation

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T11:05:39.353454Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:05:39.353454Z digest=sha256:e38d35a569716df66e05a071ded388623788b67b32a5622e4481a98bbb6df0a3

Observation 1102a049-928a-4b00-8f73-b05afd3ba384 · inbound

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision cites this paper.

Generating 6DoF Object Manipulation Trajectories from Action Description in Egocentric Vision RT-Affordance: Affordances are Versatile Intermediate Representations for Robot Manipulation

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-07T11:03:01.797316Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:03:01.797316Z digest=sha256:4f9e58bb4cc6381342865e67c354ee1e76f0a84609c66463df43a6a599830155

Observation 54ebc49d-9fa4-43e9-b1e1-27180a0e8fe3 · inbound

VLA-OS: Structuring and Dissecting Planning Representations and Paradigms in Vision-Language-Action Models cites this paper.

VLA-OS: Structuring and Dissecting Planning Representations and Paradigms in Vision-Language-Action Models RT-Affordance: Affordances are Versatile Intermediate Representations for Robot Manipulation

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-15T19:10:27.292201Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:10:27.292201Z digest=sha256:24a835bedee66cd430ad22b14420785c5f4be1ceaa26d0d47d930cce6454ae22

Observation 4985e8db-a4d8-44b2-b854-37eb7dad527d · inbound

AimBot: A Simple Auxiliary Visual Cue to Enhance Spatial Awareness of Visuomotor Policies cites this paper.

AimBot: A Simple Auxiliary Visual Cue to Enhance Spatial Awareness of Visuomotor Policies RT-Affordance: Affordances are Versatile Intermediate Representations for Robot Manipulation

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-05T21:44:59.288272Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T21:44:59.288272Z digest=sha256:4fab98fb80f44f16c52efeb0ac477e6b1e1e6664b184514011b93b654c4f802c

Observation 8f8fadbd-8a5a-42f5-833a-911185d0387d · inbound

Spatial Traces: Enhancing VLA Models with Spatial-Temporal Understanding cites this paper.

Spatial Traces: Enhancing VLA Models with Spatial-Temporal Understanding RT-Affordance: Affordances are Versatile Intermediate Representations for Robot Manipulation

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-15T17:34:25.756414Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:34:25.756414Z digest=sha256:12ba440e13332f3509e1ea4073ff79fda121409f490e51aac6ab6736510db873

Observation c26ddd38-e1c3-4657-81ee-c973305896ab · inbound

Embodied-R1: Reinforced Embodied Reasoning for General Robotic Manipulation cites this paper.

Embodied-R1: Reinforced Embodied Reasoning for General Robotic Manipulation RT-Affordance: Affordances are Versatile Intermediate Representations for Robot Manipulation

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-18T22:06:51.627881Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-18T22:04:34.235731Z digest=sha256:ee401ff19bfa361fb771d1493fb2128a9a1101773d888c7822ad9d2c4b7c92bc

Observation 534d43bc-8900-453f-8af0-42a1514f91ba · inbound

O$^3$Afford: One-Shot 3D Object-to-Object Affordance Grounding for Generalizable Robotic Manipulation cites this paper.

O$^3$Afford: One-Shot 3D Object-to-Object Affordance Grounding for Generalizable Robotic Manipulation RT-Affordance: Affordances are Versatile Intermediate Representations for Robot Manipulation

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-04T23:55:43.205205Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T23:55:43.205205Z digest=sha256:5c46cf7ace32299b0dfa20e7bcefad30fc4cfea76d1bbc6c0f97cfa1c068a79a

Observation 43a8f672-2fce-4ae8-9d6c-765790ea6735 · inbound

VLBiMan: Vision-Language Anchored One-Shot Demonstration Enables Generalizable Bimanual Robotic Manipulation cites this paper.

VLBiMan: Vision-Language Anchored One-Shot Demonstration Enables Generalizable Bimanual Robotic Manipulation RT-Affordance: Affordances are Versatile Intermediate Representations for Robot Manipulation

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-18T13:51:25.703777Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-18T13:47:30.943307Z digest=sha256:02c62fcf7b3eb36ffcd6db4abb000c06e946917822cd9730272ec296d5b37366

Observation 4a6b3248-a3a5-4bdc-a601-113b4545baf7 · inbound

InternVLA-M1: A Spatially Guided Vision-Language-Action Framework for Generalist Robot Policy cites this paper.

InternVLA-M1: A Spatially Guided Vision-Language-Action Framework for Generalist Robot Policy RT-Affordance: Affordances are Versatile Intermediate Representations for Robot Manipulation

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-14T20:09:39.851482Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-14T20:09:39.677347Z digest=sha256:b7446ff94623aec601c86cace090fc8e143b945c5135fd4a96b353027a613430

Observation 437faf3a-f7b1-4fbd-a727-44ec03b1bac4 · inbound

LACY: A Vision-Language Model-based Language-Action Cycle for Self-Improving Robotic Manipulation cites this paper.

LACY: A Vision-Language Model-based Language-Action Cycle for Self-Improving Robotic Manipulation RT-Affordance: Affordances are Versatile Intermediate Representations for Robot Manipulation

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-25T07:35:27.866968Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-25T07:34:55.240907Z digest=sha256:686f0d425504da7f6b19a2795495a27544b59609f46c45f603ba9080a8284903

Observation 4505ac3e-c727-44cb-bddb-7c72aee0f127 · inbound

Context-Dependent Affordance Computation in Vision-Language Models cites this paper.

Context-Dependent Affordance Computation in Vision-Language Models RT-Affordance: Affordances are Versatile Intermediate Representations for Robot Manipulation

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-02T23:33:58.233101Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T23:33:58.233101Z digest=sha256:ee6cf45038e561a10295542092d0c52d7c10b250558dc32852643574f4a39658

Observation 5dd4a60f-90cf-4a70-9d2c-a14a90929648 · inbound

RecoverFormer: End-to-End Contact-Aware Recovery for Humanoid Robots cites this paper.

RecoverFormer: End-to-End Contact-Aware Recovery for Humanoid Robots RT-Affordance: Affordances are Versatile Intermediate Representations for Robot Manipulation

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-11T19:36:15.487005Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-08T11:22:24.927334Z digest=sha256:ca7d8126ea041da5784d83bee914008da5a2a17be32ab5cc65c28bad84c9779e

Observation ac6f4619-371b-480e-bc66-8a6050429c33 · inbound

Affordance Agent Harness: Verification-Gated Skill Orchestration cites this paper.

Affordance Agent Harness: Verification-Gated Skill Orchestration RT-Affordance: Affordances are Versatile Intermediate Representations for Robot Manipulation

Reference 47

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T16:06:33.885055Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-09T18:40:53.380512Z digest=sha256:fee32488fa4e86151a150a01b1237b6652bdb20d3df9fc09383ad4fac24b51d8

Observation c85f8f85-7b30-4ad3-82ba-7de17366bc86 · inbound

Affordance Agent Harness: Verification-Gated Skill Orchestration cites this paper.

Affordance Agent Harness: Verification-Gated Skill Orchestration RT-Affordance: Affordances are Versatile Intermediate Representations for Robot Manipulation

Reference 47

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T04:10:55.813105Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-11T01:55:07.248106Z digest=sha256:4878d6a7fd7a31c6bd4f69266875cbb8aa29eb1c231d6fdf2669433f3f037765

Observation c18e144a-27e8-429f-8c08-7b656dee87de · inbound

Bridging the Embodiment Gap: Disentangled Cross-Embodiment Video Editing cites this paper.

Bridging the Embodiment Gap: Disentangled Cross-Embodiment Video Editing RT-Affordance: Affordances are Versatile Intermediate Representations for Robot Manipulation

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-12T11:01:31.522365Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-07T15:44:25.021995Z digest=sha256:8c9a0cb107991471e74f88e3cda9e9f505f65ce86f105d5ec1f7ae6d41b17dd6

Observation af0ede9f-05bd-4daa-b470-895c8034381e · inbound

TAP-VLA: Tactile Annotation Prompting for Vision Language Action Models cites this paper.

TAP-VLA: Tactile Annotation Prompting for Vision Language Action Models RT-Affordance: Affordances are Versatile Intermediate Representations for Robot Manipulation

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-06-30T13:24:40.647723Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-30T09:10:55.016700Z digest=sha256:79bb7412a1d36dd62fd69d348fc3715fef0254006146922a59211a67fb25280e