Pith. sign in

Paper Citation Record · LEDGER

Grounding by Remembering: Cross-Scene and In-Scene Memory for 3D Functional Affordances

As of 6 August 2026, this Paper Citation Record lists 32 of 32 outbound references and 0 inbound Pith citation observations for arXiv:2605.11616.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2605.11616 v1

Coverage vector

measured 32 of 32 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-13T01:55:47.136350Z

measured 32 of 32 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

32 of 32 outbound references displayed

  • verified exact5
  • verified fuzzy20
  • unresolved6
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 71ac32d7-4e92-4b49-8e8c-33b60c45e84c · outbound

This paper cites Referit3d: Neural listeners for fine-grained 3D object identification in real-world scenes.

Grounding by Remembering: Cross-Scene and In-Scene Memory for 3D Functional Affordances Referit3d: Neural listeners for fine-grained 3D object identification in real-world scenes

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T14:02:51.612744Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-13T01:55:47.136350Z digest=sha256:4be1386623c151519c84bf357da28c785d1e085254aca3ba851d99031681fa29

Observation 43bba1dd-7904-483d-98ca-03249c41ce08 · outbound

This paper cites Qwen2.5-VL Technical Report.

Grounding by Remembering: Cross-Scene and In-Scene Memory for 3D Functional Affordances Qwen2.5-VL Technical Report

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-05-13T01:57:05.024500Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-13T01:55:47.136350Z digest=sha256:296d79e95eaf06c313e9cc8be6275da5c9ea83dbc0de25a483b2520d60c763e3

Observation 30679490-8267-4466-b8a7-44868303cbe6 · outbound

This paper cites an unresolved cited work.

Grounding by Remembering: Cross-Scene and In-Scene Memory for 3D Functional Affordances Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-05-13T14:02:51.667733Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-13T01:55:47.136350Z digest=sha256:fcb6ac33c0d5ef4af7a5b46051043cc4435cfd60daae55050c2222a7ac977a06

Observation b0b5e929-519e-4f25-8237-7d129684c488 · outbound

This paper cites SAM 3: Segment Anything with Concepts.

Grounding by Remembering: Cross-Scene and In-Scene Memory for 3D Functional Affordances SAM 3: Segment Anything with Concepts

Reference 4

Resolution
verified exact
local_arxiv, observed 2026-05-13T01:57:05.041722Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-13T01:55:47.136350Z digest=sha256:ce64122e02aecc516ba3688a0e649845cbdc66f77ffe30d6d0e761b23ce1a13c

Observation 4d85d5dd-9e29-4a3d-a5d2-b6f71dc094f8 · outbound

This paper cites ScanRefer: 3D object localization in RGB-D scans using natural language.

Grounding by Remembering: Cross-Scene and In-Scene Memory for 3D Functional Affordances ScanRefer: 3D object localization in RGB-D scans using natural language

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T14:02:51.669722Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-13T01:55:47.136350Z digest=sha256:98c15549a7d9ec089baa834153eddb39109dbe95dc262f871401fa1b0c02f272

Observation 0c5f9b7b-8eeb-41d7-90a7-ae98c313ebc4 · outbound

This paper cites Functional- ity understanding and segmentation in 3d scenes.

Grounding by Remembering: Cross-Scene and In-Scene Memory for 3D Functional Affordances Functional- ity understanding and segmentation in 3d scenes

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T14:02:51.658672Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-13T01:55:47.136350Z digest=sha256:33952a29d14aa2c47e0d50fb0a3ef50e8c9d7ec4652014e5ebd03040473361b6

Observation 91993a69-36e3-44c5-ac51-35aa2b8e3e6b · outbound

This paper cites SceneFun3D: Fine-grained functionality and affordance understanding in 3D scenes.

Grounding by Remembering: Cross-Scene and In-Scene Memory for 3D Functional Affordances SceneFun3D: Fine-grained functionality and affordance understanding in 3D scenes

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T14:02:51.656384Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-13T01:55:47.136350Z digest=sha256:b7522e8523661108e4bcbd2f9a7f8aaa5c040f72ea88873aa2a289b7ffdf8de2

Observation 2233442e-4301-4e09-bd69-eee70028092f · outbound

This paper cites 3D affordancenet: A benchmark for visual object affordance understanding.

Grounding by Remembering: Cross-Scene and In-Scene Memory for 3D Functional Affordances 3D affordancenet: A benchmark for visual object affordance understanding

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T14:02:51.665810Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-13T01:55:47.136350Z digest=sha256:e8e1a99e0f845e54e4477c30b0407bf13c89a8d1d952f6d8410780dacfc0700b

Observation 0758f5ec-3a7f-4ae3-8efd-8abd08926d09 · outbound

This paper cites Learning 2d invariant affordance knowledge for 3d affordance grounding.

Grounding by Remembering: Cross-Scene and In-Scene Memory for 3D Functional Affordances Learning 2d invariant affordance knowledge for 3d affordance grounding

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T14:02:51.671656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-13T01:55:47.136350Z digest=sha256:6ac01a414d872bca3d44113c8a0a538f0ed412f4887eab6c68aeee4625872ac4

Observation c2a60e94-c9e7-4603-9565-70fdf1dd19da · outbound

This paper cites Task-aware 3d affordance segmentation via 2d guidance and geometric refinement.

Grounding by Remembering: Cross-Scene and In-Scene Memory for 3D Functional Affordances Task-aware 3d affordance segmentation via 2d guidance and geometric refinement

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T14:02:51.673612Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-13T01:55:47.136350Z digest=sha256:33e9930b76b75db25e114d77af05e4a0a3064a5e98fc63d175dac0d5ad565259

Observation 64959bbc-5c0a-49c7-8999-802b7cb345d5 · outbound

This paper cites Segment3d: Learning fine-grained class-agnostic 3d segmentation without manual labels.

Grounding by Remembering: Cross-Scene and In-Scene Memory for 3D Functional Affordances Segment3d: Learning fine-grained class-agnostic 3d segmentation without manual labels

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T14:02:51.651063Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-13T01:55:47.136350Z digest=sha256:b9adf484c3a16974f893080eb9ee6684d603d0b129ae82d8a4b1f886a6583f65

Observation 013b017a-74f3-4c7a-8ca0-fd8a252bca29 · outbound

This paper cites OpenIns3D: Snap and lookup for 3D open-vocabulary instance segmentation.

Grounding by Remembering: Cross-Scene and In-Scene Memory for 3D Functional Affordances OpenIns3D: Snap and lookup for 3D open-vocabulary instance segmentation

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T14:02:51.642587Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-13T01:55:47.136350Z digest=sha256:866381f3b318dc9506a514ecd894a837395a8838138028427f4a2b5789867764

Observation 79b004ce-f568-4bfe-bbfc-82e5a9ffb4e8 · outbound

This paper cites ConceptFusion: Open-set multimodal 3D mapping.

Grounding by Remembering: Cross-Scene and In-Scene Memory for 3D Functional Affordances ConceptFusion: Open-set multimodal 3D mapping

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T14:02:51.639395Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-13T01:55:47.136350Z digest=sha256:6dd81f61aef4c987ab7c6cd8f8e1c2f908ff27b68780fb70f47f8e230f1c3240

Observation 2686a365-f0d8-431b-a8de-c61d9b44afa9 · outbound

This paper cites LERF: Language embedded radiance fields.

Grounding by Remembering: Cross-Scene and In-Scene Memory for 3D Functional Affordances LERF: Language embedded radiance fields

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T14:02:51.646944Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-13T01:55:47.136350Z digest=sha256:9c90cc2e88468095130343fdb6b647d7033abfff35d42c1554812375cb1d717f

Observation 5a496579-9d83-4021-8807-4a19bafb189d · outbound

This paper cites Berg, Wan-Yen Lo, Piotr Dollár, and Ross Girshick.

Grounding by Remembering: Cross-Scene and In-Scene Memory for 3D Functional Affordances Berg, Wan-Yen Lo, Piotr Dollár, and Ross Girshick

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T14:02:51.648984Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-13T01:55:47.136350Z digest=sha256:6157b02e827051253892937e5983ae653dbab2fd37aa28be91a1306ca7211aa9

Observation 488a197a-eb2b-49cf-b881-c9112f97e06f · outbound

This paper cites Where2Act: From pixels to actions for articulated 3D objects.

Grounding by Remembering: Cross-Scene and In-Scene Memory for 3D Functional Affordances Where2Act: From pixels to actions for articulated 3D objects

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T14:02:51.653249Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-13T01:55:47.136350Z digest=sha256:8de2f4ef0aab17e721b9e90d6c3405b585470666cfe6b9544871b55d4edd2f56

Observation 70443655-d390-4d94-adfc-168b1ff210b5 · outbound

This paper cites OpenScene: 3D scene understanding with open vocabularies.

Grounding by Remembering: Cross-Scene and In-Scene Memory for 3D Functional Affordances OpenScene: 3D scene understanding with open vocabularies

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T14:02:51.637329Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-13T01:55:47.136350Z digest=sha256:4f9edfbdc448c80d51563dfce29c474d464f85f6c74e9057109f3763f6e1093e

Observation dc68a5f7-7383-47b7-9d98-bbbf31cb1a9f · outbound

This paper cites SAM 2: Segment Anything in Images and Videos.

Grounding by Remembering: Cross-Scene and In-Scene Memory for 3D Functional Affordances SAM 2: Segment Anything in Images and Videos

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-05-13T01:57:05.035670Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-13T01:55:47.136350Z digest=sha256:693ad2c1a038594e5a30580007dbc86dfa7f517826c46388a56d2619ce039a10

Observation 4963e30c-2f5e-4b1b-b54c-b9a0924506ed · outbound

This paper cites Griffin, Matthias Nießner, Federico Tombari, and Francis Engelmann.

Grounding by Remembering: Cross-Scene and In-Scene Memory for 3D Functional Affordances Griffin, Matthias Nießner, Federico Tombari, and Francis Engelmann

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T14:02:51.675861Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-13T01:55:47.136350Z digest=sha256:5300671568a062c76d5682ea43166bc48b3f3c54fb1787e82f6d6fe86575efc5

Observation 9c9513a4-2fbd-4247-adab-04d218143d3d · outbound

This paper cites SegGPT: Segmenting everything in context.

Grounding by Remembering: Cross-Scene and In-Scene Memory for 3D Functional Affordances SegGPT: Segmenting everything in context

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T14:02:51.677604Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-13T01:55:47.136350Z digest=sha256:e1699a759c5b02f2e1f6c8be2d592f505a0cd091d0b814fdab22cc9aa973b301

Observation 42663b47-fb45-47eb-ab3f-6709161e5892 · outbound

This paper cites Affordbot: 3d fine-grained embodied reasoning via multimodal large language models.

Grounding by Remembering: Cross-Scene and In-Scene Memory for 3D Functional Affordances Affordbot: 3d fine-grained embodied reasoning via multimodal large language models

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-13T01:57:05.030466Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-13T01:55:47.136350Z digest=sha256:d2ce17c0185e715adf8cf2f2c94d3e2c8c80c1e66a00fdba917b4c1117be009d

Observation 2f078beb-8b34-4aa5-bed3-6e430cf00567 · outbound

This paper cites 3d-gres: Generalized 3d referring expression segmentation.

Grounding by Remembering: Cross-Scene and In-Scene Memory for 3D Functional Affordances 3d-gres: Generalized 3d referring expression segmentation

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T14:02:51.681295Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-13T01:55:47.136350Z digest=sha256:bee98809bca5462bf526a5b87a08f4a1417a5429cc5865aa7fad74a25ea11bb7

Observation 03db3964-2480-4ab7-a19f-96fc92f3a3ca · outbound

This paper cites MVGGT: Multimodal visual geometry grounded transformer for multiview 3D referring expression segmentation.

Grounding by Remembering: Cross-Scene and In-Scene Memory for 3D Functional Affordances MVGGT: Multimodal visual geometry grounded transformer for multiview 3D referring expression segmentation

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-13T01:57:05.018368Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-13T01:55:47.136350Z digest=sha256:0bbdfaad3560c48ecab37a32008012f9df7b91b4fdfdc4dd4c6842f8f7a99935

Observation f3464c78-e69a-4a05-966e-7be46b951574 · outbound

This paper cites 11 It must be explicitly mentioned; use "None" if unavailable.

Grounding by Remembering: Cross-Scene and In-Scene Memory for 3D Functional Affordances 11 It must be explicitly mentioned; use "None" if unavailable

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T14:02:51.630493Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-13T01:55:47.136350Z digest=sha256:e2bd5b9ae4e9890a5502719b5f74db77b171b3e32daa8002d26abc7ede2447c8

Observation f33b5f2d-0bff-4beb-a665-dc376ea4e3bc · outbound

This paper cites an unresolved cited work.

Grounding by Remembering: Cross-Scene and In-Scene Memory for 3D Functional Affordances Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-05-13T14:02:51.619728Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-13T01:55:47.136350Z digest=sha256:fb7478acf05b5a2c5603d066ff07101f452cd2943bed65981e782b7c5701b006

Observation 9011e93a-1dea-4594-a815-a79b2ed01ab4 · outbound

This paper cites an unresolved cited work.

Grounding by Remembering: Cross-Scene and In-Scene Memory for 3D Functional Affordances Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-05-13T14:02:51.625370Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-13T01:55:47.136350Z digest=sha256:1a79afae5baf357fdc61b7236fab6d8f1afb9e87631201e0250f3f7d70ff444a

Observation 51e7eb97-8ab7-453c-83f7-38237fc72b3b · outbound

This paper cites an unresolved cited work.

Grounding by Remembering: Cross-Scene and In-Scene Memory for 3D Functional Affordances Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-05-13T14:02:51.617799Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-13T01:55:47.136350Z digest=sha256:d0726b0cfc38f87853369252a4f11dc9b9976efe9a3e3a3a1b36bf8c4fbd7169

Observation d340a52b-b1f9-4e34-8011-8262a75e097d · outbound

This paper cites {instruction}.

Grounding by Remembering: Cross-Scene and In-Scene Memory for 3D Functional Affordances {instruction}

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T14:02:51.610870Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-13T01:55:47.136350Z digest=sha256:68116520689b6f64077d98f686a5dde0363354999a130eabff209d501da7c606

Observation eb2f76ac-f657-4cf9-91f1-7d845098fa25 · outbound

This paper cites an unresolved cited work.

Grounding by Remembering: Cross-Scene and In-Scene Memory for 3D Functional Affordances Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-05-13T14:02:51.609105Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-13T01:55:47.136350Z digest=sha256:388ae4144399fd5bbef06890f917491e79bb8df367081ef3551b582291b96082

Observation 1ac60a6d-85b3-4794-b17f-05da9cabdf11 · outbound

This paper cites an unresolved cited work.

Grounding by Remembering: Cross-Scene and In-Scene Memory for 3D Functional Affordances Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-05-13T14:02:51.605038Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-13T01:55:47.136350Z digest=sha256:4f736a2ab2c2f42a33d1b8fa0f01fb63c79bf7687ceb263be700b393686f2b08

Observation f1e6ffd6-a29e-43e0-8522-a3e7fd524c2c · outbound

This paper cites we were unable to find the license for the dataset we used.

Grounding by Remembering: Cross-Scene and In-Scene Memory for 3D Functional Affordances we were unable to find the license for the dataset we used

Reference 31

Resolution
malformed identifier
raw_fallback, observed 2026-05-13T14:02:51.679397Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-13T01:55:47.136350Z digest=sha256:f448d2920810639343c83819e506ca26c834ec66bfa9a57ebad9e9093d22693b

Observation e698f2b6-414f-46f0-84c9-cc102d7fa6be · outbound

This paper cites Guidelines: • The answer [N/A] means that the paper does not involve crowdsourcing nor research with human subjects.

Grounding by Remembering: Cross-Scene and In-Scene Memory for 3D Functional Affordances Guidelines: • The answer [N/A] means that the paper does not involve crowdsourcing nor research with human subjects

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T14:02:51.607340Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-13T01:55:47.136350Z digest=sha256:747c826e862b685d0a476276957dc1774a9825d55933a1e655f56ac30356ce0c

Pith citing papers

No inbound Pith citation observations are available.