Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T04:22:04.066726Z
Paper Citation Record · LEDGER
As of 14 August 2026, this Paper Citation Record lists 55 of 55 outbound references and 1 inbound Pith citation observation for arXiv:2412.01550.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T04:22:04.066726Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-11T22:49:08.968786Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-11T22:49:09.109028Z
55 of 55 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation ab8e3d1b-a633-4657-af7b-cea2f5bb4f09 · outbound
SeqAfford: Sequential 3D Affordance Reasoning via Multimodal Large Language Model GPT-4 Technical Report
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a937ec8b-c65a-4d26-add7-504d835e4de8 · outbound
SeqAfford: Sequential 3D Affordance Reasoning via Multimodal Large Language Model Flamingo: a visual language model for few-shot learning
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fb8853da-f817-41b1-9c71-204d10ba00fc · outbound
SeqAfford: Sequential 3D Affordance Reasoning via Multimodal Large Language Model 3D-AffordanceLLM: Harnessing Large Language Models for Open-Vocabulary Affordance Detection in 3D Worlds
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ba620139-edf7-4a7e-ab42-c056edbf3db2 · outbound
SeqAfford: Sequential 3D Affordance Reasoning via Multimodal Large Language Model 3D-TAFS: A Training-free Framework for 3D Affordance Segmentation
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation faa3f16d-433c-414b-b3fc-4b4c039def66 · outbound
SeqAfford: Sequential 3D Affordance Reasoning via Multimodal Large Language Model Flashattention: Fast and memory-efficient exact attention with io-awareness
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 9659b92e-b005-43ea-a0cb-c595f645da6f · outbound
SeqAfford: Sequential 3D Affordance Reasoning via Multimodal Large Language Model Scene- fun3d: fine-grained functionality and affordance understand- ing in 3d scenes
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 4ab097ee-b741-4d9e-aafd-71af3453c2cd · outbound
SeqAfford: Sequential 3D Affordance Reasoning via Multimodal Large Language Model 3d affordancenet: A benchmark for visual object affordance understanding
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation ac39b2a5-e6fb-4928-bf02-2139eb936263 · outbound
SeqAfford: Sequential 3D Affordance Reasoning via Multimodal Large Language Model Learning 2d invariant affordance knowledge for 3d affordance ground- ing
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cf87460f-5780-472e-ac89-754453a834ab · outbound
SeqAfford: Sequential 3D Affordance Reasoning via Multimodal Large Language Model 3d-llm: Injecting the 3d world into large language models
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9797d96f-474d-4909-97fc-7b433f886771 · outbound
SeqAfford: Sequential 3D Affordance Reasoning via Multimodal Large Language Model LoRA: Low-Rank Adaptation of Large Language Models
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 444fe85c-5a16-4e63-8f04-37f80e7cd82b · outbound
SeqAfford: Sequential 3D Affordance Reasoning via Multimodal Large Language Model Robo-abc: Affordance generalization beyond categories via semantic correspondence for robot ma- nipulation
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 69943d3b-05b3-4f17-ba23-39d1e283c041 · outbound
SeqAfford: Sequential 3D Affordance Reasoning via Multimodal Large Language Model Bert: Pre-training of deep bidirectional transform- ers for language understanding
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation eeefe02c-7f12-4951-83f5-b56c7c5e341c · outbound
SeqAfford: Sequential 3D Affordance Reasoning via Multimodal Large Language Model Lisa: Reasoning segmentation via large language model
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation adbd5ecc-11f6-4f97-90b6-493d3d0a3e6d · outbound
SeqAfford: Sequential 3D Affordance Reasoning via Multimodal Large Language Model One-shot open affordance learning with foundation models
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation eeff70d7-cba7-4186-9b64-adc2f9042cdd · outbound
SeqAfford: Sequential 3D Affordance Reasoning via Multimodal Large Language Model Referring transformer: A one-step approach to multi-task visual grounding
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8bf1e2e4-155e-41fa-8f70-c34c6940b220 · outbound
SeqAfford: Sequential 3D Affordance Reasoning via Multimodal Large Language Model Laso: Language-guided affordance seg- mentation on 3d object
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 6bf88876-59e8-4037-9d2d-f0dde4f82d93 · outbound
SeqAfford: Sequential 3D Affordance Reasoning via Multimodal Large Language Model Gres: Gen- eralized referring expression segmentation
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 49669313-8f62-4f44-adcf-60bf0c46eba9 · outbound
SeqAfford: Sequential 3D Affordance Reasoning via Multimodal Large Language Model Visual instruction tuning
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 646cfba1-47ec-4448-a8de-f6e128e61dcf · outbound
SeqAfford: Sequential 3D Affordance Reasoning via Multimodal Large Language Model Openshape: Scaling up 3d shape representation towards open-world understanding
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation bae82463-d3b1-490d-90aa-0b6d388c2c04 · outbound
SeqAfford: Sequential 3D Affordance Reasoning via Multimodal Large Language Model RoBERTa: A Robustly Optimized BERT Pretraining Approach
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8a180739-6da4-4d7b-ba22-f36db095805f · outbound
SeqAfford: Sequential 3D Affordance Reasoning via Multimodal Large Language Model InternGPT: Solving Vision-Centric Tasks by Interacting with ChatGPT Beyond Language
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4b540fe5-7ffd-4e99-998e-1a25ac8b0bdf · outbound
SeqAfford: Sequential 3D Affordance Reasoning via Multimodal Large Language Model Auc: a misleading measure of the performance of predictive distribution models
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 4c78e5e9-fee6-4399-93c8-2f7dce28c78f · outbound
SeqAfford: Sequential 3D Affordance Reasoning via Multimodal Large Language Model Decoupled Weight Decay Regularization
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bf8cfec3-38cd-4bf5-b2c9-5db1fb9ca42d · outbound
SeqAfford: Sequential 3D Affordance Reasoning via Multimodal Large Language Model GEAL: Generalizable 3D Affordance Learning with Cross-Modal Consistency
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation deef5323-929c-4b82-b5b3-8bd734e117ce · outbound
SeqAfford: Sequential 3D Affordance Reasoning via Multimodal Large Language Model 3d-sps: Single-stage 3d visual grounding via referred point progressive selection
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 8ffaf169-f5f0-4703-abc9-638c71c4321b · outbound
SeqAfford: Sequential 3D Affordance Reasoning via Multimodal Large Language Model Partnet: A large- scale benchmark for fine-grained and hierarchical part-level 3d object understanding
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 5a12ca22-527f-420c-9ecb-d9d7196c618a · outbound
SeqAfford: Sequential 3D Affordance Reasoning via Multimodal Large Language Model O2o-afford: Annotation-free large-scale object- object affordance learning
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation d5376ed9-0cf7-4a83-a290-622e4646770e · outbound
SeqAfford: Sequential 3D Affordance Reasoning via Multimodal Large Language Model RT-Affordance: Affordances are Versatile Intermediate Representations for Robot Manipulation
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 50671ee5-14f9-40c1-a36a-229c1c433451 · outbound
SeqAfford: Sequential 3D Affordance Reasoning via Multimodal Large Language Model Open-vocabulary affordance detection in 3d point clouds
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation cfad631b-d865-42f1-84f1-e1ae302718c2 · outbound
SeqAfford: Sequential 3D Affordance Reasoning via Multimodal Large Language Model Where2explore: Few-shot affordance learning for unseen novel categories of articulated objects
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 68a1afd0-09ae-4652-8a1f-3f49512cda7d · outbound
SeqAfford: Sequential 3D Affordance Reasoning via Multimodal Large Language Model Training language models to follow instructions with human feedback.Advances in neural information processing systems, 35:27730–27744,
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 34129aa7-d87b-478f-b684-de823d1b2d94 · outbound
SeqAfford: Sequential 3D Affordance Reasoning via Multimodal Large Language Model Pointnet++: Deep hierarchical feature learning on point sets in a metric space
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 68c35e8d-34d8-43b9-a628-0deea8da1a67 · outbound
SeqAfford: Sequential 3D Affordance Reasoning via Multimodal Large Language Model Contrast with reconstruct: Contrastive 3d representation learning guided by generative pretraining
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 04be1169-c03b-451b-b780-fdaffc1f3d72 · outbound
SeqAfford: Sequential 3D Affordance Reasoning via Multimodal Large Language Model ShapeLLM: Universal 3D Object Understanding for Embodied Interaction
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ff3c5759-9ed8-4cc2-a132-7188a2bd74c3 · outbound
SeqAfford: Sequential 3D Affordance Reasoning via Multimodal Large Language Model Affordancellm: Grounding affordance from vision language models
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 94cb39af-65f2-439f-9e8f-587d903f6120 · outbound
SeqAfford: Sequential 3D Affordance Reasoning via Multimodal Large Language Model Learning transferable visual models from natural language supervi- sion
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 74d02153-9498-411f-a560-79a114b7f42d · outbound
SeqAfford: Sequential 3D Affordance Reasoning via Multimodal Large Language Model Optimizing intersection- over-union in deep neural networks for image segmentation
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 47ba5652-44dd-42a8-a062-adf68c30b165 · outbound
SeqAfford: Sequential 3D Affordance Reasoning via Multimodal Large Language Model GREAT: Geometry-Intention Collaborative Inference for Open-Vocabulary 3D Object Affordance Grounding
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 21476325-1ecd-4d30-85a7-f6deaa41a147 · outbound
SeqAfford: Sequential 3D Affordance Reasoning via Multimodal Large Language Model Color indexing
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation a00a30f4-ee81-498e-8a6d-585f8fc04bbb · outbound
SeqAfford: Sequential 3D Affordance Reasoning via Multimodal Large Language Model LLaMA: Open and Efficient Foundation Language Models
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 790adcfb-82c4-43d9-9b4e-9ea0058dcc99 · outbound
SeqAfford: Sequential 3D Affordance Reasoning via Multimodal Large Language Model Visionllm: Large language model is also an open- ended decoder for vision-centric tasks
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d8947f22-a89c-46ae-a9a5-02726637a8f5 · outbound
SeqAfford: Sequential 3D Affordance Reasoning via Multimodal Large Language Model Dynamic graph cnn for learning on point clouds
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 35681b06-9f66-45a8-b805-4fb98dd501fe · outbound
SeqAfford: Sequential 3D Affordance Reasoning via Multimodal Large Language Model Advantages of the mean absolute error (mae) over the root mean square error (rmse) in assessing average model performance
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation c05242bf-45db-48ad-9743-6e7cc79538d2 · outbound
SeqAfford: Sequential 3D Affordance Reasoning via Multimodal Large Language Model Learning environment-aware affor- dance for 3d articulated object manipulation under occlu- sions
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 67b0b7ab-fe0e-47c4-8985-11ff6c8f4cb8 · outbound
SeqAfford: Sequential 3D Affordance Reasoning via Multimodal Large Language Model AffordDP: Generalizable Diffusion Policy with Transferable Affordance
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a00cb060-c011-4fa1-826a-1702af9ba9bc · outbound
SeqAfford: Sequential 3D Affordance Reasoning via Multimodal Large Language Model PartAfford: Part-level Affordance Discovery from 3D Objects
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f78edc63-97e3-40f5-8742-2c85b156b9a5 · outbound
SeqAfford: Sequential 3D Affordance Reasoning via Multimodal Large Language Model Weakly-supervised affordance grounding guided by part-level semantic priors
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 41a64d97-2b2f-4ee0-a5ce-84dfd9281f06 · outbound
SeqAfford: Sequential 3D Affordance Reasoning via Multimodal Large Language Model PointLLM: Empowering Large Language Models to Understand Point Clouds
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 23d12af1-3457-46e9-84a6-21082957c3d2 · outbound
SeqAfford: Sequential 3D Affordance Reasoning via Multimodal Large Language Model Ulip: Learning a unified representation of language, images, and point clouds for 3d understanding
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 44b14df7-2f30-471e-bc22-79e192377a99 · outbound
SeqAfford: Sequential 3D Affordance Reasoning via Multimodal Large Language Model Grounding 3d object affordance from 2d interactions in images
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 46b23a12-9091-4d5a-b482-5f9e51f2648a · outbound
SeqAfford: Sequential 3D Affordance Reasoning via Multimodal Large Language Model Lemon: Learning 3d human-object interac- tion relation from 2d images
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation f69d034d-0738-4355-993f-2de9a35a7dec · outbound
SeqAfford: Sequential 3D Affordance Reasoning via Multimodal Large Language Model Ferret: Refer and Ground Anything Anywhere at Any Granularity
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 63edb73d-6da7-4125-9340-83cc90a7a3c0 · outbound
SeqAfford: Sequential 3D Affordance Reasoning via Multimodal Large Language Model Uni3d: A unified baseline for multi-dataset 3d object detection
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation a846904f-5e78-4b50-b971-f9ef41629ef4 · outbound
SeqAfford: Sequential 3D Affordance Reasoning via Multimodal Large Language Model GPT4RoI: Instruction Tuning Large Language Model on Region-of-Interest
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cb685245-b92f-4eb9-a991-16602a54c21e · outbound
SeqAfford: Sequential 3D Affordance Reasoning via Multimodal Large Language Model MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 85b2a812-ccc1-438d-bc66-f05c37773df8 · inbound
AffordDP: Generalizable Diffusion Policy with Transferable Affordance SeqAfford: Sequential 3D Affordance Reasoning via Multimodal Large Language Model
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.