Pith. sign in

Paper Citation Record · LEDGER

SayPlan: Grounding Large Language Models using 3D Scene Graphs for Scalable Robot Task Planning

As of 4 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 27 inbound Pith citation observations for arXiv:2307.06135.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2307.06135 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 27 of 27 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00

measured 27 of 27 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-04T04:53:32.653060Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-11T00:27:51.706200Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 1f0cbea0-5450-4561-9a44-1f10f2905b5d · inbound

LLM+P: Empowering Large Language Models with Optimal Planning Proficiency cites this paper.

LLM+P: Empowering Large Language Models with Optimal Planning Proficiency SayPlan: Grounding Large Language Models using 3D Scene Graphs for Scalable Robot Task Planning

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-05-14T18:36:18.527186Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T18:36:18.342052Z digest=sha256:3efa1ed0fdb574d0d370c15e9df44d36d3516250e13ae5fae5e4a19c6856260c

Observation 464401c7-3695-4562-99ac-6916fec9d32d · inbound

The Rise and Potential of Large Language Model Based Agents: A Survey cites this paper.

The Rise and Potential of Large Language Model Based Agents: A Survey SayPlan: Grounding Large Language Models using 3D Scene Graphs for Scalable Robot Task Planning

Reference 261

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:47:51.201250Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-11T10:47:44.152066Z digest=sha256:85859d5aa3d799f209d454a4f2dc600db86d0b253697795f098d0f434aab79b5

Observation 1b6baad1-4323-4e0d-8a14-f2aac69a4c17 · inbound

Towards Robust Surgical Automation via Digital Twin Representations from Foundation Models cites this paper.

Towards Robust Surgical Automation via Digital Twin Representations from Foundation Models SayPlan: Grounding Large Language Models using 3D Scene Graphs for Scalable Robot Task Planning

Reference 58

Resolution
verified exact
arxiv_id, observed 2026-05-23T20:23:24.908107Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-23T20:19:20.382009Z digest=sha256:3b2afec523cc4d026c7a98e02e08a972691f7d2ff11ef1e5547480ab32dfc635

Observation 7060d48a-054d-4266-bbb6-05ed2016c2fa · inbound

InternVLA-M1: A Spatially Guided Vision-Language-Action Framework for Generalist Robot Policy cites this paper.

InternVLA-M1: A Spatially Guided Vision-Language-Action Framework for Generalist Robot Policy SayPlan: Grounding Large Language Models using 3D Scene Graphs for Scalable Robot Task Planning

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-14T20:09:39.885407Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T20:09:39.677347Z digest=sha256:9ce9be7b55de9b7856d50c4dcbb9f3874c85ce3bd3e7f0aae3d0d4f33d5384af

Observation 89991576-c3ef-45da-a690-a09aa40e04de · inbound

BridgeEQA: Virtual Embodied Agents for Real Bridge Inspections cites this paper.

BridgeEQA: Virtual Embodied Agents for Real Bridge Inspections SayPlan: Grounding Large Language Models using 3D Scene Graphs for Scalable Robot Task Planning

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-05-17T21:42:07.520279Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-17T21:40:51.426489Z digest=sha256:b5500182e0f05983130b336ca9c70af9de494b10b9e0f5f9bc92c2add70a7462

Observation 7183b726-50d0-43b6-bdf9-c7758403bdf7 · inbound

VLM4VLA: Revisiting Vision-Language-Models in Vision-Language-Action Models cites this paper.

VLM4VLA: Revisiting Vision-Language-Models in Vision-Language-Action Models SayPlan: Grounding Large Language Models using 3D Scene Graphs for Scalable Robot Task Planning

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-03T12:30:33.851410Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:30:33.851410Z digest=sha256:44774b367579e78bb9b4fef4bd62f1cf2e157023c5a5f981ad0bd6b096f0ffa3

Observation cdb88895-7386-4f05-b932-c22efcb1f1ef · inbound

ST-BiBench: Benchmarking Multi-Stream Multimodal Coordination in Bimanual Embodied Tasks for MLLMs cites this paper.

ST-BiBench: Benchmarking Multi-Stream Multimodal Coordination in Bimanual Embodied Tasks for MLLMs SayPlan: Grounding Large Language Models using 3D Scene Graphs for Scalable Robot Task Planning

Reference 92

Resolution
verified exact
arxiv_id, observed 2026-05-16T06:00:40.551146Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-16T06:00:02.043029Z digest=sha256:6990410a386795e529e6ae2567308321aa099d34815c31ad7778890f80ed51c7

Observation 59f8d4d0-f2b5-4e8d-8775-fb221f3857ad · inbound

Relational Semantic Reasoning on 3D Scene Graphs for Open World Interactive Object Search cites this paper.

Relational Semantic Reasoning on 3D Scene Graphs for Open World Interactive Object Search SayPlan: Grounding Large Language Models using 3D Scene Graphs for Scalable Robot Task Planning

Reference 25

Resolution
unresolved
no resolver link, observed 2026-07-15T14:22:08.027014Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-15T14:22:08.027014Z digest=sha256:103c2bd1c67c6372dc432f7f104afe31f339d9283f9bf20be001d70e3d4ebf96

Observation d95ca078-d4c7-4d08-b803-8202abbc102d · inbound

KITE: Keyframe-Indexed Tokenized Evidence for VLM-Based Robot Failure Analysis cites this paper.

KITE: Keyframe-Indexed Tokenized Evidence for VLM-Based Robot Failure Analysis SayPlan: Grounding Large Language Models using 3D Scene Graphs for Scalable Robot Task Planning

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-10T23:55:51.408336Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T18:47:25.412497Z digest=sha256:e3cadbb2bda4c5b46024090674f6ec84b946cc1d5296a1a32e0feae78356e3ba

Observation 29292972-e6f0-478c-a6e7-67beb1167c28 · inbound

Long-Horizon Manipulation via Trace-Conditioned VLA Planning cites this paper.

Long-Horizon Manipulation via Trace-Conditioned VLA Planning SayPlan: Grounding Large Language Models using 3D Scene Graphs for Scalable Robot Task Planning

Reference 54

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T14:41:38.912572Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-09T21:10:03.554504Z digest=sha256:9973aa654c5ab11d03161b2e85edb629e3485fcd522693c0015b6079b530f963

Observation c7bad49e-a12e-4f48-a022-c9aebabb12f8 · inbound

When Robots Do the Chores: A Benchmark and Agent for Long-Horizon Household Task Execution cites this paper.

When Robots Do the Chores: A Benchmark and Agent for Long-Horizon Household Task Execution SayPlan: Grounding Large Language Models using 3D Scene Graphs for Scalable Robot Task Planning

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-15T01:59:38.797026Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-15T01:59:03.623345Z digest=sha256:5e48926e95050b5ba5af985a29725484fadc4964d5ffa6dd0f8678b7ffe70e28

Observation 89dc4626-a02e-418e-aa0a-afa44978b6d3 · inbound

When Robots Do the Chores: A Benchmark and Agent for Long-Horizon Household Task Execution cites this paper.

When Robots Do the Chores: A Benchmark and Agent for Long-Horizon Household Task Execution SayPlan: Grounding Large Language Models using 3D Scene Graphs for Scalable Robot Task Planning

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-20T21:29:03.260859Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-20T21:25:34.754045Z digest=sha256:872b75b10d08cd12a5bb27e5e88a8ffc589db4a0b46762ba91f1a90917a363de

Observation 090d9f00-0746-45ee-a5f0-15f9a2207ca0 · inbound

TaskGround: Structured Executable Task Inference for Full-Scene Household Reasoning cites this paper.

TaskGround: Structured Executable Task Inference for Full-Scene Household Reasoning SayPlan: Grounding Large Language Models using 3D Scene Graphs for Scalable Robot Task Planning

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-20T11:03:13.964857Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-20T10:53:44.014259Z digest=sha256:09f4df105db430cc810b3401e8368be6e9dd1649a51de4559b31fa013127a00d

Observation 887ac83c-40d1-4a87-ab18-3fde5e506a30 · inbound

Fixed External Cameras as Common Prior Maps for Active 3D Scene Graph Generation cites this paper.

Fixed External Cameras as Common Prior Maps for Active 3D Scene Graph Generation SayPlan: Grounding Large Language Models using 3D Scene Graphs for Scalable Robot Task Planning

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-20T09:58:10.975206Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-20T09:57:22.716399Z digest=sha256:1c7bc9e8afcdbe3b67e5c53847ba792962d2dbfadd9d431ebe3bd84fa17b3747

Observation cf6780ab-221f-45f3-b515-fd0efc268bd9 · inbound

RGB-only Active 3D Scene Graph Generation for Indoor Mobile Robots cites this paper.

RGB-only Active 3D Scene Graph Generation for Indoor Mobile Robots SayPlan: Grounding Large Language Models using 3D Scene Graphs for Scalable Robot Task Planning

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-20T09:53:15.852931Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-20T09:53:01.437258Z digest=sha256:895d721f448b4a33cc303e4d5d313ace259eca50675e84ed5980b884f4294a5f

Observation 54efcc15-1a77-4860-9105-58df4d41336b · inbound

Beyond Waypoints: Dual-Heatmap Grounding for Cross-Embodiment Semantic Navigation cites this paper.

Beyond Waypoints: Dual-Heatmap Grounding for Cross-Embodiment Semantic Navigation SayPlan: Grounding Large Language Models using 3D Scene Graphs for Scalable Robot Task Planning

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-20T05:58:05.375427Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-20T05:53:31.627150Z digest=sha256:496eeda575a8df87dccb37c82a0b9b1c6d6bea1489020a45e7125217c18cc60d

Observation 40858956-f687-4756-b479-87760731ff17 · inbound

VEOcc: Voxel-Centric Online Semantic Occupancy Prediction For Embodied Scene Understanding cites this paper.

VEOcc: Voxel-Centric Online Semantic Occupancy Prediction For Embodied Scene Understanding SayPlan: Grounding Large Language Models using 3D Scene Graphs for Scalable Robot Task Planning

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-06-30T12:34:39.623503Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-30T12:06:52.443736Z digest=sha256:86a271240a9982307f115d314f62cfab555826823f6c6253675c3f075dc4466c

Observation 9d986ea5-3112-4d4b-a061-b36e7f4167e2 · inbound

The Sword, Shield, and Achilles' Heel: Characterizing the Linguistic Inductive Bias of Large Language Models for Spatial Reasoning in Navigation Planning cites this paper.

The Sword, Shield, and Achilles' Heel: Characterizing the Linguistic Inductive Bias of Large Language Models for Spatial Reasoning in Navigation Planning SayPlan: Grounding Large Language Models using 3D Scene Graphs for Scalable Robot Task Planning

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:36:08.359447Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T22:23:24.536258Z digest=sha256:5ab9cb93d239bcc3e69bdf83435fa79d56540e337699d3059fcbdcff928bde4d

Observation 9442d808-1f24-4934-886f-86016d9064e3 · inbound

ANCHOR: Agentic Noise Creation Framework for Human Simulation and Denoising Recommendation cites this paper.

ANCHOR: Agentic Noise Creation Framework for Human Simulation and Denoising Recommendation SayPlan: Grounding Large Language Models using 3D Scene Graphs for Scalable Robot Task Planning

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-07-02T15:27:04.858486Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-27T23:54:14.214208Z digest=sha256:609e6f504280b26ca6a82e6772fdbd4f98bd2c38b8035e263f7e269d1cbf17d2

Observation 8310fa09-88b6-440e-990c-70e7f4f4ff58 · inbound

ANCHOR: Agentic Noise Creation Framework for Human Simulation and Denoising Recommendation cites this paper.

ANCHOR: Agentic Noise Creation Framework for Human Simulation and Denoising Recommendation SayPlan: Grounding Large Language Models using 3D Scene Graphs for Scalable Robot Task Planning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-04T04:53:32.653060Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T04:53:32.653060Z digest=sha256:64f9255383e126153922be0524f46896c7eb2bed2978241a932ea53b87a483d0

Observation d44b5d46-b785-4e94-bfe7-405d1aacb510 · inbound

AlloSpatial: Agentic Harness Framework for Spatial Reasoning in Foundation Models cites this paper.

AlloSpatial: Agentic Harness Framework for Spatial Reasoning in Foundation Models SayPlan: Grounding Large Language Models using 3D Scene Graphs for Scalable Robot Task Planning

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-07-03T00:27:30.262195Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-27T17:08:22.609095Z digest=sha256:204b17d64bce9f4694318f85f6a255588b8c80ac1ba5cfe24a7352fc4f3fe54f

Observation 3da9c0b3-edce-430c-8fb3-878f74744026 · inbound

ASCII Art Turns LLMs into VLA Controllers cites this paper.

ASCII Art Turns LLMs into VLA Controllers SayPlan: Grounding Large Language Models using 3D Scene Graphs for Scalable Robot Task Planning

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-07-04T06:19:38.773223Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-26T14:32:30.091437Z digest=sha256:fcceb54cfbfde97240443ec377c950141a26c99e26beabb73f02cb8bd112a839

Observation f6124dd5-3be3-486f-9d8a-e72b718fcf0f · inbound

Plan Right, Then Plan Tight: Symbolic RL for Efficient Embodied Reasoning cites this paper.

Plan Right, Then Plan Tight: Symbolic RL for Efficient Embodied Reasoning SayPlan: Grounding Large Language Models using 3D Scene Graphs for Scalable Robot Task Planning

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T10:35:42.485449Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-01T05:20:28.149651Z digest=sha256:a3dcb740f5d8ab24e50a44542ecd50433c98bbdaaa7ea957ec38efd9f244a2ee

Observation 59caeb1f-a6a7-49f4-83dd-6722e517e5b3 · inbound

Hypothesis-driven Model Expansion under Uncertainty for Open-World Robot Planning cites this paper.

Hypothesis-driven Model Expansion under Uncertainty for Open-World Robot Planning SayPlan: Grounding Large Language Models using 3D Scene Graphs for Scalable Robot Task Planning

Reference 49

Resolution
verified exact
local_arxiv, observed 2026-07-11T00:27:51.729634Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-11T00:18:08.485089Z digest=sha256:57ff862ef772530955633544194b7bd7107a4fddacecd3d1f2d2a1a2af05a5d5

Observation e786ef67-e19e-4240-af3e-c6d0c3053370 · inbound

CAIRN: Cross-Room 3D Scene Understanding with Topology-Aware Large Multimodal Models cites this paper.

CAIRN: Cross-Room 3D Scene Understanding with Topology-Aware Large Multimodal Models SayPlan: Grounding Large Language Models using 3D Scene Graphs for Scalable Robot Task Planning

Reference 45

Resolution
verified exact
local_arxiv, observed 2026-07-08T02:44:27.736477Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-07-08T02:44:00.608590Z digest=sha256:66250a3e647ebac9dc6413aa94d1414e3ebbade314c06ca8f815e3a7b6c76a01

Observation 98e9f947-7003-48a4-a3df-5864d3af193b · inbound

CAIRN: Cross-Room 3D Scene Understanding with Topology-Aware Large Multimodal Models cites this paper.

CAIRN: Cross-Room 3D Scene Understanding with Topology-Aware Large Multimodal Models SayPlan: Grounding Large Language Models using 3D Scene Graphs for Scalable Robot Task Planning

Reference 45

Resolution
unresolved
no resolver link, observed 2026-07-14T16:00:13.133298Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T16:00:13.133298Z digest=sha256:3a118eeecb2c7886638bf8853b6090adf0d49e08d9d38c6a0ce14a60154a6b1d

Observation 3969548e-c33a-478b-bcde-b60800432059 · inbound

LEACL: LLM-Enhanced Automatic Curriculum Learning for Reinforcement Learning in Long-Horizon Manipulation Tasks cites this paper.

LEACL: LLM-Enhanced Automatic Curriculum Learning for Reinforcement Learning in Long-Horizon Manipulation Tasks SayPlan: Grounding Large Language Models using 3D Scene Graphs for Scalable Robot Task Planning

Reference 16

Resolution
unresolved
no resolver link, observed 2026-07-30T20:27:56.042565Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T20:27:56.042565Z digest=sha256:249d9d2ad85a3f448d7ccb3946f41130b19f3dee0ccbeb8fcfd75ce4189e1c86