Pith. sign in

Paper Citation Record · LEDGER

Multimodal Procedural Planning via Dual Text-Image Prompting

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2305.01795.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2305.01795 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 5 of 5 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:12:01.336852Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-29T21:23:58.797233Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 938f6f8e-35d5-4c83-b31a-74051c30a35f · inbound

VoxPoser: Composable 3D Value Maps for Robotic Manipulation with Language Models cites this paper.

VoxPoser: Composable 3D Value Maps for Robotic Manipulation with Language Models Multimodal Procedural Planning via Dual Text-Image Prompting

Reference 71

Resolution
verified exact
arxiv_id, observed 2026-05-13T08:57:22.550628Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T08:57:22.299028Z digest=sha256:01c6b616702dd0ff48cf243ca91b5826fab5e9d4ea3f916623668e978b28c4a2

Observation b3930db6-c874-4184-86c9-27cd59e1ede5 · inbound

Large Language Models for Planning: A Comprehensive and Systematic Survey cites this paper.

Large Language Models for Planning: A Comprehensive and Systematic Survey Multimodal Procedural Planning via Dual Text-Image Prompting

Reference 159

Resolution
unresolved
no resolver link, observed 2026-08-07T14:12:01.336852Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:12:01.336852Z digest=sha256:f93d829723a5251c6dc29a9376687abcb9c1fd9c58b7ed651ae5a36365105b6e

Observation 95dc4e28-f74a-4ed7-a129-993b44200739 · inbound

Making VLMs More Robot-Friendly: Self-Critical Distillation of Low-Level Procedural Reasoning cites this paper.

Making VLMs More Robot-Friendly: Self-Critical Distillation of Low-Level Procedural Reasoning Multimodal Procedural Planning via Dual Text-Image Prompting

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T18:30:47.177417Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:30:47.177417Z digest=sha256:e8a1b569d34eda2a07c64e7f05c40b985dbd601a7b5d77a47b1010ccd28462fb

Observation 381d6a6b-ef6f-4752-9839-5cb1e98f4cf1 · inbound

LLaPa: A Vision-Language Model Framework for Counterfactual-Aware Procedural Planning cites this paper.

LLaPa: A Vision-Language Model Framework for Counterfactual-Aware Procedural Planning Multimodal Procedural Planning via Dual Text-Image Prompting

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T18:21:47.730503Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:21:47.730503Z digest=sha256:ccd39b3b84acaec4625081bd1b497197c83c1e031307c1e9741f9520fd897c29

Observation e9336091-a2b8-44e7-bd69-fbcc9c90dda9 · inbound

RePlan-Bot: Multi-Level Replanning for Embodied Instruction Following cites this paper.

RePlan-Bot: Multi-Level Replanning for Embodied Instruction Following Multimodal Procedural Planning via Dual Text-Image Prompting

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-06-29T21:23:58.798746Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-29T21:21:30.990215Z digest=sha256:be4772d798112f7b7b2f3011363e7317430cb56d552e0f827913e5894c389533