Pith. sign in

Paper Citation Record · LEDGER

Visual Programming: Compositional visual reasoning without training

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 13 inbound Pith citation observations for arXiv:2211.11559.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2211.11559 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 13 of 13 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 13 of 13 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T14:33:32.978926Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T19:56:10.525466Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 42e98bda-38d9-4163-a74f-459177b6fbbd · inbound

ViperGPT: Visual Inference via Python Execution for Reasoning cites this paper.

ViperGPT: Visual Inference via Python Execution for Reasoning Visual Programming: Compositional visual reasoning without training

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-17T18:15:14.565790Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-17T18:15:14.382011Z digest=sha256:3bcf2ac3e359fd8f35e815cb5fae24f62d395e9ba0aab4fd7a44434f6db430bc

Observation 8c83b403-160f-405f-9f48-0c91077ad0af · inbound

HuggingGPT: Solving AI Tasks with ChatGPT and its Friends in Hugging Face cites this paper.

HuggingGPT: Solving AI Tasks with ChatGPT and its Friends in Hugging Face Visual Programming: Compositional visual reasoning without training

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-14T00:06:45.617624Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-14T00:06:45.440071Z digest=sha256:24e8b1ced47f7aef386e5a5f7fc66301ef99783a30b0802fd306316bd18d27f2

Observation 1f21bfbc-ee16-4bac-8f20-251339993da7 · inbound

Visual Instruction Tuning cites this paper.

Visual Instruction Tuning Visual Programming: Compositional visual reasoning without training

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-11T08:22:03.757506Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-11T08:22:03.403362Z digest=sha256:332c8fb1c258914ed155099b597b0655bc511e0c64b93f8ae4efb611d6bcb4ec

Observation f0a27056-2b7e-4fe8-b71c-96721284a634 · inbound

LLaMA-Adapter V2: Parameter-Efficient Visual Instruction Model cites this paper.

LLaMA-Adapter V2: Parameter-Efficient Visual Instruction Model Visual Programming: Compositional visual reasoning without training

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-15T08:41:04.873384Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T08:41:04.743886Z digest=sha256:f2c76adca71b79c580423b417e3199f496984598d4c608651608070a14f401f0

Observation f8603a51-e249-4eaa-9a65-c1ee18503d9a · inbound

What to Say and When to Say it: Live Fitness Coaching as a Testbed for Situated Interaction cites this paper.

What to Say and When to Say it: Live Fitness Coaching as a Testbed for Situated Interaction Visual Programming: Compositional visual reasoning without training

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-23T22:53:33.145856Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-23T22:51:08.753650Z digest=sha256:625dc3d1a553a99b2e9716ea7c19d481810d785a410946c8ad6d0e6e43b26555

Observation 212acf97-e428-47db-ad61-4957bd06036c · inbound

Grounded Reinforcement Learning for Visual Reasoning cites this paper.

Grounded Reinforcement Learning for Visual Reasoning Visual Programming: Compositional visual reasoning without training

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-22T01:05:52.237592Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-22T01:05:18.801388Z digest=sha256:58cfe9b37a7c46449eb4aa05972544e5dfc36a6c5c1abc96a9480bdfb3525a20

Observation d3a5bc0b-4cf7-4701-9138-2d6a1f413096 · inbound

Augmented Vision-Language Models: A Systematic Review cites this paper.

Augmented Vision-Language Models: A Systematic Review Visual Programming: Compositional visual reasoning without training

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T14:33:32.978926Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:33:32.978926Z digest=sha256:b4ba13e6d27636212579580f0dd4697697529ebe300cf486172089fcc8ef3bbe

Observation acbf79e9-a453-445c-81e2-2cf1ccce4be1 · inbound

Multimodal Video Emotion Recognition with Reliable Reasoning Priors cites this paper.

Multimodal Video Emotion Recognition with Reliable Reasoning Priors Visual Programming: Compositional visual reasoning without training

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T12:18:30.875408Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:18:30.875408Z digest=sha256:0980e7ea02d7cc56806e2322087b9d083c0e3c7ff1a0d4c99af55f3b0be63863

Observation 843ab866-a909-45f7-825e-7ecfb8bfde60 · inbound

Reinforced Visual Perception with Tools cites this paper.

Reinforced Visual Perception with Tools Visual Programming: Compositional visual reasoning without training

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-05T12:27:04.935858Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T12:27:04.935858Z digest=sha256:db30aabbf4c3862588e14c031e1ab3281f858b033e1ff153bc6b1296365ef8ad

Observation 62e0ba91-b313-4084-988c-b66173ef2b45 · inbound

Visual Funnel: Resolving Contextual Blindness in Multimodal Large Language Models cites this paper.

Visual Funnel: Resolving Contextual Blindness in Multimodal Large Language Models Visual Programming: Compositional visual reasoning without training

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-16T23:48:41.916710Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-16T23:47:08.562575Z digest=sha256:25f9322afec20abd89420c701130d762d0f6aea5f5fb166afd79e573055629ff

Observation 3f6da490-24b7-4415-b73c-b8595729239d · inbound

MAG-3D: Multi-Agent Grounded Reasoning for 3D Understanding cites this paper.

MAG-3D: Multi-Agent Grounded Reasoning for 3D Understanding Visual Programming: Compositional visual reasoning without training

Reference 16

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T06:51:23.156390Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T17:25:31.097385Z digest=sha256:693a7d86ac3efead7265872d1b7f15be9ba06b1b1d424f37c6ba3d797d1dfd5f

Observation 498614d5-cf49-4cb6-97b1-05fcf713af9b · inbound

VESTA: Visual Exploration with Statistical Tool Agents cites this paper.

VESTA: Visual Exploration with Statistical Tool Agents Visual Programming: Compositional visual reasoning without training

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:56:10.526893Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-28T21:58:11.339217Z digest=sha256:5ded756c8e118474b431a51597474c6de8b6df866c62ba7fdffd65a1e54e9757

Observation 0359ec15-96f7-433d-a80b-280fa4e0a837 · inbound

CanvasAgent: Enabling Complex Image Creation and Editing via Visual Tool Orchestration cites this paper.

CanvasAgent: Enabling Complex Image Creation and Editing via Visual Tool Orchestration Visual Programming: Compositional visual reasoning without training

Reference 6

Resolution
unresolved
no resolver link, observed 2026-07-11T15:24:53.016199Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T15:24:53.016199Z digest=sha256:e7563794346099dd82b4e6e7bf36efaafae41bf7eaa0aa9dc7981d89b23446a8