Pith. sign in

Paper Citation Record · LEDGER

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning

As of 12 August 2026, this Paper Citation Record lists 100 of 129 outbound references and 0 inbound Pith citation observations for arXiv:2608.09682.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.09682 v1

Coverage vector

measured 100 of 129 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T12:53:14.503030Z

measured 100 of 100 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

100 of 129 outbound references displayed

  • verified exact3
  • verified fuzzy0
  • unresolved97
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 21d126e9-bc08-49be-ac7a-28c7872f3451 · outbound

This paper cites Latent reasoning with supervised thinking states.arXiv preprint arXiv:2602.08332, 2026.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Latent reasoning with supervised thinking states.arXiv preprint arXiv:2602.08332, 2026

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:13.922601Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:13.922601Z digest=sha256:36481cb086c8a8e7039d8568cce085601dc915fbbb3b184cfe9c5dfec8965a20

Observation 3b40b537-9d4c-4dae-87c7-72e0cbaead1d · outbound

This paper cites Acloserlookatbiasandchain-of-thought faithfulness of large (vision) language models.arXiv preprint arXiv:2505.23945, 2025.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Acloserlookatbiasandchain-of-thought faithfulness of large (vision) language models.arXiv preprint arXiv:2505.23945, 2025

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:13.930040Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:13.930040Z digest=sha256:ba3868b1745f2bfbdaabcba9cd3393d4c242e936b80ba8c50d52530b143f6c29

Observation 54c530f6-4381-45b3-8282-d55da1520d61 · outbound

This paper cites An Image is Worth 1/2 Tokens After Layer 2: Plug-and-Play Inference Acceleration for Large Vision-Language Models.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning An Image is Worth 1/2 Tokens After Layer 2: Plug-and-Play Inference Acceleration for Large Vision-Language Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:13.935715Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:13.935715Z digest=sha256:fe90523eae6534add29387a9b23bbcf57451939a6d598e3b75960a1bdf5a1b7d

Observation 4127ec8f-cb29-49ef-bf11-6a0e20ae4c03 · outbound

This paper cites MMStar: An evaluator-centric benchmark for multi-modal large language models, 2024.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning MMStar: An evaluator-centric benchmark for multi-modal large language models, 2024

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:13.941754Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:13.941754Z digest=sha256:c5b016521f57f3b2ed356299ebc7776465bf4c20a8a1f847cc6201f62cd55d8b

Observation e3a06ba7-1a84-4d28-9062-55bd6db2c470 · outbound

This paper cites Perception before reasoning: Two-stage reinforcement learning for visual reasoning.arXiv preprint arXiv:2509.13031, 2025.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Perception before reasoning: Two-stage reinforcement learning for visual reasoning.arXiv preprint arXiv:2509.13031, 2025

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:13.947142Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:13.947142Z digest=sha256:83d1b686def0c8b2651dcc997d08e7776cc7e4b687a4a42e39c1d00025051eb5

Observation c1d6048d-a793-41d2-8a99-295b545a062e · outbound

This paper cites Reasoning Models Don't Always Say What They Think.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Reasoning Models Don't Always Say What They Think

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:13.952698Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:13.952698Z digest=sha256:d271f6daef64fd34a2f4eea737b9bbf2a5ab8a9f829df1e5284717cbf6962a31

Observation 527f98d2-c979-4acc-b913-fbf7eec197af · outbound

This paper cites v1: Learning to Point Visual Tokens for Multimodal Grounded Reasoning.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning v1: Learning to Point Visual Tokens for Multimodal Grounded Reasoning

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:13.959442Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:13.959442Z digest=sha256:ac74b6e18b0b394f903850cafa1b5bb87b3c57f2a531f6a0dadd09e4df9086f7

Observation 4b2e1502-e09d-4db2-af0d-9f9e3250c7e9 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:13.964796Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:13.964796Z digest=sha256:c2e78f78f408a28b58ea68973ad9e352cfc1b608176352194b824b0d07726b13

Observation 509a382a-9523-41ba-98f9-4c79aa25a0ef · outbound

This paper cites Molmo and PixMo: Open Weights and Open Data for State-of-the-Art Vision-Language Models.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Molmo and PixMo: Open Weights and Open Data for State-of-the-Art Vision-Language Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:13.970205Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:13.970205Z digest=sha256:77186af30f86d5066dde1148a3c4b7f72a0c94567d4dfda801e33fb4d69660d5

Observation 93107529-ff1e-4d89-8ff6-a0d9f5579f93 · outbound

This paper cites Virgo: A Preliminary Exploration on Reproducing o1-like MLLM.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Virgo: A Preliminary Exploration on Reproducing o1-like MLLM

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:13.978385Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:13.978385Z digest=sha256:8c6565acfc81108e701c8ed0b7ceb2cc4362ab45d7be80bb7628b00e905feb72

Observation 65d3ed44-1f15-4843-b40c-8316c5ac469c · outbound

This paper cites Revisiting the necessity of lengthy chain-of-thought in vision-centric reasoning.arXiv preprint arXiv:2511.22586, 2025.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Revisiting the necessity of lengthy chain-of-thought in vision-centric reasoning.arXiv preprint arXiv:2511.22586, 2025

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:13.984603Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:13.984603Z digest=sha256:414a613e27c1ff74178fc7695096708c08fa7560f0fef7627999291f7bab2fa1

Observation be44b9db-1f68-4022-8868-c22a5ce9043b · outbound

This paper cites VLMEvalKit: An open-source toolkit for evaluating large multi-modality models.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning VLMEvalKit: An open-source toolkit for evaluating large multi-modality models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:13.990131Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:13.990131Z digest=sha256:072dba2993488ab3aced7c8cb5e37eae550209fa25e6d72bea9106fc783848af

Observation 7931aa8e-29c5-4a9f-9949-92b81a812407 · outbound

This paper cites GRIT: Teaching MLLMs to Think with Images.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning GRIT: Teaching MLLMs to Think with Images

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:13.996266Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:13.996266Z digest=sha256:fbbc115cdac5239e600e0a0df54b5be1568ce9530a09e7781f275c8201bb6e34

Observation 9056dd3f-3141-49d4-8937-2bf9ac597373 · outbound

This paper cites Reward Shaping to Mitigate Reward Hacking in RLHF.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Reward Shaping to Mitigate Reward Hacking in RLHF

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.002043Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.002043Z digest=sha256:d3f47017c9541a1c6bf41f4ddfca44da0d57f96958f0c60928d744e9306172aa

Observation 195d31a5-11c3-492e-af18-b14ebf648ec6 · outbound

This paper cites BLINK: Multimodal Large Language Models Can See but Not Perceive.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning BLINK: Multimodal Large Language Models Can See but Not Perceive

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.008737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.008737Z digest=sha256:6cce266a04bd167ac086e88e25f6e052761e107939e37264413cea9fc85bbad3

Observation 64d7c8d5-a9f5-446f-ab58-b427e8915449 · outbound

This paper cites Thinking with deltas: Incen- tivizingreinforcementlearningviadifferentialvisualreasoningpolicy.arXivpreprintarXiv:2601.06801, 2026.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Thinking with deltas: Incen- tivizingreinforcementlearningviadifferentialvisualreasoningpolicy.arXivpreprintarXiv:2601.06801, 2026

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.016201Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.016201Z digest=sha256:936eb82dc3bdefc8d708c447050e9abf6709fb3d4bf82d03865e4e3b6440e019

Observation 068e1c91-7dab-46d4-be3a-3a5502407ccd · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Gemini: A Family of Highly Capable Multimodal Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.021564Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.021564Z digest=sha256:c98ec5ae94b7d012ec0da0e680bb79ea6c17af8dcc889abeecf2dacad20e3b4b

Observation a332fe5b-83bd-4f84-a92e-1559c14595f7 · outbound

This paper cites GLM-5V-Turbo: Toward a Native Foundation Model for Multimodal Agents.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning GLM-5V-Turbo: Toward a Native Foundation Model for Multimodal Agents

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.026958Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.026958Z digest=sha256:e371e943ada0d3ffb19f174d821dba1b225c0ba9ae3bca727761ba9761ba8689

Observation 0456dc97-cbd6-472e-9331-b839f796af49 · outbound

This paper cites Visual programming: Compositional visual reasoning without training.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Visual programming: Compositional visual reasoning without training

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.032344Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.032344Z digest=sha256:ef73695d83ccde1278a9ceb125d88d27c6dc8fa107b3f3005d26e8ec5774366d

Observation e15c4967-d4e1-4555-9a7d-b39829d8404d · outbound

This paper cites Training Large Language Models to Reason in a Continuous Latent Space.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Training Large Language Models to Reason in a Continuous Latent Space

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.037446Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.037446Z digest=sha256:745622e18498dfe35a9e095b3cb8035c629bf2577f5d7c3c9ff9f8363df02faf

Observation 661bcc98-115d-42d1-af70-47b2e98f38f0 · outbound

This paper cites DeepEyesV2: Toward Agentic Multimodal Model.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning DeepEyesV2: Toward Agentic Multimodal Model

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.046256Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.046256Z digest=sha256:dfa20c1935af683b964bcd552014ef2e4896c14b83c3259cc3965c9e8914904e

Observation caeb8699-53c6-4525-ae64-b3f61ae401e4 · outbound

This paper cites Hollon, and Bryan Wang.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Hollon, and Bryan Wang

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.052496Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.052496Z digest=sha256:756c275047a608a5a1b8a098346b523b1d1b2eeb91f46bed1cf18b428c41facb

Observation 24edadf1-b96c-40e0-b450-1d1055e80bc6 · outbound

This paper cites Visual Sketchpad: Sketching as a Visual Chain of Thought for Multimodal Language Models.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Visual Sketchpad: Sketching as a Visual Chain of Thought for Multimodal Language Models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.059008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.059008Z digest=sha256:59fbe65ce4087b0b61cb9935a3eabec0ddac9da583270f306f711f7b8e34ec86

Observation 2688c6b2-61ba-4f3a-b127-b6d76bfd8826 · outbound

This paper cites VerlTool: Towards holistic agentic reinforcement learning with tool use.arXiv preprint arXiv:2509.01055, 2025.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning VerlTool: Towards holistic agentic reinforcement learning with tool use.arXiv preprint arXiv:2509.01055, 2025

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.065378Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.065378Z digest=sha256:64d22b5a428ccadb84e60aff594a5e0db04e2934d86a73679dc6f87abf145625

Observation 59f00b1d-851b-48ac-9c49-cf03c313ff85 · outbound

This paper cites Kimi K2.5: Visual Agentic Intelligence.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Kimi K2.5: Visual Agentic Intelligence

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.072026Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.072026Z digest=sha256:c616d20bc821d0184746aa689ee65318d0001f804ce8ceec8c90c12fead16006

Observation b105647b-87f3-4fb7-82bf-6f0446626b1f · outbound

This paper cites Mini-o3: Scaling Up Reasoning Patterns and Interaction Turns for Visual Search.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Mini-o3: Scaling Up Reasoning Patterns and Interaction Turns for Visual Search

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.077712Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.077712Z digest=sha256:f3b95a7f29b7abe319bae821b1d4c107c609b2836b54769b997dc12f05a686d9

Observation 2897de1e-e8b0-45dd-9850-acec5ccdf6bc · outbound

This paper cites Measuring Faithfulness in Chain-of-Thought Reasoning.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Measuring Faithfulness in Chain-of-Thought Reasoning

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.083137Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.083137Z digest=sha256:a34f826c799ca0073d101864d115fb73f41bd04469cf6fad6b5bf3a2b30d1942

Observation 35cf9fd5-68e4-40d2-972e-a5679640d9ca · outbound

This paper cites Zebra-cot: A dataset for interleaved vision language reasoning.arXiv preprint arXiv:2507.16746, 2025.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Zebra-cot: A dataset for interleaved vision language reasoning.arXiv preprint arXiv:2507.16746, 2025

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.088153Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.088153Z digest=sha256:e671d146748bbbbabff07e4ab42e85ad905c2a6c869fb2e8be61f8bf28aff3e7

Observation 76d54afa-1dd1-45c7-bd61-5613d7a3532b · outbound

This paper cites Tir-bench: A comprehensive benchmark for agentic thinking-with-images reasoning.arXiv preprint arXiv:2511.01833, 2025.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Tir-bench: A comprehensive benchmark for agentic thinking-with-images reasoning.arXiv preprint arXiv:2511.01833, 2025

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.092965Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.092965Z digest=sha256:c7b4c3515caba2cb1bc936eeba5b0d2d326c0f72086fc5f7336aa72aa9636606

Observation 6155944d-2c98-4ebf-ac10-b83d52418bbb · outbound

This paper cites Seg-Zero: Reasoning-Chain Guided Segmentation via Cognitive Reinforcement.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Seg-Zero: Reasoning-Chain Guided Segmentation via Cognitive Reinforcement

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.097596Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.097596Z digest=sha256:5bc1c673b0055d01f1460f4f01281c86c5ce40c1043aa5f979fcad9258b3d54f

Observation 8c4c7087-0bf3-4c1c-ba1f-f214947de793 · outbound

This paper cites On the faithfulness of visual thinking: Measurement and enhancement.arXiv preprint arXiv:2510.23482, 2025.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning On the faithfulness of visual thinking: Measurement and enhancement.arXiv preprint arXiv:2510.23482, 2025

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.102763Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.102763Z digest=sha256:968063793b513b1a651d5239e95a6dbd3dbfa9a40d4610da4a3eabe3540b51af

Observation fed3ace4-7f7a-4068-a766-df90f1678d4d · outbound

This paper cites Chain-of-Spot: Interactive Reasoning Improves Large Vision-Language Models.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Chain-of-Spot: Interactive Reasoning Improves Large Vision-Language Models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.107674Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.107674Z digest=sha256:08b53d2129d6d1164bd9a431273e4719614e458bd4ae05f218d61bef7b31dad7

Observation 62e485d8-49de-4c3a-9605-ae3b6734b188 · outbound

This paper cites Chameleon: Plug-and-play compositional reasoning with large language models.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Chameleon: Plug-and-play compositional reasoning with large language models

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.113186Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.113186Z digest=sha256:f9b7efc19ff99c8faed8d08721907ee8cca1e9b24398e83a24627154ac9cc0cd

Observation 9004a401-c7d2-49cc-ba01-b07b1b464f24 · outbound

This paper cites Reasoning Models Can Be Effective Without Thinking.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Reasoning Models Can Be Effective Without Thinking

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.118396Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.118396Z digest=sha256:6be7d4857be01d0cb5435f37ba44c021af5c7a8a1aad4a868934fe69377dfe60

Observation 87e7c981-a900-4a4f-9443-7f23e41f7676 · outbound

This paper cites What Does Vision Tool-Use Reinforcement Learning Really Learn? Disentangling Tool-Induced and Intrinsic Effects for Crop-and-Zoom.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning What Does Vision Tool-Use Reinforcement Learning Really Learn? Disentangling Tool-Induced and Intrinsic Effects for Crop-and-Zoom

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.125300Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.125300Z digest=sha256:ef4cca464b1f4b1ab54e833b6abb26bd3bb44524fdb36628a39b140e338a68e7

Observation 8c9b6ca7-8c46-4542-b286-5c2210ec5976 · outbound

This paper cites ChartQA: A Benchmark for Question Answering about Charts with Visual and Logical Reasoning.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning ChartQA: A Benchmark for Question Answering about Charts with Visual and Logical Reasoning

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.131378Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.131378Z digest=sha256:800fbab03b86f75b3578c79ab539e2bb191c4f5ef299ba0e3e0bbd0db7524f70

Observation d915d33d-2552-490f-b362-54b7e010fd32 · outbound

This paper cites GSM-Symbolic: Understanding the Limitations of Mathematical Reasoning in Large Language Models.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning GSM-Symbolic: Understanding the Limitations of Mathematical Reasoning in Large Language Models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.137144Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.137144Z digest=sha256:bd4075f794a268bf5d9fd56c483f85498dbb7b1fb0714979328094a2612a70c4

Observation 38d5594e-04c8-45a8-9b98-56a44e4122d7 · outbound

This paper cites Show Your Work: Scratchpads for Intermediate Computation with Language Models.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Show Your Work: Scratchpads for Intermediate Computation with Language Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.143566Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.143566Z digest=sha256:33e0042b9c8366de077ef1809334b51b8ea1b3cf18ea4471abda81b24659679b

Observation ab05b758-cb33-47e7-92ea-c69c2c1dc7d4 · outbound

This paper cites GPT-4 Technical Report.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning GPT-4 Technical Report

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.149920Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.149920Z digest=sha256:8761f17ea5217275a8e42222152b72f7eae15dc37df5a38a1f6dc9d450d47b68

Observation 29b9e7f1-eedf-448c-beb2-59ffb461b6e5 · outbound

This paper cites Thinking with images.https://openai.com/index/thinking-with-images/, 2025.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Thinking with images.https://openai.com/index/thinking-with-images/, 2025

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.156357Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.156357Z digest=sha256:4280f32068d2dca55e644800766df0507a5b9a6639a2f6a3b12559cb56ce7862

Observation 7a9f1ba0-82a1-4a24-ae45-3cfb98252cde · outbound

This paper cites Do MLLMs really see it: Reinforcing visual attention in multimodal LLMs.arXiv preprint arXiv:2602.08241, 2026.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Do MLLMs really see it: Reinforcing visual attention in multimodal LLMs.arXiv preprint arXiv:2602.08241, 2026

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.161693Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.161693Z digest=sha256:168d1e2a309f3b664dc79dfa59ce4a77d6c5fc0553d874088dd032c935fc1b81

Observation c9a368e4-f37a-4ffd-b1d6-20ad3b5f427f · outbound

This paper cites CogCoM: A Visual Language Model with Chain-of-Manipulations Reasoning.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning CogCoM: A Visual Language Model with Chain-of-Manipulations Reasoning

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.169132Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.169132Z digest=sha256:1498bb882b313f4cb72f016acb5f3589307b5248176a73153337efef38590ef1

Observation d47f3c4d-f832-4c6b-a745-c38dbd5f852b · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.175515Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.175515Z digest=sha256:6721f1c34584b64a47fa896994e5ddda9945f1eaa2d0f15fe6f54b40234796af

Observation 25ad21de-9c81-4709-b03a-9f0f39708c29 · outbound

This paper cites Qwen2.5-VL Technical Report.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Qwen2.5-VL Technical Report

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.182177Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.182177Z digest=sha256:ac90a802f29a04f8b97a77ed715a6e4e07804deb09651b9ec12181c18d59a4f3

Observation 8c0aada3-d18d-4ebe-83af-b05f2c1bd462 · outbound

This paper cites Qwen3-VL Technical Report.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Qwen3-VL Technical Report

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.188571Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.188571Z digest=sha256:86eae1efd8a8eedf5f7ff47180c34fadad827d2e2b56b24f4916143ff8d86391

Observation fa21b7df-5558-4e33-8163-a62a77acc62f · outbound

This paper cites Vision language models are blind: Failing to translate detailed visual features into words.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Vision language models are blind: Failing to translate detailed visual features into words

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.194557Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.194557Z digest=sha256:24fa8c75f6db889e3aca37d23ca52f43da166a2788e14f1f349331f0adadfee3

Observation b553977f-9a03-4e10-9724-7816c539dd03 · outbound

This paper cites Grounded Reinforcement Learning for Visual Reasoning.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Grounded Reinforcement Learning for Visual Reasoning

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.200723Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.200723Z digest=sha256:7c54cb959899f67a4cdab70d43cb4cb90e1062dfab1ae950edca7e0a09cc9579

Observation 71aed7c4-6d94-4238-a052-da220eee155a · outbound

This paper cites Toolformer: Language models can teach themselves to use tools.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Toolformer: Language models can teach themselves to use tools

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.206226Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.206226Z digest=sha256:02242072fbc8975c3882891121471678ca4298d52f65819f74d3b985733fb869

Observation eb554f4a-9642-4522-ac1f-13538d74085c · outbound

This paper cites Visual CoT: Advancing Multi-Modal Language Models with a Comprehensive Dataset and Benchmark for Chain-of-Thought Reasoning.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Visual CoT: Advancing Multi-Modal Language Models with a Comprehensive Dataset and Benchmark for Chain-of-Thought Reasoning

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.211374Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.211374Z digest=sha256:dcc3497d5b8fb0a4d4c16001cca969bfbf5859a5a39a0ceda8ebcc4bf95b6dcf

Observation 36606029-2231-4dbc-87b7-538e2926826a · outbound

This paper cites HuggingGPT: Solving AI tasks with ChatGPT and its friends in Hugging Face.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning HuggingGPT: Solving AI tasks with ChatGPT and its friends in Hugging Face

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.216761Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.216761Z digest=sha256:adee50622222ce9ff9eae8fa155d87bc6ae0c1a366b67e554ddbd2dd4ad10fb2

Observation d80610d9-2730-4bd8-bb41-1c6a246f93d8 · outbound

This paper cites HybridFlow: A Flexible and Efficient RLHF Framework.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning HybridFlow: A Flexible and Efficient RLHF Framework

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.221402Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.221402Z digest=sha256:bacc006ada2308e499304942ce42ef3e67b640994af0543711f15ed1681d0d4f

Observation 7039400c-ff0f-4515-a00c-e38454fb4b39 · outbound

This paper cites Reflexion: Language agents with verbal reinforcement learning.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Reflexion: Language agents with verbal reinforcement learning

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.226408Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.226408Z digest=sha256:849cc7b7a1412e15cbe353b66e0c644a21dfef2d39ceee25925251d0d1e06270

Observation bd7094a6-99ef-4afb-b609-325ea8259745 · outbound

This paper cites Breaking the Chain: A Causal Analysis of LLM Faithfulness to Intermediate Structures.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Breaking the Chain: A Causal Analysis of LLM Faithfulness to Intermediate Structures

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.231434Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.231434Z digest=sha256:b24f58387478d82f3a251ebb492b58a644c19d90bfba1d3ffe4527ab5d2a79e8

Observation 2f895714-0304-4781-afac-8325cbee7e5e · outbound

This paper cites OpenThinkIMG: Learning to Think with Images via Visual Tool Reinforcement Learning.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning OpenThinkIMG: Learning to Think with Images via Visual Tool Reinforcement Learning

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.237048Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.237048Z digest=sha256:55e8b4729e721d1ec5f4435a8dee97fa90bb3227d55926145f4147a294ea64b7

Observation 6adf3fcd-a92f-4e84-9f6c-f35f8fdca13d · outbound

This paper cites Thinking with Images for Multimodal Reasoning: Foundations, Methods, and Future Frontiers.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Thinking with Images for Multimodal Reasoning: Foundations, Methods, and Future Frontiers

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.243692Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.243692Z digest=sha256:3d3df68dee01054f40ec0e01d3d3098f3492656381086e5481f9350303ea15ba

Observation 808134bd-c4aa-41ce-8ced-e241ff856bee · outbound

This paper cites When thinking hurts: Mitigating visual forgetting via frame repetition.arXiv preprint arXiv:2603.16256, 2026.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning When thinking hurts: Mitigating visual forgetting via frame repetition.arXiv preprint arXiv:2603.16256, 2026

Reference 56

Resolution
verified exact
raw_fallback, observed 2026-08-11T12:53:16.994849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T12:53:14.249540Z digest=sha256:932039619bd466805fb2903a103e6d412b6c349037fa7ba3790e11945afb75dd

Observation 33cfd61d-327a-47bf-831c-7f2f998d7ef9 · outbound

This paper cites FACT-E: Causality-Inspired Evaluation for Trustworthy Chain-of-Thought Reasoning.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning FACT-E: Causality-Inspired Evaluation for Trustworthy Chain-of-Thought Reasoning

Reference 57

Resolution
verified exact
local_arxiv, observed 2026-08-11T12:53:16.902776Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T12:53:14.254925Z digest=sha256:350c0aff950f20a8270d2111aaa3e107819273098ff0552802fded013212834c

Observation a6faa050-844b-4cd5-9dc9-5a37b4bcd364 · outbound

This paper cites ViperGPT: Visual inference via python execution for reasoning.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning ViperGPT: Visual inference via python execution for reasoning

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.260904Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.260904Z digest=sha256:c2e34e769adcecaac6675d68894f81466e5d522755230f1d2a95636ccd0de909

Observation c664eea2-5da5-491d-a904-af12e5d86968 · outbound

This paper cites CV-Bench: A computer vision benchmark for evaluating visual perception in multimodal language models.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning CV-Bench: A computer vision benchmark for evaluating visual perception in multimodal language models

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.266490Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.266490Z digest=sha256:a1db71e9a42a64adb3fc765fc7546f6c0183e3dd2003624f718cf9ffd30668be

Observation 984d86a0-d722-4cdd-a983-895defd786b0 · outbound

This paper cites an unresolved cited work.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Unresolved cited work

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.271924Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.271924Z digest=sha256:c89313382412dd8ac694896d881df4f8beb1a1ad9cd911486d8a2cba2e5d18b8

Observation cb4caa7c-52de-4d80-8dba-fa757ff24771 · outbound

This paper cites Journey before destination: Visual faithfulness in slow thinking.arXiv preprint arXiv:2512.12218, 2025.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Journey before destination: Visual faithfulness in slow thinking.arXiv preprint arXiv:2512.12218, 2025

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.277132Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.277132Z digest=sha256:104d6df464a0fcbde5012e69871dd501d67250de1492788895f9783ede868a35

Observation aa0df45c-08c5-43d0-8474-d7483c923177 · outbound

This paper cites GeoEyes: On-demand visual focusing for ultra-high-resolution remote sensing.arXiv preprint arXiv:2602.14201, 2026.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning GeoEyes: On-demand visual focusing for ultra-high-resolution remote sensing.arXiv preprint arXiv:2602.14201, 2026

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.282612Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.282612Z digest=sha256:d73e3e517b72176998186986559edd390ac34ffc783dcb17d5e720343ce8dc1f

Observation 0fcc6dce-7fe1-493b-953e-1838a6248d78 · outbound

This paper cites Pixel Reasoner: Incentivizing Pixel-Space Reasoning with Curiosity-Driven Reinforcement Learning.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Pixel Reasoner: Incentivizing Pixel-Space Reasoning with Curiosity-Driven Reinforcement Learning

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.289490Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.289490Z digest=sha256:03139b0e7103d4d9d93f65fd24d551454561778a0ba41b1175d4edb813bcf60d

Observation 431c9e90-1e5c-4855-8d17-1096829a3f45 · outbound

This paper cites PLaT: Latent chain-of-thought as planning: Decoupling reasoning from verbalization.arXiv preprint arXiv:2601.21358, 2026.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning PLaT: Latent chain-of-thought as planning: Decoupling reasoning from verbalization.arXiv preprint arXiv:2601.21358, 2026

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.295877Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.295877Z digest=sha256:e54d552619819b5b60474240368ded917f4714a79d7faabab57883675d4c3a1b

Observation 13f7fe4e-f9e3-4b26-924d-51d354c8d659 · outbound

This paper cites VAGEN: Reinforcing world model reasoning for multi-turn VLM agents.arXiv preprint arXiv:2510.16907, 2025.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning VAGEN: Reinforcing world model reasoning for multi-turn VLM agents.arXiv preprint arXiv:2510.16907, 2025

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.301424Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.301424Z digest=sha256:735f08e67767ef0c6d2c0c5058da80af2fcfebb7558ab2ddf7e2bf638adc68e5

Observation 155234b5-f117-43e0-ac3f-685adee465c4 · outbound

This paper cites Plan-and-solve prompting: Improving zero-shot chain-of-thought reasoning by large language models.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Plan-and-solve prompting: Improving zero-shot chain-of-thought reasoning by large language models

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.307438Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.307438Z digest=sha256:ff667a0f1980bc87d031c7bd5a3d375aa3456bc6cd7ef14ef24f1505147a4e52

Observation ad4de20a-fa87-4e81-a3f1-935df8556f00 · outbound

This paper cites A practitioner’s guide to multi-turn agentic reinforcement learning.arXiv preprint arXiv:2510.01132, 2025.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning A practitioner’s guide to multi-turn agentic reinforcement learning.arXiv preprint arXiv:2510.01132, 2025

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.313080Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.313080Z digest=sha256:426d680ae672d3be2591510ba29979d121656c9ea4684774379d6066023ea0ef

Observation 7e8456bd-7ef2-4fdb-86cb-b3b5d1e618a7 · outbound

This paper cites Divide, Conquer and Combine: A Training-Free Framework for High-Resolution Image Perception in Multimodal Large Language Models.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Divide, Conquer and Combine: A Training-Free Framework for High-Resolution Image Perception in Multimodal Large Language Models

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.318402Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.318402Z digest=sha256:1ee135c5610759ec3c79ce867c116781165eada4511654aef1d97bb172b06efc

Observation f574102a-4bd5-4cf0-aaab-eeb6bc7aa745 · outbound

This paper cites Self-consistency improves chain of thought reasoning in language models.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Self-consistency improves chain of thought reasoning in language models

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.324074Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.324074Z digest=sha256:86ab3d367fec0b89373820fef1f9ec2baeb24cf30773cf170184101eda861d60

Observation 11876701-c04c-48cb-9602-4d5af4eb6f62 · outbound

This paper cites Simple o3: Towards Interleaved Vision-Language Reasoning.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Simple o3: Towards Interleaved Vision-Language Reasoning

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.330057Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.330057Z digest=sha256:e3ada83b772ba52ef2d4cadfef3f463e9f256dad92a79707252d5d16e1dea5f5

Observation 672af4c4-9a5f-40d2-abde-ad12cb180fd6 · outbound

This paper cites RAGEN: Understanding Self-Evolution in LLM Agents via Multi-Turn Reinforcement Learning.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning RAGEN: Understanding Self-Evolution in LLM Agents via Multi-Turn Reinforcement Learning

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.335916Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.335916Z digest=sha256:8c60339b21e4c29d0e33615f6d9bc4714693634e04b519fa19ca35d0fd175d39

Observation 58bc7e5f-585b-444a-b826-0f46909038fb · outbound

This paper cites CharXiv: Charting Gaps in Realistic Chart Understanding in Multimodal LLMs.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning CharXiv: Charting Gaps in Realistic Chart Understanding in Multimodal LLMs

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.341547Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.341547Z digest=sha256:6d64c9605a7f3db7129c1cc45a13ba61edf5c8dcedb864c44e8db8a85cc70b71

Observation f6834941-8cc1-470f-bd19-d04bb69d18b4 · outbound

This paper cites V-FAT: Benchmarking visual fidelity against text-bias.arXiv preprint arXiv:2601.04897, 2025.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning V-FAT: Benchmarking visual fidelity against text-bias.arXiv preprint arXiv:2601.04897, 2025

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.346906Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.346906Z digest=sha256:7e2d27dbde994c3f77b128f64571192cc3d5ed5613ff75edb62621b40f20686c

Observation 75cb0d3c-04e6-46b5-a8e7-454ba0800c7b · outbound

This paper cites Chi, Quoc V.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Chi, Quoc V

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.352064Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.352064Z digest=sha256:164c73b9fa39f0b1cdd3ff9ba912956f276ae3b7b7a013300d1dd055c88149f8

Observation 8614fd79-8787-43d5-83b9-21608da229f4 · outbound

This paper cites Zooming without zooming: Region- to-image distillation for fine-grained multimodal perception.arXiv preprint arXiv:2602.11858, 2026.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Zooming without zooming: Region- to-image distillation for fine-grained multimodal perception.arXiv preprint arXiv:2602.11858, 2026

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.357126Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.357126Z digest=sha256:52e7650034125cbf84b1e386745ff3b9f9d48d9ab0eb9c5c2a54ba69dc6677d4

Observation 10131f61-7daf-46c1-82f6-f0f69697db60 · outbound

This paper cites Visual generation unlocks human-like reasoning through multimodal world models.arXiv preprint arXiv:2601.19834, 2026.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Visual generation unlocks human-like reasoning through multimodal world models.arXiv preprint arXiv:2601.19834, 2026

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.362216Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.362216Z digest=sha256:432c4dc24550d0e855c09961ec6ddf945366c87d4dd5fd672d0cbbc934b453c1

Observation 812332c9-4101-4395-ad7a-532e7f106164 · outbound

This paper cites VTool-R1: VLMs learn to think with images via reinforcement learning on multimodal tool use.arXiv preprint arXiv:2505.19255, 2025.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning VTool-R1: VLMs learn to think with images via reinforcement learning on multimodal tool use.arXiv preprint arXiv:2505.19255, 2025

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.367385Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.367385Z digest=sha256:6db090aec657ab9a9570b461722b3ca968629965d2295d120959b6d8f2a0b59c

Observation c1358137-74bd-4884-b10f-9687c7390116 · outbound

This paper cites V*: Guided Visual Search as a Core Mechanism in Multimodal LLMs.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning V*: Guided Visual Search as a Core Mechanism in Multimodal LLMs

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.372654Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.372654Z digest=sha256:5cdd3227627c5f0bc2140c037eb550c69d29084163262b81a786c08926ea6a82

Observation 39fdb19d-0830-48e0-bb5b-af6e36d48e4a · outbound

This paper cites Tool-augmented policy optimization.arXiv preprint arXiv:2510.07038, 2025.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Tool-augmented policy optimization.arXiv preprint arXiv:2510.07038, 2025

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.378336Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.378336Z digest=sha256:b04497e1f08bb46ec0c1c05e81d2c6d32795babfd0c10e07f2f1e599d2989a6c

Observation d2e8fa83-702f-44e8-bdb0-0b664623ccfd · outbound

This paper cites Reasoning or Reciting? Exploring the Capabilities and Limitations of Language Models Through Counterfactual Tasks.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Reasoning or Reciting? Exploring the Capabilities and Limitations of Language Models Through Counterfactual Tasks

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.384423Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.384423Z digest=sha256:329445502aba604959a32e6655c7eaa8770f584ab43b4eef85a0d823b8ff67d5

Observation e0bf9a99-100f-4ea1-aa7d-15c558030ec7 · outbound

This paper cites LLaVA-CoT: Let Vision Language Models Reason Step-by-Step.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning LLaVA-CoT: Let Vision Language Models Reason Step-by-Step

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.390585Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.390585Z digest=sha256:5e2847161a5f026fa6f8dd84cf5b5eda07878472663773797f4fc9715f502ee3

Observation 68cb5366-89ba-4258-9f1a-99d967af8c4c · outbound

This paper cites Act Wisely: Cultivating Meta-Cognitive Tool Use in Agentic Multimodal Models.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Act Wisely: Cultivating Meta-Cognitive Tool Use in Agentic Multimodal Models

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.396078Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.396078Z digest=sha256:faeb8bf24fd976a205f4431ad842a66501594e3220164be1885cc4864f7949cd

Observation 51316cec-d858-4612-af50-38963eb96922 · outbound

This paper cites Set-of-Mark Prompting Unleashes Extraordinary Visual Grounding in GPT-4V.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Set-of-Mark Prompting Unleashes Extraordinary Visual Grounding in GPT-4V

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.403481Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.403481Z digest=sha256:6fd663d3e03cd877e071dcc71be871e91a2717f929e6e78a65680b0f2f4e0714

Observation c22ace32-3b90-4012-b5f8-9e466331462e · outbound

This paper cites Thinking in Space: How Multimodal Large Language Models See, Remember, and Recall Spaces.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Thinking in Space: How Multimodal Large Language Models See, Remember, and Recall Spaces

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.409420Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.409420Z digest=sha256:1e1f963297e196ee63e155e8469f8411d7aef2b9df61a447a696eadab634b854

Observation 5e606f51-b262-42b5-8073-9e895eb919f4 · outbound

This paper cites VisionThink: Smart and Efficient Vision Language Model via Reinforcement Learning.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning VisionThink: Smart and Efficient Vision Language Model via Reinforcement Learning

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.415257Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.415257Z digest=sha256:4941e439608db4f2c01a9f17e76fc3c88808de8d883a9d3fd7a0c413c1addde9

Observation af5e4430-ad7c-4533-b666-a57b79c4825d · outbound

This paper cites Look-Back: Implicit Visual Re-focusing in MLLM Reasoning.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Look-Back: Implicit Visual Re-focusing in MLLM Reasoning

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.422649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.422649Z digest=sha256:78b1b1aa7c794c42d4b14d6f53eb477b953e1fe05e42d48588ea3b891c031603

Observation c46a5d03-1cee-489d-98bd-52fb90f9d5b4 · outbound

This paper cites Walk the Talk: Bridging the Reasoning-Action Gap for Thinking with Images via Multimodal Agentic Policy Optimization.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Walk the Talk: Bridging the Reasoning-Action Gap for Thinking with Images via Multimodal Agentic Policy Optimization

Reference 87

Resolution
verified exact
local_arxiv, observed 2026-08-11T12:53:15.665161Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T12:53:14.428800Z digest=sha256:bedce7c5f612c957c5512afee6923f9219d8551160c2558ad31ff3061e57c48e

Observation 6dc2fa7b-8836-4961-aa4b-e74d4e6e85d4 · outbound

This paper cites Thinking with images via self-calling agent.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Thinking with images via self-calling agent

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.434700Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.434700Z digest=sha256:6c99a830b9894093d4a94704b7006f430b7d95fcb5619ac750ab93c345b147d0

Observation ebbc4565-7328-45a9-9e9e-c80012b9984e · outbound

This paper cites Tree of thoughts: Deliberate problem solving with large language models.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Tree of thoughts: Deliberate problem solving with large language models

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.440324Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.440324Z digest=sha256:4df54f7245c418a324d6d8f44a5780d70ba4e3cd8824531ecf9ba2f65928b7ba

Observation 6e03bf6e-d596-4191-833a-6f05b55e852a · outbound

This paper cites React: Synergizing reasoning and acting in language models.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning React: Synergizing reasoning and acting in language models

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.445645Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.445645Z digest=sha256:c1be46188c1b705b4161096072cabb4e76a2daf6607bc1332d722ceed22ef3b0

Observation bb86f08f-a4d6-4b04-937e-b32d0aaac4c4 · outbound

This paper cites an unresolved cited work.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Unresolved cited work

Reference 91

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.450944Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.450944Z digest=sha256:9d2bdeedd3f4634f51c26e7e41da6c3b6ebe98293dadd93caaa138bdd822a2ee

Observation 93488707-e50d-4f90-8a7d-95b8d82a412d · outbound

This paper cites ProRL agent: Rollout-as-a-service for reinforcement learning training of multi-turn LLM agents.arXiv preprint arXiv:2603.18815, 2026.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning ProRL agent: Rollout-as-a-service for reinforcement learning training of multi-turn LLM agents.arXiv preprint arXiv:2603.18815, 2026

Reference 92

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.456716Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.456716Z digest=sha256:96e3231b17e5cb368340cf553835668fbe8dfc9e528a8a9be4969bc9f0522d68

Observation 1da67c1b-6c4a-472f-93ea-6f3d8f059e62 · outbound

This paper cites MM-CoT: A benchmark for probing visual chain-of-thought reasoning.arXiv preprint arXiv:2512.08228, 2025.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning MM-CoT: A benchmark for probing visual chain-of-thought reasoning.arXiv preprint arXiv:2512.08228, 2025

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.462467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.462467Z digest=sha256:fbec21d68cf352f23824e8592ed50d2a003e9df2b62823652984d2513809f2c3

Observation 1cde152f-f8de-45dd-8e0e-f00f4a94f582 · outbound

This paper cites LLaVA-Mini: Efficient Image and Video Large Multimodal Models with One Vision Token.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning LLaVA-Mini: Efficient Image and Video Large Multimodal Models with One Vision Token

Reference 94

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.468007Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.468007Z digest=sha256:e8affee220553000d61af987bcc9dc788ba53f7bc954c0b72b6f7fad78dd0b94

Observation ef2d5d0a-b1e9-435e-9982-f1221d152fad · outbound

This paper cites MME-RealWorld: Could Your Multimodal LLM Challenge High-Resolution Real-World Scenarios that are Difficult for Humans?.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning MME-RealWorld: Could Your Multimodal LLM Challenge High-Resolution Real-World Scenarios that are Difficult for Humans?

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.474606Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.474606Z digest=sha256:a66714f7c2a9dcb8b48d580f239caa28fd5e695db505302158dc0be2b49a2142

Observation c12feb7b-3ba7-47e1-a0bc-5cb150c124f2 · outbound

This paper cites Thyme: Think Beyond Images.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Thyme: Think Beyond Images

Reference 96

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.480733Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.480733Z digest=sha256:26df420690769cec4a8b569e2fe7afee4f5c0c77e7d03bf87516ded7f0a3d81e

Observation 4fb149dd-5d29-49f4-a574-6c1c48b07001 · outbound

This paper cites Skywork-r1v4: Toward agentic multimodal intelligence through interleaved thinking with images and deepresearch.arXiv preprint arXiv:2512.02395, 2025.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Skywork-r1v4: Toward agentic multimodal intelligence through interleaved thinking with images and deepresearch.arXiv preprint arXiv:2512.02395, 2025

Reference 97

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.486447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.486447Z digest=sha256:9ce04f5153c8d51067c761413b5828e84b2c669e0fc93afc94a86ab1332d3be9

Observation 3a376dd0-d944-496d-b3da-cffe00b77a20 · outbound

This paper cites CM2: Reinforcement learning with checklist rewards for multi-turn and multi-step agentic tool use.arXiv preprint arXiv:2602.12268, 2026.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning CM2: Reinforcement learning with checklist rewards for multi-turn and multi-step agentic tool use.arXiv preprint arXiv:2602.12268, 2026

Reference 98

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.491968Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.491968Z digest=sha256:147356f91591a8df608b106a3d003b473750be9b65c20769da4c5b8b78895899

Observation abebf7fd-87a9-4b5f-ad26-07e75afc7002 · outbound

This paper cites Multimodal Chain-of-Thought Reasoning in Language Models.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Multimodal Chain-of-Thought Reasoning in Language Models

Reference 99

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.497323Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.497323Z digest=sha256:5ebe9019c24b55aa82c2a048bcd3e04ec9188982f5f151b12556dbfb48b92a87

Observation 754fb110-bb38-4730-85a6-4d1842ee219b · outbound

This paper cites On Robustness and Chain-of-Thought Consistency of RL-Finetuned VLMs.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning On Robustness and Chain-of-Thought Consistency of RL-Finetuned VLMs

Reference 100

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.503030Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.503030Z digest=sha256:e57978d8eb70edd85eb887255c7d31dfd9903f9735b23fe62a90ce17a80865d2

Pith citing papers

No inbound Pith citation observations are available.