Pith. sign in

Paper Citation Record · LEDGER

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning

As of 12 August 2026, this Paper Citation Record lists 100 of 129 outbound references and 0 inbound Pith citation observations for arXiv:2608.09682.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.09682 v1

Coverage vector

measured 100 of 129 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T12:53:14.503030Z

measured 100 of 100 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

100 of 129 outbound references displayed

  • verified exact3
  • verified fuzzy0
  • unresolved97
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 21d126e9-bc08-49be-ac7a-28c7872f3451 · outbound

This paper cites Latent reasoning with supervised thinking states.arXiv preprint arXiv:2602.08332, 2026.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Latent reasoning with supervised thinking states.arXiv preprint arXiv:2602.08332, 2026

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:13.922601Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:13.922601Z digest=sha256:555d0f61e87baeddda61de6ff42232361dda89531c3b9d0bca1e1ead38bf3b3c

Observation 3b40b537-9d4c-4dae-87c7-72e0cbaead1d · outbound

This paper cites Acloserlookatbiasandchain-of-thought faithfulness of large (vision) language models.arXiv preprint arXiv:2505.23945, 2025.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Acloserlookatbiasandchain-of-thought faithfulness of large (vision) language models.arXiv preprint arXiv:2505.23945, 2025

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:13.930040Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:13.930040Z digest=sha256:d774f73b79bb17eedeab297d015a920cf4f4470f5e68c23d239fbde9512871a0

Observation 54c530f6-4381-45b3-8282-d55da1520d61 · outbound

This paper cites An Image is Worth 1/2 Tokens After Layer 2: Plug-and-Play Inference Acceleration for Large Vision-Language Models.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning An Image is Worth 1/2 Tokens After Layer 2: Plug-and-Play Inference Acceleration for Large Vision-Language Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:13.935715Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:13.935715Z digest=sha256:e65a33ec668729d99d4a1985b10d2fa1d76c620e84be9792618c197c645bab3f

Observation 4127ec8f-cb29-49ef-bf11-6a0e20ae4c03 · outbound

This paper cites MMStar: An evaluator-centric benchmark for multi-modal large language models, 2024.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning MMStar: An evaluator-centric benchmark for multi-modal large language models, 2024

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:13.941754Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:13.941754Z digest=sha256:6b0b5ebfc84341ed89169f4b14e0900a1b15a8920207037a3d4e1650505b0060

Observation e3a06ba7-1a84-4d28-9062-55bd6db2c470 · outbound

This paper cites Perception before reasoning: Two-stage reinforcement learning for visual reasoning.arXiv preprint arXiv:2509.13031, 2025.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Perception before reasoning: Two-stage reinforcement learning for visual reasoning.arXiv preprint arXiv:2509.13031, 2025

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:13.947142Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:13.947142Z digest=sha256:ed0931e20f274772836998df2adce52da04809a5244ee832fe9b40e3616c4125

Observation c1d6048d-a793-41d2-8a99-295b545a062e · outbound

This paper cites Reasoning Models Don't Always Say What They Think.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Reasoning Models Don't Always Say What They Think

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:13.952698Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:13.952698Z digest=sha256:4743582a5daa2ad02ff59049a769e447b21dcee913bb9b5b6f2d588ebdbc2bce

Observation 527f98d2-c979-4acc-b913-fbf7eec197af · outbound

This paper cites v1: Learning to Point Visual Tokens for Multimodal Grounded Reasoning.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning v1: Learning to Point Visual Tokens for Multimodal Grounded Reasoning

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:13.959442Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:13.959442Z digest=sha256:533178c1789d50fbae3919864873e2675cb5dfabcceba5f9d676a9baaf1bb757

Observation 4b2e1502-e09d-4db2-af0d-9f9e3250c7e9 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:13.964796Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:13.964796Z digest=sha256:b4895bb754e2fe43a28ccc1a22ab1cdc56a86db6e81467cab09c94b8dad46618

Observation 509a382a-9523-41ba-98f9-4c79aa25a0ef · outbound

This paper cites Molmo and PixMo: Open Weights and Open Data for State-of-the-Art Vision-Language Models.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Molmo and PixMo: Open Weights and Open Data for State-of-the-Art Vision-Language Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:13.970205Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:13.970205Z digest=sha256:6c8b04311f0609a5b7ab254c503b7cf08b92f9c66fdfcfb0995fbd7ef2898315

Observation 93107529-ff1e-4d89-8ff6-a0d9f5579f93 · outbound

This paper cites Virgo: A Preliminary Exploration on Reproducing o1-like MLLM.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Virgo: A Preliminary Exploration on Reproducing o1-like MLLM

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:13.978385Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:13.978385Z digest=sha256:1b58a57f98216fab3ed2cb40b5231581d9e9ef620b516309fd85f122c9ff868c

Observation 65d3ed44-1f15-4843-b40c-8316c5ac469c · outbound

This paper cites Revisiting the necessity of lengthy chain-of-thought in vision-centric reasoning.arXiv preprint arXiv:2511.22586, 2025.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Revisiting the necessity of lengthy chain-of-thought in vision-centric reasoning.arXiv preprint arXiv:2511.22586, 2025

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:13.984603Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:13.984603Z digest=sha256:a51ca3283db52a930f8fbc117e39ae63ebdb724d92915fb76a84b3a5a355d428

Observation be44b9db-1f68-4022-8868-c22a5ce9043b · outbound

This paper cites VLMEvalKit: An open-source toolkit for evaluating large multi-modality models.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning VLMEvalKit: An open-source toolkit for evaluating large multi-modality models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:13.990131Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:13.990131Z digest=sha256:cd1c64727dc4f5db21084844d5f4645d15a81238fed3483ddf4dba981a911f2f

Observation 7931aa8e-29c5-4a9f-9949-92b81a812407 · outbound

This paper cites GRIT: Teaching MLLMs to Think with Images.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning GRIT: Teaching MLLMs to Think with Images

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:13.996266Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:13.996266Z digest=sha256:ce10c049e266a1fc8de06bed70e204a9988a64dab695ccad9988e24a31c08b72

Observation 9056dd3f-3141-49d4-8937-2bf9ac597373 · outbound

This paper cites Reward Shaping to Mitigate Reward Hacking in RLHF.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Reward Shaping to Mitigate Reward Hacking in RLHF

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.002043Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.002043Z digest=sha256:c6b9c515390bd840464cebb81a10432a4ef72947da8589b317fd986b533e2269

Observation 195d31a5-11c3-492e-af18-b14ebf648ec6 · outbound

This paper cites BLINK: Multimodal Large Language Models Can See but Not Perceive.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning BLINK: Multimodal Large Language Models Can See but Not Perceive

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.008737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.008737Z digest=sha256:ca7e6478a9798208eefbbb93ea0c14b66da7bf6e87cedeb71e3492b10ab011c4

Observation 64d7c8d5-a9f5-446f-ab58-b427e8915449 · outbound

This paper cites Thinking with deltas: Incen- tivizingreinforcementlearningviadifferentialvisualreasoningpolicy.arXivpreprintarXiv:2601.06801, 2026.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Thinking with deltas: Incen- tivizingreinforcementlearningviadifferentialvisualreasoningpolicy.arXivpreprintarXiv:2601.06801, 2026

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.016201Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.016201Z digest=sha256:9cced6feabf1a2510a31fe68d9fb59fe11310dc06f08216a618a85b1414fa4d7

Observation 068e1c91-7dab-46d4-be3a-3a5502407ccd · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Gemini: A Family of Highly Capable Multimodal Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.021564Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.021564Z digest=sha256:ea77151e7f01cd560c2425adc1083ed55a58299aae5c3427087a2c3d04e93c92

Observation a332fe5b-83bd-4f84-a92e-1559c14595f7 · outbound

This paper cites GLM-5V-Turbo: Toward a Native Foundation Model for Multimodal Agents.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning GLM-5V-Turbo: Toward a Native Foundation Model for Multimodal Agents

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.026958Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.026958Z digest=sha256:c3fd604a1ae4a569c7faef7459a66e9192a6a3684e85a496db0ea51d6e67371d

Observation 0456dc97-cbd6-472e-9331-b839f796af49 · outbound

This paper cites Visual programming: Compositional visual reasoning without training.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Visual programming: Compositional visual reasoning without training

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.032344Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.032344Z digest=sha256:a10d7235eaf50f720922fcc8fb5d6c46d8e851d3ba4e051276ded1dd2c9da009

Observation e15c4967-d4e1-4555-9a7d-b39829d8404d · outbound

This paper cites Training Large Language Models to Reason in a Continuous Latent Space.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Training Large Language Models to Reason in a Continuous Latent Space

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.037446Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.037446Z digest=sha256:53690a94c58050abf363448fd19df95fa4e4cc6580f71db64733a6f7f4a93935

Observation 661bcc98-115d-42d1-af70-47b2e98f38f0 · outbound

This paper cites DeepEyesV2: Toward Agentic Multimodal Model.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning DeepEyesV2: Toward Agentic Multimodal Model

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.046256Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.046256Z digest=sha256:624bae17249ab15570911f13570829510bf6d53ce3184c4ab002992f22347505

Observation caeb8699-53c6-4525-ae64-b3f61ae401e4 · outbound

This paper cites Hollon, and Bryan Wang.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Hollon, and Bryan Wang

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.052496Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.052496Z digest=sha256:6ae54b73d39e6c690c0a00d21e4ec0b58aa0b3fe23de85f7a93dd537d4c5dfd4

Observation 24edadf1-b96c-40e0-b450-1d1055e80bc6 · outbound

This paper cites Visual Sketchpad: Sketching as a Visual Chain of Thought for Multimodal Language Models.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Visual Sketchpad: Sketching as a Visual Chain of Thought for Multimodal Language Models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.059008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.059008Z digest=sha256:27354e0fdab877c95eb4788e34b51166a57ed23dea650b46b3868b662fa55461

Observation 2688c6b2-61ba-4f3a-b127-b6d76bfd8826 · outbound

This paper cites VerlTool: Towards holistic agentic reinforcement learning with tool use.arXiv preprint arXiv:2509.01055, 2025.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning VerlTool: Towards holistic agentic reinforcement learning with tool use.arXiv preprint arXiv:2509.01055, 2025

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.065378Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.065378Z digest=sha256:c80f4309fd859fa633c7210e1a96b6fb00f7c128707510d3be605dc408a91edb

Observation 59f00b1d-851b-48ac-9c49-cf03c313ff85 · outbound

This paper cites Kimi K2.5: Visual Agentic Intelligence.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Kimi K2.5: Visual Agentic Intelligence

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.072026Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.072026Z digest=sha256:2018804fe5b7f7f0ea481ddf1bf922cdd6033defcdfd5a00bf7a78d74a1ad9a0

Observation b105647b-87f3-4fb7-82bf-6f0446626b1f · outbound

This paper cites Mini-o3: Scaling Up Reasoning Patterns and Interaction Turns for Visual Search.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Mini-o3: Scaling Up Reasoning Patterns and Interaction Turns for Visual Search

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.077712Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.077712Z digest=sha256:b3fb83be97ee80e9ebf15a654989e4a68f62519da8e2759b50a8d0dd8204f398

Observation 2897de1e-e8b0-45dd-9850-acec5ccdf6bc · outbound

This paper cites Measuring Faithfulness in Chain-of-Thought Reasoning.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Measuring Faithfulness in Chain-of-Thought Reasoning

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.083137Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.083137Z digest=sha256:04ba0bf25a13a839157484f5751bc7afbaecc2717bf71ca69e9799ef0e2143f7

Observation 35cf9fd5-68e4-40d2-972e-a5679640d9ca · outbound

This paper cites Zebra-cot: A dataset for interleaved vision language reasoning.arXiv preprint arXiv:2507.16746, 2025.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Zebra-cot: A dataset for interleaved vision language reasoning.arXiv preprint arXiv:2507.16746, 2025

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.088153Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.088153Z digest=sha256:9c182d6abb497d55aec2210915ddb92a5ea010adf049940e31d165f1dd8ade76

Observation 76d54afa-1dd1-45c7-bd61-5613d7a3532b · outbound

This paper cites Tir-bench: A comprehensive benchmark for agentic thinking-with-images reasoning.arXiv preprint arXiv:2511.01833, 2025.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Tir-bench: A comprehensive benchmark for agentic thinking-with-images reasoning.arXiv preprint arXiv:2511.01833, 2025

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.092965Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.092965Z digest=sha256:bb980da91e3a3f1c36c0574bc0deaeb88b842f774762451279e212e6ea495d95

Observation 6155944d-2c98-4ebf-ac10-b83d52418bbb · outbound

This paper cites Seg-Zero: Reasoning-Chain Guided Segmentation via Cognitive Reinforcement.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Seg-Zero: Reasoning-Chain Guided Segmentation via Cognitive Reinforcement

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.097596Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.097596Z digest=sha256:979a51fb9c05b8322abf6e6e06b43e792cc830ec8a1e2e8eef507d1f52bee805

Observation 8c4c7087-0bf3-4c1c-ba1f-f214947de793 · outbound

This paper cites On the faithfulness of visual thinking: Measurement and enhancement.arXiv preprint arXiv:2510.23482, 2025.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning On the faithfulness of visual thinking: Measurement and enhancement.arXiv preprint arXiv:2510.23482, 2025

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.102763Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.102763Z digest=sha256:d4ef4f11b299224146d831238c0c72190609c74cb944c9c59aa6f4772726593c

Observation fed3ace4-7f7a-4068-a766-df90f1678d4d · outbound

This paper cites Chain-of-Spot: Interactive Reasoning Improves Large Vision-Language Models.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Chain-of-Spot: Interactive Reasoning Improves Large Vision-Language Models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.107674Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.107674Z digest=sha256:532968fe34badbf9e7cfe503c3148c6104ec9c53b1740a7d2789387b39d7a65b

Observation 62e485d8-49de-4c3a-9605-ae3b6734b188 · outbound

This paper cites Chameleon: Plug-and-play compositional reasoning with large language models.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Chameleon: Plug-and-play compositional reasoning with large language models

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.113186Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.113186Z digest=sha256:cb6a789fffc4985d3e7abcb7360117673ba9823eb25850034cb58dee5c3e5989

Observation 9004a401-c7d2-49cc-ba01-b07b1b464f24 · outbound

This paper cites Reasoning Models Can Be Effective Without Thinking.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Reasoning Models Can Be Effective Without Thinking

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.118396Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.118396Z digest=sha256:901db4e91bbb88f48f084dbbe1cf71e6d6e216044d794f71fabbba191b6c9daf

Observation 87e7c981-a900-4a4f-9443-7f23e41f7676 · outbound

This paper cites What Does Vision Tool-Use Reinforcement Learning Really Learn? Disentangling Tool-Induced and Intrinsic Effects for Crop-and-Zoom.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning What Does Vision Tool-Use Reinforcement Learning Really Learn? Disentangling Tool-Induced and Intrinsic Effects for Crop-and-Zoom

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.125300Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.125300Z digest=sha256:303957964762b2a2ec97e082b8c11406cb83a36a6de4a00a12fc46ec4f9f16ae

Observation 8c9b6ca7-8c46-4542-b286-5c2210ec5976 · outbound

This paper cites ChartQA: A Benchmark for Question Answering about Charts with Visual and Logical Reasoning.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning ChartQA: A Benchmark for Question Answering about Charts with Visual and Logical Reasoning

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.131378Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.131378Z digest=sha256:749babc4cc14f13eed842f4ea8ee233f07521e9d3bd415a20e55098b6656c1fb

Observation d915d33d-2552-490f-b362-54b7e010fd32 · outbound

This paper cites GSM-Symbolic: Understanding the Limitations of Mathematical Reasoning in Large Language Models.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning GSM-Symbolic: Understanding the Limitations of Mathematical Reasoning in Large Language Models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.137144Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.137144Z digest=sha256:3fbd5fe27ada6277935c25e331f04cf7699b1d415a658b75382d87504ed6a6b0

Observation 38d5594e-04c8-45a8-9b98-56a44e4122d7 · outbound

This paper cites Show Your Work: Scratchpads for Intermediate Computation with Language Models.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Show Your Work: Scratchpads for Intermediate Computation with Language Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.143566Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.143566Z digest=sha256:8c08be943758eaf98ec67993ce7e08700f32c34aebdcc084a0043cb47a026bbd

Observation ab05b758-cb33-47e7-92ea-c69c2c1dc7d4 · outbound

This paper cites GPT-4 Technical Report.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning GPT-4 Technical Report

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.149920Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.149920Z digest=sha256:04b270f8e968e213c7eaf3854836e2a118fb2fe64f23dcde372b647e1fcde1ee

Observation 29b9e7f1-eedf-448c-beb2-59ffb461b6e5 · outbound

This paper cites Thinking with images.https://openai.com/index/thinking-with-images/, 2025.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Thinking with images.https://openai.com/index/thinking-with-images/, 2025

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.156357Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.156357Z digest=sha256:e56638826587a244fc2d6ac1a7b89924c7010b2aef0a49d0194837584e2ba9be

Observation 7a9f1ba0-82a1-4a24-ae45-3cfb98252cde · outbound

This paper cites Do MLLMs really see it: Reinforcing visual attention in multimodal LLMs.arXiv preprint arXiv:2602.08241, 2026.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Do MLLMs really see it: Reinforcing visual attention in multimodal LLMs.arXiv preprint arXiv:2602.08241, 2026

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.161693Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.161693Z digest=sha256:8c111447c8b65afa9712814404c6ad11a289c42919c63dd92cbbc7e6ae70ffab

Observation c9a368e4-f37a-4ffd-b1d6-20ad3b5f427f · outbound

This paper cites CogCoM: A Visual Language Model with Chain-of-Manipulations Reasoning.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning CogCoM: A Visual Language Model with Chain-of-Manipulations Reasoning

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.169132Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.169132Z digest=sha256:df0d50964843d46f972edd4fbae1ba8119f7639c69d8b1cbdde465d183f239f6

Observation d47f3c4d-f832-4c6b-a745-c38dbd5f852b · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.175515Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.175515Z digest=sha256:0dae1c6841d0b7428ac5b48d59e1b2f6f94af75778955fcc4549064de3aec776

Observation 25ad21de-9c81-4709-b03a-9f0f39708c29 · outbound

This paper cites Qwen2.5-VL Technical Report.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Qwen2.5-VL Technical Report

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.182177Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.182177Z digest=sha256:75ef10fb6293546652db58c43e3040a0d384001f10fdbb0a208533c95a943f3d

Observation 8c0aada3-d18d-4ebe-83af-b05f2c1bd462 · outbound

This paper cites Qwen3-VL Technical Report.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Qwen3-VL Technical Report

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.188571Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.188571Z digest=sha256:82f138ed94c52a66bfc5df0e25a99c942c7dcf78e306138a6b385ac4562bdb91

Observation fa21b7df-5558-4e33-8163-a62a77acc62f · outbound

This paper cites Vision language models are blind: Failing to translate detailed visual features into words.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Vision language models are blind: Failing to translate detailed visual features into words

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.194557Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.194557Z digest=sha256:53f5be1ec6e83b093ca200204b354ad307f46e41f740500f1a055afff373e542

Observation b553977f-9a03-4e10-9724-7816c539dd03 · outbound

This paper cites Grounded Reinforcement Learning for Visual Reasoning.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Grounded Reinforcement Learning for Visual Reasoning

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.200723Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.200723Z digest=sha256:012d02ead6a16e64c7872d52d560b27528d86d4d5ffac6421e7b51343a48aa49

Observation 71aed7c4-6d94-4238-a052-da220eee155a · outbound

This paper cites Toolformer: Language models can teach themselves to use tools.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Toolformer: Language models can teach themselves to use tools

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.206226Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.206226Z digest=sha256:9098ccd2e58239bd4c1663d73797704b703082edba2f621aeb75c4c6abdf7687

Observation eb554f4a-9642-4522-ac1f-13538d74085c · outbound

This paper cites Visual CoT: Advancing Multi-Modal Language Models with a Comprehensive Dataset and Benchmark for Chain-of-Thought Reasoning.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Visual CoT: Advancing Multi-Modal Language Models with a Comprehensive Dataset and Benchmark for Chain-of-Thought Reasoning

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.211374Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.211374Z digest=sha256:d9a0e03c0f6eef087a98ba6ba944fcbe089a22a185553006a19ced38a9958fbd

Observation 36606029-2231-4dbc-87b7-538e2926826a · outbound

This paper cites HuggingGPT: Solving AI tasks with ChatGPT and its friends in Hugging Face.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning HuggingGPT: Solving AI tasks with ChatGPT and its friends in Hugging Face

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.216761Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.216761Z digest=sha256:2020a059cf7493143af9ea132698b1e67add24672a8ad9bf91f1382817e08855

Observation d80610d9-2730-4bd8-bb41-1c6a246f93d8 · outbound

This paper cites HybridFlow: A Flexible and Efficient RLHF Framework.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning HybridFlow: A Flexible and Efficient RLHF Framework

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.221402Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.221402Z digest=sha256:6a634dce81d9aed275203173cbc4d3056f580f2cbc2f8c1ed7aee6d08749c5dd

Observation 7039400c-ff0f-4515-a00c-e38454fb4b39 · outbound

This paper cites Reflexion: Language agents with verbal reinforcement learning.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Reflexion: Language agents with verbal reinforcement learning

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.226408Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.226408Z digest=sha256:eb58420100ab5bbd9a13a75977ebdcc3e4a5c6abad9cdb834e4ec7db44fcecf8

Observation bd7094a6-99ef-4afb-b609-325ea8259745 · outbound

This paper cites Breaking the Chain: A Causal Analysis of LLM Faithfulness to Intermediate Structures.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Breaking the Chain: A Causal Analysis of LLM Faithfulness to Intermediate Structures

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.231434Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.231434Z digest=sha256:f253dc368d6c46d94ecc779232ad8f8c2d101214f2d8420c11de0e8e84293d45

Observation 2f895714-0304-4781-afac-8325cbee7e5e · outbound

This paper cites OpenThinkIMG: Learning to Think with Images via Visual Tool Reinforcement Learning.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning OpenThinkIMG: Learning to Think with Images via Visual Tool Reinforcement Learning

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.237048Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.237048Z digest=sha256:938c39597901a478a637d3f496693e646057c75bcfaf354bf67fac7699d63d79

Observation 6adf3fcd-a92f-4e84-9f6c-f35f8fdca13d · outbound

This paper cites Thinking with Images for Multimodal Reasoning: Foundations, Methods, and Future Frontiers.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Thinking with Images for Multimodal Reasoning: Foundations, Methods, and Future Frontiers

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.243692Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.243692Z digest=sha256:58f40732ffc0cba3916eeadedb38400db10139da4d88e928847bb7b89f207eb1

Observation 808134bd-c4aa-41ce-8ced-e241ff856bee · outbound

This paper cites When thinking hurts: Mitigating visual forgetting via frame repetition.arXiv preprint arXiv:2603.16256, 2026.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning When thinking hurts: Mitigating visual forgetting via frame repetition.arXiv preprint arXiv:2603.16256, 2026

Reference 56

Resolution
verified exact
raw_fallback, observed 2026-08-11T12:53:16.994849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T12:53:14.249540Z digest=sha256:1d161158a1ff8f58ad658dfa800ac49586dacf74d8dd04ea3e41c15d16579409

Observation 33cfd61d-327a-47bf-831c-7f2f998d7ef9 · outbound

This paper cites FACT-E: Causality-Inspired Evaluation for Trustworthy Chain-of-Thought Reasoning.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning FACT-E: Causality-Inspired Evaluation for Trustworthy Chain-of-Thought Reasoning

Reference 57

Resolution
verified exact
local_arxiv, observed 2026-08-11T12:53:16.902776Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T12:53:14.254925Z digest=sha256:7a1d742df0d5646d9ab5253d84d0fda1aa6952e26bbf7609491dfde587445fca

Observation a6faa050-844b-4cd5-9dc9-5a37b4bcd364 · outbound

This paper cites ViperGPT: Visual inference via python execution for reasoning.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning ViperGPT: Visual inference via python execution for reasoning

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.260904Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.260904Z digest=sha256:66eadb35d69d011e11d8ea36f05670558f94bb3a2645c43a2ea0f8784becac99

Observation c664eea2-5da5-491d-a904-af12e5d86968 · outbound

This paper cites CV-Bench: A computer vision benchmark for evaluating visual perception in multimodal language models.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning CV-Bench: A computer vision benchmark for evaluating visual perception in multimodal language models

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.266490Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.266490Z digest=sha256:95b3c72456743596092633c6c18da16d2af8f150b9e286792a8595ece5a164b8

Observation 984d86a0-d722-4cdd-a983-895defd786b0 · outbound

This paper cites an unresolved cited work.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Unresolved cited work

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.271924Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.271924Z digest=sha256:dc80f1a67c1701ebe33b629fbfead7e0dc2ac0043c2b9f20eb3070a038885319

Observation cb4caa7c-52de-4d80-8dba-fa757ff24771 · outbound

This paper cites Journey before destination: Visual faithfulness in slow thinking.arXiv preprint arXiv:2512.12218, 2025.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Journey before destination: Visual faithfulness in slow thinking.arXiv preprint arXiv:2512.12218, 2025

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.277132Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.277132Z digest=sha256:d5ad88701775282dc76ad2bcf0be5dd93e695bc8ad5060f5a2c379c5a10f4824

Observation aa0df45c-08c5-43d0-8474-d7483c923177 · outbound

This paper cites GeoEyes: On-demand visual focusing for ultra-high-resolution remote sensing.arXiv preprint arXiv:2602.14201, 2026.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning GeoEyes: On-demand visual focusing for ultra-high-resolution remote sensing.arXiv preprint arXiv:2602.14201, 2026

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.282612Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.282612Z digest=sha256:238ef40893b51c8f98240fcc83608f0c487ffdafd2b2f1caf04279b5fd7e1c35

Observation 0fcc6dce-7fe1-493b-953e-1838a6248d78 · outbound

This paper cites Pixel Reasoner: Incentivizing Pixel-Space Reasoning with Curiosity-Driven Reinforcement Learning.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Pixel Reasoner: Incentivizing Pixel-Space Reasoning with Curiosity-Driven Reinforcement Learning

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.289490Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.289490Z digest=sha256:e7624d20413e94ca6293a3acfa0e4ee5700c62f5e7d40df2992f9c08ad2cd252

Observation 431c9e90-1e5c-4855-8d17-1096829a3f45 · outbound

This paper cites PLaT: Latent chain-of-thought as planning: Decoupling reasoning from verbalization.arXiv preprint arXiv:2601.21358, 2026.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning PLaT: Latent chain-of-thought as planning: Decoupling reasoning from verbalization.arXiv preprint arXiv:2601.21358, 2026

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.295877Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.295877Z digest=sha256:efb98a2ae153004332ec3c565be06af61b35e65a80038234b82ce0f21e4d527a

Observation 13f7fe4e-f9e3-4b26-924d-51d354c8d659 · outbound

This paper cites VAGEN: Reinforcing world model reasoning for multi-turn VLM agents.arXiv preprint arXiv:2510.16907, 2025.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning VAGEN: Reinforcing world model reasoning for multi-turn VLM agents.arXiv preprint arXiv:2510.16907, 2025

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.301424Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.301424Z digest=sha256:bfb86644a677924248314ee9c077ad1c6efb9846dfa40f8078574b54bb24a1df

Observation 155234b5-f117-43e0-ac3f-685adee465c4 · outbound

This paper cites Plan-and-solve prompting: Improving zero-shot chain-of-thought reasoning by large language models.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Plan-and-solve prompting: Improving zero-shot chain-of-thought reasoning by large language models

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.307438Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.307438Z digest=sha256:79ebd51fcb73b49e3b28770d5f7cbc415b0d78df8de7c6604142260c453ec3f6

Observation ad4de20a-fa87-4e81-a3f1-935df8556f00 · outbound

This paper cites A practitioner’s guide to multi-turn agentic reinforcement learning.arXiv preprint arXiv:2510.01132, 2025.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning A practitioner’s guide to multi-turn agentic reinforcement learning.arXiv preprint arXiv:2510.01132, 2025

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.313080Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.313080Z digest=sha256:34900f6eeffa5a627e447b4148109beb74c549ad8ad000d704e22f3e776ed037

Observation 7e8456bd-7ef2-4fdb-86cb-b3b5d1e618a7 · outbound

This paper cites Divide, Conquer and Combine: A Training-Free Framework for High-Resolution Image Perception in Multimodal Large Language Models.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Divide, Conquer and Combine: A Training-Free Framework for High-Resolution Image Perception in Multimodal Large Language Models

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.318402Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.318402Z digest=sha256:57499db1a07ff1250acde824f91624d1496054c026453e7638403a4a89fb3bcf

Observation f574102a-4bd5-4cf0-aaab-eeb6bc7aa745 · outbound

This paper cites Self-consistency improves chain of thought reasoning in language models.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Self-consistency improves chain of thought reasoning in language models

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.324074Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.324074Z digest=sha256:c1a3c4cac8d3cc88a7ff3d56d3c170d4cd69aedbdda4ebc0a48f51fe2724909a

Observation 11876701-c04c-48cb-9602-4d5af4eb6f62 · outbound

This paper cites Simple o3: Towards Interleaved Vision-Language Reasoning.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Simple o3: Towards Interleaved Vision-Language Reasoning

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.330057Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.330057Z digest=sha256:25ccce06afcd962857c2ee7583413206e621417a15cb45d442a89b304d760c27

Observation 672af4c4-9a5f-40d2-abde-ad12cb180fd6 · outbound

This paper cites RAGEN: Understanding Self-Evolution in LLM Agents via Multi-Turn Reinforcement Learning.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning RAGEN: Understanding Self-Evolution in LLM Agents via Multi-Turn Reinforcement Learning

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.335916Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.335916Z digest=sha256:8dfab2ae92cb9d4a8ddf047ecd887317637bf741e2e7384c62ef6775c65c3797

Observation 58bc7e5f-585b-444a-b826-0f46909038fb · outbound

This paper cites CharXiv: Charting Gaps in Realistic Chart Understanding in Multimodal LLMs.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning CharXiv: Charting Gaps in Realistic Chart Understanding in Multimodal LLMs

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.341547Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.341547Z digest=sha256:5134a8490cf1995917455764d786501529d43e143ce8d8254f20ca00993e5c56

Observation f6834941-8cc1-470f-bd19-d04bb69d18b4 · outbound

This paper cites V-FAT: Benchmarking visual fidelity against text-bias.arXiv preprint arXiv:2601.04897, 2025.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning V-FAT: Benchmarking visual fidelity against text-bias.arXiv preprint arXiv:2601.04897, 2025

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.346906Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.346906Z digest=sha256:d96c69e74064775bb5fb9ad077f400f55a58d7f80f706b2fe254c703487a868d

Observation 75cb0d3c-04e6-46b5-a8e7-454ba0800c7b · outbound

This paper cites Chi, Quoc V.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Chi, Quoc V

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.352064Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.352064Z digest=sha256:5e5ccf925c845359fc3f6221e81c3eb506b9f2e6237d2e43c2ca2d5de2591d39

Observation 8614fd79-8787-43d5-83b9-21608da229f4 · outbound

This paper cites Zooming without zooming: Region- to-image distillation for fine-grained multimodal perception.arXiv preprint arXiv:2602.11858, 2026.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Zooming without zooming: Region- to-image distillation for fine-grained multimodal perception.arXiv preprint arXiv:2602.11858, 2026

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.357126Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.357126Z digest=sha256:941c1ee443480d7f7be06dc5002abf8ce91ac7a874ab610f8e2b29c0ab82be34

Observation 10131f61-7daf-46c1-82f6-f0f69697db60 · outbound

This paper cites Visual generation unlocks human-like reasoning through multimodal world models.arXiv preprint arXiv:2601.19834, 2026.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Visual generation unlocks human-like reasoning through multimodal world models.arXiv preprint arXiv:2601.19834, 2026

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.362216Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.362216Z digest=sha256:49ebeb91323a9c66768914db370a9e501ff5512b032bffc991eb4a4c15bd41b5

Observation 812332c9-4101-4395-ad7a-532e7f106164 · outbound

This paper cites VTool-R1: VLMs learn to think with images via reinforcement learning on multimodal tool use.arXiv preprint arXiv:2505.19255, 2025.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning VTool-R1: VLMs learn to think with images via reinforcement learning on multimodal tool use.arXiv preprint arXiv:2505.19255, 2025

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.367385Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.367385Z digest=sha256:93487e0d14a4f16f306675fd56d8b5f5c0cd5210bd32d7b97681421a1d59f26b

Observation c1358137-74bd-4884-b10f-9687c7390116 · outbound

This paper cites V*: Guided Visual Search as a Core Mechanism in Multimodal LLMs.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning V*: Guided Visual Search as a Core Mechanism in Multimodal LLMs

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.372654Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.372654Z digest=sha256:2afebbb14d6ab29ad2eb6f6fcdc8f78f1e1eee8fe881f3dfc83b620507add1ad

Observation 39fdb19d-0830-48e0-bb5b-af6e36d48e4a · outbound

This paper cites Tool-augmented policy optimization.arXiv preprint arXiv:2510.07038, 2025.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Tool-augmented policy optimization.arXiv preprint arXiv:2510.07038, 2025

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.378336Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.378336Z digest=sha256:dea9c4c9f04b4795994c3e6d8fc80abd3523d58c95383617262eebce74f1d077

Observation d2e8fa83-702f-44e8-bdb0-0b664623ccfd · outbound

This paper cites Reasoning or Reciting? Exploring the Capabilities and Limitations of Language Models Through Counterfactual Tasks.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Reasoning or Reciting? Exploring the Capabilities and Limitations of Language Models Through Counterfactual Tasks

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.384423Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.384423Z digest=sha256:ad829d59813c75bb03f785fb8dc41dbeaf6180bb8a2642257e48bf126bcd7a0d

Observation e0bf9a99-100f-4ea1-aa7d-15c558030ec7 · outbound

This paper cites LLaVA-CoT: Let Vision Language Models Reason Step-by-Step.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning LLaVA-CoT: Let Vision Language Models Reason Step-by-Step

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.390585Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.390585Z digest=sha256:916ada6535224117080a6d82745e26d60610d8aef718e48c06241cf979a12ca7

Observation 68cb5366-89ba-4258-9f1a-99d967af8c4c · outbound

This paper cites Act Wisely: Cultivating Meta-Cognitive Tool Use in Agentic Multimodal Models.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Act Wisely: Cultivating Meta-Cognitive Tool Use in Agentic Multimodal Models

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.396078Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.396078Z digest=sha256:f99f8f20b3b922a0a8663cbdc7a6d63ad22adbb9b12258db2327102780eef016

Observation 51316cec-d858-4612-af50-38963eb96922 · outbound

This paper cites Set-of-Mark Prompting Unleashes Extraordinary Visual Grounding in GPT-4V.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Set-of-Mark Prompting Unleashes Extraordinary Visual Grounding in GPT-4V

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.403481Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.403481Z digest=sha256:5d7c98c93c1f85838023168a5806c78273b80308dced1ad2020e3a45cd328a38

Observation c22ace32-3b90-4012-b5f8-9e466331462e · outbound

This paper cites Thinking in Space: How Multimodal Large Language Models See, Remember, and Recall Spaces.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Thinking in Space: How Multimodal Large Language Models See, Remember, and Recall Spaces

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.409420Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.409420Z digest=sha256:f0cd3e4db895f655aeceb2604859adc4955835de59e2875f07f7f8c8517f88c1

Observation 5e606f51-b262-42b5-8073-9e895eb919f4 · outbound

This paper cites VisionThink: Smart and Efficient Vision Language Model via Reinforcement Learning.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning VisionThink: Smart and Efficient Vision Language Model via Reinforcement Learning

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.415257Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.415257Z digest=sha256:085e23c893335c8c2401ddcec15e90a81d4de18e1fd9ceb7dc9b506c94398e37

Observation af5e4430-ad7c-4533-b666-a57b79c4825d · outbound

This paper cites Look-Back: Implicit Visual Re-focusing in MLLM Reasoning.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Look-Back: Implicit Visual Re-focusing in MLLM Reasoning

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.422649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.422649Z digest=sha256:457c56a42da9a7d3df29a8136824b4973472d42a503d6873ca0ecfa1cf3cd485

Observation c46a5d03-1cee-489d-98bd-52fb90f9d5b4 · outbound

This paper cites Walk the Talk: Bridging the Reasoning-Action Gap for Thinking with Images via Multimodal Agentic Policy Optimization.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Walk the Talk: Bridging the Reasoning-Action Gap for Thinking with Images via Multimodal Agentic Policy Optimization

Reference 87

Resolution
verified exact
local_arxiv, observed 2026-08-11T12:53:15.665161Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T12:53:14.428800Z digest=sha256:5bc1d5e77d5a5f53076098b0b8079aacd3874fc049ff7701f6a08bcc264ce1bd

Observation 6dc2fa7b-8836-4961-aa4b-e74d4e6e85d4 · outbound

This paper cites Thinking with images via self-calling agent.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Thinking with images via self-calling agent

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.434700Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.434700Z digest=sha256:b6856b3d412f3fe772ff3029bc525a4b2ae8e688af9648d1fef528cfe4130732

Observation ebbc4565-7328-45a9-9e9e-c80012b9984e · outbound

This paper cites Tree of thoughts: Deliberate problem solving with large language models.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Tree of thoughts: Deliberate problem solving with large language models

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.440324Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.440324Z digest=sha256:b8ce07653bb1e55e6bb55167d38e3aa4ff6728e385ec57620e2d59d557f709bd

Observation 6e03bf6e-d596-4191-833a-6f05b55e852a · outbound

This paper cites React: Synergizing reasoning and acting in language models.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning React: Synergizing reasoning and acting in language models

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.445645Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.445645Z digest=sha256:43929d5be6e42231e239e07fee3f508c5805d64e9ac7f0b6752a3617bcc04fed

Observation bb86f08f-a4d6-4b04-937e-b32d0aaac4c4 · outbound

This paper cites an unresolved cited work.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Unresolved cited work

Reference 91

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.450944Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.450944Z digest=sha256:576f1641414c5b08865ba5f29638533b9caf0480416ca5c252327754c25502d8

Observation 93488707-e50d-4f90-8a7d-95b8d82a412d · outbound

This paper cites ProRL agent: Rollout-as-a-service for reinforcement learning training of multi-turn LLM agents.arXiv preprint arXiv:2603.18815, 2026.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning ProRL agent: Rollout-as-a-service for reinforcement learning training of multi-turn LLM agents.arXiv preprint arXiv:2603.18815, 2026

Reference 92

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.456716Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.456716Z digest=sha256:e3ce6507379b96ab56072ddbc07d2227a37f3b6d29601993973cae843d23bf00

Observation 1da67c1b-6c4a-472f-93ea-6f3d8f059e62 · outbound

This paper cites MM-CoT: A benchmark for probing visual chain-of-thought reasoning.arXiv preprint arXiv:2512.08228, 2025.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning MM-CoT: A benchmark for probing visual chain-of-thought reasoning.arXiv preprint arXiv:2512.08228, 2025

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.462467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.462467Z digest=sha256:ca3d1ad9a551052487b4a7b8d7c28af86fd0149d5f18793be1d07efb4e97b310

Observation 1cde152f-f8de-45dd-8e0e-f00f4a94f582 · outbound

This paper cites LLaVA-Mini: Efficient Image and Video Large Multimodal Models with One Vision Token.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning LLaVA-Mini: Efficient Image and Video Large Multimodal Models with One Vision Token

Reference 94

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.468007Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.468007Z digest=sha256:c59473561761407af1e0e1833622f6f197bccb2f9032be7aed1fb642fe7ad03f

Observation ef2d5d0a-b1e9-435e-9982-f1221d152fad · outbound

This paper cites MME-RealWorld: Could Your Multimodal LLM Challenge High-Resolution Real-World Scenarios that are Difficult for Humans?.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning MME-RealWorld: Could Your Multimodal LLM Challenge High-Resolution Real-World Scenarios that are Difficult for Humans?

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.474606Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.474606Z digest=sha256:8d2efe79c56bf53f6ccad86881773531da649d831b95ccde76a88888bda392c2

Observation c12feb7b-3ba7-47e1-a0bc-5cb150c124f2 · outbound

This paper cites Thyme: Think Beyond Images.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Thyme: Think Beyond Images

Reference 96

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.480733Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.480733Z digest=sha256:95b000f0b1e849b9937e00a1a1d315b9f5555f9a74f7447955b6fa5388538889

Observation 4fb149dd-5d29-49f4-a574-6c1c48b07001 · outbound

This paper cites Skywork-r1v4: Toward agentic multimodal intelligence through interleaved thinking with images and deepresearch.arXiv preprint arXiv:2512.02395, 2025.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Skywork-r1v4: Toward agentic multimodal intelligence through interleaved thinking with images and deepresearch.arXiv preprint arXiv:2512.02395, 2025

Reference 97

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.486447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.486447Z digest=sha256:f8685dddd6316342c516883e66ae8f2d11a064fdd3329c453c05943853cc1c6a

Observation 3a376dd0-d944-496d-b3da-cffe00b77a20 · outbound

This paper cites CM2: Reinforcement learning with checklist rewards for multi-turn and multi-step agentic tool use.arXiv preprint arXiv:2602.12268, 2026.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning CM2: Reinforcement learning with checklist rewards for multi-turn and multi-step agentic tool use.arXiv preprint arXiv:2602.12268, 2026

Reference 98

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.491968Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.491968Z digest=sha256:ca376cf6366693b9334f028d61365e54b406e0e2821d2dc7a78a14b80e5df31d

Observation abebf7fd-87a9-4b5f-ad26-07e75afc7002 · outbound

This paper cites Multimodal Chain-of-Thought Reasoning in Language Models.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Multimodal Chain-of-Thought Reasoning in Language Models

Reference 99

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.497323Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.497323Z digest=sha256:f2a4cbc3b5c696a52ca1227784d0cb2671c82255325d3d3f0ac4e603611e14b0

Observation 754fb110-bb38-4730-85a6-4d1842ee219b · outbound

This paper cites On Robustness and Chain-of-Thought Consistency of RL-Finetuned VLMs.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning On Robustness and Chain-of-Thought Consistency of RL-Finetuned VLMs

Reference 100

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.503030Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.503030Z digest=sha256:bff806c93b2e686db2f4bebd5c300863e267d00d0190d1e81028525f57435544

Pith citing papers

No inbound Pith citation observations are available.