Pith. sign in

Paper Citation Record · LEDGER

Deeper Thought, Weaker Aim: Understanding and Mitigating Perceptual Impairment during Reasoning in Multimodal Large Language Models

As of 4 August 2026, this Paper Citation Record lists 30 of 30 outbound references and 0 inbound Pith citation observations for arXiv:2603.14184.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2603.14184 v2

Coverage vector

measured 30 of 30 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-21T11:56:29.743281Z

measured 30 of 30 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

30 of 30 outbound references displayed

  • verified exact13
  • verified fuzzy16
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c8ceeb52-d5a6-4cfe-b961-c23dec8313cb · outbound

This paper cites Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond.

Deeper Thought, Weaker Aim: Understanding and Mitigating Perceptual Impairment during Reasoning in Multimodal Large Language Models Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-21T12:00:04.307365Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-21T11:56:29.743281Z digest=sha256:2240c80cfed30989450e73484a2051789b9a3837a09bce05dc7be5db852d9328

Observation 25e5159c-2406-4143-a096-dc8bf4e2d87e · outbound

This paper cites Qwen2.5-VL Technical Report.

Deeper Thought, Weaker Aim: Understanding and Mitigating Perceptual Impairment during Reasoning in Multimodal Large Language Models Qwen2.5-VL Technical Report

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-05-21T12:00:04.283277Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-21T11:56:29.743281Z digest=sha256:f57e4b44d7ce1e37957c9f5e223f146876130c19f97278aee145a4d31b1add3b

Observation a88192c8-3607-41f2-b797-1527c5301047 · outbound

This paper cites Ddot: A derivative-directed dual-decoder ordinary differential equation transformer for dynamic system modeling.

Deeper Thought, Weaker Aim: Understanding and Mitigating Perceptual Impairment during Reasoning in Multimodal Large Language Models Ddot: A derivative-directed dual-decoder ordinary differential equation transformer for dynamic system modeling

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T12:00:04.901809Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-21T11:56:29.743281Z digest=sha256:848217214cf4a9ad17f13f7568bf3e46ac229633e64336640913bb3f785cba37

Observation 0b2daca1-8bb6-4bf6-9794-10325ce6e90e · outbound

This paper cites Are we on the right way for evaluating large vision-language models?Advances in Neural Informa- tion Processing Systems, 37:27056–27087.

Deeper Thought, Weaker Aim: Understanding and Mitigating Perceptual Impairment during Reasoning in Multimodal Large Language Models Are we on the right way for evaluating large vision-language models?Advances in Neural Informa- tion Processing Systems, 37:27056–27087

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T12:00:04.926683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-21T11:56:29.743281Z digest=sha256:c530700a5bc40920fcf06acff91a0694ddf452359774f1fd1156e13c6c435c54

Observation 5f949651-fe6d-40d7-877a-78a09810b314 · outbound

This paper cites Interleaved-modal chain-of-thought.

Deeper Thought, Weaker Aim: Understanding and Mitigating Perceptual Impairment during Reasoning in Multimodal Large Language Models Interleaved-modal chain-of-thought

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T12:00:04.899743Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-21T11:56:29.743281Z digest=sha256:69c4f35b8dcac671dce0ee1b4ed56a04c6a0a3549e300809b01807f24b8f5d28

Observation d2ee727d-9616-4a24-a027-7d2d37b25fe7 · outbound

This paper cites Gemini: Our most capable multi- modal ai model yet.https://deepmind.google/ technologies/gemini/.

Deeper Thought, Weaker Aim: Understanding and Mitigating Perceptual Impairment during Reasoning in Multimodal Large Language Models Gemini: Our most capable multi- modal ai model yet.https://deepmind.google/ technologies/gemini/

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T12:00:04.890661Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-21T11:56:29.743281Z digest=sha256:1dec98ced98ebdae449ccfce8f65d43477a36e706eb57445310f68479bd2eeb2

Observation ea006f23-8a73-4f8c-9c7e-4a473ee9ae13 · outbound

This paper cites Hallusionbench: an advanced diagnos- tic suite for entangled language hallucination and visual il- lusion in large vision-language models.

Deeper Thought, Weaker Aim: Understanding and Mitigating Perceptual Impairment during Reasoning in Multimodal Large Language Models Hallusionbench: an advanced diagnos- tic suite for entangled language hallucination and visual il- lusion in large vision-language models

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T12:00:04.895058Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-21T11:56:29.743281Z digest=sha256:85c80828dd053663c32e46bbb217c8845a8212f67cf5afa8fc2b1821eaeaec96

Observation 019e877f-d788-46fd-a323-8e61986d8f3c · outbound

This paper cites See What You Are Told: Visual Attention Sink in Large Multimodal Models.

Deeper Thought, Weaker Aim: Understanding and Mitigating Perceptual Impairment during Reasoning in Multimodal Large Language Models See What You Are Told: Visual Attention Sink in Large Multimodal Models

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-21T12:00:04.294537Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-21T11:56:29.743281Z digest=sha256:bdf31a1fb5166a4830e72356ef5387a4f9c772dfdbff8842430692cfbd5fb7b1

Observation 1ecc98c0-b5b0-4f55-9e15-3c83e705dec0 · outbound

This paper cites Blip: Bootstrapping language-image pre-training for unified vision-language understanding and generation.

Deeper Thought, Weaker Aim: Understanding and Mitigating Perceptual Impairment during Reasoning in Multimodal Large Language Models Blip: Bootstrapping language-image pre-training for unified vision-language understanding and generation

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T12:00:04.897202Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-21T11:56:29.743281Z digest=sha256:01a7042e2f950839ef815d2cfd66233c8ff52adf1208108680457f15b85bffc2

Observation 873458e3-a56d-4bd4-a4d4-ed0ab4552919 · outbound

This paper cites Improved Visual-Spatial Reasoning via R1-Zero-Like Training.

Deeper Thought, Weaker Aim: Understanding and Mitigating Perceptual Impairment during Reasoning in Multimodal Large Language Models Improved Visual-Spatial Reasoning via R1-Zero-Like Training

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-21T12:00:04.278492Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-21T11:56:29.743281Z digest=sha256:03922fcf0fb483919212a151589a5a199c12f0088d2fcac8f1930d9eeb83f88a

Observation d60cb9bd-3110-4c39-a443-a367a20b05ae · outbound

This paper cites Let’s verify step by step.

Deeper Thought, Weaker Aim: Understanding and Mitigating Perceptual Impairment during Reasoning in Multimodal Large Language Models Let’s verify step by step

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T12:00:04.904787Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-21T11:56:29.743281Z digest=sha256:7bac50ba44269f3851388d57135aaf6e499a3a103c0c22b9f0ea207ee7918dfc

Observation 38f12b71-985f-4a35-82ab-9fae04cde827 · outbound

This paper cites Program Induction by Rationale Generation : Learning to Solve and Explain Algebraic Word Problems.

Deeper Thought, Weaker Aim: Understanding and Mitigating Perceptual Impairment during Reasoning in Multimodal Large Language Models Program Induction by Rationale Generation : Learning to Solve and Explain Algebraic Word Problems

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-05-21T12:00:04.303572Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-21T11:56:29.743281Z digest=sha256:8e5eddc64f077b083f29da67aa7f29bbe2501c948b9ddf8df92fbe27f4f02914

Observation b4ae380b-8e2d-4288-a5e7-958be10e4c51 · outbound

This paper cites More Thinking, Less Seeing? Assessing Amplified Hallucination in Multimodal Reasoning Models.

Deeper Thought, Weaker Aim: Understanding and Mitigating Perceptual Impairment during Reasoning in Multimodal Large Language Models More Thinking, Less Seeing? Assessing Amplified Hallucination in Multimodal Reasoning Models

Reference 13

Resolution
metadata mismatch
arxiv_id, observed 2026-05-21T12:00:04.289571Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-21T11:56:29.743281Z digest=sha256:6c5ccfa96b865ba137c2e42221d14db1d8c37e6c0c0f63a7bd3c6fcd78c30ff7

Observation 53512ddb-d9ff-4a21-b7c9-634a86f55c66 · outbound

This paper cites Visual instruction tuning.Advances in neural information processing systems, 36:34892–34916.

Deeper Thought, Weaker Aim: Understanding and Mitigating Perceptual Impairment during Reasoning in Multimodal Large Language Models Visual instruction tuning.Advances in neural information processing systems, 36:34892–34916

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T12:00:04.892804Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-21T11:56:29.743281Z digest=sha256:94ce03dc21bedad3ffc210af4e4190c23dea85dd98374d2d6cd9122f7ec1a0ab

Observation 5e6842ff-7553-47ff-a60f-94d630df5961 · outbound

This paper cites SCP-116K: A High-Quality Problem-Solution Dataset and a Generalized Pipeline for Automated Extraction in the Higher Education Science Domain.

Deeper Thought, Weaker Aim: Understanding and Mitigating Perceptual Impairment during Reasoning in Multimodal Large Language Models SCP-116K: A High-Quality Problem-Solution Dataset and a Generalized Pipeline for Automated Extraction in the Higher Education Science Domain

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-21T12:00:04.299034Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-21T11:56:29.743281Z digest=sha256:efb2a619b87d3631bcb3094e6c50ccb0542b2ca80fdaf1251f274efad556c780

Observation 5b3334e5-1a7f-4ebc-9ef3-b8be329e7342 · outbound

This paper cites Learn to explain: Multimodal reasoning via thought chains for science question answering.Advances in Neural Information Processing Systems, 35:2507–2521.

Deeper Thought, Weaker Aim: Understanding and Mitigating Perceptual Impairment during Reasoning in Multimodal Large Language Models Learn to explain: Multimodal reasoning via thought chains for science question answering.Advances in Neural Information Processing Systems, 35:2507–2521

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T12:00:04.886741Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-21T11:56:29.743281Z digest=sha256:e184e43e45b54e4d488084c00307c1cfdbc04112b3c232464a8476670ef73295

Observation 4338ff4a-26a0-4ecd-871b-088210bc443c · outbound

This paper cites MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning.

Deeper Thought, Weaker Aim: Understanding and Mitigating Perceptual Impairment during Reasoning in Multimodal Large Language Models MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-05-21T12:00:04.245706Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-21T11:56:29.743281Z digest=sha256:7d03784671e3e2521aa49f8ccb34e5865e6bcbd416d0afe0f1229cabb240af2d

Observation 517b1b64-d8cb-443b-891b-bc9a2966cc34 · outbound

This paper cites Ocean-r1: An open and generaliz- able large vision-language model enhanced by reinforcement learning.

Deeper Thought, Weaker Aim: Understanding and Mitigating Perceptual Impairment during Reasoning in Multimodal Large Language Models Ocean-r1: An open and generaliz- able large vision-language model enhanced by reinforcement learning

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T12:00:04.882099Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-21T11:56:29.743281Z digest=sha256:a6c77cb5d01943b599ab184b27b9dc653e5dec118ad51d12a46091c8216cceb1

Observation 1ead2075-e5c3-4043-b642-c902b3f78553 · outbound

This paper cites Compositional chain-of-thought prompting for large multimodal models.

Deeper Thought, Weaker Aim: Understanding and Mitigating Perceptual Impairment during Reasoning in Multimodal Large Language Models Compositional chain-of-thought prompting for large multimodal models

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T12:00:04.877763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-21T11:56:29.743281Z digest=sha256:a6f2d6b12f748e5669232bf1e569789fd36f388170418979a7d2b137c4e46c84

Observation 9c7e8226-6581-465e-ad1e-01ee43312b87 · outbound

This paper cites Towards vqa models that can read.

Deeper Thought, Weaker Aim: Understanding and Mitigating Perceptual Impairment during Reasoning in Multimodal Large Language Models Towards vqa models that can read

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T12:00:04.888729Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-21T11:56:29.743281Z digest=sha256:b749aceb8dd77b50244dc56730590413d794f27ac3aceb519355f9df23eb4ee8

Observation 1cf2d3f2-db35-4d2b-abf2-004867d2b9c4 · outbound

This paper cites Aligning Large Multimodal Models with Factually Augmented RLHF.

Deeper Thought, Weaker Aim: Understanding and Mitigating Perceptual Impairment during Reasoning in Multimodal Large Language Models Aligning Large Multimodal Models with Factually Augmented RLHF

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-05-21T12:00:04.255052Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-21T11:56:29.743281Z digest=sha256:1e2596fbc1274b6c711d86e8745a2c33ce4185c39a093927cace40a7a4815f11

Observation 2e9dbaba-656e-401c-89f5-05acc341f068 · outbound

This paper cites Qwen3 technical report.

Deeper Thought, Weaker Aim: Understanding and Mitigating Perceptual Impairment during Reasoning in Multimodal Large Language Models Qwen3 technical report

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T12:00:04.875868Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-21T11:56:29.743281Z digest=sha256:f83ae3c0d41af826384438fc39983be631e0566a5bb43b91016b3a0ec8410629

Observation 700af3fe-858e-4958-acda-e4c4b7113f97 · outbound

This paper cites Mea- suring multimodal mathematical reasoning with math-vision dataset.Advances in Neural Information Processing Sys- tems, 37:95095–95169.

Deeper Thought, Weaker Aim: Understanding and Mitigating Perceptual Impairment during Reasoning in Multimodal Large Language Models Mea- suring multimodal mathematical reasoning with math-vision dataset.Advances in Neural Information Processing Sys- tems, 37:95095–95169

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T12:00:04.880086Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-21T11:56:29.743281Z digest=sha256:d361cef148adb216ccdbcdd2adaf6288ec5d84b76bb480b92e271ce233bba3d4

Observation c784c8ef-ae77-401b-bc6e-512f950c0994 · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

Deeper Thought, Weaker Aim: Understanding and Mitigating Perceptual Impairment during Reasoning in Multimodal Large Language Models Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-05-21T12:00:04.268445Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-21T11:56:29.743281Z digest=sha256:b3f3d61b5d651857d991ffaac6230b69e5a14034663840768d717334559eec53

Observation c71b368f-6819-4534-9280-31d87dd2d01d · outbound

This paper cites SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement.

Deeper Thought, Weaker Aim: Understanding and Mitigating Perceptual Impairment during Reasoning in Multimodal Large Language Models SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-21T12:00:04.251113Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-21T11:56:29.743281Z digest=sha256:6dceda2aceb33b7ebecf3e6abdec3e2bde5686b4d222e1e727d6198d71ad9a7e

Observation 681abfc9-ffaf-4b59-83bd-fc09f848cde4 · outbound

This paper cites Haloquest: A visual hallucination dataset for advancing multimodal reasoning.

Deeper Thought, Weaker Aim: Understanding and Mitigating Perceptual Impairment during Reasoning in Multimodal Large Language Models Haloquest: A visual hallucination dataset for advancing multimodal reasoning

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T12:00:04.874020Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-21T11:56:29.743281Z digest=sha256:3d4550d50b7a6acfdabfff11bf5640b4a6655f4b5d84c5f25019ccd69646ea22

Observation 1e13eb74-28a8-4020-a17c-808c9eb0b88d · outbound

This paper cites Thinking in space: How mul- timodal large language models see, remember, and recall spaces.

Deeper Thought, Weaker Aim: Understanding and Mitigating Perceptual Impairment during Reasoning in Multimodal Large Language Models Thinking in space: How mul- timodal large language models see, remember, and recall spaces

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T12:00:04.884370Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-21T11:56:29.743281Z digest=sha256:5d6a9c22aee96e96b6be2ccae676c964727ba8bbcc338b3f76c687ff66f00ea3

Observation 022b4dd0-4592-4df5-b411-62449c6d8592 · outbound

This paper cites Gsm8k-v: Can vision language models solve grade school math word problems in visual contexts.

Deeper Thought, Weaker Aim: Understanding and Mitigating Perceptual Impairment during Reasoning in Multimodal Large Language Models Gsm8k-v: Can vision language models solve grade school math word problems in visual contexts

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-21T12:00:04.263872Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-21T11:56:29.743281Z digest=sha256:261eaee3eaa8f35f9daaad86e4b5615d6ae09a55882e78817e547f58519400e2

Observation 78da70f9-51ff-4322-b493-1722d99f08da · outbound

This paper cites MLLMs Know Where to Look: Training-free Perception of Small Visual Details with Multimodal LLMs.

Deeper Thought, Weaker Aim: Understanding and Mitigating Perceptual Impairment during Reasoning in Multimodal Large Language Models MLLMs Know Where to Look: Training-free Perception of Small Visual Details with Multimodal LLMs

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-21T12:00:04.273457Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-21T11:56:29.743281Z digest=sha256:7dcc2c93debdd45817710b5427a25f3ef89b07e78fd42e8c74568840092426a4

Observation 35150de5-752a-4584-9fb9-26114febec26 · outbound

This paper cites DeepEyes: Incentivizing "Thinking with Images" via Reinforcement Learning.

Deeper Thought, Weaker Aim: Understanding and Mitigating Perceptual Impairment during Reasoning in Multimodal Large Language Models DeepEyes: Incentivizing "Thinking with Images" via Reinforcement Learning

Reference 31

Resolution
verified exact
local_arxiv, observed 2026-05-21T12:00:04.259058Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-21T11:56:29.743281Z digest=sha256:52617b67012c9150e23b3f50372a78d39739aea0361e8e223e4b0e49e8da4afe

Pith citing papers

No inbound Pith citation observations are available.