Pith. sign in

Paper Citation Record · LEDGER

Boosting the Generalization and Reasoning of Vision Language Models with Curriculum Reinforcement Learning

As of 12 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 27 inbound Pith citation observations for arXiv:2503.07065.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2503.07065 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 27 of 27 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 27 of 27 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:42:50.648770Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T15:09:54.970421Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation c68ae557-5d86-4d4b-a033-12e69baae016 · inbound

Multimodal Chain-of-Thought Reasoning: A Comprehensive Survey cites this paper.

Multimodal Chain-of-Thought Reasoning: A Comprehensive Survey Boosting the Generalization and Reasoning of Vision Language Models with Curriculum Reinforcement Learning

Reference 245

Resolution
verified exact
arxiv_id, observed 2026-05-15T17:18:53.556644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-15T17:18:52.996467Z digest=sha256:ed9c8cd3db067ef787b1a6e3ccf85b9f88de6958a8cc8d599b6b515260abce44

Observation 2c866938-5ec6-4e64-b313-5723a5276481 · inbound

VideoChat-R1: Enhancing Spatio-Temporal Perception via Reinforcement Fine-Tuning cites this paper.

VideoChat-R1: Enhancing Spatio-Temporal Perception via Reinforcement Fine-Tuning Boosting the Generalization and Reasoning of Vision Language Models with Curriculum Reinforcement Learning

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-15T20:56:07.673376Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-15T20:56:07.247122Z digest=sha256:acd33c4923419b0406b60c5f3de1e8330decbe1c7a36472a79c2982934517c29

Observation 87cfbd50-de8b-49ec-b945-027c72db7606 · inbound

VLM-R1: A Stable and Generalizable R1-style Large Vision-Language Model cites this paper.

VLM-R1: A Stable and Generalizable R1-style Large Vision-Language Model Boosting the Generalization and Reasoning of Vision Language Models with Curriculum Reinforcement Learning

Reference 14

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T01:13:57.417184Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-13T01:13:57.368874Z digest=sha256:83d881ea199e1e552979456e4b10fab8f713370d4123805d4bcf9ad043faafe3

Observation ef240fc9-f479-47ad-bb22-f648706214c8 · inbound

UniVG-R1: Reasoning Guided Universal Visual Grounding with Reinforcement Learning cites this paper.

UniVG-R1: Reasoning Guided Universal Visual Grounding with Reinforcement Learning Boosting the Generalization and Reasoning of Vision Language Models with Curriculum Reinforcement Learning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T15:42:50.648770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:42:50.648770Z digest=sha256:9b9b4aa217cd170729e8f69cd48c8776fa5ae8a1be40895d2c4d4c894a9dfde6

Observation 49a03a3e-854f-4b06-978e-666e8a225aa5 · inbound

Reinforcement Fine-Tuning Powers Reasoning Capability of Multimodal Large Language Models cites this paper.

Reinforcement Fine-Tuning Powers Reasoning Capability of Multimodal Large Language Models Boosting the Generalization and Reasoning of Vision Language Models with Curriculum Reinforcement Learning

Reference 121

Resolution
unresolved
no resolver link, observed 2026-08-07T14:31:19.391570Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:31:19.391570Z digest=sha256:ea8c01e8051d4ed9ef8cdb51fbe7da88cbc5c9e17073ea120ed52213e9ecba91

Observation c15369b9-d485-4683-a15b-39a08f08b63a · inbound

Align and Surpass Human Camouflaged Perception: Visual Refocus Reinforcement Fine-Tuning cites this paper.

Align and Surpass Human Camouflaged Perception: Visual Refocus Reinforcement Fine-Tuning Boosting the Generalization and Reasoning of Vision Language Models with Curriculum Reinforcement Learning

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T14:13:52.196940Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:13:52.196940Z digest=sha256:5a0b3d4ba93080bc3f4de0be00cd344cbf7cd0ba7d7b5cbfa50f48f6e4c5beb2

Observation 5fd999b5-d794-44e5-8d07-b494ae53ca2f · inbound

ZeroGUI: Automating Online GUI Learning at Zero Human Cost cites this paper.

ZeroGUI: Automating Online GUI Learning at Zero Human Cost Boosting the Generalization and Reasoning of Vision Language Models with Curriculum Reinforcement Learning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T12:41:21.033702Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:41:21.033702Z digest=sha256:440a81c94e73d6f2b0b6762f735d1e6a0ce5d23aaa873bbc9921d87bce138441

Observation 6348801c-8c22-4cb9-8943-fd8dbb1ab974 · inbound

WeThink: Toward General-purpose Vision-Language Reasoning via Reinforcement Learning cites this paper.

WeThink: Toward General-purpose Vision-Language Reasoning via Reinforcement Learning Boosting the Generalization and Reasoning of Vision Language Models with Curriculum Reinforcement Learning

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-07T05:27:05.136689Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:27:05.136689Z digest=sha256:3fe97a83262bf600672934fe956ca41670a145ac3b290aee28274a38e9c0635e

Observation 8a7a9780-6cb9-4c16-a4e5-72318cf4db33 · inbound

ChartReasoner: Code-Driven Modality Bridging for Long-Chain Reasoning in Chart Question Answering cites this paper.

ChartReasoner: Code-Driven Modality Bridging for Long-Chain Reasoning in Chart Question Answering Boosting the Generalization and Reasoning of Vision Language Models with Curriculum Reinforcement Learning

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-07T04:38:32.144672Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:38:32.144672Z digest=sha256:b3bfeef01efe846bbfc6d481a00225c7783e01fcef59a4ddc6a94b5d3dfe7a79

Observation 6acd7ba0-3709-48da-8847-e86b78f7681a · inbound

MM-R5: MultiModal Reasoning-Enhanced ReRanker via Reinforcement Learning for Document Retrieval cites this paper.

MM-R5: MultiModal Reasoning-Enhanced ReRanker via Reinforcement Learning for Document Retrieval Boosting the Generalization and Reasoning of Vision Language Models with Curriculum Reinforcement Learning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T00:57:01.730564Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T00:57:01.730564Z digest=sha256:5eaeb57695cbe0a2ecfd7641cb6bac984811725eff36fb04d7a6400239c0a1b7

Observation 0c20257e-3e3b-4776-aef8-ab908f7d19ba · inbound

Improving the Reasoning of Multi-Image Grounding in MLLMs via Reinforcement Learning cites this paper.

Improving the Reasoning of Multi-Image Grounding in MLLMs via Reinforcement Learning Boosting the Generalization and Reasoning of Vision Language Models with Curriculum Reinforcement Learning

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-05-19T06:52:07.977161Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-19T06:50:02.607136Z digest=sha256:1311a889c0f19a3d64be0111f73cf849251ba02e63a5073b4fc00847b575f42e

Observation 06c3a419-31f2-4e3d-8846-29766c49ee89 · inbound

SyncLoop: A Multimodal Dual-Loop Framework for Self-Improving Mathematical Reasoning cites this paper.

SyncLoop: A Multimodal Dual-Loop Framework for Self-Improving Mathematical Reasoning Boosting the Generalization and Reasoning of Vision Language Models with Curriculum Reinforcement Learning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T15:12:01.785071Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:12:01.785071Z digest=sha256:432ab469952e12f49484a7e6d2380b1695c2852ad54cdf71f0e8b71007125b5e

Observation f3a197cc-5b96-4ff5-ae41-bd22d2977949 · inbound

Empowering Multimodal LLMs with External Tools: A Comprehensive Survey cites this paper.

Empowering Multimodal LLMs with External Tools: A Comprehensive Survey Boosting the Generalization and Reasoning of Vision Language Models with Curriculum Reinforcement Learning

Reference 141

Resolution
unresolved
no resolver link, observed 2026-08-05T20:28:55.647921Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:28:55.647921Z digest=sha256:f02e12c7196feb32151db34e665948c5ef152c4de5ba81bc67c663a6e3d3e968

Observation 81da2cfc-2953-4d37-9707-0803e844cfca · inbound

Aligning Large Vision-Language Models by Deep Reinforcement Learning and Direct Preference Optimization cites this paper.

Aligning Large Vision-Language Models by Deep Reinforcement Learning and Direct Preference Optimization Boosting the Generalization and Reasoning of Vision Language Models with Curriculum Reinforcement Learning

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-04T23:11:33.680238Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T23:11:33.680238Z digest=sha256:ec5b46ea71368a4c8c7b60a7ac1091828e3572d12605b72effd04cdd7dd37189

Observation 994cb6c9-e3dc-4725-8a93-657b246f0fd7 · inbound

Reinforcement Learning Meets Large Language Models: A Survey of Advancements and Applications Across the LLM Lifecycle cites this paper.

Reinforcement Learning Meets Large Language Models: A Survey of Advancements and Applications Across the LLM Lifecycle Boosting the Generalization and Reasoning of Vision Language Models with Curriculum Reinforcement Learning

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-04T16:07:28.711477Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T16:07:28.711477Z digest=sha256:e7060324682fa2f0ba0df95c702740269f934e23c52468a6ac46d3a89cb6e542

Observation 7fbcfb2d-1706-4683-aed3-2eb00a406eee · inbound

Rethinking Reward Signals in Video GRPO: When Scores Become Targets cites this paper.

Rethinking Reward Signals in Video GRPO: When Scores Become Targets Boosting the Generalization and Reasoning of Vision Language Models with Curriculum Reinforcement Learning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-03T20:36:05.524886Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T20:36:05.524886Z digest=sha256:8dfa0602c5418e890021800f452d604c6dd4b791ac8ca27a74e201bbcc3aa42c

Observation be8d7fde-2bbf-49a6-83d9-053626d88997 · inbound

RL-RIG: A Generative Spatial Reasoner via Intrinsic Reflection cites this paper.

RL-RIG: A Generative Spatial Reasoner via Intrinsic Reflection Boosting the Generalization and Reasoning of Vision Language Models with Curriculum Reinforcement Learning

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-15T20:36:35.334320Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-15T20:33:09.627731Z digest=sha256:618e53f05df46029cef09974072a7a6b3de881ffd7d1ade8b176c076015e3431

Observation b4429a10-cffb-411e-9217-1b9d4f2b86c7 · inbound

Curr-RLCER:Curriculum Reinforcement Learning For Coherence Explainable Recommendation cites this paper.

Curr-RLCER:Curriculum Reinforcement Learning For Coherence Explainable Recommendation Boosting the Generalization and Reasoning of Vision Language Models with Curriculum Reinforcement Learning

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:35:49.933305Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-10T19:43:11.384787Z digest=sha256:e7a3b554f3faba9ebc029ed61fe6d4bd41bfedf9d301a083e6dee2a7ae359bd6

Observation 4c59afd4-f3ee-41e1-bbb5-a5d8f8f33cc3 · inbound

S-GRPO: Unified Post-Training for Large Vision-Language Models cites this paper.

S-GRPO: Unified Post-Training for Large Vision-Language Models Boosting the Generalization and Reasoning of Vision Language Models with Curriculum Reinforcement Learning

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-10T08:22:37.495818Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-10T08:18:59.432486Z digest=sha256:4c925c99c7aaed8ab256e1f9b0c14de2222edab3978de88cc233e1afadcc2951

Observation 80d34535-592e-4bde-815e-ce842d8991ec · inbound

S-GRPO: Unified Post-Training for Large Vision-Language Models cites this paper.

S-GRPO: Unified Post-Training for Large Vision-Language Models Boosting the Generalization and Reasoning of Vision Language Models with Curriculum Reinforcement Learning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-02T16:10:54.025142Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T16:10:54.025142Z digest=sha256:bfd395ce12f47da2361bbafd07594b03e62a81d606d597139bc5df28defcf1be

Observation d1b1a740-f640-4f1f-9599-3afc69605ee2 · inbound

DARE: Difficulty-Adaptive Reinforcement Learning with Co-Evolved Difficulty Estimation cites this paper.

DARE: Difficulty-Adaptive Reinforcement Learning with Co-Evolved Difficulty Estimation Boosting the Generalization and Reasoning of Vision Language Models with Curriculum Reinforcement Learning

Reference 10

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T03:16:19.353516Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-12T03:12:49.428954Z digest=sha256:1b93c9a283d71c59a2f27a0c894067a4ec9af5a550dd6edf5acda714c7df1e05

Observation e1d6424c-6e84-42cb-abdf-bcddad31fd48 · inbound

RadThinking: A Dataset for Longitudinal Clinical Reasoning in Radiology cites this paper.

RadThinking: A Dataset for Longitudinal Clinical Reasoning in Radiology Boosting the Generalization and Reasoning of Vision Language Models with Curriculum Reinforcement Learning

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-05-12T03:56:21.641871Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-05-12T03:55:55.359488Z digest=sha256:c6e11ec9c2c8b0375ab1b8673a8d093f2219365e47d0ba7c113e8db8ae94dc13

Observation 22522478-2499-4884-aed2-626270644382 · inbound

D$^2$Evo: Dual Difficulty-Aware Self-Evolution for Data-Efficient Reinforcement Learning cites this paper.

D$^2$Evo: Dual Difficulty-Aware Self-Evolution for Data-Efficient Reinforcement Learning Boosting the Generalization and Reasoning of Vision Language Models with Curriculum Reinforcement Learning

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-05-19T20:22:45.283893Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-05-19T20:21:48.657926Z digest=sha256:4293e5fa6b3e71ea9beac576d8137d6bee1e95bf034b226b137b1482debc46ce

Observation 7cb9933f-e0ff-4e72-9b76-03f0233c8ed6 · inbound

Towards One-to-Many Temporal Grounding cites this paper.

Towards One-to-Many Temporal Grounding Boosting the Generalization and Reasoning of Vision Language Models with Curriculum Reinforcement Learning

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-07-02T12:16:57.688274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-06-28T02:11:48.455492Z digest=sha256:e644dcf49b61dc90a0dc1868f2288ede73f970ca8d42cd022a3a49c34e5708f8

Observation e7c2e331-f3c0-4d96-ab6b-e6e9a403e464 · inbound

Stage-1 Controls the Entropy Regime, Not the Outcome cites this paper.

Stage-1 Controls the Entropy Regime, Not the Outcome Boosting the Generalization and Reasoning of Vision Language Models with Curriculum Reinforcement Learning

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-07-03T00:17:29.710973Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-27T17:16:39.814082Z digest=sha256:19ba497c1d4572b682dd32ac14b734ec4ea3014511e057e5909f6a15192801d7

Observation efe7a114-88d6-46da-9f60-3ac4fa777060 · inbound

Dense Reward for Multi-View 3D Reasoning with Global Maps and Local Views cites this paper.

Dense Reward for Multi-View 3D Reasoning with Global Maps and Local Views Boosting the Generalization and Reasoning of Vision Language Models with Curriculum Reinforcement Learning

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-07-04T10:39:45.368196Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-26T08:38:46.044079Z digest=sha256:7805171996d35a51322e6f31a97e234e9e4b00c461e3fb44507ba76a7b3d6657

Observation f5baae0b-b052-45ef-9647-37a64218d44a · inbound

From Structure to Synergy: A Survey of Vision-Language Perception Paradigm Evolution in Multimodal Large Language Models cites this paper.

From Structure to Synergy: A Survey of Vision-Language Perception Paradigm Evolution in Multimodal Large Language Models Boosting the Generalization and Reasoning of Vision Language Models with Curriculum Reinforcement Learning

Reference 176

Resolution
verified exact
arxiv_id, observed 2026-07-04T15:09:54.971976Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-06-26T01:50:54.242508Z digest=sha256:b59145acac2a98038a437808c4099d28ae35c614b4bf152cf3aa024e2491a441