Pith. sign in

Paper Citation Record · LEDGER

Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models

As of 8 August 2026, this Paper Citation Record lists 35 of 35 outbound references and 4 inbound Pith citation observations for arXiv:2505.23091.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.23091 v3

Coverage vector

measured 35 of 35 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:59:23.141014Z

measured 39 of 39 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-03T16:53:43.427625Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T10:39:46.064161Z

Reference resolution

35 of 35 outbound references displayed

  • verified exact0
  • verified fuzzy5
  • unresolved30
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 44d759aa-6212-40d0-bb87-aa29433b17ce · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:19.315667Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:19.315667Z digest=sha256:159768e97035f93af8948d2b39c8644fac636799f46c34d4241aa0f57fd94b30

Observation 65a29c40-e7d4-4c9a-b8de-0221cfd36bf8 · outbound

This paper cites Qwen2 Technical Report.

Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models Qwen2 Technical Report

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:19.431681Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:19.431681Z digest=sha256:29eb4b3c897f54314e26a470cd33410fae3753afb4c34439e64b291377a32568

Observation 9cb0006b-35d7-4660-ba07-8f10ad333518 · outbound

This paper cites OpenAI o1 System Card.

Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models OpenAI o1 System Card

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:19.523089Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:19.523089Z digest=sha256:01e89e79c1e6f2f70a62ddfba84a864bcd379334d4cfbbb8693ce735c4378d72

Observation c01072d5-3cee-4766-bfb9-26bbf4b826a7 · outbound

This paper cites Step-DPO: Step-wise Preference Optimization for Long-chain Reasoning of LLMs.

Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models Step-DPO: Step-wise Preference Optimization for Long-chain Reasoning of LLMs

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:19.624700Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:19.624700Z digest=sha256:7a45a36df2c8c85fb7cc545f412a512e9a8ebd4314f6cffddd1c2b968b6f8d36

Observation 0c95002a-6c4f-4b32-9bdb-c7906314ef78 · outbound

This paper cites Kosaraju, Y.

Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models Kosaraju, Y

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:25.681023Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:59:19.764183Z digest=sha256:b504c9c69c3825311d2b440c209810df8d00f25613b7d0ee210f7be3a2f4f24f

Observation 19d83683-c19d-4022-ad09-68fcb74e33fb · outbound

This paper cites an unresolved cited work.

Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:59:25.531225Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:59:19.906443Z digest=sha256:34e53db959032d0b989454a60ce0057ab8df71ae6e673419fa8e35e8684390c9

Observation 7711fb21-8c83-49c6-8593-de6873b052a7 · outbound

This paper cites an unresolved cited work.

Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:59:25.393970Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:59:20.054294Z digest=sha256:fdd6b0bbcb1f54119a7afc31c319b68471f328dfeeb5d03af794a52fa10beccc

Observation 042b01ec-2aec-4b7e-b9d0-fc7cabbb8249 · outbound

This paper cites an unresolved cited work.

Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models Unresolved cited work

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:20.185816Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:20.185816Z digest=sha256:382a635faf4b5a55587fbc067c64202f4c35a9e5999f9f481680eb86c6a8e99b

Observation 39c6486a-1a54-499d-a35c-aff0c2dd67a6 · outbound

This paper cites Vision-R1: Incentivizing Reasoning Capability in Multimodal Large Language Models.

Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models Vision-R1: Incentivizing Reasoning Capability in Multimodal Large Language Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:20.280238Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:20.280238Z digest=sha256:1ab0b822c8a7c6539312b3da5da8d841a793a0922cc79ff90d42debee191a641

Observation dc87b2d7-493c-4505-8140-ad931d9fce0b · outbound

This paper cites Donahue, P.

Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models Donahue, P

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:25.277908Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:59:20.410033Z digest=sha256:bbdd006918be96f0170014c0f713f17ab2f3d6e8a88c4b19b538f7887cd6b4cf

Observation 654ae170-2358-4d99-874b-ea1b6f4f9f42 · outbound

This paper cites an unresolved cited work.

Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:59:25.183689Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:59:20.492336Z digest=sha256:585dd1d3c0cbc034c8021e64c8137234d6028991050423801bd5e20d3fa018c8

Observation bbc362b5-e8fe-4c70-8155-8d80da3ff91d · outbound

This paper cites Improve Vision Language Model Chain-of-thought Reasoning.

Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models Improve Vision Language Model Chain-of-thought Reasoning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:20.586668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:20.586668Z digest=sha256:9e65834c5617149ba8dd837e613fa08f6d18085ae1d4028b1e3829764964540d

Observation 55887e46-2bf9-4415-9ab0-450044554533 · outbound

This paper cites an unresolved cited work.

Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:59:25.067268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:59:20.666995Z digest=sha256:8fc794c0932803d960e821cbff809e6d733b59e47f8be4180d7a7047f30197cf

Observation 31096d2e-7c27-41ac-a0ea-5011229d128f · outbound

This paper cites R1-Onevision: Advancing Generalized Multimodal Reasoning through Cross-Modal Formalization.

Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models R1-Onevision: Advancing Generalized Multimodal Reasoning through Cross-Modal Formalization

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:20.808202Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:20.808202Z digest=sha256:fc2a10898ccc879bd2f5f605f229035adf4498988f34412726c9b6e5e4e260d4

Observation 1f49ad59-bbbc-40c1-8dd9-2b7c2e7e9b40 · outbound

This paper cites X-Reasoner: Towards Generalizable Reasoning Across Modalities and Domains.

Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models X-Reasoner: Towards Generalizable Reasoning Across Modalities and Domains

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:20.896118Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:20.896118Z digest=sha256:8d1b8a6a019a710a2c88af5c2269134856dde36235ced8bdab24d181302783f6

Observation e4014433-8b02-47a4-84e6-9bccf0045322 · outbound

This paper cites LMM-R1: Empowering 3B LMMs with Strong Reasoning Abilities Through Two-Stage Rule-Based RL.

Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models LMM-R1: Empowering 3B LMMs with Strong Reasoning Abilities Through Two-Stage Rule-Based RL

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:20.994541Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:20.994541Z digest=sha256:f67b7e70f33742b8d529930e1146bed31895d8e8369ba09bd0ce3a1e1a629be7

Observation 5415c9ba-dfb6-4ca7-98f3-f57147b36ef6 · outbound

This paper cites an unresolved cited work.

Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:59:24.879439Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:59:21.107805Z digest=sha256:339c64809e1252dd05c5bc779e34795c2f5731de61c7f94387dd9fc1f7ac4e39

Observation be86a894-3054-417b-b73c-fc6fb900a8d4 · outbound

This paper cites Competence-based Curriculum Learning for Neural Machine Translation.

Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models Competence-based Curriculum Learning for Neural Machine Translation

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:21.225659Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:21.225659Z digest=sha256:cf8e54066eb8ce54ca64540ab8a1e28f3444f08688e6b7560cb590a8c27f7c1f

Observation 975416c1-4126-4765-a05e-f68e8aa722fc · outbound

This paper cites Proximal Policy Optimization Algorithms.

Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models Proximal Policy Optimization Algorithms

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:21.343445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:21.343445Z digest=sha256:0f0f548c0f124b86e4aa1ae3a7cb7a661bf06241771f7e3064d80afb3821de23

Observation dfb51c60-7593-4a3c-a539-9c0f8e795978 · outbound

This paper cites OmniCaptioner: One Captioner to Rule Them All.

Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models OmniCaptioner: One Captioner to Rule Them All

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:21.482694Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:21.482694Z digest=sha256:7a7ce2d9a61bfaa5132a5147f5f08b71589e5fa48c6f83ba6deb7538c5cb3559

Observation e0a63174-7431-42af-8c92-dd95f1880b22 · outbound

This paper cites Qwen2.5-VL Technical Report.

Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models Qwen2.5-VL Technical Report

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:21.585019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:21.585019Z digest=sha256:828c9a3b9b0eaad2b7a1355f1c7824ad61f445c68b80a5bf7614afee6496edb3

Observation 699b119a-0813-4c66-8920-4c58e0225a34 · outbound

This paper cites an unresolved cited work.

Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:59:24.706564Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:59:21.742628Z digest=sha256:aa6472ff791d6ea64fa72b72b708e91f77533bf3bfaca6decafd00591d02c538

Observation c93cfb6a-aad6-46df-8592-84352ab045ac · outbound

This paper cites VL-Rethinker: Incentivizing Self-Reflection of Vision-Language Models with Reinforcement Learning.

Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models VL-Rethinker: Incentivizing Self-Reflection of Vision-Language Models with Reinforcement Learning

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:21.841015Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:21.841015Z digest=sha256:53715fd956bdeda3f0bef1871da1ee431b07c7c358b856ecb3726ea51b97118c

Observation 40d1d414-ca7a-4e9a-8d5a-2ee63473299f · outbound

This paper cites Light-R1: Curriculum SFT, DPO and RL for Long COT from Scratch and Beyond.

Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models Light-R1: Curriculum SFT, DPO and RL for Long COT from Scratch and Beyond

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:21.956041Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:21.956041Z digest=sha256:5fec4044c268b50e2b1b97ca60735073b1fed2373db0df03db57b76e35f307de

Observation deb10754-20f1-4f67-b05d-d636b3530df1 · outbound

This paper cites Zhang, W.

Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models Zhang, W

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:24.519630Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:59:22.076981Z digest=sha256:2e1e26f7ca3ecc713deee2d4f64f929f1ad20ec9449eb53d8307d788f9a2f5da

Observation 97be1d3a-03fc-41f3-a6e9-d4326a340a24 · outbound

This paper cites Jiang, Y.

Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models Jiang, Y

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:24.272710Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:59:22.208111Z digest=sha256:27cb2194a1d575d0a7d08fb9d325ffe1c7a7ecb442f200129bb89c2f5c26d896

Observation 83d0c1a8-6092-48b4-858e-ac0f90b51c9e · outbound

This paper cites an unresolved cited work.

Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models Unresolved cited work

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:22.296769Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:22.296769Z digest=sha256:1b3e4610dc7ab2bd19fe315039987744bbe77cf33177fe50e0bdb34b2c8ea998

Observation 8f977acf-2f75-4a43-af1b-84457d8b858c · outbound

This paper cites an unresolved cited work.

Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:59:24.038461Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:59:22.365820Z digest=sha256:8d929f2e44a616ab1477a6a197c0ff7f6f55aeeacaf26620406958ed74c9c405

Observation 795d9b8f-73d6-49f7-b812-0e8a96d55842 · outbound

This paper cites Bansal, T.

Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models Bansal, T

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:59:23.844769Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:59:22.459073Z digest=sha256:15d8dc9af2966f13e2ed742fca24042890a69a653bc9821c8821749454a75d9e

Observation 1ca04bda-1f1b-44e7-b44d-3c3c52e73e67 · outbound

This paper cites GPT-4o System Card.

Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models GPT-4o System Card

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:22.596946Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:22.596946Z digest=sha256:0d4980bc7d9eb46f548b7ccd4bfae61cc60e7d6143f34ebafe3cfc14b204e2a4

Observation 7da94750-5079-4bce-b70c-4ae33764ce60 · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:22.692295Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:22.692295Z digest=sha256:216c44923fb55fd642a3aecfa06318547ae5f54ae7e0b8ba345061f7bc727f2a

Observation 3e0959b3-67cf-4dbc-b435-144b138447c1 · outbound

This paper cites Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search.

Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models Mulberry: Empowering MLLM with o1-like Reasoning and Reflection via Collective Monte Carlo Tree Search

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:22.779575Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:22.779575Z digest=sha256:e7a7a8bad7e66fcc52831e1502524bfe94a93874a6be256a4e5e76cd99f3c7de

Observation 064d151c-f6d6-48f8-abe0-6684ee452d39 · outbound

This paper cites Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling.

Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:22.902200Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:22.902200Z digest=sha256:8174a1e93d7b6694c97597616938289e3d8e575dca244880bfaa176e3256894f

Observation 93489210-a930-4f3e-b293-6f091b59bba1 · outbound

This paper cites MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning.

Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:23.018390Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:23.018390Z digest=sha256:da674bc973525ff7523d388a9c00cbff3bfc8b49460172afeadba4ae6c94521a

Observation 8f16e8fb-f784-4672-a0c3-8e0d2dff52f7 · outbound

This paper cites an unresolved cited work.

Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models Unresolved cited work

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T12:59:23.141014Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:59:23.141014Z digest=sha256:32426e6e1c218fa5e39cb4f36e82cf5165d7b3955495aab341ebaf9b907df5fd

Pith citing papers

Observation 12bd2d1f-7a14-4920-a638-28601499118e · inbound

Latent Visual Reasoning cites this paper.

Latent Visual Reasoning Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models

Reference 16

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T18:41:30.420970Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-15T18:41:30.307521Z digest=sha256:f1fb81a5f4c7cc65b6af4b95f9c59909f2b56b6f08b595fd1f027d3f581e8c5f

Observation 600bdcfc-05fa-415f-b1e7-da58c3897ad4 · inbound

D$^2$Evo: Dual Difficulty-Aware Self-Evolution for Data-Efficient Reinforcement Learning cites this paper.

D$^2$Evo: Dual Difficulty-Aware Self-Evolution for Data-Efficient Reinforcement Learning Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models

Reference 54

Resolution
verified exact
arxiv_id, observed 2026-05-19T20:22:45.278394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-19T20:21:48.657926Z digest=sha256:c5170d5559f8c4229cbeadb9735c809f4b4cf0af0d7b33e2bf5880a7b85ee8c9

Observation 3bfa07d7-4e8b-4106-8fa1-53a296c7ad16 · inbound

VeriEvol: Scaling Multimodal Mathematical Reasoning via Verifiable Evol-Instruct cites this paper.

VeriEvol: Scaling Multimodal Mathematical Reasoning via Verifiable Evol-Instruct Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models

Reference 26

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T10:39:46.065840Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-26T08:34:27.719022Z digest=sha256:5030c489ead8c5e4472d9de2efae03128cbb04aa009cab328ce39b4ba9175870

Observation 8c55e748-8cfe-4f10-84db-cdea0c04c937 · inbound

LASER: A Corrective Lens for LVLMs via Visual Attention Preservation and Sink Suppression cites this paper.

LASER: A Corrective Lens for LVLMs via Visual Attention Preservation and Sink Suppression Infi-MMR: Curriculum-based Unlocking Multimodal Reasoning via Phased Reinforcement Learning in Multimodal Small Language Models

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-07-03T16:58:42.492709Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-07-03T16:53:43.427625Z digest=sha256:d7c947a448fc14b7555417022f7a0706f24ce5127daa2b4c40553825114c2f66