Pith. sign in

Paper Citation Record · LEDGER

MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

As of 23 July 2026, this Paper Citation Record lists 27 of 27 outbound references and 90 inbound Pith citation observations for arXiv:2503.07365.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2503.07365 v2

Coverage vector

measured 27 of 27 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-13T09:59:58.388186Z

measured 117 of 117 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-07-23T06:31:01.910684+00:00

measured 90 of 90 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-14T23:22:13.847876Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-10T12:15:01.137692Z

Reference resolution

27 of 27 outbound references displayed

  • verified exact0
  • verified fuzzy24
  • unresolved0
  • parse uncertain2
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 3a71f8c4-904f-4441-9c92-7e8d7835a192 · outbound

This paper cites an unresolved cited work.

MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning Unresolved cited work

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T09:59:58.427192Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-13T09:59:58.388186Z digest=sha256:44f383ad6a5f41543a70ddfdf6d5debacf640fbd90ee5f93ef26697ec7b28195

Observation 836cbd1c-f2d0-47ff-b755-4f4844ffa578 · outbound

This paper cites an unresolved cited work.

MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning Unresolved cited work

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T09:59:58.403070Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-13T09:59:58.388186Z digest=sha256:09cbc958a22efde1e0cecdd1e62b5ff92461a8af4175a34ef9cd9a2fd86d0861

Observation 09f1c007-8d5a-4479-989e-65b7aa4b930e · outbound

This paper cites First, find two non-parallel vectors in the plane A1C1D.

MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning First, find two non-parallel vectors in the plane A1C1D

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T09:59:58.405206Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-13T09:59:58.388186Z digest=sha256:1cd3e8280d49275c0b364cf837360747a57894d1ff28e2653e5b8103bec798d8

Observation 6cc61028-e20f-4b02-bb1c-bfee54bf6467 · outbound

This paper cites • To find the normal vector to the plane A1C1D, we use the cross product of two vectors in the plane: −−→A1D = (1−0,0−0,1−0) = (1 ,0,1), −−→C1D = (0−0,1−0,1−0) = (0 ,1,1).

MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning • To find the normal vector to the plane A1C1D, we use the cross product of two vectors in the plane: −−→A1D = (1−0,0−0,1−0) = (1 ,0,1), −−→C1D = (0−0,1−0,1−0) = (0 ,1,1)

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T09:59:58.407325Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-13T09:59:58.388186Z digest=sha256:43c243e787693c590c5be65913d7be3a5e103b997f9eaf1b352e07358d1cb35f

Observation 048fdcaf-ef6f-467d-bee6-4933d52ed586 · outbound

This paper cites • Let M = (x, y,0) be a point in the plane ABCD (since M∈plane ABCD , the z-coordinate of M is 0).

MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning • Let M = (x, y,0) be a point in the plane ABCD (since M∈plane ABCD , the z-coordinate of M is 0)

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T09:59:58.409432Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-13T09:59:58.388186Z digest=sha256:2054e9f96c8f21daf704f00438f1bdbdbe6c67c697139e88eb8fbc68e4109405

Observation 2cc12086-2e55-47f5-b3e8-3d5751e60411 · outbound

This paper cites • The length of D1M is given by the magnitude of−−−→D1M:|−−−→D1M|= √ 12 + 12 + (−1)2 =√1 + 1 + 1 =√ 3.

MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning • The length of D1M is given by the magnitude of−−−→D1M:|−−−→D1M|= √ 12 + 12 + (−1)2 =√1 + 1 + 1 =√ 3

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T09:59:58.411903Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-13T09:59:58.388186Z digest=sha256:8f2d4cb0dd62622c2d9666fc5f9bd5699c4c55249a84c0b7152d75c937f61266

Observation 4fb9efae-4285-4119-a8f5-0b0a5d879d66 · outbound

This paper cites The spring force remains 60 N (the initial force that balances the weight of A).

MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning The spring force remains 60 N (the initial force that balances the weight of A)

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T09:59:58.413987Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-13T09:59:58.388186Z digest=sha256:8153fc6d51a6672c15fb49fafcf282a0f2aff6570eb12eb74b975356f8c1a550

Observation 41f7dc3a-7872-4aa3-bb80-797f5243743b · outbound

This paper cites • The force between A and B is the force B exerts on A, which is the weight of B: FB = mBg = 4 kg· 10 m/s2 = 40 N.

MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning • The force between A and B is the force B exerts on A, which is the weight of B: FB = mBg = 4 kg· 10 m/s2 = 40 N

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T09:59:58.415997Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-13T09:59:58.388186Z digest=sha256:da8ae33cada115e4faf451c1b219eb547881d6318bfc3f255d2936c7df7f6990

Observation 2b8a0cc7-4cab-4d03-82cf-0f19c9a608b7 · outbound

This paper cites The net force on A causes an upward acceleration a, which is the same as the acceleration of B because they are in contact.

MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning The net force on A causes an upward acceleration a, which is the same as the acceleration of B because they are in contact

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T09:59:58.417972Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-13T09:59:58.388186Z digest=sha256:d820c1b3f188b1a7c178e1618274e33aad3e5187acc16c083d02e25af6b7db5a

Observation ff4bdf6b-dc27-461b-9659-28291827bbb8 · outbound

This paper cites an unresolved cited work.

MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning Unresolved cited work

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T09:59:58.419835Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-13T09:59:58.388186Z digest=sha256:19c7752dec3f4ba623abb385190bc14b3ee17b3d71608a6cde81a8146cfe07b2

Observation b85cf59e-5817-421a-b0a5-524be969417c · outbound

This paper cites an unresolved cited work.

MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning Unresolved cited work

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T09:59:58.421680Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-13T09:59:58.388186Z digest=sha256:16074802e709f724355f4dbcada2604af67e7256be77a67bebc5705fa8db8bd1

Observation f6237feb-96d2-42a0-9251-fbdeb3eccab4 · outbound

This paper cites • The net force on the system is the extra weight from B: Fnet, total = 40 N.

MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning • The net force on the system is the extra weight from B: Fnet, total = 40 N

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T09:59:58.423658Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-13T09:59:58.388186Z digest=sha256:c45fe359d42307f2164c3d5a3d04ae59710204a9f9fe45c68c88193b11cc361d

Observation 2287ebaf-e5ee-4a43-a07d-4a634f16d4cf · outbound

This paper cites <answer>The answer is C.

MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning <answer>The answer is C

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T09:59:58.425390Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-13T09:59:58.388186Z digest=sha256:6046ddabd0b7d32e66d704ff1118b7b4fd72661f7a86f8f8fc0d7b6cb0c758b7

Observation 4320e996-092b-44ee-8a9a-21a012bfe60c · outbound

This paper cites • The reaction steps are.

MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning • The reaction steps are

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T09:59:58.450020Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-13T09:59:58.388186Z digest=sha256:939221c16a579c9ec5398a8cc380827b74ce7c12f9e75656afbd900366e25bfb

Observation da8d42b2-4b81-44fc-b53f-f0ffc8e9f70f · outbound

This paper cites an unresolved cited work.

MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning Unresolved cited work

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T09:59:58.428967Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-13T09:59:58.388186Z digest=sha256:1286e06addabf8b84c96cb603df626918c17b0cd2ccd4f678f1aed2a4eb28fba

Observation 3571e21f-6b0c-4408-8901-1e8bcdd21ce7 · outbound

This paper cites an unresolved cited work.

MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning Unresolved cited work

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T09:59:58.430709Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-13T09:59:58.388186Z digest=sha256:43c2ffb3fb1724d2950c359708d94d8dccc7ab39f061a1755a0c0f13448bbbd2

Observation 3f04d6f3-8441-4017-8313-91ad66015ee3 · outbound

This paper cites The solution is acidic.

MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning The solution is acidic

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T09:59:58.432361Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-13T09:59:58.388186Z digest=sha256:1f456c4cb5030babb66aea6fb73ecd8bda98ec4af78e886b6282a7c972e22e28

Observation 950715b3-9c92-4743-8837-28cbf38b8be9 · outbound

This paper cites MM-EUREKA-32B To solve this problem, we need to analyze the titration of a weak dicarboxylic acid H2C2O4 (oxalate) with a strong base NaOH.

MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning MM-EUREKA-32B To solve this problem, we need to analyze the titration of a weak dicarboxylic acid H2C2O4 (oxalate) with a strong base NaOH

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T09:59:58.434389Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-13T09:59:58.388186Z digest=sha256:9de9d32064f3486474cfcb823096369a49bb3d8e227f0e991e7737dce5897dc5

Observation 52f44150-e1d5-4524-8b96-b45f58161245 · outbound

This paper cites an unresolved cited work.

MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning Unresolved cited work

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T09:59:58.436068Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-13T09:59:58.388186Z digest=sha256:f5d96e627d8ab05c9603b763edfa854241162588c3b2ffd44d799659735560f9

Observation 153b6b76-21a1-4010-baa0-31bfd97ae07f · outbound

This paper cites • At this point, the first equivalence point is reached.

MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning • At this point, the first equivalence point is reached

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T09:59:58.437890Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-13T09:59:58.388186Z digest=sha256:a8ab070c73167f219630841c2b8ebf5ebf6a49bbacb59a2287a9b4714c654798

Observation 8a848e1d-5f40-40ce-a464-13bd2b99095e · outbound

This paper cites an unresolved cited work.

MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning Unresolved cited work

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T09:59:58.439602Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-13T09:59:58.388186Z digest=sha256:359074815611399a59c0efe9bbeb5f79d86c0cd403cc482755cdc559822c6002

Observation 66606d7f-ef66-4b00-9703-59f74e19408e · outbound

This paper cites an unresolved cited work.

MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning Unresolved cited work

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T09:59:58.441296Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-13T09:59:58.388186Z digest=sha256:56f2a92a86c9e96d90ab363f58dee57bc1614f38422e2cd6ef75281ece4e9948

Observation b70c9eae-aa11-4a74-8b01-1e06b2c60f9b · outbound

This paper cites <answer>The answer is D.

MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning <answer>The answer is D

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T09:59:58.442820Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-13T09:59:58.388186Z digest=sha256:8c514891543daee04ae26a1cf2f2660ac36958ca9bc46abff3607e6076741cec

Observation d87ecf62-d979-45b0-80dd-4f5f0bbe705a · outbound

This paper cites • The reaction steps are.

MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning • The reaction steps are

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T09:59:58.444596Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-13T09:59:58.388186Z digest=sha256:7711c661db9523870e1a1453ff8f204ce413b20112a5c7eafdbf822cae715f58

Observation 0c480990-8698-4d33-b033-d7a0c5460f15 · outbound

This paper cites an unresolved cited work.

MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning Unresolved cited work

Reference 25

Resolution
parse uncertain
raw_fallback, observed 2026-05-13T09:59:58.446600Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-13T09:59:58.388186Z digest=sha256:b2ef404c2f1a898ae1c1f234c62e99f04b3bbc86b1fddae4f86ac0cd71ded391

Observation 67c6f9b3-35e9-44bb-bd1b-690a34b07cd7 · outbound

This paper cites an unresolved cited work.

MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning Unresolved cited work

Reference 26

Resolution
parse uncertain
raw_fallback, observed 2026-05-13T09:59:58.448244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-13T09:59:58.388186Z digest=sha256:07f761ccb40233f84127154b52fe9594db819124067c30bf0c244cf535736d2f

Observation ac5e316f-f07d-4ae4-8d06-3118342fad7c · outbound

This paper cites MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning.

MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 27

Resolution
metadata mismatch
local_arxiv, observed 2026-05-13T09:59:58.400863Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-13T09:59:58.388186Z digest=sha256:6ca9b4bac864fbde3346576e3312571b0437a37e1055a233f3fc4b17ca0239de

Pith citing papers

Observation d106f1ec-50d4-4477-bcf9-b10ec5d78280 · inbound

From System 1 to System 2: A Survey of Reasoning Large Language Models cites this paper.

From System 1 to System 2: A Survey of Reasoning Large Language Models MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 257

Resolution
verified exact
arxiv_id, observed 2026-05-13T01:36:24.446735Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-13T01:36:23.845366Z digest=sha256:e0a46618f66153aaf6571c6c2c2edb7785a67772f99d0eee771e87715462608d

Observation ac5e316f-f07d-4ae4-8d06-3118342fad7c · inbound

MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning cites this paper.

MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 27

Resolution
metadata mismatch
local_arxiv, observed 2026-05-13T09:59:58.400863Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-13T09:59:58.388186Z digest=sha256:6ca9b4bac864fbde3346576e3312571b0437a37e1055a233f3fc4b17ca0239de

Observation b080b906-6d82-4b71-b32b-ef10b95b99d9 · inbound

R1-VL: Learning to Reason with Multimodal Large Language Models via Step-wise Group Relative Policy Optimization cites this paper.

R1-VL: Learning to Reason with Multimodal Large Language Models via Step-wise Group Relative Policy Optimization MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 27

Resolution
verified exact
local_arxiv, observed 2026-05-16T15:04:22.870184Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-16T15:04:22.690503Z digest=sha256:6fc36563f132c3a6673857d3f7a03ea965ffb68ad91e753defd8a5737e0d4c0c

Observation eb2ce270-e81c-495b-b521-a160656a4591 · inbound

Time-R1: Post-Training Large Vision Language Model for Temporal Video Grounding cites this paper.

Time-R1: Post-Training Large Vision Language Model for Temporal Video Grounding MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 35

Resolution
verified exact
local_arxiv, observed 2026-05-17T02:40:06.512877Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-17T02:40:06.454859Z digest=sha256:e8a166e972b476998811dcd0dddc45c039b1e38c4f0541a0cff9bf825e8092ef

Observation eb0a7a8f-d6ef-4bbc-861e-41faac02966f · inbound

OpenVLThinker: Complex Vision-Language Reasoning via Iterative SFT-RL Cycles cites this paper.

OpenVLThinker: Complex Vision-Language Reasoning via Iterative SFT-RL Cycles MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 48

Resolution
verified exact
local_arxiv, observed 2026-05-19T06:59:03.214557Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-19T06:59:03.112252Z digest=sha256:0d27b18c4ed7a4d7aa8a1da7835a6a60c318f08c1aa1349e253304af20ba9a40

Observation e80f9667-1b61-4f9f-a389-1ecc50b2b97f · inbound

UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning cites this paper.

UI-R1: Enhancing Efficient Action Prediction of GUI Agents by Reinforcement Learning MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-05-16T11:02:41.384566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-16T11:02:41.335059Z digest=sha256:220eb06e9b6c7f17377e61500a4b6441d252561f1d466e12d5461b5589c5d72c

Observation ff7a526d-feeb-4460-bd3f-958d0d85ffd4 · inbound

VLM-R1: A Stable and Generalizable R1-style Large Vision-Language Model cites this paper.

VLM-R1: A Stable and Generalizable R1-style Large Vision-Language Model MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-05-13T01:13:57.456642Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-13T01:13:57.368874Z digest=sha256:55234a4a24c41cd1b5ba731d3a6feaf07ca6dfdc9fa36b816cf050fb40c5516c

Observation c6b4050b-173d-4dca-8512-1f906f47cc38 · inbound

GUI-R1 : A Generalist R1-Style Vision-Language Action Model For GUI Agents cites this paper.

GUI-R1 : A Generalist R1-Style Vision-Language Action Model For GUI Agents MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 20

Resolution
verified exact
local_arxiv, observed 2026-05-15T02:10:58.077914Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-15T02:10:57.976448Z digest=sha256:5170549a886c53a14ec97e14ef8729c70961984804c255cc01d9eb3ba8c57dd6

Observation 18282351-f8ca-4585-bb5d-05bcc850b501 · inbound

DeepEyes: Incentivizing "Thinking with Images" via Reinforcement Learning cites this paper.

DeepEyes: Incentivizing "Thinking with Images" via Reinforcement Learning MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-11T14:42:56.935854Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-11T14:42:56.565621Z digest=sha256:fc5fefd1ece226bbae5f451e158afcda73112e9a99bcbc12ba297489f070b563

Observation 34c58fa3-ceb0-42be-b349-0f0dd00aad9e · inbound

v1: Learning to Point Visual Tokens for Multimodal Grounded Reasoning cites this paper.

v1: Learning to Point Visual Tokens for Multimodal Grounded Reasoning MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 4

Resolution
metadata mismatch
local_arxiv, observed 2026-05-19T12:37:17.460305Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-19T12:36:32.030301Z digest=sha256:50ee73149e9967b7fd4bbe65528d7bbdd0b6aaec302a8fb5fc714c21370e409a

Observation 632efe2e-7af5-4252-9ccb-b06c0f4659c4 · inbound

Lingshu: A Generalist Foundation Model for Unified Multimodal Medical Understanding and Reasoning cites this paper.

Lingshu: A Generalist Foundation Model for Unified Multimodal Medical Understanding and Reasoning MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 6

Resolution
metadata mismatch
local_arxiv, observed 2026-05-15T11:10:59.675331Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-07-11T11:50:26.030339Z digest=sha256:b6bd0953faffedaa2417a1554175cc0bb4f00be8c8ef863a7251b2c41c89ee97

Observation 2f4c5199-f532-4876-ab9f-85090c7a7114 · inbound

Lingshu: A Generalist Foundation Model for Unified Multimodal Medical Understanding and Reasoning cites this paper.

Lingshu: A Generalist Foundation Model for Unified Multimodal Medical Understanding and Reasoning MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 7

Resolution
metadata mismatch
local_arxiv, observed 2026-05-15T11:10:59.682821Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-07-11T11:50:26.030339Z digest=sha256:1c8ecd97092918b80ecd4a2898c9a0e5615806f6e38bb0f3d19d95c2d8301577

Observation 9ca5680f-280b-46ed-94b6-e4bbf8decb29 · inbound

RealSR-R1: Reinforcement Learning for Real-World Image Super-Resolution with Vision-Language Chain-of-Thought cites this paper.

RealSR-R1: Reinforcement Learning for Real-World Image Super-Resolution with Vision-Language Chain-of-Thought MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 37

Resolution
verified exact
local_arxiv, observed 2026-05-19T08:33:02.141432Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-19T08:32:20.566798Z digest=sha256:92a82ca2c834ac679cf66b0e67c57b18073a9c9832cc2090f9e3fb82f358cf3c

Observation 89bfe261-63ba-4bc5-ac2f-b5aed3907ad6 · inbound

High-Resolution Visual Reasoning via Multi-Turn Grounding-Based Reinforcement Learning cites this paper.

High-Resolution Visual Reasoning via Multi-Turn Grounding-Based Reinforcement Learning MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 27

Resolution
verified exact
local_arxiv, observed 2026-05-19T06:12:07.062177Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-19T06:10:57.219445Z digest=sha256:a5dcf99f75df177d98208332ac8d1d3a7e26ba7fdc7d299bfa7019b7f023c952

Observation d72a3571-b1cd-4393-b4f1-333dcafb43b8 · inbound

Perception-Aware Policy Optimization for Multimodal Reasoning cites this paper.

Perception-Aware Policy Optimization for Multimodal Reasoning MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-05-19T05:12:04.803695Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=arxiv_source observed=2026-05-19T05:11:54.685897Z digest=sha256:9341196745e4fc29e10f58e76ff59eb345e4648d26c34e09e34add6918ba633f

Observation 547a05fb-e74d-4b3a-ae21-7cd9bbccebf7 · inbound

InternVL3.5: Advancing Open-Source Multimodal Models in Versatility, Reasoning, and Efficiency cites this paper.

InternVL3.5: Advancing Open-Source Multimodal Models in Versatility, Reasoning, and Efficiency MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 90

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:58:59.219806Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-10T11:58:58.660564Z digest=sha256:3a69dde8b39e09f2e56531fe2791a031fa75a19756af31559b566b620cc18cff

Observation ccef3016-2af6-4d21-82e4-97fc9798ed8a · inbound

Mini-o3: Scaling Up Reasoning Patterns and Interaction Turns for Visual Search cites this paper.

Mini-o3: Scaling Up Reasoning Patterns and Interaction Turns for Visual Search MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 26

Resolution
verified exact
local_arxiv, observed 2026-05-18T01:17:55.540566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-18T01:17:55.500268Z digest=sha256:bf2550573228f2b75bfd81d5f4475a138a479dd947ef975d0b65f2bf901364d0

Observation 9f1c250c-57b9-4e58-8e2c-4a66ec98a420 · inbound

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning cites this paper.

Mixture-of-Visual-Thoughts: Exploring Context-Adaptive Reasoning Mode Selection for General Visual Reasoning MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 28

Resolution
verified exact
local_arxiv, observed 2026-05-18T13:36:25.130294Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=arxiv_source observed=2026-05-18T13:33:12.508639Z digest=sha256:6f4cb347585c0ba3a93949a5c7617803a1a6ecfeb02b5ccf746b811636f5cfc1

Observation 649da9e0-17f2-49a3-beb2-0d99f651384e · inbound

Learning to Pose Problems: Reasoning-Driven and Solver-Adaptive Data Synthesis cites this paper.

Learning to Pose Problems: Reasoning-Driven and Solver-Adaptive Data Synthesis MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 11

Resolution
verified exact
local_arxiv, observed 2026-05-17T23:00:25.467935Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-17T22:59:48.260121Z digest=sha256:20c3647022b0968ef5c9a0dcd34ef835571bf66223f7609afab520ea9445928d

Observation 7c9778be-37ba-4697-82dc-55c8863c0f2b · inbound

REVISOR: Beyond Textual Reflection, Towards Multimodal Introspective Reasoning in Long-Form Video Understanding cites this paper.

REVISOR: Beyond Textual Reflection, Towards Multimodal Introspective Reasoning in Long-Form Video Understanding MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 30

Resolution
verified exact
local_arxiv, observed 2026-05-17T22:20:22.848543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-17T22:19:36.366837Z digest=sha256:fb7b53c538fae3b876543b52c04b060fe1beff82016f752a2a0c68ae62a0791b

Observation e830ca06-9952-4c82-8162-2386324ae9e8 · inbound

LongVT: Incentivizing "Thinking with Long Videos" via Native Tool Calling cites this paper.

LongVT: Incentivizing "Thinking with Long Videos" via Native Tool Calling MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 31

Resolution
verified exact
local_arxiv, observed 2026-05-22T12:31:32.224788Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-22T12:26:35.347190Z digest=sha256:a785f75fb1008ca32c00c663ba8815d04500cb3292a24b8c1461888d1fe6fb89

Observation d323f4f9-73d2-4707-bd6c-ede1f7656b52 · inbound

OneThinker: All-in-one Reasoning Model for Image and Video cites this paper.

OneThinker: All-in-one Reasoning Model for Image and Video MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 54

Resolution
verified exact
local_arxiv, observed 2026-05-17T02:11:26.523868Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-17T02:09:39.820651Z digest=sha256:4e3936f41719b21872c273cab92af33c4115d34a9c568ce6c2285cdbbe16503d

Observation 4dac5239-bf64-498a-97cd-e09d49350e56 · inbound

Addressing Overthinking in Large Vision-Language Models via Gated Perception-Reasoning Optimization cites this paper.

Addressing Overthinking in Large Vision-Language Models via Gated Perception-Reasoning Optimization MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 8

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T16:08:04.387844Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-16T16:06:52.052810Z digest=sha256:68d3f3f96fe58a3eed7392f3a88ae483f695cfd7bc1efd0edbba0656b465c150

Observation 070e8464-9c02-4fd7-8b6e-badcda997b1b · inbound

Omni-R1: Towards the Unified Generative Paradigm for Multimodal Reasoning cites this paper.

Omni-R1: Towards the Unified Generative Paradigm for Multimodal Reasoning MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-16T14:37:59.995943Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-16T14:37:05.402850Z digest=sha256:a7e197dbd73afc0c42440c1fc737b5d624f8c41f79009c4528f6fd3078c46c8f

Observation 7801e326-0fa9-4a3d-a4d7-978e3e75fda7 · inbound

CodePercept: Code-Grounded Visual STEM Perception for MLLMs cites this paper.

CodePercept: Code-Grounded Visual STEM Perception for MLLMs MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 32

Resolution
unresolved
no resolver link, observed 2026-07-14T23:22:13.847876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T23:22:13.847876Z digest=sha256:fb757043460d3f257e626dfb01b22c6f6c9a928d2a9d14643ca5a24c28581f8e

Observation 4338ff4a-26a0-4ecd-871b-088210bc443c · inbound

Deeper Thought, Weaker Aim: Understanding and Mitigating Perceptual Impairment during Reasoning in Multimodal Large Language Models cites this paper.

Deeper Thought, Weaker Aim: Understanding and Mitigating Perceptual Impairment during Reasoning in Multimodal Large Language Models MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-05-21T12:00:04.245706Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-21T11:56:29.743281Z digest=sha256:cff9e57ca484aa4112830967005ace1bcfbbda396a6ab8bbffd84f13ac5bf406

Observation 0018b85f-cfdb-474d-9a5d-8c9e78dffd63 · inbound

Saliency-R1: Enforcing Interpretable and Faithful Vision-language Reasoning via Saliency-map Alignment Reward cites this paper.

Saliency-R1: Enforcing Interpretable and Faithful Vision-language Reasoning via Saliency-map Alignment Reward MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:20:47.855452Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-10T19:59:19.379119Z digest=sha256:8c314645a86f5cbf0c39565affc1e83e303397802bda0d7d86f29b54fc6455f5

Observation be0a26de-8b65-49be-9108-a5c2ae0f5079 · inbound

Visually-Guided Policy Optimization for Multimodal Reasoning cites this paper.

Visually-Guided Policy Optimization for Multimodal Reasoning MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 4

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T06:51:24.226347Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-10T17:25:26.321714Z digest=sha256:76c0f01265442301d48597a6cf42400d656cf31a2fe09f86f3a647aa67ebcf46

Observation de8e13cb-5ca7-42c6-87b6-f8ba24ac6b75 · inbound

Visually-Guided Policy Optimization for Multimodal Reasoning cites this paper.

Visually-Guided Policy Optimization for Multimodal Reasoning MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 4

Resolution
metadata mismatch
local_arxiv, observed 2026-05-25T06:45:26.226255Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-25T06:41:32.231185Z digest=sha256:7878d3427549578d1be9b606dbc70b92950705387c81758d28e1333025cbc593

Observation a53d2266-4228-4da2-9647-3ea6696e3b78 · inbound

VL-Calibration: Decoupled Confidence Calibration for Large Vision-Language Models Reasoning cites this paper.

VL-Calibration: Decoupled Confidence Calibration for Large Vision-Language Models Reasoning MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T06:51:14.033427Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-10T17:25:44.453060Z digest=sha256:296a5ca73b56d99a31372cd2ed41dd8c1d2802aa8596ebdb149f942dd8833550

Observation 3576c718-e196-4517-96de-efac34d2fb7e · inbound

Cognitive Pivot Points and Visual Anchoring: Unveiling and Rectifying Hallucinations in Multimodal Reasoning Models cites this paper.

Cognitive Pivot Points and Visual Anchoring: Unveiling and Rectifying Hallucinations in Multimodal Reasoning Models MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 101

Resolution
verified exact
arxiv_id, observed 2026-05-11T09:25:58.791392Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-10T16:03:15.222571Z digest=sha256:4777034d112dca1d85b7f0cce6d60d2b121f20606bd463e84a79a94875ba549e

Observation 5634b9ab-0285-412a-8aca-e12f02287f2c · inbound

Cognitive Pivot Points and Visual Anchoring: Unveiling and Rectifying Hallucinations in Multimodal Reasoning Models cites this paper.

Cognitive Pivot Points and Visual Anchoring: Unveiling and Rectifying Hallucinations in Multimodal Reasoning Models MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 114

Resolution
unresolved
no resolver link, observed 2026-07-12T22:48:45.647588Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T22:48:45.647588Z digest=sha256:c19af96e42eb1e1d5895c8cd838119e22bc766bb7b3452c06ceee73e1bb6deac

Observation 70bb657a-fc7a-4dbb-bc57-444d0fea4e25 · inbound

AnySlot: Goal-Conditioned Vision-Language-Action Policies for Zero-Shot Slot-Level Placement cites this paper.

AnySlot: Goal-Conditioned Vision-Language-Action Policies for Zero-Shot Slot-Level Placement MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-07-12T22:40:19.969447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T22:40:19.969447Z digest=sha256:df3a7396578d473284353ec2c43c922b604c9a35b1bea244a584134015231e22

Observation 68e156f8-3506-4f3e-8848-42321e8028e4 · inbound

SignReasoner: Compositional Reasoning for Complex Traffic Sign Understanding via Functional Structure Units cites this paper.

SignReasoner: Compositional Reasoning for Complex Traffic Sign Understanding via Functional Structure Units MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-11T09:31:04.614361Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-10T15:58:48.213961Z digest=sha256:973850d1e09ca494d876d36cc70e709ff8cacd99a17e78e370c580f52e7c15e2

Observation 2a440aac-07c5-4b49-9c2a-9f012952295e · inbound

DocSeeker: Structured Visual Reasoning with Evidence Grounding for Long Document Understanding cites this paper.

DocSeeker: Structured Visual Reasoning with Evidence Grounding for Long Document Understanding MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-11T09:05:58.437658Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-10T16:16:58.889065Z digest=sha256:47417695bb96decdc16c75a7e48183a1f5655f75a523d1fe549f664ae408ff7c

Observation 563a0d19-6571-484d-979a-9217c6d21d6b · inbound

DocSeeker: Structured Visual Reasoning with Evidence Grounding for Long Document Understanding cites this paper.

DocSeeker: Structured Visual Reasoning with Evidence Grounding for Long Document Understanding MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-12T06:26:24.468847Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-12T04:17:55.318813Z digest=sha256:e93922279a7fa058b54baef68a9af50a78a5005f22e07a877ac6d8baff24685f

Observation 1d7fb70e-9f12-4abc-a220-8e22687d2130 · inbound

DT2IT-MRM: Debiased Preference Construction and Iterative Training for Multimodal Reward Modeling cites this paper.

DT2IT-MRM: Debiased Preference Construction and Iterative Training for Multimodal Reward Modeling MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-11T13:26:04.028435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-10T01:45:30.001398Z digest=sha256:d983c65a61ac4d290ce0b0e5a6ccdb1965c50dcc38d5e4ed8cb3cfff504414b5

Observation d468757c-1a7d-4194-8c31-c2238b417063 · inbound

SSL-R1: Self-Supervised Visual Reinforcement Post-Training for Multimodal Large Language Models cites this paper.

SSL-R1: Self-Supervised Visual Reinforcement Post-Training for Multimodal Large Language Models MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-05-11T13:36:08.597995Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-10T01:23:32.849326Z digest=sha256:a0357ca4d2786f04d8bf0630b5d0a35b7dc3c62054979f81c08b4ad512acb53c

Observation 8a43710e-1551-4cf1-b9cc-5ae2f5c1851e · inbound

Measure Twice, Click Once: Co-evolving Proposer and Visual Critic via Reinforcement Learning for GUI Grounding cites this paper.

Measure Twice, Click Once: Co-evolving Proposer and Visual Critic via Reinforcement Learning for GUI Grounding MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 57

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T14:16:05.947860Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=arxiv_source observed=2026-05-09T23:05:05.251150Z digest=sha256:428068bc7fb17de603b0d72d6c992cbb88052235a29c7e73fdc28131a3e97258

Observation 71bf1991-f2e9-4fe4-a8d2-818a29b30d61 · inbound

Supermassive Black Hole Winds in X-rays: SUBWAYS IV. Tracing Radio Emission and Unveiling the Role of Winds cites this paper.

Supermassive Black Hole Winds in X-rays: SUBWAYS IV. Tracing Radio Emission and Unveiling the Role of Winds MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-07-12T18:39:55.906516Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T18:39:55.906516Z digest=sha256:7941676a5fed453c30a309a1a967ce7ca1895e5e4d6e858b6de44d06ddf2d973

Observation e5e0c1af-f34b-4a29-a283-b2e3fab16fc2 · inbound

S1-VL: Scientific Multimodal Reasoning Model with Thinking-with-Images cites this paper.

S1-VL: Scientific Multimodal Reasoning Model with Thinking-with-Images MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-09T22:34:07.608449Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-09T22:26:03.080281Z digest=sha256:5aa56fda4b7f189ab3ff11303281a3133574e9c8bac1fc60b5f60533a72cb8eb

Observation 80217b3c-c117-4588-b3cd-6ff29f2946bd · inbound

CharTide: Data-Centric Chart-to-Code Generation via Tri-Perspective Tuning and Inquiry-Driven Evolution cites this paper.

CharTide: Data-Centric Chart-to-Code Generation via Tri-Perspective Tuning and Inquiry-Driven Evolution MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-11T19:06:10.300648Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=arxiv_source observed=2026-05-08T12:40:34.424341Z digest=sha256:970045de4a6b1c213e6edc265b66f3f47e83f68f13dd2ed5b73affa33940f055

Observation 7ab68e46-77fe-4b83-9d54-f68508c45d50 · inbound

CGC: Compositional Grounded Contrast for Fine-Grained Multi-Image Understanding cites this paper.

CGC: Compositional Grounded Contrast for Fine-Grained Multi-Image Understanding MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-11T19:16:08.319738Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-08T12:26:01.568507Z digest=sha256:503ea6c976d9ea620bd4171ffe819214561974ff5ef0341f2c18fe7225f53e04

Observation 8893315d-6adc-4639-9782-245356dbe89d · inbound

Persistent Visual Memory: Sustaining Perception for Deep Generation in LVLMs cites this paper.

Persistent Visual Memory: Sustaining Perception for Deep Generation in LVLMs MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-05-11T16:01:22.966186Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-09T18:53:06.494640Z digest=sha256:8a8a475290bc3351a3853684289b14fa17f837d50ee463b41a19d5bf7976d2ee

Observation 0197f5bc-b250-4550-9a20-0b925aafa808 · inbound

Persistent Visual Memory: Sustaining Perception for Deep Generation in LVLMs cites this paper.

Persistent Visual Memory: Sustaining Perception for Deep Generation in LVLMs MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-05-11T01:50:51.684916Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-11T01:49:15.136031Z digest=sha256:d350e48955c213f3c84d3e6894f4bb7cc44f1b149d4282c73c9493a2f753e022

Observation bff07238-3343-47e6-9047-a56e93f7539a · inbound

SeePhys Pro: Diagnosing Modality Transfer and Blind-Training Effects in Multimodal RLVR for Physics Reasoning cites this paper.

SeePhys Pro: Diagnosing Modality Transfer and Blind-Training Effects in Multimodal RLVR for Physics Reasoning MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-12T05:11:22.832689Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-12T05:07:39.571227Z digest=sha256:2ebb49d9ed53c4e00468f9ec009a8a401de49e94f42955832dfbeafed18c9f3c

Observation 59ed2494-a958-4be5-9957-d0dfc468257a · inbound

SeePhys Pro: Diagnosing Modality Transfer and Blind-Training Effects in Multimodal RLVR for Physics Reasoning cites this paper.

SeePhys Pro: Diagnosing Modality Transfer and Blind-Training Effects in Multimodal RLVR for Physics Reasoning MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-13T07:47:31.599131Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-13T07:43:50.633779Z digest=sha256:94c9865d0f5094ef54ad2bb0a42b8e3457b25c1f58c16ca69664a8e11829903d

Observation be6655e0-9e39-45fa-9041-5e8900fad38a · inbound

Reflection Anchors for Propagation-Aware Visual Retention in Long-Chain Multimodal Reasoning cites this paper.

Reflection Anchors for Propagation-Aware Visual Retention in Long-Chain Multimodal Reasoning MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-12T07:41:26.438867Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-12T02:24:12.349405Z digest=sha256:267073530354f5ba39023a12ac4e18239dec18f99f5c826fd9b3fd2d23e5d9df

Observation 08c7088f-7578-40eb-a52a-c65fe8c58f83 · inbound

RadThinking: A Dataset for Longitudinal Clinical Reasoning in Radiology cites this paper.

RadThinking: A Dataset for Longitudinal Clinical Reasoning in Radiology MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 82

Resolution
verified exact
arxiv_id, observed 2026-05-12T03:56:21.547786Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-12T03:55:55.359488Z digest=sha256:2779666d976333bddd7951545b18316d40d59491b1fdcd97d98b480e455c9ef3

Observation a313740c-6d2b-4da4-b5d2-beacd9776ef4 · inbound

20/20 Vision Language Models: A Prescription for Better VLMs through Data Curation Alone cites this paper.

20/20 Vision Language Models: A Prescription for Better VLMs through Data Curation Alone MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-13T02:57:09.530432Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-13T02:52:43.674969Z digest=sha256:adfaa5bc5e212fb1e016bd94fa73130fb17dc380fdbbdc9e6182942d7dd4106b

Observation 35f9e795-bc48-4901-a008-b0dc0e76ee57 · inbound

20/20 Vision Language Models: A Prescription for Better VLMs through Data Curation Alone cites this paper.

20/20 Vision Language Models: A Prescription for Better VLMs through Data Curation Alone MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 26

Resolution
verified exact
local_arxiv, observed 2026-05-14T21:29:28.612395Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-14T21:28:37.680681Z digest=sha256:ae3a18f171a54261db117e8626e657d24ae4669a2ee584b4587960aa3540a573

Observation c58556cb-2d01-40a3-903b-f6e4c2f42d4e · inbound

MM-OptBench: A Solver-Grounded Benchmark for Multimodal Optimization Modeling cites this paper.

MM-OptBench: A Solver-Grounded Benchmark for Multimodal Optimization Modeling MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-13T05:52:22.130598Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-13T05:50:25.876554Z digest=sha256:fd6cf82c4a442a66e3750f9c2a5f7e84224968ff00580a3ceb04aa403dd84fc5

Observation baff30dd-9809-4b37-8e51-a429d0107a81 · inbound

Physics-R1: An Audited Olympiad Corpus and Recipe for Visual Physics Reasoning cites this paper.

Physics-R1: An Audited Olympiad Corpus and Recipe for Visual Physics Reasoning MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 32

Resolution
metadata mismatch
local_arxiv, observed 2026-05-15T05:39:47.799619Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=arxiv_source observed=2026-05-15T05:35:32.806871Z digest=sha256:f5927910c40e8d976255564d3b871f9906fc48acb38d60831dc27140e4f4dee3

Observation 25b2168b-c9fa-45ff-9c08-72bd5ed06bae · inbound

CurveBench: A Benchmark for Exact Topological Reasoning over Nested Jordan Curves cites this paper.

CurveBench: A Benchmark for Exact Topological Reasoning over Nested Jordan Curves MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 43

Resolution
metadata mismatch
local_arxiv, observed 2026-05-15T05:39:48.018520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=arxiv_source observed=2026-05-15T05:35:02.980473Z digest=sha256:1af5cfa66dee397d695c210bc5854cc0973697c0386ba95d7078831c54c87645

Observation a6ce0054-2f28-450c-8cf5-265d66b1989b · inbound

CurveBench: A Benchmark for Exact Topological Reasoning over Nested Jordan Curves cites this paper.

CurveBench: A Benchmark for Exact Topological Reasoning over Nested Jordan Curves MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 43

Resolution
metadata mismatch
local_arxiv, observed 2026-05-20T20:43:43.457084Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=arxiv_source observed=2026-05-20T20:40:47.750678Z digest=sha256:e65fa1fd75fe9aac15378a4d11a3bf25537ec60760f797e2666241074c4573f9

Observation 18adee5e-c786-42d9-9da3-0af6fb2cc62a · inbound

ATLAS: Agentic or Latent Visual Reasoning? One Word is Enough for Both cites this paper.

ATLAS: Agentic or Latent Visual Reasoning? One Word is Enough for Both MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-05-15T03:14:53.442872Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-15T03:09:58.411261Z digest=sha256:d05b04d5155e66d53f55f09332202f1684932ff40742a26c759970adf1da605a

Observation 93ae7cbe-87a3-460f-8006-6d501706aac9 · inbound

VideoSeeker: Incentivizing Instance-level Video Understanding via Native Agentic Tool Invocation cites this paper.

VideoSeeker: Incentivizing Instance-level Video Understanding via Native Agentic Tool Invocation MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-05-20T19:38:56.081767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-20T19:37:09.244578Z digest=sha256:35c9bf73ad810f897cb7bc87df3b7b717eb3854ddb375e3345675cac3a5e0ce2

Observation 0182cff3-22c6-4a08-83cd-71f34c364ca9 · inbound

Reasoning Portability: Guiding Continual Learning for MLLMs in the RLVR Era cites this paper.

Reasoning Portability: Guiding Continual Learning for MLLMs in the RLVR Era MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 33

Resolution
verified exact
local_arxiv, observed 2026-05-20T13:33:19.426928Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-20T13:29:11.509966Z digest=sha256:24b94340635f780b4d5aac7f5c7fb43cb7ec643f6ca533a67f8a116d76355742

Observation e7c77d9d-0b59-4783-9f93-7b21d5ab3875 · inbound

Beyond Mode Collapse: Distribution Matching for Diverse Reasoning cites this paper.

Beyond Mode Collapse: Distribution Matching for Diverse Reasoning MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 24

Resolution
verified exact
local_arxiv, observed 2026-05-20T05:33:04.122817Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-20T05:30:37.685873Z digest=sha256:34690fb0fb424181b55ff4598787cf1b4cafccda2dd86ce07de19f4f2777e17f

Observation 3cafac0a-ed74-437a-83cf-3c9b7680abb6 · inbound

ParaVT: Taming the Tool Prior Paradox for Parallel Tool Use in Agentic Video Reinforcement Learning cites this paper.

ParaVT: Taming the Tool Prior Paradox for Parallel Tool Use in Agentic Video Reinforcement Learning MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 19

Resolution
verified exact
local_arxiv, observed 2026-05-21T07:34:02.599365Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-21T07:32:12.180233Z digest=sha256:1250ff7deba06a04bfc8f009ff52f83b1814280b350e0fe46e51b07577cef835

Observation d8dd16e2-9b44-41e6-8459-ed5c2c9ff29c · inbound

ParaVT: Taming the Tool Prior Paradox for Parallel Tool Use in Agentic Video Reinforcement Learning cites this paper.

ParaVT: Taming the Tool Prior Paradox for Parallel Tool Use in Agentic Video Reinforcement Learning MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 19

Resolution
verified exact
local_arxiv, observed 2026-05-22T09:01:19.375108Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-22T08:59:28.405218Z digest=sha256:f8791da26c0c8bd75a433cff9491ffdff8ddf6ef394ec98e0b74f67a3fa66098

Observation 0d640ca5-e639-4282-84db-798777699631 · inbound

RISE: Reliable Improvement in Self-Evolving Vision-Language Models cites this paper.

RISE: Reliable Improvement in Self-Evolving Vision-Language Models MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 29

Resolution
verified exact
local_arxiv, observed 2026-05-21T05:39:40.539633Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-05-21T05:38:26.590720Z digest=sha256:4ff1d6b4872c566354956931fa0187c1bf3055bfee751d787bf51d4b033e1527

Observation f445b463-0e16-451c-92da-d8f9a8a9f21e · inbound

RISE: Reliable Improvement in Self-Evolving Vision-Language Models cites this paper.

RISE: Reliable Improvement in Self-Evolving Vision-Language Models MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 33

Resolution
verified exact
local_arxiv, observed 2026-06-30T17:24:57.557692Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-06-30T17:19:09.596950Z digest=sha256:83a5e44692acc581a50031f1d92ff496f34fae551bf4cb7add6de911b2271e15

Observation 3eb5d7c6-347f-44f1-a663-d799db2f9929 · inbound

AnE: Pushing the Reasoning Frontier of Multimodal LLMs via Anchor Evolution cites this paper.

AnE: Pushing the Reasoning Frontier of Multimodal LLMs via Anchor Evolution MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 28

Resolution
verified exact
local_arxiv, observed 2026-06-30T00:14:04.566054Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-06-29T22:56:39.504430Z digest=sha256:3e93c2814c2979b3ef1e6334c0c4fb7076e50114133ccb4be6a5e3d45e6d7c0c

Observation 2b55436d-9e9f-463b-b3ca-c549ba681931 · inbound

Guidance Contrastive Token Credit Assignment for Discrete Policy Optimization cites this paper.

Guidance Contrastive Token Credit Assignment for Discrete Policy Optimization MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-06-29T09:03:15.841633Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-06-29T08:59:03.375697Z digest=sha256:5468f30c1516645aeab4a9e2015ca9cd43cd243a5d00d341d97804a9442447d5

Observation 90ac795b-396c-40be-ad33-9297b3c7df41 · inbound

TRON: Targeted Rule-Verifiable Online Environments for Visual Reasoning RL cites this paper.

TRON: Targeted Rule-Verifiable Online Environments for Visual Reasoning RL MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-07-01T22:56:19.941500Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-06-28T14:55:07.045625Z digest=sha256:3e1264c59fa3789f19de8028fca60cb2163d13c16cbee3290aaf27b4f1be38cb

Observation 5131f710-9667-42a7-8666-a27d01e7ad09 · inbound

Do Multimodal Agents Really Benefit from Tool Use? A Systematic Study of Capability Gains cites this paper.

Do Multimodal Agents Really Benefit from Tool Use? A Systematic Study of Capability Gains MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-07-01T22:56:19.492685Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-06-28T14:58:05.652366Z digest=sha256:c1d78db6f31e709b887dfdbc5e7e17229f74d882fe57c6ad92f0cae708639a8c

Observation 32ba73ab-d104-43d0-babd-63afae66ff27 · inbound

Thinking Past the Answer: Evaluating Harmful Overthinking in Large Reasoning Models cites this paper.

Thinking Past the Answer: Evaluating Harmful Overthinking in Large Reasoning Models MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 23

Resolution
verified exact
local_arxiv, observed 2026-07-01T23:36:22.521589Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-06-28T14:14:42.351048Z digest=sha256:770a1b5b0a7d4e9d6c7dd9e0c26a14da5df3308617efb2e21df9f440f0b5f8d1

Observation 4229c721-5e93-4af6-b054-a9275174b9c2 · inbound

Entropy Is Not Enough: Unlocking Effective Reinforcement Learning for Visual Reasoning via Vision-Anchored Token Selection cites this paper.

Entropy Is Not Enough: Unlocking Effective Reinforcement Learning for Visual Reasoning via Vision-Anchored Token Selection MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 5

Resolution
metadata mismatch
local_arxiv, observed 2026-07-02T03:26:29.896019Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-06-28T09:58:34.557353Z digest=sha256:7685fcb230544eb2dcb0bc970ec00d81aa4bb4dc33c2297605aa3913c0d71b71

Observation 7d3f8c16-8239-42bc-9276-d402f12ff860 · inbound

Teaching the Way, Not the Answer: Privileged Tutoring Distillation for Multimodal Policy Optimization cites this paper.

Teaching the Way, Not the Answer: Privileged Tutoring Distillation for Multimodal Policy Optimization MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-07-02T19:07:17.958313Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-06-27T21:40:02.133241Z digest=sha256:f9b7492b774cfcc3362bc9cdb2cbfb4be3966d2223b9e395697e7b183731261d

Observation 0da650e3-2895-47db-aac8-16f5ee3933fc · inbound

Artificial Intelligence for Mathematical Reasoning: An Integrated Survey of Language Models, Neuro-symbolic Systems, and Verified Discovery cites this paper.

Artificial Intelligence for Mathematical Reasoning: An Integrated Survey of Language Models, Neuro-symbolic Systems, and Verified Discovery MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 265

Resolution
verified exact
local_arxiv, observed 2026-07-02T22:47:26.085491Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-06-27T18:39:44.696961Z digest=sha256:be9b123b8b2173cbbb194852f04c2ff2ce0b267ca0e19a7e1040c434c257efcc

Observation 090470bf-19d1-4688-8b40-8bf67ea1ad39 · inbound

Are Reasoning Vision-Language Models Robust to Semantic Visual Distractions? cites this paper.

Are Reasoning Vision-Language Models Robust to Semantic Visual Distractions? MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 8

Resolution
metadata mismatch
local_arxiv, observed 2026-07-02T23:57:28.900713Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=arxiv_source observed=2026-06-27T17:36:42.185657Z digest=sha256:04bba2cf6a7bc23c1e80ed0c7ffa2399f28585bdfa5f7499eedbaa0a99c8302e

Observation 2fb5c0c7-2e0e-4299-8956-adb5b2924553 · inbound

From Shortcuts to Reasoning: Robust Post-Training of Theory of Mind with Reinforcement Learning cites this paper.

From Shortcuts to Reasoning: Robust Post-Training of Theory of Mind with Reinforcement Learning MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 73

Resolution
metadata mismatch
local_arxiv, observed 2026-07-02T23:57:28.168303Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=arxiv_source observed=2026-06-27T17:42:38.122144Z digest=sha256:163a2d4dc195d3488e4c1b3008e2963fc063067a6d530b15bc7a493356ee2e83

Observation c4b25dfe-6d47-4d5f-ba36-57f5a5b40810 · inbound

Optical Reasoning: Rethinking Images as an Expressive Reasoning Medium Beyond Text cites this paper.

Optical Reasoning: Rethinking Images as an Expressive Reasoning Medium Beyond Text MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 5

Resolution
verified exact
local_arxiv, observed 2026-07-03T01:17:30.472048Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-06-27T16:45:17.380015Z digest=sha256:7d752baf8e42963a710e30d7348ec5f77b4e5f0abc6a65977d970d1e00da168c

Observation 039986df-7720-4119-b96a-909484917f2e · inbound

Kwai Keye-VL-2.0 Technical Report cites this paper.

Kwai Keye-VL-2.0 Technical Report MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 38

Resolution
metadata mismatch
local_arxiv, observed 2026-07-03T04:27:37.089298Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=arxiv_source observed=2026-06-27T13:53:10.352603Z digest=sha256:6ed6e59c54f6b6489408c364ad03c4c2adc557048ef2e471f8570f01fbaa2d8b

Observation 652d7e85-fd80-4b70-823e-57e1dbbd0806 · inbound

Reinforcing Dual-Path Reasoning in Spatial Vision Language Models cites this paper.

Reinforcing Dual-Path Reasoning in Spatial Vision Language Models MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 93

Resolution
metadata mismatch
local_arxiv, observed 2026-07-03T20:08:55.656797Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=arxiv_source observed=2026-06-27T01:42:30.005911Z digest=sha256:427d4c413919cb39f5af128cdb1c4afbdbc750f9205b27bb3cc46eb6a944e750

Observation a616722b-3732-417f-9d41-f7d3b8af0923 · inbound

MathVis-Fine: Aligning Visual Supervision with Necessity via Progressive Dependency-Guided Training for Multimodal Mathematical Reasoning cites this paper.

MathVis-Fine: Aligning Visual Supervision with Necessity via Progressive Dependency-Guided Training for Multimodal Mathematical Reasoning MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 50

Resolution
metadata mismatch
local_arxiv, observed 2026-07-03T20:18:57.806657Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=arxiv_source observed=2026-06-27T01:23:40.564561Z digest=sha256:f99c0e0f60bc28b759f3313bc198c04ec1c6ab6f964549730e9cc702cae0b12f

Observation 00411777-b82a-4fa1-a992-d5a2b3500713 · inbound

Zone of Proximal Policy Optimization: Teacher in Prompts, Not Gradients cites this paper.

Zone of Proximal Policy Optimization: Teacher in Prompts, Not Gradients MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 101

Resolution
verified exact
local_arxiv, observed 2026-07-03T20:48:56.197288Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-06-27T01:08:52.981296Z digest=sha256:b810b1aaf2209c31ecb93fc6bfe1d2005c3518fcbf67130e1c1d769455f4e163

Observation a1d49a41-fc77-4f96-a84a-0e1b55d15284 · inbound

VeriEvol: Scaling Multimodal Mathematical Reasoning via Verifiable Evol-Instruct cites this paper.

VeriEvol: Scaling Multimodal Mathematical Reasoning via Verifiable Evol-Instruct MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 27

Resolution
verified exact
local_arxiv, observed 2026-07-04T10:39:46.071297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-06-26T08:34:27.719022Z digest=sha256:30e85d97bf7d13d48e824452aa5684631712d73fea67a2d8c1d1d06bbcb18c9e

Observation f3b8c087-6610-40b7-82ff-53746ef56cb9 · inbound

AIR: Adaptive Interleaved Reasoning with Code in MLLMs cites this paper.

AIR: Adaptive Interleaved Reasoning with Code in MLLMs MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 3

Resolution
metadata mismatch
local_arxiv, observed 2026-07-04T10:09:44.994232Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-06-26T09:06:38.001604Z digest=sha256:756db19fb372cd6fabf99e55b249053bfa1f7db692caeeb2b588ef16caae7475

Observation 8e4a0d87-8533-4ee1-bb61-02e1eaabc64a · inbound

Latent Visual States for Efficient Multimodal Reasoning cites this paper.

Latent Visual States for Efficient Multimodal Reasoning MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 27

Resolution
verified exact
local_arxiv, observed 2026-07-04T16:29:57.257089Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-06-26T00:38:11.619574Z digest=sha256:912bb7dda5b531e3820e4136268de977c70b66fb51111d37c4f2da1d947078e5

Observation ac036d50-75a8-49de-8013-edbd21aac58e · inbound

Omni-Perception Policy Optimization for Multimodal Emotion Reasoning cites this paper.

Omni-Perception Policy Optimization for Multimodal Emotion Reasoning MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 32

Resolution
metadata mismatch
local_arxiv, observed 2026-07-04T19:20:06.479816Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=arxiv_source observed=2026-06-25T21:31:38.450382Z digest=sha256:c48b7cd8f978dbc5ef10e8154fc85cc07abfc487662f1b1b86f7207cb82d0417

Observation 14bfe3fd-5aa8-4dd0-98d8-9b993342d188 · inbound

MER-R1: Multimodal Emotion Reasoning via Slow-Fast Thinking Synergy cites this paper.

MER-R1: Multimodal Emotion Reasoning via Slow-Fast Thinking Synergy MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-06-29T18:23:51.426185Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-06-29T05:06:09.216428Z digest=sha256:3a47173a2743255c4e94d63a72ae3712ad4a2f281cff91de41cb960f894250dd

Observation 7d257338-c638-4c95-9cbe-fce05404ceeb · inbound

Consistency as Inductive Bias: Learning Cross-View Invariance for Robust Multimodal Reasoning cites this paper.

Consistency as Inductive Bias: Learning Cross-View Invariance for Robust Multimodal Reasoning MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 27

Resolution
verified exact
local_arxiv, observed 2026-06-30T06:14:19.274647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-06-30T06:10:41.786328Z digest=sha256:83f5a523641a2f3a7ac3dae0f3fa6f3b309f47475cef8f1fbe045b9061804c87

Observation fd45020e-2810-47af-9660-11e716c0f1af · inbound

Before Thinking, Learn to Decide: Proactive Routing for Efficient Visual Reasoning cites this paper.

Before Thinking, Learn to Decide: Proactive Routing for Efficient Visual Reasoning MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 34

Resolution
metadata mismatch
local_arxiv, observed 2026-06-30T13:34:41.165153Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-06-30T05:55:19.083517Z digest=sha256:760a67242367fb22357fab0188dbb7e98539c10122382d8826fc19efea9879ef

Observation 390bc795-91fa-4f65-be3b-f83c0c417a6c · inbound

LASER: A Corrective Lens for LVLMs via Visual Attention Preservation and Sink Suppression cites this paper.

LASER: A Corrective Lens for LVLMs via Visual Attention Preservation and Sink Suppression MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 41

Resolution
metadata mismatch
local_arxiv, observed 2026-07-03T16:58:42.471872Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-07-03T16:53:43.427625Z digest=sha256:d248deaa016ea97f33c6db6e64c2be172f69613e6d036ed00b1fc9f8a03c1491

Observation 3998862a-bbb5-4c75-beea-34963e447f6d · inbound

DELTAVID: Enhancing Fine-Grained Spatiotemporal Perception with Cross-Video Differences cites this paper.

DELTAVID: Enhancing Fine-Grained Spatiotemporal Perception with Cross-Video Differences MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 37

Resolution
unresolved
no resolver link, observed 2026-07-12T11:31:14.532101Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T11:31:14.532101Z digest=sha256:373488d51f182f1c722d8d3d4c2e833b4c44f83eabf469edd972b07c3e8b2586

Observation b2b5c310-7b30-43ad-966b-f420ac2204fe · inbound

MentalThink: Shaping Thoughts in Mental SVG World cites this paper.

MentalThink: Shaping Thoughts in Mental SVG World MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 299

Resolution
unresolved
no resolver link, observed 2026-07-12T01:50:59.184754Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-12T01:50:59.184754Z digest=sha256:49a0183b2065e0cf29826be9382d081d4cc5ea42f11af6342aaefd031f874c0a

Observation a97e6e85-043e-48ea-b7ae-ff42b99c110d · inbound

Bridging Interleaved Multi-Modal Reasoning as a Unified Decision Process cites this paper.

Bridging Interleaved Multi-Modal Reasoning as a Unified Decision Process MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-07-12T00:14:19.104498Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:14:19.104498Z digest=sha256:baf00ba94b7ec70697c802a4a8e48a90d9f3b7884b16854804e6cc24a1c44f15

Observation b600dee4-f4a6-4258-9d8c-f2576be339ce · inbound

OpenCoF: Learning to Reason Through Video Generation cites this paper.

OpenCoF: Learning to Reason Through Video Generation MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 27

Resolution
verified exact
local_arxiv, observed 2026-07-10T01:46:40.973659Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-23T06:31:01.910684+00:00.

source=pdf_text observed=2026-07-10T01:43:37.265723Z digest=sha256:4e6260a00b433bf2cb257219747aa850e7a9f8efaff488ee79374618029001e3