Pith. sign in

Paper Citation Record · LEDGER

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs

As of 22 August 2026, this Paper Citation Record lists 96 of 96 outbound references and 13 inbound Pith citation observations for arXiv:2505.15804.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.15804 v3

Coverage vector

measured 96 of 96 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:15:42.409411Z

measured 109 of 109 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 13 of 13 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T23:21:13.340790Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T15:09:55.322058Z

Reference resolution

96 of 96 outbound references displayed

  • verified exact0
  • verified fuzzy6
  • unresolved90
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation dcdf50c0-212e-472b-95d0-367bc152ff3f · outbound

This paper cites Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T15:15:40.818259Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:15:40.818259Z digest=sha256:db72fa98c31ecf8fbc741b90b2d65d428cd8131ae817d136c7f65011efb3086d

Observation 841067a9-ed1a-46ba-a0ee-2455528740e4 · outbound

This paper cites Pixtral 12B.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Pixtral 12B

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T15:15:40.872572Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:15:40.872572Z digest=sha256:dced2c0145c2b7f810e46d656776285a1fc144bc437d9fad57333e77a8159df8

Observation ca00110c-34f1-4ad6-9275-c01dd8fc5f2f · outbound

This paper cites Qwen Technical Report.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Qwen Technical Report

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T15:15:40.949358Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:15:40.949358Z digest=sha256:dfbd8e860c0d7b171e24ea230f6c769224408c1ad4b1810c16c352a24096ada5

Observation ea2d13b7-7fab-409e-bd83-53ac10ced950 · outbound

This paper cites Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T15:15:41.008408Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:15:41.008408Z digest=sha256:7afc1c0f0a4d0fe110a05b19a10ff78db66444071306eebcb8d9e645c366b9ea

Observation 212cd02f-0c13-4f1d-adda-1d442496413e · outbound

This paper cites Qwen2.5-VL Technical Report.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Qwen2.5-VL Technical Report

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T15:15:41.082457Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:15:41.082457Z digest=sha256:a5419cf4d96c42b70743ae0db2e2345fe5d5cb7fb4d7e767c650bd3887bba879

Observation bd785b30-6e62-4a18-8f3d-bdbe30a01885 · outbound

This paper cites R1-v: Reinforcing super generaliza- tion ability in vision-language models with less than $3.https://github.com/Deep-Agent/ R1-V, 2025.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs R1-v: Reinforcing super generaliza- tion ability in vision-language models with less than $3.https://github.com/Deep-Agent/ R1-V, 2025

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T15:15:41.132137Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:15:41.132137Z digest=sha256:62fa3f655d85c3c1fda3e920d8230e603930539e44d3a7e107b3d61d3abbfd36

Observation c57b4766-fc0e-46f8-b0e5-4965091fba79 · outbound

This paper cites Are We on the Right Way for Evaluating Large Vision-Language Models?.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Are We on the Right Way for Evaluating Large Vision-Language Models?

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T15:15:41.240256Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:15:41.240256Z digest=sha256:2b2c53fc08623db15de21f2e224a242526fe9b1fdb44dc1e7efb550405afb92e

Observation 7fadc10a-2785-49c0-aedb-977e046774f4 · outbound

This paper cites Janus-Pro: Unified Multimodal Understanding and Generation with Data and Model Scaling.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Janus-Pro: Unified Multimodal Understanding and Generation with Data and Model Scaling

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T15:15:41.254728Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:15:41.254728Z digest=sha256:e24f69563c9426ccf8b6c88750cffa873b367af71a17a943e13ae16ae2f58582

Observation b1d57b38-53f5-49c4-a16d-20c220232ee3 · outbound

This paper cites Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T15:15:41.300013Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:15:41.300013Z digest=sha256:945fefa8bb70962583d18f9cc18127afc194afb2c4b9682df2584c28a45f1622

Observation 675891a1-c4a2-47a4-9012-e0581005d856 · outbound

This paper cites How far are we to gpt-4v? closing the gap to commercial multimodal models with open-source suites.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs How far are we to gpt-4v? closing the gap to commercial multimodal models with open-source suites

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T15:15:41.309847Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:15:41.309847Z digest=sha256:a8ac19ebcaab53e9b0f08f109a7249edb383821c8b39ff6fa757a4650b1b1d86

Observation a081f643-7ecc-473a-9653-2dedd8b9ab6a · outbound

This paper cites Internvl: Scaling up vision foundation models and aligning for generic visual-linguistic tasks.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Internvl: Scaling up vision foundation models and aligning for generic visual-linguistic tasks

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T15:15:41.364691Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:15:41.364691Z digest=sha256:12ed50189f9d86aed1ae301320ce93051eb81a1237c9f1a8ea7f03ca87ecb976

Observation 293ee2da-eb8e-4f1b-82cc-912464bb340f · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Training Verifiers to Solve Math Word Problems

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T15:15:41.409185Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:15:41.409185Z digest=sha256:663eb43a2cdb68116f10adc810a843dab772eec93ff0cabccc92a5a7ca7b479b

Observation a4f0255e-e449-42b3-8f9f-641679fe6bcd · outbound

This paper cites Sophiavl-r1: Reinforcing mllms reasoning with thinking reward.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Sophiavl-r1: Reinforcing mllms reasoning with thinking reward

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T15:15:41.466292Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:15:41.466292Z digest=sha256:b3b209b5d4bbe15bd6062a7908bba23919942dd48149f0fd11abab984ee28d22

Observation 1901736a-b61f-43ec-9751-af79e224f2da · outbound

This paper cites Video-R1: Reinforcing Video Reasoning in MLLMs.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Video-R1: Reinforcing Video Reasoning in MLLMs

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T15:15:41.500371Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:15:41.500371Z digest=sha256:b843d4723323d328e9b1fd4eb64b9e5c29fa27589e50108c546fedd3c2536a1d

Observation 50844e53-ff55-471e-b50c-1fd8da55cd7c · outbound

This paper cites Video-MME: The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video Analysis.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Video-MME: The First-Ever Comprehensive Evaluation Benchmark of Multi-modal LLMs in Video Analysis

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T15:15:41.535470Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:15:41.535470Z digest=sha256:106ddcfe5a0428c659fce6a9b8e4cbf4965ed0a891611e0499f65ac915f0588a

Observation 5d73b1f5-bc43-45ea-8adc-38778701ecfd · outbound

This paper cites MME-Survey: A Comprehensive Survey on Evaluation of Multimodal LLMs.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs MME-Survey: A Comprehensive Survey on Evaluation of Multimodal LLMs

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T15:15:41.573252Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:15:41.573252Z digest=sha256:045a708c61c0d99ff733fca058d53ea5c451dd3147f907063f0a799b95f70ae5

Observation f68052e0-3672-428a-a044-8e74c7dce9f9 · outbound

This paper cites Geneval: An object-focused framework for evaluating text-to-image alignment.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Geneval: An object-focused framework for evaluating text-to-image alignment

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T15:15:41.605318Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:15:41.605318Z digest=sha256:9b8f1053f11e85078cba87f436411eda5994d93617d6b2001356ad7b88d26900

Observation 37009068-0b55-4422-bbfa-5d26f070f066 · outbound

This paper cites The Llama 3 Herd of Models.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs The Llama 3 Herd of Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T15:15:41.644523Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:15:41.644523Z digest=sha256:fe5858ba0a66250e8f67ed1832a7b948d27b3c7db8fbb2c1f527d4f8ff4c4e79

Observation 9e9b71cb-00c7-43cd-b130-f5b43043dd0e · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T15:15:41.685199Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:15:41.685199Z digest=sha256:bfc47b6328fac1243a7158933c5966cf5cde8aa85388d6f4cd335f6dba83942d

Observation a7e2a54a-18fa-44ae-8c03-ba17156ca081 · outbound

This paper cites Measuring Massive Multitask Language Understanding.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Measuring Massive Multitask Language Understanding

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T15:15:41.719514Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:15:41.719514Z digest=sha256:a22ddf16c924e52ef7bdd2b2b0dcc53db82450afacb4d5b45d0b07ca0194cda7

Observation 60642aa0-19f7-4be2-9edb-31f111399fa1 · outbound

This paper cites Measuring Mathematical Problem Solving With the MATH Dataset.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Measuring Mathematical Problem Solving With the MATH Dataset

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T15:15:41.755343Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:15:41.755343Z digest=sha256:ea0667c4630a98ff3d99e29902396590eb23a433bf08d9014c6a145271ac212d

Observation 31f0ea88-2bad-4749-9571-c93c37a31657 · outbound

This paper cites Transformation driven visual reasoning.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Transformation driven visual reasoning

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T15:15:41.794764Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:15:41.794764Z digest=sha256:6154f68f3ba840e29e28c759404102a0ccf5a274dacfae207798bf458d14b680

Observation 917f6bee-c864-4648-aedc-76a5bb7d8451 · outbound

This paper cites ELLA: Equip Diffusion Models with LLM for Enhanced Semantic Alignment.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs ELLA: Equip Diffusion Models with LLM for Enhanced Semantic Alignment

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T15:15:41.837834Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:15:41.837834Z digest=sha256:cee455bcfecb0b5a7dfecb2bd3a3930fc61fe5b16638372622297fda03e2cd8c

Observation f85bfca2-c3c0-4ca3-addc-00913ef40ebb · outbound

This paper cites GPT-4o System Card.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs GPT-4o System Card

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T15:15:41.877810Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:15:41.877810Z digest=sha256:40ea201468e8edcd7693260d6bdb39aa9fead95122d0bb2f67ee694dab31129d

Observation d30f8d17-a1a3-411d-9322-29d2704458c1 · outbound

This paper cites Omnispatial: Towards comprehensive spatial reasoning benchmark for vision language models.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Omnispatial: Towards comprehensive spatial reasoning benchmark for vision language models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T15:15:41.919164Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:15:41.919164Z digest=sha256:c112d4493e8b2145b7626b282b311f7c97b876576bad1ef7ba7bbe430d52b5a8

Observation 20186f00-b6df-410e-90fb-6bb84a7c64b3 · outbound

This paper cites Gonzalez, Hao Zhang, and Ion Stoica.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Gonzalez, Hao Zhang, and Ion Stoica

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T15:15:41.956284Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:15:41.956284Z digest=sha256:eb8ea84b2e43da189c94a99647a8e01ebb263b8192e155b93a436d4823b8ca0b

Observation 0a109d7e-59ea-4cba-8521-f65ccbe551f3 · outbound

This paper cites LLaVA-OneVision: Easy Visual Task Transfer.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs LLaVA-OneVision: Easy Visual Task Transfer

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T15:15:42.027467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:15:42.027467Z digest=sha256:e7a89c09dd809977fddee8aa785f1e6895bd6f4bc581f08e734876e765f96096

Observation 5e666ddb-be36-425a-86fd-cf359b6216de · outbound

This paper cites LLaVA-NeXT-Interleave: Tackling Multi-image, Video, and 3D in Large Multimodal Models.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs LLaVA-NeXT-Interleave: Tackling Multi-image, Video, and 3D in Large Multimodal Models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T15:15:42.075306Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:15:42.075306Z digest=sha256:c612a26fb4a2bf9480706f19b9ba2a35a436f3544646bff43d04831c3dc2132c

Observation d582b926-e633-4030-af2b-ea3b7dd4edb3 · outbound

This paper cites Mvbench: A comprehensive multi-modal video understanding benchmark.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Mvbench: A comprehensive multi-modal video understanding benchmark

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T15:15:42.096080Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:15:42.096080Z digest=sha256:b7965369439eb29181f3f1ebc9876f1431ba3913f5d2c227c2d9e53c6cc50d3d

Observation caddb07a-2955-4f0f-b28c-415281230c70 · outbound

This paper cites Temporal Sampling for Forgotten Reasoning in LLMs.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Temporal Sampling for Forgotten Reasoning in LLMs

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T15:15:42.129693Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:15:42.129693Z digest=sha256:09f6a3211e389dc6a8b2f87b8c4f20b757684d2b0f076b632ce5d56cc7a56c80

Observation fa3718f8-5a0e-4349-bf35-85bc59e9dbb5 · outbound

This paper cites Sws: Self-aware weakness-driven problem synthesis in reinforcement learning for llm reasoning, 2025.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Sws: Self-aware weakness-driven problem synthesis in reinforcement learning for llm reasoning, 2025

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T15:15:42.134756Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:15:42.134756Z digest=sha256:23ca3463bf35a45cfbcefe473972bb371d800b86f3a643721cce70bada776ddc

Observation 44d8611a-5f30-48b5-bde4-b16373eea687 · outbound

This paper cites Improved baselines with visual instruction tuning.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Improved baselines with visual instruction tuning

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T15:15:42.139301Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:15:42.139301Z digest=sha256:60bdd47796c9214b3954dce304a9c67199292fee9b62a4666927c5d0a26ce666

Observation eec8507e-f0c7-4735-a8fe-825ecadb582b · outbound

This paper cites Visual instruction tuning.Advances in neural information processing systems, 36:34892–34916, 2023.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Visual instruction tuning.Advances in neural information processing systems, 36:34892–34916, 2023

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T15:15:42.143883Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:15:42.143883Z digest=sha256:db6d31fa9c5a0f68f7c4f230707d5898fd4d6daa0ac18472aa4bf603d80e7363

Observation 58f1947f-c76c-4444-8630-cb9fb76e6631 · outbound

This paper cites IR3D-Bench: Evaluating Vision-Language Model Scene Understanding as Agentic Inverse Rendering.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs IR3D-Bench: Evaluating Vision-Language Model Scene Understanding as Agentic Inverse Rendering

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T15:15:42.147646Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:15:42.147646Z digest=sha256:1d0583e393bef8a900dc835d658cd65f77df835c566784dd9eb7c8d6911223d1

Observation fdf7e325-6210-4536-bd71-a9c819e8432e · outbound

This paper cites DeepSeek-VL: Towards Real-World Vision-Language Understanding.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs DeepSeek-VL: Towards Real-World Vision-Language Understanding

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T15:15:42.152250Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:15:42.152250Z digest=sha256:ab632f481912b5f46d8294e1c07e9860d56a1210adc279178f655f5bf5e7eb94

Observation 25d24bf1-19d6-4345-88ee-c4c100b35c1f · outbound

This paper cites Egoschema: A diagnostic benchmark for very long-form video language understanding.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Egoschema: A diagnostic benchmark for very long-form video language understanding

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T15:15:42.156102Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:15:42.156102Z digest=sha256:063cb3cf2a402da62d7c1a97d45bf3a4654a7bc29008d6688a69ebf19b8b3039

Observation 0c8831b1-3d61-4d0e-8dc8-4c15ce522f78 · outbound

This paper cites Docvqa: A dataset for vqa on document images.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Docvqa: A dataset for vqa on document images

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T15:15:42.160189Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:15:42.160189Z digest=sha256:f6793a444a7310404dda8310e0b79a6bdfaf25fcfd56eb29823ccbe171f5832c

Observation cb33604e-ab4d-4479-992f-5a4abef9a15e · outbound

This paper cites ivispar– an interactive visual-spatial reasoning benchmark for vlms.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs ivispar– an interactive visual-spatial reasoning benchmark for vlms

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T15:15:42.164590Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:15:42.164590Z digest=sha256:080833163ab827b01e136f94995391ae3ec0f32a6a1b0bb1ac136468d7d5eb97

Observation f5b910e8-6cbf-4963-97ab-65f3af45f796 · outbound

This paper cites MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs MM-Eureka: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T15:15:42.168926Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:15:42.168926Z digest=sha256:31ba7a12b291bd667f544683f1836cf62b1afed1bf6a580a6144baaefc1fde11

Observation e0443644-5b99-45d8-aca1-fcc9c1fdd016 · outbound

This paper cites LMM-R1: Empowering 3B LMMs with Strong Reasoning Abilities Through Two-Stage Rule-Based RL.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs LMM-R1: Empowering 3B LMMs with Strong Reasoning Abilities Through Two-Stage Rule-Based RL

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T15:15:42.173319Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:15:42.173319Z digest=sha256:3c23a7ca3b37d4e23b0afa268050bf509c036404fc08dd2c818b8e38a91fbdd7

Observation 2bf1dcc4-e754-41a0-83dd-12fce5eed07f · outbound

This paper cites Learning transferable visual models from natural language supervision.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Learning transferable visual models from natural language supervision

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T15:15:42.177791Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:15:42.177791Z digest=sha256:6ec21b810fc3166ef1367cb60096e8d73d7f210a9f7de68757ecb70ee5f7fc5d

Observation 735331ce-7784-402a-accb-e2dadbf6ddb1 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T15:15:42.182014Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:15:42.182014Z digest=sha256:c73dc22c4b933fd93ead1475525ba2862a342e30f36729e5dea18576f1010af2

Observation c8e46b05-bd85-4e0f-b50b-6367451fa9d2 · outbound

This paper cites VLM-R1: A Stable and Generalizable R1-style Large Vision-Language Model.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs VLM-R1: A Stable and Generalizable R1-style Large Vision-Language Model

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T15:15:42.185632Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:15:42.185632Z digest=sha256:21cc79a4a4395e99ac3ba47dd172dfb2a754a44ad4e955234ef463f965c6cf2c

Observation dd0adcd0-9995-410b-bd76-a4e7d4c46316 · outbound

This paper cites Maniplvm-r1: Reinforcement learning for reasoning in embodied manipulation with large vision-language models, 2025.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Maniplvm-r1: Reinforcement learning for reasoning in embodied manipulation with large vision-language models, 2025

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T15:15:42.189660Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:15:42.189660Z digest=sha256:fad28eba12f6a050e7608c7c6b3c644764da629528c3e04a2104e1c7a7f1a3d5

Observation e12c1220-c471-427b-a623-b9aab9688493 · outbound

This paper cites Generative multimodal models are in-context learners.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Generative multimodal models are in-context learners

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T15:15:42.193280Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:15:42.193280Z digest=sha256:4e02ed89da6118f3cdee2b2f41475e5bbca1a5c34be720452c4ebd302754964b

Observation 1cb7b0e2-9c83-43a7-9cce-1d876cd68385 · outbound

This paper cites EVA-CLIP: Improved Training Techniques for CLIP at Scale.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs EVA-CLIP: Improved Training Techniques for CLIP at Scale

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T15:15:42.196929Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:15:42.196929Z digest=sha256:801c628aaa5c6065a47c743ad456df609cd7441cfb8bbb176ee53cf6cf7d2a61

Observation bcf325dd-e680-401c-9b78-362e0746a881 · outbound

This paper cites LEGO-Puzzles: How Good Are MLLMs at Multi-Step Spatial Reasoning?.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs LEGO-Puzzles: How Good Are MLLMs at Multi-Step Spatial Reasoning?

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T15:15:42.201162Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:15:42.201162Z digest=sha256:210d717ee074b0553617141c3f479d23a21f98095a745fd1560a0e4f214b3f75

Observation 44a1aefd-278f-4292-b814-a934fda0d30e · outbound

This paper cites Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T15:15:42.205991Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:15:42.205991Z digest=sha256:8c963ae39fa79d54357dc5c82a0fce30a38760ead8af6fdf6f72103b501ee0d3

Observation 4d813ab3-513b-41b2-98d7-a677c7f9c0d0 · outbound

This paper cites Internvl2: Better than the best—expanding performance boundaries of open-source multimodal models with the progressive scaling strategy, 2024.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Internvl2: Better than the best—expanding performance boundaries of open-source multimodal models with the progressive scaling strategy, 2024

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T15:15:42.210026Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:15:42.210026Z digest=sha256:bbfca0d43af66524c7dec7c38dcb509f014403c4443936f5aecaf9c76c161b63

Observation 8b76ae6b-8b34-42b9-8a24-c21f19a89c97 · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs LLaMA: Open and Efficient Foundation Language Models

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T15:15:42.214143Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:15:42.214143Z digest=sha256:733060e8090d759f07cbe88eb7e7d45d4e9d01846c8a60a1f2d2f4e386f378ac

Observation 95cbe8c0-37a0-440d-834d-f0fb00ad8cf2 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T15:15:42.218603Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:15:42.218603Z digest=sha256:effa2c5f9a7eee132d3cac65c6d469df984041a064a5e5798fa35897e204e4a1

Observation 489e5f84-5ebd-474e-ba7e-77519bc44865 · outbound

This paper cites Solidgeo: Measuring multimodal spatial math reasoning in solid geometry, 2025.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Solidgeo: Measuring multimodal spatial math reasoning in solid geometry, 2025

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:15:43.631470Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T15:15:42.222331Z digest=sha256:cc091a24c6a0d1b9bcc6386125d74c7d937a024e8ce49419c90137643df71968

Observation dc84bc8d-8da3-4aa6-a097-e5ddbc760436 · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T15:15:42.226384Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:15:42.226384Z digest=sha256:e46da82a81f2804d98afde02398aff4a755a235cbab1265ddd4849c133275757

Observation 444d3501-957b-412e-94e2-d863ecb8ec32 · outbound

This paper cites LVBench: An Extreme Long Video Understanding Benchmark.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs LVBench: An Extreme Long Video Understanding Benchmark

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-07T15:15:42.230745Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:15:42.230745Z digest=sha256:2d132dcd07a1eaea12f60705412540946e0323f54c259fbcb89ed83c5346af4a

Observation 232b0cc2-325e-4178-9a77-6c71c261ab59 · outbound

This paper cites Emu3: Next-Token Prediction is All You Need.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Emu3: Next-Token Prediction is All You Need

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T15:15:42.234497Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:15:42.234497Z digest=sha256:25f83faac72b488db07f6050e0d3945b9c9dd47e734e451da45502049ed56861

Observation b6686393-a273-418f-857c-304c2427e61a · outbound

This paper cites Chain-of-thought prompting elicits reasoning in large language models.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Chain-of-thought prompting elicits reasoning in large language models

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-07T15:15:42.243110Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:15:42.243110Z digest=sha256:be006bb42a620d7f3f0f77a02e10f6bd91beb21d30467b2b365babd50586a62e

Observation 164c55ef-ee9c-4dad-b383-3b7f377d05c5 · outbound

This paper cites CMATH: Can Your Language Model Pass Chinese Elementary School Math Test?.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs CMATH: Can Your Language Model Pass Chinese Elementary School Math Test?

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-07T15:15:42.248418Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:15:42.248418Z digest=sha256:d71ba4fdea67fdc04ac4b3fe4b6bae97aead3b331ee39f74ec6038a43d670310

Observation 0d081f85-7c69-48b1-8c65-71e19db85da3 · outbound

This paper cites Logic-RL: Unleashing LLM Reasoning with Rule-Based Reinforcement Learning.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Logic-RL: Unleashing LLM Reasoning with Rule-Based Reinforcement Learning

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-07T15:15:42.253671Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:15:42.253671Z digest=sha256:21721060aebee2dedd30b3ebe7d2280e58836c2395f8a324d048647eca715466

Observation e8c930ac-37a6-4783-9ef2-ee21b8180066 · outbound

This paper cites LLaVA-CoT: Let Vision Language Models Reason Step-by-Step.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs LLaVA-CoT: Let Vision Language Models Reason Step-by-Step

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-07T15:15:42.258330Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:15:42.258330Z digest=sha256:12f1e88c82afc433bce39a643bd31952b5d124ed68854a35fec5768afd032642

Observation 51606719-fb28-47d4-ad15-36ff7ef06158 · outbound

This paper cites RedStar: Does Scaling Long-CoT Data Unlock Better Slow-Reasoning Systems?.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs RedStar: Does Scaling Long-CoT Data Unlock Better Slow-Reasoning Systems?

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-07T15:15:42.262101Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:15:42.262101Z digest=sha256:531127722134a6f70181e431afee876fbc0bccd661f8348959a03d8c21e0cace

Observation ddc43479-a4c1-4e89-9009-4e231dfb0a77 · outbound

This paper cites Qwen2.5 Technical Report.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Qwen2.5 Technical Report

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-07T15:15:42.266320Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:15:42.266320Z digest=sha256:f0e2ef232faba5e81e65c35066904865f25a2bccdec15c633bfc22c8804c55b4

Observation a6d0a0dc-20a9-4b37-b7ad-11ab42174270 · outbound

This paper cites DeepCritic: Deliberate Critique with Large Language Models.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs DeepCritic: Deliberate Critique with Large Language Models

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-07T15:15:42.269555Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:15:42.269555Z digest=sha256:2b750d8f5ca1b04092ee36c4c439665bf3e67918ce60a5fb6bb3bce5e4e3054c

Observation 35e4b4e5-238b-4731-a32d-1a80201f4aea · outbound

This paper cites Towards thinking-optimal scaling of test-time compute for llm reasoning.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Towards thinking-optimal scaling of test-time compute for llm reasoning

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-07T15:15:42.273216Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:15:42.273216Z digest=sha256:1969d40a396bea6ae72569b3c43f2b57fe1b9fcd34a6dc3cba56fb1cc5e17d61

Observation 901e1fc9-dabd-4011-8231-b168b2820735 · outbound

This paper cites MiniCPM-V: A GPT-4V Level MLLM on Your Phone.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs MiniCPM-V: A GPT-4V Level MLLM on Your Phone

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-07T15:15:42.281622Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:15:42.281622Z digest=sha256:94f12b5ec40969bf38e75ac24fd02d1ee1c2478ec2bc34ce880d922058c42c89

Observation e23c3376-8607-4cfd-b530-4cdb6610b899 · outbound

This paper cites mPLUG-Owl3: Towards Long Image-Sequence Understanding in Multi-Modal Large Language Models.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs mPLUG-Owl3: Towards Long Image-Sequence Understanding in Multi-Modal Large Language Models

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-07T15:15:42.285220Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:15:42.285220Z digest=sha256:491be86cedb90bc5df12962bc7172e34e06e57eb14e537a714280b9b2779feaa

Observation 519b80c0-16a8-4165-b300-b17059383537 · outbound

This paper cites MM-Vet: Evaluating Large Multimodal Models for Integrated Capabilities.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs MM-Vet: Evaluating Large Multimodal Models for Integrated Capabilities

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-07T15:15:42.289839Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:15:42.289839Z digest=sha256:84f660e584669c5146e91dc8fe5c1b3a4e5d05bcf3b569942e99bbee4964ffa2

Observation e1520b89-e6ad-4e5d-95f4-2128216ee265 · outbound

This paper cites Mmmu: A massive multi-discipline multimodal understanding and reasoning benchmark for expert agi.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Mmmu: A massive multi-discipline multimodal understanding and reasoning benchmark for expert agi

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-07T15:15:42.293769Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:15:42.293769Z digest=sha256:751a909d02ca1d8d13bfa3cee731cb52d9a9392354646c77ebe180e8da259007

Observation ba8898d5-c2c4-4162-b806-bda2c55114b5 · outbound

This paper cites AGIEval: A Human-Centric Benchmark for Evaluating Foundation Models.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs AGIEval: A Human-Centric Benchmark for Evaluating Foundation Models

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-07T15:15:42.297149Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:15:42.297149Z digest=sha256:baa5141900fb655c18c927f093053a89b4c614bc46f69e349e3e94c1693d69aa

Observation 4d24babf-3c90-4ee8-8d36-656b4965946f · outbound

This paper cites MLVU: Benchmarking Multi-task Long Video Understanding.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs MLVU: Benchmarking Multi-task Long Video Understanding

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-07T15:15:42.301134Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:15:42.301134Z digest=sha256:e4f7fba26c8dcae760bb12d8965c838a55f564b6df9de5af89a405575dd4028b

Observation f67027c9-4458-48db-9a08-8c0e0be4206f · outbound

This paper cites InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-07T15:15:42.305127Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:15:42.305127Z digest=sha256:7157577c6531c5a702a61d9c7a99c9806fd1073fb287b78279992a0790654d3d

Observation 096c5082-9f0f-4141-9b5d-e7cca97dc84f · outbound

This paper cites We employ Qwen2.5-VL-7B as our base model and utilize vLLM [26] as the training framework.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs We employ Qwen2.5-VL-7B as our base model and utilize vLLM [26] as the training framework

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:15:43.605647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T15:15:42.309967Z digest=sha256:caf7afdae8a1e6530c6bab80edc35b30aa738e8eef57b7f0022e0a59c6cc0488

Observation 244a5890-5c75-4fe9-899d-5cf984edde73 · outbound

This paper cites cube", "sphere.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs cube", "sphere

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:15:43.592351Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T15:15:42.314386Z digest=sha256:bbe031502edc7468242d6f4199fe660d2eb1effaa4381d1660fba83cfc3685d3

Observation db0c6a97-a730-43a0-a38f-f16d909cd050 · outbound

This paper cites large" (radius = 6),.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs large" (radius = 6),

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:15:43.581594Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T15:15:42.318647Z digest=sha256:5cd63a081f9badcda1e88193e5cf68506d41d810dce4f938e538b84ac5d537e8

Observation 18b1c8d7-ee0a-4965-9833-d33fe965697d · outbound

This paper cites gray", "red.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs gray", "red

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:15:43.571255Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T15:15:42.322551Z digest=sha256:b34e02d51ccd1264821855956ae710d91e61d32cf6da33c44bd6cacbed4e3e12

Observation c323fc22-ff28-4c8c-91f5-c36482bf9bc0 · outbound

This paper cites rubber",.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs rubber",

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:15:43.560229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T15:15:42.326696Z digest=sha256:0fac4ef71df1fc3fe729a692b3942e1bfecc838a7917e69bddcc2d1ae37b1296

Observation 873d6753-d86f-4d12-84b2-2cbb8531affb · outbound

This paper cites an unresolved cited work.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Unresolved cited work

Reference 77

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:15:43.548789Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T15:15:42.331251Z digest=sha256:3352585f21ad54591127679fed5cc5d673df544313422a86b2781be829970093

Observation 11a8e450-aaaf-4057-9dbc-be23bb60b513 · outbound

This paper cites an unresolved cited work.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Unresolved cited work

Reference 78

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:15:43.537164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T15:15:42.335528Z digest=sha256:1317166b9db93a5f26ce687f4142dd81cd7d2be55226e1968aa9ed28fa0361c0

Observation 0f7bfbb6-b6d6-44cc-870d-d6db07df822b · outbound

This paper cites an unresolved cited work.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Unresolved cited work

Reference 79

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:15:43.525219Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T15:15:42.339882Z digest=sha256:011d333fade2f559f2cacdc1af13fdb7d695a69f1f04155f70735adc4bfae1c1

Observation 065dac49-2284-4996-aa2a-0163cf1d7964 · outbound

This paper cites an unresolved cited work.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Unresolved cited work

Reference 80

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:15:43.513574Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T15:15:42.344086Z digest=sha256:43d663d8886b69bc3f8e759f94bfcbc06e9006414a71e0c4fc02de5cffce706b

Observation 4a827d13-3c5d-461a-9659-77a7cccc9283 · outbound

This paper cites an unresolved cited work.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Unresolved cited work

Reference 81

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:15:43.496484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T15:15:42.347883Z digest=sha256:ed7f1a6f8ffe4127123a5feab3c982aa840574d54f0db51c7f409235075599c8

Observation e93abd86-47ce-4df2-8ffb-6a464d6f5781 · outbound

This paper cites an unresolved cited work.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Unresolved cited work

Reference 82

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:15:43.483698Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T15:15:42.351668Z digest=sha256:59018f29869ea41ffe2449fef0a2504df12e122d4c41a5d7a4ff2f96b4cb32ce

Observation e66144b6-0ba6-492e-b06e-c3d0a3b6f0b9 · outbound

This paper cites an unresolved cited work.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Unresolved cited work

Reference 83

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:15:43.471860Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T15:15:42.355256Z digest=sha256:0850a022e5976f245c6ad9b9f1ec5e7f941e94de8dea8a0dc3f0e69cb340aef0

Observation cfb56c9a-21dc-4c78-a0d7-f11722c27f22 · outbound

This paper cites an unresolved cited work.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Unresolved cited work

Reference 84

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:15:43.460565Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T15:15:42.359404Z digest=sha256:badae3258e98e04f4c5ca3358d22f95bcc68660e2310ed28656ee2758f0eea49

Observation ab463e98-fd8e-4f31-a007-2314901443fc · outbound

This paper cites an unresolved cited work.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Unresolved cited work

Reference 85

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:15:43.440298Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T15:15:42.363581Z digest=sha256:4b85891a87f1803a4b4302b11240735c073bf2cb26b43eedb9db4146e45b3d58

Observation 29dbecf5-bbf3-42fb-9721-69d2c6ec808e · outbound

This paper cites an unresolved cited work.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Unresolved cited work

Reference 86

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:15:43.426319Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T15:15:42.367119Z digest=sha256:8a006a9fce34fd8cd3b0efb5dd2fdf7ca79e577032b9a8ced2550efb58176955

Observation 61be8813-fb52-4503-92fa-cd11e7f434f1 · outbound

This paper cites an unresolved cited work.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Unresolved cited work

Reference 87

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:15:43.411313Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T15:15:42.370418Z digest=sha256:0f701405eb029212767aee45803ff1f69af0b23d76a3513986350dd6a949f7a4

Observation 183f12b8-9608-4408-9b86-3b50e85ac1a1 · outbound

This paper cites an unresolved cited work.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Unresolved cited work

Reference 88

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:15:43.397344Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T15:15:42.374440Z digest=sha256:af0195c5eb328f9ce790cf5374a3f766924623bd010cafa498cfea5bbc50f2e3

Observation d5820b93-e0f5-4ddf-a0a6-ed01e031abe1 · outbound

This paper cites an unresolved cited work.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Unresolved cited work

Reference 89

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:15:43.382882Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T15:15:42.378262Z digest=sha256:8a7626559f8a09a6b800b6470e1ceaa6929810dae4f4036e95bfa9b2edd10a09

Observation a87e34da-cf3f-4178-8d74-735548c44d4d · outbound

This paper cites an unresolved cited work.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Unresolved cited work

Reference 90

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:15:43.371422Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T15:15:42.381824Z digest=sha256:b487e8c7d256fc3e09827173cca3169080076186547d6ea72d5591494d863bd5

Observation 9c93ac4e-86e4-4858-8a80-58a2ab6fa953 · outbound

This paper cites an unresolved cited work.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Unresolved cited work

Reference 91

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:15:43.359597Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T15:15:42.386297Z digest=sha256:cd0c3f32fe87141f22a9bf714e0adf4040610ed2b560d777a61ffb504301a1b7

Observation ca901aa9-f866-445d-be0f-d3bf6cc630ed · outbound

This paper cites an unresolved cited work.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Unresolved cited work

Reference 92

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:15:43.347914Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T15:15:42.390123Z digest=sha256:113de3b4c61dc08abf0dd862a642053119e8fe24cee42e0940ca827a1e633451

Observation 808d1f98-d957-4809-8e18-8245b27918ed · outbound

This paper cites an unresolved cited work.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Unresolved cited work

Reference 93

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:15:43.334435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T15:15:42.393717Z digest=sha256:98a93a685d0e3b8b5fae09e95fcfdfd08180789bc8cb61f1ecda08637bdd0389

Observation 46eb0fbf-5a8a-4bda-aa91-b8eae4758204 · outbound

This paper cites an unresolved cited work.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Unresolved cited work

Reference 94

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:15:43.319013Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T15:15:42.397927Z digest=sha256:fec9a16b953850e3b4c980af1925d734370b0f3c64ce9bb3c9b97cfd0bd04c4d

Observation 20b72e1b-ea82-4c20-8442-c828b79e361e · outbound

This paper cites an unresolved cited work.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Unresolved cited work

Reference 95

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:15:43.307761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T15:15:42.402199Z digest=sha256:c16ce70d23bcbfcbab572a273b31be28ca0d90430112913605871693df2a8e71

Observation 9a894983-d8ce-402e-b519-68e6b63834c0 · outbound

This paper cites an unresolved cited work.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Unresolved cited work

Reference 96

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:15:43.294524Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T15:15:42.405954Z digest=sha256:2590737e9c117cd416a328fcc4af0ac11b2940ebc4cc730610b697a2bdf212f9

Observation 6335bd32-dda7-42fb-b6b2-ff2428698243 · outbound

This paper cites an unresolved cited work.

STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs Unresolved cited work

Reference 97

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:15:43.245130Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-07T15:15:42.409411Z digest=sha256:162afbc7ba19fefed10f888e8dfc103019a8bb9f267f7904adc2c845240378d5

Pith citing papers

Observation 4cb3a065-2600-492c-a285-090f70fc77d3 · inbound

From System 1 to System 2: A Survey of Reasoning Large Language Models cites this paper.

From System 1 to System 2: A Survey of Reasoning Large Language Models STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs

Reference 297

Resolution
verified exact
arxiv_id, observed 2026-05-13T01:36:24.521889Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-13T01:36:23.845366Z digest=sha256:f12b721b0a2e099a85b77f2da2a147e4b2aa76d659233ec0c548fb0287b19c7e

Observation cad82a6b-df4e-4c32-8a85-f002afae1123 · inbound

Video-R1: Reinforcing Video Reasoning in MLLMs cites this paper.

Video-R1: Reinforcing Video Reasoning in MLLMs STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-12T09:43:00.404330Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-12T09:43:00.208065Z digest=sha256:0500aa2f6e1a128d962d8d2ab921e9017c8592f49a71ff95a3a306fee7798947

Observation af4c9a95-d7ed-4833-b603-264f4541f191 · inbound

Perception, Reason, Think, and Plan: A Survey on Large Multimodal Reasoning Models cites this paper.

Perception, Reason, Think, and Plan: A Survey on Large Multimodal Reasoning Models STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs

Reference 273

Resolution
unresolved
no resolver link, observed 2026-08-15T23:21:13.340790Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T23:21:13.340790Z digest=sha256:fe4a3cbb02a7f4d788ad03c06a2dbf3a2e2557b6d630174e316f362e94d39b76

Observation 23c52071-9374-4517-9e4b-b5b62d315cf2 · inbound

The Landscape of Agentic Reinforcement Learning for LLMs: A Survey cites this paper.

The Landscape of Agentic Reinforcement Learning for LLMs: A Survey STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs

Reference 233

Resolution
verified exact
arxiv_id, observed 2026-05-18T19:21:48.322627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-18T19:19:36.427337Z digest=sha256:7060fefdfc255cd6947a79f2e0d35473128eadb49e8cec01b868d28e25d0a2b0

Observation c0c118ee-ba2c-4943-91d0-197b451d7758 · inbound

VIDEOP2R: Video Understanding from Perception to Reasoning cites this paper.

VIDEOP2R: Video Understanding from Perception to Reasoning STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-17T22:25:22.690136Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-17T22:24:41.760120Z digest=sha256:110d112e6714b68b0888c6e149b0499247a670a7d50936358dc218b582957bc6

Observation c5ef8303-4565-4db1-a785-e21994e09dc6 · inbound

OneThinker: All-in-one Reasoning Model for Image and Video cites this paper.

OneThinker: All-in-one Reasoning Model for Image and Video STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-17T02:11:26.499729Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-17T02:09:39.820651Z digest=sha256:09f0b99d9fb3546aa99bd052d18008cf8b5422d2314ac2a00f91dac7cfe00ed2

Observation 0b1a2ca8-b126-4ea0-a620-aba6a117bbea · inbound

CamReasoner: Reinforcing Camera Movement Understanding via Structured Spatial Reasoning cites this paper.

CamReasoner: Reinforcing Camera Movement Understanding via Structured Spatial Reasoning STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-16T10:02:42.413324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-16T10:02:20.477517Z digest=sha256:389facb0f70b36cc8b2e397c148a2426cc4822f9f3a52fd782b8e96d3c35dcc8

Observation f5fe546e-47f0-4f33-b2df-f9ef43d8d2e8 · inbound

Fully Spiking Neural Networks with Target Awareness for Energy-Efficient UAV Tracking cites this paper.

Fully Spiking Neural Networks with Target Awareness for Energy-Efficient UAV Tracking STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs

Reference 20

Resolution
unresolved
no resolver link, observed 2026-07-13T16:55:20.099628Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T16:55:20.099628Z digest=sha256:815319c38eeabf1a0d51c6b1d910870cf825ba60e7d3c7b03e480a44f14c76d5

Observation e5b02838-3488-422e-b562-c080880d485d · inbound

Learning to Focus and Precise Cropping: A Reinforcement Learning Framework with Information Gaps and Grounding Loss for MLLMs cites this paper.

Learning to Focus and Precise Cropping: A Reinforcement Learning Framework with Information Gaps and Grounding Loss for MLLMs STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-14T21:38:00.985595Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-14T21:35:12.859669Z digest=sha256:12a481fb10eb4134f9dd3708ad4901d4c397d78cdd4193ffb82e41e3cc1063f6

Observation 9ae53715-7a4f-4529-9a94-fbda2d4735b4 · inbound

Graph-to-Frame RAG: Visual-Space Knowledge Fusion for Training-Free and Auditable Video Reasoning cites this paper.

Graph-to-Frame RAG: Visual-Space Knowledge Fusion for Training-Free and Auditable Video Reasoning STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:30:52.525638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-10T19:46:16.975267Z digest=sha256:7d038e0643b394b5fe507ac834ed955e68ae378df6311da6233e5881eea05d3f

Observation dd904fd6-2380-4591-afc2-19ec70eb6245 · inbound

Reinforce to Learn, Elect to Reason: A Dual Paradigm for Video Reasoning cites this paper.

Reinforce to Learn, Elect to Reason: A Dual Paradigm for Video Reasoning STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:35:52.573891Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-10T19:40:41.642852Z digest=sha256:257db4bc4aae5d5765134fde3160772345720fa3f6113dd5cfadd16905624d0f

Observation a1bcdc9a-4721-4290-a3e5-7842709a2325 · inbound

Beyond Semantic Search: Towards Referential Anchoring in Composed Image Retrieval cites this paper.

Beyond Semantic Search: Towards Referential Anchoring in Composed Image Retrieval STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-10T22:35:50.543718Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-10T19:42:23.224746Z digest=sha256:70a516325b080e05352ce756a7c992902e719dcb5e6a9573af3d1a7b135d4c0d

Observation 76164de6-9435-4e6a-a1ef-2695f34cf3de · inbound

From Structure to Synergy: A Survey of Vision-Language Perception Paradigm Evolution in Multimodal Large Language Models cites this paper.

From Structure to Synergy: A Survey of Vision-Language Perception Paradigm Evolution in Multimodal Large Language Models STAR-R1: Spatial TrAnsformation Reasoning by Reinforcing Multimodal LLMs

Reference 192

Resolution
verified exact
arxiv_id, observed 2026-07-04T15:09:55.323884Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-26T01:50:54.242508Z digest=sha256:6f95ea590561db347c47a79519aa13e569ea08ffd42fe9c84560d7ece66a93ee