Pith. sign in

Paper Citation Record · LEDGER

Perception-R1: Pioneering Perception Policy with Reinforcement Learning

As of 22 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 48 inbound Pith citation observations for arXiv:2504.07954.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.07954 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 48 of 48 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 48 of 48 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T05:12:18.669805Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-05T11:41:02.675605Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation ae8d8474-6395-471d-9ed2-dad6e7b3dcb0 · inbound

Reinforced MLLM: A Survey on RL-Based Reasoning in Multimodal Large Language Models cites this paper.

Reinforced MLLM: A Survey on RL-Based Reasoning in Multimodal Large Language Models Perception-R1: Pioneering Perception Policy with Reinforcement Learning

Reference 134

Resolution
unresolved
no resolver link, observed 2026-08-16T05:12:18.669805Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T05:12:18.669805Z digest=sha256:9551d26ceb0c284435494743c13e277488f72a9f7e7768e5fc0096e08c400048

Observation 3363e78c-dbd3-49b5-a566-eea2ce310682 · inbound

UniVG-R1: Reasoning Guided Universal Visual Grounding with Reinforcement Learning cites this paper.

UniVG-R1: Reasoning Guided Universal Visual Grounding with Reinforcement Learning Perception-R1: Pioneering Perception Policy with Reinforcement Learning

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-07T15:42:56.467361Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:42:56.467361Z digest=sha256:fcdf8575b26eec9c7522553797ebaf7807ec0c849d67f884347719c4a070a508

Observation 131b1d12-7e74-488f-9639-5a66260a5684 · inbound

RePrompt: Reasoning-Augmented Reprompting for Text-to-Image Generation via Reinforcement Learning cites this paper.

RePrompt: Reasoning-Augmented Reprompting for Text-to-Image Generation via Reinforcement Learning Perception-R1: Pioneering Perception Policy with Reinforcement Learning

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-07T14:49:07.789675Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:49:07.789675Z digest=sha256:583b9f88a6ca9bb4669efe7f177555cafbada261326d46642a2a9e84226a0039

Observation 54ee01f6-ba4b-4c0a-bea4-e51e7d7f9665 · inbound

Reinforcement Fine-Tuning Powers Reasoning Capability of Multimodal Large Language Models cites this paper.

Reinforcement Fine-Tuning Powers Reasoning Capability of Multimodal Large Language Models Perception-R1: Pioneering Perception Policy with Reinforcement Learning

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-07T14:31:16.116604Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:31:16.116604Z digest=sha256:f580d9b817743566e466e663255b8c13d7e1a7233a77262b73724b02a6883f37

Observation 512576cf-8bdd-4828-88db-3d113f7ce37f · inbound

VRAG-RL: Empower Vision-Perception-Based RAG for Visually Rich Information Understanding via Iterative Reasoning with Reinforcement Learning cites this paper.

VRAG-RL: Empower Vision-Perception-Based RAG for Visually Rich Information Understanding via Iterative Reasoning with Reinforcement Learning Perception-R1: Pioneering Perception Policy with Reinforcement Learning

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T13:24:54.235680Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:24:54.235680Z digest=sha256:e6e79b3f788d42cc546d9c1e19160ddb8903b141426fb64398b21fd7bedd740d

Observation d68c3269-b1d9-48f9-8ace-48cd4b1d7387 · inbound

Advancing Multimodal Reasoning via Reinforcement Learning with Cold Start cites this paper.

Advancing Multimodal Reasoning via Reinforcement Learning with Cold Start Perception-R1: Pioneering Perception Policy with Reinforcement Learning

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-07T13:12:58.487945Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:12:58.487945Z digest=sha256:75be24433d3d027cedc6339e15d18e4e4e53417492a9ea4624519bbb05ed570a

Observation 3a2ae015-0241-454b-8d78-6ab933a81cb4 · inbound

AlphaOne: Reasoning Models Thinking Slow and Fast at Test Time cites this paper.

AlphaOne: Reasoning Models Thinking Slow and Fast at Test Time Perception-R1: Pioneering Perception Policy with Reinforcement Learning

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:46.108172Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:35:46.108172Z digest=sha256:bb18ffe4e7856c11c1b546afa58bf1c20930d9c98710531f0d2cfd7850daa68d

Observation f6575182-2a03-4c9a-a160-4fad195fd962 · inbound

Rex-Thinker: Grounded Object Referring via Chain-of-Thought Reasoning cites this paper.

Rex-Thinker: Grounded Object Referring via Chain-of-Thought Reasoning Perception-R1: Pioneering Perception Policy with Reinforcement Learning

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-07T10:55:14.875009Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:55:14.875009Z digest=sha256:6f94d3626c619d79f22974d26ca8bd1a53415c6ccc5492f051e115afa2af6bad

Observation d120e7ae-903b-489a-aa89-14f38d5d8a2f · inbound

WeThink: Toward General-purpose Vision-Language Reasoning via Reinforcement Learning cites this paper.

WeThink: Toward General-purpose Vision-Language Reasoning via Reinforcement Learning Perception-R1: Pioneering Perception Policy with Reinforcement Learning

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-07T05:27:05.125750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:27:05.125750Z digest=sha256:75c350e945b8d96e4325a9c3076bcfd362283ce25eaab6a056286310e8c77f4a

Observation b2f7a081-d9ec-4414-bb4c-2ecd80e87187 · inbound

Efficient Medical VIE via Reinforcement Learning cites this paper.

Efficient Medical VIE via Reinforcement Learning Perception-R1: Pioneering Perception Policy with Reinforcement Learning

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-15T20:07:32.911052Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:07:32.911052Z digest=sha256:59e68820a6957f31d1964de4f0336812b414d14d89d64718c2d870693f626e5a

Observation 318971f5-3c8e-4986-9e14-17882b99da2e · inbound

PeRL: Permutation-Enhanced Reinforcement Learning for Interleaved Vision-Language Reasoning cites this paper.

PeRL: Permutation-Enhanced Reinforcement Learning for Interleaved Vision-Language Reasoning Perception-R1: Pioneering Perception Policy with Reinforcement Learning

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T00:13:38.613601Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:13:38.613601Z digest=sha256:eac58495928a83eede18c1c6c764296bce495e21bad2a036ae9775e3d16cb523

Observation b5c2021d-9678-414a-904f-2e5d2f55041b · inbound

Improving the Reasoning of Multi-Image Grounding in MLLMs via Reinforcement Learning cites this paper.

Improving the Reasoning of Multi-Image Grounding in MLLMs via Reinforcement Learning Perception-R1: Pioneering Perception Policy with Reinforcement Learning

Reference 47

Resolution
metadata mismatch
arxiv_id, observed 2026-05-19T06:52:07.981908Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-19T06:50:02.607136Z digest=sha256:dd0fdcf30eb7c0030f7971ae7c1e4b947baf6eae27f77ac61b4c44d00fbc945d

Observation ca86a00a-d62e-4365-a261-53a173b04924 · inbound

StructVRM: Aligning Multimodal Reasoning with Structured and Verifiable Reward Models cites this paper.

StructVRM: Aligning Multimodal Reasoning with Structured and Verifiable Reward Models Perception-R1: Pioneering Perception Policy with Reinforcement Learning

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-05T23:29:16.461188Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:29:16.461188Z digest=sha256:2638353116bbea024bffea19f2d021b5cb0794ad6aaa06b6178e56a9f1bffc0c

Observation 05bd08eb-7222-4cb6-880d-87c3c8d19e68 · inbound

An Explainable Machine Learning Framework for Railway Predictive Maintenance using Data Streams from the Metro Operator of Portugal cites this paper.

An Explainable Machine Learning Framework for Railway Predictive Maintenance using Data Streams from the Metro Operator of Portugal Perception-R1: Pioneering Perception Policy with Reinforcement Learning

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-05T23:25:49.286524Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:25:49.286524Z digest=sha256:08bcc87e773128d2b15dd35e684069add7ec58d8e68d71b8d8b11df55e22224a

Observation b47c1fe3-3739-4914-b810-0b8a7b30ce6a · inbound

Beyond Reasoning Gains: Mitigating General-Capability Forgetting in Large Reasoning Models cites this paper.

Beyond Reasoning Gains: Mitigating General-Capability Forgetting in Large Reasoning Models Perception-R1: Pioneering Perception Policy with Reinforcement Learning

Reference 107

Resolution
unresolved
no resolver link, observed 2026-08-04T08:15:58.361229Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T08:15:58.361229Z digest=sha256:42dc83e7e60eb7c424d07abb03537d2e84969582eb2b9d78e55c441fd68c725c

Observation 09e4aa05-6496-48db-a02d-4fc0d6edfcf2 · inbound

VIDEOP2R: Video Understanding from Perception to Reasoning cites this paper.

VIDEOP2R: Video Understanding from Perception to Reasoning Perception-R1: Pioneering Perception Policy with Reinforcement Learning

Reference 64

Resolution
verified exact
arxiv_id, observed 2026-05-17T22:25:22.654708Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-17T22:24:41.760120Z digest=sha256:8e51b501c171bae40e1a14ff77760241637a772a9bc034d4d672954903dedfa9

Observation 2865d547-2a01-4168-afe4-567ee1a15d04 · inbound

OneThinker: All-in-one Reasoning Model for Image and Video cites this paper.

OneThinker: All-in-one Reasoning Model for Image and Video Perception-R1: Pioneering Perception Policy with Reinforcement Learning

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-17T02:11:26.599580Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-17T02:09:39.820651Z digest=sha256:447c49750ce77d6d832f25563551fe5d7873a767fc9afb1b0a15826e06275045

Observation 2dd413a9-b3c8-492c-8eac-b7f26b238764 · inbound

RefBench-PRO: Perceptual and Reasoning Oriented Benchmark for Referring Expression Comprehension cites this paper.

RefBench-PRO: Perceptual and Reasoning Oriented Benchmark for Referring Expression Comprehension Perception-R1: Pioneering Perception Policy with Reinforcement Learning

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-03T18:15:08.183639Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:15:08.183639Z digest=sha256:9ccc434aaee37f965dc3754b1cb183bcb9546862281b453db3f55c5026273c47

Observation 71fe76fd-1a14-4f4b-bde2-9582ca997613 · inbound

ProAgent: Harnessing On-Demand Sensory Contexts for Proactive LLM Agent Systems in the Wild cites this paper.

ProAgent: Harnessing On-Demand Sensory Contexts for Proactive LLM Agent Systems in the Wild Perception-R1: Pioneering Perception Policy with Reinforcement Learning

Reference 61

Resolution
verified exact
arxiv_id, observed 2026-05-17T00:41:24.002284Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-17T00:38:54.001341Z digest=sha256:2e497f1503dc454dd36b781eb530f96e0ba751434c7cd05858c01de4d3209224

Observation 7a2cbe96-a7ea-458a-8649-21e4cc364f6a · inbound

Zoom-IQA: Image Quality Assessment with Reliable Region-Aware Reasoning cites this paper.

Zoom-IQA: Image Quality Assessment with Reliable Region-Aware Reasoning Perception-R1: Pioneering Perception Policy with Reinforcement Learning

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-03T12:31:20.389170Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:31:20.389170Z digest=sha256:1d963beabe3e9769b6ffa875e5c9093727251b956ab5d40547a63b7c6da52fb6

Observation 1d764c15-923c-4e85-99a5-88fa76f4428e · inbound

Can Textual Reasoning Improve the Performance of MLLMs on Fine-grained Visual Classification? cites this paper.

Can Textual Reasoning Improve the Performance of MLLMs on Fine-grained Visual Classification? Perception-R1: Pioneering Perception Policy with Reinforcement Learning

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-05-16T14:53:00.616722Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-16T14:51:26.368439Z digest=sha256:50c6a2cc469725bfbed02e2a96a1c7e5905cbed748094bc38c9b126de05888d0

Observation 134255b6-36c8-4b8f-ab2e-40551bd5d901 · inbound

CamReasoner: Reinforcing Camera Movement Understanding via Structured Spatial Reasoning cites this paper.

CamReasoner: Reinforcing Camera Movement Understanding via Structured Spatial Reasoning Perception-R1: Pioneering Perception Policy with Reinforcement Learning

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-05-16T10:02:42.508887Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-16T10:02:20.477517Z digest=sha256:3333cdd6e7e2045a56e8693517a7f58a32d603ca8d39a4686b3c016a1f9ce7dc

Observation 1453a838-9af6-498b-a1f3-d2f3aca6e466 · inbound

PaLMR: Towards Faithful Visual Reasoning via Multimodal Process Alignment cites this paper.

PaLMR: Towards Faithful Visual Reasoning via Multimodal Process Alignment Perception-R1: Pioneering Perception Policy with Reinforcement Learning

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-02T19:57:33.519473Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T19:57:33.519473Z digest=sha256:a6242dbd828e7000bdd5de0ef9d90d77a8cf9916e2b52fda229c41d4198569bc

Observation 653f93b5-ec73-4dff-97e5-21730460aae8 · inbound

CodePercept: Code-Grounded Visual STEM Perception for MLLMs cites this paper.

CodePercept: Code-Grounded Visual STEM Perception for MLLMs Perception-R1: Pioneering Perception Policy with Reinforcement Learning

Reference 58

Resolution
unresolved
no resolver link, observed 2026-07-14T23:22:13.847876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T23:22:13.847876Z digest=sha256:b899eddade2bae720f0a77386f467ab7a4fd270c5638765c2d0d675fcd2f3f1a

Observation 83b88eaf-9ed7-4802-aeb5-e0b6c180f8db · inbound

V-Reflection: Transforming MLLMs from Passive Observers to Active Interrogators cites this paper.

V-Reflection: Transforming MLLMs from Passive Observers to Active Interrogators Perception-R1: Pioneering Perception Policy with Reinforcement Learning

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-05-13T23:58:28.586668Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-13T23:57:47.657243Z digest=sha256:c1005c434a7c7c6d292c7ad2e9a3b5dec3bcb3adc02f354855597f35de09b2b2

Observation 03360d8f-d0ee-411e-aba6-b93696797b5d · inbound

Act Wisely: Cultivating Meta-Cognitive Tool Use in Agentic Multimodal Models cites this paper.

Act Wisely: Cultivating Meta-Cognitive Tool Use in Agentic Multimodal Models Perception-R1: Pioneering Perception Policy with Reinforcement Learning

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-05-11T00:20:53.791165Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-10T18:35:21.514502Z digest=sha256:376be54e5197533c4aeb7ad430adf4e789b0d844dbb545d667629489096b2ecd

Observation 471b1c4a-5dc7-4667-8d48-6679a0729125 · inbound

Steadily moving semi-infinite fracture in plane poroelasticity cites this paper.

Steadily moving semi-infinite fracture in plane poroelasticity Perception-R1: Pioneering Perception Policy with Reinforcement Learning

Reference 110

Resolution
verified exact
local_arxiv, observed 2026-07-05T11:41:02.677026Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-07-05T11:39:05.686584Z digest=sha256:4df00bba5e9fef60d525e365fbbbee5d9ebfd9258f4dc1dd9f79395518d9aa23

Observation 1b3b8c61-45ed-4b6e-b32b-58230ffa91b3 · inbound

XEmbodied: A Foundation Model with Enhanced Geometric and Physical Cues for Large-Scale Embodied Environments cites this paper.

XEmbodied: A Foundation Model with Enhanced Geometric and Physical Cues for Large-Scale Embodied Environments Perception-R1: Pioneering Perception Policy with Reinforcement Learning

Reference 110

Resolution
verified exact
arxiv_id, observed 2026-05-10T05:51:10.303502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-10T05:46:36.865150Z digest=sha256:f3555110dd7b412da5b99752c2e5b7699c9691fef39e3ff71cafd9b2ed5aed1f

Observation 013d3648-2590-407f-9943-9e8bc8b7be78 · inbound

SSL-R1: Self-Supervised Visual Reinforcement Post-Training for Multimodal Large Language Models cites this paper.

SSL-R1: Self-Supervised Visual Reinforcement Post-Training for Multimodal Large Language Models Perception-R1: Pioneering Perception Policy with Reinforcement Learning

Reference 78

Resolution
verified exact
arxiv_id, observed 2026-05-11T13:36:08.633735Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-10T01:23:32.849326Z digest=sha256:40c3c79d0553b6108391399954fdfc07ad8245635e28549ed1a16bdebb321810

Observation 754de6ce-fdbc-4d95-8492-d22ba9e68312 · inbound

CGC: Compositional Grounded Contrast for Fine-Grained Multi-Image Understanding cites this paper.

CGC: Compositional Grounded Contrast for Fine-Grained Multi-Image Understanding Perception-R1: Pioneering Perception Policy with Reinforcement Learning

Reference 54

Resolution
verified exact
arxiv_id, observed 2026-05-11T19:16:08.342294Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-08T12:26:01.568507Z digest=sha256:05c0e4017940b3a9ecaa0ab08c9e40d79122f1ce836fb7b08cda8dcfde08bc2c

Observation 3bc4e86f-2a6c-4127-9d6a-d4392f44e1d4 · inbound

Injecting Distributional Awareness into MLLMs via Reinforcement Learning for Deep Imbalanced Regression cites this paper.

Injecting Distributional Awareness into MLLMs via Reinforcement Learning for Deep Imbalanced Regression Perception-R1: Pioneering Perception Policy with Reinforcement Learning

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-11T16:51:09.309286Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-09T14:36:29.666730Z digest=sha256:8ff8a29406cfaeed6b8d69dc65faeda218ec537a578d56be5326291154539083

Observation 9cc4566b-bd7b-40b5-8c5a-6c53122a417a · inbound

Injecting Distributional Awareness into MLLMs via Reinforcement Learning for Deep Imbalanced Regression cites this paper.

Injecting Distributional Awareness into MLLMs via Reinforcement Learning for Deep Imbalanced Regression Perception-R1: Pioneering Perception Policy with Reinforcement Learning

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-12T05:51:26.126862Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-12T04:52:09.685243Z digest=sha256:40799fb5dafa498c641b19727db0ebf76f03015493ae0bcb6b0a56b9dd22bfc1

Observation a9c3f915-f5ba-46bf-b7f4-de634d8f2efa · inbound

Self-Consistent Latent Reasoning: Long Latent Sequence Reasoning for Vision-Language Model cites this paper.

Self-Consistent Latent Reasoning: Long Latent Sequence Reasoning for Vision-Language Model Perception-R1: Pioneering Perception Policy with Reinforcement Learning

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-05-13T07:17:29.071938Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-13T07:14:48.918959Z digest=sha256:168f276f0e6cac46820d080375659f8a7598723a1982094cef091abb538ccc85

Observation aede9686-4b49-444d-97bf-70e3bd2a780f · inbound

Self-Consistent Latent Reasoning: Long Latent Sequence Reasoning for Vision-Language Model cites this paper.

Self-Consistent Latent Reasoning: Long Latent Sequence Reasoning for Vision-Language Model Perception-R1: Pioneering Perception Policy with Reinforcement Learning

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-05-14T21:48:00.502847Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-14T21:47:50.595481Z digest=sha256:9aa6e93f84e18f6e8254c0338648d7652ff9c4bbc6ba6bba657ec487c94d43a5

Observation 739c202d-54a7-48c9-a7a5-0a7ededea585 · inbound

CurveBench: A Benchmark for Exact Topological Reasoning over Nested Jordan Curves cites this paper.

CurveBench: A Benchmark for Exact Topological Reasoning over Nested Jordan Curves Perception-R1: Pioneering Perception Policy with Reinforcement Learning

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-15T05:39:48.044228Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-15T05:35:02.980473Z digest=sha256:a2f556a66950367fb4167b436097e1e0be1ceb2bb2df3df4c243f5e19d06ee84

Observation c1c34f59-9dc7-4120-afc6-d10ba06562ba · inbound

CurveBench: A Benchmark for Exact Topological Reasoning over Nested Jordan Curves cites this paper.

CurveBench: A Benchmark for Exact Topological Reasoning over Nested Jordan Curves Perception-R1: Pioneering Perception Policy with Reinforcement Learning

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-20T20:43:43.444374Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-20T20:40:47.750678Z digest=sha256:013e99a5dc8920310a74b43c9e52c1ee3f5d4fa82a193aed7433a3e95330dad6

Observation 048032d9-4f39-43d3-b943-7d2ca08caca8 · inbound

VideoSeeker: Incentivizing Instance-level Video Understanding via Native Agentic Tool Invocation cites this paper.

VideoSeeker: Incentivizing Instance-level Video Understanding via Native Agentic Tool Invocation Perception-R1: Pioneering Perception Policy with Reinforcement Learning

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-05-20T19:38:56.151767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-20T19:37:09.244578Z digest=sha256:88f1179482080991233a66ea8ffc449aa3bf0d927c7fcc6798fffd2644a35795

Observation 772dcefb-41e5-4bf0-9feb-5315db61efca · inbound

Reasoning Portability: Guiding Continual Learning for MLLMs in the RLVR Era cites this paper.

Reasoning Portability: Guiding Continual Learning for MLLMs in the RLVR Era Perception-R1: Pioneering Perception Policy with Reinforcement Learning

Reference 58

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T13:33:19.317425Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-20T13:29:11.509966Z digest=sha256:5b061aeec93484fb998774b9030d5238fb588179e03cb1571bb3edb10de76847

Observation c023d8d4-1963-497e-8158-a53ef0397e09 · inbound

CaMo: Camera Motion Grounded Evaluation and Training for Vision-Language Models cites this paper.

CaMo: Camera Motion Grounded Evaluation and Training for Vision-Language Models Perception-R1: Pioneering Perception Policy with Reinforcement Learning

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-05-20T05:28:04.573557Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-20T05:27:30.938311Z digest=sha256:9e2915dcbbb6df892c1e3b88355f5a43a4fca1e72b7085b201b42f0f4519d5ae

Observation 8c666529-1689-4752-93a3-7448513c3cd5 · inbound

Guidance Contrastive Token Credit Assignment for Discrete Policy Optimization cites this paper.

Guidance Contrastive Token Credit Assignment for Discrete Policy Optimization Perception-R1: Pioneering Perception Policy with Reinforcement Learning

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-06-29T09:03:15.876522Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-29T08:59:03.375697Z digest=sha256:0496d6bc70b8e5f07c4e5cad956699813461dae9fbdb743ca1289bba30b289dc

Observation 3c63fc7a-baf4-4113-8d65-f3e6dcbb0f01 · inbound

From Structure to Synergy: A Survey of Vision-Language Perception Paradigm Evolution in Multimodal Large Language Models cites this paper.

From Structure to Synergy: A Survey of Vision-Language Perception Paradigm Evolution in Multimodal Large Language Models Perception-R1: Pioneering Perception Policy with Reinforcement Learning

Reference 171

Resolution
verified exact
arxiv_id, observed 2026-07-04T15:09:54.977454Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-26T01:50:54.242508Z digest=sha256:bf24f3107567447c867c27302fa756669761d94e2299b1d2ffd7482298a9e744

Observation 00dc9687-0d54-45e4-9029-c3c156ae5d36 · inbound

Perceive-to-Reason: Decoupling Perception and Reasoning for Fine-Grained Visual Reasoning cites this paper.

Perceive-to-Reason: Decoupling Perception and Reasoning for Fine-Grained Visual Reasoning Perception-R1: Pioneering Perception Policy with Reinforcement Learning

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-07-02T13:26:58.415910Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-07-02T13:24:17.538850Z digest=sha256:51aaf382cc9a81b3a8f20ff95654ec23280f4e6805463c3d5b67fa8b734c0bea

Observation 35d974e9-fbad-4f29-9edf-19989a11d0e1 · inbound

DELTAVID: Enhancing Fine-Grained Spatiotemporal Perception with Cross-Video Differences cites this paper.

DELTAVID: Enhancing Fine-Grained Spatiotemporal Perception with Cross-Video Differences Perception-R1: Pioneering Perception Policy with Reinforcement Learning

Reference 57

Resolution
unresolved
no resolver link, observed 2026-07-12T11:31:14.532101Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T11:31:14.532101Z digest=sha256:5673bb55ff00e5192b05f11cfd94ccb436eb5a84ca01ca1dcbed67c489e78a92

Observation 54621a2c-7471-45b4-94fc-9b72a41f2ae0 · inbound

Stop Thinking, Start Looking: Efficient Post-Training for Multimodal Document Question Answering via Reasoning-Free Alignment cites this paper.

Stop Thinking, Start Looking: Efficient Post-Training for Multimodal Document Question Answering via Reasoning-Free Alignment Perception-R1: Pioneering Perception Policy with Reinforcement Learning

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-02T01:25:28.694126Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T01:25:28.694126Z digest=sha256:1c3162fc6bc33dbe6ca6433afefe472fff3f3868db52fee2305220dbf402d1cd

Observation 141dcf71-7f39-4c77-a1a5-63b7bb52163e · inbound

IoU-PD: IoU-Aware Privileged Distillation for Visual Grounding with Multimodal Large Language Models cites this paper.

IoU-PD: IoU-Aware Privileged Distillation for Visual Grounding with Multimodal Large Language Models Perception-R1: Pioneering Perception Policy with Reinforcement Learning

Reference 120

Resolution
unresolved
no resolver link, observed 2026-08-01T22:32:28.980145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T22:32:28.980145Z digest=sha256:dc44fa763e1e18ca818c40d7dbbd7174ff903f5e8801991ef17b17ca5fea332b

Observation 8cb68f21-faca-4452-8d73-7fbfea0e70dc · inbound

IoU-PD: IoU-Aware Privileged Distillation for Visual Grounding with Multimodal Large Language Models cites this paper.

IoU-PD: IoU-Aware Privileged Distillation for Visual Grounding with Multimodal Large Language Models Perception-R1: Pioneering Perception Policy with Reinforcement Learning

Reference 120

Resolution
unresolved
no resolver link, observed 2026-08-04T04:21:09.633302Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T04:21:09.633302Z digest=sha256:1e01c551ab6b562274a52685ebc14b3a902473ec36b4a78f8b0c5df60954d9d0

Observation 86241d39-3a7a-48cf-b4fd-83c5afcb81af · inbound

Hi-Token: Hierarchical Coordinate Tokenization for Generative Visual Grounding cites this paper.

Hi-Token: Hierarchical Coordinate Tokenization for Generative Visual Grounding Perception-R1: Pioneering Perception Policy with Reinforcement Learning

Reference 118

Resolution
unresolved
no resolver link, observed 2026-08-05T18:36:14.869453Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T18:36:14.869453Z digest=sha256:844064dc1ffe2431d168596f18a6e63c8f867eea6e0388e8c4ed9569f2ae7bb8

Observation 4189cd70-4d04-4c72-8181-5dc7d919a096 · inbound

SCOUT: Unlocking Enhanced Spatial Reasoning via Structured Chain-of-Thought and Multi-Objective Process Reward cites this paper.

SCOUT: Unlocking Enhanced Spatial Reasoning via Structured Chain-of-Thought and Multi-Objective Process Reward Perception-R1: Pioneering Perception Policy with Reinforcement Learning

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-16T00:18:18.138194Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T00:18:18.138194Z digest=sha256:ef687f7bdd97b7a80d272aac50cbf92d8f69e8ec299301211ecb44f58533e3e5