Pith. sign in

Paper Citation Record · LEDGER

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning

As of 5 August 2026, this Paper Citation Record lists 40 of 40 outbound references and 1 inbound Pith citation observation for arXiv:2606.08035.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2606.08035 v1

Coverage vector

measured 40 of 40 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-06-27T20:08:29.208550Z

measured 41 of 41 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-02T03:31:46.679205Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

40 of 40 outbound references displayed

  • verified exact22
  • verified fuzzy0
  • unresolved13
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch5

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 582e515b-bd11-4609-84e1-890ce1bd32b3 · outbound

This paper cites arXiv preprint arXiv:2510.06477 , year=.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning arXiv preprint arXiv:2510.06477 , year=

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T20:47:23.065620Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:cedb63bdcce490feca575a8647761eec6a2caf9b363b9256d1804ed6923a9e87

Observation a676f31b-ce4b-4224-88e2-6982e9703f15 · outbound

This paper cites Vlmevalkit: An open-source toolkit for evaluating large multi-modality mod- els.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning Vlmevalkit: An open-source toolkit for evaluating large multi-modality mod- els

Reference 2

Resolution
unresolved
no resolver link, observed 2026-06-27T20:08:29.208550Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:33b0138a43dfa3fe266701a261d79fc988cabdef04939e465b221ab6ebf40af8

Observation 09e8f39a-82e3-40cd-802b-7efa230a74ef · outbound

This paper cites MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-07-02T20:47:23.092686Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:c27eb55f1513e9909e55e83d311cafd2101b41f1693b6ea7ddc03fcce0817bb2

Observation a1c8aef3-7b27-44b4-a0fd-5067f7784eeb · outbound

This paper cites Soft Adaptive Policy Optimization.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning Soft Adaptive Policy Optimization

Reference 4

Resolution
verified exact
local_arxiv, observed 2026-07-02T20:47:23.102053Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:fe6910417efe3a75e31f4eb1cfc69293aa4427457fd983e9f303327ff7018413

Observation 34c691ca-64e5-4090-8c06-2767c0c8131a · outbound

This paper cites Hallusionbench: An advanced diagnostic suite for entangled language hallucination and visual illusion in large vision-language models.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning Hallusionbench: An advanced diagnostic suite for entangled language hallucination and visual illusion in large vision-language models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-06-27T20:08:29.208550Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:fe9dbac5c696be8c6eca653924d2ef47ff49eb5848824d1b8ca8c342778d7ea6

Observation 56956140-54b6-4060-8376-7f7e53d74533 · outbound

This paper cites Martin, Ming-Ming Cheng, and Shi-Min Hu.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning Martin, Ming-Ming Cheng, and Shi-Min Hu

Reference 6

Resolution
verified exact
doi, observed 2026-06-27T20:11:12.978750Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:c1117ec66977ebf6d19f5e65184a254ea5e344af29daa04763adc95eed68021f

Observation b98a9876-c02b-4d23-9bf3-6a3a62f52207 · outbound

This paper cites Visual attention methods in deep learning: An in-depth survey.Information Fusion, 108:102417, 2024.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning Visual attention methods in deep learning: An in-depth survey.Information Fusion, 108:102417, 2024

Reference 7

Resolution
unresolved
no resolver link, observed 2026-06-27T20:08:29.208550Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:ba4ea05da75a510dd82514cf8aacff1f82b768a28214909628f4fe4f9847a2d6

Observation 3a9bff08-a24e-4be2-a0d7-9af113012c43 · outbound

This paper cites URLhttps://www.sciencedirect.com/science/ article/pii/S1566253524001957.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning URLhttps://www.sciencedirect.com/science/ article/pii/S1566253524001957

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-06-27T20:11:12.974033Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:bfa3dd68510bf0457c4ff89630c7536f01d4d8166757f238cb36fe7edd557f81

Observation 2fa58f46-4bb1-4717-bf55-e5cb53abe2a9 · outbound

This paper cites Interpretable visual reasoning: A survey.Image and Vision Computing, 112:104194, 2021.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning Interpretable visual reasoning: A survey.Image and Vision Computing, 112:104194, 2021

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-06-27T20:11:12.976948Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:1282eb407d60346ad9535dff342af0e0cfc6b9d63f656f45e7b878773c74f97f

Observation 0b97bbb4-d8eb-41ba-a587-19603332a4e5 · outbound

This paper cites Distill visual chart reasoning ability from llms to mllms.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning Distill visual chart reasoning ability from llms to mllms

Reference 10

Resolution
unresolved
no resolver link, observed 2026-06-27T20:08:29.208550Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:628ee9cdffa0cff8f433b48f3ed1a22eef00ef0be5419b96b9261c79cfc61502

Observation 912b9290-22b4-437c-91c1-24863a2fa035 · outbound

This paper cites Spotlight on token perception for multimodal reinforcement learning.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning Spotlight on token perception for multimodal reinforcement learning

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-07-02T20:47:23.084893Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:3ab2a150a4a175746fe44ba8e3f77dc3c1e7196a12e1718f01176ac58c4b5a4c

Observation 0ddf6bf1-bf5a-499c-8f53-50cda13ae7da · outbound

This paper cites Credit where it is due: Cross-modality connectivity drives precise reinforcement learning for mllm reasoning, 2026.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning Credit where it is due: Cross-modality connectivity drives precise reinforcement learning for mllm reasoning, 2026

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-07-02T20:47:23.073667Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:a3c45c58a110e64b0d3c0739da257d2ad4377c5fc70ebe4073b99854aaa1827d

Observation bcf55865-2ecb-4bf6-907e-250b4f2f899e · outbound

This paper cites Explain Before You Answer: A Survey on Compositional Visual Reasoning.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning Explain Before You Answer: A Survey on Compositional Visual Reasoning

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-07-09T01:19:37.179138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:8d5761e22bde7c1e77b02bffa43ce64a906205507273b1a81a337eda68c7a8ba

Observation 2651b951-387c-4c00-9ddd-c24f483a219a · outbound

This paper cites an unresolved cited work.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning Unresolved cited work

Reference 14

Resolution
verified exact
doi, observed 2026-06-27T20:11:12.980820Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:75c614ae10182ac338a65745a1e683a777bb95f4dd11c7664087ab4a45bb6448

Observation b43941b8-f4d3-45ff-aa04-d7e46507aa91 · outbound

This paper cites What does rl improve for visual reasoning? a frankenstein-style analysis,.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning What does rl improve for visual reasoning? a frankenstein-style analysis,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-06-27T20:08:29.208550Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:fc743d1ec86b73410fe67e823c3c1525b384e3e7937973745410c5546eb013b9

Observation 8576971e-b4bd-4132-b323-596baafc1934 · outbound

This paper cites an unresolved cited work.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning Unresolved cited work

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-07-02T20:47:23.087569Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:31a27f82de2f985cba2c088bc6ca1c9b442cb493787a2133490d55bdca9d27da

Observation ded3bdf1-a0d6-49a3-ba14-8c779f7a77e2 · outbound

This paper cites Critical tokens matter: Token-level contrastive estimation enhances LLM’s reasoning capability.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning Critical tokens matter: Token-level contrastive estimation enhances LLM’s reasoning capability

Reference 17

Resolution
unresolved
no resolver link, observed 2026-06-27T20:08:29.208550Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:8a76533859a536748429e8cbba03446a33d3e5bf3d1c4675e0511d0a1b36b201

Observation 9cfe4eb9-6a13-42d8-af07-916feb82fcb7 · outbound

This paper cites Mmbench: Is your multi-modal model an all-around player? InEuropean conference on computer vision, pages 216–233.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning Mmbench: Is your multi-modal model an all-around player? InEuropean conference on computer vision, pages 216–233

Reference 18

Resolution
unresolved
no resolver link, observed 2026-06-27T20:08:29.208550Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:0f4af78807a1882520ba7cb926dcdebedf518e8a49fc9fa5bbd51300a57fac4e

Observation c728fae0-79da-4506-b42c-f0597fa79a34 · outbound

This paper cites an unresolved cited work.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning Unresolved cited work

Reference 19

Resolution
unresolved
no resolver link, observed 2026-06-27T20:08:29.208550Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:962db4aaa4726d6d51fd6bf59d9cfc39abc2e36b05bdb241e1b5743a7a23835f

Observation f866d8bf-8b5f-4da2-b9c5-4d640f93ea7a · outbound

This paper cites Cppo: Contrastive perception for vision language policy optimization.arXiv preprint arXiv:XXXX.XXXXX, 2026.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning Cppo: Contrastive perception for vision language policy optimization.arXiv preprint arXiv:XXXX.XXXXX, 2026

Reference 20

Resolution
unresolved
no resolver link, observed 2026-06-27T20:08:29.208550Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:db570b442235591b4701ef2bcef9024ed26896e892d2bb91368d6acd02a70fe8

Observation 24c960c3-c814-498b-a57a-e0d9cb1d8607 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-07-02T20:47:23.086545Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:2ad62f0046e0958596f49dc64ee4fc7a15dab0f91deb8d200d017d240ac56057

Observation 69b5d094-513b-422e-8731-8e498e982fc1 · outbound

This paper cites VLM-R1: A Stable and Generalizable R1-style Large Vision-Language Model.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning VLM-R1: A Stable and Generalizable R1-style Large Vision-Language Model

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-07-02T20:47:23.090192Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:e70e743d5234392229a94f923f4dc787af515d9f147ffce3d091814545593355

Observation 00d8e3da-6333-4448-82c2-3944ed297dc0 · outbound

This paper cites arXiv preprint arXiv:2511.00916 , year=.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning arXiv preprint arXiv:2511.00916 , year=

Reference 23

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T20:47:23.095852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:7c7ea747091e8fe75d9d05a1984acaa199aa3c5c515b258092a25b94e5c8f29c

Observation d44c2315-de41-47d5-8808-fb4ade00f4bf · outbound

This paper cites Terrascope: Pixel-grounded visual reasoning for earth observation.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning Terrascope: Pixel-grounded visual reasoning for earth observation

Reference 24

Resolution
unresolved
no resolver link, observed 2026-06-27T20:08:29.208550Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:a347dcd72419da3630e5a4bf169b47b1b51acc373aeaf43094ab94e6de830cf9

Observation ea5a36e1-cc8a-46e7-87dc-23c50209f99a · outbound

This paper cites Attention Residuals.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning Attention Residuals

Reference 25

Resolution
metadata mismatch
local_arxiv, observed 2026-07-02T20:47:23.055792Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:128580edb32292bfce8bff4832f59892b252895659c91bffdf3a77cf1892eba2

Observation 540c9ab2-f33f-4a51-b708-a774e13b85b0 · outbound

This paper cites Mllm can see? dynamic correction decoding for hallucination mitigation.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning Mllm can see? dynamic correction decoding for hallucination mitigation

Reference 26

Resolution
unresolved
no resolver link, observed 2026-06-27T20:08:29.208550Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:90acea3e29e41bae206e955f357fbd3d552305696e90b85d2b2be1c3f8af3f4e

Observation 9b775aed-d7b4-4397-a070-2e5098f2278a · outbound

This paper cites Measuring multimodal mathematical reasoning with math-vision dataset.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning Measuring multimodal mathematical reasoning with math-vision dataset

Reference 27

Resolution
unresolved
no resolver link, observed 2026-06-27T20:08:29.208550Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:6ba189b24487b42fcecfe7d1bd57606cff2565f1b5bbacc2d9227ccff03503c9

Observation 054e8c7e-f438-48cc-b2d1-32f7f80c9388 · outbound

This paper cites SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-07-02T20:47:23.075925Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:f2f5035c44661ca3da47e574cc2c9ae218039f43245eb93d10290eba62fa10cd

Observation b231af4a-3a9f-479c-bf24-841ad247d51a · outbound

This paper cites Latent Space Chain-of-Embedding Enables Output-free LLM Self-Evaluation.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning Latent Space Chain-of-Embedding Enables Output-free LLM Self-Evaluation

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-07-02T20:47:23.068117Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:7619003a56b8cc3c93f053a72c8d321c52be48cada1557d14f9cbe89eb4b9081

Observation ffcadad2-aec4-47fc-afe0-0ef93208dae1 · outbound

This paper cites Perception-Aware Policy Optimization for Multimodal Reasoning.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning Perception-Aware Policy Optimization for Multimodal Reasoning

Reference 30

Resolution
verified exact
local_arxiv, observed 2026-07-02T20:47:23.104445Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:e88511607da6e1b4859cde9461cea3b89ab7fe24ebc4118039b7c01630fa7b2c

Observation c50f43f9-1964-4adc-a361-49b8fd08c196 · outbound

This paper cites VisuLogic: A Benchmark for Evaluating Visual Reasoning in Multi-modal Large Language Models.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning VisuLogic: A Benchmark for Evaluating Visual Reasoning in Multi-modal Large Language Models

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-07-02T20:47:23.099001Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:a5e7f0ef5e6b2c842fc2ed4ba482a01e5dd26f5eabe3c03e8ddd6d5df7a5e300

Observation cd8caa39-9c92-487b-b70b-1771377efa11 · outbound

This paper cites BioProBench: A Corpus and Benchmark for Biological Protocol Reasoning in Autonomous Science.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning BioProBench: A Corpus and Benchmark for Biological Protocol Reasoning in Autonomous Science

Reference 32

Resolution
metadata mismatch
arxiv_id, observed 2026-07-28T02:22:20.781065Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:825ed73745b11d861de67971371b2705504f355445224af8dfd3e234cc55afdc

Observation 4b0b781d-be45-438e-a111-41b6c7d9fcb5 · outbound

This paper cites DAPO: An Open-Source LLM Reinforcement Learning System at Scale.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning DAPO: An Open-Source LLM Reinforcement Learning System at Scale

Reference 33

Resolution
verified exact
local_arxiv, observed 2026-07-02T20:47:23.073078Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:f66c93db05f70276775a46d5422584d5b90b8a37a10fa2eed35f3ab49cd096b4

Observation a8e2813e-a815-4476-8a2a-e98ae7d2016c · outbound

This paper cites Token-level Direct Preference Optimization.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning Token-level Direct Preference Optimization

Reference 34

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T20:47:23.097979Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:149e2913155023b23a8f19ec494203fa8d3f739d5de552de66d36afe7eac1e07

Observation 0837cd07-6d75-47b0-a2a7-eeb8d6e6209e · outbound

This paper cites Perceptual-evidence anchored reinforced learning for multimodal reasoning.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning Perceptual-evidence anchored reinforced learning for multimodal reasoning

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-07-02T20:47:23.042929Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:112ff55b748758a49f007ce7b6d82bf336e769e4657b50ad1e0ec0f5486cb3d9

Observation 6fb61329-55b6-4ac2-96a2-70ba13836a0f · outbound

This paper cites Thinking With Videos: Multimodal Tool-Augmented Reinforcement Learning for Long Video Reasoning.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning Thinking With Videos: Multimodal Tool-Augmented Reinforcement Learning for Long Video Reasoning

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-07-02T20:47:23.078516Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:9a26e5ee7be1ef1e9d288761ea0401048e5a1d6b28c6f095ff109f51d5423bbd

Observation d0e0f70c-5ecc-4ac2-b0ad-9f99da2cbee1 · outbound

This paper cites MLLMs Know Where to Look: Training-free Perception of Small Visual Details with Multimodal LLMs.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning MLLMs Know Where to Look: Training-free Perception of Small Visual Details with Multimodal LLMs

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-07-02T20:47:23.070679Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:26c90874db0a76161febc2ab9cb9334d65e2e3c772011a2f8c9441dd4cc94946

Observation b175a55f-ddcb-477a-aeb9-7766730dae7b · outbound

This paper cites MathVerse: Does Your Multi-modal LLM Truly See the Diagrams in Visual Math Problems?.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning MathVerse: Does Your Multi-modal LLM Truly See the Diagrams in Visual Math Problems?

Reference 38

Resolution
verified exact
local_arxiv, observed 2026-07-02T20:47:23.084111Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:ee41239342bb48bed9c3a5eb00ce9f11ee989005298ae78ac8e044cd5db8ce78

Observation c818b8f7-dd24-4ceb-94b3-2f3cbee5dd23 · outbound

This paper cites Group Sequence Policy Optimization.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning Group Sequence Policy Optimization

Reference 39

Resolution
verified exact
local_arxiv, observed 2026-07-02T20:47:23.067963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:2832acf90907429026ecd9963d3f25cb4ff29574420e7d0e8cf645942e24d0f6

Observation 913cf652-56c7-4950-806c-e7793cf6d6d7 · outbound

This paper cites since angle ADE is 80 ◦.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning since angle ADE is 80 ◦

Reference 40

Resolution
unresolved
no resolver link, observed 2026-06-27T20:08:29.208550Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:402d3fb872f7af77fd28168315d6c702790c181077d91a4269ef8a5a72cf8019

Pith citing papers

Observation d320dcc6-4118-4dc5-b9f3-1d40a5ed7a78 · inbound

Towards Enhancing 3D Spatial Reasoning in Medical Multimodal Large Language Models cites this paper.

Towards Enhancing 3D Spatial Reasoning in Medical Multimodal Large Language Models DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-02T03:31:46.679205Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T03:31:46.679205Z digest=sha256:91f5d96ad00d2e4db45c3c034ddfbb7a66aab298166c52548bb155bfc9858cc2