Pith. sign in

Paper Citation Record · LEDGER

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning

As of 7 August 2026, this Paper Citation Record lists 40 of 40 outbound references and 1 inbound Pith citation observation for arXiv:2606.08035.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2606.08035 v1

Coverage vector

measured 40 of 40 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-06-27T20:08:29.208550Z

measured 41 of 41 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-02T03:31:46.679205Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

40 of 40 outbound references displayed

  • verified exact22
  • verified fuzzy0
  • unresolved13
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch5

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 582e515b-bd11-4609-84e1-890ce1bd32b3 · outbound

This paper cites arXiv preprint arXiv:2510.06477 , year=.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning arXiv preprint arXiv:2510.06477 , year=

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T20:47:23.065620Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:9e4ea75745eb59d54f597ac363aea0996abbc9b44748de2a996bcd044fea2396

Observation a676f31b-ce4b-4224-88e2-6982e9703f15 · outbound

This paper cites Vlmevalkit: An open-source toolkit for evaluating large multi-modality mod- els.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning Vlmevalkit: An open-source toolkit for evaluating large multi-modality mod- els

Reference 2

Resolution
unresolved
no resolver link, observed 2026-06-27T20:08:29.208550Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:3ff5fe38449139f9e2f5719e267530422cb777122948fae87dde8579d667223a

Observation 09e8f39a-82e3-40cd-802b-7efa230a74ef · outbound

This paper cites MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-07-02T20:47:23.092686Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:20867254b227112d855995884e9ffcca38a6f615266a26643b7206a87a3cabd1

Observation a1c8aef3-7b27-44b4-a0fd-5067f7784eeb · outbound

This paper cites Soft Adaptive Policy Optimization.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning Soft Adaptive Policy Optimization

Reference 4

Resolution
verified exact
local_arxiv, observed 2026-07-02T20:47:23.102053Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:d26fe28988a9b06000e3ee7963eb662624c7722c382acd727cdd671ff110879d

Observation 34c691ca-64e5-4090-8c06-2767c0c8131a · outbound

This paper cites Hallusionbench: An advanced diagnostic suite for entangled language hallucination and visual illusion in large vision-language models.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning Hallusionbench: An advanced diagnostic suite for entangled language hallucination and visual illusion in large vision-language models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-06-27T20:08:29.208550Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:83cdd3ea5f80114bba60f240b387d3cadf4447eea59fb32401a700aae68fdb89

Observation 56956140-54b6-4060-8376-7f7e53d74533 · outbound

This paper cites Martin, Ming-Ming Cheng, and Shi-Min Hu.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning Martin, Ming-Ming Cheng, and Shi-Min Hu

Reference 6

Resolution
verified exact
doi, observed 2026-06-27T20:11:12.978750Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:983884d31e7a8c04ea05611915113e6c6ead38d2392b6c0be1c1cff40764e8df

Observation b98a9876-c02b-4d23-9bf3-6a3a62f52207 · outbound

This paper cites Visual attention methods in deep learning: An in-depth survey.Information Fusion, 108:102417, 2024.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning Visual attention methods in deep learning: An in-depth survey.Information Fusion, 108:102417, 2024

Reference 7

Resolution
unresolved
no resolver link, observed 2026-06-27T20:08:29.208550Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:1706a8ea5a1551943f49e12e778381108f0829192806e8e900bbd30d43c337c0

Observation 3a9bff08-a24e-4be2-a0d7-9af113012c43 · outbound

This paper cites URLhttps://www.sciencedirect.com/science/ article/pii/S1566253524001957.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning URLhttps://www.sciencedirect.com/science/ article/pii/S1566253524001957

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-06-27T20:11:12.974033Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:45b0f93bed36643cf66d50fe284e1a4ec3b2c7f46606aa4f0bba7e63d7eeae6a

Observation 2fa58f46-4bb1-4717-bf55-e5cb53abe2a9 · outbound

This paper cites Interpretable visual reasoning: A survey.Image and Vision Computing, 112:104194, 2021.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning Interpretable visual reasoning: A survey.Image and Vision Computing, 112:104194, 2021

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-06-27T20:11:12.976948Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:6f56a6b1ccef4f5c877bcff7635350d6edc51915c507244f5154677336fef5de

Observation 0b97bbb4-d8eb-41ba-a587-19603332a4e5 · outbound

This paper cites Distill visual chart reasoning ability from llms to mllms.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning Distill visual chart reasoning ability from llms to mllms

Reference 10

Resolution
unresolved
no resolver link, observed 2026-06-27T20:08:29.208550Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:fa6ac02114ee0cd1e5f40d2178f0070233aed09692182541371f41cd72398f1d

Observation 912b9290-22b4-437c-91c1-24863a2fa035 · outbound

This paper cites Spotlight on token perception for multimodal reinforcement learning.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning Spotlight on token perception for multimodal reinforcement learning

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-07-02T20:47:23.084893Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:ebd5ebfaa1e247a41df31f05eb8eeead5afd4b47d7d207463597b5d76794f521

Observation 0ddf6bf1-bf5a-499c-8f53-50cda13ae7da · outbound

This paper cites Credit where it is due: Cross-modality connectivity drives precise reinforcement learning for mllm reasoning, 2026.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning Credit where it is due: Cross-modality connectivity drives precise reinforcement learning for mllm reasoning, 2026

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-07-02T20:47:23.073667Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:978d1f875b3d6972fa9e32ec6202b3fd1baff230cb3a9dafba132e56735e33e4

Observation bcf55865-2ecb-4bf6-907e-250b4f2f899e · outbound

This paper cites Explain Before You Answer: A Survey on Compositional Visual Reasoning.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning Explain Before You Answer: A Survey on Compositional Visual Reasoning

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-07-09T01:19:37.179138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:a2d734f3406b663d56edfe45960fdbe48a10d9bca7db9af12ce694110d7077b9

Observation 2651b951-387c-4c00-9ddd-c24f483a219a · outbound

This paper cites an unresolved cited work.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning Unresolved cited work

Reference 14

Resolution
verified exact
doi, observed 2026-06-27T20:11:12.980820Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:368bb2b2ef4e15d4204d0ae0584c847e0696c6bd2861ca8e743de8ea5eea8e1b

Observation b43941b8-f4d3-45ff-aa04-d7e46507aa91 · outbound

This paper cites What does rl improve for visual reasoning? a frankenstein-style analysis,.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning What does rl improve for visual reasoning? a frankenstein-style analysis,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-06-27T20:08:29.208550Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:864d3070727cf6f93b8f1284d25976aeacdf14c3cb61cfedd75e58c9a38b300b

Observation 8576971e-b4bd-4132-b323-596baafc1934 · outbound

This paper cites an unresolved cited work.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning Unresolved cited work

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-07-02T20:47:23.087569Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:7d8c0946491669bc1a74619cee32f5bbdf3210836bb6572271891d9ef3fb24f4

Observation ded3bdf1-a0d6-49a3-ba14-8c779f7a77e2 · outbound

This paper cites Critical tokens matter: Token-level contrastive estimation enhances LLM’s reasoning capability.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning Critical tokens matter: Token-level contrastive estimation enhances LLM’s reasoning capability

Reference 17

Resolution
unresolved
no resolver link, observed 2026-06-27T20:08:29.208550Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:c3a61815513f9d1d05a30a60ce5a882b3651f7230af128014e43e9a7c682b7b5

Observation 9cfe4eb9-6a13-42d8-af07-916feb82fcb7 · outbound

This paper cites Mmbench: Is your multi-modal model an all-around player? InEuropean conference on computer vision, pages 216–233.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning Mmbench: Is your multi-modal model an all-around player? InEuropean conference on computer vision, pages 216–233

Reference 18

Resolution
unresolved
no resolver link, observed 2026-06-27T20:08:29.208550Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:d44af4935f0cdd0a27218b6e6a7012ce0d0895ffe478eda1a98d7eb838b094b3

Observation c728fae0-79da-4506-b42c-f0597fa79a34 · outbound

This paper cites an unresolved cited work.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning Unresolved cited work

Reference 19

Resolution
unresolved
no resolver link, observed 2026-06-27T20:08:29.208550Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:4133b2923157513a9fd07d24e5de227196fb4cf40602805f463c2ae8fa5eeec0

Observation f866d8bf-8b5f-4da2-b9c5-4d640f93ea7a · outbound

This paper cites Cppo: Contrastive perception for vision language policy optimization.arXiv preprint arXiv:XXXX.XXXXX, 2026.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning Cppo: Contrastive perception for vision language policy optimization.arXiv preprint arXiv:XXXX.XXXXX, 2026

Reference 20

Resolution
unresolved
no resolver link, observed 2026-06-27T20:08:29.208550Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:4237425344d26f04ff16c1eb744d4de5b69f94aa2ad9dc925bcaf2205e447299

Observation 24c960c3-c814-498b-a57a-e0d9cb1d8607 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-07-02T20:47:23.086545Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:7c7623edf3bfc376867b99f5ee0589b77d7fdaac7052c452e5f063c487ff11ac

Observation 69b5d094-513b-422e-8731-8e498e982fc1 · outbound

This paper cites VLM-R1: A Stable and Generalizable R1-style Large Vision-Language Model.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning VLM-R1: A Stable and Generalizable R1-style Large Vision-Language Model

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-07-02T20:47:23.090192Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:ccf7a2cc39af5497cd96bd8a5688b2d3ee1938223593080c335aca098e8c36da

Observation 00d8e3da-6333-4448-82c2-3944ed297dc0 · outbound

This paper cites arXiv preprint arXiv:2511.00916 , year=.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning arXiv preprint arXiv:2511.00916 , year=

Reference 23

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T20:47:23.095852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:43a0f2745ee22d06a81f150c01323740e8d5aa1d9777a147fe0dda74cdaa6b4a

Observation d44c2315-de41-47d5-8808-fb4ade00f4bf · outbound

This paper cites Terrascope: Pixel-grounded visual reasoning for earth observation.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning Terrascope: Pixel-grounded visual reasoning for earth observation

Reference 24

Resolution
unresolved
no resolver link, observed 2026-06-27T20:08:29.208550Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:fa6ae305a4e5a781e8b65c1b48d834daa66dbbf23215461a6d9cd6f907db989f

Observation ea5a36e1-cc8a-46e7-87dc-23c50209f99a · outbound

This paper cites Attention Residuals.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning Attention Residuals

Reference 25

Resolution
metadata mismatch
local_arxiv, observed 2026-07-02T20:47:23.055792Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:e167b747726e01b4b3d6914bb5eca7c816c4843439012e9b70c4d73c5b03ea6e

Observation 540c9ab2-f33f-4a51-b708-a774e13b85b0 · outbound

This paper cites Mllm can see? dynamic correction decoding for hallucination mitigation.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning Mllm can see? dynamic correction decoding for hallucination mitigation

Reference 26

Resolution
unresolved
no resolver link, observed 2026-06-27T20:08:29.208550Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:cea5f2d826c8dd08a33b17635fa36d8371331d9dbaa461219c8651534e468439

Observation 9b775aed-d7b4-4397-a070-2e5098f2278a · outbound

This paper cites Measuring multimodal mathematical reasoning with math-vision dataset.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning Measuring multimodal mathematical reasoning with math-vision dataset

Reference 27

Resolution
unresolved
no resolver link, observed 2026-06-27T20:08:29.208550Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:98fdfa6b413132986dca6eed28bf209b89a24147773e94bcfa8bcff519ba6fe6

Observation 054e8c7e-f438-48cc-b2d1-32f7f80c9388 · outbound

This paper cites SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning SoTA with Less: MCTS-Guided Sample Selection for Data-Efficient Visual Reasoning Self-Improvement

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-07-02T20:47:23.075925Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:123ec6886bf40932a873dedd29e857aadb35214c8747b1f37a98453975e1bb33

Observation b231af4a-3a9f-479c-bf24-841ad247d51a · outbound

This paper cites Latent Space Chain-of-Embedding Enables Output-free LLM Self-Evaluation.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning Latent Space Chain-of-Embedding Enables Output-free LLM Self-Evaluation

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-07-02T20:47:23.068117Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:20c050d6263ee93bdce92396b5416f258d75332a82637ada466040b975a15303

Observation ffcadad2-aec4-47fc-afe0-0ef93208dae1 · outbound

This paper cites Perception-Aware Policy Optimization for Multimodal Reasoning.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning Perception-Aware Policy Optimization for Multimodal Reasoning

Reference 30

Resolution
verified exact
local_arxiv, observed 2026-07-02T20:47:23.104445Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:e08667b90f2ac8e4fbbefe9aca30286297e6bf4ed544836568fa0a887b42078b

Observation c50f43f9-1964-4adc-a361-49b8fd08c196 · outbound

This paper cites VisuLogic: A Benchmark for Evaluating Visual Reasoning in Multi-modal Large Language Models.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning VisuLogic: A Benchmark for Evaluating Visual Reasoning in Multi-modal Large Language Models

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-07-02T20:47:23.099001Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:ae0b1d3c9bb56f0d8a46726adca8cf11694ce19f94cbbe8a2abc573a95f6c482

Observation cd8caa39-9c92-487b-b70b-1771377efa11 · outbound

This paper cites BioProBench: A Corpus and Benchmark for Biological Protocol Reasoning in Autonomous Science.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning BioProBench: A Corpus and Benchmark for Biological Protocol Reasoning in Autonomous Science

Reference 32

Resolution
metadata mismatch
arxiv_id, observed 2026-07-28T02:22:20.781065Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:d909c6abd45434dd90056cc0373e18e3d7b9f212cf7f09fa7401771de33fc33b

Observation 4b0b781d-be45-438e-a111-41b6c7d9fcb5 · outbound

This paper cites DAPO: An Open-Source LLM Reinforcement Learning System at Scale.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning DAPO: An Open-Source LLM Reinforcement Learning System at Scale

Reference 33

Resolution
verified exact
local_arxiv, observed 2026-07-02T20:47:23.073078Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:bc97254ae616b159ceac4b3eba260da055ed00e5067d7994982b379665d62d75

Observation a8e2813e-a815-4476-8a2a-e98ae7d2016c · outbound

This paper cites Token-level Direct Preference Optimization.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning Token-level Direct Preference Optimization

Reference 34

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T20:47:23.097979Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:09679e9879cdaff0d144b62b7bc4f6b4dacbe1607e73cf10574a01a8a1df38dd

Observation 0837cd07-6d75-47b0-a2a7-eeb8d6e6209e · outbound

This paper cites Perceptual-evidence anchored reinforced learning for multimodal reasoning.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning Perceptual-evidence anchored reinforced learning for multimodal reasoning

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-07-02T20:47:23.042929Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:96cb019bce3a2c49b02c03011d20f7854ffd1c0c823ff240eeefd4298285d991

Observation 6fb61329-55b6-4ac2-96a2-70ba13836a0f · outbound

This paper cites Thinking With Videos: Multimodal Tool-Augmented Reinforcement Learning for Long Video Reasoning.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning Thinking With Videos: Multimodal Tool-Augmented Reinforcement Learning for Long Video Reasoning

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-07-02T20:47:23.078516Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:4e2dd78a6519eb15dab023a237b2aed2c56fc173bc4122e86d8d854c14276514

Observation d0e0f70c-5ecc-4ac2-b0ad-9f99da2cbee1 · outbound

This paper cites MLLMs Know Where to Look: Training-free Perception of Small Visual Details with Multimodal LLMs.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning MLLMs Know Where to Look: Training-free Perception of Small Visual Details with Multimodal LLMs

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-07-02T20:47:23.070679Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:19a28212c5a9668cb40fd9195e3d61f06776acb443890e6f9366da2d3b4c976f

Observation b175a55f-ddcb-477a-aeb9-7766730dae7b · outbound

This paper cites MathVerse: Does Your Multi-modal LLM Truly See the Diagrams in Visual Math Problems?.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning MathVerse: Does Your Multi-modal LLM Truly See the Diagrams in Visual Math Problems?

Reference 38

Resolution
verified exact
local_arxiv, observed 2026-07-02T20:47:23.084111Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:6c60735d62ec2ee8cc47084bba7d250b1c2c2583d140bd0a7dd4fe2f44cf88ab

Observation c818b8f7-dd24-4ceb-94b3-2f3cbee5dd23 · outbound

This paper cites Group Sequence Policy Optimization.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning Group Sequence Policy Optimization

Reference 39

Resolution
verified exact
local_arxiv, observed 2026-07-02T20:47:23.067963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:a77cc93ac66fd42a08606c787b697d8781d1305383caa375a37d9511322282be

Observation 913cf652-56c7-4950-806c-e7793cf6d6d7 · outbound

This paper cites since angle ADE is 80 ◦.

DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning since angle ADE is 80 ◦

Reference 40

Resolution
unresolved
no resolver link, observed 2026-06-27T20:08:29.208550Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-27T20:08:29.208550Z digest=sha256:063f2880a690721ad00d3864445f1517e4796a01cf1a0e485dd85aa8361878b7

Pith citing papers

Observation d320dcc6-4118-4dc5-b9f3-1d40a5ed7a78 · inbound

Towards Enhancing 3D Spatial Reasoning in Medical Multimodal Large Language Models cites this paper.

Towards Enhancing 3D Spatial Reasoning in Medical Multimodal Large Language Models DyCo-RL: Dynamic Cross-Modal Coordination for Visual Reasoning

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-02T03:31:46.679205Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T03:31:46.679205Z digest=sha256:472b1ce766599fee89022b2a3eeab11245c491170ff84afc3bbe28f0a7644710