Pith. sign in

Paper Citation Record · LEDGER

Skywork R1V2: Multimodal Hybrid Reinforcement Learning for Reasoning

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 18 inbound Pith citation observations for arXiv:2504.16656.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.16656 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 18 of 18 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 18 of 18 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:01:37.530069Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T20:10:08.032034Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 43041165-440b-400b-96f9-d1d2cafa7e9f · inbound

R1-ShareVL: Incentivizing Reasoning Capability of Multimodal Large Language Models via Share-GRPO cites this paper.

R1-ShareVL: Incentivizing Reasoning Capability of Multimodal Large Language Models via Share-GRPO Skywork R1V2: Multimodal Hybrid Reinforcement Learning for Reasoning

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T15:01:37.530069Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:01:37.530069Z digest=sha256:3ff5246f8d32988bb853c4dc134625bc77599180cd4d9e7f4b43b3f273f96dd0

Observation 00eac3aa-c4fa-40f9-ac6b-2a9297bbe228 · inbound

Reinforcement Fine-Tuning Powers Reasoning Capability of Multimodal Large Language Models cites this paper.

Reinforcement Fine-Tuning Powers Reasoning Capability of Multimodal Large Language Models Skywork R1V2: Multimodal Hybrid Reinforcement Learning for Reasoning

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-07T14:31:15.321786Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:31:15.321786Z digest=sha256:f11ff918f5c248cf5c3f0bf8fc5f997579ebb3b89c57bf477e6975c3988eef5e

Observation c52f5b08-33f5-4921-a7eb-fdd7fb77eefd · inbound

Seeing is Believing, but How Much? A Comprehensive Analysis of Verbalized Calibration in Vision-Language Models cites this paper.

Seeing is Believing, but How Much? A Comprehensive Analysis of Verbalized Calibration in Vision-Language Models Skywork R1V2: Multimodal Hybrid Reinforcement Learning for Reasoning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T14:01:22.483974Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:01:22.483974Z digest=sha256:1871087a640b5172862be172213e7b56208d854878b634d0ed27b14f7602421a

Observation 7a5d3b8e-0064-4738-9167-8b751905a84a · inbound

MedBookVQA: A Systematic and Comprehensive Medical Benchmark Derived from Open-Access Book cites this paper.

MedBookVQA: A Systematic and Comprehensive Medical Benchmark Derived from Open-Access Book Skywork R1V2: Multimodal Hybrid Reinforcement Learning for Reasoning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T12:01:03.055536Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:01:03.055536Z digest=sha256:8b00fbd14ff00811c442cf24034aa11c91e084703442d3b3dcb3da6b6d71b5fc

Observation b797d668-c259-444f-899a-0e6aa87ddfc8 · inbound

Rex-Thinker: Grounded Object Referring via Chain-of-Thought Reasoning cites this paper.

Rex-Thinker: Grounded Object Referring via Chain-of-Thought Reasoning Skywork R1V2: Multimodal Hybrid Reinforcement Learning for Reasoning

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T10:55:14.826127Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:55:14.826127Z digest=sha256:10fc4a7636a940d525900e18bac517bb043f637fc64e0c78d0dc11ee94bb91d0

Observation c9527092-abfe-40c2-90fc-066cb510a9a0 · inbound

WeThink: Toward General-purpose Vision-Language Reasoning via Reinforcement Learning cites this paper.

WeThink: Toward General-purpose Vision-Language Reasoning via Reinforcement Learning Skywork R1V2: Multimodal Hybrid Reinforcement Learning for Reasoning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T05:27:04.960499Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:27:04.960499Z digest=sha256:f16d34927d57a4d23e1af51784a5f51638d1e2529ef6c66ce8743ed5ec20d397

Observation 822e85af-b74d-4cbe-8011-983260eafac0 · inbound

FinLMM-R1: Enhancing Financial Reasoning in LMM through Scalable Data and Reward Design cites this paper.

FinLMM-R1: Enhancing Financial Reasoning in LMM through Scalable Data and Reward Design Skywork R1V2: Multimodal Hybrid Reinforcement Learning for Reasoning

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T00:40:40.201848Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:40:40.201848Z digest=sha256:3de671bf4f4905aa1d06d2994e9fd4bc1c87c8a0e5ce83e869302c6e0404a88f

Observation 37c69bd4-167e-4749-ae2e-08166d75dcbc · inbound

Skywork-R1V3 Technical Report cites this paper.

Skywork-R1V3 Technical Report Skywork R1V2: Multimodal Hybrid Reinforcement Learning for Reasoning

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T19:14:05.468540Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:14:05.468540Z digest=sha256:1834388e7032c3ed27472f349a79d17d9d59d2408e1d7706007161082f692ab6

Observation f68c5350-190b-4d9d-8bb9-663afbc9c051 · inbound

Perception-Aware Policy Optimization for Multimodal Reasoning cites this paper.

Perception-Aware Policy Optimization for Multimodal Reasoning Skywork R1V2: Multimodal Hybrid Reinforcement Learning for Reasoning

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-19T05:12:04.894974Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-19T05:11:54.685897Z digest=sha256:1f1871f2152b0faf0708357060e3f94eb8ab8b21ddec4e77fae86772b54923e6

Observation 1c3f7eaa-e072-4070-a191-14f807f18d2d · inbound

The Synergy Dilemma of Long-CoT SFT and RL: Investigating Post-Training Techniques for Reasoning VLMs cites this paper.

The Synergy Dilemma of Long-CoT SFT and RL: Investigating Post-Training Techniques for Reasoning VLMs Skywork R1V2: Multimodal Hybrid Reinforcement Learning for Reasoning

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T18:43:13.461715Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:43:13.461715Z digest=sha256:48e9b2daaf8b99f41ea9fc983558d3aa44d2e69f7af2c578c3704957901bebd4

Observation 7557b985-2c27-4fa3-a62c-bb6b5b8b145b · inbound

VL-Cogito: Progressive Curriculum Reinforcement Learning for Advanced Multimodal Reasoning cites this paper.

VL-Cogito: Progressive Curriculum Reinforcement Learning for Advanced Multimodal Reasoning Skywork R1V2: Multimodal Hybrid Reinforcement Learning for Reasoning

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-06T11:35:16.929602Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T11:35:16.929602Z digest=sha256:f770fe947aac0c54e7dead730862f88254f39098a78b32eedebcc276b3d6a3ec

Observation 4c08be35-c821-4880-8015-d56a3704094c · inbound

InternVL3.5: Advancing Open-Source Multimodal Models in Versatility, Reasoning, and Efficiency cites this paper.

InternVL3.5: Advancing Open-Source Multimodal Models in Versatility, Reasoning, and Efficiency Skywork R1V2: Multimodal Hybrid Reinforcement Learning for Reasoning

Reference 138

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:58:58.831501Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T11:58:58.660564Z digest=sha256:05b46a3e1917801af31e32e8e33d25e6a0d3244132e54240adcf75e502a245c7

Observation b2df9d6a-be8e-4b2f-ba4f-e856480b8e1c · inbound

The Landscape of Agentic Reinforcement Learning for LLMs: A Survey cites this paper.

The Landscape of Agentic Reinforcement Learning for LLMs: A Survey Skywork R1V2: Multimodal Hybrid Reinforcement Learning for Reasoning

Reference 60

Resolution
verified exact
arxiv_id, observed 2026-05-18T19:21:48.436079Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-18T19:19:36.427337Z digest=sha256:9f3292702f2bd4135791b0bb0579bf2df1cfb92e6433fa63a502ae6fffe352fd

Observation d23576fd-95c2-4705-993c-d1c6faa8a2bc · inbound

Boosting Reasoning in Large Multimodal Models via Activation Replay cites this paper.

Boosting Reasoning in Large Multimodal Models via Activation Replay Skywork R1V2: Multimodal Hybrid Reinforcement Learning for Reasoning

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-17T05:09:03.913660Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-17T05:05:48.682057Z digest=sha256:0cd89c1fa8b557286a07187264a6ffc5ac4949c74d90b0d7a18353e8ec9cd67a

Observation b7ae3f4e-507d-498b-8fa1-8a22ed821df6 · inbound

MMCORE: MultiModal COnnection with Representation Aligned Latent Embeddings cites this paper.

MMCORE: MultiModal COnnection with Representation Aligned Latent Embeddings Skywork R1V2: Multimodal Hybrid Reinforcement Learning for Reasoning

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-10T03:29:21.666566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T03:27:50.144706Z digest=sha256:833ec4f594427a21d74f827525653526dfc6a9543341bcd109860a6b7b9f7706

Observation 58fdb66b-61b1-44cb-b81d-fe39515eeb57 · inbound

Structured Role-Aware Policy Optimization for Multimodal Reasoning cites this paper.

Structured Role-Aware Policy Optimization for Multimodal Reasoning Skywork R1V2: Multimodal Hybrid Reinforcement Learning for Reasoning

Reference 6

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T04:25:58.316033Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-11T01:24:20.286120Z digest=sha256:4f947d41876cdb118bd91fab158c58090ce51e120168e289772d4dc3988e1367

Observation d3a16a42-9743-40dd-bb18-c9f6530313e1 · inbound

VeriEvol: Scaling Multimodal Mathematical Reasoning via Verifiable Evol-Instruct cites this paper.

VeriEvol: Scaling Multimodal Mathematical Reasoning via Verifiable Evol-Instruct Skywork R1V2: Multimodal Hybrid Reinforcement Learning for Reasoning

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-07-04T10:39:46.028994Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-26T08:34:27.719022Z digest=sha256:e33071c6f88bd4a2cca1bbc2f96743ae95f8786e0efb35d6fe1ad4658dc3263a

Observation 19cae604-2808-4b8b-bb41-6ef22e2c502e · inbound

OracleAnalyser: Analysing Implicit Semantics of Oracle Bone Scripts through MLLMs with Post-training cites this paper.

OracleAnalyser: Analysing Implicit Semantics of Oracle Bone Scripts through MLLMs with Post-training Skywork R1V2: Multimodal Hybrid Reinforcement Learning for Reasoning

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-07-04T20:10:08.034694Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-25T20:36:16.927195Z digest=sha256:57c0fb711e99ec20d0fbdbc06a5434088a480c05464c988c242e90a2b1b6176e