Pith. sign in

Paper Citation Record · LEDGER

Skywork-VL Reward: An Effective Reward Model for Multimodal Understanding and Reasoning

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 9 inbound Pith citation observations for arXiv:2505.07263.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.07263 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 9 of 9 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 9 of 9 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:31:22.244781Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T08:25:33.331848Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 29c752e8-4d25-4393-9110-51be7b734697 · inbound

Reinforcement Fine-Tuning Powers Reasoning Capability of Multimodal Large Language Models cites this paper.

Reinforcement Fine-Tuning Powers Reasoning Capability of Multimodal Large Language Models Skywork-VL Reward: An Effective Reward Model for Multimodal Understanding and Reasoning

Reference 148

Resolution
unresolved
no resolver link, observed 2026-08-07T14:31:22.244781Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:31:22.244781Z digest=sha256:9d369d4510aa9e8cf8f3d84271c534800324e293ca849e05f63ab0121fc4dac9

Observation daea2ec7-3631-4e7d-a0ea-a91d883a4ca9 · inbound

Skywork-R1V3 Technical Report cites this paper.

Skywork-R1V3 Technical Report Skywork-VL Reward: An Effective Reward Model for Multimodal Understanding and Reasoning

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T19:14:05.742695Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:14:05.742695Z digest=sha256:a99c40d075a8a211ecf0f558423ff1f695b41793175d876f0b6fde010e05d8fe

Observation afe50853-a83f-4c34-9d08-5af78b86a8a9 · inbound

TAR: Temporal Anchor-Constrained Reasoning for Video Temporal Grounding cites this paper.

TAR: Temporal Anchor-Constrained Reasoning for Video Temporal Grounding Skywork-VL Reward: An Effective Reward Model for Multimodal Understanding and Reasoning

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-05T22:01:25.998214Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:01:25.998214Z digest=sha256:a9b4e9526b31cea2f5109379d67066cd373c1a04281f7ae15980f7e5ca957ff6

Observation 061fc6a0-f652-4dc6-9d5b-c112bd9a4b5b · inbound

Visual Preference Optimization with Rubric Rewards cites this paper.

Visual Preference Optimization with Rubric Rewards Skywork-VL Reward: An Effective Reward Model for Multimodal Understanding and Reasoning

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-11T09:56:00.954463Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T15:45:52.980881Z digest=sha256:800004004de40aecf13ce8e1f63cc5d9c5e46c7d24d46ab2f8bddfebbd38e751

Observation 55746557-662c-4734-8315-55f47fd05c61 · inbound

DT2IT-MRM: Debiased Preference Construction and Iterative Training for Multimodal Reward Modeling cites this paper.

DT2IT-MRM: Debiased Preference Construction and Iterative Training for Multimodal Reward Modeling Skywork-VL Reward: An Effective Reward Model for Multimodal Understanding and Reasoning

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-05-11T13:26:03.881700Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T01:45:30.001398Z digest=sha256:05fa26b4f3dd24ad7e25c081de969e60e3c4eb7ef55548e8a6c9abdab6fb9498

Observation 8da7362c-25cc-4882-92e1-9c33fa08d5be · inbound

How to Guide Your Flow: Few-Step Alignment via Flow Map Reward Guidance cites this paper.

How to Guide Your Flow: Few-Step Alignment via Flow Map Reward Guidance Skywork-VL Reward: An Effective Reward Model for Multimodal Understanding and Reasoning

Reference 62

Resolution
verified exact
arxiv_id, observed 2026-05-12T09:46:26.660162Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-07T09:24:42.020782Z digest=sha256:5924ae35c21ff7ac1e28ccb9841926c33383f206ccef0c5534cd3579ae0b5c3f

Observation 072a0e18-81d2-4d77-b783-e6f9fc6f7f6a · inbound

How to Guide Your Flow: Few-Step Alignment via Flow Map Reward Guidance cites this paper.

How to Guide Your Flow: Few-Step Alignment via Flow Map Reward Guidance Skywork-VL Reward: An Effective Reward Model for Multimodal Understanding and Reasoning

Reference 62

Resolution
verified exact
arxiv_id, observed 2026-05-21T09:29:57.206393Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-21T09:25:19.439423Z digest=sha256:1a33cc2d8a9e403f76cac79218e28c4a8ffc6cb77470ba25f7e9a06671cc16d6

Observation 68e5921d-2ec6-4578-8812-980a02904756 · inbound

How to Guide Your Flow: Few-Step Alignment via Flow Map Reward Guidance cites this paper.

How to Guide Your Flow: Few-Step Alignment via Flow Map Reward Guidance Skywork-VL Reward: An Effective Reward Model for Multimodal Understanding and Reasoning

Reference 64

Resolution
verified exact
arxiv_id, observed 2026-07-01T08:25:33.334411Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-07-01T08:19:37.044714Z digest=sha256:e24dce1f459a1897db7fabf8071b4e765a48cc8120e06e7f8c58c062e9cdfedc

Observation e45cb040-cb95-4b16-8068-23c9f4e0fb6a · inbound

Video Understanding Reward Modeling: A Robust Benchmark and Performant Reward Models cites this paper.

Video Understanding Reward Modeling: A Robust Benchmark and Performant Reward Models Skywork-VL Reward: An Effective Reward Model for Multimodal Understanding and Reasoning

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-11T03:45:58.370830Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-11T02:18:20.880231Z digest=sha256:28bdd107164282537d0ba8b5c6c3b1315a925f9f1fba672e3a8f62348b658571