Pith. sign in

Paper Citation Record · LEDGER

ClaimDiff-RL: Fine-Grained Caption Reinforcement Learning through Visual Claim Comparison

As of 5 August 2026, this Paper Citation Record lists 50 of 50 outbound references and 0 inbound Pith citation observations for arXiv:2605.20278.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2605.20278 v2

Coverage vector

measured 50 of 50 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-06-30T17:57:47.409741Z

measured 50 of 50 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

50 of 50 outbound references displayed

  • verified exact15
  • verified fuzzy19
  • unresolved13
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch3

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 202b597d-3aaf-4a2b-a1e1-ba64a0e7b169 · outbound

This paper cites SPICE: Semantic Propositional Image Caption Evaluation.

ClaimDiff-RL: Fine-Grained Caption Reinforcement Learning through Visual Claim Comparison SPICE: Semantic Propositional Image Caption Evaluation

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-06-30T18:04:58.328227Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T17:57:47.409741Z digest=sha256:3b8f5899c5b95c68b51f35e52224e3ff5fa95c533102d738f60c86cf24a49bda

Observation 164d6e08-f470-49df-8dd9-c0388e2b5717 · outbound

This paper cites Qwen3-VL Technical Report.

ClaimDiff-RL: Fine-Grained Caption Reinforcement Learning through Visual Claim Comparison Qwen3-VL Technical Report

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-06-30T18:04:58.322968Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T17:57:47.409741Z digest=sha256:141e85625d1ca4d708c49a00acc83ec749188527688adfbf63fcb0963842865d

Observation d5969fc1-a13b-4b36-b50b-58ba14c051dd · outbound

This paper cites Meteor: An automatic metric for mt evaluation with improved correlation with human judgments.

ClaimDiff-RL: Fine-Grained Caption Reinforcement Learning through Visual Claim Comparison Meteor: An automatic metric for mt evaluation with improved correlation with human judgments

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T06:44:40.863740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T17:57:47.409741Z digest=sha256:aebaf0a8cf65feebcdc75f333b0320ea69e6efaa05e0864733f21aed73b6ed70

Observation cb6a31f9-1896-4b26-b5ed-1b1f5a00630f · outbound

This paper cites CLAIR: Evaluating Image Captions with Large Language Models.

ClaimDiff-RL: Fine-Grained Caption Reinforcement Learning through Visual Claim Comparison CLAIR: Evaluating Image Captions with Large Language Models

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-06-30T18:04:58.325835Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T17:57:47.409741Z digest=sha256:301341645bafa828fd4b35c19911f9b1d7b12b3cb8092bf5054d81e4e219195f

Observation 85767527-49cc-4d97-b275-de06a3520f70 · outbound

This paper cites Simplevqa: Multimodal factuality evaluation for multimodal large language models.2025 IEEE/CVF International Conference on Computer Vision (ICCV), pages 4637–4646.

ClaimDiff-RL: Fine-Grained Caption Reinforcement Learning through Visual Claim Comparison Simplevqa: Multimodal factuality evaluation for multimodal large language models.2025 IEEE/CVF International Conference on Computer Vision (ICCV), pages 4637–4646

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T06:44:40.821858Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T17:57:47.409741Z digest=sha256:ba7a6f08f191d993f376ab992807cceb2d89f4d335c4c2d0f50aa07bcad2357e

Observation 01a168a1-f316-41bb-815b-d6fce0b9e254 · outbound

This paper cites Benchmarking and Improving Detail Image Caption.

ClaimDiff-RL: Fine-Grained Caption Reinforcement Learning through Visual Claim Comparison Benchmarking and Improving Detail Image Caption

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-06-30T18:04:58.330839Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T17:57:47.409741Z digest=sha256:86db6341927c301e75d818aee50cbc8ac1cc88ed1d9678addc77f3ff3084fd2c

Observation 2c7155a6-e2c7-4136-89cb-c535243115f1 · outbound

This paper cites OCRBench v2: An Improved Benchmark for Evaluating Large Multimodal Models on Visual Text Localization and Reasoning.

ClaimDiff-RL: Fine-Grained Caption Reinforcement Learning through Visual Claim Comparison OCRBench v2: An Improved Benchmark for Evaluating Large Multimodal Models on Visual Text Localization and Reasoning

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-06-30T18:04:58.316832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T17:57:47.409741Z digest=sha256:92c12dfa49fa75d01a601c53eb5547643774b004f0fc90086661bd1e4c07385f

Observation b280e90a-091a-4308-a785-3ce2c8c263d9 · outbound

This paper cites BLINK: Multimodal Large Language Models Can See but Not Perceive.

ClaimDiff-RL: Fine-Grained Caption Reinforcement Learning through Visual Claim Comparison BLINK: Multimodal Large Language Models Can See but Not Perceive

Reference 8

Resolution
metadata mismatch
local_arxiv, observed 2026-06-30T18:04:58.341972Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T17:57:47.409741Z digest=sha256:67a4c384d29f7d809afaf11e5fffef31f6f73718ad7769d38e63ea809ea560d4

Observation cc80062b-c6f6-46c5-be8d-15b48b4c6989 · outbound

This paper cites DataComp: In search of the next generation of multimodal datasets.

ClaimDiff-RL: Fine-Grained Caption Reinforcement Learning through Visual Claim Comparison DataComp: In search of the next generation of multimodal datasets

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-06-30T18:04:58.323685Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T17:57:47.409741Z digest=sha256:5dfe4bc3728659885e50484ef13d2a9914f56f96ebd52314efd67ce191514090

Observation e8527aef-a42b-4e14-bd4e-8a0d96871f3a · outbound

This paper cites an unresolved cited work.

ClaimDiff-RL: Fine-Grained Caption Reinforcement Learning through Visual Claim Comparison Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-07-08T06:44:40.816420Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T17:57:47.409741Z digest=sha256:a640123a6e9a17a241fe4cd14cdf8b4bf353570e064481e0c9c2f2ded567cc09

Observation 33398531-6074-4991-b97c-3551cf8118f2 · outbound

This paper cites CLIPScore: A Reference-free Evaluation Metric for Image Captioning.

ClaimDiff-RL: Fine-Grained Caption Reinforcement Learning through Visual Claim Comparison CLIPScore: A Reference-free Evaluation Metric for Image Captioning

Reference 11

Resolution
verified exact
local_arxiv, observed 2026-06-30T18:04:58.339542Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T17:57:47.409741Z digest=sha256:6bfc798d0cd67db747ee6b9fdb8b38bb0e48d9b5049b9cc5fc18cb658f770416

Observation 1d5ead44-912d-415a-9ff1-b58cfc52d902 · outbound

This paper cites VIEScore: Towards Explainable Metrics for Conditional Image Synthesis Evaluation.

ClaimDiff-RL: Fine-Grained Caption Reinforcement Learning through Visual Claim Comparison VIEScore: Towards Explainable Metrics for Conditional Image Synthesis Evaluation

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-06-30T18:04:58.344776Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T17:57:47.409741Z digest=sha256:5cebeba68e455b1acfda27d97782ece4415d4279fcfa8e37ec63c8b829ef171d

Observation b771662a-7a68-4334-82be-da0249707b07 · outbound

This paper cites Prometheus- vision: Vision-language model as a judge for fine-grained evaluation.

ClaimDiff-RL: Fine-Grained Caption Reinforcement Learning through Visual Claim Comparison Prometheus- vision: Vision-language model as a judge for fine-grained evaluation

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T06:44:40.833705Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T17:57:47.409741Z digest=sha256:1a2f981faab371eb3dcfbb82fc072c6c59fb6289fc9c1e862480d283a506749c

Observation ec678de2-6e99-4496-b25e-4ea439317eac · outbound

This paper cites Evaluating object hallucination in large vision-language models.

ClaimDiff-RL: Fine-Grained Caption Reinforcement Learning through Visual Claim Comparison Evaluating object hallucination in large vision-language models

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T06:44:40.840321Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T17:57:47.409741Z digest=sha256:7617c07c58c04eb7d39a259976124ce2f882b1ca09862df27b99e46d2bf82e21

Observation 0258ed1f-f829-4ffb-8b77-265bda24120a · outbound

This paper cites Describe Anything: Detailed Localized Image and Video Captioning.

ClaimDiff-RL: Fine-Grained Caption Reinforcement Learning through Visual Claim Comparison Describe Anything: Detailed Localized Image and Video Captioning

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-06-30T18:04:58.306030Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T17:57:47.409741Z digest=sha256:050dac4a001c164fba48cb7810ecddb6e4ea27d040ebef00e77b2c6cf967f42f

Observation e535b3bf-6bd4-4b87-90a9-2a5aa3d8087f · outbound

This paper cites Capability: A comprehensive visual caption benchmark for evaluating both correctness and thoroughness.

ClaimDiff-RL: Fine-Grained Caption Reinforcement Learning through Visual Claim Comparison Capability: A comprehensive visual caption benchmark for evaluating both correctness and thoroughness

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T06:44:40.796583Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T17:57:47.409741Z digest=sha256:e42d7147c25e10473f4f35532330b742c52afc9e428bbe7f2e8d04e05f0b6530

Observation 5a921b0b-c5c9-40a2-b73a-66146222511e · outbound

This paper cites Training language models to follow instructions with human feedback.

ClaimDiff-RL: Fine-Grained Caption Reinforcement Learning through Visual Claim Comparison Training language models to follow instructions with human feedback

Reference 17

Resolution
metadata mismatch
local_arxiv, observed 2026-06-30T18:04:58.328638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T17:57:47.409741Z digest=sha256:0c0c4dccb268a6c692ef1f62e93cfcc90a1466d972238561f63e9852d0e3dd77

Observation ac60f457-c5f4-4484-a788-20441bb6f57e · outbound

This paper cites Bleu: a method for automatic evaluation of machine translation.

ClaimDiff-RL: Fine-Grained Caption Reinforcement Learning through Visual Claim Comparison Bleu: a method for automatic evaluation of machine translation

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T06:44:40.823012Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T17:57:47.409741Z digest=sha256:6b35aa8e76ee96a48d542ea8932a783297c9d1484760ffb573cdfd4db2c0bca4

Observation f326c86f-c7ea-4f6f-b2ee-9ea29065b9f5 · outbound

This paper cites Direct Preference Optimization: Your Language Model is Secretly a Reward Model.

ClaimDiff-RL: Fine-Grained Caption Reinforcement Learning through Visual Claim Comparison Direct Preference Optimization: Your Language Model is Secretly a Reward Model

Reference 19

Resolution
metadata mismatch
local_arxiv, observed 2026-06-30T18:04:58.336874Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T17:57:47.409741Z digest=sha256:4e6198107f1b0e6285e63a7542fe0628f0a0bd8beef6c650ed1dd7c386f58ec8

Observation 5e2ffbba-33e4-4ab9-83ed-2cab41ff948f · outbound

This paper cites Rennie, Etienne Marcheret, Youssef Mroueh, Jarret Ross, and Vaibhava Goel.

ClaimDiff-RL: Fine-Grained Caption Reinforcement Learning through Visual Claim Comparison Rennie, Etienne Marcheret, Youssef Mroueh, Jarret Ross, and Vaibhava Goel

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T06:44:40.818871Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T17:57:47.409741Z digest=sha256:d19f971a7e3f69129dc97d020fa85f07ae6c13f9dec114c21e739290e497b3bb

Observation 477a2f88-8086-4b5a-a52d-a13134a89063 · outbound

This paper cites Object hallucination in image captioning.

ClaimDiff-RL: Fine-Grained Caption Reinforcement Learning through Visual Claim Comparison Object hallucination in image captioning

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T06:44:40.793572Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T17:57:47.409741Z digest=sha256:b70dd719acc5b68493e6d09cad42eab3929aaa0bb803181ca2a92003ebfd92e1

Observation 890af3bf-eaf9-49db-a635-fbd3ca5f14f7 · outbound

This paper cites Laion-5b: An open large-scale dataset for training next generation image-text models.

ClaimDiff-RL: Fine-Grained Caption Reinforcement Learning through Visual Claim Comparison Laion-5b: An open large-scale dataset for training next generation image-text models

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T06:44:40.824504Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T17:57:47.409741Z digest=sha256:9c9d8357298574cf46a1fb6b08cafa03c668a0b37901e9052ab28dcf5f21012c

Observation 3c368f15-bf23-4746-b2ba-c99d14459c07 · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

ClaimDiff-RL: Fine-Grained Caption Reinforcement Learning through Visual Claim Comparison DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 23

Resolution
verified exact
local_arxiv, observed 2026-06-30T18:04:58.326219Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T17:57:47.409741Z digest=sha256:fe39dd4b23268a7a28c3fa6319e2f9f9dd6e5cafa972ba501117ac71259b0231

Observation 9a919184-911f-4f75-b596-cff3a269ae64 · outbound

This paper cites an unresolved cited work.

ClaimDiff-RL: Fine-Grained Caption Reinforcement Learning through Visual Claim Comparison Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-07-08T06:44:40.807066Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T17:57:47.409741Z digest=sha256:3202b9a8185e8b0609e14e4609f2113c445bc0b3e145952d7b6e7de0c433e617

Observation 497e6852-f409-4cf6-88f2-77ef7d44cbdd · outbound

This paper cites Aligning Large Multimodal Models with Factually Augmented RLHF.

ClaimDiff-RL: Fine-Grained Caption Reinforcement Learning through Visual Claim Comparison Aligning Large Multimodal Models with Factually Augmented RLHF

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-06-30T18:04:58.347388Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T17:57:47.409741Z digest=sha256:2c93bba49dccd8c3bcff802dbe25e1b21926819c29bcec94f6285b2a070cd14c

Observation fc3499c6-148d-44f5-a8d0-5b8a641f26bb · outbound

This paper cites an unresolved cited work.

ClaimDiff-RL: Fine-Grained Caption Reinforcement Learning through Visual Claim Comparison Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-07-08T06:44:40.813496Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T17:57:47.409741Z digest=sha256:e0740722cc7cc4a48113081dfe0223aee813013630cf3f0d356a705c94c8cfb0

Observation 5b7ac8b4-3ce2-40fe-8706-c2bcbf3cf0c9 · outbound

This paper cites Lawrence Zitnick, and Devi Parikh.

ClaimDiff-RL: Fine-Grained Caption Reinforcement Learning through Visual Claim Comparison Lawrence Zitnick, and Devi Parikh

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T06:44:40.810687Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T17:57:47.409741Z digest=sha256:7152bdd4281c9c779988344ff58ccaf628bc234fda765751a3147e1c999098a1

Observation 23609eb5-ccdf-4b80-8646-33284dee9746 · outbound

This paper cites Grasp any region: Towards precise, contextual pixel understanding for multimodal llms.ArXiv, abs/2510.18876.

ClaimDiff-RL: Fine-Grained Caption Reinforcement Learning through Visual Claim Comparison Grasp any region: Towards precise, contextual pixel understanding for multimodal llms.ArXiv, abs/2510.18876

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-06-30T18:04:58.350347Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T17:57:47.409741Z digest=sha256:ef3ebd1761319e0914a47340ebac436b27f5ea4a0ef2da1b0aca8098d3eccb26

Observation f5651ffc-427a-40e7-877f-47917c8181a4 · outbound

This paper cites Divide, Conquer and Combine: A Training-Free Framework for High-Resolution Image Perception in Multimodal Large Language Models.

ClaimDiff-RL: Fine-Grained Caption Reinforcement Learning through Visual Claim Comparison Divide, Conquer and Combine: A Training-Free Framework for High-Resolution Image Perception in Multimodal Large Language Models

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-06-30T18:04:58.341749Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T17:57:47.409741Z digest=sha256:1a652d955554b89d3bc7d5de53c5ed56cd9528fc052eecb4540930491c121f60

Observation 3bc28391-656a-4359-b517-46c94055bba6 · outbound

This paper cites Vicrit: A verifiable reinforcement learning proxy task for visual perception in vlms.

ClaimDiff-RL: Fine-Grained Caption Reinforcement Learning through Visual Claim Comparison Vicrit: A verifiable reinforcement learning proxy task for visual perception in vlms

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T06:44:40.849294Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T17:57:47.409741Z digest=sha256:c1ba2a518f214db0c689b1dc41ed9ae309e2973b16c71f333c1686a7396af26f

Observation 4a2adb6e-e098-4fc4-a0c6-eae351e5f508 · outbound

This paper cites Realworldqa.

ClaimDiff-RL: Fine-Grained Caption Reinforcement Learning through Visual Claim Comparison Realworldqa

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T06:44:40.868803Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T17:57:47.409741Z digest=sha256:5574087fbca94e363d386064afe8d110b478ec3ad33a541e15440e7aaf009d6b

Observation f5cfb4a4-bb06-483d-825a-2f36eedf8620 · outbound

This paper cites Caprl: Stimulating dense image caption capabilities via reinforcement learning.

ClaimDiff-RL: Fine-Grained Caption Reinforcement Learning through Visual Claim Comparison Caprl: Stimulating dense image caption capabilities via reinforcement learning

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T06:44:40.838788Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T17:57:47.409741Z digest=sha256:c00eeb4558e6394b6d16309cee3fa2f72d6baabc476ff6965859cec12eebdcf8

Observation 41943c0c-0556-415c-a68b-4f9e3379d8a3 · outbound

This paper cites CaptionQA: Is Your Caption as Useful as the Image Itself?.

ClaimDiff-RL: Fine-Grained Caption Reinforcement Learning through Visual Claim Comparison CaptionQA: Is Your Caption as Useful as the Image Itself?

Reference 33

Resolution
verified exact
local_arxiv, observed 2026-06-30T18:04:58.336063Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T17:57:47.409741Z digest=sha256:a5c02c4c747afe22c61cf1990c5635cd2eea3810dcf3a481b7922a5b6fb931b2

Observation cf9e781d-3689-4981-85ce-7589f9d91bf6 · outbound

This paper cites Sc-captioner: Improving image captioning with self-correction by reinforcement learning.

ClaimDiff-RL: Fine-Grained Caption Reinforcement Learning through Visual Claim Comparison Sc-captioner: Improving image captioning with self-correction by reinforcement learning

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T06:44:40.851512Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T17:57:47.409741Z digest=sha256:6e132c2f969cc286361e6438cbafb0df3085f08aa3f34d9aed6f4872a79e10b8

Observation b675ee7c-31ab-47da-8984-70f9b25599a6 · outbound

This paper cites Please describe this 18 image in detail.

ClaimDiff-RL: Fine-Grained Caption Reinforcement Learning through Visual Claim Comparison Please describe this 18 image in detail

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T06:44:40.854116Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T17:57:47.409741Z digest=sha256:da1a003a672bb364948c8839aef0c9b5b01974c88398ddad58e2f2b01bc58da2

Observation 360a0039-0231-450e-a164-63d966dc384b · outbound

This paper cites Important rules: - A correct detail in the actor caption should not be penalized merely because it is absent from the reference.

ClaimDiff-RL: Fine-Grained Caption Reinforcement Learning through Visual Claim Comparison Important rules: - A correct detail in the actor caption should not be penalized merely because it is absent from the reference

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T06:44:40.858879Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T17:57:47.409741Z digest=sha256:b5ccedb33c2143c0773195033d529776f5d2aecb18f3bdae5acd66b386dd0025

Observation 2629b8ea-2788-4f68-ac03-0b63bdfc5004 · outbound

This paper cites an unresolved cited work.

ClaimDiff-RL: Fine-Grained Caption Reinforcement Learning through Visual Claim Comparison Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-07-08T06:44:40.846727Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T17:57:47.409741Z digest=sha256:81a8e778564519a47d99000958811244bf3e3b9f24cd93e1ca60379ed7a0cfaf

Observation 5c3f46b1-39af-4e9e-8f7f-ec20ed4d3080 · outbound

This paper cites an unresolved cited work.

ClaimDiff-RL: Fine-Grained Caption Reinforcement Learning through Visual Claim Comparison Unresolved cited work

Reference 42

Resolution
unresolved
raw_fallback, observed 2026-07-08T06:44:40.849050Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T17:57:47.409741Z digest=sha256:58897b765c860ece245d39ceaa5d5d901be3ce5a98d6be8e3b03b1edfabb6c2f

Observation 6c8ba550-db97-4844-bf03-6a5359cd3cd6 · outbound

This paper cites an unresolved cited work.

ClaimDiff-RL: Fine-Grained Caption Reinforcement Learning through Visual Claim Comparison Unresolved cited work

Reference 43

Resolution
unresolved
raw_fallback, observed 2026-07-08T06:44:40.856579Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T17:57:47.409741Z digest=sha256:282fcf34061b9781689828e2e7f4e2a08c911ac8913ecbfb37c99d2ad233b35e

Observation bea42471-b43a-46ab-88e8-8ebaac4b82b2 · outbound

This paper cites an unresolved cited work.

ClaimDiff-RL: Fine-Grained Caption Reinforcement Learning through Visual Claim Comparison Unresolved cited work

Reference 44

Resolution
unresolved
raw_fallback, observed 2026-07-08T06:44:40.844257Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T17:57:47.409741Z digest=sha256:2ec30516b270a5313b60e44a7678788d74174cd890c77871d4042405a4379f78

Observation b4f1c1d0-305a-4b00-ac77-1f1c5ab34474 · outbound

This paper cites Important rules: - Penalize hallucination more than omission.

ClaimDiff-RL: Fine-Grained Caption Reinforcement Learning through Visual Claim Comparison Important rules: - Penalize hallucination more than omission

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T06:44:40.866027Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T17:57:47.409741Z digest=sha256:ceba0ebc6452d4f26a720501338c4ece3250011b1f87c990411d6b55d1e78114

Observation 22912e29-8dba-443e-aedf-1ee9104a1d55 · outbound

This paper cites A contradiction is a candidate claim that conflicts with the reference.

ClaimDiff-RL: Fine-Grained Caption Reinforcement Learning through Visual Claim Comparison A contradiction is a candidate claim that conflicts with the reference

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T06:44:40.874624Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T17:57:47.409741Z digest=sha256:7330c7fe0bcb89b5d865a0a3a5af409ba299ba9852c3604f29f2b58185b3b19b

Observation 73a9e897-14fe-4b70-91c7-cc3d2f63d527 · outbound

This paper cites missing" and is_hallucination=false Strict mapping rules for is_hallucination: - type=.

ClaimDiff-RL: Fine-Grained Caption Reinforcement Learning through Visual Claim Comparison missing" and is_hallucination=false Strict mapping rules for is_hallucination: - type=

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-07-08T06:44:40.845935Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T17:57:47.409741Z digest=sha256:6f51a1810f4efcd7a9b52d1e532e497283efe017b026d8ffd5ff8f47f7f436a0

Observation 0b646652-8957-455d-aadc-ae93c556291e · outbound

This paper cites an unresolved cited work.

ClaimDiff-RL: Fine-Grained Caption Reinforcement Learning through Visual Claim Comparison Unresolved cited work

Reference 48

Resolution
unresolved
raw_fallback, observed 2026-07-08T06:44:40.835033Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T17:57:47.409741Z digest=sha256:16fa49ae5fd6cd10fde5261cdd9d718a5033a2dbda364dbf5ed75d75a7520233

Observation a32f7731-a9f5-4357-92ba-0511daca04ab · outbound

This paper cites an unresolved cited work.

ClaimDiff-RL: Fine-Grained Caption Reinforcement Learning through Visual Claim Comparison Unresolved cited work

Reference 49

Resolution
unresolved
raw_fallback, observed 2026-07-08T06:44:40.841705Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T17:57:47.409741Z digest=sha256:d81c7c38b050a54af69dac800ad4a8431fab457756267e7db9c68253be9a9931

Observation fcaa91af-7613-40ec-ba4c-421c1779e95d · outbound

This paper cites an unresolved cited work.

ClaimDiff-RL: Fine-Grained Caption Reinforcement Learning through Visual Claim Comparison Unresolved cited work

Reference 50

Resolution
unresolved
raw_fallback, observed 2026-07-08T06:44:40.876872Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T17:57:47.409741Z digest=sha256:7c6de9db128fe823c6e72f39423a010ac2ee3caeea198c9911e19f2342ecaf44

Observation e9b7f836-959c-41ba-bba0-89d14d1f8d39 · outbound

This paper cites an unresolved cited work.

ClaimDiff-RL: Fine-Grained Caption Reinforcement Learning through Visual Claim Comparison Unresolved cited work

Reference 51

Resolution
unresolved
raw_fallback, observed 2026-07-08T06:44:40.861105Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T17:57:47.409741Z digest=sha256:e242334c50d89a24b60c48120e42d467209d2c309c6fd3806140d64037e7293a

Observation 8910c7f5-d178-4a79-8fa3-c2fd74224fb5 · outbound

This paper cites an unresolved cited work.

ClaimDiff-RL: Fine-Grained Caption Reinforcement Learning through Visual Claim Comparison Unresolved cited work

Reference 52

Resolution
unresolved
raw_fallback, observed 2026-07-08T06:44:40.871663Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T17:57:47.409741Z digest=sha256:926a0c02772194538c1ac9c72b83f5db68cb526c3ef9f9f23f0c7969c9f2bdd3

Observation 1e0f0f2b-fd64-4f0c-802e-0f288db7db3c · outbound

This paper cites an unresolved cited work.

ClaimDiff-RL: Fine-Grained Caption Reinforcement Learning through Visual Claim Comparison Unresolved cited work

Reference 53

Resolution
unresolved
raw_fallback, observed 2026-07-08T06:44:40.829938Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T17:57:47.409741Z digest=sha256:fb4bb3e381544a1df85516678f2619194384ee09d8b8718229d0a12ba679e0db

Observation 2ec98524-8838-4b6b-895e-705ab8a2ef56 · outbound

This paper cites claims": [ {.

ClaimDiff-RL: Fine-Grained Caption Reinforcement Learning through Visual Claim Comparison claims": [ {

Reference 54

Resolution
verified exact
arxiv_id, observed 2026-06-30T18:04:58.338821Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T17:57:47.409741Z digest=sha256:3154336a9267e223c82ebdd25b8e05500aa04e933ad79f5638414c76b0d0b2a0

Pith citing papers

No inbound Pith citation observations are available.