Pith. sign in

Paper Citation Record · LEDGER

CaReBench: A Fine-Grained Benchmark for Video Captioning and Retrieval

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2501.00513.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.00513 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 5 of 5 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:14:45.772749Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T00:17:28.977461Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 262fe47c-c479-434d-ad18-cf9fbb39a787 · inbound

VideoChat-R1: Enhancing Spatio-Temporal Perception via Reinforcement Fine-Tuning cites this paper.

VideoChat-R1: Enhancing Spatio-Temporal Perception via Reinforcement Fine-Tuning CaReBench: A Fine-Grained Benchmark for Video Captioning and Retrieval

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-15T20:56:07.717041Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T20:56:07.247122Z digest=sha256:34b81ee0ae3eefba9da5037d97408029cc562ca9f137781b17e2862a8d313bd7

Observation ff5b82ca-a31b-43ae-aebe-5614bf80505d · inbound

Modality Curation: Building Universal Embeddings for Advanced Multimodal Information Retrieval cites this paper.

Modality Curation: Building Universal Embeddings for Advanced Multimodal Information Retrieval CaReBench: A Fine-Grained Benchmark for Video Captioning and Retrieval

Reference 96

Resolution
unresolved
no resolver link, observed 2026-08-07T14:14:45.772749Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:14:45.772749Z digest=sha256:5968686655db51cfb3577661395b4dc804b12b170ee8c12858f46c51d1b57ca8

Observation c42b6749-3580-4a62-b6eb-d81860a6ce66 · inbound

VideoCap-R1: Enhancing MLLMs for Video Captioning via Structured Thinking cites this paper.

VideoCap-R1: Enhancing MLLMs for Video Captioning via Structured Thinking CaReBench: A Fine-Grained Benchmark for Video Captioning and Retrieval

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T11:40:45.357734Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:40:45.357734Z digest=sha256:9735e0c6cfe11fdda18e7491a61f850ec31609f41fc4371713ddb3d5554365e8

Observation e0a0f027-d7dc-45d1-a39f-a41bddca0301 · inbound

CapRL++: Unified Reinforcement Learning with Verifiable Rewards for Dense Image and Video Captioning cites this paper.

CapRL++: Unified Reinforcement Learning with Verifiable Rewards for Dense Image and Video Captioning CaReBench: A Fine-Grained Benchmark for Video Captioning and Retrieval

Reference 89

Resolution
verified exact
arxiv_id, observed 2026-07-03T00:17:28.980461Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T17:21:38.543724Z digest=sha256:7149f4605740ac318c150bdca5abe2a5bcf40fa69f5cb20cf60f8642bfcbe791

Observation 403360a5-3898-4e15-b88c-8971218457de · inbound

PercepCap: Video Captioner with Structured Spatio-Temporal Perception cites this paper.

PercepCap: Video Captioner with Structured Spatio-Temporal Perception CaReBench: A Fine-Grained Benchmark for Video Captioning and Retrieval

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-01T10:02:03.257559Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T10:02:03.257559Z digest=sha256:150839c8d505bc9327253976e8894ea24ef55874d77eefca36b7635ea4cd30ed