Pith. sign in

Paper Citation Record · LEDGER

WorldQA: Multimodal World Knowledge in Videos through Long-Chain Reasoning

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 4 inbound Pith citation observations for arXiv:2405.03272.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2405.03272 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 4 of 4 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T19:03:06.462504Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T23:06:21.362709Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation d9882c54-9125-4896-870d-a8a50bc149c4 · inbound

Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos cites this paper.

Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos WorldQA: Multimodal World Knowledge in Videos through Long-Chain Reasoning

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-05-14T00:32:41.207259Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-14T00:32:41.059558Z digest=sha256:969b8efd66d5bc09af5084ac7fd97de512e4a3bacd61a734ac14cf23dc42c0af

Observation adb17a9c-1322-40b5-aec0-5be0cdf6e821 · inbound

Multimodal Chain-of-Thought Reasoning: A Comprehensive Survey cites this paper.

Multimodal Chain-of-Thought Reasoning: A Comprehensive Survey WorldQA: Multimodal World Knowledge in Videos through Long-Chain Reasoning

Reference 175

Resolution
verified exact
arxiv_id, observed 2026-05-15T17:18:53.297298Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-15T17:18:52.996467Z digest=sha256:292b0a9d404b7eea46afc62a8404d2076839d619303c8b0ce5e1e8ae29c2a34a

Observation 9d63197b-3105-4dd8-b02b-0a042dbec751 · inbound

HumanPCR: Probing MLLM Capabilities in Diverse Human-Centric Scenes cites this paper.

HumanPCR: Probing MLLM Capabilities in Diverse Human-Centric Scenes WorldQA: Multimodal World Knowledge in Videos through Long-Chain Reasoning

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-05T19:03:06.462504Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:03:06.462504Z digest=sha256:5b97793dd2da1508bc93b6a055c5897c508111646b870e26745b27828917c063

Observation 6f764bd1-c41f-4b79-b7b9-0cde6e78e852 · inbound

AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs cites this paper.

AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs WorldQA: Multimodal World Knowledge in Videos through Long-Chain Reasoning

Reference 35

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T23:06:21.364975Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-28T14:36:53.295540Z digest=sha256:e367edf7c4bac727d23f435348e655bc41b77c3ef23cf37465a087e6ab40e0d3