Pith. sign in

Paper Citation Record · LEDGER

WorldQA: Multimodal World Knowledge in Videos through Long-Chain Reasoning

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2405.03272.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2405.03272 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:05:19.768115Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T23:06:21.362709Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation d9882c54-9125-4896-870d-a8a50bc149c4 · inbound

Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos cites this paper.

Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos WorldQA: Multimodal World Knowledge in Videos through Long-Chain Reasoning

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-05-14T00:32:41.207259Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-14T00:32:41.059558Z digest=sha256:5d738a6ddbaae22f5c6af02f94c7d012e82c4f43bdbb8f3ec6512ba2778c6ba5

Observation adb17a9c-1322-40b5-aec0-5be0cdf6e821 · inbound

Multimodal Chain-of-Thought Reasoning: A Comprehensive Survey cites this paper.

Multimodal Chain-of-Thought Reasoning: A Comprehensive Survey WorldQA: Multimodal World Knowledge in Videos through Long-Chain Reasoning

Reference 175

Resolution
verified exact
arxiv_id, observed 2026-05-15T17:18:53.297298Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T17:18:52.996467Z digest=sha256:3eea4d061330be6248472872182edbdb23c1c868c26c0ff210bc2de372e2fcc8

Observation 04bcbfa7-9314-4b0c-83ac-9331cbf40313 · inbound

Evaluating MLLMs with Multimodal Multi-image Reasoning Benchmark cites this paper.

Evaluating MLLMs with Multimodal Multi-image Reasoning Benchmark WorldQA: Multimodal World Knowledge in Videos through Long-Chain Reasoning

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T11:05:19.768115Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:05:19.768115Z digest=sha256:c2a97e12ce3de46de8b0126ca90dde2c076047fa5d53489a63c758ced6498ddb

Observation 6a7746b3-9374-467e-9557-a23862a0ad83 · inbound

VideoMathQA: Benchmarking Mathematical Reasoning via Multimodal Understanding in Videos cites this paper.

VideoMathQA: Benchmarking Mathematical Reasoning via Multimodal Understanding in Videos WorldQA: Multimodal World Knowledge in Videos through Long-Chain Reasoning

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T10:25:36.351474Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:25:36.351474Z digest=sha256:bf6f457bac185a13aed3e939f4c405ef637331bcd848b949547b865d0d793b3d

Observation 9d63197b-3105-4dd8-b02b-0a042dbec751 · inbound

HumanPCR: Probing MLLM Capabilities in Diverse Human-Centric Scenes cites this paper.

HumanPCR: Probing MLLM Capabilities in Diverse Human-Centric Scenes WorldQA: Multimodal World Knowledge in Videos through Long-Chain Reasoning

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-05T19:03:06.462504Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:03:06.462504Z digest=sha256:5e689caffeaec065fe7d05e9a518ef181e779c3afd7392eef5907a8c363e7eca

Observation 6f764bd1-c41f-4b79-b7b9-0cde6e78e852 · inbound

AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs cites this paper.

AVI-Bench: Toward Human-like Audio-Visual Intelligence of Omni-MLLMs WorldQA: Multimodal World Knowledge in Videos through Long-Chain Reasoning

Reference 35

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T23:06:21.364975Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-28T14:36:53.295540Z digest=sha256:8b1f18f4df83b9e3c5f075a7d5693b481a301c26b881ad32896c5bd297c55e4e