Pith. sign in

Paper Citation Record · LEDGER

ViBe: A Text-to-Video Benchmark for Evaluating Hallucination in Large Multimodal Models

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 3 inbound Pith citation observations for arXiv:2411.10867.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.10867 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 3 of 3 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:09:00.950330Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-16T04:21:29.768609Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation bfe830b5-088e-4fdb-91af-422c53f1b5e6 · inbound

Dynamic-I2V: Exploring Image-to-Video Generation Models via Multimodal LLM cites this paper.

Dynamic-I2V: Exploring Image-to-Video Generation Models via Multimodal LLM ViBe: A Text-to-Video Benchmark for Evaluating Hallucination in Large Multimodal Models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:00.950330Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:00.950330Z digest=sha256:d662291a6774ad23f37e8ebf2adf165c7c1381b92e99af9196d421ed362192aa

Observation 0820b5a3-e805-40ee-adce-0c285e84c68a · inbound

FIFA: Unified Faithfulness Evaluation Framework for Text-to-Video and Video-to-Text Generation cites this paper.

FIFA: Unified Faithfulness Evaluation Framework for Text-to-Video and Video-to-Text Generation ViBe: A Text-to-Video Benchmark for Evaluating Hallucination in Large Multimodal Models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T19:06:15.846144Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:06:15.846144Z digest=sha256:e0a321b50a9bff57b5c99fcc69bb98af100586ff99c91540241aeff1c14e2662

Observation 844b1adc-4d2f-45fb-8d31-d5cd3e3a88d7 · inbound

Molmo2: Open Weights and Data for Vision-Language Models with Video Understanding and Grounding cites this paper.

Molmo2: Open Weights and Data for Vision-Language Models with Video Understanding and Grounding ViBe: A Text-to-Video Benchmark for Evaluating Hallucination in Large Multimodal Models

Reference 124

Resolution
verified exact
arxiv_id, observed 2026-05-16T04:21:29.771928Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-16T04:21:29.526008Z digest=sha256:003aad3b678f4e58ea229987829c9402f2e1ae588c37c35426b0340fc6adb4b4