Pith. sign in

Paper Citation Record · LEDGER

MiraData: A Large-Scale Video Dataset with Long Durations and Structured Captions

As of 19 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 19 inbound Pith citation observations for arXiv:2407.06358.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2407.06358 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 19 of 19 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 19 of 19 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T11:46:40.567769Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T10:48:02.975279Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 803514b4-b1ad-446b-b1be-b2fa3e211f81 · inbound

MovieBench: A Hierarchical Movie Level Dataset for Long Video Generation cites this paper.

MovieBench: A Hierarchical Movie Level Dataset for Long Video Generation MiraData: A Large-Scale Video Dataset with Long Durations and Structured Captions

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-12T14:53:11.647456Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:53:11.647456Z digest=sha256:d341342c1608b64aae0d89c6ac8e724244defd2f0efba6e2b335c0ee124fa86c

Observation ecb673a5-47a4-4bd5-815c-1231256f4f9b · inbound

SALOVA: Segment-Augmented Long Video Assistant for Targeted Retrieval and Routing in Long-Form Video Analysis cites this paper.

SALOVA: Segment-Augmented Long Video Assistant for Targeted Retrieval and Routing in Long-Form Video Analysis MiraData: A Large-Scale Video Dataset with Long Durations and Structured Captions

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-12T13:31:04.168046Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:31:04.168046Z digest=sha256:882482621fe6be6b35818cd8df0435e1b763d95df3162d3df33188aab4c8902a

Observation c8f53c0f-a6d3-40cf-9677-aec4f9e82a04 · inbound

Trajectory Attention for Fine-grained Video Motion Control cites this paper.

Trajectory Attention for Fine-grained Video Motion Control MiraData: A Large-Scale Video Dataset with Long Durations and Structured Captions

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-12T10:21:23.907337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:21:23.907337Z digest=sha256:16e3c822c0faecb8229a1b345017e1394c5f1e82c02b7486d7cd9e3f7b48a9b8

Observation 832b02b6-ff8c-42e2-bc4a-a6c3a52ced10 · inbound

VISTA: Enhancing Long-Duration and High-Resolution Video Understanding by Video Spatiotemporal Augmentation cites this paper.

VISTA: Enhancing Long-Duration and High-Resolution Video Understanding by Video Spatiotemporal Augmentation MiraData: A Large-Scale Video Dataset with Long Durations and Structured Captions

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-12T04:56:01.245267Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T04:56:01.245267Z digest=sha256:9013914244c5d644a4cb0175569a726c14e19ed5ec599e5f82a23fdbe0f2e178

Observation 6fecf06f-b317-4ca0-a273-e80656179d7d · inbound

Mind the Time: Temporally-Controlled Multi-Event Video Generation cites this paper.

Mind the Time: Temporally-Controlled Multi-Event Video Generation MiraData: A Large-Scale Video Dataset with Long Durations and Structured Captions

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-11T20:55:02.000168Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:55:02.000168Z digest=sha256:165a1cff99f72e169896cc797849748ae9bae705c7f729db7c0c7f3a61fa3b7a

Observation 8ce0e290-1ac3-40ad-934d-864ee0974ff2 · inbound

Mojito: Motion Trajectory and Intensity Control for Video Generation cites this paper.

Mojito: Motion Trajectory and Intensity Control for Video Generation MiraData: A Large-Scale Video Dataset with Long Durations and Structured Captions

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-11T17:25:51.481030Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:25:51.481030Z digest=sha256:5819e5d8d5c04dc25e2d2c93e3147f5f80e5735fc8fd764ac8fdff8e027eb8b4

Observation 5af02d9c-2cc1-411a-8e94-3e020a7ddb98 · inbound

InstanceCap: Improving Text-to-Video Generation via Instance-aware Structured Caption cites this paper.

InstanceCap: Improving Text-to-Video Generation via Instance-aware Structured Caption MiraData: A Large-Scale Video Dataset with Long Durations and Structured Captions

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-11T17:10:21.932432Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T17:10:21.932432Z digest=sha256:414a796a050897a62738bade0baa84104ad277261a289ec98916332a5f03f709

Observation 42c39777-63e8-4890-bceb-1e2291c77520 · inbound

Owl-1: Omni World Model for Consistent Long Video Generation cites this paper.

Owl-1: Omni World Model for Consistent Long Video Generation MiraData: A Large-Scale Video Dataset with Long Durations and Structured Captions

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-11T16:58:05.504241Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:58:05.504241Z digest=sha256:c0ec09d2f6a867e466221ba6f80e444e8b0c22ee5400751d6707b1fadfeab9c5

Observation 3f0d0446-ea3c-409a-9b5f-67fd84e4c6d0 · inbound

PVC: Progressive Visual Token Compression for Unified Image and Video Processing in Large Vision-Language Models cites this paper.

PVC: Progressive Visual Token Compression for Unified Image and Video Processing in Large Vision-Language Models MiraData: A Large-Scale Video Dataset with Long Durations and Structured Captions

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-11T16:56:42.405514Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:56:42.405514Z digest=sha256:db7a659a242038e00b1b3129d1b0d920c3c52357fcd119ba0f4315b59a225ec4

Observation ebc306c6-34c7-4484-8efc-8ab750767b79 · inbound

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity cites this paper.

LinGen: Towards High-Resolution Minute-Length Text-to-Video Generation with Linear Computational Complexity MiraData: A Large-Scale Video Dataset with Long Durations and Structured Captions

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-11T16:43:08.291243Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T16:43:08.291243Z digest=sha256:dd79775c99d9a8e44320432a0d7e1db24a0d9e3ee703addae3edbb39e9ef0d7c

Observation 249d3d11-45b4-4f11-af20-b6aca5a58216 · inbound

VideoMaker: Zero-shot Customized Video Generation with the Inherent Force of Video Diffusion Models cites this paper.

VideoMaker: Zero-shot Customized Video Generation with the Inherent Force of Video Diffusion Models MiraData: A Large-Scale Video Dataset with Long Durations and Structured Captions

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-11T00:10:16.550221Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T00:10:16.550221Z digest=sha256:b4e8a5d4ec72d66da6e8b5851cd2492468f088a23f22838a98e379dbfb485c91

Observation f41c7ea8-baff-483e-98b6-afeb880af81c · inbound

VideoChat-Flash: Hierarchical Compression for Long-Context Video Modeling cites this paper.

VideoChat-Flash: Hierarchical Compression for Long-Context Video Modeling MiraData: A Large-Scale Video Dataset with Long Durations and Structured Captions

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-18T04:02:43.396104Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-18T04:02:43.261543Z digest=sha256:38e44534c7a1a2d4186e74e46573311fe750d5fb841dd21252a1da5a2c9e13c7

Observation 0464e479-621f-4718-9ce0-85088e4a6c8c · inbound

Video Latent Flow Matching: Optimal Polynomial Projections for Video Interpolation and Extrapolation cites this paper.

Video Latent Flow Matching: Optimal Polynomial Projections for Video Interpolation and Extrapolation MiraData: A Large-Scale Video Dataset with Long Durations and Structured Captions

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-09T18:51:12.489164Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:51:12.489164Z digest=sha256:ede0e4bc5df279b92d2a51b77b1551b798f9da6652cd793d05a398a114dfa2e4

Observation aa3150f1-ec66-4108-8b90-a27ce11fbc2d · inbound

Goku: Flow Based Video Generative Foundation Models cites this paper.

Goku: Flow Based Video Generative Foundation Models MiraData: A Large-Scale Video Dataset with Long Durations and Structured Captions

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-08T21:07:32.322663Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T21:07:32.322663Z digest=sha256:79198f6698976d0063cc1eb0d0dfc4e053a6faee017b217337312c369bc9525b

Observation 8ccdcbef-ee78-40f2-afed-7201174e90cd · inbound

Video-MMLU: A Massive Multi-Discipline Lecture Understanding Benchmark cites this paper.

Video-MMLU: A Massive Multi-Discipline Lecture Understanding Benchmark MiraData: A Large-Scale Video Dataset with Long Durations and Structured Captions

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-16T11:46:40.567769Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T11:46:40.567769Z digest=sha256:7fead24630d0ae1a1bee8f357c43066c70ab98649f157991862369b04adf6db5

Observation 81cb9156-82e7-46a4-b478-d22e37d2b0f9 · inbound

TUNA: Comprehensive Fine-grained Temporal Understanding Evaluation on Dense Dynamic Videos cites this paper.

TUNA: Comprehensive Fine-grained Temporal Understanding Evaluation on Dense Dynamic Videos MiraData: A Large-Scale Video Dataset with Long Durations and Structured Captions

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T14:03:00.523964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:03:00.523964Z digest=sha256:d8bda19cbad5e30faa65637c5dc03e828f585d554dd92b03ef4dca9a095f5114

Observation 8bebd46a-7809-49b1-82df-feac7956dab0 · inbound

Video World Models with Long-term Spatial Memory cites this paper.

Video World Models with Long-term Spatial Memory MiraData: A Large-Scale Video Dataset with Long Durations and Structured Captions

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T10:28:38.940419Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:28:38.940419Z digest=sha256:d2f19aae7d3f66d3a5bef9ca23e1c15c376950c152cb074f10a1a9d2a5499479

Observation 595b59ab-a4f6-42df-89ff-33596ea41bb7 · inbound

InternVideo3: Agentify Foundation Models with Multimodal Contextual Reasoning cites this paper.

InternVideo3: Agentify Foundation Models with Multimodal Contextual Reasoning MiraData: A Large-Scale Video Dataset with Long Durations and Structured Captions

Reference 271

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T10:48:02.977160Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-06-27T09:48:27.652901Z digest=sha256:069e98d3ad94930ea56d50cf94bfdc82501e9c3577cef0f50e21331ef65b8421

Observation 89eab255-b671-42bc-82e4-20375aa5c10b · inbound

Visko Orbis 1.0: A Live Model for Real-Time Interactive Long Video Generation cites this paper.

Visko Orbis 1.0: A Live Model for Real-Time Interactive Long Video Generation MiraData: A Large-Scale Video Dataset with Long Durations and Structured Captions

Reference 35

Resolution
unresolved
no resolver link, observed 2026-07-30T23:47:57.782044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T23:47:57.782044Z digest=sha256:41f99e02f125ae333c03d13e82eb138c88cfcf5353653080a6c572b5911ca54c