Pith. sign in

Paper Citation Record · LEDGER

VBench: Comprehensive Benchmark Suite for Video Generative Models

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 32 inbound Pith citation observations for arXiv:2311.17982.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2311.17982 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 32 of 32 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 32 of 32 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:59:30.690138Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T16:29:56.629606Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 283d2ce1-df0e-4109-a5c1-ad967052b041 · inbound

CameraCtrl: Enabling Camera Control for Text-to-Video Generation cites this paper.

CameraCtrl: Enabling Camera Control for Text-to-Video Generation VBench: Comprehensive Benchmark Suite for Video Generative Models

Reference 123

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T02:06:23.651555Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-13T02:06:23.410241Z digest=sha256:55f2b67b17b0ab1ab82f6714fd92dbc96e92328e0481820167594f3eeb1e96cc

Observation b7350057-e307-4923-8ef1-279ead804b89 · inbound

VideoPhy: Evaluating Physical Commonsense for Video Generation cites this paper.

VideoPhy: Evaluating Physical Commonsense for Video Generation VBench: Comprehensive Benchmark Suite for Video Generative Models

Reference 40

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T11:34:37.754990Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-20T11:34:37.599691Z digest=sha256:3b7b740b9c7c7991051372ec281d52ea9df9f9c6fe72f8530f6cb60db37902a8

Observation 65e3c255-602b-4982-bb4b-57471c45dd2c · inbound

OpenS2V-Nexus: A Detailed Benchmark and Million-Scale Dataset for Subject-to-Video Generation cites this paper.

OpenS2V-Nexus: A Detailed Benchmark and Million-Scale Dataset for Subject-to-Video Generation VBench: Comprehensive Benchmark Suite for Video Generative Models

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T13:59:30.690138Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:59:30.690138Z digest=sha256:a32421089a74c022c695e178387a079c2b60de1b8f5c12a0f46c450c4a4d6831

Observation f320bd9e-5f10-4597-96fe-2cd2ae9cb08e · inbound

PAROAttention: Pattern-Aware ReOrdering for Efficient Sparse and Quantized Attention in Visual Generation Models cites this paper.

PAROAttention: Pattern-Aware ReOrdering for Efficient Sparse and Quantized Attention in Visual Generation Models VBench: Comprehensive Benchmark Suite for Video Generative Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T23:51:05.341299Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:51:05.341299Z digest=sha256:b81754a582c1d3eedcd9b95835d6e9641efa55f3357421b2dc61a14381446a45

Observation 975c9967-78f9-4993-b917-94f9c112bf94 · inbound

FreeLong++: Training-Free Long Video Generation via Multi-band SpectralFusion cites this paper.

FreeLong++: Training-Free Long Video Generation via Multi-band SpectralFusion VBench: Comprehensive Benchmark Suite for Video Generative Models

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-06T21:28:22.113060Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:28:22.113060Z digest=sha256:ec671c65d46917ef3adcb5d60a6371bda0dc49a275fdd1db15a379f0a54a63da

Observation b26e608d-5d8e-415d-9b8f-a33390aec125 · inbound

AIGVE-MACS: Unified Multi-Aspect Commenting and Scoring Model for AI-Generated Video Evaluation cites this paper.

AIGVE-MACS: Unified Multi-Aspect Commenting and Scoring Model for AI-Generated Video Evaluation VBench: Comprehensive Benchmark Suite for Video Generative Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T21:03:00.930314Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T21:03:00.930314Z digest=sha256:a690fff2e8995a03aacf8fbafc585dbf9d1af33602d71ebfdc5327ea7142159c

Observation 6d39ec04-f441-49c5-a6e8-d54c75feb9c7 · inbound

CineScale: Free Lunch in High-Resolution Cinematic Visual Generation cites this paper.

CineScale: Free Lunch in High-Resolution Cinematic Visual Generation VBench: Comprehensive Benchmark Suite for Video Generative Models

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-05T17:46:36.701746Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:46:36.701746Z digest=sha256:b1e7e5ad04dbac0320a332573ac592d8a3d23e6580a48f0f0be099420e1b2c61

Observation ba92620b-78f7-4168-89e8-e4ed17d4dff0 · inbound

Forecast then Calibrate: Feature Caching as ODE for Efficient Diffusion Transformers cites this paper.

Forecast then Calibrate: Feature Caching as ODE for Efficient Diffusion Transformers VBench: Comprehensive Benchmark Suite for Video Generative Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-05T17:31:13.039875Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:31:13.039875Z digest=sha256:11bfa3d103c144dfab1d3cb37a4a81049388b95e406730465fa49bc4e52c3d8f

Observation 24b92ba4-8eb7-4a3e-897e-9e3f3a89ca43 · inbound

From Sound to Sight: Towards AI-authored Music Videos cites this paper.

From Sound to Sight: Towards AI-authored Music Videos VBench: Comprehensive Benchmark Suite for Video Generative Models

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-05T18:23:29.510645Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T18:23:29.510645Z digest=sha256:55304530490cb4987e6b0d7fd84b7da84d6326bf77f302238b1ca9fa981b9be3

Observation 49a6debb-ce2a-46bd-b2f7-c62df7640491 · inbound

DisCa: Accelerating Video Diffusion Transformers with Distillation-Compatible Learnable Feature Caching cites this paper.

DisCa: Accelerating Video Diffusion Transformers with Distillation-Compatible Learnable Feature Caching VBench: Comprehensive Benchmark Suite for Video Generative Models

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-16T07:30:44.430379Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-16T07:28:08.873654Z digest=sha256:2a286884b263c7537144db60f30bc0341aa828119bb07b39354b01c56c45e039

Observation 69ca81e5-dea0-4174-a95b-4b7c3bd884e8 · inbound

VIGOR: VIdeo Geometry-Oriented Reward for Temporal Generative Alignment cites this paper.

VIGOR: VIdeo Geometry-Oriented Reward for Temporal Generative Alignment VBench: Comprehensive Benchmark Suite for Video Generative Models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-07-13T23:54:26.281624Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T23:54:26.281624Z digest=sha256:d834391fb253e82f91a21c19a3396c28ef9244920544b9e568d87820801640ed

Observation 46b8e4a2-eab5-46c8-a002-519834d0ef6d · inbound

VERTIGO: Visual Preference Optimization for Cinematic Camera Trajectory Generation cites this paper.

VERTIGO: Visual Preference Optimization for Cinematic Camera Trajectory Generation VBench: Comprehensive Benchmark Suite for Video Generative Models

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-13T21:18:17.206695Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T21:14:42.021240Z digest=sha256:9d3acb844c1b8469630b3dd9f278563efce3e79452fe06e5f4396817cedab8ed

Observation 1e54b2de-a5d0-4052-a6a5-e031f45e2e4d · inbound

Physics-Aware Video Instance Removal Benchmark cites this paper.

Physics-Aware Video Instance Removal Benchmark VBench: Comprehensive Benchmark Suite for Video Generative Models

Reference 24

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T22:55:52.029383Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T19:27:33.716719Z digest=sha256:83d78d40790e051f6ad079955c48042576f38f1e7d0f158664f42178ea6c7bf1

Observation 08a00a97-2c66-40ea-a8df-38ecee816c8c · inbound

Quantitative Video World Model Evaluation for Geometric-Consistency cites this paper.

Quantitative Video World Model Evaluation for Geometric-Consistency VBench: Comprehensive Benchmark Suite for Video Generative Models

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-15T03:14:53.150583Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-15T03:11:03.060052Z digest=sha256:4b843bdbc763923017605115712f29e26a28bac85aef16a37be11ef413ee7c97

Observation ec4eaf1e-fde4-4155-a19b-fd9f2f8a5595 · inbound

Dynamic Video Generation: Shaping Video Generation Across Time and Space cites this paper.

Dynamic Video Generation: Shaping Video Generation Across Time and Space VBench: Comprehensive Benchmark Suite for Video Generative Models

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-21T05:43:58.822618Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-21T05:42:41.925474Z digest=sha256:919155364a54411ee5e4b95d6ddbb152de048e9883cb70029a254e19bb0fa3ac

Observation 42021878-01df-428b-9091-0775df5a1778 · inbound

Video as Natural Augmentation: Towards Unified AI-Generated Image and Video Detection cites this paper.

Video as Natural Augmentation: Towards Unified AI-Generated Image and Video Detection VBench: Comprehensive Benchmark Suite for Video Generative Models

Reference 54

Resolution
metadata mismatch
arxiv_id, observed 2026-05-22T08:01:15.636806Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T08:00:21.829807Z digest=sha256:923b57cce9de2cc23a1ffe4643f29513642e301934d85ff58dfff1799c6bcc74

Observation 67b9ecea-7823-4ee3-b1f7-afd59ce0431b · inbound

One Sentence, One Drama: Personalized Short-Form Drama Generation via Multi-Agent Systems cites this paper.

One Sentence, One Drama: Personalized Short-Form Drama Generation via Multi-Agent Systems VBench: Comprehensive Benchmark Suite for Video Generative Models

Reference 23

Resolution
metadata mismatch
arxiv_id, observed 2026-05-22T06:54:42.324684Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T06:51:58.805848Z digest=sha256:a24bbd14f9806cbfcdcb13b5529a3e872f83e59134fbe7ab8b309f454151381a

Observation a37f25cf-e205-4755-b6fe-250976085359 · inbound

CRONOS: Benchmarking Counterfactual Physical Consistency in Video Models cites this paper.

CRONOS: Benchmarking Counterfactual Physical Consistency in Video Models VBench: Comprehensive Benchmark Suite for Video Generative Models

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-25T04:40:23.327008Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-25T04:39:22.400458Z digest=sha256:11d3ad27b93a4b2eca09dd32252ac27eeb5fb5343c262d429624e86d19f90fa6

Observation a8d6cb9d-7812-49ad-b878-856beee0be97 · inbound

Quantized Keys Steal Attention: Bias Correction for KV-Cache Compression in Video Diffusion cites this paper.

Quantized Keys Steal Attention: Bias Correction for KV-Cache Compression in Video Diffusion VBench: Comprehensive Benchmark Suite for Video Generative Models

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-06-29T22:54:01.104736Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-29T22:50:05.037051Z digest=sha256:95705e6114f92935611e1fe75269da7c6ae348d58b5d21c602fa0b059b7dd659

Observation 8a40c1b3-a967-4bc2-a2ef-8c05fbe18c4b · inbound

MiraBench: Evaluating Action-Conditioned Reliability in Robotic World Models cites this paper.

MiraBench: Evaluating Action-Conditioned Reliability in Robotic World Models VBench: Comprehensive Benchmark Suite for Video Generative Models

Reference 21

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T07:53:13.817117Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-29T07:46:28.469913Z digest=sha256:c278922d322a2587c56b900c6f1c5edf3eeb73aa20c36942666462c0ce0f619a

Observation 3eb93132-6784-4cbf-82fd-c9a4fdb597d9 · inbound

Collaborative Few-Step Distillation and Low-Bit Quantization for Wan2.2 Dual-Expert Video Diffusion Models cites this paper.

Collaborative Few-Step Distillation and Low-Bit Quantization for Wan2.2 Dual-Expert Video Diffusion Models VBench: Comprehensive Benchmark Suite for Video Generative Models

Reference 9

Resolution
metadata mismatch
arxiv_id, observed 2026-06-28T19:02:34.323250Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T18:58:13.799615Z digest=sha256:625c4b752e29ebd34e7f5f9a6fdaf1f79d31409a2b824c766df4c4655572ac6e

Observation 479b067c-9e9e-43bb-a012-7db1d3602ff8 · inbound

VideoWeaver: Evaluating and Evolving Skills for Agentic Long Video Generation cites this paper.

VideoWeaver: Evaluating and Evolving Skills for Agentic Long Video Generation VBench: Comprehensive Benchmark Suite for Video Generative Models

Reference 47

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T20:57:23.167851Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-27T20:04:10.711238Z digest=sha256:6d3297d2a59b66458e281ba9a637ee0b09033b20b15e762b618129cb526d1d49

Observation fa4a83e0-bb1e-44a8-9f2d-e11521952691 · inbound

Geometry-Instructed Video Editing cites this paper.

Geometry-Instructed Video Editing VBench: Comprehensive Benchmark Suite for Video Generative Models

Reference 8

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T16:29:56.631281Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-26T00:43:03.861412Z digest=sha256:63f15b013081bb9f8275f1f6ad623f17f8ee138dfb523f48fe162c2c0fb308c3

Observation c0c26be5-7419-4e15-8acd-4ea7b13bf3d2 · inbound

HAT-4D: Lifting Monocular Video for 4D Multi-Object Interactions via Human-Agent Collaboration cites this paper.

HAT-4D: Lifting Monocular Video for 4D Multi-Object Interactions via Human-Agent Collaboration VBench: Comprehensive Benchmark Suite for Video Generative Models

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-07-01T17:05:50.609528Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-29T04:18:02.341742Z digest=sha256:5a99a32d6180620d8b1a0100ae938e9e63c270cdeefaa1f89fc563749594cf38

Observation 8fcf203d-ba52-440a-8681-59efbdfeb580 · inbound

HandsOnWorld: Unconstrained Egocentric Video Generation with Camera-Disentangled Hand Control cites this paper.

HandsOnWorld: Unconstrained Egocentric Video Generation with Camera-Disentangled Hand Control VBench: Comprehensive Benchmark Suite for Video Generative Models

Reference 24

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T15:58:37.451051Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-07-03T15:52:18.371702Z digest=sha256:a9e8f7327900027fcea6a84b7d6604d80b73f31fdd572fbde81795eb206a3347

Observation aa56e97a-3fbb-4207-a3be-6a820a498ea2 · inbound

Hallo4D: Multi-Modal Hallucination Mitigation for Consistent Spatio-Temporal Generation cites this paper.

Hallo4D: Multi-Modal Hallucination Mitigation for Consistent Spatio-Temporal Generation VBench: Comprehensive Benchmark Suite for Video Generative Models

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-02T06:23:30.601553Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T06:23:30.601553Z digest=sha256:b42eabb9afcf6c20bff5f2bbca49e50255990f539df553c95d4a6593caaba51d

Observation e24ef7ab-211a-4e86-ad7d-e272c3d2d538 · inbound

CODA: Algorithm-Hardware Co-design for Edge Video Diffusion via NMP-Enabled Compute-Cache Operator Disaggregation cites this paper.

CODA: Algorithm-Hardware Co-design for Edge Video Diffusion via NMP-Enabled Compute-Cache Operator Disaggregation VBench: Comprehensive Benchmark Suite for Video Generative Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-02T00:49:35.218788Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T00:49:35.218788Z digest=sha256:817edca09be04119ff7b6101c6f0ce17bf818c990e933c0d0544acafa5667a94

Observation 3b530f4f-58b1-4eb3-9c83-250294dab735 · inbound

PhysAgent: Reflective Agentic Physics Control for Physically Plausible Video Generation cites this paper.

PhysAgent: Reflective Agentic Physics Control for Physically Plausible Video Generation VBench: Comprehensive Benchmark Suite for Video Generative Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-01T22:27:52.162212Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T22:27:52.162212Z digest=sha256:13b3a1518df062a0e2182ebcdbf2892410dd2ea73c22d0887df016d9e0204f2b

Observation aa745742-9a59-4141-b743-e57a877b974c · inbound

Adaptive Identity Anchoring: Closed-Loop Keyframe Placement for Synthetic Paired Supervision in Video Face Swapping cites this paper.

Adaptive Identity Anchoring: Closed-Loop Keyframe Placement for Synthetic Paired Supervision in Video Face Swapping VBench: Comprehensive Benchmark Suite for Video Generative Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-01T07:31:28.790940Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T07:31:28.790940Z digest=sha256:0190c77b64d835fb48b027377d246405b9fda85475564d1023e7639837af4d11

Observation b031b815-db42-4831-bed9-68efd86f7cf6 · inbound

CachedSearch: Training-Free Cached Exploration for Test-Time Search in Video Diffusion cites this paper.

CachedSearch: Training-Free Cached Exploration for Test-Time Search in Video Diffusion VBench: Comprehensive Benchmark Suite for Video Generative Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-01T03:30:23.925493Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:30:23.925493Z digest=sha256:1ac9d34b51bce291b7b0850ba24c6a97b74035f68b01079a408e478c080571bc

Observation 0fc5390d-b41b-4fe7-95f1-3b7a61c5623b · inbound

Parallel Decoding Distillation for Fast Image and Video Generation cites this paper.

Parallel Decoding Distillation for Fast Image and Video Generation VBench: Comprehensive Benchmark Suite for Video Generative Models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-01T00:57:38.892549Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T00:57:38.892549Z digest=sha256:919d95cef71de23ef6ec1dac5a1a0e2fbe2739f10202f9241c8a933c6de5307b

Observation 6fb25780-0e1c-49e6-ae45-12a2e1e7c3de · inbound

Extended Field of View Analysis for VideoGAN-based Trajectory Generation cites this paper.

Extended Field of View Analysis for VideoGAN-based Trajectory Generation VBench: Comprehensive Benchmark Suite for Video Generative Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-04T09:55:32.510610Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:55:32.510610Z digest=sha256:30c6d0ebc702e697439cb690b8ca776abe439b967888169a6fe0ea2fc98c033a