Pith. sign in

Paper Citation Record · LEDGER

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection

As of 4 August 2026, this Paper Citation Record lists 59 of 59 outbound references and 1 inbound Pith citation observation for arXiv:2512.00336.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2512.00336 v3

Coverage vector

measured 59 of 59 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-17T03:48:26.807495Z

measured 60 of 60 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-01T02:09:50.577740Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

59 of 59 outbound references displayed

  • verified exact20
  • verified fuzzy38
  • unresolved1
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 385ea3f2-394b-4b69-942f-8a78c5f7a5a4 · outbound

This paper cites write newline.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection write newline

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T03:48:58.682602Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:9efd86d2b50755bcee49441f0abd49bbd1af70503a25f7ba1c061b3f6f6355c5

Observation be46b26c-0823-4a42-b745-796776d3d43e · outbound

This paper cites Cosmos World Foundation Model Platform for Physical AI.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection Cosmos World Foundation Model Platform for Physical AI

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-05-17T03:48:58.485045Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:4efafd40185dea8503090c94d1e3cd90534a2a66f04f738f19c92a4ef44d6c86

Observation bfb3ff66-0d22-46e4-99bc-01b76015ee21 · outbound

This paper cites Pixverse: Ai video generation platform.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection Pixverse: Ai video generation platform

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T03:48:58.748751Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:492696a63648419efd93db28b683f6117d05e0645f21fae8be6392b9abc52fcf

Observation 06801ed0-8e3d-42a7-a8cc-015c18139320 · outbound

This paper cites Wan (tongyi wanxiang): Ai video generation platform.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection Wan (tongyi wanxiang): Ai video generation platform

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T03:48:58.689467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:165b482dcbc2542d7a678dfb38276c0d751cd17f23f2cee4507f0f2dcf0fdf1f

Observation ba1ad8be-6252-440e-b739-2ae62638602e · outbound

This paper cites Ai-generated video detection via spatial-temporal anomaly learning.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection Ai-generated video detection via spatial-temporal anomaly learning

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T03:48:58.686307Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:8552380bfd20a7dd78db50b317d257655da88fa5b53478943ff5332ff68329a1

Observation fb233442-266a-4ef8-9531-6a1c30fc9d98 · outbound

This paper cites R., Christodorescu, M., Datta, A., Feizi, S., et al.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection R., Christodorescu, M., Datta, A., Feizi, S., et al

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T03:48:58.760771Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:f3501d420ad1572e0cf107482b842096ea57e9d447a140e3f648612901007d6e

Observation f9debaa9-ecb0-43ec-a26f-ce1b534b8a3a · outbound

This paper cites Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-05-17T03:48:58.498245Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:e4a5eb2f2b87d09ece2e122d1aaedf0290cdba1124b0c63a1a438042c1ab1611

Observation 9e456cce-2586-4711-beb9-c400eef1481a · outbound

This paper cites and Dolan, W.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection and Dolan, W

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T03:48:58.743124Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:965d52f1699343c5c906fd0cc98c64ae0cefdecd5d8ebfd9c82499534a1f16c9

Observation 5c2d038c-5b25-4881-9559-0cdb87202e9c · outbound

This paper cites Diffute: Universal text editing diffusion model.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection Diffute: Universal text editing diffusion model

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T03:48:58.718885Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:fa0b8da86a6ddefaa28bacb4e35939e292c839af11bddff39ea277be44e420d6

Observation 5f9987b1-f6a8-4b79-884d-a082faac9e62 · outbound

This paper cites DeMamba: AI-Generated Video Detection on Million-Scale GenVideo Benchmark.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection DeMamba: AI-Generated Video Detection on Million-Scale GenVideo Benchmark

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-17T03:48:58.490202Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:368cb818eee6fe3bc361454d3d393051b3bb238be097357f5b85a6c3df86bf59

Observation a283af88-5db6-41a9-b208-470110ee1944 · outbound

This paper cites Videocrafter2: Overcoming data limitations for high-quality video diffusion models.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection Videocrafter2: Overcoming data limitations for high-quality video diffusion models

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T03:48:58.669045Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:10f087384079336f978614aece484b8f406a783f81dea6cce3833b4c945e609e

Observation b2177093-0791-4c79-9f92-221d3a59a587 · outbound

This paper cites HuMo: Human-Centric Video Generation via Collaborative Multi-Modal Conditioning.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection HuMo: Human-Centric Video Generation via Collaborative Multi-Modal Conditioning

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-17T03:48:58.476215Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:341d8cb5774a64ef0ffa7f29914b7b883e197a2bf40097cd7e24b5f33edcf08d

Observation aff5c502-e45a-46f6-8ca0-34a01eb77b0b · outbound

This paper cites TalkVid: A Large-Scale Diversified Dataset for Audio-Driven Talking Head Synthesis.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection TalkVid: A Large-Scale Diversified Dataset for Audio-Driven Talking Head Synthesis

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-17T03:48:58.439681Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:fa07c23ad1587fa28943e9055367aa775402aa3154264ccd2fc9a51428c2ffd7

Observation 6b415433-1011-4cb0-91c8-4ba9b555151f · outbound

This paper cites GenWorld: Towards Detecting AI-generated Real-world Simulation Videos.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection GenWorld: Towards Detecting AI-generated Real-world Simulation Videos

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-17T03:48:58.508275Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:45dbf9102ad954de73346718c82e399630beac43a58a9ef919719750e434e8b7

Observation f94fbd65-2ad7-41ee-9536-01bb94e8e7e5 · outbound

This paper cites K., Ishii, M., Hayakawa, A., Shibuya, T., Schwing, A., and Mitsufuji, Y.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection K., Ishii, M., Hayakawa, A., Shibuya, T., Schwing, A., and Mitsufuji, Y

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T03:48:58.729168Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:6ca8602defcb5cd3ea829f799f1316f5ac2d88a9d17da2fef393f5d9e7c8a302

Observation 9d832905-4a76-41b1-86c7-d6383882e2dd · outbound

This paper cites Deepseek ai chat platform.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection Deepseek ai chat platform

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T03:48:58.698974Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:439b50a6cd411fc6afdc65520a16713dbf4e88e5e2453453a431903ec2da5b68

Observation cbfb7b39-9a4d-4015-a19b-eec7d01089aa · outbound

This paper cites The DeepFake Detection Challenge (DFDC) Dataset.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection The DeepFake Detection Challenge (DFDC) Dataset

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-05-17T03:48:58.449861Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:129004a06a4affe404783fd5884d28de3ac62ed64716a257c3bacce8a26a5dc0

Observation d17fb73d-1e4e-4406-85f2-089e4660f65f · outbound

This paper cites Privacy and security concerns in generative AI : A comprehensive survey.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection Privacy and security concerns in generative AI : A comprehensive survey

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T03:48:58.710603Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:4c7f99634b1edec5279def95d8b712190e8756019be98f715ed24dd1f6fa71cd

Observation 274e194d-73f5-4ddb-8c66-12fdeaf162c7 · outbound

This paper cites Haiper AI : AI video generation platform.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection Haiper AI : AI video generation platform

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T03:48:58.757722Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:1b2cd94d7b3a3ad345d33cc62e5ffd771e8576089b6e142b03a03528a0126222

Observation fadd186e-dc12-4b94-aa28-d98c3e458f1a · outbound

This paper cites Denoising diffusion probabilistic models.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection Denoising diffusion probabilistic models

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T03:48:58.763921Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:8bd678571f408427048ed3def0ad0cf2c7061bc9d37587d0f4168f42bbef82b5

Observation c2ebd8ed-355e-449c-bf08-2d4498977f32 · outbound

This paper cites VBench : Comprehensive benchmark suite for video generative models.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection VBench : Comprehensive benchmark suite for video generative models

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T03:48:58.776552Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:dd0cb94525f7e986e526ce16ff08c1f11539dc61a762240a8a89bd186cc86fd1

Observation a4bbcb05-04ed-4b7a-bde3-d240acb30e22 · outbound

This paper cites Speech-forensics: Towards comprehensive synthetic speech dataset establishment and analysis.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection Speech-forensics: Towards comprehensive synthetic speech dataset establishment and analysis

Reference 22

Resolution
verified exact
doi, observed 2026-05-17T03:48:58.222809Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:72b7d0a0e726de309773a3bfce2339dc82893743a7ce389f316ff19b48224b69

Observation efb550e6-1126-44ec-857e-4afc930f274e · outbound

This paper cites Jiying ai: Video generation platform.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection Jiying ai: Video generation platform

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T03:48:58.773431Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:e9a47e5beea61a3453138dbcfe88b2bd5c27c43a306e650eb86a8099fc83a891

Observation 91bd9d9e-d018-4d1d-91b8-fa56f35463cb · outbound

This paper cites Spoofceleb: Speech deepfake detection and SASV in the wild.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection Spoofceleb: Speech deepfake detection and SASV in the wild

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T03:48:58.770380Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:40be759efbea34b8d42d9fed3f44226f4a3fb77b3216de8d5c9f7743b18e1802

Observation 8a75ae79-6716-4ef4-8cad-b3bbd923a34c · outbound

This paper cites FakeAVCeleb: A Novel Audio-Video Multimodal Deepfake Dataset.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection FakeAVCeleb: A Novel Audio-Video Multimodal Deepfake Dataset

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-17T03:48:58.454672Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:ab5b606689238562e1e46b3feaa833d97480d313d85fdd890527727c6c1f6b7f

Observation 91a5a70c-1c20-4e7d-b3c2-d66fe2279a9c · outbound

This paper cites Kling ai: Advanced video generation platform.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection Kling ai: Advanced video generation platform

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T03:48:58.678930Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:e28b33a070f92c81c45830d6560de482a348393b536134023f3a942dd698ba49

Observation cb96bf76-a2d5-4192-948b-fa6beb9e0373 · outbound

This paper cites Kling ai: Advanced video generation platform.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection Kling ai: Advanced video generation platform

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T03:48:58.740978Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:f8a276c30c666cc5c7285c323a4acd27541dd3b440c1f4dce538c8fef4f1b9e1

Observation 0854a0a2-e6ca-4e51-b458-91b31e781a56 · outbound

This paper cites VideoPoet: A Large Language Model for Zero-Shot Video Generation.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection VideoPoet: A Large Language Model for Zero-Shot Video Generation

Reference 28

Resolution
verified exact
local_arxiv, observed 2026-05-17T03:48:58.444755Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:fce4731218b81f1f702539f7f23ee5fb9e33e7d89de1e38549635f6f0dc89de0

Observation 27f0263c-bbde-44e6-b260-79b9f16520ca · outbound

This paper cites Snapfusion: Text-to-image diffusion model on mobile devices within two seconds.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection Snapfusion: Text-to-image diffusion model on mobile devices within two seconds

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T03:48:58.704696Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:cc51898f84410a589b704577d413c7f096c6db3f04933f5db15ebf0fbe91afc1

Observation c03309e9-ed0d-475e-8b3b-19fe2c112312 · outbound

This paper cites Stiv: Scalable text and image conditioned video generation.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection Stiv: Scalable text and image conditioned video generation

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T03:48:58.701958Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:59d1443eda31245b286003509b103d6120c596a87a5880e707160125e019a6cb

Observation 511b0684-6fdc-4b10-8744-d6a284b04226 · outbound

This paper cites Evalcrafter: Benchmarking and evaluating large video generation models.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection Evalcrafter: Benchmarking and evaluating large video generation models

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T03:48:58.734095Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:5335bcd79289fe2082c3210a5f031e70504780f63c199064d349bf097ff33b5e

Observation 06dfe0d5-1661-4759-8e92-498e5a3668d7 · outbound

This paper cites Uve: Are mllms uni- fied evaluators for ai-generated videos?.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection Uve: Are mllms uni- fied evaluators for ai-generated videos?

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-17T03:48:58.434459Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:4c973f707fa2ad6c38f821ff4802fb88d07207a14fe64fafd6971f223c6a4c56

Observation 373a8b8c-e438-4957-9e27-4af6ee801ae4 · outbound

This paper cites DeCoF : Generated video detection via frame consistency: The first benchmark dataset.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection DeCoF : Generated video detection via frame consistency: The first benchmark dataset

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T03:48:58.731700Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:fb4ad2570c52e592158aad3a4b02735cc6fc2633c3ad0ea9d7028b2f81c02000

Observation 01b51975-af4e-457d-a69d-7620548f457b · outbound

This paper cites Moonvalley: Ai video generation platform.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection Moonvalley: Ai video generation platform

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T03:48:58.672210Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:9bf3a1bd399eecf65be9f816433173b69d1774b50485bb67aaa4b4c42fb1312e

Observation de74674a-4403-4655-8f83-3db6bf4b7de4 · outbound

This paper cites OpenVid-1M: A Large-Scale High-Quality Dataset for Text-to-video Generation.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection OpenVid-1M: A Large-Scale High-Quality Dataset for Text-to-video Generation

Reference 35

Resolution
verified exact
local_arxiv, observed 2026-05-17T03:48:58.480719Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:f5440b2bdcd1d8b7ca80ca63503e3465f26e2ccf3563385422c76b387ad3e0d5

Observation 9163091b-7f8a-46ac-9cb4-336cddaaa5c5 · outbound

This paper cites Genvidbench: A challenging benchmark for detecting ai- generated video.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection Genvidbench: A challenging benchmark for detecting ai- generated video

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-05-17T03:48:58.518103Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:ad49cfd537343ae5f76e36c2981fe7481fe65b7c4528a02f511c0f63359915e3

Observation 49270c42-31b2-4728-82f8-aceaf78d57e4 · outbound

This paper cites an unresolved cited work.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-05-17T03:48:58.663956Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:3b2e51a8f30cc5814c2f09407389c25ea57afba28e359c19bf9f233b1caf6e1d

Observation 1ea498d5-e9f5-47c1-898f-f05538fb07cd · outbound

This paper cites Sora: Creating video from text.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection Sora: Creating video from text

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T03:48:58.745699Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:a4cb1dfdf6bd25dba5386339a4a71204019b4dbef59897a1df6779d5f1a23d0e

Observation 1c1cd15f-6ef9-420e-bcde-4cc2aa13ad2e · outbound

This paper cites Sora: Creating video from text.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection Sora: Creating video from text

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T03:48:58.767119Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:648b84c93abb2ac0298c656cb8250e6d3b5ff37ea953117ed2dce090b9e04b4d

Observation 4f57e052-4437-44d8-9169-69ba2369f3cf · outbound

This paper cites Pika: Ai video generation platform.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection Pika: Ai video generation platform

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T03:48:58.707582Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:63c245e1d0abf7a9443e6a82032e6fe932b53933457875f601ef26d0e0b0857c

Observation 253e9e22-cc2c-42fa-b68f-6643fe72e3f9 · outbound

This paper cites Runwayml: Ai video generation platform.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection Runwayml: Ai video generation platform

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T03:48:58.726841Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:92aeaf1f83c91945f089dfea18ac3c84b90ccab7d12e844c558a9282456fef9f

Observation 50efffe9-2666-45d3-a15b-0ec52b44af34 · outbound

This paper cites HunyuanVideo-Foley: Multimodal Diffusion with Representation Alignment for High-Fidelity Foley Audio Generation.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection HunyuanVideo-Foley: Multimodal Diffusion with Representation Alignment for High-Fidelity Foley Audio Generation

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-17T03:48:58.471609Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:3d5dbb1c1543e57813de4c78e5137b37540a52ad0e46112e9d72527100b6c494

Observation dea81e3e-a84a-4672-8059-f183746fb1a6 · outbound

This paper cites On learning multi-modal forgery representation for diffusion generated video detection.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection On learning multi-modal forgery representation for diffusion generated video detection

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T03:48:58.724267Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:d89536ed340b6b02f54de6fab36f3d45f2de8861e77f067e42e9c5531a85347a

Observation 01912514-e6ff-413a-932d-456ba25297a3 · outbound

This paper cites AudioX: A Unified Framework for Anything-to-Audio Generation.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection AudioX: A Unified Framework for Anything-to-Audio Generation

Reference 44

Resolution
verified exact
local_arxiv, observed 2026-05-17T03:48:58.464790Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:7fa7bef19ce746908378614d7a7d2dd63af18dc536755aa34d73f55253f46a90

Observation 9b6a0c37-3867-4399-b1ea-385bc8941c5d · outbound

This paper cites Noisee ai: Ai music video generation platform.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection Noisee ai: Ai music video generation platform

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T03:48:58.721738Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:762b85d818dbabef5fc04a136f387556ef6e9012ba6cc3e9a2625fbcbd44443a

Observation ce423e2f-c8e2-4e01-88b7-b1ff7a154525 · outbound

This paper cites Veo3 ai: Advanced video generation platform.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection Veo3 ai: Advanced video generation platform

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T03:48:58.675390Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:8562b3e46ba71124b4bcbbc0004969db3f0ad6c617894852ddb657ce4f0166a9

Observation 1e84c879-3455-45d6-95db-6520b75a922e · outbound

This paper cites Vidu: Ultra-realistic video generation model.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection Vidu: Ultra-realistic video generation model

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T03:48:58.779587Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:670df3f6d02ad2c9fcc72d8db14f07546ea00b52ab0c8649cb9bff41784f6879

Observation e12fc069-f194-412e-a5ad-7646c49b25ec · outbound

This paper cites Viva video ai: Ai video generation platform.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection Viva video ai: Ai video generation platform

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T03:48:58.737094Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:137f217c25b8365859f5a4d23e748efa83c38e07f0a38c6b71524c33b7278257

Observation 8490cdfd-32df-451d-8825-b2c8c225a84b · outbound

This paper cites Emu3: Next-Token Prediction is All You Need.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection Emu3: Next-Token Prediction is All You Need

Reference 49

Resolution
verified exact
local_arxiv, observed 2026-05-17T03:48:58.429877Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:fe60472a9c0ef8bfd8c49749ef234fb8d1fb9016bcea5bb2469457f20105f658

Observation aee86242-127c-4615-a89d-8336e02dae48 · outbound

This paper cites InternVid: A Large-scale Video-Text Dataset for Multimodal Understanding and Generation.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection InternVid: A Large-scale Video-Text Dataset for Multimodal Understanding and Generation

Reference 50

Resolution
verified exact
local_arxiv, observed 2026-05-17T03:48:58.512894Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:2c46ebe7b40a8908d1aa8c9f360bd8a49baa873f49c360e1173e25dcfae87879

Observation bfb542af-99c8-4afa-ae7d-4972022864fd · outbound

This paper cites Lavie: High-quality video generation with cascaded latent diffusion models.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection Lavie: High-quality video generation with cascaded latent diffusion models

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T03:48:58.751676Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:c2f85708b3ae04717ff5ff9ea8f1e39a44dd7034b1177b632be448e3df692c77

Observation 99ec872a-2f0c-4b73-aa44-141f08650ff9 · outbound

This paper cites Art-v: Auto-regressive text-to-video generation with diffusion models.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection Art-v: Auto-regressive text-to-video generation with diffusion models

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T03:48:58.695929Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:02c6cd16d60fca0d2f1441369f175a4a482d16a9f98b805271ca020a020fc779

Observation 40aa8e99-3fd4-4cdf-93de-2182c11273bb · outbound

This paper cites UGC-VideoCaptioner : An omni ugc video detail caption model and new benchmarks.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection UGC-VideoCaptioner : An omni ugc video detail caption model and new benchmarks

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-05-17T03:48:58.503202Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:c4e81de883c0a14a82cda54cf3812bacf281742adfc199e5d1940a70ac4e16a2

Observation c245160a-4d9f-4606-b9ad-5db345a98370 · outbound

This paper cites MSR-VTT : A large video description dataset for bridging video and language.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection MSR-VTT : A large video description dataset for bridging video and language

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T03:48:58.716384Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:43d077501b05d015c86d2e6cebbd81a889ef63fe89891c688eb0295363aee9bd

Observation c7a10a39-8adf-49be-ae4f-c6ea8c30eaa1 · outbound

This paper cites Adding conditional control to text-to-image diffusion models.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection Adding conditional control to text-to-image diffusion models

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T03:48:58.692452Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:d34ed1e84242135c17e2bdde154f334f5e14c3e7855946d3fdd020431ad4a1ed

Observation 270cdb8c-cd4a-449e-b745-efa3bffbc071 · outbound

This paper cites FoleyCrafter: Bring Silent Videos to Life with Lifelike and Synchronized Sounds.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection FoleyCrafter: Bring Silent Videos to Life with Lifelike and Synchronized Sounds

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-05-17T03:48:58.459651Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:c2244cb5f8a459c96d06b9bf1b228266c2d59b00503df9bdcd779eba220d8d24

Observation 7f26f15f-5020-4141-ab57-831089f2ca56 · outbound

This paper cites Occworld: Learning a 3d occupancy world model for autonomous driving.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection Occworld: Learning a 3d occupancy world model for autonomous driving

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T03:48:58.713578Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:8561e867022c6ce5b71651de7300742b547463391e69c712a434187de5dfa90f

Observation 65f086f1-3f17-4604-8f29-e6c908eb2083 · outbound

This paper cites Open-Sora: Democratizing Efficient Video Production for All.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection Open-Sora: Democratizing Efficient Video Production for All

Reference 58

Resolution
verified exact
local_arxiv, observed 2026-05-17T03:48:58.522778Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:d295b22dc4620bb2adb751cfd9c816dc2ef2d26dcb8046c32e8517b6ac2cb84d

Observation 91170183-a9de-4f74-aa59-24d9f9cb0632 · outbound

This paper cites Harmonyset: A comprehensive dataset for understanding video-music semantic alignment and temporal synchronization.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection Harmonyset: A comprehensive dataset for understanding video-music semantic alignment and temporal synchronization

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T03:48:58.754656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:94ff550cd3076272313055b72503d162fa255cdd1964387eb08d8ac8e3af6a84

Pith citing papers

Observation 8a8c9257-e712-4914-bca4-bfb95e7a82f4 · inbound

Less is More: Modality-Decoupling for General AIGC Audio-Video Detection cites this paper.

Less is More: Modality-Decoupling for General AIGC Audio-Video Detection MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-01T02:09:50.577740Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:09:50.577740Z digest=sha256:9fcac53ce8ed87d7a988cd21c25de1d0e8949aecc632e50180a745caa3122608