Pith. sign in

Paper Citation Record · LEDGER

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection

As of 10 August 2026, this Paper Citation Record lists 59 of 59 outbound references and 1 inbound Pith citation observation for arXiv:2512.00336.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2512.00336 v3

Coverage vector

measured 59 of 59 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-17T03:48:26.807495Z

measured 60 of 60 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-01T02:09:50.577740Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

59 of 59 outbound references displayed

  • verified exact20
  • verified fuzzy38
  • unresolved1
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 385ea3f2-394b-4b69-942f-8a78c5f7a5a4 · outbound

This paper cites write newline.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection write newline

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T03:48:58.682602Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:153b1a09be01152f54a82a0cc7da7fafb06d388a4b8d2aa4701d0be8530dcd7a

Observation be46b26c-0823-4a42-b745-796776d3d43e · outbound

This paper cites Cosmos World Foundation Model Platform for Physical AI.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection Cosmos World Foundation Model Platform for Physical AI

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-05-17T03:48:58.485045Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:e9faec738a4c82336fcc07c880df3ae1bdfe0d52936cf19de04661ad0a4fd84c

Observation bfb3ff66-0d22-46e4-99bc-01b76015ee21 · outbound

This paper cites Pixverse: Ai video generation platform.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection Pixverse: Ai video generation platform

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T03:48:58.748751Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:0e9ac569382f863f12b588a960eab5e19e4d160cf88ddde3eb3ddd408aeda888

Observation 06801ed0-8e3d-42a7-a8cc-015c18139320 · outbound

This paper cites Wan (tongyi wanxiang): Ai video generation platform.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection Wan (tongyi wanxiang): Ai video generation platform

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T03:48:58.689467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:4cbe6adb579b67fa05f9de55f4296b69b7a39aef5c3516a2dad79352b77b03a1

Observation ba1ad8be-6252-440e-b739-2ae62638602e · outbound

This paper cites Ai-generated video detection via spatial-temporal anomaly learning.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection Ai-generated video detection via spatial-temporal anomaly learning

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T03:48:58.686307Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:7ba2db30083883e8ab45988076d3a119f74efdc874159fb780358efd04f4a57a

Observation fb233442-266a-4ef8-9531-6a1c30fc9d98 · outbound

This paper cites R., Christodorescu, M., Datta, A., Feizi, S., et al.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection R., Christodorescu, M., Datta, A., Feizi, S., et al

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T03:48:58.760771Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:87aa197722e7a64e2fc60d0c32326d7f85fe6d976c95566bc673cb4788510596

Observation f9debaa9-ecb0-43ec-a26f-ce1b534b8a3a · outbound

This paper cites Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-05-17T03:48:58.498245Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:1b5e702be50ba272f898de8bf0fd80a3529a0710fde13aac3590f5b99838b9ad

Observation 9e456cce-2586-4711-beb9-c400eef1481a · outbound

This paper cites and Dolan, W.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection and Dolan, W

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T03:48:58.743124Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:641da3a4b53c78d20ad2f1976937c99d6aad0179b6e244e4ba2b52ee7be9202a

Observation 5c2d038c-5b25-4881-9559-0cdb87202e9c · outbound

This paper cites Diffute: Universal text editing diffusion model.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection Diffute: Universal text editing diffusion model

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T03:48:58.718885Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:a17c0d6a046b0a588e4223b2ad33f72733886c5ecfe907c8b21220eb9a317678

Observation 5f9987b1-f6a8-4b79-884d-a082faac9e62 · outbound

This paper cites DeMamba: AI-Generated Video Detection on Million-Scale GenVideo Benchmark.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection DeMamba: AI-Generated Video Detection on Million-Scale GenVideo Benchmark

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-17T03:48:58.490202Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:83873b5b95be086b85246acfd6579394e1358881ab9d29ceef989eca4c5c51f8

Observation a283af88-5db6-41a9-b208-470110ee1944 · outbound

This paper cites Videocrafter2: Overcoming data limitations for high-quality video diffusion models.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection Videocrafter2: Overcoming data limitations for high-quality video diffusion models

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T03:48:58.669045Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:dcad008c18ffa2ea0fcaa45697d7536bac7c95372a0cf6ab4b616b2e3bffe4d6

Observation b2177093-0791-4c79-9f92-221d3a59a587 · outbound

This paper cites HuMo: Human-Centric Video Generation via Collaborative Multi-Modal Conditioning.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection HuMo: Human-Centric Video Generation via Collaborative Multi-Modal Conditioning

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-17T03:48:58.476215Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:029a3ae1c0393f1aeb59094f21536ba8fe4c9f28b022c0f46631cad457514b36

Observation aff5c502-e45a-46f6-8ca0-34a01eb77b0b · outbound

This paper cites TalkVid: A Large-Scale Diversified Dataset for Audio-Driven Talking Head Synthesis.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection TalkVid: A Large-Scale Diversified Dataset for Audio-Driven Talking Head Synthesis

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-17T03:48:58.439681Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:e871662aa31988e7e367d80ed90b8d808e2ee9e5d5debb8743b366e2959e616a

Observation 6b415433-1011-4cb0-91c8-4ba9b555151f · outbound

This paper cites GenWorld: Towards Detecting AI-generated Real-world Simulation Videos.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection GenWorld: Towards Detecting AI-generated Real-world Simulation Videos

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-17T03:48:58.508275Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:8667e33c0e4af999ac3182ee911f478ac1a077335824b4ee94ab4a62b3f835b5

Observation f94fbd65-2ad7-41ee-9536-01bb94e8e7e5 · outbound

This paper cites K., Ishii, M., Hayakawa, A., Shibuya, T., Schwing, A., and Mitsufuji, Y.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection K., Ishii, M., Hayakawa, A., Shibuya, T., Schwing, A., and Mitsufuji, Y

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T03:48:58.729168Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:fe3e320ad9c092ca157c98f8b041068a9961dd65f3c43270787bc657ee0b3723

Observation 9d832905-4a76-41b1-86c7-d6383882e2dd · outbound

This paper cites Deepseek ai chat platform.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection Deepseek ai chat platform

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T03:48:58.698974Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:545fa8fdb08886c72e14c4326f4b3e54bd90df703a953e79f1eba3426dc461ee

Observation cbfb7b39-9a4d-4015-a19b-eec7d01089aa · outbound

This paper cites The DeepFake Detection Challenge (DFDC) Dataset.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection The DeepFake Detection Challenge (DFDC) Dataset

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-05-17T03:48:58.449861Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:6f89b4ec984e13ab14552b262be159f17ea8627b2c1cfae1c87df4cf463d252d

Observation d17fb73d-1e4e-4406-85f2-089e4660f65f · outbound

This paper cites Privacy and security concerns in generative AI : A comprehensive survey.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection Privacy and security concerns in generative AI : A comprehensive survey

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T03:48:58.710603Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:f49fb23eb5c00fa3e14c3191d83e66458ba09bd7af761fb4bb82f7b75f81c30b

Observation 274e194d-73f5-4ddb-8c66-12fdeaf162c7 · outbound

This paper cites Haiper AI : AI video generation platform.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection Haiper AI : AI video generation platform

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T03:48:58.757722Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:8f3a8b0065b65112f86dc47a2835ce5c1d4a0592074bb2453a965b9a6eec48b6

Observation fadd186e-dc12-4b94-aa28-d98c3e458f1a · outbound

This paper cites Denoising diffusion probabilistic models.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection Denoising diffusion probabilistic models

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T03:48:58.763921Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:e326dc5676d1ac67132974f03615255ece65582bbf0b5ec9f21f7aa8a7f61246

Observation c2ebd8ed-355e-449c-bf08-2d4498977f32 · outbound

This paper cites VBench : Comprehensive benchmark suite for video generative models.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection VBench : Comprehensive benchmark suite for video generative models

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T03:48:58.776552Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:67fbf5ec54d86bd3d983d6ee2585dec62ca41b3d25c4825be927a660a5d0fbdc

Observation a4bbcb05-04ed-4b7a-bde3-d240acb30e22 · outbound

This paper cites Speech-forensics: Towards comprehensive synthetic speech dataset establishment and analysis.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection Speech-forensics: Towards comprehensive synthetic speech dataset establishment and analysis

Reference 22

Resolution
verified exact
doi, observed 2026-05-17T03:48:58.222809Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:c7919cfb75d2fbbed8589a0e156adb8c44912faa54b1ae2aadc2dcb63fb6d6f8

Observation efb550e6-1126-44ec-857e-4afc930f274e · outbound

This paper cites Jiying ai: Video generation platform.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection Jiying ai: Video generation platform

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T03:48:58.773431Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:d6da4b29e1d417a5929546d134089af6a5daf109499b479c70f0882e531a82aa

Observation 91bd9d9e-d018-4d1d-91b8-fa56f35463cb · outbound

This paper cites Spoofceleb: Speech deepfake detection and SASV in the wild.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection Spoofceleb: Speech deepfake detection and SASV in the wild

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T03:48:58.770380Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:2cb2c4a90e94309edebf176def61766aaf3995031c5a11dd23862fcd252fd30e

Observation 8a75ae79-6716-4ef4-8cad-b3bbd923a34c · outbound

This paper cites FakeAVCeleb: A Novel Audio-Video Multimodal Deepfake Dataset.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection FakeAVCeleb: A Novel Audio-Video Multimodal Deepfake Dataset

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-17T03:48:58.454672Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:3c9ad94eb39de95ae277b015711b29691a2c986853b3ca4c0c5e3af841a8fc8d

Observation 91a5a70c-1c20-4e7d-b3c2-d66fe2279a9c · outbound

This paper cites Kling ai: Advanced video generation platform.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection Kling ai: Advanced video generation platform

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T03:48:58.678930Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:354512d5893e26cc8a93224d888a7b2f974f2ac98d2ba3a64f7a74e4838784ba

Observation cb96bf76-a2d5-4192-948b-fa6beb9e0373 · outbound

This paper cites Kling ai: Advanced video generation platform.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection Kling ai: Advanced video generation platform

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T03:48:58.740978Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:267f18393a5a3caaeec76cf76ec8b42b960a3b8c6e14f91777a0be154131ad76

Observation 0854a0a2-e6ca-4e51-b458-91b31e781a56 · outbound

This paper cites VideoPoet: A Large Language Model for Zero-Shot Video Generation.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection VideoPoet: A Large Language Model for Zero-Shot Video Generation

Reference 28

Resolution
verified exact
local_arxiv, observed 2026-05-17T03:48:58.444755Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:3e5e840cda145b06c3d41186c72d174cbc6a24960c8bcb579a0b8cf8b5f8bbcd

Observation 27f0263c-bbde-44e6-b260-79b9f16520ca · outbound

This paper cites Snapfusion: Text-to-image diffusion model on mobile devices within two seconds.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection Snapfusion: Text-to-image diffusion model on mobile devices within two seconds

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T03:48:58.704696Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:7e608580a1cbe2a07fb3d7bdf20d5ad2cae43ebd98440882a99989d4694fea7e

Observation c03309e9-ed0d-475e-8b3b-19fe2c112312 · outbound

This paper cites Stiv: Scalable text and image conditioned video generation.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection Stiv: Scalable text and image conditioned video generation

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T03:48:58.701958Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:97ad97030c982ace77cbeda3c8d0de8392148f38086437a297485abca0097b7b

Observation 511b0684-6fdc-4b10-8744-d6a284b04226 · outbound

This paper cites Evalcrafter: Benchmarking and evaluating large video generation models.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection Evalcrafter: Benchmarking and evaluating large video generation models

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T03:48:58.734095Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:ded89f276a2b754f4106f902ac1f71431013e9344e4022e6e35107600f1e5abd

Observation 06dfe0d5-1661-4759-8e92-498e5a3668d7 · outbound

This paper cites Uve: Are mllms uni- fied evaluators for ai-generated videos?.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection Uve: Are mllms uni- fied evaluators for ai-generated videos?

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-17T03:48:58.434459Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:6ffa6abe9722d31a0c1071aa7e236a7d29cf8bde9fa3f1a52c8d29d2945c6bcb

Observation 373a8b8c-e438-4957-9e27-4af6ee801ae4 · outbound

This paper cites DeCoF : Generated video detection via frame consistency: The first benchmark dataset.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection DeCoF : Generated video detection via frame consistency: The first benchmark dataset

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T03:48:58.731700Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:40454333e10224bfd622308e7a6ef06a63ced1187d50a9ba593b4afef2176e91

Observation 01b51975-af4e-457d-a69d-7620548f457b · outbound

This paper cites Moonvalley: Ai video generation platform.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection Moonvalley: Ai video generation platform

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T03:48:58.672210Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:84db272152f3e89845b352c4ea21bf51e82c53674ff5e5fb6d6f475737945783

Observation de74674a-4403-4655-8f83-3db6bf4b7de4 · outbound

This paper cites OpenVid-1M: A Large-Scale High-Quality Dataset for Text-to-video Generation.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection OpenVid-1M: A Large-Scale High-Quality Dataset for Text-to-video Generation

Reference 35

Resolution
verified exact
local_arxiv, observed 2026-05-17T03:48:58.480719Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:afeb9a5ccd5555f78082b8e5d3fa979fb36915a2ed211c7c3ea62b02f65e5122

Observation 9163091b-7f8a-46ac-9cb4-336cddaaa5c5 · outbound

This paper cites Genvidbench: A challenging benchmark for detecting ai- generated video.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection Genvidbench: A challenging benchmark for detecting ai- generated video

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-05-17T03:48:58.518103Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:2297e7fd17b2c2c3de37348f6f48f32757c9572be2a2ac7af6d5cd5bde3f7f0f

Observation 49270c42-31b2-4728-82f8-aceaf78d57e4 · outbound

This paper cites an unresolved cited work.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-05-17T03:48:58.663956Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:f5500487e7a3f1f2895b24238c96f16ba0d4311b41bde8a2df88ec9e7f8a008f

Observation 1ea498d5-e9f5-47c1-898f-f05538fb07cd · outbound

This paper cites Sora: Creating video from text.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection Sora: Creating video from text

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T03:48:58.745699Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:d09be82cc8a4b71305a5968ef2ccbce62c32d0bd4a1a9e3900a3c0aa37dc3faa

Observation 1c1cd15f-6ef9-420e-bcde-4cc2aa13ad2e · outbound

This paper cites Sora: Creating video from text.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection Sora: Creating video from text

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T03:48:58.767119Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:7b39e15e27366cec2d83c1b6fc6d21b3483389e01c5603cb5357bfc94e130b77

Observation 4f57e052-4437-44d8-9169-69ba2369f3cf · outbound

This paper cites Pika: Ai video generation platform.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection Pika: Ai video generation platform

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T03:48:58.707582Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:5766443f0216c2a81005448d42f70d1ea885747c56d1f007958237e19989f44f

Observation 253e9e22-cc2c-42fa-b68f-6643fe72e3f9 · outbound

This paper cites Runwayml: Ai video generation platform.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection Runwayml: Ai video generation platform

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T03:48:58.726841Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:b788b21732144ed58a59ab7617c57a4604d2dbc97993b6eb03f307b1ceceb907

Observation 50efffe9-2666-45d3-a15b-0ec52b44af34 · outbound

This paper cites HunyuanVideo-Foley: Multimodal Diffusion with Representation Alignment for High-Fidelity Foley Audio Generation.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection HunyuanVideo-Foley: Multimodal Diffusion with Representation Alignment for High-Fidelity Foley Audio Generation

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-17T03:48:58.471609Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:71c803191a6ae9ba7931188d45e492fc260b4e8e807ded7dd2f80090b4830b60

Observation dea81e3e-a84a-4672-8059-f183746fb1a6 · outbound

This paper cites On learning multi-modal forgery representation for diffusion generated video detection.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection On learning multi-modal forgery representation for diffusion generated video detection

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T03:48:58.724267Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:5b464a0ca48230838aecf275a1745a8acbcf3a68783a08d42c523662694a866b

Observation 01912514-e6ff-413a-932d-456ba25297a3 · outbound

This paper cites AudioX: A Unified Framework for Anything-to-Audio Generation.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection AudioX: A Unified Framework for Anything-to-Audio Generation

Reference 44

Resolution
verified exact
local_arxiv, observed 2026-05-17T03:48:58.464790Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:766a5e720d42b81ee7cf6abd6c3da526834e6b71850c7edb52ac6c87b0d5d085

Observation 9b6a0c37-3867-4399-b1ea-385bc8941c5d · outbound

This paper cites Noisee ai: Ai music video generation platform.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection Noisee ai: Ai music video generation platform

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T03:48:58.721738Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:dbd5fe4271e50e56c30579dc5ee666c49d6c85fec2b6c47d73881a90a7c7a768

Observation ce423e2f-c8e2-4e01-88b7-b1ff7a154525 · outbound

This paper cites Veo3 ai: Advanced video generation platform.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection Veo3 ai: Advanced video generation platform

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T03:48:58.675390Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:ec9927c9c42c24eb35bb06ed976d8bc37bdccc83415c07b460b7c6324dcc22dd

Observation 1e84c879-3455-45d6-95db-6520b75a922e · outbound

This paper cites Vidu: Ultra-realistic video generation model.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection Vidu: Ultra-realistic video generation model

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T03:48:58.779587Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:60d0ef669da80949668b7d012c7c2ee03e3c2f105c04cfaddcca4b5ec198036c

Observation e12fc069-f194-412e-a5ad-7646c49b25ec · outbound

This paper cites Viva video ai: Ai video generation platform.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection Viva video ai: Ai video generation platform

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T03:48:58.737094Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:9bb94d66ca2113db57667b5bdc43479ffa780d26c127b4567cae0dc374d90107

Observation 8490cdfd-32df-451d-8825-b2c8c225a84b · outbound

This paper cites Emu3: Next-Token Prediction is All You Need.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection Emu3: Next-Token Prediction is All You Need

Reference 49

Resolution
verified exact
local_arxiv, observed 2026-05-17T03:48:58.429877Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:bc70fb7a59fee1f2ffef801c5ae2beffe01fbbe5c18f8991c26f0688164816af

Observation aee86242-127c-4615-a89d-8336e02dae48 · outbound

This paper cites InternVid: A Large-scale Video-Text Dataset for Multimodal Understanding and Generation.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection InternVid: A Large-scale Video-Text Dataset for Multimodal Understanding and Generation

Reference 50

Resolution
verified exact
local_arxiv, observed 2026-05-17T03:48:58.512894Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:530634eab0c342ac4fa52396b81523f021171828cf347ae2746f4280c0cbf16d

Observation bfb542af-99c8-4afa-ae7d-4972022864fd · outbound

This paper cites Lavie: High-quality video generation with cascaded latent diffusion models.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection Lavie: High-quality video generation with cascaded latent diffusion models

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T03:48:58.751676Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:03f9d6fe6978aa0a2f6fa7ace2cb0b7cd7927a054bcf26b59f2609241058a771

Observation 99ec872a-2f0c-4b73-aa44-141f08650ff9 · outbound

This paper cites Art-v: Auto-regressive text-to-video generation with diffusion models.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection Art-v: Auto-regressive text-to-video generation with diffusion models

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T03:48:58.695929Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:9f5370fe393cc99f2ba0337e9955cb6dd5b9e8883c297d353d47cf85c49df512

Observation 40aa8e99-3fd4-4cdf-93de-2182c11273bb · outbound

This paper cites UGC-VideoCaptioner : An omni ugc video detail caption model and new benchmarks.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection UGC-VideoCaptioner : An omni ugc video detail caption model and new benchmarks

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-05-17T03:48:58.503202Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:9313a1524dcd0cc589f456403ad8f621355b051d7c29f461c57c30f2423b45f6

Observation c245160a-4d9f-4606-b9ad-5db345a98370 · outbound

This paper cites MSR-VTT : A large video description dataset for bridging video and language.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection MSR-VTT : A large video description dataset for bridging video and language

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T03:48:58.716384Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:638c4f3c4de24f57e658adf13487a28becc3648ad019e3f7172791dca77693b6

Observation c7a10a39-8adf-49be-ae4f-c6ea8c30eaa1 · outbound

This paper cites Adding conditional control to text-to-image diffusion models.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection Adding conditional control to text-to-image diffusion models

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T03:48:58.692452Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:a3ff6ad06c2ab238133a3f3b270d92f71a1c3eeea2f5c8118b994916cb655d37

Observation 270cdb8c-cd4a-449e-b745-efa3bffbc071 · outbound

This paper cites FoleyCrafter: Bring Silent Videos to Life with Lifelike and Synchronized Sounds.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection FoleyCrafter: Bring Silent Videos to Life with Lifelike and Synchronized Sounds

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-05-17T03:48:58.459651Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:03d11962d77f251fd60f8e45fb222a2420a5758fc4bd129215f08ce6ca5fe772

Observation 7f26f15f-5020-4141-ab57-831089f2ca56 · outbound

This paper cites Occworld: Learning a 3d occupancy world model for autonomous driving.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection Occworld: Learning a 3d occupancy world model for autonomous driving

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T03:48:58.713578Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:270240437f57ef144820f4d57747d181b41444d4eea1130bb13d2046499caa9b

Observation 65f086f1-3f17-4604-8f29-e6c908eb2083 · outbound

This paper cites Open-Sora: Democratizing Efficient Video Production for All.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection Open-Sora: Democratizing Efficient Video Production for All

Reference 58

Resolution
verified exact
local_arxiv, observed 2026-05-17T03:48:58.522778Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:04eb6146082cab63d09ab89826cdcf3d32fd6140aa12ef52805e1b29fecd34bd

Observation 91170183-a9de-4f74-aa59-24d9f9cb0632 · outbound

This paper cites Harmonyset: A comprehensive dataset for understanding video-music semantic alignment and temporal synchronization.

MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection Harmonyset: A comprehensive dataset for understanding video-music semantic alignment and temporal synchronization

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-05-17T03:48:58.754656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-17T03:48:26.807495Z digest=sha256:4fa60d04cddad68f51370cf6ca3b3f80300d3130da7ae90c4e9135ce081cf5a2

Pith citing papers

Observation 8a8c9257-e712-4914-bca4-bfb95e7a82f4 · inbound

Less is More: Modality-Decoupling for General AIGC Audio-Video Detection cites this paper.

Less is More: Modality-Decoupling for General AIGC Audio-Video Detection MVAD: A Benchmark Dataset for Multimodal AI-Generated Video-Audio Detection

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-01T02:09:50.577740Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:09:50.577740Z digest=sha256:8983a9daaa5e962026e82cadbe43214f48db91ba9d490bab78a061876a102c28