Pith. sign in

Paper Citation Record · LEDGER

Youku-mPLUG: A 10 Million Large-scale Chinese Video-Language Dataset for Pre-training and Benchmarks

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2306.04362.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2306.04362 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T21:12:23.223434Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-13T17:18:01.355607Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 55a46a9f-563a-445b-ac35-67455fb18c71 · inbound

Survey on AI-Generated Media Detection: From Non-MLLM to MLLM cites this paper.

Survey on AI-Generated Media Detection: From Non-MLLM to MLLM Youku-mPLUG: A 10 Million Large-scale Chinese Video-Language Dataset for Pre-training and Benchmarks

Reference 188

Resolution
unresolved
no resolver link, observed 2026-08-08T21:12:23.223434Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T21:12:23.223434Z digest=sha256:bb864464cc19b22a6855e7da19d6770908c1bb8572578ad8d6c99bc718b3a10c

Observation 50918665-cf97-4c2f-9f92-a8529f81cb66 · inbound

OmniEval: A Benchmark for Evaluating Omni-modal Models with Visual, Auditory, and Textual Inputs cites this paper.

OmniEval: A Benchmark for Evaluating Omni-modal Models with Visual, Auditory, and Textual Inputs Youku-mPLUG: A 10 Million Large-scale Chinese Video-Language Dataset for Pre-training and Benchmarks

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-06T22:41:11.661900Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:41:11.661900Z digest=sha256:711b2e766012e6e1009a244d2f06850fef1c2afa8b345a7de09ceaa1d0880815

Observation db9c54da-b929-4ebf-97e6-73ec42ea623b · inbound

Physics-Driven Spatiotemporal Modeling for AI-Generated Video Detection cites this paper.

Physics-Driven Spatiotemporal Modeling for AI-Generated Video Detection Youku-mPLUG: A 10 Million Large-scale Chinese Video-Language Dataset for Pre-training and Benchmarks

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-04T10:54:29.611312Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T10:54:29.611312Z digest=sha256:ef209ab9ec8d927efa07aaa9cdb4fed0a188689bf2d69a693525c36059ccb940

Observation 01c26340-723a-4cad-affb-4265dcf14459 · inbound

ATSS: Detecting AI-Generated Videos via Anomalous Temporal Self-Similarity cites this paper.

ATSS: Detecting AI-Generated Videos via Anomalous Temporal Self-Similarity Youku-mPLUG: A 10 Million Large-scale Chinese Video-Language Dataset for Pre-training and Benchmarks

Reference 57

Resolution
verified exact
orphan_title_repair, observed 2026-05-13T17:18:57.108015Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-13T17:15:14.018979Z digest=sha256:4017b4eefb46e0b182f99abac291a79ce7efaebfb9b85df7d5aea7de1f201c60

Observation dc350964-d9d8-4713-97ff-e3027c07d8f4 · inbound

CMTA: Leveraging Cross-Modal Temporal Artifacts for Generalizable AI-Generated Video Detection cites this paper.

CMTA: Leveraging Cross-Modal Temporal Artifacts for Generalizable AI-Generated Video Detection Youku-mPLUG: A 10 Million Large-scale Chinese Video-Language Dataset for Pre-training and Benchmarks

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-05-11T15:16:10.034191Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-09T20:22:00.777817Z digest=sha256:e57d9edc509eabb824abf29b05e45a4c70777f4404f8358024ba494625e0d514

Observation 6b2f5801-d3f0-4c5b-8f9a-f7097b7e2736 · inbound

From Static Analysis to Audience Dissemination: A Training-Free Multimodal Controversy Detection Multi-Agent Framework cites this paper.

From Static Analysis to Audience Dissemination: A Training-Free Multimodal Controversy Detection Multi-Agent Framework Youku-mPLUG: A 10 Million Large-scale Chinese Video-Language Dataset for Pre-training and Benchmarks

Reference 25

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T15:16:06.762039Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-09T20:26:09.112189Z digest=sha256:f2ef51b8bc9d41bc730e10bd7682fb4ac26a9c7b4f527dfc65c1d40ca4966db3