Pith. sign in

Paper Citation Record · LEDGER

Content-Rich AIGC Video Quality Assessment via Intricate Text Alignment and Motion-Aware Consistency

As of 9 August 2026, this Paper Citation Record lists 19 of 19 outbound references and 6 inbound Pith citation observations for arXiv:2502.04076.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.04076 v1

Coverage vector

measured 19 of 19 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-08T23:40:21.599948Z

measured 25 of 25 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:51:34.599824Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T16:39:57.276136Z

Reference resolution

19 of 19 outbound references displayed

  • verified exact0
  • verified fuzzy2
  • unresolved16
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c2a8e74a-7113-4f37-adbf-25e50092632d · outbound

This paper cites GPT-4 Technical Report.

Content-Rich AIGC Video Quality Assessment via Intricate Text Alignment and Motion-Aware Consistency GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-08T23:40:21.526498Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T23:40:21.526498Z digest=sha256:29d1d14dbe44c7ac4aad1d30c902fa21b4295f9d7339784c8ea6e956dd24ebb3

Observation be168710-22c4-40bb-809c-a770244efa32 · outbound

This paper cites VideoCrafter1: Open Diffusion Models for High-Quality Video Generation.

Content-Rich AIGC Video Quality Assessment via Intricate Text Alignment and Motion-Aware Consistency VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-08T23:40:21.538548Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T23:40:21.538548Z digest=sha256:fd8434bdc59509964a1344a5d1ed1b8d12722bc44e0dab0b7e532e3a8082ef32

Observation e8eaeb3f-7de1-4d52-9224-d1bab1684c4d · outbound

This paper cites Huang, Z., He, Y ., Yu, J., Zhang, F., Si, C., Jiang, Y ., Zhang, Y ., Wu, T., Jin, Q., Chanpaisit, N., Wang, Y ., Chen, X., Wang, L., Lin, D., Qiao, Y ., and Liu, Z.

Content-Rich AIGC Video Quality Assessment via Intricate Text Alignment and Motion-Aware Consistency Huang, Z., He, Y ., Yu, J., Zhang, F., Si, C., Jiang, Y ., Zhang, Y ., Wu, T., Jin, Q., Chanpaisit, N., Wang, Y ., Chen, X., Wang, L., Lin, D., Qiao, Y ., and Liu, Z

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-08T23:40:21.542777Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T23:40:21.542777Z digest=sha256:4826d5575bf8029bb8fa8f93c98ad6c2091b3fd00712964adda394710ff0e503

Observation 249154e6-5495-43f8-8529-cc6f2f248189 · outbound

This paper cites Subjective-Aligned Dataset and Metric for Text-to-Video Quality Assessment.

Content-Rich AIGC Video Quality Assessment via Intricate Text Alignment and Motion-Aware Consistency Subjective-Aligned Dataset and Metric for Text-to-Video Quality Assessment

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-08T23:40:21.554288Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T23:40:21.554288Z digest=sha256:3164db11a6edd221ec8a9a9b4e5703e3b606bbda431c69430197bb3b0362db12

Observation 0cd71f66-bd3c-40e0-9bbb-0cbd466d8d3a · outbound

This paper cites 10948109.

Content-Rich AIGC Video Quality Assessment via Intricate Text Alignment and Motion-Aware Consistency 10948109

Reference 9

Resolution
malformed identifier
no resolver link, observed 2026-08-08T23:40:21.557939Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T23:40:21.557939Z digest=sha256:0d0dc355c64cebcb0eb08f254002964e8ed0dd98b027fb974fa66905d69a8afb

Observation 016ad169-e761-4ded-a56e-f41c735fe44b · outbound

This paper cites FETV: A Benchmark for Fine-Grained Evaluation of Open-Domain Text-to-Video Generation.

Content-Rich AIGC Video Quality Assessment via Intricate Text Alignment and Motion-Aware Consistency FETV: A Benchmark for Fine-Grained Evaluation of Open-Domain Text-to-Video Generation

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-08T23:40:21.561264Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T23:40:21.561264Z digest=sha256:d30c711f52848ff92347a4d3a1b47eac97a7be30b98447e8759f9e2cf40f3d73

Observation 910d71c2-770b-42ff-bb88-a2651194a62a · outbound

This paper cites Exploring AIGC Video Quality: A Focus on Visual Harmony, Video-Text Consistency and Domain Distribution Gap.

Content-Rich AIGC Video Quality Assessment via Intricate Text Alignment and Motion-Aware Consistency Exploring AIGC Video Quality: A Focus on Visual Harmony, Video-Text Consistency and Domain Distribution Gap

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-08T23:40:21.565103Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T23:40:21.565103Z digest=sha256:796bafb96bc6076d29366ab59c42880783d7b34fa074beebfe2c045278cea97f

Observation abfdd187-1a43-4bf7-bce2-6ed78f0b9b96 · outbound

This paper cites T2V-CompBench: A Comprehensive Benchmark for Compositional Text-to-video Generation.

Content-Rich AIGC Video Quality Assessment via Intricate Text Alignment and Motion-Aware Consistency T2V-CompBench: A Comprehensive Benchmark for Compositional Text-to-video Generation

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-08T23:40:21.573733Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T23:40:21.573733Z digest=sha256:f992add7f766dc4656b1f79288029dadf61fef78ae88f2342ebf8b94fb7b3e3c

Observation 38913820-e6c5-4ca7-9bd0-cb3b5104e22d · outbound

This paper cites and Deng, J.

Content-Rich AIGC Video Quality Assessment via Intricate Text Alignment and Motion-Aware Consistency and Deng, J

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T23:40:21.782648Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T23:40:21.578004Z digest=sha256:2fc3a025714f986714aa67bd98091e0321db8f25cba63a3f37fdc975c77d0edc

Observation 82f6c8db-c73e-429c-8c24-c231878f1b91 · outbound

This paper cites Towards Accurate Generative Models of Video: A New Metric & Challenges.

Content-Rich AIGC Video Quality Assessment via Intricate Text Alignment and Motion-Aware Consistency Towards Accurate Generative Models of Video: A New Metric & Challenges

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-08T23:40:21.581583Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T23:40:21.581583Z digest=sha256:909b906090d5bd07fcbb5a41ac702cb26229d62a6fc94f0ef21c47d6645d0e9d

Observation a8a32d52-2eda-4422-9962-52baae7d614f · outbound

This paper cites CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer.

Content-Rich AIGC Video Quality Assessment via Intricate Text Alignment and Motion-Aware Consistency CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-08T23:40:21.592874Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T23:40:21.592874Z digest=sha256:b79c0ef1a77a59eb68bcbbcbc01d1bbe5679c481641078fde8f82664c38b8ea8

Observation 8cf20e18-74b9-46a5-9924-b4439dadbd04 · outbound

This paper cites The Dawn of Video Generation: Preliminary Explorations with SORA-like Models.

Content-Rich AIGC Video Quality Assessment via Intricate Text Alignment and Motion-Aware Consistency The Dawn of Video Generation: Preliminary Explorations with SORA-like Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-08T23:40:21.596386Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T23:40:21.596386Z digest=sha256:62fbbdfe9ed41c93b3d6c0e931b201868872a9ceb714d12090bcc8afd0e376f7

Observation 6cb0ee77-9ec0-4717-8e75-9b406f2599c7 · outbound

This paper cites Long-CLIP: Unlocking the Long-Text Capability of CLIP.

Content-Rich AIGC Video Quality Assessment via Intricate Text Alignment and Motion-Aware Consistency Long-CLIP: Unlocking the Long-Text Capability of CLIP

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-08T23:40:21.599948Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T23:40:21.599948Z digest=sha256:ae883db250cfad482d1eb474c82f2e9779f9cd53c01fbadb2ceb568764ee2848

Observation 19de4fc0-0c87-4f11-8229-4f01711a7157 · outbound

This paper cites The Kinetics Human Action Video Dataset.

Content-Rich AIGC Video Quality Assessment via Intricate Text Alignment and Motion-Aware Consistency The Kinetics Human Action Video Dataset

Reference 2000

Resolution
unresolved
no resolver link, observed 2026-08-08T23:40:21.550393Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T23:40:21.550393Z digest=sha256:f1a627ac6c26208f14783b472fa1f3b75d6e554df083cb5595d7e85a2a846c8e

Observation 9e5bb564-1f1a-457c-b9a4-e4bb6ab4ed61 · outbound

This paper cites ModelScope Text-to-Video Technical Report.

Content-Rich AIGC Video Quality Assessment via Intricate Text Alignment and Motion-Aware Consistency ModelScope Text-to-Video Technical Report

Reference 2018

Resolution
unresolved
no resolver link, observed 2026-08-08T23:40:21.585302Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T23:40:21.585302Z digest=sha256:8f95909624c9d5abef1f4081965ab630b7033b8e6753c98282300e130c26b23c

Observation e702f49b-d926-4fa1-aa05-02d04b8a92bc · outbound

This paper cites Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets.

Content-Rich AIGC Video Quality Assessment via Intricate Text Alignment and Motion-Aware Consistency Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-08T23:40:21.534636Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T23:40:21.534636Z digest=sha256:6010163e79b28f609a2981dbbf6751f372babd1f7acb6b3f16d13f9a9edd4f4f

Observation a58988eb-ec31-4e82-ad1e-c789bae44716 · outbound

This paper cites Q-Align: Teaching LMMs for Visual Scoring via Discrete Text-Defined Levels.

Content-Rich AIGC Video Quality Assessment via Intricate Text Alignment and Motion-Aware Consistency Q-Align: Teaching LMMs for Visual Scoring via Discrete Text-Defined Levels

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-08T23:40:21.589149Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T23:40:21.589149Z digest=sha256:255781591e219f6ba23142e957594841b547b4e632dcc63a0e881b170176b889

Observation 30a4a48c-c2a8-4d19-9885-9a379cbcdd0a · outbound

This paper cites De- tecting deep-fake videos from appearance and behavior.

Content-Rich AIGC Video Quality Assessment via Intricate Text Alignment and Motion-Aware Consistency De- tecting deep-fake videos from appearance and behavior

Reference 2023

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T23:40:21.794447Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-08T23:40:21.530843Z digest=sha256:389a0af84cb979e006d5db46cc988606a4639d238d2ae7b39e6ca269a8c40d53

Observation 336e50a1-1a3d-46db-be39-bd18ea234aeb · outbound

This paper cites HunyuanVideo: A Systematic Framework For Large Video Generative Models.

Content-Rich AIGC Video Quality Assessment via Intricate Text Alignment and Motion-Aware Consistency HunyuanVideo: A Systematic Framework For Large Video Generative Models

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-08T23:40:21.546574Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T23:40:21.546574Z digest=sha256:f07335d14b96ded55a7c91f0f0914564454c590478c47e1aa55dcd0c9d1c352f

Pith citing papers

Observation 5fc68969-9ec8-418d-a931-28abb29414e6 · inbound

More Thinking, Less Seeing? Assessing Amplified Hallucination in Multimodal Reasoning Models cites this paper.

More Thinking, Less Seeing? Assessing Amplified Hallucination in Multimodal Reasoning Models Content-Rich AIGC Video Quality Assessment via Intricate Text Alignment and Motion-Aware Consistency

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:34.599824Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:51:34.599824Z digest=sha256:782cef0ef6afb3386f0f75bfee7c57eb66717e23327abf9b5d4f2b20e5499080

Observation 1a583657-983a-4c61-81a8-6814acaa4142 · inbound

MASS: Motion-Aware Spatial-Temporal Grounding for Physics Reasoning and Comprehension in Vision-Language Models cites this paper.

MASS: Motion-Aware Spatial-Temporal Grounding for Physics Reasoning and Comprehension in Vision-Language Models Content-Rich AIGC Video Quality Assessment via Intricate Text Alignment and Motion-Aware Consistency

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-05-17T05:59:08.657072Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-17T05:55:11.495430Z digest=sha256:ef7356fa0dfb212037734535aab374ea55bfc792a518bbed9cd0475959b48242

Observation 4872d589-ca48-4feb-b1bd-f9cd0f64585b · inbound

Comparison Drives Preference: Reference-Aware Modeling for AI-Generated Video Quality Assessment cites this paper.

Comparison Drives Preference: Reference-Aware Modeling for AI-Generated Video Quality Assessment Content-Rich AIGC Video Quality Assessment via Intricate Text Alignment and Motion-Aware Consistency

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-05-10T06:31:30.717061Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T06:27:43.395596Z digest=sha256:553dd29ededf9fd314ee818ca8ae4e09ef0cfc1d0e6eb098bbce810bc0dae5be

Observation fa1bf4dd-37e6-41c9-81ea-bc7a5c8936cd · inbound

PhyGround: Benchmarking Physical Reasoning in Generative World Models cites this paper.

PhyGround: Benchmarking Physical Reasoning in Generative World Models Content-Rich AIGC Video Quality Assessment via Intricate Text Alignment and Motion-Aware Consistency

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-05-12T05:21:28.553668Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-12T05:17:30.010064Z digest=sha256:288048c787fb07e847c7f34cc31d3d4d9966be0b828c804d685ddf539fd9ffdf

Observation 9c4b37f6-b502-40e9-b2e8-cfd7b8aae281 · inbound

TailorMind: Towards Preference-Aligned Multimodal Content Generation cites this paper.

TailorMind: Towards Preference-Aligned Multimodal Content Generation Content-Rich AIGC Video Quality Assessment via Intricate Text Alignment and Motion-Aware Consistency

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T10:59:46.786194Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-26T08:14:07.112511Z digest=sha256:c9905c7c99a4dbfb1611d1d300029da9e863974637dc95027f07b14b40bb5406

Observation 8a7c33c9-1046-466c-9033-12cb3fbd7f91 · inbound

Navigating User Behavior toward Personalized Multimodal Generation cites this paper.

Navigating User Behavior toward Personalized Multimodal Generation Content-Rich AIGC Video Quality Assessment via Intricate Text Alignment and Motion-Aware Consistency

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T16:39:57.277627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-26T00:26:36.300071Z digest=sha256:f95a82a6cd5b6a0437273f6e099327ce0db32ba24fcc0956f0647a37e290acba