Pith. sign in

Paper Citation Record · LEDGER

VideoScore: Building Automatic Metrics to Simulate Fine-grained Human Feedback for Video Generation

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 29 inbound Pith citation observations for arXiv:2406.15252.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2406.15252 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 29 of 29 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 29 of 29 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T10:40:44.156703Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T15:45:48.888277Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 42495389-73cf-457f-a1e7-20efcb1aeacd · inbound

Towards World Simulator: Crafting Physical Commonsense-Based Benchmark for Video Generation cites this paper.

Towards World Simulator: Crafting Physical Commonsense-Based Benchmark for Video Generation VideoScore: Building Automatic Metrics to Simulate Fine-grained Human Feedback for Video Generation

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-18T14:40:00.111640Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-18T14:39:59.870039Z digest=sha256:d566041880891b0a63ec49882942f5ac8d9df56045f33bccd631131e05e3b920

Observation 4c271c58-00d9-4d51-89d6-8c207850fcca · inbound

DOLLAR: Few-Step Video Generation via Distillation and Latent Reward Optimization cites this paper.

DOLLAR: Few-Step Video Generation via Distillation and Latent Reward Optimization VideoScore: Building Automatic Metrics to Simulate Fine-grained Human Feedback for Video Generation

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-23T07:02:41.973532Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-23T06:57:50.897865Z digest=sha256:00321d21b20f876e708d80bb03c555042e90049a27db175e32e21d9482e0a764

Observation adff459a-b84b-46dc-b336-189ff2d2d82a · inbound

Do generative video models understand physical principles? cites this paper.

Do generative video models understand physical principles? VideoScore: Building Automatic Metrics to Simulate Fine-grained Human Feedback for Video Generation

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-05-20T12:47:05.922143Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-20T12:47:05.825659Z digest=sha256:dfdb49e916f8770e28e705ce0f9645fa934b723390415f6956c030017e34f338

Observation 86f0bdce-4432-4328-9c2e-6ade9c058368 · inbound

Improving Video Generation with Human Feedback cites this paper.

Improving Video Generation with Human Feedback VideoScore: Building Automatic Metrics to Simulate Fine-grained Human Feedback for Video Generation

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-13T15:30:02.653134Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T15:30:02.578430Z digest=sha256:f73fef795c2211a5a4e71c1478d43d26b6cf964bc59376f0beadb64bd24d3db6

Observation 33bd3cf3-9ee6-432b-ad22-f8d3d3ab9209 · inbound

Unified Reward Model for Multimodal Understanding and Generation cites this paper.

Unified Reward Model for Multimodal Understanding and Generation VideoScore: Building Automatic Metrics to Simulate Fine-grained Human Feedback for Video Generation

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-14T00:44:30.808338Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-14T00:44:30.558048Z digest=sha256:3d9c370b4fb3c801eee44ad0a1493fe1c3cd8c56cfb4a101bc9ea14263af9caa

Observation 254274c1-bab8-4519-a71c-5e0dd76505c2 · inbound

DanceGRPO: Unleashing GRPO on Visual Generation cites this paper.

DanceGRPO: Unleashing GRPO on Visual Generation VideoScore: Building Automatic Metrics to Simulate Fine-grained Human Feedback for Video Generation

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-05-11T22:28:27.476623Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-11T22:28:24.929046Z digest=sha256:c5bbe70ee445d55fd29558945e94bbcd62d8ef47f11a2df03d7f0634b827bcdf

Observation 29cdfd3c-1b3f-48eb-a91c-d9b440d1da04 · inbound

Towards Holistic Visual Quality Assessment of AI-Generated Videos: A LLM-Based Multi-Dimensional Evaluation Model cites this paper.

Towards Holistic Visual Quality Assessment of AI-Generated Videos: A LLM-Based Multi-Dimensional Evaluation Model VideoScore: Building Automatic Metrics to Simulate Fine-grained Human Feedback for Video Generation

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T10:40:44.156703Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:40:44.156703Z digest=sha256:29c7276b90540d8336bbc9c8ff076f161952cc90a4aadcc08b3f336201dde097

Observation be29dfc3-72d3-426d-bcdf-c9f4e5149870 · inbound

GigaVideo-1: Advancing Video Generation via Automatic Feedback with 4 GPU-Hours Fine-Tuning cites this paper.

GigaVideo-1: Advancing Video Generation via Automatic Feedback with 4 GPU-Hours Fine-Tuning VideoScore: Building Automatic Metrics to Simulate Fine-grained Human Feedback for Video Generation

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T04:29:39.531704Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:29:39.531704Z digest=sha256:7b15372065711d209b8dd559a024a7278af47fb5cc0994843f39d52660d5db3d

Observation ccbba998-7073-4628-b767-729d3c09c679 · inbound

AlignHuman: Improving Motion and Fidelity via Timestep-Segment Preference Optimization for Audio-Driven Human Animation cites this paper.

AlignHuman: Improving Motion and Fidelity via Timestep-Segment Preference Optimization for Audio-Driven Human Animation VideoScore: Building Automatic Metrics to Simulate Fine-grained Human Feedback for Video Generation

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T04:55:28.726179Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:55:28.726179Z digest=sha256:f2626539c726a457581d0332b9e8675ca365dee7a0372c40955ce39617c3bb6e

Observation c454e54b-d1bd-49c1-852d-f1b4e5d86501 · inbound

Fake it till You Make it: Reward Modeling as Discriminative Prediction cites this paper.

Fake it till You Make it: Reward Modeling as Discriminative Prediction VideoScore: Building Automatic Metrics to Simulate Fine-grained Human Feedback for Video Generation

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T00:30:41.631422Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:30:41.631422Z digest=sha256:64afcec18f80abc55b04d9eede8fdbd3868e680c3748bafebf2a1b3e08578547

Observation 402c126e-f787-4653-afe9-1576a29eff35 · inbound

AIGVE-MACS: Unified Multi-Aspect Commenting and Scoring Model for AI-Generated Video Evaluation cites this paper.

AIGVE-MACS: Unified Multi-Aspect Commenting and Scoring Model for AI-Generated Video Evaluation VideoScore: Building Automatic Metrics to Simulate Fine-grained Human Feedback for Video Generation

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T21:03:00.640203Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T21:03:00.640203Z digest=sha256:3f5e4f4fbf41d7e4d9c70311a4145be645a1d40a249970a794c583eccb28d0e8

Observation a2f3e10d-edde-4bb5-9854-aa29c9c39f57 · inbound

MedGen: Unlocking Medical Video Generation by Scaling Granularly-annotated Medical Videos cites this paper.

MedGen: Unlocking Medical Video Generation by Scaling Granularly-annotated Medical Videos VideoScore: Building Automatic Metrics to Simulate Fine-grained Human Feedback for Video Generation

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T19:25:08.523862Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:25:08.523862Z digest=sha256:060c49d0a7cb16752b87f7c263b08dc501d240d26390589b5fa9ccbf3d306ca0

Observation 2c2f0c50-0fd6-480d-ad51-af2ed04f2177 · inbound

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation cites this paper.

ADIEE: Automatic Dataset Creation and Scorer for Instruction-Guided Image Editing Evaluation VideoScore: Building Automatic Metrics to Simulate Fine-grained Human Feedback for Video Generation

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:39.761303Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:39.761303Z digest=sha256:f40bd554cf5315c342b08eb36163a85a0dedad08b063a0c36283365052bedbba

Observation aae6831f-afa6-472b-b66f-2d8592347138 · inbound

"PhyWorldBench": A Comprehensive Evaluation of Physical Realism in Text-to-Video Models cites this paper.

"PhyWorldBench": A Comprehensive Evaluation of Physical Realism in Text-to-Video Models VideoScore: Building Automatic Metrics to Simulate Fine-grained Human Feedback for Video Generation

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T16:30:51.136313Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:30:51.136313Z digest=sha256:acef1cd30ec0d2f0ab0eaa358b35f95b4e62b56156fa537c3e28f895c3510141

Observation ee107e7e-ec9e-49aa-abba-d72598cc329b · inbound

Runtime Failure Hunting for Physics Engine Based Software Systems: How Far Can We Go? cites this paper.

Runtime Failure Hunting for Physics Engine Based Software Systems: How Far Can We Go? VideoScore: Building Automatic Metrics to Simulate Fine-grained Human Feedback for Video Generation

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T12:09:26.767123Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T12:09:26.767123Z digest=sha256:f0104098e2a5779c5193654019141efee664690d3c4d98f416dae929ab4bc9cb

Observation 22539301-87c3-4ef8-91cb-5d08fdaf0ed7 · inbound

Robust Single-Stage Fully Sparse 3D Object Detection via Detachable Latent Diffusion cites this paper.

Robust Single-Stage Fully Sparse 3D Object Detection via Detachable Latent Diffusion VideoScore: Building Automatic Metrics to Simulate Fine-grained Human Feedback for Video Generation

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T04:36:13.428878Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T04:36:13.428878Z digest=sha256:680a4e918e70537f6330ad9d329011729d0b1eb7fc8e683b18b696cb73c2b430

Observation 9daed3f1-6769-4f2b-af1e-6b23ecdb6f14 · inbound

FantasyTalking2: Timestep-Layer Adaptive Preference Optimization for Audio-Driven Portrait Animation cites this paper.

FantasyTalking2: Timestep-Layer Adaptive Preference Optimization for Audio-Driven Portrait Animation VideoScore: Building Automatic Metrics to Simulate Fine-grained Human Feedback for Video Generation

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-05T20:06:46.273086Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:06:46.273086Z digest=sha256:f1434a83fcbc077a7e145838979bf719c1f4e77f6cf6d9cc684e9a0537d8ce55

Observation 38b5cae5-8a68-411d-9d4e-a5f6033d5569 · inbound

GeneVA: A Dataset of Human Annotations for Generative Text to Video Artifacts cites this paper.

GeneVA: A Dataset of Human Annotations for Generative Text to Video Artifacts VideoScore: Building Automatic Metrics to Simulate Fine-grained Human Feedback for Video Generation

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-04T20:09:21.219231Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T20:09:21.219231Z digest=sha256:84c7c96593adcdea643fdab4518855d7cc2b9e252fa06ec240a279eca2f48b3e

Observation 1998247a-f52c-4469-a7e0-02fb9ebd2ff5 · inbound

RAPO++: Cross-Stage Prompt Optimization for Text-to-Video Generation via Data Alignment and Test-Time Scaling cites this paper.

RAPO++: Cross-Stage Prompt Optimization for Text-to-Video Generation via Data Alignment and Test-Time Scaling VideoScore: Building Automatic Metrics to Simulate Fine-grained Human Feedback for Video Generation

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-18T05:15:54.498099Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-18T05:13:42.934115Z digest=sha256:547c805162160031db81a3677cf1d14ec186892cc4eaff0625e7a2cc9c410626

Observation 54d9b552-4606-48de-8c64-f9e60dde1205 · inbound

MASS: Motion-Aware Spatial-Temporal Grounding for Physics Reasoning and Comprehension in Vision-Language Models cites this paper.

MASS: Motion-Aware Spatial-Temporal Grounding for Physics Reasoning and Comprehension in Vision-Language Models VideoScore: Building Automatic Metrics to Simulate Fine-grained Human Feedback for Video Generation

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-17T05:59:08.552890Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-17T05:55:11.495430Z digest=sha256:197d03d5f9b4a5a51b73f88774f0e028f47f4f4c7dbfd0a3034b6aa8356fbee6

Observation 2fb21064-138e-4a63-80c1-4d01e3b50093 · inbound

Seeing What Matters: Visual Preference Policy Optimization for Visual Generation cites this paper.

Seeing What Matters: Visual Preference Policy Optimization for Visual Generation VideoScore: Building Automatic Metrics to Simulate Fine-grained Human Feedback for Video Generation

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-21T18:44:18.974203Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-21T18:42:29.327151Z digest=sha256:c5c91f449fc06ed599437490e382acc7134ebc88270bfbdf669d47a4e32407ce

Observation 97a9ad42-829b-4166-85db-26d8bde655d4 · inbound

Zoom-IQA: Image Quality Assessment with Reliable Region-Aware Reasoning cites this paper.

Zoom-IQA: Image Quality Assessment with Reliable Region-Aware Reasoning VideoScore: Building Automatic Metrics to Simulate Fine-grained Human Feedback for Video Generation

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-03T12:31:13.136990Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:31:13.136990Z digest=sha256:668a5a2031e31cd45664ba4558296548c0dbe1d40dafe7b937900ad98643d03d

Observation c9665ed3-4b2b-4ad7-9ce5-ef91a69b4ee7 · inbound

How Far Are Video Models from True Multimodal Reasoning? cites this paper.

How Far Are Video Models from True Multimodal Reasoning? VideoScore: Building Automatic Metrics to Simulate Fine-grained Human Feedback for Video Generation

Reference 23

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T12:51:04.183684Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T02:44:52.920816Z digest=sha256:3f6d5db7912bc87fb829ee2a93f673d030bb6653ffdf6f5636c2741e0274f7b5

Observation 1649c847-af35-432c-b465-af1b0e80e1db · inbound

Learning to Credit the Right Steps: Objective-aware Process Optimization for Visual Generation cites this paper.

Learning to Credit the Right Steps: Objective-aware Process Optimization for Visual Generation VideoScore: Building Automatic Metrics to Simulate Fine-grained Human Feedback for Video Generation

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-11T13:16:03.101352Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T02:10:26.580964Z digest=sha256:747feec96e5ee442c498cba2af59741207337577d7b2ddfe211d1616c0eed1cf

Observation 5b801cdc-3f3a-4afc-aac7-31f92adf255e · inbound

HuM-Eval: A Coarse-to-Fine Framework for Human-Centric Video Evaluation cites this paper.

HuM-Eval: A Coarse-to-Fine Framework for Human-Centric Video Evaluation VideoScore: Building Automatic Metrics to Simulate Fine-grained Human Feedback for Video Generation

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-11T23:31:15.290963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-07T16:53:42.069330Z digest=sha256:abe2bc6f2309faddf97851da9772bf1514a7d415a7df6d67679fedb0c4692eaa

Observation a66b5498-1e68-4eeb-a9ef-cb79201c070e · inbound

PhyMotion: Structured 3D Motion Reward for Physics-Grounded Human Video Generation cites this paper.

PhyMotion: Structured 3D Motion Reward for Physics-Grounded Human Video Generation VideoScore: Building Automatic Metrics to Simulate Fine-grained Human Feedback for Video Generation

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-15T02:49:41.459894Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-15T02:49:21.291716Z digest=sha256:a73a8ecdbf321a7c2188525764d81e5a93a242a347574166f259a055225d96dd

Observation 668a0ba2-e968-4a46-92e7-0de69e8ddcd5 · inbound

A Good Talk Does not Look Like a Summary, It Teaches You! Measuring Takeaways from Paper-to-Video Talks cites this paper.

A Good Talk Does not Look Like a Summary, It Teaches You! Measuring Takeaways from Paper-to-Video Talks VideoScore: Building Automatic Metrics to Simulate Fine-grained Human Feedback for Video Generation

Reference 68

Resolution
verified exact
arxiv_id, observed 2026-07-01T15:45:48.889899Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-30T00:59:50.478584Z digest=sha256:60c6c15f778c406cddbb1106c34156f347041b2761ff5f4d0271ddcadc925cbd

Observation a49cb5eb-3175-49f8-baef-8c2d6a57e06e · inbound

Reinforcement Learning: From Algorithms To Foundation Models cites this paper.

Reinforcement Learning: From Algorithms To Foundation Models VideoScore: Building Automatic Metrics to Simulate Fine-grained Human Feedback for Video Generation

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-01T17:45:04.144003Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T17:45:04.144003Z digest=sha256:a4ebba3859bd04792a8d20c1f8ef26e7ba05c9b9d0ae4af0969aa6326d496e5e

Observation e56cd1b3-8bd0-4e82-bd48-9ede5254a173 · inbound

CachedSearch: Training-Free Cached Exploration for Test-Time Search in Video Diffusion cites this paper.

CachedSearch: Training-Free Cached Exploration for Test-Time Search in Video Diffusion VideoScore: Building Automatic Metrics to Simulate Fine-grained Human Feedback for Video Generation

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-01T03:30:23.679425Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T03:30:23.679425Z digest=sha256:56351ec15d24b83cea1439d5b902b00746b246819a34cd47dadd07fd292059bb