Pith. sign in

Paper Citation Record · LEDGER

Towards A Better Metric for Text-to-Video Generation

As of 18 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 13 inbound Pith citation observations for arXiv:2401.07781.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2401.07781 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 13 of 13 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 13 of 13 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-12T14:26:22.078469Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T08:17:45.921041Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation b12f0e9c-2bdb-4259-ae5f-bed25f8e91d2 · inbound

Neuro-Symbolic Evaluation of Text-to-Video Models using Formal Verification cites this paper.

Neuro-Symbolic Evaluation of Text-to-Video Models using Formal Verification Towards A Better Metric for Text-to-Video Generation

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-12T14:26:22.078469Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T14:26:22.078469Z digest=sha256:7633df0b26480087ddf7346dfc16aead4656cb786c87b5ecf3825074a1cfded3

Observation c3a86d49-5613-41b7-80cf-553565c35d36 · inbound

HANDI: Hand-Centric Text-and-Image Conditioned Video Generation cites this paper.

HANDI: Hand-Centric Text-and-Image Conditioned Video Generation Towards A Better Metric for Text-to-Video Generation

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-11T21:45:10.147322Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T21:45:10.147322Z digest=sha256:6bb96c964d15fa0fd7fd9339b748cb1b5d25a5fdbbacf338c2dab80ada124434

Observation 75ac7204-a380-491e-8ef3-f137b7f14726 · inbound

Track4Gen: Teaching Video Diffusion Models to Track Points Improves Video Generation cites this paper.

Track4Gen: Teaching Video Diffusion Models to Track Points Improves Video Generation Towards A Better Metric for Text-to-Video Generation

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-11T20:10:47.626159Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:10:47.626159Z digest=sha256:7127fe7b21aa43753e4fb88617270c2965f39bcbcfed8500544b8bf8b4eabcc5

Observation 646c1d86-a489-4340-ba33-775ab3af110b · inbound

VideoDPO: Omni-Preference Alignment for Video Diffusion Generation cites this paper.

VideoDPO: Omni-Preference Alignment for Video Diffusion Generation Towards A Better Metric for Text-to-Video Generation

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-11T12:29:09.671977Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:29:09.671977Z digest=sha256:136d680c95cb2c5608bc6acaa318ea740d22fd1d3a4a4a4df3c9fb2d3a968050

Observation 2bc85961-b5da-4802-9ecb-18e9262208b7 · inbound

SafeMVDrive: Multi-view Safety-Critical Driving Video Synthesis in the Real World Domain cites this paper.

SafeMVDrive: Multi-view Safety-Critical Driving Video Synthesis in the Real World Domain Towards A Better Metric for Text-to-Video Generation

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:52.645716Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:52.645716Z digest=sha256:496d02f7f2b19cdffd735165558212fc1c7227e212a90c9de11ab75847782202

Observation c48c9115-7c42-4700-92cb-3b8a15840695 · inbound

OpenS2V-Nexus: A Detailed Benchmark and Million-Scale Dataset for Subject-to-Video Generation cites this paper.

OpenS2V-Nexus: A Detailed Benchmark and Million-Scale Dataset for Subject-to-Video Generation Towards A Better Metric for Text-to-Video Generation

Reference 101

Resolution
unresolved
no resolver link, observed 2026-08-07T13:59:35.615058Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:59:35.615058Z digest=sha256:eb73b190263a5e9b1b8ca6ee8550838147441e756029563233947cc95ec9cd6c

Observation d94fbeeb-4f90-4105-8084-6e11203b2c68 · inbound

FIFA: Unified Faithfulness Evaluation Framework for Text-to-Video and Video-to-Text Generation cites this paper.

FIFA: Unified Faithfulness Evaluation Framework for Text-to-Video and Video-to-Text Generation Towards A Better Metric for Text-to-Video Generation

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T19:06:15.889647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T19:06:15.889647Z digest=sha256:25f4435c6f5b6d7a6ea79180e8d243e7c77fd5aa9dea46f139fc0e8266527463

Observation e648002b-7a69-4abb-8ff1-b27fb60f6bd6 · inbound

Seeing Clearly, Forgetting Deeply: Revisiting Fine-Tuned Video Generators for Driving Simulation cites this paper.

Seeing Clearly, Forgetting Deeply: Revisiting Fine-Tuned Video Generators for Driving Simulation Towards A Better Metric for Text-to-Video Generation

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-05T17:19:20.415679Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:19:20.415679Z digest=sha256:7fbe8406afc84cc8aa0ff7ce83e163f5e214d0e6686a07bc2c37b7dcd983700b

Observation 00d1c516-17f7-4edc-ab3c-1c61dd76310f · inbound

When Numbers Speak: Aligning Textual Numerals and Visual Instances in Text-to-Video Diffusion Models cites this paper.

When Numbers Speak: Aligning Textual Numerals and Visual Instances in Text-to-Video Diffusion Models Towards A Better Metric for Text-to-Video Generation

Reference 64

Resolution
verified exact
arxiv_id, observed 2026-05-11T00:10:52.960076Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-10T18:39:26.420463Z digest=sha256:29a34518872e5b05d150ccb430bdf4aedb051dbad19135b6fe97971ac100e818

Observation d82f33b4-5749-48de-9e74-4788ad43eb00 · inbound

ARGUS: Stacked Multi-View Identity Mosaic Injection for Subject-Preserving Video Generation cites this paper.

ARGUS: Stacked Multi-View Identity Mosaic Injection for Subject-Preserving Video Generation Towards A Better Metric for Text-to-Video Generation

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-07-03T08:17:45.922557Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-27T10:46:56.871174Z digest=sha256:9637e55ae301fc0b8c74a06271f9cbeef7ecb316105ca06f0bfe7f37a2291b96

Observation 8739ceaa-b1c5-42ac-b2c1-4e45991434c3 · inbound

A Physics-Grounded Benchmark for Multi-Agent Dynamics in World Models cites this paper.

A Physics-Grounded Benchmark for Multi-Agent Dynamics in World Models Towards A Better Metric for Text-to-Video Generation

Reference 62

Resolution
verified exact
arxiv_id, observed 2026-06-30T09:54:34.937252Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-30T09:51:45.272494Z digest=sha256:0ff319c2224788da7704de311a469f06df955441fa550931d69aaccc30d73a6a

Observation 0beb3a96-74c9-4d75-83ef-09340d4870a0 · inbound

A Physics-Grounded Benchmark for Multi-Agent Dynamics in World Models cites this paper.

A Physics-Grounded Benchmark for Multi-Agent Dynamics in World Models Towards A Better Metric for Text-to-Video Generation

Reference 62

Resolution
unresolved
no resolver link, observed 2026-07-12T11:18:58.558989Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T11:18:58.558989Z digest=sha256:f6969e4d3d28e721ca4f1328a6a25687cc80a08e12e2d0a30f180050c32fb6b2

Observation 71475724-7945-4ddc-a12c-9e501990faf3 · inbound

ParticleGen: A Multi-Agent System for Particle Effects Generation cites this paper.

ParticleGen: A Multi-Agent System for Particle Effects Generation Towards A Better Metric for Text-to-Video Generation

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-04T02:48:05.322308Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T02:48:05.322308Z digest=sha256:40e6ae8d3e42f9fa4c1fc31f54852e96a9f08925784682947c27ebc2f8195f42