Pith. sign in

Paper Citation Record · LEDGER

Identity-Preserving Text-to-Video Generation by Frequency Decomposition

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 18 inbound Pith citation observations for arXiv:2411.17440.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.17440 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 18 of 18 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 18 of 18 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:59:36.479084Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T08:49:42.422443Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 157cdf74-9c98-48e7-9520-9f6312c91447 · inbound

Learning Zero-Shot Subject-Driven Video Generation Using 1% Compute cites this paper.

Learning Zero-Shot Subject-Driven Video Generation Using 1% Compute Identity-Preserving Text-to-Video Generation by Frequency Decomposition

Reference 57

Resolution
verified exact
arxiv_id, observed 2026-05-22T17:51:54.418911Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-22T17:51:41.947939Z digest=sha256:1724242e410c5e89a5055f366931ecbdf98302902c5df50a0a579f6f94a976a4

Observation a0f306ac-9e16-44a7-9521-3256679dafe1 · inbound

ImgEdit: A Unified Image Editing Dataset and Benchmark cites this paper.

ImgEdit: A Unified Image Editing Dataset and Benchmark Identity-Preserving Text-to-Video Generation by Frequency Decomposition

Reference 81

Resolution
verified exact
arxiv_id, observed 2026-05-12T18:17:45.442169Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-12T18:17:45.123690Z digest=sha256:6c6f054227fe1f13361de5b8a1e47d4e55bd0db8ad58dbdc17aaa75442afb729

Observation ab3fa46f-dd86-49fc-afe4-7e1a6d1947b2 · inbound

OpenS2V-Nexus: A Detailed Benchmark and Million-Scale Dataset for Subject-to-Video Generation cites this paper.

OpenS2V-Nexus: A Detailed Benchmark and Million-Scale Dataset for Subject-to-Video Generation Identity-Preserving Text-to-Video Generation by Frequency Decomposition

Reference 113

Resolution
unresolved
no resolver link, observed 2026-08-07T13:59:36.479084Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:59:36.479084Z digest=sha256:708f348b8e11d5e753736e4d1aa5aab71f8d4b5f22f701596ae473ce75bcebe5

Observation 46a1ad32-721b-48e2-a3d0-d42d1fac1549 · inbound

UniWorld-V1: High-Resolution Semantic Encoders for Unified Visual Understanding and Generation cites this paper.

UniWorld-V1: High-Resolution Semantic Encoders for Unified Visual Understanding and Generation Identity-Preserving Text-to-Video Generation by Frequency Decomposition

Reference 52

Resolution
verified exact
arxiv_id, observed 2026-05-12T17:34:27.070137Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-12T17:34:26.951644Z digest=sha256:1baa6dd62a79c64421756e5b668e05e382cb86593cbf254d87dea27b2584d0d6

Observation aaadd1b3-ad36-4399-88df-ec3aae29cbb2 · inbound

UNIC: Unified In-Context Video Editing cites this paper.

UNIC: Unified In-Context Video Editing Identity-Preserving Text-to-Video Generation by Frequency Decomposition

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T10:51:42.872478Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:51:42.872478Z digest=sha256:b5ec88c268c9cb4208de3b73aba821f662634af3f2a3efbf67d1aeb9a2fd8ad2

Observation fef7acec-22d0-4247-bda7-769ee6ae4dbd · inbound

Follow-Your-Creation: Empowering 4D Creation through Video Inpainting cites this paper.

Follow-Your-Creation: Empowering 4D Creation through Video Inpainting Identity-Preserving Text-to-Video Generation by Frequency Decomposition

Reference 96

Resolution
unresolved
no resolver link, observed 2026-08-07T10:43:26.982019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:43:26.982019Z digest=sha256:a7c956251a17663037db4136115e0b4344ad727cd43a0c4bfbc26f4247ece8b9

Observation c3761732-b039-47c0-a26f-1e59f3b149d6 · inbound

PolyVivid: Vivid Multi-Subject Video Generation with Cross-Modal Interaction and Enhancement cites this paper.

PolyVivid: Vivid Multi-Subject Video Generation with Cross-Modal Interaction and Enhancement Identity-Preserving Text-to-Video Generation by Frequency Decomposition

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T05:30:33.052194Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:30:33.052194Z digest=sha256:1888004c9bfd4549672cb8bd0426540fb027bde310b5cc869fde5791e40533e9

Observation 241d36e5-0297-4afd-97b0-cd103d00fd43 · inbound

DreamActor-H1: High-Fidelity Human-Product Demonstration Video Generation via Motion-designed Diffusion Transformers cites this paper.

DreamActor-H1: High-Fidelity Human-Product Demonstration Video Generation via Motion-designed Diffusion Transformers Identity-Preserving Text-to-Video Generation by Frequency Decomposition

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-07T04:27:39.711058Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:27:39.711058Z digest=sha256:f161992cb4219855ecd3df00c08758e4129e4f675ace69c9f02f36e89a3c4513

Observation 75be7455-bf0e-4288-ad58-2f692377874b · inbound

SynMotion: Semantic-Visual Adaptation for Motion Customized Video Generation cites this paper.

SynMotion: Semantic-Visual Adaptation for Motion Customized Video Generation Identity-Preserving Text-to-Video Generation by Frequency Decomposition

Reference 101

Resolution
verified exact
arxiv_id, observed 2026-05-19T08:02:10.514105Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-19T07:59:49.398271Z digest=sha256:98450ca87de312dc73b632c11ee1fa0e2f075b40d73110adaabc324176b81089

Observation 9a45fe12-994c-4630-b9ce-94e3317f18ec · inbound

Tora2: Motion and Appearance Customized Diffusion Transformer for Multi-Entity Video Generation cites this paper.

Tora2: Motion and Appearance Customized Diffusion Transformer for Multi-Entity Video Generation Identity-Preserving Text-to-Video Generation by Frequency Decomposition

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T19:19:32.838413Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:19:32.838413Z digest=sha256:5cf5c273e88d092f8c28db13f6cdcb661cf42040618c11989db53d3f47e1f4c4

Observation 368fc191-8da6-4ce5-9ed0-67e9b7be3445 · inbound

A Summer Meridional Subsurface Temperature Dipole Mode in the South China Sea cites this paper.

A Summer Meridional Subsurface Temperature Dipole Mode in the South China Sea Identity-Preserving Text-to-Video Generation by Frequency Decomposition

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T04:44:59.127465Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T04:44:59.127465Z digest=sha256:5f1019b04dbfde285ec808ce33d401389466aa93c55a5a1c685b32c07376c449

Observation 1d4434e8-b090-4277-b226-97d0e06fddcc · inbound

Evolution of Video Generative Foundations cites this paper.

Evolution of Video Generative Foundations Identity-Preserving Text-to-Video Generation by Frequency Decomposition

Reference 196

Resolution
verified exact
arxiv_id, observed 2026-05-11T00:05:51.584551Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T18:41:38.616611Z digest=sha256:45316a102d22a173aa11aeb8d6a57cf8187cde21fd7800030331bb568029ad9e

Observation 0975307d-bb4c-460d-ae3a-0e738b08c3fc · inbound

Script-a-Video: Deep Structured Audio-visual Captions via Factorized Streams and Relational Grounding cites this paper.

Script-a-Video: Deep Structured Audio-visual Captions via Factorized Streams and Relational Grounding Identity-Preserving Text-to-Video Generation by Frequency Decomposition

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-11T11:11:03.483661Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T15:07:45.595260Z digest=sha256:37936e6ed269ac328d3cd8969b80fb928bf86df54b0371a0f64f3890ce6ce77a

Observation 609aa3f1-5b4b-42f1-a784-2f903f394879 · inbound

MoZoo:Unleashing Video Diffusion power in animal fur and muscle simulation cites this paper.

MoZoo:Unleashing Video Diffusion power in animal fur and muscle simulation Identity-Preserving Text-to-Video Generation by Frequency Decomposition

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-05-15T06:55:10.368181Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T06:53:13.152677Z digest=sha256:097ffde953947c19a14626b942d2ed32e0ed5fc19c48ac836c7aacd64c2d5a1e

Observation beaaa26e-a4a8-4af6-88e3-1f09932a983f · inbound

ElasticDiT: Efficient Diffusion Transformers via Elastic Architecture and Sparse Attention for High-Resolution Image Generation on Mobile Devices cites this paper.

ElasticDiT: Efficient Diffusion Transformers via Elastic Architecture and Sparse Attention for High-Resolution Image Generation on Mobile Devices Identity-Preserving Text-to-Video Generation by Frequency Decomposition

Reference 82

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T20:03:43.795138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-20T20:00:27.987481Z digest=sha256:1857951b6c93e4410e7a3048bd17b48d6eb5ded19ae2f9157713ae1ab44f7a3a

Observation 22a7e0d9-e635-464f-ba6d-a50f53319213 · inbound

Beyond Skeletons: Learning Animation Directly from Driving Videos with Same2X Training Strategy cites this paper.

Beyond Skeletons: Learning Animation Directly from Driving Videos with Same2X Training Strategy Identity-Preserving Text-to-Video Generation by Frequency Decomposition

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-07-02T16:27:09.308647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T22:41:06.600948Z digest=sha256:9461e320d9d64a4088f3f2cf1ead28cff2da20c4261d7452071b7810df4d306f

Observation 5a9005f6-bf64-40aa-a5a7-0e02b26ee519 · inbound

A Comprehensive Ecosystem for Open-Domain Customized Video Generation cites this paper.

A Comprehensive Ecosystem for Open-Domain Customized Video Generation Identity-Preserving Text-to-Video Generation by Frequency Decomposition

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-07-03T10:17:57.902889Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T10:06:33.572182Z digest=sha256:2b87aada024880e7b7238fffde01f08c710e8b2e18636d74e46c0f6e876b2920

Observation 2819a73a-0c3c-428b-8e6b-a5cfa6e6ec2d · inbound

Customizing Video Portraits via Identity-ActionDecoupling cites this paper.

Customizing Video Portraits via Identity-ActionDecoupling Identity-Preserving Text-to-Video Generation by Frequency Decomposition

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T08:49:42.423833Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-26T10:56:11.553110Z digest=sha256:919777fe3c04158cb6ee3b2544d14e5fd93b1552b2d7bd00a53e0cf43cf238b2