Pith. sign in

Paper Citation Record · LEDGER

ControlVideo: Training-free Controllable Text-to-Video Generation

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 36 inbound Pith citation observations for arXiv:2305.13077.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2305.13077 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 36 of 36 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 36 of 36 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T10:09:43.533968Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T17:40:00.876998Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 3370a45b-0ee5-45ce-a2d2-727b4f1223a5 · inbound

VideoCrafter1: Open Diffusion Models for High-Quality Video Generation cites this paper.

VideoCrafter1: Open Diffusion Models for High-Quality Video Generation ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 61

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T21:40:44.165165Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-14T21:40:43.956642Z digest=sha256:dd3e277d97b82c16e4fb96d9b20c1bfe1ba80788c038cfc2d4a6c628fb13f37a

Observation 1b7005b5-5792-41ba-ba8b-260cf0009ecf · inbound

CameraCtrl: Enabling Camera Control for Text-to-Video Generation cites this paper.

CameraCtrl: Enabling Camera Control for Text-to-Video Generation ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 168

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T02:06:23.772419Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-13T02:06:23.410241Z digest=sha256:cbc519dc3765d423b098bebac459d81bcc1f51c321f676661f4cdcccba6c8992

Observation bfe7b1ff-0811-4196-aed4-0880a4927512 · inbound

Self-Correcting Text-to-Video Generation with Misalignment Detection and Localized Refinement cites this paper.

Self-Correcting Text-to-Video Generation with Misalignment Detection and Localized Refinement ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 59

Resolution
verified exact
arxiv_id, observed 2026-05-23T08:25:29.374103Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-23T08:25:01.468957Z digest=sha256:9a0d022fc7ed17a6e9ad2fbb385580b49b5ed926135679dda0284d2d358f1b73

Observation 0d711ce7-1dd0-4996-bb42-fb2b2a989cef · inbound

AnyCharV: Bootstrap Controllable Character Video Generation with Fine-to-Coarse Guidance cites this paper.

AnyCharV: Bootstrap Controllable Character Video Generation with Fine-to-Coarse Guidance ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-08T10:09:43.533968Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T10:09:43.533968Z digest=sha256:14c4278b66f15b9fe811ec608dc360466df2f74d60957c4d630be598dec15157

Observation 56b581cd-a0ec-44b8-941d-f65361bd4b69 · inbound

Scene-Action Prompt Fusion for Coherent Text-to-Video Storytelling cites this paper.

Scene-Action Prompt Fusion for Coherent Text-to-Video Storytelling ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 43

Resolution
metadata mismatch
arxiv_id, observed 2026-05-23T00:12:17.866070Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-23T00:07:39.286486Z digest=sha256:529b767e06c1e75866f0f2a35c8dba82eaf9b478a8df486424b520630b3317f4

Observation df02ffec-6c20-4990-8811-a3ce28a0ec0c · inbound

DriVerse: Navigation World Model for Driving Simulation via Multimodal Trajectory Prompting and Motion Alignment cites this paper.

DriVerse: Navigation World Model for Driving Simulation via Multimodal Trajectory Prompting and Motion Alignment ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 75

Resolution
metadata mismatch
arxiv_id, observed 2026-05-22T17:51:54.771490Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-22T17:50:59.797593Z digest=sha256:03a264553cd0124c4e273e3316ac68c9824fa77abf473b78509c08aeb65d8941

Observation 67c04337-b401-4f40-a231-f54b8c7dfaf6 · inbound

Character-Centered Dialogue Generation from Scene-Level Prompts cites this paper.

Character-Centered Dialogue Generation from Scene-Level Prompts ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 70

Resolution
metadata mismatch
arxiv_id, observed 2026-05-22T13:34:53.728719Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-22T13:31:43.083678Z digest=sha256:1d80d6d0fb6197122d8d0a48d0efcbf7a815b98b4ca8784f41016114af870a26

Observation e40a1cf4-b60b-4a78-ad8a-2157fc8515e1 · inbound

EF-VI: Enhancing End-Frame Injection for Video Inbetweening cites this paper.

EF-VI: Enhancing End-Frame Injection for Video Inbetweening ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T13:44:18.599871Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:44:18.599871Z digest=sha256:dbe97e229cdd09c611962adb229b87bfddfde648ca02c5cacdbf75ac7fc0d427

Observation e804f18c-d55f-4e65-b0fd-2ad121e584c6 · inbound

Interactive Video Generation via Domain Adaptation cites this paper.

Interactive Video Generation via Domain Adaptation ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:30.703332Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:35:30.703332Z digest=sha256:604f119c32826135069ebe6f68cf20acc2cb5974730fc6c034062f6861722033

Observation c33bb063-84d5-43bb-b491-bd7ae33294ef · inbound

Dual-Expert Consistency Model for Efficient and High-Quality Video Generation cites this paper.

Dual-Expert Consistency Model for Efficient and High-Quality Video Generation ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-07T11:14:10.431319Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:14:10.431319Z digest=sha256:ba42c80cc1866f59be665b37c50e16c1ee96cad42270b52b72aa8d2d6df616b6

Observation 25f3bfd1-c6ac-4ec6-bb99-0234563cd95d · inbound

UNIC: Unified In-Context Video Editing cites this paper.

UNIC: Unified In-Context Video Editing ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T10:51:42.944866Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:51:42.944866Z digest=sha256:382ead4c1e4bb3e42105e49907d7e90ef271a3d31e5ad8730cad8c192f68119c

Observation 3cd75832-ee52-4205-a91f-c9d317662a89 · inbound

Controllable Coupled Image Generation via Diffusion Models cites this paper.

Controllable Coupled Image Generation via Diffusion Models ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-07T05:54:41.801190Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:54:41.801190Z digest=sha256:6dcd86b3e54ea4b90190757db7ea8e2ee414b26d1672be105d27b52706b8c580

Observation 9a89f376-e36c-4602-be0a-3a09ca30806b · inbound

When Distillation Breaks Motion Control: Restoring Generative Trajectories for Fast Video Generators cites this paper.

When Distillation Breaks Motion Control: Restoring Generative Trajectories for Fast Video Generators ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T23:12:53.181724Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:12:53.181724Z digest=sha256:f53aacd6c285f7c36f958aa9bbb516a6d4bdacbfee46b4739a3cd18f9cf6d9f4

Observation ef7e3720-4104-411f-b090-21da33ad60e7 · inbound

DFVEdit: Conditional Delta Flow Vector for Zero-shot Video Editing cites this paper.

DFVEdit: Conditional Delta Flow Vector for Zero-shot Video Editing ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T22:42:45.185458Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:42:45.185458Z digest=sha256:b5f00736e9e0736ebe930674d51dc0abfacf87b3012acca597fd932a117aaf5d

Observation 1dc34cde-99a8-4f00-8daa-3face3f52316 · inbound

AnyI2V: Animating Any Conditional Image with Motion Control cites this paper.

AnyI2V: Animating Any Conditional Image with Motion Control ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-06T20:25:29.057301Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:25:29.057301Z digest=sha256:e588d6d224d9d5199ceffe78c1edbb60f61fef23a753096368360e33905dd782

Observation 5c8d1a88-dd44-494c-ab66-5b0295f7f59c · inbound

HairShifter: Consistent and High-Fidelity Video Hair Transfer via Anchor-Guided Animation cites this paper.

HairShifter: Consistent and High-Fidelity Video Hair Transfer via Anchor-Guided Animation ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-06T16:44:40.462538Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:44:40.462538Z digest=sha256:5f699ab51f652382f68a2e7467d780f70cb2404c59a8a377606f03b339f16364

Observation 6f4480d7-b3f2-4914-bd8d-0eb6e57c9371 · inbound

Encapsulated Composition of Text-to-Image and Text-to-Video Models for High-Quality Video Synthesis cites this paper.

Encapsulated Composition of Text-to-Image and Text-to-Video Models for High-Quality Video Synthesis ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T16:24:07.699049Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:24:07.699049Z digest=sha256:c29f129282d4ec25613bf551d035e4b3ad43b4ae04ba8e37a0ce5da175e9c36f

Observation dd113f57-ff7a-4c4d-b99b-9a8bc628d593 · inbound

EndoControlMag: Robust Endoscopic Vascular Motion Magnification with Periodic Reference Resetting and Hierarchical Tissue-aware Dual-Mask Control cites this paper.

EndoControlMag: Robust Endoscopic Vascular Motion Magnification with Periodic Reference Resetting and Hierarchical Tissue-aware Dual-Mask Control ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-06T15:40:22.509935Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:40:22.509935Z digest=sha256:76850424ee3f0bcb7d65297aa40e14e5b90d1af07eb4c509b78e247f7b87b843

Observation d9f94ade-3628-4304-8e12-02d3839a66c2 · inbound

LongVie: Multimodal-Guided Controllable Ultra-Long Video Generation cites this paper.

LongVie: Multimodal-Guided Controllable Ultra-Long Video Generation ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T04:17:49.372651Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T04:17:49.372651Z digest=sha256:914c19dbd0b48aa061a45b71637bf4375c121c7c093189163617f136d599087e

Observation e6718536-ffea-46ce-9414-f0b541d54caf · inbound

Seeing Clearly, Forgetting Deeply: Revisiting Fine-Tuned Video Generators for Driving Simulation cites this paper.

Seeing Clearly, Forgetting Deeply: Revisiting Fine-Tuned Video Generators for Driving Simulation ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-05T17:19:21.556731Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:19:21.556731Z digest=sha256:c26984c3950d9aad0ae97e683099322e7d8efda5d898c2de4ada06835e8d87b6

Observation 33fcba22-8bcf-47a2-9635-729e3000a360 · inbound

Look Beyond: Two-Stage Scene View Generation via Panorama and Video Diffusion cites this paper.

Look Beyond: Two-Stage Scene View Generation via Panorama and Video Diffusion ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-05T13:15:16.447451Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T13:15:16.447451Z digest=sha256:02f25c4e66315c43ed072869d2a35ddf6c3fbeb8d288f1385ae09cf49dd2f761

Observation 6dd3ca95-4ddd-4d81-a4da-806ec7375f14 · inbound

DUDE: Diffusion-Based Unsupervised Cross-Domain Image Retrieval cites this paper.

DUDE: Diffusion-Based Unsupervised Cross-Domain Image Retrieval ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-05T10:21:49.797769Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:21:49.797769Z digest=sha256:82a35e4db14c28ead9ce9259aedec45a5b52247f63ffe01eef9a8b244204c225

Observation 6037ebff-dcfa-4145-8f01-9d0bbf3fc588 · inbound

ANYPORTAL: Zero-Shot Consistent Video Background Replacement cites this paper.

ANYPORTAL: Zero-Shot Consistent Video Background Replacement ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-04T22:13:51.564360Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T22:13:51.564360Z digest=sha256:7eb2a2945ec6d3353162da0074fe8a28afcb318c516a23c521c24c7f857a2774

Observation 4cabcea0-7acc-4ff7-8bd0-3f2a0a6ffbb5 · inbound

Measuring 3D Spatial Geometric Consistency in Dynamic Video Generation cites this paper.

Measuring 3D Spatial Geometric Consistency in Dynamic Video Generation ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 34

Resolution
unresolved
no resolver link, observed 2026-07-13T22:13:53.383917Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T22:13:53.383917Z digest=sha256:3fe0594a274aa655728778515c70b42b409c2674a4c5bf6c1bad5d591b1ac8b2

Observation 675cfb74-fe6f-442a-8074-f8124120e8f2 · inbound

Efficient Video Diffusion Models: Advancements and Challenges cites this paper.

Efficient Video Diffusion Models: Advancements and Challenges ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 190

Resolution
verified exact
arxiv_id, observed 2026-05-10T09:03:25.814675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T08:28:29.706249Z digest=sha256:efe1b47ef78f4f4eb711b591b03aae7a5515c4dc6752ab6c4e6a4ffdd1eda6e8

Observation 03b31827-bccf-4474-a905-46759d11b26f · inbound

FlowAnchor: Stabilizing the Editing Signal for Inversion-Free Video Editing cites this paper.

FlowAnchor: Stabilizing the Editing Signal for Inversion-Free Video Editing ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 39

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T19:06:11.376775Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-08T12:35:20.900453Z digest=sha256:fc636c572d4c363865730f6ec8929fdf35c80c327d0b6543aa8338b37f99d67a

Observation c85e5773-ddb0-4596-b94c-6b3160b42968 · inbound

SignVerse-2M: A Two-Million-Clip Pose-Native Universe of 55+ Sign Languages cites this paper.

SignVerse-2M: A Two-Million-Clip Pose-Native Universe of 55+ Sign Languages ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 21

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T09:56:04.781751Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T15:43:36.954114Z digest=sha256:aaf40625459bde6939c2dd2090212473cb62decb005689ea7ade34d2301b0fa4

Observation 0aef9bb5-9781-42a0-a22e-763b09c5e1a1 · inbound

Functionalization via Structure Completion and Motion Rectification cites this paper.

Functionalization via Structure Completion and Motion Rectification ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 248

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T12:28:17.085157Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-20T12:25:07.157086Z digest=sha256:9d72d230eea40d469ce077b09b7ff344c12cde3acea498707898b88a7e86def5

Observation 5baab3bb-e47e-45ee-8c36-b1f657caf21c · inbound

DeltaCam: Differential Intrinsic Camera Modeling for Video Generation cites this paper.

DeltaCam: Differential Intrinsic Camera Modeling for Video Generation ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-06-30T11:44:37.834716Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T11:44:08.734390Z digest=sha256:3b54fc3ed9f22ed5621161a76968e6d8b947dcf4ea3ef019fc5ea29356a78eb9

Observation ff382f18-c057-4347-98e2-b392fcdb53f8 · inbound

TeleMorpher: Toward Robust Simultaneous Motion-Location Editing cites this paper.

TeleMorpher: Toward Robust Simultaneous Motion-Location Editing ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 38

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T03:29:29.519115Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-26T18:05:53.625489Z digest=sha256:74a3df72f102917e590f3cc325807d95d485a0e8c2459ec0303fd0b42d049d63

Observation 0247e34f-c2f6-4e24-9e30-601382f083ee · inbound

IDAG-Edit: Multi-Object Video Editing via Instance-Decoupled Attention and Guidance cites this paper.

IDAG-Edit: Multi-Object Video Editing via Instance-Decoupled Attention and Guidance ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 9

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T07:49:39.347979Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-26T12:43:34.724121Z digest=sha256:a4995122474b16c3dce492c5c8a884e4b9f3cd4ce9c51b4f972e10540b74248b

Observation 8f24b24f-a6e2-4a85-8625-0ccf4ae8f43e · inbound

CineCap: Structured Reasoning with Spatio-Temporal Anchors for Cinematographic Video Captioning cites this paper.

CineCap: Structured Reasoning with Spatio-Temporal Anchors for Cinematographic Video Captioning ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-07-04T17:40:00.878645Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-25T23:29:24.520537Z digest=sha256:9f697b18672e621f037a7fc065a7704a290fe7f3ce68cb13a13f48405ffc7ba8

Observation 5daf2a82-62b9-45e7-9513-392c273391f4 · inbound

Directing the World: Fast Autoregressive Video Generation with Compositional Human-Camera Control cites this paper.

Directing the World: Fast Autoregressive Video Generation with Compositional Human-Camera Control ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 24

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T18:43:51.113091Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-29T05:03:56.624536Z digest=sha256:49c2db93ccd57cdd281f68b8f15ecd09a7300e156b0b42d6eab335094561fd9d

Observation 2a6077bc-ce8c-4bd3-8e5d-f6b9aabbaa8d · inbound

WorldDirector: Building Controllable World Simulators with Persistent Dynamic Memory cites this paper.

WorldDirector: Building Controllable World Simulators with Persistent Dynamic Memory ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 52

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T14:28:31.056068Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-07-03T14:28:24.922144Z digest=sha256:4eccf626567ee8b23d9ea4056e21b08883679be0e47bcf3ab1e0348ec0c689e4

Observation 5560d96c-3f2a-404a-9bd7-27c00ae454c3 · inbound

Multi-View Face and Gesture Animation with Dynamic Gaussians cites this paper.

Multi-View Face and Gesture Animation with Dynamic Gaussians ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 114

Resolution
unresolved
no resolver link, observed 2026-08-06T18:11:14.917337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:11:14.917337Z digest=sha256:9a29801e84516dd8bd2a9a9b974c764e8da707e81825c4b29d48b9d3253fcfb8

Observation 75f096b5-c924-4e3a-88c0-dc1b88acc952 · inbound

EmoWorld: A Decoupled Affective Field for Controllable Emotional Video Generation cites this paper.

EmoWorld: A Decoupled Affective Field for Controllable Emotional Video Generation ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T12:03:35.762749Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:03:35.762749Z digest=sha256:fad175a928f6edb276999df81b807c01901de929eb8b44941b116a35a8983d4f