Pith. sign in

Paper Citation Record · LEDGER

ControlVideo: Training-free Controllable Text-to-Video Generation

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 42 inbound Pith citation observations for arXiv:2305.13077.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2305.13077 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 42 of 42 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 42 of 42 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-10T22:17:25.180547Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T17:40:00.876998Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 3370a45b-0ee5-45ce-a2d2-727b4f1223a5 · inbound

VideoCrafter1: Open Diffusion Models for High-Quality Video Generation cites this paper.

VideoCrafter1: Open Diffusion Models for High-Quality Video Generation ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 61

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T21:40:44.165165Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-14T21:40:43.956642Z digest=sha256:be7fb03a8aea72634ef3ba8dd2876074daf1a6eb65b00bad8fead37212bd28e7

Observation 1b7005b5-5792-41ba-ba8b-260cf0009ecf · inbound

CameraCtrl: Enabling Camera Control for Text-to-Video Generation cites this paper.

CameraCtrl: Enabling Camera Control for Text-to-Video Generation ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 168

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T02:06:23.772419Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-13T02:06:23.410241Z digest=sha256:91af429210f89302ed8c542080e3da6a67f54cc6a5e45c37985c9c2fe3f71224

Observation bfe7b1ff-0811-4196-aed4-0880a4927512 · inbound

Self-Correcting Text-to-Video Generation with Misalignment Detection and Localized Refinement cites this paper.

Self-Correcting Text-to-Video Generation with Misalignment Detection and Localized Refinement ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 59

Resolution
verified exact
arxiv_id, observed 2026-05-23T08:25:29.374103Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-23T08:25:01.468957Z digest=sha256:2586097e02718a35df9eb3b7d5a18c3821db194d55c3b54587828dfc1c3f1dd7

Observation 1b41b852-b17a-4be5-9bea-3fc585cc3021 · inbound

TDM: Temporally-Consistent Diffusion Model for All-in-One Real-World Video Restoration cites this paper.

TDM: Temporally-Consistent Diffusion Model for All-in-One Real-World Video Restoration ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-10T22:17:25.180547Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T22:17:25.180547Z digest=sha256:14a98f437b0327fa0477c550353354ccd4bedc1ee6f4e8515b633a2683cd0e25

Observation af673c95-9d2b-444d-9661-b666a1270a9e · inbound

FramePainter: Endowing Interactive Image Editing with Video Diffusion Priors cites this paper.

FramePainter: Endowing Interactive Image Editing with Video Diffusion Priors ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-10T20:33:55.713364Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:33:55.713364Z digest=sha256:16482dfa71143fbc7dfcff5d010f7d069d55ffbaa575f64b2fd6fc3c38869f58

Observation 27c27c54-7a09-4279-a0c9-c8000260a2f9 · inbound

VideoWorld: Exploring Knowledge Learning from Unlabeled Videos cites this paper.

VideoWorld: Exploring Knowledge Learning from Unlabeled Videos ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-10T19:49:20.143466Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T19:49:20.143466Z digest=sha256:c7afc05765cce090015325e25381c7abce1026ddd8f43097eedf134b96001f8d

Observation 55a93a03-dca6-45e3-a772-355c38ec9552 · inbound

DiffuEraser: A Diffusion Model for Video Inpainting cites this paper.

DiffuEraser: A Diffusion Model for Video Inpainting ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-10T19:25:45.268321Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T19:25:45.268321Z digest=sha256:351d3f9a94f14b6ef8056ccbae7c226f6090971053c5d69bf64006854c067b55

Observation 7c8ead4a-2144-429a-8e7a-8c6963282a0c · inbound

Video Depth Anything: Consistent Depth Estimation for Super-Long Videos cites this paper.

Video Depth Anything: Consistent Depth Estimation for Super-Long Videos ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-10T17:16:42.721276Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:16:42.721276Z digest=sha256:83b751f1803715c0eebce20ca2bb4a411440503f0b05e3a23a906eef450e1503

Observation 162719ec-b930-4feb-9a5c-4f9162495ce4 · inbound

Separate Motion from Appearance: Customizing Motion via Customizing Text-to-Video Diffusion Models cites this paper.

Separate Motion from Appearance: Customizing Motion via Customizing Text-to-Video Diffusion Models ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-10T11:17:24.896071Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T11:17:24.896071Z digest=sha256:d71de389b3d6791bfc85f986b78dc25975c10987affd405fead787817e2eddc6

Observation 0d711ce7-1dd0-4996-bb42-fb2b2a989cef · inbound

AnyCharV: Bootstrap Controllable Character Video Generation with Fine-to-Coarse Guidance cites this paper.

AnyCharV: Bootstrap Controllable Character Video Generation with Fine-to-Coarse Guidance ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-08T10:09:43.533968Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T10:09:43.533968Z digest=sha256:71538cd8cbbebec22088a9ef8bbad37a948c51fa4ac20cc265787922b5de6b83

Observation 56b581cd-a0ec-44b8-941d-f65361bd4b69 · inbound

Scene-Action Prompt Fusion for Coherent Text-to-Video Storytelling cites this paper.

Scene-Action Prompt Fusion for Coherent Text-to-Video Storytelling ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 43

Resolution
metadata mismatch
arxiv_id, observed 2026-05-23T00:12:17.866070Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-23T00:07:39.286486Z digest=sha256:9701aa01b9f302cc486d5e7e68c4e9c16d66c407538fc8e8b6caa89044817294

Observation df02ffec-6c20-4990-8811-a3ce28a0ec0c · inbound

DriVerse: Navigation World Model for Driving Simulation via Multimodal Trajectory Prompting and Motion Alignment cites this paper.

DriVerse: Navigation World Model for Driving Simulation via Multimodal Trajectory Prompting and Motion Alignment ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 75

Resolution
metadata mismatch
arxiv_id, observed 2026-05-22T17:51:54.771490Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-22T17:50:59.797593Z digest=sha256:feebd4fa160d0d86d7e829977c26137dfbeabe98bc561683d0a2dba75744b6d5

Observation 67c04337-b401-4f40-a231-f54b8c7dfaf6 · inbound

Character-Centered Dialogue Generation from Scene-Level Prompts cites this paper.

Character-Centered Dialogue Generation from Scene-Level Prompts ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 70

Resolution
metadata mismatch
arxiv_id, observed 2026-05-22T13:34:53.728719Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-22T13:31:43.083678Z digest=sha256:f95bf8c3ee3944b1cb079e03e19fbc9d51a427033dbcf68a3c2da386f50a30ef

Observation e40a1cf4-b60b-4a78-ad8a-2157fc8515e1 · inbound

EF-VI: Enhancing End-Frame Injection for Video Inbetweening cites this paper.

EF-VI: Enhancing End-Frame Injection for Video Inbetweening ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T13:44:18.599871Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:44:18.599871Z digest=sha256:dbe97e229cdd09c611962adb229b87bfddfde648ca02c5cacdbf75ac7fc0d427

Observation e804f18c-d55f-4e65-b0fd-2ad121e584c6 · inbound

Interactive Video Generation via Domain Adaptation cites this paper.

Interactive Video Generation via Domain Adaptation ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-07T12:35:30.703332Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:35:30.703332Z digest=sha256:604f119c32826135069ebe6f68cf20acc2cb5974730fc6c034062f6861722033

Observation c33bb063-84d5-43bb-b491-bd7ae33294ef · inbound

Dual-Expert Consistency Model for Efficient and High-Quality Video Generation cites this paper.

Dual-Expert Consistency Model for Efficient and High-Quality Video Generation ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-07T11:14:10.431319Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:14:10.431319Z digest=sha256:db56f971c1a9bfc7f7e952672e9de86b89b9f51dca4c7ac25fc057a730cb60c7

Observation 25f3bfd1-c6ac-4ec6-bb99-0234563cd95d · inbound

UNIC: Unified In-Context Video Editing cites this paper.

UNIC: Unified In-Context Video Editing ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T10:51:42.944866Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:51:42.944866Z digest=sha256:382ead4c1e4bb3e42105e49907d7e90ef271a3d31e5ad8730cad8c192f68119c

Observation 3cd75832-ee52-4205-a91f-c9d317662a89 · inbound

Controllable Coupled Image Generation via Diffusion Models cites this paper.

Controllable Coupled Image Generation via Diffusion Models ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-07T05:54:41.801190Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:54:41.801190Z digest=sha256:6dcd86b3e54ea4b90190757db7ea8e2ee414b26d1672be105d27b52706b8c580

Observation 9a89f376-e36c-4602-be0a-3a09ca30806b · inbound

When Distillation Breaks Motion Control: Restoring Generative Trajectories for Fast Video Generators cites this paper.

When Distillation Breaks Motion Control: Restoring Generative Trajectories for Fast Video Generators ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T23:12:53.181724Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:12:53.181724Z digest=sha256:f53aacd6c285f7c36f958aa9bbb516a6d4bdacbfee46b4739a3cd18f9cf6d9f4

Observation ef7e3720-4104-411f-b090-21da33ad60e7 · inbound

DFVEdit: Conditional Delta Flow Vector for Zero-shot Video Editing cites this paper.

DFVEdit: Conditional Delta Flow Vector for Zero-shot Video Editing ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T22:42:45.185458Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:42:45.185458Z digest=sha256:b5f00736e9e0736ebe930674d51dc0abfacf87b3012acca597fd932a117aaf5d

Observation 1dc34cde-99a8-4f00-8daa-3face3f52316 · inbound

AnyI2V: Animating Any Conditional Image with Motion Control cites this paper.

AnyI2V: Animating Any Conditional Image with Motion Control ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-06T20:25:29.057301Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:25:29.057301Z digest=sha256:e588d6d224d9d5199ceffe78c1edbb60f61fef23a753096368360e33905dd782

Observation 5c8d1a88-dd44-494c-ab66-5b0295f7f59c · inbound

HairShifter: Consistent and High-Fidelity Video Hair Transfer via Anchor-Guided Animation cites this paper.

HairShifter: Consistent and High-Fidelity Video Hair Transfer via Anchor-Guided Animation ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-06T16:44:40.462538Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:44:40.462538Z digest=sha256:02c6d8401b5c2d5c196a318ecd8c60928f205f1cb9670428d2f6275c506cc8dc

Observation 6f4480d7-b3f2-4914-bd8d-0eb6e57c9371 · inbound

Encapsulated Composition of Text-to-Image and Text-to-Video Models for High-Quality Video Synthesis cites this paper.

Encapsulated Composition of Text-to-Image and Text-to-Video Models for High-Quality Video Synthesis ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T16:24:07.699049Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:24:07.699049Z digest=sha256:d76454b9ba7cb61547f463df873ad9e0ac0451915ac7ee1969f5771028ab5e95

Observation dd113f57-ff7a-4c4d-b99b-9a8bc628d593 · inbound

EndoControlMag: Robust Endoscopic Vascular Motion Magnification with Periodic Reference Resetting and Hierarchical Tissue-aware Dual-Mask Control cites this paper.

EndoControlMag: Robust Endoscopic Vascular Motion Magnification with Periodic Reference Resetting and Hierarchical Tissue-aware Dual-Mask Control ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-06T15:40:22.509935Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:40:22.509935Z digest=sha256:76850424ee3f0bcb7d65297aa40e14e5b90d1af07eb4c509b78e247f7b87b843

Observation d9f94ade-3628-4304-8e12-02d3839a66c2 · inbound

LongVie: Multimodal-Guided Controllable Ultra-Long Video Generation cites this paper.

LongVie: Multimodal-Guided Controllable Ultra-Long Video Generation ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T04:17:49.372651Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T04:17:49.372651Z digest=sha256:914c19dbd0b48aa061a45b71637bf4375c121c7c093189163617f136d599087e

Observation e6718536-ffea-46ce-9414-f0b541d54caf · inbound

Seeing Clearly, Forgetting Deeply: Revisiting Fine-Tuned Video Generators for Driving Simulation cites this paper.

Seeing Clearly, Forgetting Deeply: Revisiting Fine-Tuned Video Generators for Driving Simulation ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-05T17:19:21.556731Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:19:21.556731Z digest=sha256:c26984c3950d9aad0ae97e683099322e7d8efda5d898c2de4ada06835e8d87b6

Observation 33fcba22-8bcf-47a2-9635-729e3000a360 · inbound

Look Beyond: Two-Stage Scene View Generation via Panorama and Video Diffusion cites this paper.

Look Beyond: Two-Stage Scene View Generation via Panorama and Video Diffusion ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-05T13:15:16.447451Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T13:15:16.447451Z digest=sha256:02f25c4e66315c43ed072869d2a35ddf6c3fbeb8d288f1385ae09cf49dd2f761

Observation 6dd3ca95-4ddd-4d81-a4da-806ec7375f14 · inbound

DUDE: Diffusion-Based Unsupervised Cross-Domain Image Retrieval cites this paper.

DUDE: Diffusion-Based Unsupervised Cross-Domain Image Retrieval ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-05T10:21:49.797769Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:21:49.797769Z digest=sha256:82a35e4db14c28ead9ce9259aedec45a5b52247f63ffe01eef9a8b244204c225

Observation 6037ebff-dcfa-4145-8f01-9d0bbf3fc588 · inbound

ANYPORTAL: Zero-Shot Consistent Video Background Replacement cites this paper.

ANYPORTAL: Zero-Shot Consistent Video Background Replacement ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-04T22:13:51.564360Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T22:13:51.564360Z digest=sha256:1d7073432837ba2600e45858f9671b5dbedbd0aeee79643d6e8f5d6572f4d86b

Observation 4cabcea0-7acc-4ff7-8bd0-3f2a0a6ffbb5 · inbound

Measuring 3D Spatial Geometric Consistency in Dynamic Video Generation cites this paper.

Measuring 3D Spatial Geometric Consistency in Dynamic Video Generation ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 34

Resolution
unresolved
no resolver link, observed 2026-07-13T22:13:53.383917Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T22:13:53.383917Z digest=sha256:3fe0594a274aa655728778515c70b42b409c2674a4c5bf6c1bad5d591b1ac8b2

Observation 675cfb74-fe6f-442a-8074-f8124120e8f2 · inbound

Efficient Video Diffusion Models: Advancements and Challenges cites this paper.

Efficient Video Diffusion Models: Advancements and Challenges ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 190

Resolution
verified exact
arxiv_id, observed 2026-05-10T09:03:25.814675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T08:28:29.706249Z digest=sha256:087cb1771b130db332cb4fa174cf21b751b2b4389b696ff84fa8ec0a082576e5

Observation 03b31827-bccf-4474-a905-46759d11b26f · inbound

FlowAnchor: Stabilizing the Editing Signal for Inversion-Free Video Editing cites this paper.

FlowAnchor: Stabilizing the Editing Signal for Inversion-Free Video Editing ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 39

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T19:06:11.376775Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-08T12:35:20.900453Z digest=sha256:0edc1a5e3f3c48293fba5cc300cc6ca99f1e0d636a62c1e08e9a649aa1e8df7d

Observation c85e5773-ddb0-4596-b94c-6b3160b42968 · inbound

SignVerse-2M: A Two-Million-Clip Pose-Native Universe of 55+ Sign Languages cites this paper.

SignVerse-2M: A Two-Million-Clip Pose-Native Universe of 55+ Sign Languages ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 21

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T09:56:04.781751Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T15:43:36.954114Z digest=sha256:27d5281624d521c8781b91dd1ff9338417ed68b78195c0931a3488039310bd7d

Observation 0aef9bb5-9781-42a0-a22e-763b09c5e1a1 · inbound

Functionalization via Structure Completion and Motion Rectification cites this paper.

Functionalization via Structure Completion and Motion Rectification ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 248

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T12:28:17.085157Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-20T12:25:07.157086Z digest=sha256:817bbd1389713dfbb2b7f1f582054d4c198fcc2744ff2eaeafdb92990e4b9d13

Observation 5baab3bb-e47e-45ee-8c36-b1f657caf21c · inbound

DeltaCam: Differential Intrinsic Camera Modeling for Video Generation cites this paper.

DeltaCam: Differential Intrinsic Camera Modeling for Video Generation ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-06-30T11:44:37.834716Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T11:44:08.734390Z digest=sha256:f7dfe3216c84342e9c4dba811223df1cca727d3563cb9550b1fd0f4f085657fc

Observation ff382f18-c057-4347-98e2-b392fcdb53f8 · inbound

TeleMorpher: Toward Robust Simultaneous Motion-Location Editing cites this paper.

TeleMorpher: Toward Robust Simultaneous Motion-Location Editing ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 38

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T03:29:29.519115Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-26T18:05:53.625489Z digest=sha256:c696da5b7aa4272d195690ea991480e8309f12e48a98658a10d4d0488405b49e

Observation 0247e34f-c2f6-4e24-9e30-601382f083ee · inbound

IDAG-Edit: Multi-Object Video Editing via Instance-Decoupled Attention and Guidance cites this paper.

IDAG-Edit: Multi-Object Video Editing via Instance-Decoupled Attention and Guidance ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 9

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T07:49:39.347979Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-26T12:43:34.724121Z digest=sha256:60ba1b10a38aeb12cba83e1b076c724e479b46e649e1e8fa6cd7e4eccff0cb67

Observation 8f24b24f-a6e2-4a85-8625-0ccf4ae8f43e · inbound

CineCap: Structured Reasoning with Spatio-Temporal Anchors for Cinematographic Video Captioning cites this paper.

CineCap: Structured Reasoning with Spatio-Temporal Anchors for Cinematographic Video Captioning ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-07-04T17:40:00.878645Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-25T23:29:24.520537Z digest=sha256:ac3d68abb9eccc5091f8f7a8fdf40a0e79e65d1227b7b492f4fc5661d58cc83c

Observation 5daf2a82-62b9-45e7-9513-392c273391f4 · inbound

Directing the World: Fast Autoregressive Video Generation with Compositional Human-Camera Control cites this paper.

Directing the World: Fast Autoregressive Video Generation with Compositional Human-Camera Control ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 24

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T18:43:51.113091Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-29T05:03:56.624536Z digest=sha256:0589e60d090caaf8ab0c45ce5e6d2f80859ea105b8123ea82dad59d4bce375a5

Observation 2a6077bc-ce8c-4bd3-8e5d-f6b9aabbaa8d · inbound

WorldDirector: Building Controllable World Simulators with Persistent Dynamic Memory cites this paper.

WorldDirector: Building Controllable World Simulators with Persistent Dynamic Memory ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 52

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T14:28:31.056068Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-03T14:28:24.922144Z digest=sha256:c8bc9b3ab899185f746506e55b8f636c1faefbcb141ae275cd5ba40afe992617

Observation 5560d96c-3f2a-404a-9bd7-27c00ae454c3 · inbound

Multi-View Face and Gesture Animation with Dynamic Gaussians cites this paper.

Multi-View Face and Gesture Animation with Dynamic Gaussians ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 114

Resolution
unresolved
no resolver link, observed 2026-08-06T18:11:14.917337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T18:11:14.917337Z digest=sha256:5fad009bd4089d503e38b8a64e4fbe4da39d082efad80052595b111fd4968473

Observation 75f096b5-c924-4e3a-88c0-dc1b88acc952 · inbound

EmoWorld: A Decoupled Affective Field for Controllable Emotional Video Generation cites this paper.

EmoWorld: A Decoupled Affective Field for Controllable Emotional Video Generation ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T12:03:35.762749Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:03:35.762749Z digest=sha256:db08d1ca32f9299718e6c5e7d379f2d24a0ef379375c0fd1dcf36a4bb22b13d0