Pith. sign in

Paper Citation Record · LEDGER

VideoDirectorGPT: Consistent Multi-scene Video Generation via LLM-Guided Planning

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 28 inbound Pith citation observations for arXiv:2309.15091.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2309.15091 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 28 of 28 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 28 of 28 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T18:51:27.528264Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T14:59:55.973287Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation d0afbd36-c360-440f-8a0c-eaa9d78ac62a · inbound

Self-Correcting Text-to-Video Generation with Misalignment Detection and Localized Refinement cites this paper.

Self-Correcting Text-to-Video Generation with Misalignment Detection and Localized Refinement VideoDirectorGPT: Consistent Multi-scene Video Generation via LLM-Guided Planning

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-23T08:25:29.379696Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-23T08:25:01.468957Z digest=sha256:66eceb205ea7657b57d76d343ef83724e637cc984145fbf3bcb83f801ef5e485

Observation 1ea95042-a158-4c30-a6d4-89cc9b8d2c9d · inbound

Scene-Action Prompt Fusion for Coherent Text-to-Video Storytelling cites this paper.

Scene-Action Prompt Fusion for Coherent Text-to-Video Storytelling VideoDirectorGPT: Consistent Multi-scene Video Generation via LLM-Guided Planning

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-23T00:12:17.857765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-23T00:07:39.286486Z digest=sha256:f3ed6a10c2e8c5a95b5fec4e9820e7d153f7eac083de85f380493b6813571383

Observation e9449504-7d8a-4256-af88-17a596d64b25 · inbound

Character-Centered Dialogue Generation from Scene-Level Prompts cites this paper.

Character-Centered Dialogue Generation from Scene-Level Prompts VideoDirectorGPT: Consistent Multi-scene Video Generation via LLM-Guided Planning

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-05-22T13:34:53.653993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-22T13:31:43.083678Z digest=sha256:f5a8417210c3d4d11065ac729afc086d1f66c45b1e36c470d356f46bebfcadd2

Observation 75a20dc8-5800-4d8b-9cbc-dcbd3b4c8144 · inbound

A Survey on Long-Video Storytelling Generation: Architectures, Consistency, and Cinematic Quality cites this paper.

A Survey on Long-Video Storytelling Generation: Architectures, Consistency, and Cinematic Quality VideoDirectorGPT: Consistent Multi-scene Video Generation via LLM-Guided Planning

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-06T18:51:27.528264Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:51:27.528264Z digest=sha256:5b72d7646665a74b9f2307af25086d9f733ca631f405e841f2847a740c79ad9d

Observation 8fac7790-d4ac-4ad2-bb8e-81e0db60d796 · inbound

Enhancing Scene Transition Awareness in Video Generation via Post-Training cites this paper.

Enhancing Scene Transition Awareness in Video Generation via Post-Training VideoDirectorGPT: Consistent Multi-scene Video Generation via LLM-Guided Planning

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-06T14:43:03.268029Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:43:03.268029Z digest=sha256:fc9711e36071d5935f829972b97f3ff5705b068608a7f4b5804f3f7895abf537

Observation 666964f2-506d-4ca3-860c-6dda68be8baa · inbound

Test-time Prompt Refinement for Text-to-Image Models cites this paper.

Test-time Prompt Refinement for Text-to-Image Models VideoDirectorGPT: Consistent Multi-scene Video Generation via LLM-Guided Planning

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T15:01:52.294824Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:01:52.294824Z digest=sha256:c1df04a4cbe711f8eaa78bf901bdbe2822ba3d9c0fb1384445a7e4f14cf51464

Observation c3d95465-f087-40fb-8cfe-de98495c3b1e · inbound

A Survey on Evaluating Quality and Trustworthiness in LLM-Generated Data cites this paper.

A Survey on Evaluating Quality and Trustworthiness in LLM-Generated Data VideoDirectorGPT: Consistent Multi-scene Video Generation via LLM-Guided Planning

Reference 128

Resolution
unresolved
no resolver link, observed 2026-08-03T08:15:23.553433Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T08:15:23.553433Z digest=sha256:6080e03039478270c480b156714d89fcc6c07a6435d2084ab3408aa90a59c152

Observation d969f506-aacb-406b-a0e5-8d626183965a · inbound

The Script is All You Need: An Agentic Framework for Long-Horizon Dialogue-to-Cinematic Video Generation cites this paper.

The Script is All You Need: An Agentic Framework for Long-Horizon Dialogue-to-Cinematic Video Generation VideoDirectorGPT: Consistent Multi-scene Video Generation via LLM-Guided Planning

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-03T08:16:22.397859Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T08:16:22.397859Z digest=sha256:5a7be9da5642bbe6ce02c5bd25f403620e68e58ebf309b617fa6ed8d64ba2e1a

Observation dd2b4661-1c51-49bb-871b-c1d4cf6c585d · inbound

StoryBlender: Inter-Shot Consistent and Editable 3D Storyboard with Spatial-temporal Dynamics cites this paper.

StoryBlender: Inter-Shot Consistent and Editable 3D Storyboard with Spatial-temporal Dynamics VideoDirectorGPT: Consistent Multi-scene Video Generation via LLM-Guided Planning

Reference 15

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T23:28:26.324324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T23:24:26.557732Z digest=sha256:c77a6a7423ccefeb3c80d7c0c1074496cc7d177075c397de9db8212a2e3b8b7f

Observation f9cc28ea-0311-4458-8971-98a32e02264d · inbound

Evolution of Video Generative Foundations cites this paper.

Evolution of Video Generative Foundations VideoDirectorGPT: Consistent Multi-scene Video Generation via LLM-Guided Planning

Reference 68

Resolution
verified exact
arxiv_id, observed 2026-05-11T00:05:51.953522Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T18:41:38.616611Z digest=sha256:02781367b01656aa9eb7490883d39d32fd3569a34f717551289ce13bde866bc6

Observation 5ad7da70-e73c-4eb7-b59c-d03ee6e8ebd3 · inbound

Authoring for Living Worlds: Tool-Constrained LLM Agents for Executable Multi-Actor Scenarios cites this paper.

Authoring for Living Worlds: Tool-Constrained LLM Agents for Executable Multi-Actor Scenarios VideoDirectorGPT: Consistent Multi-scene Video Generation via LLM-Guided Planning

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:41:04.765086Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T15:22:45.975771Z digest=sha256:eb19fe9de2e13eba6b23e0b1b8407a858b9708e5c551931d02dd121e966177db

Observation 7b0890e4-f7cb-4315-95e3-03233b799254 · inbound

Authoring for Living Worlds: Tool-Constrained LLM Agents for Executable Multi-Actor Scenarios cites this paper.

Authoring for Living Worlds: Tool-Constrained LLM Agents for Executable Multi-Actor Scenarios VideoDirectorGPT: Consistent Multi-scene Video Generation via LLM-Guided Planning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-07-15T11:37:24.415703Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-15T11:37:24.415703Z digest=sha256:4788a3fa26ef6712ec2e8071b359d436c3b94026256964036cbbea43956e307a

Observation c6a34476-9b21-43a2-8afa-3810d14f4553 · inbound

Authoring for Living Worlds: Tool-Constrained LLM Agents for Executable Multi-Actor Scenarios cites this paper.

Authoring for Living Worlds: Tool-Constrained LLM Agents for Executable Multi-Actor Scenarios VideoDirectorGPT: Consistent Multi-scene Video Generation via LLM-Guided Planning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-02T16:34:31.229647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T16:34:31.229647Z digest=sha256:e1d0315e895d58b103d6bc1a13f421177f4142c1df27d78290cf85cde2568526

Observation cebc56e3-810c-4b00-bd87-12c294657764 · inbound

Ego-InBetween: Generating Object State Transitions in Ego-Centric Videos cites this paper.

Ego-InBetween: Generating Object State Transitions in Ego-Centric Videos VideoDirectorGPT: Consistent Multi-scene Video Generation via LLM-Guided Planning

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-10T06:56:48.002468Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T05:26:48.606759Z digest=sha256:08419636bb8b384d0652ad2765b70ead1e803a145f3359863da5e9bac81264ae

Observation 0c383d40-0285-442d-817d-50c420810adc · inbound

TS-Attn: Temporal-wise Separable Attention for Multi-Event Video Generation cites this paper.

TS-Attn: Temporal-wise Separable Attention for Multi-Event Video Generation VideoDirectorGPT: Consistent Multi-scene Video Generation via LLM-Guided Planning

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-10T02:53:29.929385Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T02:45:10.577070Z digest=sha256:f18e239605d1ea44dce38eeabc3665b0e441741d7e3f815c724fd7805d36082e

Observation f70c3f07-c849-4559-bbc2-addc8850810e · inbound

CineAGI: Character-Consistent Movie Creation through LLM-Orchestrated Multi-Modal Generation and Cross-Scene Integration cites this paper.

CineAGI: Character-Consistent Movie Creation through LLM-Orchestrated Multi-Modal Generation and Cross-Scene Integration VideoDirectorGPT: Consistent Multi-scene Video Generation via LLM-Guided Planning

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-11T21:31:17.835483Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-08T05:08:37.971891Z digest=sha256:4e701a2f5d2bf2e4fba4889bdd1eff0c1e581c855c1728a363f5f667b3a4982f

Observation f365793c-848b-404e-aea4-4ffad6d3875d · inbound

Cutscene Agent: An LLM Agent Framework for Automated 3D Cutscene Generation cites this paper.

Cutscene Agent: An LLM Agent Framework for Automated 3D Cutscene Generation VideoDirectorGPT: Consistent Multi-scene Video Generation via LLM-Guided Planning

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-12T08:41:25.983406Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-07T14:01:23.894966Z digest=sha256:248d2082fc6cecfa70f180651f75a332df295a588c2cbf9160c3d7e7d20b87b6

Observation eb0ae429-ec71-4d5e-aea3-5c5c84b69a8e · inbound

PresentAgent-2: Towards Generalist Multimodal Presentation Agents cites this paper.

PresentAgent-2: Towards Generalist Multimodal Presentation Agents VideoDirectorGPT: Consistent Multi-scene Video Generation via LLM-Guided Planning

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-05-13T02:32:06.450743Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T02:29:42.157339Z digest=sha256:c6acd456c9a1783bca031a2627888e88a65e9915b83e64b6562f4ec4f1893408

Observation 5aed57c1-ab0a-456e-99e3-1faef4181eae · inbound

EntityBench: Towards Entity-Consistent Long-Range Multi-Shot Video Generation cites this paper.

EntityBench: Towards Entity-Consistent Long-Range Multi-Shot Video Generation VideoDirectorGPT: Consistent Multi-scene Video Generation via LLM-Guided Planning

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-15T03:19:44.045709Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-15T03:16:03.450742Z digest=sha256:ede44a36f8f4d659b6b58b91461ab95854aba2e4282e7130b11976376d59b9b7

Observation 32d792ba-9b03-44b9-8542-d02a27fa58e1 · inbound

Soap2Soap: Long Cinematic Video Remaking via Multi-Agent Collaboration cites this paper.

Soap2Soap: Long Cinematic Video Remaking via Multi-Agent Collaboration VideoDirectorGPT: Consistent Multi-scene Video Generation via LLM-Guided Planning

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-20T13:18:18.344479Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-20T13:15:54.413960Z digest=sha256:10d0c10b3814e929fb136d4ca465257581aed02d47fe187d2549a10e38232f43

Observation 1ec408e2-a1fe-47aa-89ea-8b6505e9ad2a · inbound

One Sentence, One Drama: Personalized Short-Form Drama Generation via Multi-Agent Systems cites this paper.

One Sentence, One Drama: Personalized Short-Form Drama Generation via Multi-Agent Systems VideoDirectorGPT: Consistent Multi-scene Video Generation via LLM-Guided Planning

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-22T06:54:42.331471Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-22T06:51:58.805848Z digest=sha256:e669d6db45e245595d05ed9cf04ed1a7a3305b8b0808998d92b71e66aa044b50

Observation 181ade3d-c326-43be-8e74-941359de4e2f · inbound

LongAV-Compass: Towards Unified Evaluation of Minute-Scale Audio-Visual Generation Across T2AV, I2AV, and V2AV cites this paper.

LongAV-Compass: Towards Unified Evaluation of Minute-Scale Audio-Visual Generation Across T2AV, I2AV, and V2AV VideoDirectorGPT: Consistent Multi-scene Video Generation via LLM-Guided Planning

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-06-29T22:54:00.787750Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-29T22:52:38.330851Z digest=sha256:4fa8c4467a7b645491ceca4fdb3ad5f7d77b5ec8e77cc2b2f20d5de914981146

Observation 2e682433-2b29-4acf-8a57-9cf0fad346ae · inbound

MTAVG-Bench 2.0: Diagnosing Failure Modes of Cinematic Expressiveness in Multi-Talker Audio-Video Generation cites this paper.

MTAVG-Bench 2.0: Diagnosing Failure Modes of Cinematic Expressiveness in Multi-Talker Audio-Video Generation VideoDirectorGPT: Consistent Multi-scene Video Generation via LLM-Guided Planning

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-06-29T13:03:26.236751Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-29T12:58:56.555335Z digest=sha256:e2c548b29800d414e224f16c589fb3a7046edcf6572d5aec05e8a9424210ab43

Observation 24c35097-d0b2-4122-b3ff-7b66a11d1f6c · inbound

Crayotter: Traceable Multi-Agent Workflows for Long-Form Video Editing cites this paper.

Crayotter: Traceable Multi-Agent Workflows for Long-Form Video Editing VideoDirectorGPT: Consistent Multi-scene Video Generation via LLM-Guided Planning

Reference 12

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T21:26:14.472237Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-28T17:03:08.229448Z digest=sha256:8e446740fc8fa8a85e37db4f6c8691750b8ff47455343405a6572d3410d70c41

Observation 6519fd17-b1b9-4161-acad-429d090792c0 · inbound

Crayotter: Traceable Multi-Agent Workflows for Long-Form Video Editing cites this paper.

Crayotter: Traceable Multi-Agent Workflows for Long-Form Video Editing VideoDirectorGPT: Consistent Multi-scene Video Generation via LLM-Guided Planning

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-02T12:39:46.844935Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T12:39:46.844935Z digest=sha256:5302cf38f9b0724825f596e52735d56b842e359b8ad67affe02e51537fdf3605

Observation aca95c1a-b086-4a40-aa32-f12640120b6f · inbound

VideoWeaver: Evaluating and Evolving Skills for Agentic Long Video Generation cites this paper.

VideoWeaver: Evaluating and Evolving Skills for Agentic Long Video Generation VideoDirectorGPT: Consistent Multi-scene Video Generation via LLM-Guided Planning

Reference 38

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T20:57:23.180126Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-27T20:04:10.711238Z digest=sha256:751545ea80f2f954961ca6ba70846b5f1fef40ed72aa845221bf6c670c6ce8f6

Observation fbd92f55-828c-4ea6-8282-72c988c4de72 · inbound

Can Image Models Imagine Time? ImageTime: A Novel Benchmark for Probing Visual World Modeling Through Spatiotemporal Consistency cites this paper.

Can Image Models Imagine Time? ImageTime: A Novel Benchmark for Probing Visual World Modeling Through Spatiotemporal Consistency VideoDirectorGPT: Consistent Multi-scene Video Generation via LLM-Guided Planning

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-07-03T04:57:38.043953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-27T13:34:25.079037Z digest=sha256:3c6f7019edf1af1d93969b660ed28e617f4449c69a3714227e1d4ee911a42c21

Observation a1a804fd-5e67-40f6-8099-3996d31379d9 · inbound

DramaDirector: Geometry-Guided Short Drama Generation cites this paper.

DramaDirector: Geometry-Guided Short Drama Generation VideoDirectorGPT: Consistent Multi-scene Video Generation via LLM-Guided Planning

Reference 29

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T14:59:55.974623Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-26T01:57:15.550335Z digest=sha256:be7824a9b618054fde6f3881d025f9aef596adab88ab59e05f79d2f405d4a146