Pith. sign in

Paper Citation Record · LEDGER

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation

As of 8 August 2026, this Paper Citation Record lists 50 of 50 outbound references and 8 inbound Pith citation observations for arXiv:2502.08639.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.08639 v1

Coverage vector

measured 50 of 50 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-08T00:02:56.890698Z

measured 58 of 58 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 8 of 8 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T10:51:43.018033Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-15T12:20:00.524134Z

Reference resolution

50 of 50 outbound references displayed

  • verified exact0
  • verified fuzzy21
  • unresolved29
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 3d7055a8-32bc-4e2b-97e1-e79d484629d5 · outbound

This paper cites Loosecontrol: Lifting controlnet for generalized depth conditioning.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation Loosecontrol: Lifting controlnet for generalized depth conditioning

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:02:57.415015Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T00:02:56.729320Z digest=sha256:1f9ab5a31879a2a9e8a9d387d37246d907a9a15dc38ce3703a00820c27cc014c

Observation 34234d4e-dc0f-4f30-8bb2-6e4825a8fa9a · outbound

This paper cites GS-DiT: Advancing Video Generation with Pseudo 4D Gaussian Fields through Efficient Dense 3D Point Tracking.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation GS-DiT: Advancing Video Generation with Pseudo 4D Gaussian Fields through Efficient Dense 3D Point Tracking

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-08T00:02:56.732859Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:02:56.732859Z digest=sha256:a02865c291ab6d605a46f8149284e675380697ab2b64da6738b898954de93a1d

Observation 3e1da12f-295f-4186-8634-1b0d399dfa8e · outbound

This paper cites Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-08T00:02:56.736420Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:02:56.736420Z digest=sha256:aaf40b9da2d957b416d82fa7087b050098c1dc7c12bf56d3aa2b3343a981e01a

Observation 2161f796-4dea-4823-92ef-0c07279a7564 · outbound

This paper cites VideoCrafter1: Open Diffusion Models for High-Quality Video Generation.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-08T00:02:56.739564Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:02:56.739564Z digest=sha256:33cda134367765f5d2108b978cd0cebe3fa5b8f430bacc1a91909c350ef765ce

Observation 91ce4814-df97-4855-aec5-2a96e9f17145 · outbound

This paper cites PixArt-$\alpha$: Fast Training of Diffusion Transformer for Photorealistic Text-to-Image Synthesis.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation PixArt-$\alpha$: Fast Training of Diffusion Transformer for Photorealistic Text-to-Image Synthesis

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-08T00:02:56.742731Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:02:56.742731Z digest=sha256:c43a6642a68e4cdaa5497b6f26b6944f02e0351fd1ba7c549df230c3f0e4cf3d

Observation 866547a4-cdd9-4e22-a574-04109815eb0b · outbound

This paper cites Control-A-Video: Controllable Text-to-Video Diffusion Models with Motion Prior and Reward Feedback Learning.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation Control-A-Video: Controllable Text-to-Video Diffusion Models with Motion Prior and Reward Feedback Learning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-08T00:02:56.746313Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:02:56.746313Z digest=sha256:f3597faee98526ef525b944a087ff8ae97470cbfa00fc50131e112de54ebd7cd

Observation 7a9b1d57-9407-4db1-a43a-1816bb7e8a48 · outbound

This paper cites Perception-as-Control: Fine-grained Controllable Image Animation with 3D-aware Motion Representation.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation Perception-as-Control: Fine-grained Controllable Image Animation with 3D-aware Motion Representation

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-08T00:02:56.750090Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:02:56.750090Z digest=sha256:b67464dd9302891f8c772509fdc129c57d2bbbe2e6832626abac208555548d34

Observation 077f6917-1086-4fdd-8eca-db4e36324122 · outbound

This paper cites Patch n’pack: Navit, a vision transformer for any aspect ratio and resolu- tion.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation Patch n’pack: Navit, a vision transformer for any aspect ratio and resolu- tion

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:02:57.405949Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T00:02:56.753614Z digest=sha256:7c90abcfafca6be551dc68df0491b2d2c5585d12aaf29a6e65085c48037b5acb

Observation 19660e4c-a8c5-4a24-9f3c-c301d28c31ed · outbound

This paper cites Scaling rectified flow transformers for high-resolution image synthesis.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation Scaling rectified flow transformers for high-resolution image synthesis

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:02:57.397458Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T00:02:56.756898Z digest=sha256:a36e2ff9a47cec90776da1a842ef51a486c89cd8ae47da163f52c767fe7071e7

Observation 1fa4bd7f-8e40-4784-a71e-cb116ca056e7 · outbound

This paper cites 3DTrajMaster: Mastering 3D Trajectory for Multi-Entity Motion in Video Generation.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation 3DTrajMaster: Mastering 3D Trajectory for Multi-Entity Motion in Video Generation

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-08T00:02:56.760277Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:02:56.760277Z digest=sha256:83fe2407c0bcccd0ebef8bccf686539586b16ae28f22fa59888c5a55f8c917c2

Observation 2fe49bbb-275f-40b7-ba3d-82fb1a129cdb · outbound

This paper cites Motion Prompting: Controlling Video Generation with Motion Trajectories.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation Motion Prompting: Controlling Video Generation with Motion Trajectories

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-08T00:02:56.763785Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:02:56.763785Z digest=sha256:2a9b6216313022fc507ce5376f81578e20118af88e22c5653e30a9ce12b32404

Observation 66444dcb-3d0e-4aa9-b310-cae0a9e0fe6c · outbound

This paper cites Diffusion as Shader: 3D-aware Video Diffusion for Versatile Video Generation Control.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation Diffusion as Shader: 3D-aware Video Diffusion for Versatile Video Generation Control

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-08T00:02:56.767177Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:02:56.767177Z digest=sha256:62da760e985b9807f3d37f854c650d43f2ef5d23e0dade8add107b612987a65f

Observation 73e66d04-f15f-4942-a413-ecec944bb694 · outbound

This paper cites Sparsectrl: Adding sparse controls to text-to-video diffusion mod- els, 2023.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation Sparsectrl: Adding sparse controls to text-to-video diffusion mod- els, 2023

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:02:57.389476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T00:02:56.770594Z digest=sha256:12cc25e368a92750828bd9fc2bfd527221fbd1c7e846dc48071c2051a4f915c3

Observation 883d6db8-5aa2-4f10-930d-442a70661dc1 · outbound

This paper cites AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-08T00:02:56.773765Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:02:56.773765Z digest=sha256:b9a778223e9e5bd4a3cbebb8c4195a5764f98c50a81e511e636e893341abee5f

Observation 5bcde63a-fca3-448f-a866-ba5cb2699e87 · outbound

This paper cites Cameractrl: Enabling camera control for text-to-video generation,.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation Cameractrl: Enabling camera control for text-to-video generation,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:02:57.381179Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T00:02:56.777309Z digest=sha256:18baf9e2bf4621bd37e9ea4cc0f455b048f602ac9dbda05c9666b5f22ce28368

Observation ee1a13c5-2739-4bf5-95d8-22c4c6bb4a1b · outbound

This paper cites Classifier-Free Diffusion Guidance.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation Classifier-Free Diffusion Guidance

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-08T00:02:56.780647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:02:56.780647Z digest=sha256:29925f6f558494476d47568ca27b99bb37e2cf8885917fd292365d1a30b33b2d

Observation 4f8aba04-1035-46bf-ab98-61c1ec602ea0 · outbound

This paper cites Denoising diffusion probabilistic models.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation Denoising diffusion probabilistic models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-08T00:02:56.784510Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:02:56.784510Z digest=sha256:0093b2e0b3aa787d1906a1974cbc0ec5c6551ed139244b28db335acb7c8c6d4a

Observation dde0967a-ce82-4fed-8da0-ae40f691a31e · outbound

This paper cites Animate anyone: Consistent and controllable image-to-video synthesis for character animation.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation Animate anyone: Consistent and controllable image-to-video synthesis for character animation

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:02:57.367946Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T00:02:56.787678Z digest=sha256:bae87f60f77cb3585e8f6922c76d7f90fc65f81054fc1ca4a6b7bb9540ac4a86

Observation 1b6582cf-f5a1-4d2b-8d8d-130447fbddf2 · outbound

This paper cites Auto-Encoding Variational Bayes.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation Auto-Encoding Variational Bayes

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-08T00:02:56.790984Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:02:56.790984Z digest=sha256:eec118e806808ebb2baf714c8030f09907e62975e035f472121aec4485776c91

Observation 546c9a69-90eb-430d-befb-d727c70ffdf8 · outbound

This paper cites Adam: A Method for Stochastic Optimization.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation Adam: A Method for Stochastic Optimization

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-08T00:02:56.794985Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:02:56.794985Z digest=sha256:24b0cc9b149fee93ddaff2a40d1f63f83f3bca800c34ebe7a1fc1f30de878578

Observation 8c1910f7-b06e-408e-9fdf-e35a5de8835d · outbound

This paper cites Onlyflow: Optical flow based motion condition- ing for video diffusion models.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation Onlyflow: Optical flow based motion condition- ing for video diffusion models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-08T00:02:56.798165Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:02:56.798165Z digest=sha256:24f60101cab52d6b4d8aa32d32d3b89b31330109070c04407953d126f69dde64

Observation de8b809f-963a-4304-8a84-6a47247d9194 · outbound

This paper cites Microsoft coco: Com- mon objects in context.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation Microsoft coco: Com- mon objects in context

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:02:57.358490Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T00:02:56.801473Z digest=sha256:6f6e5f26e40513b7137d019e62d393ff97923ea7d3e7a2f53e9b82e572009054

Observation 0f6cdc96-a0fd-4c86-a8ce-7a1dece3959d · outbound

This paper cites Flow Matching for Generative Modeling.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation Flow Matching for Generative Modeling

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-08T00:02:56.804612Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:02:56.804612Z digest=sha256:c4f1e1e841ddcc820fa436c6631acbb4c5dd7fc18f971ef4a39a989056ec1efe

Observation 217560b7-e689-4c43-b3e9-b6483542b5d3 · outbound

This paper cites Grounding DINO: Marrying DINO with Grounded Pre-Training for Open-Set Object Detection.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation Grounding DINO: Marrying DINO with Grounded Pre-Training for Open-Set Object Detection

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-08T00:02:56.807850Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:02:56.807850Z digest=sha256:edf48bb77a01c4541972cb847999474a0913c1a98dc4e20b30a14c007dd9daea

Observation 06f3bfe0-1217-4eed-a73f-36241564ad16 · outbound

This paper cites Are GANs Created Equal? A Large-Scale Study.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation Are GANs Created Equal? A Large-Scale Study

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-08T00:02:56.811116Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:02:56.811116Z digest=sha256:fed7b457d35b5bd3ef56c4f5a2373d3e0d8e4be5ba65d38ed2657919702be30c

Observation ae8b787e-d25b-470e-9254-acc6f0aa3025 · outbound

This paper cites T2i- adapter: Learning adapters to dig out more control- lable ability for text-to-image diffusion models.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation T2i- adapter: Learning adapters to dig out more control- lable ability for text-to-image diffusion models

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:02:57.349395Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T00:02:56.814377Z digest=sha256:088cd5ec293bd9dbfe397d31ba797bae557a541b01a96aaea9c6b671254eb1b5

Observation 173cba0a-faa2-4d25-8311-b572145c1921 · outbound

This paper cites Scalable diffu- sion models with transformers.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation Scalable diffu- sion models with transformers

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:02:57.340055Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T00:02:56.817478Z digest=sha256:0c14a7d8c5587db25ff873b5b8d64ce6ace1dd58ccd3ca8da823550cc3dcd2cb

Observation 2cd34242-d751-4332-898a-2c98233d2a8f · outbound

This paper cites Exploring the limits of trans- fer learning with a unified text-to-text transformer.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation Exploring the limits of trans- fer learning with a unified text-to-text transformer

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:02:57.331015Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T00:02:56.820105Z digest=sha256:8f8161f5917d9ed453458f9e3bd9d89d3801df4623e8641c06da48a2fe98a9d6

Observation 08b0fba2-6043-4402-a039-9f68a2db9b56 · outbound

This paper cites Sam 2: Segment anything in images and videos, 2024.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation Sam 2: Segment anything in images and videos, 2024

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:02:57.322060Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T00:02:56.822756Z digest=sha256:afe28b5ebaecf0bedbf01d58a6fc9c884e8df711fd8e48806a702e0412affa15

Observation 596111a1-b820-458a-b9db-c050d5ebb640 · outbound

This paper cites High- resolution image synthesis with latent diffusion mod- els.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation High- resolution image synthesis with latent diffusion mod- els

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-08T00:02:56.825250Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:02:56.825250Z digest=sha256:d74ed111b03b9e87b6d1702bb74fa0b9f8ffbea3d99ba6f5a71a6e2bf9b82f05

Observation 47fb0fe8-d5bc-4b9a-82d1-a59451c815b3 · outbound

This paper cites Ob- jects365: A large-scale, high-quality dataset for ob- ject detection.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation Ob- jects365: A large-scale, high-quality dataset for ob- ject detection

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:02:57.308114Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T00:02:56.827792Z digest=sha256:7d807dc2bf4a0d97f8a9b10997f3a94c27f43b77f51018ff894c58c1b055b3d8

Observation c7a5896a-e3f6-48bd-b63d-c88b73172609 · outbound

This paper cites Motion-i2v: Consistent and controllable image-to- video generation with explicit motion modeling.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation Motion-i2v: Consistent and controllable image-to- video generation with explicit motion modeling

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:02:57.299396Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T00:02:56.830386Z digest=sha256:7aeca750fa6b31e29676ffd5213f0fdf03fa8451f976d45dba2a3c6d2d2a84e6

Observation 629b0e59-32bc-438f-b424-cab34ff2d2b4 · outbound

This paper cites Free-Form Motion Control: Controlling the 6D Poses of Camera and Objects in Video Generation.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation Free-Form Motion Control: Controlling the 6D Poses of Camera and Objects in Video Generation

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-08T00:02:56.833092Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:02:56.833092Z digest=sha256:f0fa3e3e26f4bdcb83b0673d102ae8e430ed766e25b05fed4513173d60e181c6

Observation be80ea5b-41b7-4381-b6e5-066d0c5671bf · outbound

This paper cites Denoising Diffusion Implicit Models.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation Denoising Diffusion Implicit Models

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-08T00:02:56.838753Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:02:56.838753Z digest=sha256:7f4feb83d846fa3fafd4919f23c34d12dc1fabb4be2c3e0bf18828e0a631bd4f

Observation 3c7d1b30-4766-4a8f-af2f-2e1475961f0c · outbound

This paper cites Towards Accurate Generative Models of Video: A New Metric & Challenges.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation Towards Accurate Generative Models of Video: A New Metric & Challenges

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-08T00:02:56.841183Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:02:56.841183Z digest=sha256:e0cc8c508caac475484bd3e1c91831666c99e6bae0bdc13d0540f9953c589dc8

Observation 27f2f074-a029-4933-b0bd-77f99351cbcd · outbound

This paper cites EasyControl: Transfer ControlNet to Video Diffusion for Controllable Generation and Interpolation.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation EasyControl: Transfer ControlNet to Video Diffusion for Controllable Generation and Interpolation

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-08T00:02:56.844613Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:02:56.844613Z digest=sha256:8bc8193e1824bb38af28eeb69277b2aaded6c6a5e560c40366e90d6adb3a5638

Observation 56794fc1-ecbc-48cd-8ef2-f0693a1f0e71 · outbound

This paper cites Boxi- mator: Generating rich and controllable motions for video synthesis, 2024.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation Boxi- mator: Generating rich and controllable motions for video synthesis, 2024

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:02:57.289950Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T00:02:56.848055Z digest=sha256:326d47f9516e6cf73495eadc4db5beb39b68238d43aae366e8b9d1f0b7ba2d6b

Observation caaf6c67-f80b-4cf3-adf8-7447a10f3076 · outbound

This paper cites VideoComposer: Compo- sitional Video Synthesis with Motion Controllability,.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation VideoComposer: Compo- sitional Video Synthesis with Motion Controllability,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:02:57.280861Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T00:02:56.851016Z digest=sha256:3c43e0971a5f47b5c06d2436b707662ed97de4a6525efaa57c734e2a8880144b

Observation f987ab68-b443-46a8-98ff-ec08f495293c · outbound

This paper cites MotionCtrl: A Unified and Flexible Motion Controller for Video Generation.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation MotionCtrl: A Unified and Flexible Motion Controller for Video Generation

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-08T00:02:56.858520Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:02:56.858520Z digest=sha256:66391119194b675fbf155e25f2de4a8289305adbba0782aca6c3d3f3645be1f4

Observation 5990c761-e728-4072-988f-be861396dd37 · outbound

This paper cites Spa- tialtracker: Tracking any 2d pixels in 3d space.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation Spa- tialtracker: Tracking any 2d pixels in 3d space

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:02:57.271568Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T00:02:56.861788Z digest=sha256:82aeb32d920cc03e5ee8773738f5a8ea8190ec1fcd34b0ee8ce9f832e5a21e0d

Observation e495d943-efd2-4664-b720-b7d83fc5e850 · outbound

This paper cites Motioncanvas: Cinematic shot design with controllable image-to-video generation,.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation Motioncanvas: Cinematic shot design with controllable image-to-video generation,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:02:57.262175Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T00:02:56.865009Z digest=sha256:82dd70e3fe7c67eb738f851652d73e284958c5f6debeadc47b39220acffbc6b5

Observation 10ff7950-a9b6-44ad-b0bf-a3520130ebe3 · outbound

This paper cites Qwen2 Technical Report.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation Qwen2 Technical Report

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-08T00:02:56.868225Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:02:56.868225Z digest=sha256:4f07ef11b036892486dbc3945aafbefab08a08bced9c2f46f33cb688257f1c40

Observation 2099e773-d86d-4d89-81ed-33d389445988 · outbound

This paper cites Depth Anything V2.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation Depth Anything V2

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-08T00:02:56.871655Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:02:56.871655Z digest=sha256:ed4a69e5bb5a0874052ec399a2d1a3adf5344ab618dbaf475dd8d3ccc4f888ad

Observation dd9de1db-cd0b-40d9-83e3-618da0ac978f · outbound

This paper cites Direct-a-video: Customized video generation with user-directed camera movement and object mo- tion.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation Direct-a-video: Customized video generation with user-directed camera movement and object mo- tion

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:02:57.253247Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T00:02:56.874885Z digest=sha256:6aba95cc9b0a8eed1f00c69974c8bfaa168ac1e2b70ee48d5e2c32b648c14003

Observation 21de9dea-0a20-4a62-bd2e-0f733902e218 · outbound

This paper cites Dragnuwa: Fine-grained control in video generation by integrat- ing text, image, and trajectory, 2023.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation Dragnuwa: Fine-grained control in video generation by integrat- ing text, image, and trajectory, 2023

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:02:57.245032Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T00:02:56.877873Z digest=sha256:57a843d819e64a73983961c746aac256c231fcf06477a32c3bf375aa63dff0ca

Observation 2cf0157e-579d-4b68-ad26-7ff3b2aa69e0 · outbound

This paper cites MonST3R: A Simple Approach for Estimating Geometry in the Presence of Motion.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation MonST3R: A Simple Approach for Estimating Geometry in the Presence of Motion

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-08T00:02:56.880918Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:02:56.880918Z digest=sha256:5550ecd66a6864bb21208d5300ae04ba562e2b0b4ef0265356cf5a624d654a15

Observation d22034f3-50dc-4118-8204-d60ebeb7d1b6 · outbound

This paper cites Adding conditional control to text-to-image diffusion models, 2023.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation Adding conditional control to text-to-image diffusion models, 2023

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:02:57.236915Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T00:02:56.884172Z digest=sha256:b6716c9bf9e25ded19c80953c6c94c3e5253e2370a4f030ae083be9076ad0807

Observation 53a765ae-f14a-434c-9360-07a55f6f0aad · outbound

This paper cites Tora: Trajectory-oriented Diffusion Transformer for Video Generation.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation Tora: Trajectory-oriented Diffusion Transformer for Video Generation

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-08T00:02:56.887287Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:02:56.887287Z digest=sha256:955b62ea8040204262e2ca5a9c1859eab99e7d5452d6d423aa1853a5ec251033

Observation dece0ae7-5638-4861-a651-c91dc6052124 · outbound

This paper cites Stereo magnifica- tion: Learning view synthesis using multiplane im- ages, 2018.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation Stereo magnifica- tion: Learning view synthesis using multiplane im- ages, 2018

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:02:57.227739Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-08T00:02:56.890698Z digest=sha256:739210435e3c656c1e5d391bee69d79cb9a853774f542494647e1a9af1541657

Observation 96e55f9f-acd4-42f0-ae51-d28808703bc4 · outbound

This paper cites VideoComposer: Compositional Video Synthesis with Motion Controllability.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation VideoComposer: Compositional Video Synthesis with Motion Controllability

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-08T00:02:56.854273Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:02:56.854273Z digest=sha256:f387113c2b28f0ba0dbe9cf4c313626c89be9f0c371a7ff889b54473d87eba95

Pith citing papers

Observation dd2f96f5-700b-4d64-8bd3-9637ec7baba7 · inbound

UNIC: Unified In-Context Video Editing cites this paper.

UNIC: Unified In-Context Video Editing CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T10:51:43.018033Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:51:43.018033Z digest=sha256:c17a5daf115fa34151b8d4be44b16f8b84b66ae1a4d77d67962970772e4af404

Observation a6a5c587-8e23-4c81-b068-a1ab62738937 · inbound

CoT-lized Diffusion: Let's Reinforce T2I Generation Step-by-step cites this paper.

CoT-lized Diffusion: Let's Reinforce T2I Generation Step-by-step CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T19:55:03.552423Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:55:03.552423Z digest=sha256:1537d203464f7048a8f3a2fa37b751d401d397b2d0038a20977a24e774c64227

Observation cb7e9486-d713-4478-a10c-2d229f39f468 · inbound

T2VWorldBench: A Benchmark for Evaluating World Knowledge in Text-to-Video Generation cites this paper.

T2VWorldBench: A Benchmark for Evaluating World Knowledge in Text-to-Video Generation CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T14:43:28.546293Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:43:28.546293Z digest=sha256:a229a58a12a5abeb3c698d4d7d4b70cef30149f91e01c303a5db4f05da0efa47

Observation 72f60e5b-bac9-44d6-ad60-ac90ac3e9e29 · inbound

ToonComposer: Streamlining Cartoon Production with Generative Post-Keyframing cites this paper.

ToonComposer: Streamlining Cartoon Production with Generative Post-Keyframing CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-05T20:17:30.370400Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:17:30.370400Z digest=sha256:ef6907c7705388091dcc7287ab441bcb62084b8b40a004ba9e489dd6d5e730cf

Observation 6f1e3eeb-821f-4326-bd67-eb2f638c27d3 · inbound

Fuel Consumption in Platoons: A Literature Review cites this paper.

Fuel Consumption in Platoons: A Literature Review CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-05T20:15:52.917682Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:15:52.917682Z digest=sha256:6a060c7cedc47dbac308afb5ef7ed0522f3c1df0387aab1cd6e66156a1c518b6

Observation 7888dd77-faf9-4890-909a-f228a66bd3cf · inbound

Controllable Egocentric Video Generation via Occlusion-Aware Sparse 3D Hand Joints cites this paper.

Controllable Egocentric Video Generation via Occlusion-Aware Sparse 3D Hand Joints CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-05-15T12:20:00.527788Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T12:18:45.538658Z digest=sha256:6496686e9d5b0e420210f8661834ad508766122774a1af84d324258373973365

Observation 65d6a91f-024b-4cf5-9837-058df07a7e76 · inbound

VERTIGO: Visual Preference Optimization for Cinematic Camera Trajectory Generation cites this paper.

VERTIGO: Visual Preference Optimization for Cinematic Camera Trajectory Generation CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-05-13T21:18:17.245847Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-13T21:14:42.021240Z digest=sha256:4550efdcb77c1997d429525ffe4e51d59f01300cf913df5024ebdc2426f7861d

Observation 71fec4b0-8aba-4668-8d75-7adc083cd330 · inbound

Evolution of Video Generative Foundations cites this paper.

Evolution of Video Generative Foundations CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation

Reference 225

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T00:05:51.747555Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T18:41:38.616611Z digest=sha256:e5e61df6d99bd3e22b635634d099125bc78cb26f306431c9b75cde91b496f5e8