Pith. sign in

Paper Citation Record · LEDGER

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation

As of 17 August 2026, this Paper Citation Record lists 50 of 50 outbound references and 8 inbound Pith citation observations for arXiv:2502.08639.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.08639 v1

Coverage vector

measured 50 of 50 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-08T00:02:56.890698Z

measured 58 of 58 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 8 of 8 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T10:51:43.018033Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-15T12:20:00.524134Z

Reference resolution

50 of 50 outbound references displayed

  • verified exact0
  • verified fuzzy21
  • unresolved29
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 3d7055a8-32bc-4e2b-97e1-e79d484629d5 · outbound

This paper cites Loosecontrol: Lifting controlnet for generalized depth conditioning.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation Loosecontrol: Lifting controlnet for generalized depth conditioning

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:02:57.415015Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-08T00:02:56.729320Z digest=sha256:11b6530e8ddec0adc9c76d97703c39f442a3b0d983e2fb7fabe942e960e560e2

Observation 34234d4e-dc0f-4f30-8bb2-6e4825a8fa9a · outbound

This paper cites GS-DiT: Advancing Video Generation with Pseudo 4D Gaussian Fields through Efficient Dense 3D Point Tracking.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation GS-DiT: Advancing Video Generation with Pseudo 4D Gaussian Fields through Efficient Dense 3D Point Tracking

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-08T00:02:56.732859Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:02:56.732859Z digest=sha256:0c3a8c3114fea267bf7bda3a76fd754f2947d2f481d3384d72e664c0f6da141f

Observation 3e1da12f-295f-4186-8634-1b0d399dfa8e · outbound

This paper cites Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-08T00:02:56.736420Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:02:56.736420Z digest=sha256:a42b5d4447f6099f7ac91e9b507efff4ac82887d73eddd1f500b737c19b46b2c

Observation 2161f796-4dea-4823-92ef-0c07279a7564 · outbound

This paper cites VideoCrafter1: Open Diffusion Models for High-Quality Video Generation.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation VideoCrafter1: Open Diffusion Models for High-Quality Video Generation

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-08T00:02:56.739564Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:02:56.739564Z digest=sha256:60826ada937b58ed56fa9d8dce6cb637042e6679c8993a8867b6b8f8bc0560be

Observation 91ce4814-df97-4855-aec5-2a96e9f17145 · outbound

This paper cites PixArt-$\alpha$: Fast Training of Diffusion Transformer for Photorealistic Text-to-Image Synthesis.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation PixArt-$\alpha$: Fast Training of Diffusion Transformer for Photorealistic Text-to-Image Synthesis

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-08T00:02:56.742731Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:02:56.742731Z digest=sha256:5d760ea9d48600631441682608bf6fa7eb5e281092d807daa28c4b91c80a2587

Observation 866547a4-cdd9-4e22-a574-04109815eb0b · outbound

This paper cites Control-A-Video: Controllable Text-to-Video Diffusion Models with Motion Prior and Reward Feedback Learning.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation Control-A-Video: Controllable Text-to-Video Diffusion Models with Motion Prior and Reward Feedback Learning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-08T00:02:56.746313Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:02:56.746313Z digest=sha256:6309051752e9d980b17e45a274fa9d4bfaa7c21046d1acb6101021d3c3ecae01

Observation 7a9b1d57-9407-4db1-a43a-1816bb7e8a48 · outbound

This paper cites Perception-as-Control: Fine-grained Controllable Image Animation with 3D-aware Motion Representation.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation Perception-as-Control: Fine-grained Controllable Image Animation with 3D-aware Motion Representation

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-08T00:02:56.750090Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:02:56.750090Z digest=sha256:9ab8516b2edfc4617bef378fdcf06c1630f67c41df8ffa74855277aeaac49b34

Observation 077f6917-1086-4fdd-8eca-db4e36324122 · outbound

This paper cites Patch n’pack: Navit, a vision transformer for any aspect ratio and resolu- tion.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation Patch n’pack: Navit, a vision transformer for any aspect ratio and resolu- tion

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:02:57.405949Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-08T00:02:56.753614Z digest=sha256:81cf40f677984abbfe17c962c1da0887330294deb5ed708f91a9bc104ed62d45

Observation 19660e4c-a8c5-4a24-9f3c-c301d28c31ed · outbound

This paper cites Scaling rectified flow transformers for high-resolution image synthesis.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation Scaling rectified flow transformers for high-resolution image synthesis

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:02:57.397458Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-08T00:02:56.756898Z digest=sha256:4f93e55337d3f9a618fd376e2ace12a08ba1828b36639a04a4e15ea60ccfed39

Observation 1fa4bd7f-8e40-4784-a71e-cb116ca056e7 · outbound

This paper cites 3DTrajMaster: Mastering 3D Trajectory for Multi-Entity Motion in Video Generation.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation 3DTrajMaster: Mastering 3D Trajectory for Multi-Entity Motion in Video Generation

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-08T00:02:56.760277Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:02:56.760277Z digest=sha256:ff83e916907ae7b24f3af0526715fc9ebc5008a6b7b11c0d0ec97e2dcbe8e7cc

Observation 2fe49bbb-275f-40b7-ba3d-82fb1a129cdb · outbound

This paper cites Motion Prompting: Controlling Video Generation with Motion Trajectories.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation Motion Prompting: Controlling Video Generation with Motion Trajectories

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-08T00:02:56.763785Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:02:56.763785Z digest=sha256:53c18568fc23fdef3f4da701e8e5509a3ac5da38a8d2e0d5372dfa3aca8806ec

Observation 66444dcb-3d0e-4aa9-b310-cae0a9e0fe6c · outbound

This paper cites Diffusion as Shader: 3D-aware Video Diffusion for Versatile Video Generation Control.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation Diffusion as Shader: 3D-aware Video Diffusion for Versatile Video Generation Control

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-08T00:02:56.767177Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:02:56.767177Z digest=sha256:541b5e2d0dcc539ff07b7a52e61778d8ae1fda433dc6b1bac19b100fa24d1de4

Observation 73e66d04-f15f-4942-a413-ecec944bb694 · outbound

This paper cites Sparsectrl: Adding sparse controls to text-to-video diffusion mod- els, 2023.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation Sparsectrl: Adding sparse controls to text-to-video diffusion mod- els, 2023

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:02:57.389476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-08T00:02:56.770594Z digest=sha256:f7a69eb71b0a44885d1978d3b827bb57e0720ba00add6d6a7e1f817e8ff23c36

Observation 883d6db8-5aa2-4f10-930d-442a70661dc1 · outbound

This paper cites AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-08T00:02:56.773765Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:02:56.773765Z digest=sha256:da56c30bdb8349db95476f4c859641ca6cc57bb4452c0a53ddecf7975f831bb9

Observation 5bcde63a-fca3-448f-a866-ba5cb2699e87 · outbound

This paper cites Cameractrl: Enabling camera control for text-to-video generation,.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation Cameractrl: Enabling camera control for text-to-video generation,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:02:57.381179Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-08T00:02:56.777309Z digest=sha256:997a44f9d04251a60ee39c83996cd68828aa98cce14a55218d8755852ce19a31

Observation ee1a13c5-2739-4bf5-95d8-22c4c6bb4a1b · outbound

This paper cites Classifier-Free Diffusion Guidance.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation Classifier-Free Diffusion Guidance

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-08T00:02:56.780647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:02:56.780647Z digest=sha256:9e36698ccc7fc61120407f4052f95c18fb704b2a07720ddfc6a960481a8ef6dd

Observation 4f8aba04-1035-46bf-ab98-61c1ec602ea0 · outbound

This paper cites Denoising diffusion probabilistic models.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation Denoising diffusion probabilistic models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-08T00:02:56.784510Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:02:56.784510Z digest=sha256:fa2fe56287b90ceab4eda2fa4e8cd628aa8824b4f7b67e80dd5e102802e1930c

Observation dde0967a-ce82-4fed-8da0-ae40f691a31e · outbound

This paper cites Animate anyone: Consistent and controllable image-to-video synthesis for character animation.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation Animate anyone: Consistent and controllable image-to-video synthesis for character animation

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:02:57.367946Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-08T00:02:56.787678Z digest=sha256:8e4b7090c30d1d597ed68083057998c144f47f9d1339703fe43ba4c0f605d97f

Observation 1b6582cf-f5a1-4d2b-8d8d-130447fbddf2 · outbound

This paper cites Auto-Encoding Variational Bayes.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation Auto-Encoding Variational Bayes

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-08T00:02:56.790984Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:02:56.790984Z digest=sha256:2091ac6dabb136f38f6ab7ca6c6e081e8ff052c1bcea1b2f6caf92b51d5e34b6

Observation 546c9a69-90eb-430d-befb-d727c70ffdf8 · outbound

This paper cites Adam: A Method for Stochastic Optimization.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation Adam: A Method for Stochastic Optimization

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-08T00:02:56.794985Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:02:56.794985Z digest=sha256:da9fcebfe64fcbdaef8da72237c3487931360adc7878cc3b490ed38a302ac3a7

Observation 8c1910f7-b06e-408e-9fdf-e35a5de8835d · outbound

This paper cites Onlyflow: Optical flow based motion condition- ing for video diffusion models.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation Onlyflow: Optical flow based motion condition- ing for video diffusion models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-08T00:02:56.798165Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:02:56.798165Z digest=sha256:3c64cc48dcf8fa71e490ab01c07cf5e3a0c370873c4a9e0c010dcacf2f16429b

Observation de8b809f-963a-4304-8a84-6a47247d9194 · outbound

This paper cites Microsoft coco: Com- mon objects in context.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation Microsoft coco: Com- mon objects in context

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:02:57.358490Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-08T00:02:56.801473Z digest=sha256:d2454bb4ce8877a17100deec94aeff76cf55ad09917561b27881f1a579240e30

Observation 0f6cdc96-a0fd-4c86-a8ce-7a1dece3959d · outbound

This paper cites Flow Matching for Generative Modeling.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation Flow Matching for Generative Modeling

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-08T00:02:56.804612Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:02:56.804612Z digest=sha256:e7c817f6acf6f3c0ece4c177c5399367e61f59651087e50a6ac021b54cabad6a

Observation 217560b7-e689-4c43-b3e9-b6483542b5d3 · outbound

This paper cites Grounding DINO: Marrying DINO with Grounded Pre-Training for Open-Set Object Detection.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation Grounding DINO: Marrying DINO with Grounded Pre-Training for Open-Set Object Detection

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-08T00:02:56.807850Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:02:56.807850Z digest=sha256:845f53586e50e7e7d0232d63bf8efa052842c4a3efd6d0889aeec6bf1efabe21

Observation 06f3bfe0-1217-4eed-a73f-36241564ad16 · outbound

This paper cites Are GANs Created Equal? A Large-Scale Study.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation Are GANs Created Equal? A Large-Scale Study

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-08T00:02:56.811116Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:02:56.811116Z digest=sha256:d893802a7ada6764e6ac91951f12a6fe52a7b31177b6f7f11b41cfeb3839552b

Observation ae8b787e-d25b-470e-9254-acc6f0aa3025 · outbound

This paper cites T2i- adapter: Learning adapters to dig out more control- lable ability for text-to-image diffusion models.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation T2i- adapter: Learning adapters to dig out more control- lable ability for text-to-image diffusion models

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:02:57.349395Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-08T00:02:56.814377Z digest=sha256:03ade2108c20b5490889c2533efc49a102371a90c0c2eaadab4df7a5074eb349

Observation 173cba0a-faa2-4d25-8311-b572145c1921 · outbound

This paper cites Scalable diffu- sion models with transformers.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation Scalable diffu- sion models with transformers

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:02:57.340055Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-08T00:02:56.817478Z digest=sha256:c13545824cee2308e6713a8b70904888309daae388838f069a7a1a46d1cbf034

Observation 2cd34242-d751-4332-898a-2c98233d2a8f · outbound

This paper cites Exploring the limits of trans- fer learning with a unified text-to-text transformer.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation Exploring the limits of trans- fer learning with a unified text-to-text transformer

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:02:57.331015Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-08T00:02:56.820105Z digest=sha256:6c4151bca0d6ea0187d929be9e19cfb4bcac1403c67460c416824a8915a41fc7

Observation 08b0fba2-6043-4402-a039-9f68a2db9b56 · outbound

This paper cites Sam 2: Segment anything in images and videos, 2024.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation Sam 2: Segment anything in images and videos, 2024

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:02:57.322060Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-08T00:02:56.822756Z digest=sha256:50a94b764632b66556ac39e3a5a808222329f81150e82381eff9f8b2f31045d5

Observation 596111a1-b820-458a-b9db-c050d5ebb640 · outbound

This paper cites High- resolution image synthesis with latent diffusion mod- els.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation High- resolution image synthesis with latent diffusion mod- els

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-08T00:02:56.825250Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:02:56.825250Z digest=sha256:eaf373a45eb1a7a0d549030bc14d0de88a0b1009d962abd4a7c69ac85c284597

Observation 47fb0fe8-d5bc-4b9a-82d1-a59451c815b3 · outbound

This paper cites Ob- jects365: A large-scale, high-quality dataset for ob- ject detection.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation Ob- jects365: A large-scale, high-quality dataset for ob- ject detection

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:02:57.308114Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-08T00:02:56.827792Z digest=sha256:30392deceaca9d63b5145f022999ba48a937369d08107a26adec00edf55766d7

Observation c7a5896a-e3f6-48bd-b63d-c88b73172609 · outbound

This paper cites Motion-i2v: Consistent and controllable image-to- video generation with explicit motion modeling.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation Motion-i2v: Consistent and controllable image-to- video generation with explicit motion modeling

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:02:57.299396Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-08T00:02:56.830386Z digest=sha256:3570a27143e39ba12075ca633394bdef8496c913b9f76d487b70d7d57dbcbf77

Observation 629b0e59-32bc-438f-b424-cab34ff2d2b4 · outbound

This paper cites Free-Form Motion Control: Controlling the 6D Poses of Camera and Objects in Video Generation.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation Free-Form Motion Control: Controlling the 6D Poses of Camera and Objects in Video Generation

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-08T00:02:56.833092Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:02:56.833092Z digest=sha256:7b55ab09ba2ec87353b03a5a8bdf531b7abb7b41f87350830d99805ed073928b

Observation be80ea5b-41b7-4381-b6e5-066d0c5671bf · outbound

This paper cites Denoising Diffusion Implicit Models.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation Denoising Diffusion Implicit Models

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-08T00:02:56.838753Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:02:56.838753Z digest=sha256:2688eb0f5e59b01e6e664f0ddd87ad3725f5fe7d730937e231933949d8dfdbdd

Observation 3c7d1b30-4766-4a8f-af2f-2e1475961f0c · outbound

This paper cites Towards Accurate Generative Models of Video: A New Metric & Challenges.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation Towards Accurate Generative Models of Video: A New Metric & Challenges

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-08T00:02:56.841183Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:02:56.841183Z digest=sha256:1ea2476ef6ad361bffa9302dbe477c2a69e4f834c2084ddbdc4545aca54f0160

Observation 27f2f074-a029-4933-b0bd-77f99351cbcd · outbound

This paper cites EasyControl: Transfer ControlNet to Video Diffusion for Controllable Generation and Interpolation.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation EasyControl: Transfer ControlNet to Video Diffusion for Controllable Generation and Interpolation

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-08T00:02:56.844613Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:02:56.844613Z digest=sha256:67b76f67e13310a3670665cb97513d08b8fd6d4658c451150eb51c98b81bbf71

Observation 56794fc1-ecbc-48cd-8ef2-f0693a1f0e71 · outbound

This paper cites Boxi- mator: Generating rich and controllable motions for video synthesis, 2024.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation Boxi- mator: Generating rich and controllable motions for video synthesis, 2024

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:02:57.289950Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-08T00:02:56.848055Z digest=sha256:1447221f972bbf100942fe38b8d8636dedfa199d4b20eb48ab7e5c7368ac3af7

Observation caaf6c67-f80b-4cf3-adf8-7447a10f3076 · outbound

This paper cites VideoComposer: Compo- sitional Video Synthesis with Motion Controllability,.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation VideoComposer: Compo- sitional Video Synthesis with Motion Controllability,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:02:57.280861Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-08T00:02:56.851016Z digest=sha256:34e5dbd18c09b9bf90a61c591dcd618bdaf9490b1acd37e948d191c61cfb1306

Observation f987ab68-b443-46a8-98ff-ec08f495293c · outbound

This paper cites MotionCtrl: A Unified and Flexible Motion Controller for Video Generation.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation MotionCtrl: A Unified and Flexible Motion Controller for Video Generation

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-08T00:02:56.858520Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:02:56.858520Z digest=sha256:502028efc0190d6cf83fa0a625d804c7268be68ba4ec513ca157541d62b013a3

Observation 5990c761-e728-4072-988f-be861396dd37 · outbound

This paper cites Spa- tialtracker: Tracking any 2d pixels in 3d space.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation Spa- tialtracker: Tracking any 2d pixels in 3d space

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:02:57.271568Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-08T00:02:56.861788Z digest=sha256:5e26b1e25e65f7a2c6b42cf22dd975df1d0e199ecaba1dffc13368ca8215810a

Observation e495d943-efd2-4664-b720-b7d83fc5e850 · outbound

This paper cites Motioncanvas: Cinematic shot design with controllable image-to-video generation,.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation Motioncanvas: Cinematic shot design with controllable image-to-video generation,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:02:57.262175Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-08T00:02:56.865009Z digest=sha256:e8e7a6d2199f20ba47bde0050723a291777be4b8db4f3b920c1cb3fdf4044b59

Observation 10ff7950-a9b6-44ad-b0bf-a3520130ebe3 · outbound

This paper cites Qwen2 Technical Report.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation Qwen2 Technical Report

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-08T00:02:56.868225Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:02:56.868225Z digest=sha256:a4cbbce3bc3bc4dfb2824933ca2addb40a84172426d179d7dd17a614cff09486

Observation 2099e773-d86d-4d89-81ed-33d389445988 · outbound

This paper cites Depth Anything V2.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation Depth Anything V2

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-08T00:02:56.871655Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:02:56.871655Z digest=sha256:02d7dd1058f5b5634c7395ecfe6c730d89cde15ec150a11722b00185dda92ceb

Observation dd9de1db-cd0b-40d9-83e3-618da0ac978f · outbound

This paper cites Direct-a-video: Customized video generation with user-directed camera movement and object mo- tion.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation Direct-a-video: Customized video generation with user-directed camera movement and object mo- tion

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:02:57.253247Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-08T00:02:56.874885Z digest=sha256:ff7fc23c085dbe90b221fe2104b809e46ab8fcc530a1061938c56ca36b92a069

Observation 21de9dea-0a20-4a62-bd2e-0f733902e218 · outbound

This paper cites Dragnuwa: Fine-grained control in video generation by integrat- ing text, image, and trajectory, 2023.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation Dragnuwa: Fine-grained control in video generation by integrat- ing text, image, and trajectory, 2023

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:02:57.245032Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-08T00:02:56.877873Z digest=sha256:903cf6952007485002208e899c7fada1a1055985c4998aa0644498b43dbdd346

Observation 2cf0157e-579d-4b68-ad26-7ff3b2aa69e0 · outbound

This paper cites MonST3R: A Simple Approach for Estimating Geometry in the Presence of Motion.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation MonST3R: A Simple Approach for Estimating Geometry in the Presence of Motion

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-08T00:02:56.880918Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:02:56.880918Z digest=sha256:65386018c2a19b0bdeeaacbddee8d328e699a05980f14fcaef39ee1e22ec5904

Observation d22034f3-50dc-4118-8204-d60ebeb7d1b6 · outbound

This paper cites Adding conditional control to text-to-image diffusion models, 2023.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation Adding conditional control to text-to-image diffusion models, 2023

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:02:57.236915Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-08T00:02:56.884172Z digest=sha256:6c56f4fb3782fc152d962c5ce2a0e68df793a599df373fe7c1f767dabb056d0a

Observation 53a765ae-f14a-434c-9360-07a55f6f0aad · outbound

This paper cites Tora: Trajectory-oriented Diffusion Transformer for Video Generation.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation Tora: Trajectory-oriented Diffusion Transformer for Video Generation

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-08T00:02:56.887287Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:02:56.887287Z digest=sha256:ce1f3177336704cdd333f8d9846c7373a852b8007fcea5f33da9404a4e351c55

Observation dece0ae7-5638-4861-a651-c91dc6052124 · outbound

This paper cites Stereo magnifica- tion: Learning view synthesis using multiplane im- ages, 2018.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation Stereo magnifica- tion: Learning view synthesis using multiplane im- ages, 2018

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T00:02:57.227739Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-08T00:02:56.890698Z digest=sha256:cbd94a343a98bcebde63e0b0a91f592b22990779fe82f7a35d58f60a00d575dc

Observation 96e55f9f-acd4-42f0-ae51-d28808703bc4 · outbound

This paper cites VideoComposer: Compositional Video Synthesis with Motion Controllability.

CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation VideoComposer: Compositional Video Synthesis with Motion Controllability

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-08T00:02:56.854273Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:02:56.854273Z digest=sha256:c5a30416a3b8ab5d3fac02dc63b5573a43a52c8178d128cabae40fb88d441c77

Pith citing papers

Observation dd2f96f5-700b-4d64-8bd3-9637ec7baba7 · inbound

UNIC: Unified In-Context Video Editing cites this paper.

UNIC: Unified In-Context Video Editing CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T10:51:43.018033Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:51:43.018033Z digest=sha256:ef1ec738c6c9fa897c11d3815004ef17dc95f1bcb2ebff73d3b249efe54b1a86

Observation a6a5c587-8e23-4c81-b068-a1ab62738937 · inbound

CoT-lized Diffusion: Let's Reinforce T2I Generation Step-by-step cites this paper.

CoT-lized Diffusion: Let's Reinforce T2I Generation Step-by-step CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T19:55:03.552423Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:55:03.552423Z digest=sha256:4a34bab6ef629f4961e0b8d1251c6c4106a642634b68cccf6653586af620345d

Observation cb7e9486-d713-4478-a10c-2d229f39f468 · inbound

T2VWorldBench: A Benchmark for Evaluating World Knowledge in Text-to-Video Generation cites this paper.

T2VWorldBench: A Benchmark for Evaluating World Knowledge in Text-to-Video Generation CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T14:43:28.546293Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:43:28.546293Z digest=sha256:76f1584d56083919bcedee276d6380e4c038b31895fe9529eecc08c926cd400e

Observation 72f60e5b-bac9-44d6-ad60-ac90ac3e9e29 · inbound

ToonComposer: Streamlining Cartoon Production with Generative Post-Keyframing cites this paper.

ToonComposer: Streamlining Cartoon Production with Generative Post-Keyframing CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-05T20:17:30.370400Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:17:30.370400Z digest=sha256:a1b1ba07ce8ab1b2f371cc67e2e7d2c94c54347349238d2c8d0c25386318ef01

Observation 6f1e3eeb-821f-4326-bd67-eb2f638c27d3 · inbound

Fuel Consumption in Platoons: A Literature Review cites this paper.

Fuel Consumption in Platoons: A Literature Review CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-05T20:15:52.917682Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:15:52.917682Z digest=sha256:92928688631c033e7fa986f7049f43468ed256737f87dffcbcaf6a049b4385cd

Observation 7888dd77-faf9-4890-909a-f228a66bd3cf · inbound

Controllable Egocentric Video Generation via Occlusion-Aware Sparse 3D Hand Joints cites this paper.

Controllable Egocentric Video Generation via Occlusion-Aware Sparse 3D Hand Joints CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-05-15T12:20:00.527788Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-15T12:18:45.538658Z digest=sha256:ab575c009f191c499b3b0b9a62e58a0fd9249264ebf7cb4469c84bf376801533

Observation 65d6a91f-024b-4cf5-9837-058df07a7e76 · inbound

VERTIGO: Visual Preference Optimization for Cinematic Camera Trajectory Generation cites this paper.

VERTIGO: Visual Preference Optimization for Cinematic Camera Trajectory Generation CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-05-13T21:18:17.245847Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-13T21:14:42.021240Z digest=sha256:5f5a2fd3bf1f3959a6dff2c85131156a9320b0c0c4354748620f6c4fcc705e6e

Observation 71fec4b0-8aba-4668-8d75-7adc083cd330 · inbound

Evolution of Video Generative Foundations cites this paper.

Evolution of Video Generative Foundations CineMaster: A 3D-Aware and Controllable Framework for Cinematic Text-to-Video Generation

Reference 225

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T00:05:51.747555Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-10T18:41:38.616611Z digest=sha256:f50ef0e05b945ae28397e0276b1bb4615ac6953d56197703908f3d59b02b6352