Pith. sign in

Paper Citation Record · LEDGER

Tune-A-Video: One-Shot Tuning of Image Diffusion Models for Text-to-Video Generation

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 11 inbound Pith citation observations for arXiv:2212.11565.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2212.11565 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 11 of 11 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 11 of 11 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T13:21:50.643848Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T12:16:56.715496Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation e4eabf86-1565-4c35-b383-74844550e4c2 · inbound

InternVid: A Large-scale Video-Text Dataset for Multimodal Understanding and Generation cites this paper.

InternVid: A Large-scale Video-Text Dataset for Multimodal Understanding and Generation Tune-A-Video: One-Shot Tuning of Image Diffusion Models for Text-to-Video Generation

Reference 80

Resolution
verified exact
arxiv_id, observed 2026-05-15T06:30:22.735873Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T06:30:22.431538Z digest=sha256:7c7fb8b781e9eaf607ab31ceae1b00846765689badf55d6810fdc49ee4d131df

Observation 34030fb2-8381-4d48-be12-af5bcb12e1f6 · inbound

TokenFlow: Consistent Diffusion Features for Consistent Video Editing cites this paper.

TokenFlow: Consistent Diffusion Features for Consistent Video Editing Tune-A-Video: One-Shot Tuning of Image Diffusion Models for Text-to-Video Generation

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-17T20:17:47.102071Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-17T20:17:47.001786Z digest=sha256:7a5487f496e8287bd4c488f51ea2cb9b7781c7d52da6314627dfaa461ccf99ab

Observation f8d86f92-1bef-444b-acd0-7f7ea73c1cc6 · inbound

PixArt-$\alpha$: Fast Training of Diffusion Transformer for Photorealistic Text-to-Image Synthesis cites this paper.

PixArt-$\alpha$: Fast Training of Diffusion Transformer for Photorealistic Text-to-Image Synthesis Tune-A-Video: One-Shot Tuning of Image Diffusion Models for Text-to-Video Generation

Reference 83

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T20:38:52.844051Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-12T20:38:52.749119Z digest=sha256:25924f889d3dcb0ab0df0e05c0260bc0858257ac7763e93a1c27e258dbdfe3f3

Observation 0d484bcb-4291-4aaa-b026-56e8c6253311 · inbound

Learning Interactive Real-World Simulators cites this paper.

Learning Interactive Real-World Simulators Tune-A-Video: One-Shot Tuning of Image Diffusion Models for Text-to-Video Generation

Reference 196

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T02:15:18.705669Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-16T02:15:18.265190Z digest=sha256:d17848ff728191f538e0be70f50c9b985b37a4ea77a2a1287209652202a2f653

Observation dffbbbad-d74a-4db3-86d1-954bd0524924 · inbound

SOWing Information: Cultivating Contextual Coherence with MLLMs in Image Generation cites this paper.

SOWing Information: Cultivating Contextual Coherence with MLLMs in Image Generation Tune-A-Video: One-Shot Tuning of Image Diffusion Models for Text-to-Video Generation

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-23T08:37:44.774230Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-23T08:37:13.529654Z digest=sha256:5728bd1c1000ad51043ea3dee36cdcd8d120e1a15838ee8417759c89fbdfc7ab

Observation 838b2a32-a27b-4368-9569-ca73060f2595 · inbound

Articulate That Object Part (ATOP): 3D Part Articulation via Text and Motion Personalization cites this paper.

Articulate That Object Part (ATOP): 3D Part Articulation via Text and Motion Personalization Tune-A-Video: One-Shot Tuning of Image Diffusion Models for Text-to-Video Generation

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-08T13:21:50.643848Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T13:21:50.643848Z digest=sha256:0a950c2798d883ade11e77e97b904f363eeb8204c4b4cb9918c61852d6e1d92a

Observation 4ced25f5-cc39-442b-ba66-144c741ccb6a · inbound

Communicative Agents for Slideshow Storytelling Video Generation based on LLMs cites this paper.

Communicative Agents for Slideshow Storytelling Video Generation based on LLMs Tune-A-Video: One-Shot Tuning of Image Diffusion Models for Text-to-Video Generation

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-05T12:44:22.209865Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:44:22.209865Z digest=sha256:95bbdde599de807bf092ce3ce27894e34e0e596679eec88b2c06a8d6b2d57a95

Observation b2235b24-3c3f-4fb8-8dee-fb3cd71b8b38 · inbound

Physics-Aware Video Instance Removal Benchmark cites this paper.

Physics-Aware Video Instance Removal Benchmark Tune-A-Video: One-Shot Tuning of Image Diffusion Models for Text-to-Video Generation

Reference 26

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T22:55:52.060467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T19:27:33.716719Z digest=sha256:5db370bde8f7a071e3c5753316dbc2396b837f68e79f3ee90b42aaecf6a700e9

Observation aa04ece2-2c25-4e9d-b45d-0cd4e41d9382 · inbound

Functionalization via Structure Completion and Motion Rectification cites this paper.

Functionalization via Structure Completion and Motion Rectification Tune-A-Video: One-Shot Tuning of Image Diffusion Models for Text-to-Video Generation

Reference 249

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T12:28:17.019205Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-20T12:25:07.157086Z digest=sha256:d74efb4fc6b54b3096d62c98dfc5412da23b452b2c338d188b635b43560bc862

Observation f08f0307-3916-4e2a-9568-bf6190d05d83 · inbound

V2V-Bench: A Comprehensive Benchmark for Video-to-Video Generation Evaluation cites this paper.

V2V-Bench: A Comprehensive Benchmark for Video-to-Video Generation Evaluation Tune-A-Video: One-Shot Tuning of Image Diffusion Models for Text-to-Video Generation

Reference 7

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T12:16:56.716968Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-28T02:20:05.748127Z digest=sha256:1aaf3c96b70b147e03fc6a8d409ed4782b199e1b659e94bc5301090f2ca9feef

Observation c9712c0d-c52d-4c4d-bc23-c8b30a1082bd · inbound

OSVE: One Step Video Editing with One Step Diffusion Models cites this paper.

OSVE: One Step Video Editing with One Step Diffusion Models Tune-A-Video: One-Shot Tuning of Image Diffusion Models for Text-to-Video Generation

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-01T11:29:19.733828Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:29:19.733828Z digest=sha256:7582d51fba525a3bc4327e8c00a5fc0457933ce9ec1741a4125224a13520fde5