Pith. sign in

Paper Citation Record · LEDGER

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation

As of 13 August 2026, this Paper Citation Record lists 50 of 50 outbound references and 4 inbound Pith citation observations for arXiv:2506.03126.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.03126 v1

Coverage vector

measured 50 of 50 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:12:11.643982Z

measured 54 of 54 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-13T12:23:05.876881Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-25T02:40:14.528606Z

Reference resolution

50 of 50 outbound references displayed

  • verified exact1
  • verified fuzzy13
  • unresolved36
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 389f885a-2c05-4557-aa24-30917415e80d · outbound

This paper cites Frozen in time: A joint video and image encoder for end-to-end retrieval.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Frozen in time: A joint video and image encoder for end-to-end retrieval

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:07.596205Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:07.596205Z digest=sha256:0917cb8169b8bd7b241d42abf82539ec576f8c94859b54ab66f96d2df9c4590c

Observation dd058a13-c529-4a70-ab25-e2538c4a7612 · outbound

This paper cites Activitynet: A large-scale video benchmark for human activity understanding.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Activitynet: A large-scale video benchmark for human activity understanding

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:07.662318Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:07.662318Z digest=sha256:22ba31d84cbb2d4669b73c9a57eb1b464d08990dce010d0bd6a830d53606b045

Observation 8af38723-941f-4c8c-a6f2-d8a21683ed7c · outbound

This paper cites Panda-70m: Captioning 70m videos with multiple cross-modality teachers.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Panda-70m: Captioning 70m videos with multiple cross-modality teachers

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:07.730235Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:07.730235Z digest=sha256:59d7ce994b366b784f41c0c2a31903da949abe13717418ca1fa5725b25d9ddad

Observation da75d926-c224-400f-916a-84cd7e90880d · outbound

This paper cites Multi-subject Open-set Personalization in Video Generation.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Multi-subject Open-set Personalization in Video Generation

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:07.805412Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:07.805412Z digest=sha256:14daa0f7fd274af49fd43e66d263e33b5652b0a30d1222efc9e004d6aadc1aa5

Observation 7a32cccd-bf7b-4da4-9bb8-7731c437384d · outbound

This paper cites Internvl: Scaling up vision foundation models and aligning for generic visual-linguistic tasks.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Internvl: Scaling up vision foundation models and aligning for generic visual-linguistic tasks

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:07.888630Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:07.888630Z digest=sha256:3277b34c80840b50748ebc8e1caf43ab203b0e76960a1762124d98359dbd5390

Observation 2ad33bc0-f505-474e-b8ce-99f722b5d127 · outbound

This paper cites AnimeGamer: Infinite Anime Life Simulation with Next Game State Prediction.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation AnimeGamer: Infinite Anime Life Simulation with Next Game State Prediction

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-08-07T11:12:12.106906Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T11:12:07.996721Z digest=sha256:3b44f42a85d2269e5f728eb6f099b50d2b3e9fc7cf98698bdc851dc095e45006

Observation 42f72ecf-dc28-47aa-828b-08c5f6d867ff · outbound

This paper cites Gemini, 2024.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Gemini, 2024

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:12:14.189027Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T11:12:08.133677Z digest=sha256:aad31b232e8dc0f637d17af927cff9285d0218910a4041a47dc03ddd97814d08

Observation 10bbdf14-7a25-4c48-bb27-3789b2a20813 · outbound

This paper cites CINEMA: Coherent Multi-Subject Video Generation via MLLM-Based Guidance.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation CINEMA: Coherent Multi-Subject Video Generation via MLLM-Based Guidance

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:08.227600Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:08.227600Z digest=sha256:47c87998e6b2d4a496d7ede7ae6c2f3ff53ec39f392a148adefb4f9d18a5c8a1

Observation 374fa6d5-8d77-432a-8862-f6ea8f81c154 · outbound

This paper cites DreamSim: Learning New Dimensions of Human Visual Similarity using Synthetic Data.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation DreamSim: Learning New Dimensions of Human Visual Similarity using Synthetic Data

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:08.339057Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:08.339057Z digest=sha256:7e358d407630b340247bb7209ded820953eac21b0b0dd623c3f41eff541a5534

Observation 8a36178e-afd2-410e-a11e-03b8c7f5c3c3 · outbound

This paper cites TaleCrafter: Interactive Story Visualization with Multiple Characters.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation TaleCrafter: Interactive Story Visualization with Multiple Characters

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:08.485043Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:08.485043Z digest=sha256:51a208131309ed2a03eb16b332710dad084783b2c0fd159b7fb5af0c78b9b33e

Observation d64ba14d-5ff1-4bed-9806-cc8749ca3d35 · outbound

This paper cites The Llama 3 Herd of Models.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation The Llama 3 Herd of Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:08.630129Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:08.630129Z digest=sha256:285e098a48ca04355d9db67230950f4e12af64463f3a2d098f461bd50b78a319

Observation b90d077f-4e9f-4da3-a746-15ea5e9d7b81 · outbound

This paper cites Mix-of-show: Decentralized low-rank adaptation for multi-concept customization of diffusion models.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Mix-of-show: Decentralized low-rank adaptation for multi-concept customization of diffusion models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:08.729397Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:08.729397Z digest=sha256:511303d98696d7a05c4fc697cabf0d9e5bf0ea947be80ad14025108f55f3b008

Observation 9356d9de-e9d4-44bb-bcb6-f78c65570da9 · outbound

This paper cites ROICtrl: Boosting Instance Control for Visual Generation.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation ROICtrl: Boosting Instance Control for Visual Generation

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:08.846411Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:08.846411Z digest=sha256:e923d9af60cb32f8ffa49f77017a0f1895103fc9a18c715ba7101962c1bae707

Observation f573cba1-365b-462e-8c50-832797999e3e · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:08.909571Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:08.909571Z digest=sha256:c8adec6110a453a484ccae07af236101bd9d0f08b34021ff4bd46397b20d844b

Observation 0ffec6cb-a1a0-4858-b45a-43d583ac1ff2 · outbound

This paper cites Long Context Tuning for Video Generation.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Long Context Tuning for Video Generation

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:08.965970Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:08.965970Z digest=sha256:e04f73b93aea5bcb46c75f68f08fc9993539a9363fe350284c131111e7bcb7b6

Observation e051b2b5-5fe8-4eb4-944e-760793cdaaea · outbound

This paper cites AnyStory: Towards Unified Single and Multiple Subject Personalization in Text-to-Image Generation.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation AnyStory: Towards Unified Single and Multiple Subject Personalization in Text-to-Image Generation

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:09.048829Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:09.048829Z digest=sha256:1d666da85d46698b3d522ff2b058123213198e8076a078c16dcb49f6d1e41e4a

Observation 7fa80870-4be0-4fc5-a828-286d8f634c71 · outbound

This paper cites ID-Animator: Zero-Shot Identity-Preserving Human Video Generation.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation ID-Animator: Zero-Shot Identity-Preserving Human Video Generation

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:09.126710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:09.126710Z digest=sha256:561392a7e339e3583eb24cce04456f2fc3d67c38091b18a5c8d837e7dba9196f

Observation aaf0aeb4-19f8-4519-9e5d-d7174043c9cf · outbound

This paper cites CogVideo: Large-scale Pretraining for Text-to-Video Generation via Transformers.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation CogVideo: Large-scale Pretraining for Text-to-Video Generation via Transformers

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:09.189653Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:09.189653Z digest=sha256:7de33e695374d4663025b42ecb15b958f474378777ab9efba4c31ffd478f6652

Observation a40c66a4-0d8e-408a-a927-93464dff33f3 · outbound

This paper cites Owl-1: Omni World Model for Consistent Long Video Generation.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Owl-1: Omni World Model for Consistent Long Video Generation

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:09.266583Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:09.266583Z digest=sha256:b731f2e5af1c351f408ed105f4fa4e4c2e78edc322caa4fa388370e189ab9813

Observation bc92ba84-61f2-483c-bbe5-3ecf2d2c648f · outbound

This paper cites ConceptMaster: Multi-Concept Video Customization on Diffusion Transformer Models Without Test-Time Tuning.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation ConceptMaster: Multi-Concept Video Customization on Diffusion Transformer Models Without Test-Time Tuning

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:09.347154Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:09.347154Z digest=sha256:52f4460c0788ae4b19bc2597f473105d9260c3aca50b654b515019d0e5e68277

Observation ae659b54-8ae6-4d0f-8e35-7eb036357a6c · outbound

This paper cites TangoFlux: Super Fast and Faithful Text to Audio Generation with Flow Matching and Clap-Ranked Preference Optimization.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation TangoFlux: Super Fast and Faithful Text to Audio Generation with Flow Matching and Clap-Ranked Preference Optimization

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:09.446928Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:09.446928Z digest=sha256:107221465c51ed0856c07594cd78088445baa1097d6319c26fbcb9022a1a1f21

Observation 75efc4a0-733f-4862-8af8-bcd8407e95e1 · outbound

This paper cites AniSora: Exploring the Frontiers of Animation Video Generation in the Sora Era.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation AniSora: Exploring the Frontiers of Animation Video Generation in the Sora Era

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:09.494420Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:09.494420Z digest=sha256:ddc4ca0ba71f00dbba35fe37b6e21d79a0a07b9a650c3366f7c255f61d1551bd

Observation eb25d9f7-ac88-4b57-a016-a3eefcb212da · outbound

This paper cites Videobooth: Diffusion-based video generation with image prompts.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Videobooth: Diffusion-based video generation with image prompts

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:12:13.999178Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T11:12:09.571073Z digest=sha256:7f8786b913673d9ed69fc9f56be1b6e771521cbc534e1c928ac71167d122aff3

Observation cca380a1-eb77-432b-aa22-f71f09fbdc11 · outbound

This paper cites Miradata: A large-scale video dataset with long durations and structured captions.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Miradata: A large-scale video dataset with long durations and structured captions

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:12:13.803816Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T11:12:09.665314Z digest=sha256:2423c0cfdfe2878a267ee64c13e40969bf9c2f41843899b8d7c6fbad367af05b

Observation 35e5e471-8e47-4d3f-a3e5-44aaacf52b8a · outbound

This paper cites Animeceleb: Large-scale animation celebheads dataset for head reenactment.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Animeceleb: Large-scale animation celebheads dataset for head reenactment

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:12:13.699408Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T11:12:09.757488Z digest=sha256:79439ad9f58215fee6a0b1bef13022f917a89d9cfe79213700347ed97ed3da2a

Observation 24929bb3-2b7c-47be-be85-b7fb14f89aa9 · outbound

This paper cites Segment anything.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Segment anything

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:09.813019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:09.813019Z digest=sha256:e11de54fc4313dfffb2319501088aa2aa669375ed58e259097d122502294ce85

Observation a86e30b0-d6d0-4c61-96be-3e92e6e1f0ac · outbound

This paper cites Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:09.891514Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:09.891514Z digest=sha256:b7825f876cb4fdbc0225216119b5378e6f6debffbab2b594e8f99caed1315e4f

Observation 36f8d7e7-a113-43b8-a597-b8278264c7b6 · outbound

This paper cites Anim-director: A large multimodal model powered agent for controllable animation video generation.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Anim-director: A large multimodal model powered agent for controllable animation video generation

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:12:13.577772Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T11:12:09.969711Z digest=sha256:187a48bc4619a8e6a43dc5181c455ba2764dfd61fe410b1effe61c913f67921b

Observation 820661d1-558e-4c66-b64a-1d735b45ff39 · outbound

This paper cites Phantom: Subject-consistent video generation via cross-modal alignment.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Phantom: Subject-consistent video generation via cross-modal alignment

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:10.052472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:10.052472Z digest=sha256:aea4bbf7de48d6993dc7f896d448632c434f2cdcf1963aea41f3989b8dfb9a6d

Observation bfb470b3-7a96-427b-b7df-d31c0b71d6d1 · outbound

This paper cites Grounding dino: Marrying dino with grounded pre-training for open-set object detection.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Grounding dino: Marrying dino with grounded pre-training for open-set object detection

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:12:13.409848Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T11:12:10.127134Z digest=sha256:7b295b086c404395241ab35b41c29b3852fc0b800487a2b1befe16c166f8141d

Observation 54f3858b-f0ae-4f61-9f1b-9fe9c8f76c74 · outbound

This paper cites NVILA: Efficient Frontier Visual Language Models.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation NVILA: Efficient Frontier Visual Language Models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:10.219918Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:10.219918Z digest=sha256:c55c9bd55de2fa8cbe7806d43c77164caf16e7b0c8caba0748bc03073f5e183a

Observation ec6cd1af-2c39-4fe7-ba93-a9203c189b99 · outbound

This paper cites Videostudio: Generating consistent-content and multi-scene videos.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Videostudio: Generating consistent-content and multi-scene videos

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:12:13.247077Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T11:12:10.299362Z digest=sha256:8e02c00fc081384c296b3b68e9db9e1cce3d5f6bdab7f084bfb754bab49b291b

Observation 07fdbe30-a898-4f35-893a-7e0a49723455 · outbound

This paper cites Gpt-4o: Multimodal large language model, 2025.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Gpt-4o: Multimodal large language model, 2025

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:12:13.078271Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T11:12:10.378355Z digest=sha256:04e1386b42577573a78c4d787c50f1d3f48e4323ad9eb31dd3ce4c89c6c6b38d

Observation 8c0100f0-a661-4ea9-be03-c8bdda53aa7a · outbound

This paper cites Sakuga-42M Dataset: Scaling Up Cartoon Research.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Sakuga-42M Dataset: Scaling Up Cartoon Research

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:10.466249Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:10.466249Z digest=sha256:2573514586dc201977742b1bd63de231b3e8671bc6f0534ed5508fd0d01399fb

Observation 555b527d-8b63-4e0f-9bd6-be90d689a97a · outbound

This paper cites Scalable diffusion models with transformers.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Scalable diffusion models with transformers

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:10.547419Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:10.547419Z digest=sha256:3f923ffd92808ac5371a3c3883b9ec22c5119e1b24c2a391ebcab926e86f2228

Observation db3ff834-c87b-4a72-8b25-6c91f900c8d6 · outbound

This paper cites SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:10.655648Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:10.655648Z digest=sha256:46b0a3cfb5399c2b8dc12ff664071e636ab06881d5484851ae426f526446d257

Observation ef22b037-3bb6-4cbe-b768-acdf42d65130 · outbound

This paper cites Learning transferable visual models from natural language supervision.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Learning transferable visual models from natural language supervision

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:10.735546Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:10.735546Z digest=sha256:f593134d9367442dd91c9e19bce5252bae4e30e186bf72519b93ceebec195694

Observation 67e8c4e3-c52d-46b9-9d04-ae9478b51db7 · outbound

This paper cites Videofactory: Swap attention in spatiotemporal diffusions for text-to-video generation.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Videofactory: Swap attention in spatiotemporal diffusions for text-to-video generation

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:12:12.866309Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T11:12:10.816759Z digest=sha256:7b7e75d5eda127e1221da83773d328bad8f365e53ea2e4a5f55ee7dd16615673

Observation d9559444-4464-46cf-b125-2bc3146cfffa · outbound

This paper cites InternVid: A Large-scale Video-Text Dataset for Multimodal Understanding and Generation.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation InternVid: A Large-scale Video-Text Dataset for Multimodal Understanding and Generation

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:10.899127Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:10.899127Z digest=sha256:e47610944228afdadf297a9a3d1c8622a1e45189dbed10e2dde9834f9d3a8324

Observation a3ff8ddc-5acc-4107-97cd-f55f3767ee50 · outbound

This paper cites Dreamrunner: Fine-grained storytelling video generation with retrieval-augmented motion adaptation.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Dreamrunner: Fine-grained storytelling video generation with retrieval-augmented motion adaptation

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:10.988953Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:10.988953Z digest=sha256:c0a7836c51152b9abf1a73033f88afd740383c39f3353e8394855c00715ed7f5

Observation 6eb2fcc0-3d80-460c-8e33-0dc1fa6869a5 · outbound

This paper cites Understanding animation.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Understanding animation

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:12:12.648824Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T11:12:11.080666Z digest=sha256:0af9a470aca94b09b93745627c0648269184d25cedbccd8593a1b21c13fe824c

Observation 6a6aa371-4938-48d2-90ad-025c4183ff3d · outbound

This paper cites Automated Movie Generation via Multi-Agent CoT Planning.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Automated Movie Generation via Multi-Agent CoT Planning

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:11.166087Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:11.166087Z digest=sha256:af8e56d19688e9886beb7cb5b255ae86ddeff612f3b8149f3bf45b96784a2e7f

Observation 30381dfa-6f88-438b-a9c9-990b24562a6d · outbound

This paper cites Pandora: Towards General World Model with Natural Language Actions and Video States.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Pandora: Towards General World Model with Natural Language Actions and Video States

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:11.222249Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:11.222249Z digest=sha256:06e16c54f7aa846666c490922dcf2c18ce6deedfcf27ac96815e0f4d4a53c697

Observation f75afb69-c4ce-4946-a2e8-a6659b257310 · outbound

This paper cites Dynamicrafter: Animating open-domain images with video diffusion priors.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Dynamicrafter: Animating open-domain images with video diffusion priors

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:12:12.492949Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T11:12:11.286902Z digest=sha256:e0a50a9df44a305444658b7037bbdd494f9afb1ad83ddf22ae9294f35ed13c08

Observation c906f12f-548a-4246-a263-f63b490a5dfd · outbound

This paper cites LVD-2M: A Long-take Video Dataset with Temporally Dense Captions.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation LVD-2M: A Long-take Video Dataset with Temporally Dense Captions

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:11.338933Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:11.338933Z digest=sha256:0a93bff4144a0e0783f80e4be9944cb5502186b6108f2960eafa23a12c2549c6

Observation 50a82ba6-d370-4019-b301-e78b5afa1694 · outbound

This paper cites Vript: A video is worth thousands of words.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Vript: A video is worth thousands of words

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:12:12.358110Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T11:12:11.402600Z digest=sha256:3db6c467828d614310f90e04f240caf0dab363a115655ff9efd51b08fb189163

Observation bc59aa92-dab7-4113-ae87-2e9112b87892 · outbound

This paper cites SEED-Story: Multimodal Long Story Generation with Large Language Model.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation SEED-Story: Multimodal Long Story Generation with Large Language Model

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:11.474850Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:11.474850Z digest=sha256:b93f19a7bbf407abc3ed28a007b78e2218efb0a694a3f2be138f7699868c1b90

Observation 2be077c8-5812-4e90-beb0-51b75fdd6892 · outbound

This paper cites IP-Adapter: Text Compatible Image Prompt Adapter for Text-to-Image Diffusion Models.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation IP-Adapter: Text Compatible Image Prompt Adapter for Text-to-Image Diffusion Models

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:11.511487Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:11.511487Z digest=sha256:63e3dff6df0a0c2668d18ac7d41b0eb2a8654dcd362051015c35147cae89519a

Observation babab955-deef-496a-a763-52bb4fe2f02c · outbound

This paper cites Sa2VA: Marrying SAM2 with LLaVA for Dense Grounded Understanding of Images and Videos.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Sa2VA: Marrying SAM2 with LLaVA for Dense Grounded Understanding of Images and Videos

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:11.581622Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:11.581622Z digest=sha256:f5fcc17e04a66941022ca8648efcc494926865a201c0b9e56ce8bf6443fd7bf3

Observation 3056d301-ef7b-43d8-90cb-904e1dcc52be · outbound

This paper cites a cow is mooning.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation a cow is mooning

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:12:12.237359Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-07T11:12:11.643982Z digest=sha256:a41d927ee84e0e7b28c007c9f750e12f26df2fb6e6d4ba9f6db806d3f6bc71a6

Pith citing papers

Observation 12a9d3e5-c4e3-4a23-a319-0bec9d33f43d · inbound

Rethinking Position Embedding as a Context Controller for Multi-Reference and Multi-Shot Video Generation cites this paper.

Rethinking Position Embedding as a Context Controller for Multi-Reference and Multi-Shot Video Generation AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation

Reference 27

Resolution
unresolved
no resolver link, observed 2026-07-13T12:23:05.876881Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T12:23:05.876881Z digest=sha256:80f671ed3e9d8fae525a5cf85f00cd99df2eaf81e131bd25c83f6daefb32a8f6

Observation 03dd8ea7-b774-48bf-ab89-aa092011185b · inbound

MuSS: A Large-Scale Dataset and Cinematic Narrative Benchmark for Multi-Shot Subject-to-Video Generation cites this paper.

MuSS: A Large-Scale Dataset and Cinematic Narrative Benchmark for Multi-Shot Subject-to-Video Generation AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-11T21:11:19.182772Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-08T06:28:42.129881Z digest=sha256:a59f7ee4c911fc32122e042ac4cde3f2408cfb3c73bb54cdd9e2eae327279b16

Observation 1cf7ef51-3212-4406-ac33-bd2199c83f1a · inbound

MuSS: A Large-Scale Dataset and Cinematic Narrative Benchmark for Multi-Shot Subject-to-Video Generation cites this paper.

MuSS: A Large-Scale Dataset and Cinematic Narrative Benchmark for Multi-Shot Subject-to-Video Generation AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-12T00:51:15.171755Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-12T00:50:10.509727Z digest=sha256:8654d996165691305807cb7d8fd18a0cc5eff8dfb81df7a7537693d962738287

Observation 0495b8e7-e0f5-4b47-8fb6-7ae4308d815c · inbound

DrawVideo: Generating Long Video from Storyboard Keyframe Sketches cites this paper.

DrawVideo: Generating Long Video from Storyboard Keyframe Sketches AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-25T02:40:14.532466Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-05-25T02:40:00.873188Z digest=sha256:1859ec4cce6cfb130d49883abceac85586141de7d21c528c32018ff63eda4722