Pith. sign in

Paper Citation Record · LEDGER

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation

As of 19 August 2026, this Paper Citation Record lists 50 of 50 outbound references and 4 inbound Pith citation observations for arXiv:2506.03126.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.03126 v1

Coverage vector

measured 50 of 50 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:12:11.643982Z

measured 54 of 54 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-13T12:23:05.876881Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-25T02:40:14.528606Z

Reference resolution

50 of 50 outbound references displayed

  • verified exact1
  • verified fuzzy13
  • unresolved36
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 389f885a-2c05-4557-aa24-30917415e80d · outbound

This paper cites Frozen in time: A joint video and image encoder for end-to-end retrieval.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Frozen in time: A joint video and image encoder for end-to-end retrieval

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:07.596205Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:07.596205Z digest=sha256:eb8ce2a3790004387ee45016d07d36dc339115419707916a9819e7eabf788c3a

Observation dd058a13-c529-4a70-ab25-e2538c4a7612 · outbound

This paper cites Activitynet: A large-scale video benchmark for human activity understanding.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Activitynet: A large-scale video benchmark for human activity understanding

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:07.662318Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:07.662318Z digest=sha256:5044d581486d51123c8e15f30ebc26393338172428d5ec7b4e73c592cc0ff26b

Observation 8af38723-941f-4c8c-a6f2-d8a21683ed7c · outbound

This paper cites Panda-70m: Captioning 70m videos with multiple cross-modality teachers.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Panda-70m: Captioning 70m videos with multiple cross-modality teachers

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:07.730235Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:07.730235Z digest=sha256:5bdd8b4b50d0616e04060b14cd993c6b62060b9df4b26b189af3242f57508b67

Observation da75d926-c224-400f-916a-84cd7e90880d · outbound

This paper cites Multi-subject Open-set Personalization in Video Generation.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Multi-subject Open-set Personalization in Video Generation

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:07.805412Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:07.805412Z digest=sha256:0adbed61ed3252c6e16a1e286764324e6433abcd430a2f979c1347cb5bb7ee82

Observation 7a32cccd-bf7b-4da4-9bb8-7731c437384d · outbound

This paper cites Internvl: Scaling up vision foundation models and aligning for generic visual-linguistic tasks.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Internvl: Scaling up vision foundation models and aligning for generic visual-linguistic tasks

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:07.888630Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:07.888630Z digest=sha256:6f4bf9b0c23688326bf5d58d6b05195b1bb0e57f24deda9480311797eecb42a1

Observation 2ad33bc0-f505-474e-b8ce-99f722b5d127 · outbound

This paper cites AnimeGamer: Infinite Anime Life Simulation with Next Game State Prediction.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation AnimeGamer: Infinite Anime Life Simulation with Next Game State Prediction

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-08-07T11:12:12.106906Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T11:12:07.996721Z digest=sha256:56a57fc7b09854115321de27465eb9980827625e8bce05a963a5d2c68dadb972

Observation 42f72ecf-dc28-47aa-828b-08c5f6d867ff · outbound

This paper cites Gemini, 2024.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Gemini, 2024

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:12:14.189027Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T11:12:08.133677Z digest=sha256:7a40735ad71944e221f20048073f52b3e372f915ca06be8717011d1b40891ad5

Observation 10bbdf14-7a25-4c48-bb27-3789b2a20813 · outbound

This paper cites CINEMA: Coherent Multi-Subject Video Generation via MLLM-Based Guidance.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation CINEMA: Coherent Multi-Subject Video Generation via MLLM-Based Guidance

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:08.227600Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:08.227600Z digest=sha256:e2972c63738e4c6ac548bda9570001c287b99dee43afd874e05e297908be4195

Observation 374fa6d5-8d77-432a-8862-f6ea8f81c154 · outbound

This paper cites DreamSim: Learning New Dimensions of Human Visual Similarity using Synthetic Data.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation DreamSim: Learning New Dimensions of Human Visual Similarity using Synthetic Data

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:08.339057Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:08.339057Z digest=sha256:f00d9fa465f44a511887223804d59f11744dfe2cdcc20a2b1613ba6ca78246d6

Observation 8a36178e-afd2-410e-a11e-03b8c7f5c3c3 · outbound

This paper cites TaleCrafter: Interactive Story Visualization with Multiple Characters.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation TaleCrafter: Interactive Story Visualization with Multiple Characters

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:08.485043Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:08.485043Z digest=sha256:5a8154839d6321fbd0334c02e2b2c70dd90c490d695784257dd6232bc1da60b2

Observation d64ba14d-5ff1-4bed-9806-cc8749ca3d35 · outbound

This paper cites The Llama 3 Herd of Models.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation The Llama 3 Herd of Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:08.630129Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:08.630129Z digest=sha256:7d79452546d8cb9f681413eef9b60082bde1689b520bdea18c8010d58abbf886

Observation b90d077f-4e9f-4da3-a746-15ea5e9d7b81 · outbound

This paper cites Mix-of-show: Decentralized low-rank adaptation for multi-concept customization of diffusion models.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Mix-of-show: Decentralized low-rank adaptation for multi-concept customization of diffusion models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:08.729397Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:08.729397Z digest=sha256:a667d18eee9aca0133d25d96db1b7e6caf0da05d1a3670a35b250adb084bc9d7

Observation 9356d9de-e9d4-44bb-bcb6-f78c65570da9 · outbound

This paper cites ROICtrl: Boosting Instance Control for Visual Generation.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation ROICtrl: Boosting Instance Control for Visual Generation

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:08.846411Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:08.846411Z digest=sha256:37f2e0ded0ec00300deb96e7d6b03d144d462840e9c62560101c29286f228f65

Observation f573cba1-365b-462e-8c50-832797999e3e · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:08.909571Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:08.909571Z digest=sha256:97f85f2c95fee72570584496f8aa59ec00cb48c709e4dcb0a528f6bb8fd1bdb4

Observation 0ffec6cb-a1a0-4858-b45a-43d583ac1ff2 · outbound

This paper cites Long Context Tuning for Video Generation.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Long Context Tuning for Video Generation

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:08.965970Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:08.965970Z digest=sha256:5be60052f1aee8d0ed8c7b23a884df574b0809541325a6d43944034cfb26f839

Observation e051b2b5-5fe8-4eb4-944e-760793cdaaea · outbound

This paper cites AnyStory: Towards Unified Single and Multiple Subject Personalization in Text-to-Image Generation.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation AnyStory: Towards Unified Single and Multiple Subject Personalization in Text-to-Image Generation

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:09.048829Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:09.048829Z digest=sha256:15e73be7ae999e857e2cd862648530bb87b84b0802ebcdb48a982fb7decd6ba7

Observation 7fa80870-4be0-4fc5-a828-286d8f634c71 · outbound

This paper cites ID-Animator: Zero-Shot Identity-Preserving Human Video Generation.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation ID-Animator: Zero-Shot Identity-Preserving Human Video Generation

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:09.126710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:09.126710Z digest=sha256:d557a42c5bd1f99edaaa4757a74a2930ceb3908f158bb1141ef4915d659a5791

Observation aaf0aeb4-19f8-4519-9e5d-d7174043c9cf · outbound

This paper cites CogVideo: Large-scale Pretraining for Text-to-Video Generation via Transformers.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation CogVideo: Large-scale Pretraining for Text-to-Video Generation via Transformers

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:09.189653Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:09.189653Z digest=sha256:0c3578727f601bdbe7656a94f677e94485706915f203b74f7158ecd246c08c48

Observation a40c66a4-0d8e-408a-a927-93464dff33f3 · outbound

This paper cites Owl-1: Omni World Model for Consistent Long Video Generation.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Owl-1: Omni World Model for Consistent Long Video Generation

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:09.266583Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:09.266583Z digest=sha256:e23b3cb706a47a64848fb168ba34f91e9a762f10fa9e3866103c2086e693d739

Observation bc92ba84-61f2-483c-bbe5-3ecf2d2c648f · outbound

This paper cites ConceptMaster: Multi-Concept Video Customization on Diffusion Transformer Models Without Test-Time Tuning.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation ConceptMaster: Multi-Concept Video Customization on Diffusion Transformer Models Without Test-Time Tuning

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:09.347154Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:09.347154Z digest=sha256:0a0f90cf97110c961874fb2534f6deaf84e2815e52f9c0dbaf8856d37013cf52

Observation ae659b54-8ae6-4d0f-8e35-7eb036357a6c · outbound

This paper cites TangoFlux: Super Fast and Faithful Text to Audio Generation with Flow Matching and Clap-Ranked Preference Optimization.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation TangoFlux: Super Fast and Faithful Text to Audio Generation with Flow Matching and Clap-Ranked Preference Optimization

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:09.446928Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:09.446928Z digest=sha256:5c847b6baa47631895534f01ca28d5b8a1231284197fb792b3c92622bccc6573

Observation 75efc4a0-733f-4862-8af8-bcd8407e95e1 · outbound

This paper cites AniSora: Exploring the Frontiers of Animation Video Generation in the Sora Era.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation AniSora: Exploring the Frontiers of Animation Video Generation in the Sora Era

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:09.494420Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:09.494420Z digest=sha256:c3611cc0ee496137f69956303fc4cd340172bfaff71cdaff71c8d40841c77af0

Observation eb25d9f7-ac88-4b57-a016-a3eefcb212da · outbound

This paper cites Videobooth: Diffusion-based video generation with image prompts.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Videobooth: Diffusion-based video generation with image prompts

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:12:13.999178Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T11:12:09.571073Z digest=sha256:696ab12e1cf2bf3ed645d362371a6f268cb2ae2826a718894abc7071e2b58611

Observation cca380a1-eb77-432b-aa22-f71f09fbdc11 · outbound

This paper cites Miradata: A large-scale video dataset with long durations and structured captions.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Miradata: A large-scale video dataset with long durations and structured captions

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:12:13.803816Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T11:12:09.665314Z digest=sha256:7af0cbf610c24ba9ac7bcbfb1de3f702bb57dfcebab5d59eea8daad1243b8cdf

Observation 35e5e471-8e47-4d3f-a3e5-44aaacf52b8a · outbound

This paper cites Animeceleb: Large-scale animation celebheads dataset for head reenactment.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Animeceleb: Large-scale animation celebheads dataset for head reenactment

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:12:13.699408Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T11:12:09.757488Z digest=sha256:1969310b480fdb0df0d13f3187a8c7d4c30bef6c435c2c0cfbe6f8b2407e896c

Observation 24929bb3-2b7c-47be-be85-b7fb14f89aa9 · outbound

This paper cites Segment anything.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Segment anything

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:09.813019Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:09.813019Z digest=sha256:796966c66deaa2a77a9f29f4d392dd5aaf615361a6994bed153ba8262f350bc2

Observation a86e30b0-d6d0-4c61-96be-3e92e6e1f0ac · outbound

This paper cites Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:09.891514Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:09.891514Z digest=sha256:a6f5afc75f131490680e2e1b87aad99efaa0c4f67f903f8c4800942bac13b4f9

Observation 36f8d7e7-a113-43b8-a597-b8278264c7b6 · outbound

This paper cites Anim-director: A large multimodal model powered agent for controllable animation video generation.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Anim-director: A large multimodal model powered agent for controllable animation video generation

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:12:13.577772Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T11:12:09.969711Z digest=sha256:2efd14fa0be20aa4e4d41a18201afcac3c88631e8079faf497e761ae3bff6631

Observation 820661d1-558e-4c66-b64a-1d735b45ff39 · outbound

This paper cites Phantom: Subject-consistent video generation via cross-modal alignment.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Phantom: Subject-consistent video generation via cross-modal alignment

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:10.052472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:10.052472Z digest=sha256:bdeb81285a4d2080c543dba8f60df41bf666b54ecbef1159df7fc456dd475ae7

Observation bfb470b3-7a96-427b-b7df-d31c0b71d6d1 · outbound

This paper cites Grounding dino: Marrying dino with grounded pre-training for open-set object detection.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Grounding dino: Marrying dino with grounded pre-training for open-set object detection

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:12:13.409848Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T11:12:10.127134Z digest=sha256:a63e77958133444e9d172139045b625daa7a156663920b0876cbb9c6e9da0d40

Observation 54f3858b-f0ae-4f61-9f1b-9fe9c8f76c74 · outbound

This paper cites NVILA: Efficient Frontier Visual Language Models.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation NVILA: Efficient Frontier Visual Language Models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:10.219918Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:10.219918Z digest=sha256:23406d3d591b560cef87f41804bfd5feccb6fa99c349c530f933ef789d5f8550

Observation ec6cd1af-2c39-4fe7-ba93-a9203c189b99 · outbound

This paper cites Videostudio: Generating consistent-content and multi-scene videos.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Videostudio: Generating consistent-content and multi-scene videos

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:12:13.247077Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T11:12:10.299362Z digest=sha256:a18ee701649d9479931818200275fc91636bb0e336269053887523d205ad13c4

Observation 07fdbe30-a898-4f35-893a-7e0a49723455 · outbound

This paper cites Gpt-4o: Multimodal large language model, 2025.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Gpt-4o: Multimodal large language model, 2025

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:12:13.078271Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T11:12:10.378355Z digest=sha256:4771b66ebd9a06439121eb3d00356b93c7659014cac6d8f2ca6dc60aeabe6e5b

Observation 8c0100f0-a661-4ea9-be03-c8bdda53aa7a · outbound

This paper cites Sakuga-42M Dataset: Scaling Up Cartoon Research.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Sakuga-42M Dataset: Scaling Up Cartoon Research

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:10.466249Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:10.466249Z digest=sha256:ebfc5602cf25054276d8f9764529242b3e4945a1b3d07558916a48bdfbf4dff0

Observation 555b527d-8b63-4e0f-9bd6-be90d689a97a · outbound

This paper cites Scalable diffusion models with transformers.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Scalable diffusion models with transformers

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:10.547419Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:10.547419Z digest=sha256:21f13f7c5d18d194c9cd5bf1a16b3acea53c1fee04033548f9df34d04ea2a59d

Observation db3ff834-c87b-4a72-8b25-6c91f900c8d6 · outbound

This paper cites SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:10.655648Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:10.655648Z digest=sha256:44545bae30f308df91bd65844a2b83456e526e5a92ba3b6bc73040a5507d2724

Observation ef22b037-3bb6-4cbe-b768-acdf42d65130 · outbound

This paper cites Learning transferable visual models from natural language supervision.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Learning transferable visual models from natural language supervision

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:10.735546Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:10.735546Z digest=sha256:4087c6941b89c662b871bab1d576a76bdd2659691cb74621be873456b55874cc

Observation 67e8c4e3-c52d-46b9-9d04-ae9478b51db7 · outbound

This paper cites Videofactory: Swap attention in spatiotemporal diffusions for text-to-video generation.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Videofactory: Swap attention in spatiotemporal diffusions for text-to-video generation

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:12:12.866309Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T11:12:10.816759Z digest=sha256:3538b1d82c636f778606f756ad91163c4060db2fc86d1ee05193e158b94fa96a

Observation d9559444-4464-46cf-b125-2bc3146cfffa · outbound

This paper cites InternVid: A Large-scale Video-Text Dataset for Multimodal Understanding and Generation.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation InternVid: A Large-scale Video-Text Dataset for Multimodal Understanding and Generation

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:10.899127Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:10.899127Z digest=sha256:aca2e2a0f48dc8c99f67e6a0d2eae2cb4a78b5d4f693c1b7048616d01fe99786

Observation a3ff8ddc-5acc-4107-97cd-f55f3767ee50 · outbound

This paper cites Dreamrunner: Fine-grained storytelling video generation with retrieval-augmented motion adaptation.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Dreamrunner: Fine-grained storytelling video generation with retrieval-augmented motion adaptation

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:10.988953Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:10.988953Z digest=sha256:2e964c1a9d19ee46e82cbbc7e8212725fb7eb00cdf5a7661500e4d272f7adebe

Observation 6eb2fcc0-3d80-460c-8e33-0dc1fa6869a5 · outbound

This paper cites Understanding animation.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Understanding animation

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:12:12.648824Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T11:12:11.080666Z digest=sha256:4c3b91b048072a7c62dac547e4540b9446a03ffd74717e5b47925cb761661834

Observation 6a6aa371-4938-48d2-90ad-025c4183ff3d · outbound

This paper cites Automated Movie Generation via Multi-Agent CoT Planning.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Automated Movie Generation via Multi-Agent CoT Planning

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:11.166087Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:11.166087Z digest=sha256:0da7afa4f613ffcde4d02c316e4a527e00d8df5355673cf34e76021900665d5f

Observation 30381dfa-6f88-438b-a9c9-990b24562a6d · outbound

This paper cites Pandora: Towards General World Model with Natural Language Actions and Video States.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Pandora: Towards General World Model with Natural Language Actions and Video States

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:11.222249Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:11.222249Z digest=sha256:d7b7a8b2aeb2f2a818ba3b7d3b1df052fd177aa3c1e5f3f2c42421b64f4712e2

Observation f75afb69-c4ce-4946-a2e8-a6659b257310 · outbound

This paper cites Dynamicrafter: Animating open-domain images with video diffusion priors.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Dynamicrafter: Animating open-domain images with video diffusion priors

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:12:12.492949Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T11:12:11.286902Z digest=sha256:84b075579d1ff0fff03d4d767e158d9b73e64b5e5a2ba307d03f5f1ecdb4744f

Observation c906f12f-548a-4246-a263-f63b490a5dfd · outbound

This paper cites LVD-2M: A Long-take Video Dataset with Temporally Dense Captions.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation LVD-2M: A Long-take Video Dataset with Temporally Dense Captions

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:11.338933Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:11.338933Z digest=sha256:16a154ec334a881cc8f0297f3dae78938b3219a5c3b41cb457787a10beb39855

Observation 50a82ba6-d370-4019-b301-e78b5afa1694 · outbound

This paper cites Vript: A video is worth thousands of words.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Vript: A video is worth thousands of words

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:12:12.358110Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T11:12:11.402600Z digest=sha256:403481e5946b887b07da95be4ca821480a0fc175855d9c89302fbb67afb86c2e

Observation bc59aa92-dab7-4113-ae87-2e9112b87892 · outbound

This paper cites SEED-Story: Multimodal Long Story Generation with Large Language Model.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation SEED-Story: Multimodal Long Story Generation with Large Language Model

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:11.474850Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:11.474850Z digest=sha256:0a98736fd141b561f72382986b49dc2e3d7206022c5f27d5930bf1e8e48a39f0

Observation 2be077c8-5812-4e90-beb0-51b75fdd6892 · outbound

This paper cites IP-Adapter: Text Compatible Image Prompt Adapter for Text-to-Image Diffusion Models.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation IP-Adapter: Text Compatible Image Prompt Adapter for Text-to-Image Diffusion Models

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:11.511487Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:11.511487Z digest=sha256:2f6ea780737901a3e67fcd4994dc4dbd4e949fdf3c068eb79df55b505f51b66a

Observation babab955-deef-496a-a763-52bb4fe2f02c · outbound

This paper cites Sa2VA: Marrying SAM2 with LLaVA for Dense Grounded Understanding of Images and Videos.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation Sa2VA: Marrying SAM2 with LLaVA for Dense Grounded Understanding of Images and Videos

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T11:12:11.581622Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:12:11.581622Z digest=sha256:b2c6fe6f71dea252d3014e7607081c4c931f7971d77a3b39a0067ca4196ac53e

Observation 3056d301-ef7b-43d8-90cb-904e1dcc52be · outbound

This paper cites a cow is mooning.

AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation a cow is mooning

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:12:12.237359Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T11:12:11.643982Z digest=sha256:9a12a3b5d786ed3a232c57cd9a7dd65aff754653c30a05ab23fbc2a0a250f2c3

Pith citing papers

Observation 12a9d3e5-c4e3-4a23-a319-0bec9d33f43d · inbound

Rethinking Position Embedding as a Context Controller for Multi-Reference and Multi-Shot Video Generation cites this paper.

Rethinking Position Embedding as a Context Controller for Multi-Reference and Multi-Shot Video Generation AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation

Reference 27

Resolution
unresolved
no resolver link, observed 2026-07-13T12:23:05.876881Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T12:23:05.876881Z digest=sha256:7c6d1313a9a46ac13e0a48cdc048fdebc7d178c58dc3c33c9cb763b692bb3d51

Observation 03dd8ea7-b774-48bf-ab89-aa092011185b · inbound

MuSS: A Large-Scale Dataset and Cinematic Narrative Benchmark for Multi-Shot Subject-to-Video Generation cites this paper.

MuSS: A Large-Scale Dataset and Cinematic Narrative Benchmark for Multi-Shot Subject-to-Video Generation AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-11T21:11:19.182772Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-08T06:28:42.129881Z digest=sha256:bd7076371469b76bfd1b61ad86ef5b10ca0a773374feb25c2929a43be79fb384

Observation 1cf7ef51-3212-4406-ac33-bd2199c83f1a · inbound

MuSS: A Large-Scale Dataset and Cinematic Narrative Benchmark for Multi-Shot Subject-to-Video Generation cites this paper.

MuSS: A Large-Scale Dataset and Cinematic Narrative Benchmark for Multi-Shot Subject-to-Video Generation AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-05-12T00:51:15.171755Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-12T00:50:10.509727Z digest=sha256:ad99514ff564efeeffa723a7bc7b5de9163183baa01e5cf3813e416d61c0041d

Observation 0495b8e7-e0f5-4b47-8fb6-7ae4308d815c · inbound

DrawVideo: Generating Long Video from Storyboard Keyframe Sketches cites this paper.

DrawVideo: Generating Long Video from Storyboard Keyframe Sketches AnimeShooter: A Multi-Shot Animation Dataset for Reference-Guided Video Generation

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-25T02:40:14.532466Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-25T02:40:00.873188Z digest=sha256:597d64785ae12a8748a7a68c0bf3ad894759d63cbc5e710e6b3a9cd6309e9616