Pith. sign in

Paper Citation Record · LEDGER

Vorch-IR: Long-Form Unified Multimodal Identity Replacement Video Generation

As of 15 August 2026, this Paper Citation Record lists 15 of 15 outbound references and 0 inbound Pith citation observations for arXiv:2608.05648.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.05648 v1

Coverage vector

measured 15 of 15 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-08T04:54:58.686040Z

measured 15 of 15 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

15 of 15 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved15
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 14876000-a03e-4650-b884-99611cdddd6e · outbound

This paper cites Diffusion Forcing: Next-token Prediction Meets Full-Sequence Diffusion.

Vorch-IR: Long-Form Unified Multimodal Identity Replacement Video Generation Diffusion Forcing: Next-token Prediction Meets Full-Sequence Diffusion

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-08T04:54:58.639683Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T04:54:58.639683Z digest=sha256:3b9b4b36bf7cd8409c3085d42087f774319ab16652142d3cf4a7763be589dbe1

Observation 694d1c61-ed22-4361-b70e-37d4d2dc49a7 · outbound

This paper cites Seedance 1.0: Exploring the Boundaries of Video Generation Models.

Vorch-IR: Long-Form Unified Multimodal Identity Replacement Video Generation Seedance 1.0: Exploring the Boundaries of Video Generation Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-08T04:54:58.647292Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T04:54:58.647292Z digest=sha256:475c7558f0a9198c74aa8888ee5ff58c0f1d3162802cb87f96701887d54e521b

Observation d25c41a4-1df8-4587-af7b-1bf1dcea5886 · outbound

This paper cites LTX-2: Efficient Joint Audio-Visual Foundation Model.

Vorch-IR: Long-Form Unified Multimodal Identity Replacement Video Generation LTX-2: Efficient Joint Audio-Visual Foundation Model

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-08T04:54:58.650815Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T04:54:58.650815Z digest=sha256:e3c578c85657e4109d48bdb3a4e29872a446a65af417bfd730f03f4afdbf3df9

Observation f6fa9955-a18d-487f-8b09-5d0b3c84cff8 · outbound

This paper cites Self Forcing: Bridging the Train-Test Gap in Autoregressive Video Diffusion.

Vorch-IR: Long-Form Unified Multimodal Identity Replacement Video Generation Self Forcing: Bridging the Train-Test Gap in Autoregressive Video Diffusion

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-08T04:54:58.658009Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T04:54:58.658009Z digest=sha256:bbcf5ccdd441affd28c9967cd52d62f0dc0b2196524fc5260e45d5f72da72c6e

Observation 91ea9fcc-b054-465a-9e74-1cafd7393850 · outbound

This paper cites AnyV2V: A Tuning-Free Framework For Any Video-to-Video Editing Tasks.

Vorch-IR: Long-Form Unified Multimodal Identity Replacement Video Generation AnyV2V: A Tuning-Free Framework For Any Video-to-Video Editing Tasks

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-08T04:54:58.661523Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T04:54:58.661523Z digest=sha256:16e59b76dd58039c90011167eda31612f065c917e96c667b4a5839c536a7f0f9

Observation 9438765d-d80d-42f6-bb78-657360bcfe7b · outbound

This paper cites One-to-All Animation: Alignment-Free Character Animation and Image Pose Transfer.

Vorch-IR: Long-Form Unified Multimodal Identity Replacement Video Generation One-to-All Animation: Alignment-Free Character Animation and Image Pose Transfer

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-08T04:54:58.667551Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T04:54:58.667551Z digest=sha256:d2ed9155cf74ac33e65b1573232a3dbb97680b9bfcd0bd8ad9cbdbeea7a055a5

Observation 99330b26-b962-428f-9d6e-ed4b9a17c12a · outbound

This paper cites HunyuanVideo: A Systematic Framework For Large Video Generative Models.

Vorch-IR: Long-Form Unified Multimodal Identity Replacement Video Generation HunyuanVideo: A Systematic Framework For Large Video Generative Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-08T04:54:58.670639Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T04:54:58.670639Z digest=sha256:e6ac1a364f1a5c76445fe3c441cd11621b6ea0f398f8042b9ebf41cd15c33894

Observation 75a0b93e-7885-4c3b-be6b-3beb7251ee16 · outbound

This paper cites Unianimate-dit: Human image animation with large-scale video diffusion transformer.arXiv preprint arXiv:2504.11289,.

Vorch-IR: Long-Form Unified Multimodal Identity Replacement Video Generation Unianimate-dit: Human image animation with large-scale video diffusion transformer.arXiv preprint arXiv:2504.11289,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-08T04:54:58.677211Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T04:54:58.677211Z digest=sha256:b1b59ac3c0414da60be050c7e328336a4628795eb766703edb55d5ad3bc2e5da

Observation f01d21ae-a35f-41aa-a0cd-21ca06c4ba1f · outbound

This paper cites SCAIL-2: Unifying Controlled Character Animation with End-to-End In-Context Conditioning.

Vorch-IR: Long-Form Unified Multimodal Identity Replacement Video Generation SCAIL-2: Unifying Controlled Character Animation with End-to-End In-Context Conditioning

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-08T04:54:58.680041Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T04:54:58.680041Z digest=sha256:9eebf1691ead91b12bf618f5f58f140b394f9f567938f179b4b38e76e739a19c

Observation 5bc91c86-05b6-4636-90b4-a3ee629a747e · outbound

This paper cites SteadyDancer: Harmonized and Coherent Human Image Animation with First-Frame Preservation.

Vorch-IR: Long-Form Unified Multimodal Identity Replacement Video Generation SteadyDancer: Harmonized and Coherent Human Image Animation with First-Frame Preservation

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-08T04:54:58.683114Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T04:54:58.683114Z digest=sha256:931ee789bc0f7981741b8b0641de6cb237ee6d0e5f51cc913cc262275d132701

Observation ea3bcdc1-ed91-4ee5-aa03-278c26125255 · outbound

This paper cites Taming Teacher Forcing for Masked Autoregressive Video Generation.

Vorch-IR: Long-Form Unified Multimodal Identity Replacement Video Generation Taming Teacher Forcing for Masked Autoregressive Video Generation

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-08T04:54:58.686040Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T04:54:58.686040Z digest=sha256:a857db39691b350efee29b0a4552321dfac212cae3d6d282ec02330164e700a5

Observation 9eb543e8-d88d-4823-bb44-91125cb2dd70 · outbound

This paper cites Dreamactor-m2: Universal character image animation via spatiotemporal in-context learning.arXiv preprint arXiv:2601.21716,.

Vorch-IR: Long-Form Unified Multimodal Identity Replacement Video Generation Dreamactor-m2: Universal character image animation via spatiotemporal in-context learning.arXiv preprint arXiv:2601.21716,

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-08T04:54:58.664692Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T04:54:58.664692Z digest=sha256:e0a4340a5bce4a2745c5c4dbcae03f04f67353e329ed81fc67c1dce770807e0f

Observation e7b16efa-ece7-4586-a970-745ac168715d · outbound

This paper cites Wan: Open and Advanced Large-Scale Video Generative Models.

Vorch-IR: Long-Form Unified Multimodal Identity Replacement Video Generation Wan: Open and Advanced Large-Scale Video Generative Models

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-08T04:54:58.673754Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T04:54:58.673754Z digest=sha256:cdcce6bc1a2384862735b34cf5e6a585491eb6e22fffaa432c36453e5281a7e7

Observation 74bb284c-54b8-47aa-99d2-c80300e66aec · outbound

This paper cites SkyReels-Audio: Omni Audio-Conditioned Talking Portraits in Video Diffusion Transformers.

Vorch-IR: Long-Form Unified Multimodal Identity Replacement Video Generation SkyReels-Audio: Omni Audio-Conditioned Talking Portraits in Video Diffusion Transformers

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-08T04:54:58.643717Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T04:54:58.643717Z digest=sha256:ad5f125e6ede9e9eeb6b161ddc1767617d61255b974371884940b1bb5b82db2a

Observation 7b68b13c-0de6-45ee-807e-dc1472fd2363 · outbound

This paper cites HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation.

Vorch-IR: Long-Form Unified Multimodal Identity Replacement Video Generation HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation

Reference 2026

Resolution
unresolved
no resolver link, observed 2026-08-08T04:54:58.654243Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T04:54:58.654243Z digest=sha256:a348f6ac966773871ee12dc8fe3af2ef19f22be2196e32a4c051db9e1fc202a3

Pith citing papers

No inbound Pith citation observations are available.