Pith. sign in

Paper Citation Record · LEDGER

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation

As of 16 August 2026, this Paper Citation Record lists 100 of 122 outbound references and 6 inbound Pith citation observations for arXiv:2505.18078.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.18078 v1

Coverage vector

measured 100 of 122 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:39:40.471781Z

measured 106 of 106 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T18:29:40.891406Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T05:17:40.170733Z

Reference resolution

100 of 122 outbound references displayed

  • verified exact2
  • verified fuzzy17
  • unresolved81
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 2900a376-f8b4-4b1b-99e2-9c70f8521af7 · outbound

This paper cites Kling ai: Next-generation ai creative studio.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Kling ai: Next-generation ai creative studio

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:27.854249Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:27.854249Z digest=sha256:5822a4793b2fc35dfee22aace35512c97b67aa51eeb9c9f7282c7dfc757dbd3a

Observation 1168006f-b201-49e8-9707-34b34f9389b7 · outbound

This paper cites NIL: No-data Imitation Learning by Leveraging Pre-trained Video Diffusion Models.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation NIL: No-data Imitation Learning by Leveraging Pre-trained Video Diffusion Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:27.962933Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:27.962933Z digest=sha256:da14469447c40f38992f9694c6228a85ec81df122eccd3f873439934434dc760

Observation 600e7f26-f3fa-4eeb-b067-daa256352da9 · outbound

This paper cites Evaluating multiple object tracking performance: the clear mot metrics.EURASIP Journal on Image and Video Processing, 2008:1–10, 2008.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Evaluating multiple object tracking performance: the clear mot metrics.EURASIP Journal on Image and Video Processing, 2008:1–10, 2008

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:28.058069Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:28.058069Z digest=sha256:1cd7a2bdba825df47016a914dae1e851bcfcad3930d829e821cac2109a11bfd7

Observation e47415ba-09ac-47d7-9104-881ef902dfbf · outbound

This paper cites Gen2Act: Human Video Generation in Novel Scenarios enables Generalizable Robot Manipulation.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Gen2Act: Human Video Generation in Novel Scenarios enables Generalizable Robot Manipulation

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:28.179162Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:28.179162Z digest=sha256:7fca7de8d094ed23dfb036bbb40f3a66f35d29fee93effe9872bd93da96a211d

Observation df3c0f70-42af-4cd7-8b4c-70e5deb96fb6 · outbound

This paper cites Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:28.316171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:28.316171Z digest=sha256:477a4e64d9283c9104b15aa5b9565251fb644729c8036e09785907c88fd458df

Observation 957d04fc-d08e-4f4d-8e24-11439fcae46a · outbound

This paper cites Align your latents: High-resolution video synthesis with latent diffusion models.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Align your latents: High-resolution video synthesis with latent diffusion models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:28.432698Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:28.432698Z digest=sha256:2e3b23d79f666bceb1d23cb2f313ac899d2465ab6c0cbab87a4a33d3aecd96ad

Observation cdf151f3-4e91-4d33-902e-62499e8bd14a · outbound

This paper cites What Are You Doing? A Closer Look at Controllable Human Video Generation.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation What Are You Doing? A Closer Look at Controllable Human Video Generation

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:28.537052Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:28.537052Z digest=sha256:ea27edf088692ad3a98459cffecdf516bdc4acf20c2eb9cbea9f4eb3ca207a8a

Observation b2f98aa9-0a88-4e9a-9f9a-d44c9902f925 · outbound

This paper cites Deep video generation, prediction and completion of human action sequences.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Deep video generation, prediction and completion of human action sequences

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:28.702062Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:28.702062Z digest=sha256:1ee4a4402b8bfe25bfca21dd275b0a2816e7aae054d74697f080069f15b20c7d

Observation c227ce5b-e088-42c4-b987-b010f71ce6e2 · outbound

This paper cites Realtime multi-person 2d pose estimation using part affinity fields.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Realtime multi-person 2d pose estimation using part affinity fields

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:28.910813Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:28.910813Z digest=sha256:6fabffecd58ee37b0466faf5d9e729de097668ce11850a9a2cd8eda3b7b208a9

Observation fe4bec42-cecc-420a-8afe-80132b8a15a3 · outbound

This paper cites Everybody dance now.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Everybody dance now

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:29.012468Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:29.012468Z digest=sha256:80cf28c03c59a86e892ebb7e86cc5160c2dd90a4932c403bbddc562cb9d3d21b

Observation a11af6a5-570e-4f38-841b-c9cc3c9f17a0 · outbound

This paper cites A survey on evaluation of large language models.ACM transactions on intelligent systems and technology, 15(3):1–45, 2024.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation A survey on evaluation of large language models.ACM transactions on intelligent systems and technology, 15(3):1–45, 2024

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:29.121664Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:29.121664Z digest=sha256:4fda3dfdec21c4c91e2b18722e032d6aca777a02cdf1007252ab50e10d63f3a9

Observation 44b5296a-29b9-49a1-b2ca-279347e69f5a · outbound

This paper cites Idea23d: Collaborative lmm agents enable 3d model generation from interleaved multimodal inputs.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Idea23d: Collaborative lmm agents enable 3d model generation from interleaved multimodal inputs

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:29.276838Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:29.276838Z digest=sha256:0f7bfd5f08040dca1c6687b7351b34b274193c331b35dd758870ac0f9275ebfd

Observation 5ee3d724-3094-49fa-ae27-139edab772e8 · outbound

This paper cites Ultraman: Ultra-fast and high-resolution texture generation for 3d human reconstruction from a single image.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Ultraman: Ultra-fast and high-resolution texture generation for 3d human reconstruction from a single image

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:29.431038Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:29.431038Z digest=sha256:0d2bab0d74603a2a04bf0a507cad133198398f3fcdf4b5d27890d9a2fe2d3af2

Observation 0787fcbb-8a67-45a3-bacf-9a7be07ce9da · outbound

This paper cites Follow-Your-Canvas: Higher-Resolution Video Outpainting with Extensive Content Generation.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Follow-Your-Canvas: Higher-Resolution Video Outpainting with Extensive Content Generation

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:29.576107Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:29.576107Z digest=sha256:a8534fb6afe21d934dd9c17423292f024b3f3f93155941cdd86bd33b8ea3a8ea

Observation c6267999-d86e-4c8f-be8a-773bfa80472c · outbound

This paper cites Control3d: Towards controllable text-to-3d generation.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Control3d: Towards controllable text-to-3d generation

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:29.716317Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:29.716317Z digest=sha256:ca81f7cc7dd6842ce4bcd9ba7b07c5fe417106618307c70b3f9e637d27b9a865

Observation 02698d6c-1116-4086-82a6-f5fe84d2b928 · outbound

This paper cites Abo: Dataset and benchmarks for real-world 3d object understanding.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Abo: Dataset and benchmarks for real-world 3d object understanding

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:29.878018Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:29.878018Z digest=sha256:6b2b6b94086005f4e8205f807070ceca29d512b5f23455668e58e948715f81f7

Observation ef74ef35-882e-452e-8674-9c4390b03984 · outbound

This paper cites Arcface: Additive angular margin loss for deep face recognition.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Arcface: Additive angular margin loss for deep face recognition

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:29.971259Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:29.971259Z digest=sha256:772f585723bb436e6a78004e48a8134fecfdce24ee944d6e6af6c410a15c140e

Observation e60f2730-3902-4bdd-b390-6833cb5fff33 · outbound

This paper cites MagicPose: Realistic Human Poses and Facial Expressions Retargeting with Identity-aware Diffusion.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation MagicPose: Realistic Human Poses and Facial Expressions Retargeting with Identity-aware Diffusion

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:30.078011Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:30.078011Z digest=sha256:f374a6b2fc3e1777145c5d56622228c6896251c8b0b33e1ce2d530c5ef79bc9e

Observation 267dda78-9c90-4817-b5f6-b6137d6b873f · outbound

This paper cites Image quality assessment: Unifying structure and texture similarity.IEEE transactions on pattern analysis and machine intelligence, 44(5):2567–2581, 2020.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Image quality assessment: Unifying structure and texture similarity.IEEE transactions on pattern analysis and machine intelligence, 44(5):2567–2581, 2020

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:30.203373Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:30.203373Z digest=sha256:a1fa5d9995a9f71b97c26e55b907de0f12439cf5470044026a9b35a50195d953

Observation 3a8d5778-848c-46ab-8ac9-064bf19ec487 · outbound

This paper cites A survey of embodied ai: From simulators to research tasks.IEEE Transactions on Emerging Topics in Computational Intelligence, 6(2):230–244, 2022.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation A survey of embodied ai: From simulators to research tasks.IEEE Transactions on Emerging Topics in Computational Intelligence, 6(2):230–244, 2022

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:30.357678Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:30.357678Z digest=sha256:78682d38373f326ee76844901fc06d97486764acbe6e8274db06e4bc43acedd4

Observation 7e3136c8-8dd5-40fd-abf9-f1c597ca576e · outbound

This paper cites DreaMoving: A Human Video Generation Framework based on Diffusion Models.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation DreaMoving: A Human Video Generation Framework based on Diffusion Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:30.469507Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:30.469507Z digest=sha256:36825e12c425b3587796f955809a6fbe45da816c33da3587d1f0c54ed997b1bc

Observation a156db48-51d7-49dd-973d-9dd3e69652f7 · outbound

This paper cites Reconstructing Three-Dimensional Models of Interacting Humans.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Reconstructing Three-Dimensional Models of Interacting Humans

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:30.578026Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:30.578026Z digest=sha256:58003d03324502c00c0e198f82a635e412405ecfa11aea4cefdf388218c6bc2b

Observation 24d1dbb2-82b1-4de1-a26f-aa50f20d428f · outbound

This paper cites Iw-bench: Evaluating large multimodal models for converting image-to-web.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Iw-bench: Evaluating large multimodal models for converting image-to-web

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:30.706731Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:30.706731Z digest=sha256:3871f7008a3ea7f26082a9398daa48d7ae72a8b696312c3260a7aeb06295eb12

Observation a5f23d6e-e533-4821-bd2f-20bd4b8923be · outbound

This paper cites REPARO: Compositional 3D Assets Generation with Differentiable 3D Layout Alignment.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation REPARO: Compositional 3D Assets Generation with Differentiable 3D Layout Alignment

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:30.852424Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:30.852424Z digest=sha256:07881836c8da0aef480a1e7283e9c51f2951453b7fa69e243ffef9a15bac33a2

Observation 3a2db74a-06c7-4702-a933-436efdfd1d1a · outbound

This paper cites Controllable video generation with sparse trajectories.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Controllable video generation with sparse trajectories

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:31.000337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:31.000337Z digest=sha256:34b23a75edd13432c065d2b261e7a1a750b15a74272560193a18893b43e29e11

Observation 966235b4-d245-45a2-adbe-c8cce4d0a6cd · outbound

This paper cites Clipscore: A reference-free evaluation metric for image captioning.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Clipscore: A reference-free evaluation metric for image captioning

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:31.077064Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:31.077064Z digest=sha256:f91d7e331e414dc2c93187a78a1326e5dbc4d36add1d03ca9e064c2e86c94e1b

Observation 2ac61a93-afe7-4117-a40d-0a7215b9d292 · outbound

This paper cites Gans trained by a two time-scale update rule converge to a local nash equilibrium.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Gans trained by a two time-scale update rule converge to a local nash equilibrium

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:31.195892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:31.195892Z digest=sha256:dd739f9ea7a7ae3eebc037a056c719fb8a9ff8be68343bf1255532e664602a2f

Observation 19b36b3b-9e7e-4807-b977-68aadc3a9386 · outbound

This paper cites Video diffusion models.Advances in Neural Information Processing Systems, 35:8633–8646, 2022.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Video diffusion models.Advances in Neural Information Processing Systems, 35:8633–8646, 2022

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:31.330830Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:31.330830Z digest=sha256:57b59721fa0b3baa90eb9bd5580e122d43f108879e953dc434ee52aaa028e44b

Observation 3a338dd0-f641-4126-ad46-10370e14664c · outbound

This paper cites Animate anyone: Consistent and controllable image-to-video synthesis for character animation.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Animate anyone: Consistent and controllable image-to-video synthesis for character animation

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:31.448865Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:31.448865Z digest=sha256:e35b1d212183653abc477ec14dbd5c532eb42a81b1959a164e52d5e58fbe84b3

Observation 4bdc7f19-0ce8-4b8d-a1cd-666550d4a03d · outbound

This paper cites Make it move: controllable image-to-video generation with text descriptions.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Make it move: controllable image-to-video generation with text descriptions

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:31.551390Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:31.551390Z digest=sha256:20f586e4ab9296970d71b1f0e4d38eb89dd5a80af2ee141778b6ab5bc7130318

Observation 148131a0-b0cb-4168-8428-04eda8d52060 · outbound

This paper cites T2i-compbench: A comprehensive benchmark for open-world compositional text-to-image generation.Advances in Neural Information Processing Systems, 36:78723–78747, 2023.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation T2i-compbench: A comprehensive benchmark for open-world compositional text-to-image generation.Advances in Neural Information Processing Systems, 36:78723–78747, 2023

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:31.697935Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:31.697935Z digest=sha256:148d5df6f3087600ea8353ed473ed1b5968c703553f49dcd1b6875f1d74035b3

Observation 7c81c201-5c26-4245-a1ca-67ffbeecd471 · outbound

This paper cites Learning high fidelity depths of dressed humans by watching social media dance videos.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Learning high fidelity depths of dressed humans by watching social media dance videos

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:31.816062Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:31.816062Z digest=sha256:06df82027a4c1a286c36052ed60b136bb431545fe7e0f76e879ec380df11efc6

Observation d2601d8f-14fd-4923-a6b3-ebbbfd91843f · outbound

This paper cites Yolo by ultralytics.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Yolo by ultralytics

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:31.952185Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:31.952185Z digest=sha256:6edd038a9d04004de74dd18f76983923576794ca9e2e831f0f517970bbef12ca

Observation 7acefeae-4451-407b-95f2-9d10240b0820 · outbound

This paper cites Dreampose: Fashion image-to-video synthesis via stable diffusion.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Dreampose: Fashion image-to-video synthesis via stable diffusion

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:32.041557Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:32.041557Z digest=sha256:7b6d6535eea7cfc823cc85b71d70a52be6d552137807ad97aee5914184d924d1

Observation c32c5365-7c35-42a8-a1d0-3d0489a46a4a · outbound

This paper cites Text2video-zero: Text-to-image diffusion models are zero-shot video generators.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Text2video-zero: Text-to-image diffusion models are zero-shot video generators

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:32.189957Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:32.189957Z digest=sha256:bea2f9ea6000a8d5297ae8f5264ea224042f1de07417d865dc5f0cd85219fe57

Observation a70eeb23-462f-4a78-8dd7-4da9da819357 · outbound

This paper cites Harmony4d: A video dataset for in-the-wild close human interactions.Advances in Neural Information Processing Systems, 37:107270–107285, 2024.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Harmony4d: A video dataset for in-the-wild close human interactions.Advances in Neural Information Processing Systems, 37:107270–107285, 2024

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:32.335868Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:32.335868Z digest=sha256:821fb944e4ade615e690f7a9aca17f5083675de4453f6eb9aa7105e8e46d4167

Observation 27af0b82-7793-48ce-826a-09120dbb2c87 · outbound

This paper cites HunyuanVideo: A Systematic Framework For Large Video Generative Models.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation HunyuanVideo: A Systematic Framework For Large Video Generative Models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:32.441384Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:32.441384Z digest=sha256:c246218f2d834e396d225ac898821221a4c503ae7d393f1991ac012cf4e3bf0c

Observation 0f9e282f-c51b-434d-9c89-643ac36b716c · outbound

This paper cites World Knowledge from AI Image Generation for Robot Control.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation World Knowledge from AI Image Generation for Robot Control

Reference 38

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:39:44.390335Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:39:32.544278Z digest=sha256:951f3e758fb506a6d6427389a58ab5d3e095d17cc6a01d8f9c4ae1199c57c6ff

Observation e9fad175-df89-4ddd-9db3-16588d30dcb2 · outbound

This paper cites Collaborative video diffusion: Consistent multi-video generation with camera control.Advances in Neural Information Processing Systems, 37:16240–16271, 2024.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Collaborative video diffusion: Consistent multi-video generation with camera control.Advances in Neural Information Processing Systems, 37:16240–16271, 2024

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:32.667049Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:32.667049Z digest=sha256:f071f22b482e6569175b8cb4833bbd388d05bf06ca1adfb54865a6fa52a9ca2f

Observation 71d895b1-7f2a-435b-95e4-ea3e890f6c3f · outbound

This paper cites Interactive control of avatars animated with human motion data.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Interactive control of avatars animated with human motion data

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:32.784501Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:32.784501Z digest=sha256:1a9f5bdb6291c116d96d1430f245b623b7326afb7d079da7ec9d4b24947289ef

Observation 78117396-776c-4748-833d-7021096a11d7 · outbound

This paper cites Dispose: Disentangling pose guidance for controllable human image animation.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Dispose: Disentangling pose guidance for controllable human image animation

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:32.929080Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:32.929080Z digest=sha256:9960b1b957203ed27649eea2df53179d869f3194d1899133dddb1c6ff3e13ecd

Observation 1cea47e6-c866-4cac-9f51-982f745f81b1 · outbound

This paper cites Magicmotion: Controllable video generation with dense-to-sparse trajectory guidance.arXiv preprint arXiv:2503.16421, 2025.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Magicmotion: Controllable video generation with dense-to-sparse trajectory guidance.arXiv preprint arXiv:2503.16421, 2025

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:33.077966Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:33.077966Z digest=sha256:bf4f766231a936524fa5cd9b8210e08a8905120f75ac691554e644c5efe4629d

Observation d78f8ca3-ebb4-4682-8bc1-a86e74cae192 · outbound

This paper cites Ai choreographer: Music conditioned 3d dance generation with aist++.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Ai choreographer: Music conditioned 3d dance generation with aist++

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:33.241476Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:33.241476Z digest=sha256:691b41f48bc83f37c578289f8a27b37a46445e4c5f99abc891aee50acb9577b8

Observation 441d9249-c4c1-4ea9-ad25-9693c9ecd598 · outbound

This paper cites Lodge: A coarse to fine diffusion network for long dance generation guided by the characteristic dance primitives.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Lodge: A coarse to fine diffusion network for long dance generation guided by the characteristic dance primitives

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:33.415877Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:33.415877Z digest=sha256:f6665b181ece164ca161ab49f465ab34b4bf2450e2b39f079ce5e98b7e94a50e

Observation c8951d2e-04fe-4d05-b8e8-5767a0fe2ab1 · outbound

This paper cites Finedance: A fine-grained choreography dataset for 3d full body dance generation.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Finedance: A fine-grained choreography dataset for 3d full body dance generation

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:33.524998Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:33.524998Z digest=sha256:2143ed21607d10e105ba140f326e769c1057cb728955749421045b417f8ef24c

Observation ae3446fd-c7f8-491e-a946-312f88190ce6 · outbound

This paper cites Evaluation of text-to-video generation models: A dynamics perspective.Advances in Neural Information Processing Systems, 37:109790–109816, 2024.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Evaluation of text-to-video generation models: A dynamics perspective.Advances in Neural Information Processing Systems, 37:109790–109816, 2024

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:33.655440Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:33.655440Z digest=sha256:cec457d824008d08983a0d459d9453259732b9d2f529fee9cb55b8fa6b83c547

Observation eba97441-0e2c-47f4-b193-0346b6d5c10a · outbound

This paper cites Open-Sora Plan: Open-Source Large Video Generation Model.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Open-Sora Plan: Open-Source Large Video Generation Model

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:33.784524Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:33.784524Z digest=sha256:ef50616563a0a37ebe78e929ab781a16fbdd72c099f3cf6e15398ab71395c84a

Observation 561c3ced-6317-4037-96e2-0760abce90d3 · outbound

This paper cites Microsoft coco: Common objects in context.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Microsoft coco: Common objects in context

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:33.911487Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:33.911487Z digest=sha256:a7d30a1fe2c283fa1de34bd0787f896cac4e9b432803bd828b211eae54536e8a

Observation 5eecf7a7-32db-4588-a5a7-6b9a37951ff3 · outbound

This paper cites Rich: Robust implicit clothed humans reconstruction from multi-scale spatial cues.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Rich: Robust implicit clothed humans reconstruction from multi-scale spatial cues

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:34.080061Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:34.080061Z digest=sha256:e3a8b4f8512503ec40467ae3c0a42043fa26e1c713c5da4dfdb49fe46e4c436d

Observation 70b16e33-f959-4153-8202-71429bf008e1 · outbound

This paper cites Fr\'echet Video Motion Distance: A Metric for Evaluating Motion Consistency in Videos.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Fr\'echet Video Motion Distance: A Metric for Evaluating Motion Consistency in Videos

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:34.232980Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:34.232980Z digest=sha256:04d1e3b6843df0e38119f93879830d8a0d6596ac468ee7cf04a91b90526889eb

Observation 7c276a2c-8df3-43f3-ac91-74cc9c55b17b · outbound

This paper cites Evalcrafter: Benchmarking and evaluating large video generation models.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Evalcrafter: Benchmarking and evaluating large video generation models

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:34.359389Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:34.359389Z digest=sha256:072fc0b1497319d9334442193616d5afb815c988deccfa442f50822fac648f09

Observation 866f235c-51e3-4a8a-90fd-f5a545b88ed9 · outbound

This paper cites Fetv: A benchmark for fine-grained evaluation of open-domain text-to-video generation.Advances in Neural Information Processing Systems, 36:62352–62387, 2023.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Fetv: A benchmark for fine-grained evaluation of open-domain text-to-video generation.Advances in Neural Information Processing Systems, 36:62352–62387, 2023

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:34.496802Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:34.496802Z digest=sha256:b0c9ff1baa213e8c9511c21aab721cdfe0e2fdd49ffcb08dee14eda0b7c68ef8

Observation 55ef3d96-bbc4-445e-a9ff-abcd5bbf4852 · outbound

This paper cites Hota: A higher order metric for evaluating multi-object tracking.International Journal of Computer Vision, 129(2):548–578, 2021.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Hota: A higher order metric for evaluating multi-object tracking.International Journal of Computer Vision, 129(2):548–578, 2021

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:34.644160Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:34.644160Z digest=sha256:ca7cebf80005446f4e4bd67861409956e3ad02b69fd9e5189e83c97a21fc6d57

Observation 700ba6c6-ae3e-4c95-970e-c8e24237ad65 · outbound

This paper cites DreamActor-M1: Holistic, Expressive and Robust Human Image Animation with Hybrid Guidance.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation DreamActor-M1: Holistic, Expressive and Robust Human Image Animation with Hybrid Guidance

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:34.790939Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:34.790939Z digest=sha256:145cec17d093dcb77752ecc04f6dbca3d625d43ec094d12526bc47732c0202c7

Observation 2eecacde-9246-4bbe-862e-d50b89449f80 · outbound

This paper cites Notice of removal: Videofusion: Decomposed diffusion models for high-quality video generation.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Notice of removal: Videofusion: Decomposed diffusion models for high-quality video generation

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:34.892799Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:34.892799Z digest=sha256:547d69f7a00060d4dfe2a48d5e06214208029524032c1e654d0c9b665563bdf9

Observation bfc92c3e-e5f4-4101-aed4-172674a16189 · outbound

This paper cites Follow your pose: Pose-guided text-to-video generation using pose-free videos.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Follow your pose: Pose-guided text-to-video generation using pose-free videos

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:35.023802Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:35.023802Z digest=sha256:435381b5cbfc7d7eb9e8b2a78e7aef481ad7500ef35c3254e5f4c8713efe21a3

Observation da0f7099-f544-417d-8f8a-38e68d11bd91 · outbound

This paper cites Foundation models for video understanding: A survey.Authorea Preprints, 2024.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Foundation models for video understanding: A survey.Authorea Preprints, 2024

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:35.160080Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:35.160080Z digest=sha256:9b6fc7f2574b6d9248636624729eeb1fe198c578b3d2161e8c9efd9a86269968

Observation 9a1ae84c-733b-48a2-bfd4-2a7e08e13cfb · outbound

This paper cites Synergy and Synchrony in Couple Dances.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Synergy and Synchrony in Couple Dances

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:35.320236Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:35.320236Z digest=sha256:1d37fbb96dc9c29ad07d017d76f9a41fa7712460be23e326f38644099b961bda

Observation b9996c06-4d85-4b26-a5f0-faf946da36cf · outbound

This paper cites Benchmarking counterfactual image generation.Advances in Neural Information Processing Systems, 37:133207–133230, 2024.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Benchmarking counterfactual image generation.Advances in Neural Information Processing Systems, 37:133207–133230, 2024

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:50.790830Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:39:35.420035Z digest=sha256:895f1a2a31ae6c1b44173390bf06b2da689449b250083f46dfad7c44d99305f8

Observation 1d08ed28-1e54-4507-bcb4-c3558f9515f9 · outbound

This paper cites Efficient motion weighted spatio-temporal video ssim index.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Efficient motion weighted spatio-temporal video ssim index

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:50.621763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:39:35.511112Z digest=sha256:ba54505ecef328a6fc3ed2f498898c60831ac8f059769304d11179a16121df81

Observation 599b29c2-6318-4b10-8bb1-c61fe0763ba6 · outbound

This paper cites Sora: Creating video from text.https://openai.com/sora, 2024.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Sora: Creating video from text.https://openai.com/sora, 2024

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:50.415565Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:39:35.638475Z digest=sha256:f3ef63a1652c0f4ae7595e4ccf23ef8ed1797a9a95e6d9c7651dd4c7cf75601f

Observation 1c2473c0-16db-4e18-9f5e-9aa1cc2a423a · outbound

This paper cites ControlNeXt: Powerful and Efficient Control for Image and Video Generation.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation ControlNeXt: Powerful and Efficient Control for Image and Video Generation

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:35.757677Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:35.757677Z digest=sha256:2834e6b9660edcfc1a96ccfa3a439b7f8ba0792dde9df8d1bbe5994a174e2ea2

Observation 7c9b63c7-f018-40b3-a9ba-7c2120a5f46a · outbound

This paper cites Sdxl: Improving latent diffusion models for high-resolution image synthesis.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Sdxl: Improving latent diffusion models for high-resolution image synthesis

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:50.202694Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:39:35.914744Z digest=sha256:594ec8f99e6ab533763d460f46fe9199b3417182aa55c5177e4111ca5404af3c

Observation fe34d260-c08b-4195-ba96-e4d7d70a2aec · outbound

This paper cites DreamFusion: Text-to-3D using 2D Diffusion.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation DreamFusion: Text-to-3D using 2D Diffusion

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:36.051233Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:36.051233Z digest=sha256:63c5544630ebaa2c7bf331bc882bdcbff83f14585ccda88e4a10bc514b7a5631

Observation 8174cd73-7e13-4a04-9ccb-9620d9d92d33 · outbound

This paper cites WorldSimBench: Towards Video Generation Models as World Simulators.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation WorldSimBench: Towards Video Generation Models as World Simulators

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:36.194252Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:36.194252Z digest=sha256:eef4f7907d09eeee405fe4890f597bc219047942addde8c38995c55f9baebb52

Observation a2ad7b24-9803-44a6-8f01-9e7a52255402 · outbound

This paper cites Performance measures and a data set for multi-target, multi-camera tracking.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Performance measures and a data set for multi-target, multi-camera tracking

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:49.977423Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:39:36.300010Z digest=sha256:0983bf851433fb43dd11c242edf0e0397770c9e4f967f335bd100ddfe3258e94

Observation 986d20ab-c7c7-4611-8eeb-ad05a8af723f · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation High-resolution image synthesis with latent diffusion models

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:36.412598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:36.412598Z digest=sha256:c08607f904da6117c97a1df956422b991768697d0da6e9e8836fa8ef2453fd53

Observation f6261cbb-7924-47d6-87df-5d0653f45bff · outbound

This paper cites Motion-i2v: Consistent and controllable image-to- video generation with explicit motion modeling.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Motion-i2v: Consistent and controllable image-to- video generation with explicit motion modeling

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:49.824616Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:39:36.511573Z digest=sha256:0d5401e6b6c08a4e758f66743ff569154c9dce3bdb3e03cd0afa43e52f83331c

Observation 682b075f-1a07-4905-a8b1-4919e9eb09c4 · outbound

This paper cites Make-A-Video: Text-to-Video Generation without Text-Video Data.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Make-A-Video: Text-to-Video Generation without Text-Video Data

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:36.624194Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:36.624194Z digest=sha256:d5aa9954c13447ecaa9a3cfe6d588b6b6dffa5a93dac6f969f40892a5c9bb636

Observation fe4084ab-4cf7-48a7-9c17-4e1a95b2bb29 · outbound

This paper cites Llama Learns to Direct: DirectorLLM for Human-Centric Video Generation.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Llama Learns to Direct: DirectorLLM for Human-Centric Video Generation

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:36.773038Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:36.773038Z digest=sha256:19f1913ec1c4f580da87246fd26d837198a6beb43a89aa9a90a8c29762d807fd

Observation a735f50a-d9d7-44ff-984d-281da8418352 · outbound

This paper cites Transnet v2: An effective deep network architecture for fast shot transition detection.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Transnet v2: An effective deep network architecture for fast shot transition detection

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:49.631466Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:39:36.918213Z digest=sha256:669340705132d01cf63bae2310da5cae2e6b7fe68a410f24730f029a8a6a61a6

Observation 295ef70f-8597-41db-824d-ae06ac631d1a · outbound

This paper cites T2V-CompBench: A Comprehensive Benchmark for Compositional Text-to-video Generation.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation T2V-CompBench: A Comprehensive Benchmark for Compositional Text-to-video Generation

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:37.046057Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:37.046057Z digest=sha256:28535047790cd2b4f0bca6850df2e946dca7eb9993717cd5130649643b1cf317

Observation 17dab6e0-1d7f-4180-93a1-3b9ef8666ec9 · outbound

This paper cites Journeydb: A benchmark for generative image understanding.Advances in neural information processing systems, 36:49659–49678, 2023.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Journeydb: A benchmark for generative image understanding.Advances in neural information processing systems, 36:49659–49678, 2023

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:37.148856Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:37.148856Z digest=sha256:a99d6f08191085054104a19c248c950f9faa7692d936aeea9e38fd8088abc071

Observation 0ab2f963-912a-410f-b6ef-81fe8dfa4e11 · outbound

This paper cites DRiVE: Diffusion-based Rigging Empowers Generation of Versatile and Expressive Characters.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation DRiVE: Diffusion-based Rigging Empowers Generation of Versatile and Expressive Characters

Reference 74

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:39:44.064186Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:39:37.284377Z digest=sha256:c579f072b5619d3abc45f6f6a6df6810b956d97ae7f6ec0ce02c406ec53f995d

Observation a8539f1c-a8e2-4549-adf7-af82da83321e · outbound

This paper cites Beyond talking–generating holistic 3d human dyadic motion for communication.International Journal of Computer Vision, 133(5):2910–2926, 2025.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Beyond talking–generating holistic 3d human dyadic motion for communication.International Journal of Computer Vision, 133(5):2910–2926, 2025

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:49.423338Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:39:37.406670Z digest=sha256:210f5b774ac304a71658113dde2a32d798e8616c9ec60a18a43164288a6c288d

Observation 462060f9-a4b0-4549-9216-66fbc861a471 · outbound

This paper cites Video understanding with large language models: A survey.IEEE Transactions on Circuits and Systems for Video Technology, 2025.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Video understanding with large language models: A survey.IEEE Transactions on Circuits and Systems for Video Technology, 2025

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:37.503517Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:37.503517Z digest=sha256:ba09562022470b06a02913730da2d25bff438530496e085bdac895b855655efc

Observation 49039b20-9780-4dcf-88f4-a11b31bd5e09 · outbound

This paper cites Human Motion Diffusion Model.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Human Motion Diffusion Model

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:37.631482Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:37.631482Z digest=sha256:b5a2273b38c7c0c0562657fc9ee9d105102dd64fabb53443bc46d6f10c936bab

Observation 3db3c4ac-a43b-4926-b51d-162747b2aa91 · outbound

This paper cites EMO2: End-Effector Guided Audio-Driven Avatar Video Generation.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation EMO2: End-Effector Guided Audio-Driven Avatar Video Generation

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:37.744863Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:37.744863Z digest=sha256:4db043a2b770711cc916028e4c90c6511656421ad265138bbec768313891de33

Observation 24dcd086-b1d8-49c5-b27f-026e69609a5c · outbound

This paper cites StableAnimator: High-Quality Identity-Preserving Human Image Animation.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation StableAnimator: High-Quality Identity-Preserving Human Image Animation

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:37.877008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:37.877008Z digest=sha256:fd7addfed43651426b9d2720a9177f97cfd20c254a5f67663d9745d15b9450ce

Observation 534ead77-e2ec-49f6-9f7f-096e10dc33a1 · outbound

This paper cites Fvd: A new metric for video generation.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Fvd: A new metric for video generation

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:38.009818Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:38.009818Z digest=sha256:be085a283127cdd270f9f0d55e3aebf25fe3a3b5044f2fd770683f59d4794851

Observation 52b01249-347e-4aab-990a-adfe2027572e · outbound

This paper cites Articulated mesh animation from multi-view silhouettes.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Articulated mesh animation from multi-view silhouettes

Reference 81

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:49.170717Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:39:38.139105Z digest=sha256:eecd8e0a36ea57d82bf85171ffc8239de4f7fdc4aaa567f6f1c7a32543022f4f

Observation dfd18d8f-71fe-406c-b47d-1988373d0415 · outbound

This paper cites This&That: Language-Gesture Controlled Video Generation for Robot Planning.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation This&That: Language-Gesture Controlled Video Generation for Robot Planning

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:38.240073Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:38.240073Z digest=sha256:0a57af9b463b45ecbae6dbd1c67b4e5ad3b510cadc7306ad6efd13105088cf15

Observation e9cfa681-e7ca-4641-be4d-fa5de2ea8ff4 · outbound

This paper cites COVE: Unleashing the Diffusion Feature Correspondence for Consistent Video Editing.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation COVE: Unleashing the Diffusion Feature Correspondence for Consistent Video Editing

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:38.352118Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:38.352118Z digest=sha256:e40c0741901f816528ef06783286b61db056d97b5e0799047e5982e873a5c2e8

Observation e864da19-96b1-4111-ab4c-20661dfd5966 · outbound

This paper cites Taming Rectified Flow for Inversion and Editing.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Taming Rectified Flow for Inversion and Editing

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:38.465996Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:38.465996Z digest=sha256:dfe4f0c1f148a5c8c79f7a58ac2780d9bfd1d0e55d9320cf32e79dc021aab145

Observation 565a5ffb-4f3d-4970-8bb6-9b8050c1f3fa · outbound

This paper cites VividPose: Advancing Stable Video Diffusion for Realistic Human Image Animation.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation VividPose: Advancing Stable Video Diffusion for Realistic Human Image Animation

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:38.608579Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:38.608579Z digest=sha256:720ff4837d8e2e1def24b8aa129f54282eb342809dbb70f5d5ccdda29d9cb8d3

Observation 6119652b-97ef-4568-b9a7-3cfbd9be9051 · outbound

This paper cites Disco: Disentangled control for realistic human dance generation.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Disco: Disentangled control for realistic human dance generation

Reference 86

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:48.943110Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:39:38.726642Z digest=sha256:948110f0fa1e61df3c46aa58ea9fc9c825f3d57c1ae5a7ed08465a8fdf9477cc

Observation 65f317fb-c8d2-4a1c-beca-ac1f96bba92e · outbound

This paper cites Unianimate: Taming unified video diffusion models for consistent human image animation.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Unianimate: Taming unified video diffusion models for consistent human image animation

Reference 87

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:48.700278Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:39:38.845741Z digest=sha256:18f420c6910fee33e3c63a17e79b26aeefd81a5c2f08262d2a36ba5b02740d09

Observation faf8ef46-0828-4759-a6f9-a773ad75e8d1 · outbound

This paper cites Unianimate-dit: Human image animation with large-scale video diffusion transformer.arXiv preprint arXiv:2504.11289, 2025.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Unianimate-dit: Human image animation with large-scale video diffusion transformer.arXiv preprint arXiv:2504.11289, 2025

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:38.951240Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:38.951240Z digest=sha256:a0fdcd1fcf06c63bc1198e9cc097bfcb762469f2f58ded4c175b05f0f0e21f17

Observation c8bf9644-9762-4233-9e2e-7c962d2e1edb · outbound

This paper cites Instructavatar: Text-guided emotion and motion control for avatar generation.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Instructavatar: Text-guided emotion and motion control for avatar generation

Reference 89

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:48.495127Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:39:39.037038Z digest=sha256:7fa1e0fc2f6d74228728b2a3988b169831e6dbbe114ac4f4d7882695c359aafe

Observation e6bc5f26-51b7-48df-8410-96fa349b471d · outbound

This paper cites Bovik, H.R.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Bovik, H.R

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:39.157902Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:39.157902Z digest=sha256:db4006294f6395f05ce36898ab513cdfea52add1664facac6c66fe5f1db2584f

Observation 3ece665e-7fe3-49eb-9d56-486174428730 · outbound

This paper cites Humanvid: Demystifying training data for camera-controllable human image animation.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Humanvid: Demystifying training data for camera-controllable human image animation

Reference 91

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:48.285878Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:39:39.307743Z digest=sha256:60551367600ba01343a450ff97acb3dedaefeae9d560c820fddec9475b90d232

Observation a2ee6168-693d-40cb-8685-840a0cec5c10 · outbound

This paper cites Multi- identity human image animation with structural video diffusion.arXiv preprint arXiv:2504.04126, 2025.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Multi- identity human image animation with structural video diffusion.arXiv preprint arXiv:2504.04126, 2025

Reference 92

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:39.404847Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:39.404847Z digest=sha256:7196f574f503c507f74386bab81db07a633a8ec2656d8a563807cb455d5b2f81

Observation 4a53d9b5-b2d4-4d2c-80f0-ecc04d5e0eda · outbound

This paper cites MotionCtrl: A Unified and Flexible Motion Controller for Video Generation.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation MotionCtrl: A Unified and Flexible Motion Controller for Video Generation

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:39.527828Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:39.527828Z digest=sha256:d6b8c94ca3bcfc8d5fc1e6252783baf801af7c47adf265ddb364e3d7a4a48b30

Observation 02e6b24b-ea57-4867-9607-8d45d5c52528 · outbound

This paper cites Easyanimate: A high-performance long video generation method based on transformer architecture.arXiv preprint arXiv:2405.18991, 2024.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Easyanimate: A high-performance long video generation method based on transformer architecture.arXiv preprint arXiv:2405.18991, 2024

Reference 94

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:39.651709Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:39.651709Z digest=sha256:32616d89ae73a2cee424edc3295846fab4ec44713f3c81fdea6f28c742b5f5bf

Observation 547a547b-feb7-4de0-a803-018f84f6dd5f · outbound

This paper cites Xagen: 3d expressive human avatars generation.Advances in Neural Information Processing Systems, 36, 2024.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Xagen: 3d expressive human avatars generation.Advances in Neural Information Processing Systems, 36, 2024

Reference 95

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:48.028384Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:39:39.799853Z digest=sha256:1e924d7c856a24b0038e20741374d1131f19602555be2607a4fafeaaf677745a

Observation 85fc7bfc-08dc-4b08-9ab8-cd81ff7f3f8f · outbound

This paper cites Magicanimate: Temporally consistent human image animation using diffusion model.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Magicanimate: Temporally consistent human image animation using diffusion model

Reference 96

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:47.844906Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:39:39.935004Z digest=sha256:75c49de2e6ab1a15035f7c9446511797178786c995c299c4b5048809150a2b74

Observation af62cdba-6000-43af-b74d-f697c5675f72 · outbound

This paper cites Magicanimate: Temporally consistent human image animation using diffusion model.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Magicanimate: Temporally consistent human image animation using diffusion model

Reference 97

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:40.070249Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:40.070249Z digest=sha256:7a628e0d643401eaf0b45096bed13192180649e765cfc37ccbf9f30aab85b344

Observation 50f1e59c-f42d-4244-90e8-cc3913e8dec7 · outbound

This paper cites Human motion video generation: A survey.Authorea Preprints, 2024.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Human motion video generation: A survey.Authorea Preprints, 2024

Reference 98

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:47.723646Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:39:40.182270Z digest=sha256:46db32ed7704ec7a2f4479425edfdce174acd0ef46ce2e3753813d2546fe9214

Observation 2a91e9a4-748e-4ad8-b9f0-f62fe7a174fb · outbound

This paper cites Towards Multiple Character Image Animation Through Enhancing Implicit Decoupling.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Towards Multiple Character Image Animation Through Enhancing Implicit Decoupling

Reference 99

Resolution
unresolved
no resolver link, observed 2026-08-07T14:39:40.295999Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:39:40.295999Z digest=sha256:4aeb64c25cfbdbaf12ad7c5888d98afdac20ba437b50b535818919cbc5ca6799

Observation 6c5edda7-7994-4a5f-aa32-a100e15d477b · outbound

This paper cites Video quality assessment via gradient magnitude similarity deviation of spatial and spatiotemporal slices.

DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation Video quality assessment via gradient magnitude similarity deviation of spatial and spatiotemporal slices

Reference 100

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:39:47.474740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:39:40.471781Z digest=sha256:937afa20270950790d0b9754a9658995e3a24410ea40b122da79f34c5527025c

Pith citing papers

Observation fc12339c-6837-4ef5-b3dc-a834be500af1 · inbound

Bind-Your-Avatar: Multi-Talking-Character Video Generation with Dynamic 3D-mask-based Embedding Router cites this paper.

Bind-Your-Avatar: Multi-Talking-Character Video Generation with Dynamic 3D-mask-based Embedding Router DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-15T18:29:40.891406Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T18:29:40.891406Z digest=sha256:ffb0c66d05e4ef441c629674ffe9b90bf91a87c3feb0a7a67fce77f92948780c

Observation 5f61ef65-53ed-4d52-ac84-5f8dcce9fa77 · inbound

Go to Zero: Towards Zero-shot Motion Generation with Million-scale Data cites this paper.

Go to Zero: Towards Zero-shot Motion Generation with Million-scale Data DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T18:53:27.675148Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:53:27.675148Z digest=sha256:fb9f68c4c613a556c312728953a5168ac8c43b3de8739dbac02ea29cc5425269

Observation f3b5af5a-7f4e-483f-9830-8d65451aa8f7 · inbound

Animate-X++: Universal Character Image Animation with Dynamic Backgrounds cites this paper.

Animate-X++: Universal Character Image Animation with Dynamic Backgrounds DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-05T21:09:08.459299Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T21:09:08.459299Z digest=sha256:eecbd76bc00cec48b131c9de613ac8c48bb19158593ecec81af8a0777773959e

Observation 62c4e2ef-c1ee-42c1-88c8-37f3ddddffe0 · inbound

SocialDirector: Training-Free Social Interaction Control for Multi-Person Video Generation cites this paper.

SocialDirector: Training-Free Social Interaction Control for Multi-Person Video Generation DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-12T07:21:23.832707Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-05-12T03:28:30.471533Z digest=sha256:bdadef8c3653008ba49eaf50ad61baf93c637b40872cf75e398c51cd08be457a

Observation b5a0b152-3b45-4e4e-8e3e-791cba101049 · inbound

SCAIL-2: Unifying Controlled Character Animation with End-to-End In-Context Conditioning cites this paper.

SCAIL-2: Unifying Controlled Character Animation with End-to-End In-Context Conditioning DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation

Reference 28

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T05:17:40.172362Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-06-27T13:21:00.443497Z digest=sha256:c81805530a6a1dd447c0dd26b324945a554e446f0d29dc7cc180b5af5402a81e

Observation c0883a40-a918-4dd6-b485-326d8bfe23f3 · inbound

LogiShot: Logically Coherent Cross-Shot Video Generation cites this paper.

LogiShot: Logically Coherent Cross-Shot Video Generation DanceTogether! Identity-Preserving Multi-Person Interactive Video Generation

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-14T04:27:58.264034Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:27:58.264034Z digest=sha256:252049c925d3303be0a1032acf64320c2ee6645e31824c636902e48558391ae5