Pith. sign in

Paper Citation Record · LEDGER

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation

As of 7 August 2026, this Paper Citation Record lists 53 of 53 outbound references and 7 inbound Pith citation observations for arXiv:2506.08797.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.08797 v3

Coverage vector

measured 53 of 53 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:07:15.192008Z

measured 60 of 60 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 7 of 7 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T15:09:51.916002Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-16T23:31:21.996122Z

Reference resolution

53 of 53 outbound references displayed

  • verified exact1
  • verified fuzzy10
  • unresolved42
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation b0ca87d1-4796-4256-ad53-f730ee11e621 · outbound

This paper cites Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:14.990221Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:14.990221Z digest=sha256:dad5bf1b692815148315dfec899e046910150e460b52aa5a7d1d8bdf05cbc269

Observation ae459ba1-010d-4845-8394-1058f9fdecb7 · outbound

This paper cites Caron, H.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Caron, H

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:07:16.048518Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:07:14.994895Z digest=sha256:826d393f98c337259e79e3d6ffed863a5393b74ded47069bc2fb88e38b39f53d

Observation 0220a26c-8c11-4f9c-812a-d556b4e0d301 · outbound

This paper cites Chang, Y.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Chang, Y

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:07:16.036510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:07:14.999349Z digest=sha256:23bea27f0f535599da8594d9f134495322c1a8deca18c1cd169de4a2fbea105d

Observation f97e95c5-d0b4-4ce1-91a3-0836e6316685 · outbound

This paper cites an unresolved cited work.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:07:16.024565Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:07:15.003301Z digest=sha256:5890a82581c1cb14c053d45c37321c90019e778f5a4ac070d5062bad9d5a543e

Observation 4a7a9c56-3416-4d80-bd3f-af0a6ade30e2 · outbound

This paper cites an unresolved cited work.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:07:16.012438Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:07:15.007770Z digest=sha256:0a5bd8fdfe98a13432c7db9a0c9cb7a49b477d5faadd28d4e69a3de525ced3fd

Observation b1686fa9-01dc-4ba4-91c5-e233b32bcb2f · outbound

This paper cites VLOGGER: Multimodal Diffusion for Embodied Avatar Synthesis.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation VLOGGER: Multimodal Diffusion for Embodied Avatar Synthesis

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:15.011660Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:15.011660Z digest=sha256:3489afde41d1eea4000652d04dbc7caf8deb5449ef92af7be69f8f86f6444137

Observation ae9fa597-04f9-4565-bd5e-b4e03bb22212 · outbound

This paper cites Esser, S.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Esser, S

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:15.016076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:15.016076Z digest=sha256:e892dd39c446f9207e24940a019ed5df5879cf38c92c7c5b7b0fdea2adb18d0b

Observation de855897-b19b-43bd-b91f-bdbd81174438 · outbound

This paper cites Re-HOLD: Video Hand Object Interaction Reenactment via adaptive Layout-instructed Diffusion Model.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Re-HOLD: Video Hand Object Interaction Reenactment via adaptive Layout-instructed Diffusion Model

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-08-07T05:07:15.702938Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:07:15.020125Z digest=sha256:def651d5a9ff19d2c37205b28b8a89e4c815b10dc4f09f043d2b7381378d6e5e

Observation 1bb3af1c-2a0f-4f5f-8cbe-f1d10a43893d · outbound

This paper cites Ghosh, R.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Ghosh, R

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:07:15.991435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:07:15.024580Z digest=sha256:64ca1a5b8e490880b5851c088f12818414a37cf92171dcf5218dc1969403327d

Observation 66442418-fbac-4938-96ff-d3afea193c17 · outbound

This paper cites Heusel, H.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Heusel, H

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:07:15.980094Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:07:15.028260Z digest=sha256:f93f4e2a7c5deaffb566485be01393ba28c19e80ac22fcc8107bd619e82e85ed

Observation 3eb2ed14-aaa6-4e03-952a-3acaaf760805 · outbound

This paper cites an unresolved cited work.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:07:15.968197Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:07:15.032053Z digest=sha256:d80f7037555e27e3aaec6f288078054918bc8374deeb8916b897762a9ebd5495

Observation 85e619ce-9d97-4571-8c6b-cad9d204c591 · outbound

This paper cites Animate Anyone 2: High-Fidelity Character Image Animation with Environment Affordance.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Animate Anyone 2: High-Fidelity Character Image Animation with Environment Affordance

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:15.035885Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:15.035885Z digest=sha256:f32c1183652940c441fefdc9c0f76d4bc1d472aeb0d00ead3f54529c50295b10

Observation 65ac84c3-2cb5-4cf1-9e4c-e237b7ff3b63 · outbound

This paper cites Huang, F.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Huang, F

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:07:15.956274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:07:15.039605Z digest=sha256:10ee2cee1c434936650d62cc333e28f4f40ceb71d1c7af916ba739180cefd642

Observation b60969b4-3581-4303-999a-c620a3b768f8 · outbound

This paper cites Sonic: Shifting Focus to Global Audio Perception in Portrait Animation.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Sonic: Shifting Focus to Global Audio Perception in Portrait Animation

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:15.043271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:15.043271Z digest=sha256:827820d4f623afa50329e90548b983fd80aab87b7326dc78fc34fadb273c7c92

Observation 09c495b4-da43-45ba-8c73-a0fa6310d9ba · outbound

This paper cites Jiang, Z.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Jiang, Z

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:07:15.944542Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:07:15.047279Z digest=sha256:6467d10f69f623ca95f78cccf8ce9742630e75ebe6e433cfdcf8b8d6c05e9351

Observation fbc50b40-7c98-4fa7-864b-389ee098231c · outbound

This paper cites Jiang, Z.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Jiang, Z

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:07:15.932510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:07:15.050852Z digest=sha256:007f2af0c9b9c565d29975ed347e7514f7d184b6a317cbed791dc4ab55e53af8

Observation bd07486d-a18f-48d6-ac7a-5c469db53935 · outbound

This paper cites VACE: All-in-One Video Creation and Editing.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation VACE: All-in-One Video Creation and Editing

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:15.054368Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:15.054368Z digest=sha256:cacfcc6a708a59fce4478eca274f6c4b0a038b7d5a9b1e7e4bf91b201ebb473d

Observation ee91f300-9ad8-4abf-8bb7-d6cea612440f · outbound

This paper cites Auto-Encoding Variational Bayes.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Auto-Encoding Variational Bayes

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:15.058186Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:15.058186Z digest=sha256:76548aa757d278879177e1ec004703def64bdc21a56b516b259cd65b966c8a64

Observation 54dd8474-957c-45fe-8996-b5ec7fd95e48 · outbound

This paper cites HunyuanVideo: A Systematic Framework For Large Video Generative Models.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation HunyuanVideo: A Systematic Framework For Large Video Generative Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:15.061950Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:15.061950Z digest=sha256:8b465a46ebc35f23b34a38effb181d7adb664ef0e2447c8d2ce854761cbe0ecc

Observation ec9eb658-dbd3-48f9-a893-5ca6d3a0608a · outbound

This paper cites an unresolved cited work.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:07:15.920091Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:07:15.065656Z digest=sha256:02941e8f9063138c6f5aaa652b0a73f28be1123381b4e9a62aa757549eca06ec

Observation cfd25029-0362-4163-8bca-e628478d122c · outbound

This paper cites CyberHost: Taming Audio-driven Avatar Diffusion Model with Region Codebook Attention.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation CyberHost: Taming Audio-driven Avatar Diffusion Model with Region Codebook Attention

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:15.069517Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:15.069517Z digest=sha256:85da549875f01380904b00d75aa47fb1ff0f375c437ef5e5ad28f097481ea19e

Observation b23d7f39-b1a6-4e11-9b16-4b50e8928c45 · outbound

This paper cites OmniHuman-1: Rethinking the Scaling-Up of One-Stage Conditioned Human Animation Models.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation OmniHuman-1: Rethinking the Scaling-Up of One-Stage Conditioned Human Animation Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:15.073459Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:15.073459Z digest=sha256:52b62dd2d09e82d29d560bccf2d3eac5282a94ea80b99064ae4760956aaf26d8

Observation 0d936f60-8589-456f-872e-634135e14145 · outbound

This paper cites Flow Matching for Generative Modeling.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Flow Matching for Generative Modeling

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:15.077554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:15.077554Z digest=sha256:46e47d8690baec98953d5f57268a36b44e178151019f6a01a4070892f1c3727e

Observation a5af6cab-0e8a-4acf-a7f5-623ac1baeb75 · outbound

This paper cites an unresolved cited work.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:07:15.906640Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:07:15.081439Z digest=sha256:7335ca9ef393217b1fe5a22ee292072a2675f632b286d6c26f23b80bd3a78ebc

Observation 7d5c20ba-b5c1-4483-86d3-30b4e8471066 · outbound

This paper cites an unresolved cited work.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:07:15.894523Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:07:15.085543Z digest=sha256:58ab2294b6d88383df391e691f6ffd00f9f48ef5bf8d5d7c3c02817f0fece5df

Observation 617c4b12-378c-488b-b0ef-02e5e9dd2683 · outbound

This paper cites McFee, C.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation McFee, C

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:07:15.882539Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:07:15.089605Z digest=sha256:67781d295bb4029b355660cfe55842d3118bf957458b2920ac5f0cc5896e0667

Observation e88431f2-74ff-4469-93e9-fb36edb08c7e · outbound

This paper cites MIMO: Controllable Character Video Synthesis with Spatial Decomposed Modeling.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation MIMO: Controllable Character Video Synthesis with Spatial Decomposed Modeling

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:15.094291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:15.094291Z digest=sha256:c771624603a25614ac1cc27ed470212942d81d9bba27420c90d751a979cebcc6

Observation 65e7f9bc-1091-4f8d-a439-60a3744b5faa · outbound

This paper cites an unresolved cited work.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Unresolved cited work

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:15.098106Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:15.098106Z digest=sha256:3442adf97a5c059d9637deec12570cfe4d032edd6d8839a623c88ce0f2888f74

Observation a4743763-47f5-4b1c-8470-485a02c0b32a · outbound

This paper cites ManiVideo: Generating Hand-Object Manipulation Video with Dexterous and Generalizable Grasping.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation ManiVideo: Generating Hand-Object Manipulation Video with Dexterous and Generalizable Grasping

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:15.101686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:15.101686Z digest=sha256:9ec2ffbd3c2c3a03abe10a4eff1f9f8ec75f0ecefa3b9d16fe63254bcd16ae95

Observation 4dcea8d2-cb25-4ec2-9951-010cb855c598 · outbound

This paper cites HOI-Diff: Text-Driven Synthesis of 3D Human-Object Interactions using Diffusion Models.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation HOI-Diff: Text-Driven Synthesis of 3D Human-Object Interactions using Diffusion Models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:15.105677Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:15.105677Z digest=sha256:9daeddb63302c2b0769d5e8b204abdd2619a2bf9da52f660aff27e320b310678

Observation 0ad89ee6-9fe6-4f38-a2cf-4f596193cb7c · outbound

This paper cites Radford, J.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Radford, J

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:07:15.870529Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:07:15.109313Z digest=sha256:d01cb413353a06828deca808114dae8f9af1db73765b6bacf94009b134c22cee

Observation 7bf6a82f-9aa0-4b4f-b799-aa2eb985cab8 · outbound

This paper cites SAM 2: Segment Anything in Images and Videos.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation SAM 2: Segment Anything in Images and Videos

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:15.112832Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:15.112832Z digest=sha256:39ee62b818ef4f3f6fa31dfaad6fbaeafe8f040745e7799fe39083cbefeacf60

Observation 31de9a2b-2e04-436a-b098-dda4e6952cce · outbound

This paper cites an unresolved cited work.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:07:15.858624Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:07:15.117225Z digest=sha256:b32dc070a34097bdd7b727bd9f028d81c39f307eea079e6ba7f457a0555d89e0

Observation fd3f0e28-2d5d-45df-9678-7c1b18453a12 · outbound

This paper cites an unresolved cited work.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Unresolved cited work

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:15.120920Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:15.120920Z digest=sha256:56ba6e6620666423e53e4700d8bdb80637809f76148fe58a9cb8e9d23ac35a51

Observation 484fe3bb-2ce6-4ff8-8570-ae5d2a8f3313 · outbound

This paper cites an unresolved cited work.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:07:15.838532Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:07:15.124595Z digest=sha256:6a13755af888287c0c5c266454ade7cab8ecd66dd9c6f4bfb98188c68527713f

Observation a04d0853-748d-49a4-bd63-e9d7f45dcbd0 · outbound

This paper cites StableAnimator: High-Quality Identity-Preserving Human Image Animation.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation StableAnimator: High-Quality Identity-Preserving Human Image Animation

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:15.128129Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:15.128129Z digest=sha256:2539a7aa73497c4957eb5a1593ae02233b008c1e1f9e462e89e43a7147fef3b1

Observation 6d1904a6-f7fe-4897-bd51-2227546b86b2 · outbound

This paper cites Towards Accurate Generative Models of Video: A New Metric & Challenges.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Towards Accurate Generative Models of Video: A New Metric & Challenges

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:15.131757Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:15.131757Z digest=sha256:84e9ee0d3ced78abfb76b09040b7ff2d906c8faea544a579b62428d8557c647d

Observation d2dca16d-e51f-45be-82a3-418c67c2a856 · outbound

This paper cites Wan: Open and Advanced Large-Scale Video Generative Models.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Wan: Open and Advanced Large-Scale Video Generative Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:15.135416Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:15.135416Z digest=sha256:ba0082f3e99f2012441c405f353c6e0aa781c8f97c6d55bd2643beb141346883

Observation eae3a003-2f1b-4816-b084-79848605e6c1 · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:15.139064Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:15.139064Z digest=sha256:855266313562923e87cbad52147a9a6d94aac5411dc5ca62104a1d9c8566eb13

Observation 4838ac87-d887-4cd9-ab43-df53cbe08d6f · outbound

This paper cites an unresolved cited work.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Unresolved cited work

Reference 40

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:07:15.827007Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:07:15.142718Z digest=sha256:6e88ba57f9889dc9e292a450e7d2d7821d00e3efd53f412c43e49f58698218d9

Observation 91ca9bd0-f838-4fa4-acdd-9ef2b09d78a0 · outbound

This paper cites UniAnimate: Taming Unified Video Diffusion Models for Consistent Human Image Animation.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation UniAnimate: Taming Unified Video Diffusion Models for Consistent Human Image Animation

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:15.146443Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:15.146443Z digest=sha256:a44c75af1d83e861fab2b60667989d9f2d5fb3f15ec0c0b6388d409979d49cf2

Observation dc0a1690-b236-4020-bb5d-04958a9d41c7 · outbound

This paper cites an unresolved cited work.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Unresolved cited work

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:15.150092Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:15.150092Z digest=sha256:f5dc73caa12afa5483a7f46188b2c38223bc4a88a1a3e8dcab5371e7783a5efa

Observation 4f40aa06-c7c2-4967-9623-69976fed45b0 · outbound

This paper cites an unresolved cited work.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Unresolved cited work

Reference 43

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:07:15.815546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:07:15.154172Z digest=sha256:ee3cb3eb37502c3c2ba1dc3ef46f2448b6e0e364b9ae5dc7b9ae594f8d8a5c8c

Observation fd352863-e67c-4f0d-854d-90d3ee8b8360 · outbound

This paper cites an unresolved cited work.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Unresolved cited work

Reference 44

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:07:15.803322Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:07:15.157826Z digest=sha256:ad6dcf08f7d949f6991aa0fb1275867bdddc18c5e364a6e5da92a33879a2f42a

Observation d82759de-ba97-4831-9f09-2ce941a4833e · outbound

This paper cites AnchorCrafter: Animate Cyber-Anchors Selling Your Products via Human-Object Interacting Video Generation.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation AnchorCrafter: Animate Cyber-Anchors Selling Your Products via Human-Object Interacting Video Generation

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:15.161431Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:15.161431Z digest=sha256:8600c2601266ffd901f8503eb63300caead9ec4605857f5ce6250cb28f240c97

Observation 433e54dc-d2aa-4faf-8152-6f221b269c38 · outbound

This paper cites an unresolved cited work.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Unresolved cited work

Reference 46

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:07:15.789992Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:07:15.165245Z digest=sha256:d79356cd35500ba852e7dae8a7527898a2eb6ea96038f4193d5d53f02c601139

Observation 84bebacb-0d3f-4e78-85fb-f212ed8f2ca5 · outbound

This paper cites HOI-Swap: Swapping Objects in Videos with Hand-Object Interaction Awareness.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation HOI-Swap: Swapping Objects in Videos with Hand-Object Interaction Awareness

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:15.168770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:15.168770Z digest=sha256:981246e94f6eea5fc3bdb56379a6c2d4947f5557802db013dfa0211672366a1b

Observation 81cf7fda-803e-4bbb-9e3c-101e60cac3e8 · outbound

This paper cites an unresolved cited work.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Unresolved cited work

Reference 48

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:07:15.778193Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:07:15.172606Z digest=sha256:8f748b68a8e3098afcfca0bde84aabd763b16afcfb9fc8f40d7296da219f7471

Observation 8083b526-4c7a-4a20-a453-1c66d4e64b98 · outbound

This paper cites an unresolved cited work.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Unresolved cited work

Reference 49

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:07:15.766481Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:07:15.176525Z digest=sha256:de12aee3712bc7939c6a5b4b734095b606dca715e2e5425fbe0ad3260bb16735

Observation 12113b3d-b607-4cd5-91dd-38834a3c3a54 · outbound

This paper cites Zhang, X.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Zhang, X

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:07:15.754349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:07:15.180425Z digest=sha256:19437260c565dfd2b7d489f415df143ca94138acf5eaa9a96716d427f1c38c02

Observation 81d175ab-acfe-476a-86f2-1be1d4a54a75 · outbound

This paper cites MimicMotion: High-Quality Human Motion Video Generation with Confidence-aware Pose Guidance.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation MimicMotion: High-Quality Human Motion Video Generation with Confidence-aware Pose Guidance

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:15.184611Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:15.184611Z digest=sha256:d757af4464de727c75b50aa7a95d1b2c00d64158e6a3d9c6c3778a595d43598b

Observation 8298ced2-5661-40d2-a24b-03c834d10b85 · outbound

This paper cites Allegro: Open the Black Box of Commercial-Level Video Generation Model.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Allegro: Open the Black Box of Commercial-Level Video Generation Model

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:15.188299Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:15.188299Z digest=sha256:a350f28fe4bea30485dd388059dad1705d653e5ad1dfe148773a2adb6108002b

Observation 9891aa66-5325-4f7d-b603-e7a25078fd66 · outbound

This paper cites an unresolved cited work.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Unresolved cited work

Reference 53

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:07:15.741878Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T05:07:15.192008Z digest=sha256:0b1ed92302beaa3b71bfe6853c2573cba29de4cebb011533ec09b439839da86d

Pith citing papers

Observation d4776e48-b2a1-4f33-9ac4-7b64aa05c48d · inbound

HOComp: Interaction-Aware Human-Object Composition cites this paper.

HOComp: Interaction-Aware Human-Object Composition HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T15:09:51.916002Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:09:51.916002Z digest=sha256:99f8d5f3e0414f12c770ade1328fa76206179776d48cf6f6637a9aebd7fb11c2

Observation 6520dd86-2b53-45dc-8825-bf9b92f07a76 · inbound

VHOI: Controllable Video Generation of Human-Object Interactions from Sparse Trajectories via Motion Densification cites this paper.

VHOI: Controllable Video Generation of Human-Object Interactions from Sparse Trajectories via Motion Densification HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-07-09T01:19:36.380554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-16T23:30:14.969895Z digest=sha256:72e59f2a68606dd5880dc96a168ed93b1b89ac7a36c675bb83862df67e321389

Observation 5cc0e35d-d677-4b28-ae37-737691ad3291 · inbound

CoMoVi: Co-Generation of 3D Human Motions and Realistic Videos cites this paper.

CoMoVi: Co-Generation of 3D Human Motions and Realistic Videos HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-07-09T01:19:36.380554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-16T13:43:26.460480Z digest=sha256:c4a2fa95aafa182072628e8dda81ceeed6308fbb7912dad478b6693e2c2327c5

Observation ccc26b66-cc67-4ae2-a977-bf55d8a440c0 · inbound

EchoTorrent: Towards Swift, Sustained, and Streaming Multi-Modal Video Generation cites this paper.

EchoTorrent: Towards Swift, Sustained, and Streaming Multi-Modal Video Generation HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation

Reference 73

Resolution
verified exact
arxiv_id, observed 2026-07-09T01:19:36.380554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-15T22:20:16.320171Z digest=sha256:f5c654ec53bfa9b000c4ce936c29dbac2cf771ea14a5c41ff97b9696efe8afc2

Observation d6adda24-10da-4a15-b2b2-8317664b43b4 · inbound

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model cites this paper.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.179232Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.179232Z digest=sha256:f3b75ef4d05a3d6e8a14305f038e2707d5d40968e7c36f97d8c3d5a9ce2bf62e

Observation 85627e22-73e7-4684-b5a1-bea08e89bdd5 · inbound

OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation cites this paper.

OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-07-09T01:19:36.380554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T15:09:02.727887Z digest=sha256:53c84cb4f898d3dbf2d8c320fd93db41519ad7e5b4d8bcc69c9c012a47c565e2

Observation 96207e89-b4cf-4340-a436-2da9dfdc30b4 · inbound

StreamHOI: Interaction-aware Temporal Memory Adaptation for Streaming HOI Video Generation cites this paper.

StreamHOI: Interaction-aware Temporal Memory Adaptation for Streaming HOI Video Generation HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-01T10:37:53.829539Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T10:37:53.829539Z digest=sha256:0dc6b8c4270b633e0dd57f0a4a2682c9c1191a601acbf30c06ba20da66290148