Pith. sign in

Paper Citation Record · LEDGER

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation

As of 19 August 2026, this Paper Citation Record lists 53 of 53 outbound references and 7 inbound Pith citation observations for arXiv:2506.08797.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.08797 v3

Coverage vector

measured 53 of 53 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:07:15.192008Z

measured 60 of 60 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 7 of 7 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T15:09:51.916002Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-16T23:31:21.996122Z

Reference resolution

53 of 53 outbound references displayed

  • verified exact1
  • verified fuzzy10
  • unresolved42
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation b0ca87d1-4796-4256-ad53-f730ee11e621 · outbound

This paper cites Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:14.990221Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:14.990221Z digest=sha256:a6b1af35151e97d5c3f5367ddf0ae11533cb5d8993901b460bdb5f94b8bd3529

Observation ae459ba1-010d-4845-8394-1058f9fdecb7 · outbound

This paper cites Caron, H.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Caron, H

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:07:16.048518Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T05:07:14.994895Z digest=sha256:11b870d860d8ffc5375eb063eea6fe1f4df5302041e304533a650c6bb6a4e1b9

Observation 0220a26c-8c11-4f9c-812a-d556b4e0d301 · outbound

This paper cites Chang, Y.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Chang, Y

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:07:16.036510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T05:07:14.999349Z digest=sha256:c9739b5b10a61daaa3599199e14e9e695bbdd829fa93ab652bf3ab025e5002ea

Observation f97e95c5-d0b4-4ce1-91a3-0836e6316685 · outbound

This paper cites an unresolved cited work.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:07:16.024565Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T05:07:15.003301Z digest=sha256:2d472c3a062eba4ddbc976a95c82a8ae58b175fa119d226477bcaae56206975e

Observation 4a7a9c56-3416-4d80-bd3f-af0a6ade30e2 · outbound

This paper cites an unresolved cited work.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:07:16.012438Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T05:07:15.007770Z digest=sha256:3e6fe92200ffdee6b959f65f8d6a3a33ca605c318a28fe5d267c6a9f4d5f566a

Observation b1686fa9-01dc-4ba4-91c5-e233b32bcb2f · outbound

This paper cites VLOGGER: Multimodal Diffusion for Embodied Avatar Synthesis.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation VLOGGER: Multimodal Diffusion for Embodied Avatar Synthesis

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:15.011660Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:15.011660Z digest=sha256:cb8ddc5fa3cb74ca4ed6250f6f073f4f123167e16c64f41292ba515096075b83

Observation ae9fa597-04f9-4565-bd5e-b4e03bb22212 · outbound

This paper cites Esser, S.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Esser, S

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:15.016076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:15.016076Z digest=sha256:51f8d227b1840777a3fb1672209c34a829ed5e2b595492d09d4008a3a8be4899

Observation de855897-b19b-43bd-b91f-bdbd81174438 · outbound

This paper cites Re-HOLD: Video Hand Object Interaction Reenactment via adaptive Layout-instructed Diffusion Model.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Re-HOLD: Video Hand Object Interaction Reenactment via adaptive Layout-instructed Diffusion Model

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-08-07T05:07:15.702938Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T05:07:15.020125Z digest=sha256:64f3df31f00851017e8ccf37e9c72618d164e8dcf5c22f8892807137c124d53a

Observation 1bb3af1c-2a0f-4f5f-8cbe-f1d10a43893d · outbound

This paper cites Ghosh, R.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Ghosh, R

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:07:15.991435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T05:07:15.024580Z digest=sha256:44e082101ba622453f4061e777b6c074a4446e4747d41ad07fbb53ff842f1e3e

Observation 66442418-fbac-4938-96ff-d3afea193c17 · outbound

This paper cites Heusel, H.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Heusel, H

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:07:15.980094Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T05:07:15.028260Z digest=sha256:eb18728d85a2ef223d082bdefa81bd7c3bc456b833c1e55b5ac2d60704718fad

Observation 3eb2ed14-aaa6-4e03-952a-3acaaf760805 · outbound

This paper cites an unresolved cited work.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:07:15.968197Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T05:07:15.032053Z digest=sha256:5008f25918a6cac79c551933a4167ce7e9f2976fe952ed0e31769aac44d520c1

Observation 85e619ce-9d97-4571-8c6b-cad9d204c591 · outbound

This paper cites Animate Anyone 2: High-Fidelity Character Image Animation with Environment Affordance.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Animate Anyone 2: High-Fidelity Character Image Animation with Environment Affordance

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:15.035885Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:15.035885Z digest=sha256:eb6dbe7a37b88064f88b0161f8e2113829adc31a2f47f6ea7114fe724042452d

Observation 65ac84c3-2cb5-4cf1-9e4c-e237b7ff3b63 · outbound

This paper cites Huang, F.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Huang, F

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:07:15.956274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T05:07:15.039605Z digest=sha256:e303e1ece22a430f042a81a0c9e73ac268df5798f0b2a11761e791922636403f

Observation b60969b4-3581-4303-999a-c620a3b768f8 · outbound

This paper cites Sonic: Shifting Focus to Global Audio Perception in Portrait Animation.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Sonic: Shifting Focus to Global Audio Perception in Portrait Animation

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:15.043271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:15.043271Z digest=sha256:276c24dad6be021cf1de87309e716edfbb553bef49d9b082a5b751b1380dd4c4

Observation 09c495b4-da43-45ba-8c73-a0fa6310d9ba · outbound

This paper cites Jiang, Z.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Jiang, Z

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:07:15.944542Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T05:07:15.047279Z digest=sha256:2ed02d7e328663367d666d228e07f8f844dd51721a41185b013a5a27eb53f409

Observation fbc50b40-7c98-4fa7-864b-389ee098231c · outbound

This paper cites Jiang, Z.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Jiang, Z

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:07:15.932510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T05:07:15.050852Z digest=sha256:2db5a0220f98c05e78dc26a790d6040704caba2cf139efc8177a8ee1e5be1629

Observation bd07486d-a18f-48d6-ac7a-5c469db53935 · outbound

This paper cites VACE: All-in-One Video Creation and Editing.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation VACE: All-in-One Video Creation and Editing

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:15.054368Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:15.054368Z digest=sha256:946b926181e19b5b85aa96fbe2f586f4842acbe1336ccf7c569285f577e84ec6

Observation ee91f300-9ad8-4abf-8bb7-d6cea612440f · outbound

This paper cites Auto-Encoding Variational Bayes.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Auto-Encoding Variational Bayes

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:15.058186Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:15.058186Z digest=sha256:d405e96a4f11a4c1c863b107c0517872a8c1bb6ca35abf243d891eeb6650c8f0

Observation 54dd8474-957c-45fe-8996-b5ec7fd95e48 · outbound

This paper cites HunyuanVideo: A Systematic Framework For Large Video Generative Models.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation HunyuanVideo: A Systematic Framework For Large Video Generative Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:15.061950Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:15.061950Z digest=sha256:7d3bbd53e5a012af6a419951a27a82d8ba1de1cc7e8ec8d2156c16815a1e67b4

Observation ec9eb658-dbd3-48f9-a893-5ca6d3a0608a · outbound

This paper cites an unresolved cited work.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:07:15.920091Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T05:07:15.065656Z digest=sha256:b938dc8401060cdedef747405ca444d7f300eb84e37064ddf2ed64c9efef64d4

Observation cfd25029-0362-4163-8bca-e628478d122c · outbound

This paper cites CyberHost: Taming Audio-driven Avatar Diffusion Model with Region Codebook Attention.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation CyberHost: Taming Audio-driven Avatar Diffusion Model with Region Codebook Attention

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:15.069517Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:15.069517Z digest=sha256:ad2bf35a1c98045ea8693f995f5a6f5d4c0fffbc74b75094e159bb232f5cf0ca

Observation b23d7f39-b1a6-4e11-9b16-4b50e8928c45 · outbound

This paper cites OmniHuman-1: Rethinking the Scaling-Up of One-Stage Conditioned Human Animation Models.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation OmniHuman-1: Rethinking the Scaling-Up of One-Stage Conditioned Human Animation Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:15.073459Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:15.073459Z digest=sha256:de39a99c5712b1c83e28a0e62bae839e6d8816803e6c1d45ad876fb09c107486

Observation 0d936f60-8589-456f-872e-634135e14145 · outbound

This paper cites Flow Matching for Generative Modeling.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Flow Matching for Generative Modeling

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:15.077554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:15.077554Z digest=sha256:9951951eaa02badfc3c8c958e85a3e36dfc2cbe6a05cbab7c35d434be057bd85

Observation a5af6cab-0e8a-4acf-a7f5-623ac1baeb75 · outbound

This paper cites an unresolved cited work.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:07:15.906640Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T05:07:15.081439Z digest=sha256:09e817aa7ef823253f22cbc015ea99cede2cb4fdec9885e3280ff4350d7902c4

Observation 7d5c20ba-b5c1-4483-86d3-30b4e8471066 · outbound

This paper cites an unresolved cited work.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:07:15.894523Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T05:07:15.085543Z digest=sha256:106f81d59a6e1c614be8e7693c319c73626d7956a328b96b6721011db1f64ed1

Observation 617c4b12-378c-488b-b0ef-02e5e9dd2683 · outbound

This paper cites McFee, C.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation McFee, C

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:07:15.882539Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T05:07:15.089605Z digest=sha256:3c00d0eadf90ef6b37b40f3ec026e347ed95aa3425d3ac10ec1df0c71652c35a

Observation e88431f2-74ff-4469-93e9-fb36edb08c7e · outbound

This paper cites MIMO: Controllable Character Video Synthesis with Spatial Decomposed Modeling.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation MIMO: Controllable Character Video Synthesis with Spatial Decomposed Modeling

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:15.094291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:15.094291Z digest=sha256:505c47264efed7d554e99563ddd6dd68b89de0d1650ec84f1554664cfba13d44

Observation 65e7f9bc-1091-4f8d-a439-60a3744b5faa · outbound

This paper cites an unresolved cited work.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Unresolved cited work

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:15.098106Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:15.098106Z digest=sha256:7fe0b9cb5b7f7ac1638ccd6adeb8ac0fea46dbb3ff995353c49a9f746e2c5fa7

Observation a4743763-47f5-4b1c-8470-485a02c0b32a · outbound

This paper cites ManiVideo: Generating Hand-Object Manipulation Video with Dexterous and Generalizable Grasping.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation ManiVideo: Generating Hand-Object Manipulation Video with Dexterous and Generalizable Grasping

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:15.101686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:15.101686Z digest=sha256:06e2eb9250898ba74088760b3bc241539a0d33ffda8b80e9b4305e78868a6469

Observation 4dcea8d2-cb25-4ec2-9951-010cb855c598 · outbound

This paper cites HOI-Diff: Text-Driven Synthesis of 3D Human-Object Interactions using Diffusion Models.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation HOI-Diff: Text-Driven Synthesis of 3D Human-Object Interactions using Diffusion Models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:15.105677Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:15.105677Z digest=sha256:940a910f408c91ea375a9fde9a84651da250b9bdd76e5b032eea2ec8c13dde3b

Observation 0ad89ee6-9fe6-4f38-a2cf-4f596193cb7c · outbound

This paper cites Radford, J.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Radford, J

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:07:15.870529Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T05:07:15.109313Z digest=sha256:07872f0bc1c0fb4b550343db044ae1d513692dcfa20c3a5aead556d89519c724

Observation 7bf6a82f-9aa0-4b4f-b799-aa2eb985cab8 · outbound

This paper cites SAM 2: Segment Anything in Images and Videos.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation SAM 2: Segment Anything in Images and Videos

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:15.112832Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:15.112832Z digest=sha256:349f3da1a467fa29faa863f919a7e01bfb2c2904fbadb5fe0da58112f7ef9079

Observation 31de9a2b-2e04-436a-b098-dda4e6952cce · outbound

This paper cites an unresolved cited work.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:07:15.858624Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T05:07:15.117225Z digest=sha256:04c80b235c9d3639fd8a6092d26e955f5879bb1a0de351bcfb30e5d33f924485

Observation fd3f0e28-2d5d-45df-9678-7c1b18453a12 · outbound

This paper cites an unresolved cited work.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Unresolved cited work

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:15.120920Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:15.120920Z digest=sha256:cc5fd5a4a00619951f5fad4844899d2bda57a02042d18e0b11d67c94c153cca7

Observation 484fe3bb-2ce6-4ff8-8570-ae5d2a8f3313 · outbound

This paper cites an unresolved cited work.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:07:15.838532Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T05:07:15.124595Z digest=sha256:2cc3e06644aa48eca2e9680de2816f7f1f3e05b460c7571a8d0e2f248e530028

Observation a04d0853-748d-49a4-bd63-e9d7f45dcbd0 · outbound

This paper cites StableAnimator: High-Quality Identity-Preserving Human Image Animation.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation StableAnimator: High-Quality Identity-Preserving Human Image Animation

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:15.128129Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:15.128129Z digest=sha256:ab19e2bc8113cd88d94d92f0266750dc4a960a76bddeaa43b5c14484f2f4abe9

Observation 6d1904a6-f7fe-4897-bd51-2227546b86b2 · outbound

This paper cites Towards Accurate Generative Models of Video: A New Metric & Challenges.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Towards Accurate Generative Models of Video: A New Metric & Challenges

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:15.131757Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:15.131757Z digest=sha256:571510aa4017e8f8df8ba25c1940c06fe491f2006f8cc7ccce78f0ef45eaf6fc

Observation d2dca16d-e51f-45be-82a3-418c67c2a856 · outbound

This paper cites Wan: Open and Advanced Large-Scale Video Generative Models.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Wan: Open and Advanced Large-Scale Video Generative Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:15.135416Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:15.135416Z digest=sha256:6b126330d5ceb1781cfcf2e83f61aa97c4a7f5e56297536720d53380d82cd042

Observation eae3a003-2f1b-4816-b084-79848605e6c1 · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:15.139064Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:15.139064Z digest=sha256:3266bda0f248aa514a546fa7bc74a8c4e3fc977eecb0036a412c12c27853a1a5

Observation 4838ac87-d887-4cd9-ab43-df53cbe08d6f · outbound

This paper cites an unresolved cited work.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Unresolved cited work

Reference 40

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:07:15.827007Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T05:07:15.142718Z digest=sha256:67412232fb7f21a5da68f2f5e5ecb9b13782c5c873b5c16eddbf2c2c9d0cf9f1

Observation 91ca9bd0-f838-4fa4-acdd-9ef2b09d78a0 · outbound

This paper cites UniAnimate: Taming Unified Video Diffusion Models for Consistent Human Image Animation.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation UniAnimate: Taming Unified Video Diffusion Models for Consistent Human Image Animation

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:15.146443Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:15.146443Z digest=sha256:05af5124c8ddd6634432b15c9687942f614e557cdb986583fb25c589d7d90477

Observation dc0a1690-b236-4020-bb5d-04958a9d41c7 · outbound

This paper cites an unresolved cited work.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Unresolved cited work

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:15.150092Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:15.150092Z digest=sha256:ea298932f57cf463dc5944f652aaee4f657c09857c3cea6a98a5417318c554e4

Observation 4f40aa06-c7c2-4967-9623-69976fed45b0 · outbound

This paper cites an unresolved cited work.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Unresolved cited work

Reference 43

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:07:15.815546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T05:07:15.154172Z digest=sha256:3541fadd70722cc5a4a66a327700f876774ec83a4e511f32626effac7b617a4b

Observation fd352863-e67c-4f0d-854d-90d3ee8b8360 · outbound

This paper cites an unresolved cited work.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Unresolved cited work

Reference 44

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:07:15.803322Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T05:07:15.157826Z digest=sha256:21ed0fc61efc485f0daef84fa18e11128c6b6a2c0c826854d4fc375860d35bfe

Observation d82759de-ba97-4831-9f09-2ce941a4833e · outbound

This paper cites AnchorCrafter: Animate Cyber-Anchors Selling Your Products via Human-Object Interacting Video Generation.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation AnchorCrafter: Animate Cyber-Anchors Selling Your Products via Human-Object Interacting Video Generation

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:15.161431Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:15.161431Z digest=sha256:d3948f457439fe7a7be74edd7ceaf5d0a0c5ae0e9444820820c576eab3d6a9fe

Observation 433e54dc-d2aa-4faf-8152-6f221b269c38 · outbound

This paper cites an unresolved cited work.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Unresolved cited work

Reference 46

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:07:15.789992Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T05:07:15.165245Z digest=sha256:2d1906d79b252301e57f4aea7529546222430978d96dfcf3b5287c539bfee1ad

Observation 84bebacb-0d3f-4e78-85fb-f212ed8f2ca5 · outbound

This paper cites HOI-Swap: Swapping Objects in Videos with Hand-Object Interaction Awareness.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation HOI-Swap: Swapping Objects in Videos with Hand-Object Interaction Awareness

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:15.168770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:15.168770Z digest=sha256:2800522c35ac5ea0a84f4e27d77e1851ce4948927158fda5130e7e36d1862667

Observation 81cf7fda-803e-4bbb-9e3c-101e60cac3e8 · outbound

This paper cites an unresolved cited work.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Unresolved cited work

Reference 48

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:07:15.778193Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T05:07:15.172606Z digest=sha256:d8f97d4511475637112b47cf236e619470a5bfbc5dd20b57d806f5e1a04bed66

Observation 8083b526-4c7a-4a20-a453-1c66d4e64b98 · outbound

This paper cites an unresolved cited work.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Unresolved cited work

Reference 49

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:07:15.766481Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T05:07:15.176525Z digest=sha256:a675cd6d0577fc18b7224d272bc8104bcaca6dd3978c98f58709e3436676303e

Observation 12113b3d-b607-4cd5-91dd-38834a3c3a54 · outbound

This paper cites Zhang, X.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Zhang, X

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:07:15.754349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T05:07:15.180425Z digest=sha256:8caacc8adfe428c7f3548b9255ea71f718ef00a45ada096970964aedc18b733e

Observation 81d175ab-acfe-476a-86f2-1be1d4a54a75 · outbound

This paper cites MimicMotion: High-Quality Human Motion Video Generation with Confidence-aware Pose Guidance.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation MimicMotion: High-Quality Human Motion Video Generation with Confidence-aware Pose Guidance

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:15.184611Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:15.184611Z digest=sha256:2f66e8519eefcb42f4a01d7247a0e952d94f847b7f745cb6ba3109e033663d7a

Observation 8298ced2-5661-40d2-a24b-03c834d10b85 · outbound

This paper cites Allegro: Open the Black Box of Commercial-Level Video Generation Model.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Allegro: Open the Black Box of Commercial-Level Video Generation Model

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T05:07:15.188299Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:07:15.188299Z digest=sha256:f613f34e9eb8fd730be489897bb05fed2f902bc174583a8d2229b7f96cdab938

Observation 9891aa66-5325-4f7d-b603-e7a25078fd66 · outbound

This paper cites an unresolved cited work.

HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation Unresolved cited work

Reference 53

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:07:15.741878Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-08-07T05:07:15.192008Z digest=sha256:7e422f5b5efbf15e96552e9d81fa337833f33f0533cc886376e68fe710105c94

Pith citing papers

Observation d4776e48-b2a1-4f33-9ac4-7b64aa05c48d · inbound

HOComp: Interaction-Aware Human-Object Composition cites this paper.

HOComp: Interaction-Aware Human-Object Composition HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T15:09:51.916002Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:09:51.916002Z digest=sha256:4c1f4afe0e219f9c4b038fe9cef99694416f7499e78b1328e055d0728bc5cc9a

Observation 6520dd86-2b53-45dc-8825-bf9b92f07a76 · inbound

VHOI: Controllable Video Generation of Human-Object Interactions from Sparse Trajectories via Motion Densification cites this paper.

VHOI: Controllable Video Generation of Human-Object Interactions from Sparse Trajectories via Motion Densification HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-07-09T01:19:36.380554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-16T23:30:14.969895Z digest=sha256:300e361f25dad25a7f4592e351c8756d16260f51fd79c07f33fa3c8287d29188

Observation 5cc0e35d-d677-4b28-ae37-737691ad3291 · inbound

CoMoVi: Co-Generation of 3D Human Motions and Realistic Videos cites this paper.

CoMoVi: Co-Generation of 3D Human Motions and Realistic Videos HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-07-09T01:19:36.380554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-16T13:43:26.460480Z digest=sha256:2ff86ce3cc5c5b8c870eb36ec0da28da24565fb1611799fe04880854294874b7

Observation ccc26b66-cc67-4ae2-a977-bf55d8a440c0 · inbound

EchoTorrent: Towards Swift, Sustained, and Streaming Multi-Modal Video Generation cites this paper.

EchoTorrent: Towards Swift, Sustained, and Streaming Multi-Modal Video Generation HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation

Reference 73

Resolution
verified exact
arxiv_id, observed 2026-07-09T01:19:36.380554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-15T22:20:16.320171Z digest=sha256:bfceb33b94f8d60f680c1b0188a5f952444352c61f9d1f65566976389f8ad747

Observation d6adda24-10da-4a15-b2b2-8317664b43b4 · inbound

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model cites this paper.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.179232Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.179232Z digest=sha256:6f6418bb42b871fd6207e31fd7ad7785616da4a3cd97e02062b7b23839dfab2e

Observation 85627e22-73e7-4684-b5a1-bea08e89bdd5 · inbound

OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation cites this paper.

OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-07-09T01:19:36.380554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-10T15:09:02.727887Z digest=sha256:1a8e4e1439336607823106e16bc3c64abc6842df5c9eabfdf875068261152ca0

Observation 96207e89-b4cf-4340-a436-2da9dfdc30b4 · inbound

StreamHOI: Interaction-aware Temporal Memory Adaptation for Streaming HOI Video Generation cites this paper.

StreamHOI: Interaction-aware Temporal Memory Adaptation for Streaming HOI Video Generation HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-01T10:37:53.829539Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T10:37:53.829539Z digest=sha256:630418b808e9f28e659e27f33a17c8c6837598b63385248598a55efa590502fb