Pith. sign in

Paper Citation Record · LEDGER

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model

As of 9 August 2026, this Paper Citation Record lists 47 of 47 outbound references and 2 inbound Pith citation observations for arXiv:2603.14686.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2603.14686 v2

Coverage vector

measured 47 of 47 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-04T05:51:18.368996Z

measured 49 of 49 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-26T00:59:43.745654Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T16:09:56.485116Z

Reference resolution

47 of 47 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved47
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation b1aa708c-375e-4a52-867e-68f4f8ab4712 · outbound

This paper cites Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.098812Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.098812Z digest=sha256:0999f3f9654d2ec8ca2f9f1a0fa15967543d761df3223fec3522818ae4d24fe5

Observation e29c055a-350e-4ff8-9f65-44c65b6e6ee6 · outbound

This paper cites an unresolved cited work.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Unresolved cited work

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.105517Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.105517Z digest=sha256:800c06b3ba42f01ef552e2c31d3c3a094ebfdd23dea347ebb11f156b206a5c1a

Observation e55e8956-b8f0-4f28-8a80-3d82ee835944 · outbound

This paper cites HuMo: Human-Centric Video Generation via Collaborative Multi-Modal Conditioning.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model HuMo: Human-Centric Video Generation via Collaborative Multi-Modal Conditioning

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.110781Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.110781Z digest=sha256:a8f9aa4d5fac1954edfa4bb591e0f0d50e21926525bd722515d078d01d916ee5

Observation b510ec62-daa0-4da5-ae4a-7cb0773c249d · outbound

This paper cites an unresolved cited work.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Unresolved cited work

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.115975Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.115975Z digest=sha256:3c8305f6ff554db492b3eebf483ace1a6836386696a6a91aede78cd59651b288

Observation d18c06b5-69be-4030-9fce-1223b936dc7b · outbound

This paper cites an unresolved cited work.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Unresolved cited work

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.120659Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.120659Z digest=sha256:33d4f0251e455a7dccd8035f33eb95e268e239817bcb01c9bb022a04f609de65

Observation 9ff5ed00-7ae5-492b-b778-91a3e60a7ea9 · outbound

This paper cites an unresolved cited work.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Unresolved cited work

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.126833Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.126833Z digest=sha256:c74bb344778583fffa487a03ada39acf9007678daf8eef55cf6aad9e4235dc0d

Observation 490134c9-613a-49b7-b5bf-be4e2c19bd48 · outbound

This paper cites an unresolved cited work.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Unresolved cited work

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.138845Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.138845Z digest=sha256:838291ad9aae40ebe5fdacdd39ad928c68c66ee98e363717541d733519839149

Observation c6840a51-fe24-4d12-a38e-61f3f85afeb0 · outbound

This paper cites an unresolved cited work.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Unresolved cited work

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.143944Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.143944Z digest=sha256:3c72d2887a2880904b1435dba1c1a922e77cb432095fcf19b28460ed9230f005

Observation 038ae1d7-c724-47d8-860d-9402ba8ba5e0 · outbound

This paper cites AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.152357Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.152357Z digest=sha256:fdb3711351a2a6971feae2dcd01f9aa11910056881a35f4dd55e88d1b10aac88

Observation bb95b24b-9a80-490d-b510-3b6420cf8919 · outbound

This paper cites Prompt-to-Prompt Image Editing with Cross Attention Control.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Prompt-to-Prompt Image Editing with Cross Attention Control

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.157830Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.157830Z digest=sha256:12a9005aff923ee338e9e79ff45c606821971cca5d2466b5da2d3ffb4cdc3bc6

Observation 133b439f-3eac-4fb4-9b53-14a50f0180d9 · outbound

This paper cites Animate Anyone: Consistent and Controllable Image-to-Video Synthesis for Character Animation.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Animate Anyone: Consistent and Controllable Image-to-Video Synthesis for Character Animation

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.163164Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.163164Z digest=sha256:c7238cc00c0683b5b9f215e8e1d167fc935d5905ef4c86b9a3ac3069721d139e

Observation c97ce011-c67f-4cbf-a1d6-0403d5d9fecc · outbound

This paper cites HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.169044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.169044Z digest=sha256:a2910f9878817f7ab021fdf568d86e3e702f086728e77b19d5d381378bacf752

Observation d6adda24-10da-4a15-b2b2-8317664b43b4 · outbound

This paper cites HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.179232Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.179232Z digest=sha256:9974ce934a9fc1bbd84ea7dcbce0b440aa590f39230f7ca3444a525e7a8b6553

Observation cd5008ab-a760-4c54-a698-5794c5011dfa · outbound

This paper cites RTMW: Real-Time Multi-Person 2D and 3D Whole-body Pose Estimation.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model RTMW: Real-Time Multi-Person 2D and 3D Whole-body Pose Estimation

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.185135Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.185135Z digest=sha256:ab8de459eb2f874e3a67d0a1ddbb051f5b14136949f332cb786f21467898cf8f

Observation ce23bc03-0137-4b26-82cf-d439a1628d84 · outbound

This paper cites an unresolved cited work.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Unresolved cited work

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.192246Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.192246Z digest=sha256:1b1d37286ecb6491389e54f652a60d2f19264ea2fa0f9255c20131d120689f96

Observation b218307f-9b56-46b9-9a4b-36196cf3f4be · outbound

This paper cites HunyuanVideo: A Systematic Framework For Large Video Generative Models.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model HunyuanVideo: A Systematic Framework For Large Video Generative Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.204730Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.204730Z digest=sha256:6dcd9dd2a6a2e34c32d0202dd84729a047e3748f331ad16131b4ded664d17ac4

Observation c5b06e17-aa0b-4273-a73b-8b29735ecb1a · outbound

This paper cites Depth Anything 3: Recovering the Visual Space from Any Views.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Depth Anything 3: Recovering the Visual Space from Any Views

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.215614Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.215614Z digest=sha256:d5888636f7c9ff359203c64f22cf67dbf031f8087ebdb042a0f42d5ab9321069

Observation e21ede37-84ea-46d8-a57f-ad12acb28182 · outbound

This paper cites an unresolved cited work.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Unresolved cited work

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.221154Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.221154Z digest=sha256:01e5bac33e022233f46d419c7dfd8d60332e9c3f352e28b1a87e1f199cf6a835

Observation b16c105f-f6af-4215-ad0b-6fb297f91cdc · outbound

This paper cites Phantom: Subject-consistent video generation via cross-modal alignment.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Phantom: Subject-consistent video generation via cross-modal alignment

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.232473Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.232473Z digest=sha256:ec1c8ce547e358be10607893b074b3a71e7a45555de4f581ded2049faca95ecc

Observation 16979f16-ebb9-4677-b32a-776376fd6c4f · outbound

This paper cites Flow Straight and Fast: Learning to Generate and Transfer Data with Rectified Flow.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Flow Straight and Fast: Learning to Generate and Transfer Data with Rectified Flow

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.237773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.237773Z digest=sha256:c5144493daade35fdd6e7c6fffab5185e1b3186787547cedeeb8dee8e5f53540

Observation 04d54133-eb18-4228-a6f2-5658f12fa613 · outbound

This paper cites an unresolved cited work.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Unresolved cited work

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.243683Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.243683Z digest=sha256:d98a7391bb22f90f78e24f2f3e5cd092f5a9b2459f70e865515bbeedd726a02f

Observation f01654d5-378e-4f14-9079-a98f9a64b87e · outbound

This paper cites an unresolved cited work.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Unresolved cited work

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.255365Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.255365Z digest=sha256:8972ab1f38b0d9ba57cfa68a5366566b97c027f11cdc9d6cea8fad8c656c6777

Observation 068d8b06-fb24-4022-888d-ab858b393888 · outbound

This paper cites an unresolved cited work.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Unresolved cited work

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.260833Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.260833Z digest=sha256:099ae6307522ed23fbd8d2fb0e58f552a8fe34c94e12b20ffcafc702f48f47e9

Observation 5c40f80d-7768-49f7-a28a-f4b114bbe64c · outbound

This paper cites SAM 2: Segment Anything in Images and Videos.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model SAM 2: Segment Anything in Images and Videos

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.267443Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.267443Z digest=sha256:4fc1060f0d682ccccce60f2d12daa03ab7271f3cc82124c2783b45f7c14a4462

Observation 04532f0f-0f71-4dac-8848-862446cf5bc6 · outbound

This paper cites InProceedings of the IEEE/CVF International Conference on Computer Vision.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model InProceedings of the IEEE/CVF International Conference on Computer Vision

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.248963Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.248963Z digest=sha256:8fb202d6d7877519b0a03e9d2a731ebd2faab2dae401222ac9d6b938a77383e8

Observation 455c3f13-6a47-4e7c-af02-b4c4abf191af · outbound

This paper cites an unresolved cited work.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Unresolved cited work

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.278318Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.278318Z digest=sha256:2ee1092ae2a54c8ffbc5da3056408d4b38e39c16397037a0b65cf5e865bb6075

Observation 18119a37-0603-4057-9a07-671a26fcbc29 · outbound

This paper cites Wan: Open and Advanced Large-Scale Video Generative Models.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Wan: Open and Advanced Large-Scale Video Generative Models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.283800Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.283800Z digest=sha256:125c15f658f477a0a972994010eefa4b8829c3d82355fb9542092359cd7c1342

Observation e32f25a6-a742-480c-a4dc-8e7b60e0d4a6 · outbound

This paper cites an unresolved cited work.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Unresolved cited work

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.288763Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.288763Z digest=sha256:e4ef5949a427ab70afd2aca55ced0feedb2f91c8bb55bb2522baaf5f4372f1a3

Observation 3f551690-9f63-4a02-b25b-03ae57310c35 · outbound

This paper cites an unresolved cited work.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Unresolved cited work

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.272924Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.272924Z digest=sha256:0eaac5ca1f2c2df5ed70870c6a5dcce1a1e3b728ef23db2bfc898e195c6dd67b

Observation 6aebffb1-0f1a-44a7-86ae-ab297ce6e65e · outbound

This paper cites an unresolved cited work.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Unresolved cited work

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.300737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.300737Z digest=sha256:98c43c45f66054be1b63fe5391d10301ff30c05a5bd2aa08643cdc4d7b3f4b84

Observation 6b6c40c3-0cb7-466a-804b-f41bee43aa1f · outbound

This paper cites $\pi^3$: Permutation-Equivariant Visual Geometry Learning.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model $\pi^3$: Permutation-Equivariant Visual Geometry Learning

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.305928Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.305928Z digest=sha256:c353efe94087cee99aa826530c611ef488a2b703895d6f3d8f0b415db030c149

Observation 03da2018-497a-4bf2-9c50-105062db989f · outbound

This paper cites an unresolved cited work.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Unresolved cited work

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.311084Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.311084Z digest=sha256:bd0969988d5566e3bca878059e96b501ca8d34e55b11f0b23bf5fceb1e5df512

Observation be3e3eba-ad2f-4489-b69e-9ae609aadbb7 · outbound

This paper cites DreamActor-H1: High-Fidelity Human-Product Demonstration Video Generation via Motion-designed Diffusion Transformers.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model DreamActor-H1: High-Fidelity Human-Product Demonstration Video Generation via Motion-designed Diffusion Transformers

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.294240Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.294240Z digest=sha256:4268c2be6eea68cd0939685c7860ce0a189074b5fc98d9a2d30186941db3099f

Observation 3040c207-a712-458a-84e4-e2076723cad9 · outbound

This paper cites an unresolved cited work.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Unresolved cited work

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.321386Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.321386Z digest=sha256:ce46fcd039f5b2379f2cf47902fe177d722ec5756c9972ef735d1d84cc914f41

Observation 500fe1d5-d8c3-499e-a066-fcbdc3341779 · outbound

This paper cites an unresolved cited work.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Unresolved cited work

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.326505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.326505Z digest=sha256:ca40960ee627a7992b0c3a874f933ecc32f64f2e73a2376d63b2df486ac5233e

Observation c5a55f1f-e6f0-4b1e-bd12-009264df94b3 · outbound

This paper cites an unresolved cited work.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Unresolved cited work

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.331988Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.331988Z digest=sha256:b1e0f2dda0d92cd5078589a6f96a482b23963dbe6dca767b191cef10be5ae241

Observation 6604432f-ebaa-4d44-a338-b4e6769e4fa7 · outbound

This paper cites AnchorCrafter: Animate Cyber-Anchors Selling Your Products via Human-Object Interacting Video Generation.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model AnchorCrafter: Animate Cyber-Anchors Selling Your Products via Human-Object Interacting Video Generation

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.316887Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.316887Z digest=sha256:fe0ee7915b105b4ba6cbae229899bbe1b28e71fa217f5d74bbbdc0804aed01ef

Observation a3c75d5f-5f17-46d8-a0a3-a045909e59d5 · outbound

This paper cites CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.341928Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.341928Z digest=sha256:b64913af62038f4e7a68eebd92729c8361fe9818892a5ff156b73bdf734ed101

Observation 51ff2502-47e8-4c67-9e06-75716d625942 · outbound

This paper cites an unresolved cited work.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Unresolved cited work

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.347090Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.347090Z digest=sha256:ed3537338a3c0fae9905d3474b21653962ab138091af7c70223290094bcaaf47

Observation ac46fc77-f504-4418-8c2e-e76e0b9c3484 · outbound

This paper cites MiniCPM-V 4.5: Cooking Efficient MLLMs via Architecture, Data, and Training Recipe.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model MiniCPM-V 4.5: Cooking Efficient MLLMs via Architecture, Data, and Training Recipe

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.351672Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.351672Z digest=sha256:ddfba44ffdf7ea112b6886c2d27fee168db75d41cc2b3b7cf8e1ff14e769b850

Observation ed2c303c-9e75-4f58-9952-0a77f0615b85 · outbound

This paper cites an unresolved cited work.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Unresolved cited work

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.336435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.336435Z digest=sha256:8600f5303922d35a34503f339496316f0e1e8ff50faecd90105b01d277880540

Observation 13bfb6fa-cde1-4de3-9fb3-a6e25a97abc5 · outbound

This paper cites an unresolved cited work.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Unresolved cited work

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.362675Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.362675Z digest=sha256:c5be2438f67ae5868fc7e56ef815191c1a0be460f08f96cefc2c624fdc58c3eb

Observation bacbffde-6624-4b81-93d3-eaa7844ff749 · outbound

This paper cites an unresolved cited work.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Unresolved cited work

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.368996Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.368996Z digest=sha256:0631693f52a3dfc09706349e22dade934497db0d69fc1d3443a9da05d700c573

Observation 69f03735-c228-4fb3-b692-a0287f117ffa · outbound

This paper cites an unresolved cited work.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Unresolved cited work

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.357386Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.357386Z digest=sha256:d4531ad1600250ae2b060331c20df01efc387f0d9552babbd4af61dce1c388a3

Observation 89f85e54-5606-424c-b4e4-237bd2bbec87 · outbound

This paper cites Flow Matching for Generative Modeling.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Flow Matching for Generative Modeling

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.226041Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.226041Z digest=sha256:f32e08260ecaa6ea0b00d77df043dac56edeb810b1b90d87043f1b716583233b

Observation ed24f181-dec8-4e7a-a322-aaee1267b8d0 · outbound

This paper cites InProceedings of the IEEE/CVF conference on computer vision and pattern recognition.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model InProceedings of the IEEE/CVF conference on computer vision and pattern recognition

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.132547Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.132547Z digest=sha256:b0bcada5b24780150e80c79122199de9aa9efafadeb6e70baaa4dcc43d946074

Observation c19dee81-eaf7-4b91-be3f-d3db49139466 · outbound

This paper cites VACE: All-in-One Video Creation and Editing.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model VACE: All-in-One Video Creation and Editing

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.200198Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.200198Z digest=sha256:43a87bf38e8e29e3e1c62ce799c91536379213e18fcf466707ce8d40cba25146

Pith citing papers

Observation d380bd2e-4ecb-4882-b390-a636882722b2 · inbound

Controllable Video Object Insertion via Multiview Priors cites this paper.

Controllable Video Object Insertion via Multiview Priors MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-08-04T01:52:33.282752Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T11:44:17.033051Z digest=sha256:880bcc207d2812d095685e9ff99fc0e1725c5df6eda47d13db7eae5adb1513b1

Observation 38925e92-2ad6-4530-97bb-74311b3b8347 · inbound

VistaRef: Boosting Visual Spatial Orientation Awareness for Pointing-to-Object Detection cites this paper.

VistaRef: Boosting Visual Spatial Orientation Awareness for Pointing-to-Object Detection MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-08-04T01:52:33.282752Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-26T00:59:43.745654Z digest=sha256:6cf3807f9b72cbafb3564a7626cce32997c4f3d9f9e87cf2a66a6fd41c846eea