Pith. sign in

Paper Citation Record · LEDGER

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model

As of 7 August 2026, this Paper Citation Record lists 47 of 47 outbound references and 2 inbound Pith citation observations for arXiv:2603.14686.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2603.14686 v2

Coverage vector

measured 47 of 47 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-04T05:51:18.368996Z

measured 49 of 49 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-26T00:59:43.745654Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T16:09:56.485116Z

Reference resolution

47 of 47 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved47
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation b1aa708c-375e-4a52-867e-68f4f8ab4712 · outbound

This paper cites Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.098812Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.098812Z digest=sha256:234d03062b7f34d82ed8c637b7b7ef4d6d1d474a713cb53dde903ab5e8e2f3fc

Observation e29c055a-350e-4ff8-9f65-44c65b6e6ee6 · outbound

This paper cites an unresolved cited work.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Unresolved cited work

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.105517Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.105517Z digest=sha256:533bff815dd1bbb01a61ec34f49b055e43ff415dca85d9912866bfeaf2743cea

Observation e55e8956-b8f0-4f28-8a80-3d82ee835944 · outbound

This paper cites HuMo: Human-Centric Video Generation via Collaborative Multi-Modal Conditioning.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model HuMo: Human-Centric Video Generation via Collaborative Multi-Modal Conditioning

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.110781Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.110781Z digest=sha256:a2f3cfde7ec204289021c110425d7e94fa427ae08ebd0aa8578d6f2adc03c6a1

Observation b510ec62-daa0-4da5-ae4a-7cb0773c249d · outbound

This paper cites an unresolved cited work.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Unresolved cited work

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.115975Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.115975Z digest=sha256:cd2b57484df3b6ed4b0a5ac1e26ca78beb8556174547da094740aa27321a1eee

Observation d18c06b5-69be-4030-9fce-1223b936dc7b · outbound

This paper cites an unresolved cited work.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Unresolved cited work

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.120659Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.120659Z digest=sha256:63e50e47a4698e18348c1f39f5c73bb696e5b84922687d2d0aa8a63be2b88f37

Observation 9ff5ed00-7ae5-492b-b778-91a3e60a7ea9 · outbound

This paper cites an unresolved cited work.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Unresolved cited work

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.126833Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.126833Z digest=sha256:f8ac07ad0bc1508778548dc8e5aaa342a38df161a5d0b9512711f7041d06a089

Observation 490134c9-613a-49b7-b5bf-be4e2c19bd48 · outbound

This paper cites an unresolved cited work.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Unresolved cited work

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.138845Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.138845Z digest=sha256:fafac61c7385117bb46f148457e9106d5e50e88f76a6df193ba8e694e10c4c29

Observation c6840a51-fe24-4d12-a38e-61f3f85afeb0 · outbound

This paper cites an unresolved cited work.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Unresolved cited work

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.143944Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.143944Z digest=sha256:771cc60b67e2769b70df7cfe8946ee2532ab58fc14cf57b051e8933059abf7f7

Observation 038ae1d7-c724-47d8-860d-9402ba8ba5e0 · outbound

This paper cites AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.152357Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.152357Z digest=sha256:613aa5484ca9f6e95050782f4886ded9a84423b1d728e3d2d161fb4429f9d780

Observation bb95b24b-9a80-490d-b510-3b6420cf8919 · outbound

This paper cites Prompt-to-Prompt Image Editing with Cross Attention Control.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Prompt-to-Prompt Image Editing with Cross Attention Control

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.157830Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.157830Z digest=sha256:b0fc0405b3f86e9b3cad996be11ba18e58d4c7039dd89d036b1a1e45915ed0f3

Observation 133b439f-3eac-4fb4-9b53-14a50f0180d9 · outbound

This paper cites Animate Anyone: Consistent and Controllable Image-to-Video Synthesis for Character Animation.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Animate Anyone: Consistent and Controllable Image-to-Video Synthesis for Character Animation

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.163164Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.163164Z digest=sha256:0b3c093d507d7dc23bba34aed0c3f64851545df4ce3a32b64e98d5a43aa1dda2

Observation c97ce011-c67f-4cbf-a1d6-0403d5d9fecc · outbound

This paper cites HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model HunyuanCustom: A Multimodal-Driven Architecture for Customized Video Generation

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.169044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.169044Z digest=sha256:afca0deb8366e71435a00302d7b1888c387ad440a096f3ca36ff294e96a6f9a5

Observation d6adda24-10da-4a15-b2b2-8317664b43b4 · outbound

This paper cites HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model HunyuanVideo-HOMA: Generic Human-Object Interaction in Multimodal Driven Human Animation

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.179232Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.179232Z digest=sha256:f3b75ef4d05a3d6e8a14305f038e2707d5d40968e7c36f97d8c3d5a9ce2bf62e

Observation cd5008ab-a760-4c54-a698-5794c5011dfa · outbound

This paper cites RTMW: Real-Time Multi-Person 2D and 3D Whole-body Pose Estimation.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model RTMW: Real-Time Multi-Person 2D and 3D Whole-body Pose Estimation

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.185135Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.185135Z digest=sha256:c1a70b12d36dbe5cd3e58edecb6b354aced6b6a23dbcc52eb3098206e4e33f7d

Observation ce23bc03-0137-4b26-82cf-d439a1628d84 · outbound

This paper cites an unresolved cited work.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Unresolved cited work

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.192246Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.192246Z digest=sha256:be124366848f9930397caef72cf2b59ee672b6170bf17837163b2756fc3a70f9

Observation b218307f-9b56-46b9-9a4b-36196cf3f4be · outbound

This paper cites HunyuanVideo: A Systematic Framework For Large Video Generative Models.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model HunyuanVideo: A Systematic Framework For Large Video Generative Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.204730Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.204730Z digest=sha256:9b9e41262d8216c3d00543b05e5129f50f5cdd2b8146721ea9888fcab603ad6a

Observation c5b06e17-aa0b-4273-a73b-8b29735ecb1a · outbound

This paper cites Depth Anything 3: Recovering the Visual Space from Any Views.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Depth Anything 3: Recovering the Visual Space from Any Views

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.215614Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.215614Z digest=sha256:e36ff5e966bd6d26b797ce4d24868ab0c15fb484e4f15eaa73c2854e1db4b9c8

Observation e21ede37-84ea-46d8-a57f-ad12acb28182 · outbound

This paper cites an unresolved cited work.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Unresolved cited work

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.221154Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.221154Z digest=sha256:ac55737469904954062f30bebb0c124f50321796d40221d9a0a2b3729dfe6771

Observation b16c105f-f6af-4215-ad0b-6fb297f91cdc · outbound

This paper cites Phantom: Subject-consistent video generation via cross-modal alignment.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Phantom: Subject-consistent video generation via cross-modal alignment

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.232473Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.232473Z digest=sha256:468c10f95c43de0eed5eea53b61bd6bf579569c49dcce1a7b068b5425827083c

Observation 16979f16-ebb9-4677-b32a-776376fd6c4f · outbound

This paper cites Flow Straight and Fast: Learning to Generate and Transfer Data with Rectified Flow.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Flow Straight and Fast: Learning to Generate and Transfer Data with Rectified Flow

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.237773Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.237773Z digest=sha256:3007c05b6440b8cfb643243715ee2f26d83fdfb827fcb098eb5d5881fd1237a6

Observation 04d54133-eb18-4228-a6f2-5658f12fa613 · outbound

This paper cites an unresolved cited work.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Unresolved cited work

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.243683Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.243683Z digest=sha256:b6284b252df526f29f404ec272eb54e9e0fb1763144b3270e5c27e2686c94bc5

Observation f01654d5-378e-4f14-9079-a98f9a64b87e · outbound

This paper cites an unresolved cited work.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Unresolved cited work

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.255365Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.255365Z digest=sha256:d76348d4d991b5d4074315137ac54f46ad2690c515dd6efb64a19fa9774cf4b5

Observation 068d8b06-fb24-4022-888d-ab858b393888 · outbound

This paper cites an unresolved cited work.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Unresolved cited work

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.260833Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.260833Z digest=sha256:35c8f5b51ca9ef1ca8355f370cbf3c0537a09afd3811d6a92def80755c073c86

Observation 5c40f80d-7768-49f7-a28a-f4b114bbe64c · outbound

This paper cites SAM 2: Segment Anything in Images and Videos.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model SAM 2: Segment Anything in Images and Videos

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.267443Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.267443Z digest=sha256:97d349a7761e5a5d2b49a78165b356698efb13fb18d720e28495e821259cb204

Observation 04532f0f-0f71-4dac-8848-862446cf5bc6 · outbound

This paper cites InProceedings of the IEEE/CVF International Conference on Computer Vision.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model InProceedings of the IEEE/CVF International Conference on Computer Vision

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.248963Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.248963Z digest=sha256:d769a44edd49c6261418b802990c7cd4131b15cc205571b60cd3ee7853fc637c

Observation 455c3f13-6a47-4e7c-af02-b4c4abf191af · outbound

This paper cites an unresolved cited work.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Unresolved cited work

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.278318Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.278318Z digest=sha256:f1ff8e2e6a17d70b9a91ea233ca694d693edcd81cdff1b0a97882a779e873820

Observation 18119a37-0603-4057-9a07-671a26fcbc29 · outbound

This paper cites Wan: Open and Advanced Large-Scale Video Generative Models.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Wan: Open and Advanced Large-Scale Video Generative Models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.283800Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.283800Z digest=sha256:391ca50e1f8b86f3468841189acef82e5def281dcd22f02f5a480fa670b378ea

Observation e32f25a6-a742-480c-a4dc-8e7b60e0d4a6 · outbound

This paper cites an unresolved cited work.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Unresolved cited work

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.288763Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.288763Z digest=sha256:db621242df4036ee8da6c8bf431d5169a8f3d1d7e548d139ffc189207d5978b2

Observation 3f551690-9f63-4a02-b25b-03ae57310c35 · outbound

This paper cites an unresolved cited work.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Unresolved cited work

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.272924Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.272924Z digest=sha256:6c6433dcb97a58a0d811eac9cc6845a2476504941d08d0a61565b41992ff099f

Observation 6aebffb1-0f1a-44a7-86ae-ab297ce6e65e · outbound

This paper cites an unresolved cited work.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Unresolved cited work

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.300737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.300737Z digest=sha256:8fd1b3b4589a964ddd996695a3b210197a9eae54e224b0e60e3cadf3ef1d5b03

Observation 6b6c40c3-0cb7-466a-804b-f41bee43aa1f · outbound

This paper cites $\pi^3$: Permutation-Equivariant Visual Geometry Learning.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model $\pi^3$: Permutation-Equivariant Visual Geometry Learning

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.305928Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.305928Z digest=sha256:3af61a00a709c6239ae0ce9ce28522b73241ef52624afc95bb3102a74d39bf47

Observation 03da2018-497a-4bf2-9c50-105062db989f · outbound

This paper cites an unresolved cited work.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Unresolved cited work

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.311084Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.311084Z digest=sha256:1ad185eb975883317a0c89531f1a028e02ab87b5f5bcfba552e247cdff310c5c

Observation be3e3eba-ad2f-4489-b69e-9ae609aadbb7 · outbound

This paper cites DreamActor-H1: High-Fidelity Human-Product Demonstration Video Generation via Motion-designed Diffusion Transformers.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model DreamActor-H1: High-Fidelity Human-Product Demonstration Video Generation via Motion-designed Diffusion Transformers

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.294240Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.294240Z digest=sha256:ab819212750d8e4163af233aa7cd429da1905bc742aa76b7a503605e6e7834f4

Observation 3040c207-a712-458a-84e4-e2076723cad9 · outbound

This paper cites an unresolved cited work.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Unresolved cited work

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.321386Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.321386Z digest=sha256:66b5314696d0d525c1bcb8476f43ba84d7fff3d7593c5b91c26d7419ffd1b291

Observation 500fe1d5-d8c3-499e-a066-fcbdc3341779 · outbound

This paper cites an unresolved cited work.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Unresolved cited work

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.326505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.326505Z digest=sha256:5ccf494d18e694b64116e11d36ebae39c4c71964ca97a8a362ee6a586133058e

Observation c5a55f1f-e6f0-4b1e-bd12-009264df94b3 · outbound

This paper cites an unresolved cited work.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Unresolved cited work

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.331988Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.331988Z digest=sha256:c835b800312fe8d732d31ff114dd749fba8482dbea0fe87813c37f393a8954d4

Observation 6604432f-ebaa-4d44-a338-b4e6769e4fa7 · outbound

This paper cites AnchorCrafter: Animate Cyber-Anchors Selling Your Products via Human-Object Interacting Video Generation.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model AnchorCrafter: Animate Cyber-Anchors Selling Your Products via Human-Object Interacting Video Generation

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.316887Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.316887Z digest=sha256:70bb3f73d57cd3fe12e3c5d417b2d8c825c56c28f2c961a6150af0b5b300d706

Observation a3c75d5f-5f17-46d8-a0a3-a045909e59d5 · outbound

This paper cites CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.341928Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.341928Z digest=sha256:15908ccd919bd9823010a740088a9479fd5c4ce51deb3f549982f86d29dc959d

Observation 51ff2502-47e8-4c67-9e06-75716d625942 · outbound

This paper cites an unresolved cited work.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Unresolved cited work

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.347090Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.347090Z digest=sha256:400c1155fc17a0dd5baa18693c53b428baa9b4220614fc600cf063aee468ca74

Observation ac46fc77-f504-4418-8c2e-e76e0b9c3484 · outbound

This paper cites MiniCPM-V 4.5: Cooking Efficient MLLMs via Architecture, Data, and Training Recipe.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model MiniCPM-V 4.5: Cooking Efficient MLLMs via Architecture, Data, and Training Recipe

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.351672Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.351672Z digest=sha256:49d291ea95e66ed2410f6cab63dda1fb2a937221dce88564eefa14e4ef108acc

Observation ed2c303c-9e75-4f58-9952-0a77f0615b85 · outbound

This paper cites an unresolved cited work.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Unresolved cited work

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.336435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.336435Z digest=sha256:e5520175d9d347644709eeea31f50f2b1f1d1b2b2c452e27d0a6f654ea2ca3f2

Observation 13bfb6fa-cde1-4de3-9fb3-a6e25a97abc5 · outbound

This paper cites an unresolved cited work.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Unresolved cited work

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.362675Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.362675Z digest=sha256:de4a189848d77b0d4c132def39d33ccc68f406d381d69c78809237d006166350

Observation bacbffde-6624-4b81-93d3-eaa7844ff749 · outbound

This paper cites an unresolved cited work.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Unresolved cited work

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.368996Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.368996Z digest=sha256:2cdb61be6cb982721681a42ac1715e5bf6ab08894e1570fc0195ffd73e1e4926

Observation 69f03735-c228-4fb3-b692-a0287f117ffa · outbound

This paper cites an unresolved cited work.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Unresolved cited work

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.357386Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.357386Z digest=sha256:bd465c965580891d1d529cac181fad249f3fffbd7f15b6e9629935a55212cf44

Observation 89f85e54-5606-424c-b4e4-237bd2bbec87 · outbound

This paper cites Flow Matching for Generative Modeling.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model Flow Matching for Generative Modeling

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.226041Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.226041Z digest=sha256:7fe0723a7f3603d5ea55b760fa9b42bfc2b9b3d54a787eef0c71ffc5031f5042

Observation ed24f181-dec8-4e7a-a322-aaee1267b8d0 · outbound

This paper cites InProceedings of the IEEE/CVF conference on computer vision and pattern recognition.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model InProceedings of the IEEE/CVF conference on computer vision and pattern recognition

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.132547Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.132547Z digest=sha256:eef3d0d6f80605867b3b1e332727c40a6686b663a727988319a09d98baaead6a

Observation c19dee81-eaf7-4b91-be3f-d3db49139466 · outbound

This paper cites VACE: All-in-One Video Creation and Editing.

MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model VACE: All-in-One Video Creation and Editing

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-04T05:51:18.200198Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T05:51:18.200198Z digest=sha256:5b2b5a6318fe1272ec67f222df9ab97e52891a52e9e4ec4dff0be934d043cdad

Pith citing papers

Observation d380bd2e-4ecb-4882-b390-a636882722b2 · inbound

Controllable Video Object Insertion via Multiview Priors cites this paper.

Controllable Video Object Insertion via Multiview Priors MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-08-04T01:52:33.282752Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T11:44:17.033051Z digest=sha256:133a0bbf1ecabf97660ce5b5db072e7b2c66fa190f45034d253158d3439eae16

Observation 38925e92-2ad6-4530-97bb-74311b3b8347 · inbound

VistaRef: Boosting Visual Spatial Orientation Awareness for Pointing-to-Object Detection cites this paper.

VistaRef: Boosting Visual Spatial Orientation Awareness for Pointing-to-Object Detection MVHOI: Bridge Multi-view Condition to Complex Human-Object Interaction Video Reenactment via 3D Foundation Model

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-08-04T01:52:33.282752Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-26T00:59:43.745654Z digest=sha256:69be2b778ba26f19072066416a35350f717e650a2e8bf7358d9b20e27e077b19