Pith. sign in

Paper Citation Record · LEDGER

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency

As of 20 August 2026, this Paper Citation Record lists 47 of 47 outbound references and 4 inbound Pith citation observations for arXiv:2501.08682.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.08682 v2

Coverage vector

measured 47 of 47 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T20:23:41.484020Z

measured 51 of 51 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-27T19:00:44.243335Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

47 of 47 outbound references displayed

  • verified exact0
  • verified fuzzy33
  • unresolved14
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation 7518a99d-d4de-40b5-8042-847ad32b2daf · outbound

This paper cites Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-10T20:23:40.716948Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:23:40.716948Z digest=sha256:9492c3f20ad591c27a44bd43938235900f3aba99a600747f3e1468cfc053a6f1

Observation 550e779a-d7f2-43a7-b9d4-1e113b29ed14 · outbound

This paper cites Elucidating the design space of diffusion-based generative models,.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency Elucidating the design space of diffusion-based generative models,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:23:43.284218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T20:23:40.733448Z digest=sha256:e50e7fcb50055098e87fd1a7be7c67462ab4d23250ad6ffb329729aa952e90cc

Observation befa29aa-189f-4a4a-b5f7-b7ba7fa4cdde · outbound

This paper cites Auto-Encoding Variational Bayes.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency Auto-Encoding Variational Bayes

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-10T20:23:40.739601Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:23:40.739601Z digest=sha256:713a419db38c6503765d2f0f96b1ddee033e6f20c648c6acebe648fbeceadd7c

Observation 060082d7-6318-4df1-80e0-7eb74315284f · outbound

This paper cites Single stage virtual try-on via deformable attention flows,.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency Single stage virtual try-on via deformable attention flows,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:23:43.267878Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T20:23:40.745924Z digest=sha256:8d8bffe92a680313592e8bb8792b8e2b2de83cf5d681c7869df3a16ffb0fc91b

Observation 48563e53-3878-442d-b91e-0dc33e566a53 · outbound

This paper cites Sapiens: Foundation for human vision models,.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency Sapiens: Foundation for human vision models,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:23:43.236033Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T20:23:40.799954Z digest=sha256:9fb61ffec89888beb7798458bbf8f5ef174fb15964d4e30402e253a675b75b19

Observation 791e8503-54c9-457f-b44b-3bb8b5e9a340 · outbound

This paper cites Animate anyone: Consistent and controllable image- to-video synthesis for character animation,.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency Animate anyone: Consistent and controllable image- to-video synthesis for character animation,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:23:43.189360Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T20:23:40.846489Z digest=sha256:f113c129234d290f9a11061ff3184bf2fc3bb3efc1780f61d6d303c1482b40ef

Observation b29a7acb-c54a-473d-9087-e49e16b3010b · outbound

This paper cites Dress code: High-resolution multi-category virtual try-on,.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency Dress code: High-resolution multi-category virtual try-on,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:23:43.154725Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T20:23:40.851934Z digest=sha256:8127b865b5567a180bf303b100d8e73b82f5132bc71c42b465ceb7bc5b4e7e44

Observation c6e2d118-5801-42ba-8f0f-1947efebec8d · outbound

This paper cites Rerender a video: Zero-shot text-guided video-to-video translation,.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency Rerender a video: Zero-shot text-guided video-to-video translation,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:23:43.115482Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T20:23:40.858531Z digest=sha256:f90dc8648d1b928bea8b381328283b110a014eb3328ede4b5fc7c9f8fc5d8ef6

Observation b1cb9d71-ce1a-440f-a14e-723fc00fd856 · outbound

This paper cites Diffusion models beat gans on image synthesis,.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency Diffusion models beat gans on image synthesis,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:23:43.050704Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T20:23:40.863638Z digest=sha256:9a329ec5f9fb55e50b7bde6be211941a06014b26db92bbf8567cd32e3f6cdc5c

Observation f5bdd9a6-5a08-434c-84f0-435e5a41f38b · outbound

This paper cites Learning transferable visual models from natural language supervision,.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency Learning transferable visual models from natural language supervision,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:23:43.019924Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T20:23:40.924733Z digest=sha256:bd9080c9b5bff0354444f633a5e7023e6db08879597165b8e48b466e30ce6cac

Observation 40034013-c75a-417a-bada-db14c0134e48 · outbound

This paper cites AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-10T20:23:40.952598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:23:40.952598Z digest=sha256:2b90651a287dc6ee59d50f6b0e29e2910933a647ba331d9fc438f28698b9566b

Observation 480e9800-5630-407a-84e5-374a1174f5e9 · outbound

This paper cites OOTDiffusion: Outfitting Fusion based Latent Diffusion for Controllable Virtual Try-on.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency OOTDiffusion: Outfitting Fusion based Latent Diffusion for Controllable Virtual Try-on

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-10T20:23:40.958697Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:23:40.958697Z digest=sha256:ef1da8a4c10728c8dc03b9c7af1134906bc558b1c0a1e164c3266ff3d203cb41

Observation 71ef351b-32bf-4cc8-b46d-ffdac378aa70 · outbound

This paper cites Gpd-vvto: Preserving garment details in video virtual try- on,.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency Gpd-vvto: Preserving garment details in video virtual try- on,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:23:42.983205Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T20:23:41.021799Z digest=sha256:05b3224d028e24dc562225005f8f65a73554abdfaae70bc8d1e964840db71a9e

Observation 83084fb9-fac4-43c7-adc0-cd10c3698f56 · outbound

This paper cites Latte: Latent Diffusion Transformer for Video Generation.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency Latte: Latent Diffusion Transformer for Video Generation

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-10T20:23:41.051508Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:23:41.051508Z digest=sha256:cf8ea16c9d964b6d00a94d09d582e6ad64faa6a66d6cd3e50af0bf6cd2b427dd

Observation e010788b-c20d-48e4-9449-70010462bff5 · outbound

This paper cites High-resolution image synthesis with latent diffusion models,.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency High-resolution image synthesis with latent diffusion models,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:23:42.937876Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T20:23:41.057309Z digest=sha256:74b658fe2644be928e455263dc9aa302e0d03a66f0a1ecf13113f8fe8375fa1c

Observation 798c1a56-b1dc-4320-ac8c-04b2d4b277e7 · outbound

This paper cites VITON-DiT: Learning In-the-Wild Video Try-On from Human Dance Videos via Diffusion Transformers.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency VITON-DiT: Learning In-the-Wild Video Try-On from Human Dance Videos via Diffusion Transformers

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-10T20:23:41.063202Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:23:41.063202Z digest=sha256:274d33329b981988207fac35c9de463cbcb23d138006fd615509380b757bf2b4

Observation d3a4cd6a-4147-4605-9a4a-5bcbf9082ae6 · outbound

This paper cites Tunnel try-on: Excavating spatial- temporal tunnels for high-quality virtual try-on in videos,.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency Tunnel try-on: Excavating spatial- temporal tunnels for high-quality virtual try-on in videos,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:23:42.886479Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T20:23:41.068622Z digest=sha256:9c58bf9c0c7f20e350f4beec1a1100deb94d753954323a6d81e6c8e3bcacd46e

Observation 69f62e6b-464d-4c5b-b2a9-ba8f1a416c24 · outbound

This paper cites Clothformer: Tam- ing video virtual try-on in all module,.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency Clothformer: Tam- ing video virtual try-on in all module,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:23:42.845582Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T20:23:41.073473Z digest=sha256:1dc5fbb110c4530757d632d321c28a1b98678cf2a6a5aec699a5686cdec7e410

Observation 8e9446ca-ff3b-483e-86b0-e44e767e08ed · outbound

This paper cites Improving diffusion models for authentic virtual try-on in the wild,.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency Improving diffusion models for authentic virtual try-on in the wild,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:23:42.819282Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T20:23:41.077891Z digest=sha256:05b5bc537ca22b935e4a752adf46dfcf70cd9e0f441378386b37d22034517b76

Observation 1c3237f3-1c29-45e4-aeab-6679c72b1333 · outbound

This paper cites Stablevi- ton: Learning semantic correspondence with latent diffusion model for virtual try-on,.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency Stablevi- ton: Learning semantic correspondence with latent diffusion model for virtual try-on,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:23:42.742148Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T20:23:41.141373Z digest=sha256:3bb8a80e6524afd7697fe8a4c40665553e346db97068486b32088bc2596b1fa0

Observation 2ec9f07f-b1bf-487a-83c9-6a5aa6ad3b21 · outbound

This paper cites Dit: Self- supervised pre-training for document image transformer,.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency Dit: Self- supervised pre-training for document image transformer,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:23:42.679155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T20:23:41.207875Z digest=sha256:a770ca6221ee84d2f6ed49ecfab1126bd657039b972f2839217fd3b25d0b11c7

Observation 7ced8ebe-d9a9-4397-bc07-1699087ad35e · outbound

This paper cites Flownet: Learning optical flow with convolutional net- works,.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency Flownet: Learning optical flow with convolutional net- works,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:23:42.615996Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T20:23:41.212780Z digest=sha256:9826c6fe5b75f09f512333884391b22d485ca5b943e980ffca59e9dbcfbe4f21

Observation 572c88e9-71b4-4d9b-827c-a47eb318a226 · outbound

This paper cites Shineon: Illuminating design choices for practical video-based virtual clothing try-on,.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency Shineon: Illuminating design choices for practical video-based virtual clothing try-on,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:23:42.599265Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T20:23:41.218167Z digest=sha256:3c68de9c39ba8bed40856bd31e0c3a50dda2d05812ffb2a72c7618f8c19047a3

Observation 5d77c051-aaf5-4a6d-8522-4146242327b9 · outbound

This paper cites Mv-ton: Memory-based video virtual try-on network,.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency Mv-ton: Memory-based video virtual try-on network,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:23:42.515659Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T20:23:41.223168Z digest=sha256:e7c0c9b396fe9962833a484b305e58cc7ecb70729d7fe8d85e2bda6a53de9a46

Observation 1f4ad7c9-c713-490d-a639-36fd86fcd7b8 · outbound

This paper cites Fw-gan: Flow-navigated warping gan for video virtual try- on,.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency Fw-gan: Flow-navigated warping gan for video virtual try- on,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:23:42.482507Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T20:23:41.228081Z digest=sha256:ebd0fb06835ac678861011357bc17e835a8c15e733b90d3ad3ec863ad13dbfa2

Observation 98820a1f-14bb-422e-a67c-d46580be4b95 · outbound

This paper cites Gentron: Diffusion transformers for image and video generation,.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency Gentron: Diffusion transformers for image and video generation,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:23:42.436364Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T20:23:41.233133Z digest=sha256:517e17d9591791081d08b974665e98064631277334ce4850d6e6d94a09027a9c

Observation 5ae0294a-2efd-4915-874f-05503c61c6a3 · outbound

This paper cites DiVE: DiT-based Video Generation with Enhanced Control.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency DiVE: DiT-based Video Generation with Enhanced Control

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-10T20:23:41.238260Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:23:41.238260Z digest=sha256:6b6b694e64f80aabc0b846f9d213e2d4b0aabddf41ba58a621e67e1db186b8a1

Observation 6933059d-7d9f-4a4d-acfa-8daed6c78319 · outbound

This paper cites Sora: A Review on Background, Technology, Limitations, and Opportunities of Large Vision Models.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency Sora: A Review on Background, Technology, Limitations, and Opportunities of Large Vision Models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-10T20:23:41.243874Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:23:41.243874Z digest=sha256:66178d3a8f1451e84cec7af871068da79c4efe944440dd31fa051241629b87ea

Observation 8739ab18-7fd0-4818-be72-d880f5617086 · outbound

This paper cites NUWA-XL: Diffusion over Diffusion for eXtremely Long Video Generation.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency NUWA-XL: Diffusion over Diffusion for eXtremely Long Video Generation

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-10T20:23:41.249291Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:23:41.249291Z digest=sha256:b98ab994efab5c48fe60fb3d45264d6e5be2d97afb9e4f1b9e304fc6626c60b1

Observation 93c28c2b-f510-4270-ab7e-0afb0e787922 · outbound

This paper cites Trip: Temporal residual learning with image noise prior for image-to-video diffusion models,.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency Trip: Temporal residual learning with image noise prior for image-to-video diffusion models,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:23:42.398694Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T20:23:41.254119Z digest=sha256:eb970fec2d6315fc069c1e4941bebf88b652fd175f935a946909633e03e0337d

Observation 71dd5b71-2eb1-4d9c-b646-7440dc726f38 · outbound

This paper cites Latent Video Diffusion Models for High-Fidelity Long Video Generation.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency Latent Video Diffusion Models for High-Fidelity Long Video Generation

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-10T20:23:41.258890Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:23:41.258890Z digest=sha256:1b69e43a41461d4d47cf9ecde623ee904df1ab8b0d9442dd8d139dcae40ce8dc

Observation ce4f4262-4437-4568-b641-e7dc5013a8d3 · outbound

This paper cites Multi-concept customization of text-to-image diffu- sion,.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency Multi-concept customization of text-to-image diffu- sion,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:23:42.328075Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T20:23:41.264121Z digest=sha256:1789eb2aa05fb565032d10d9176b3574c1d3ef9ed4705df1b94526fac3fdfc2e

Observation 8ed88b5e-5cc5-43eb-9b1c-cb701fe6c125 · outbound

This paper cites Distilling Diffusion Models into Conditional GANs.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency Distilling Diffusion Models into Conditional GANs

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-10T20:23:41.268741Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:23:41.268741Z digest=sha256:f869ce3654a3b095719eb479d67dc2ce0ee726c52c70186781181214a8ac2ca4

Observation 44fece41-f004-493a-b349-5d3cf8640c70 · outbound

This paper cites Toward characteristic-preserving image-based virtual try- on network,.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency Toward characteristic-preserving image-based virtual try- on network,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:23:42.302802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T20:23:41.280868Z digest=sha256:c75185cd7229a6b5e093c6723553d45ad20736f698d32dd1ae8079ae1903be39

Observation 8374b284-19a3-4721-93c4-0c4fc3eb2d0d · outbound

This paper cites High- resolution virtual try-on with misalignment and occlusion- handled conditions,.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency High- resolution virtual try-on with misalignment and occlusion- handled conditions,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:23:42.217866Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T20:23:41.336274Z digest=sha256:d063c1b70b73ddfa27ecfd4d2c8f4c268a7601ff6d18610cf18d22d000da7670

Observation 47c9771c-594a-4c43-a5b4-72d9329e894c · outbound

This paper cites Ladi-vton: Latent diffusion textual- inversion enhanced virtual try-on,.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency Ladi-vton: Latent diffusion textual- inversion enhanced virtual try-on,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:23:42.199086Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T20:23:41.398737Z digest=sha256:99d3351d0483e557704be3408e4a080768d3a9847b355f6100f779f1fff9eea6

Observation 8c839f01-dd03-4c51-995b-338fa2e623cc · outbound

This paper cites Tam- ing the power of diffusion models for high-quality virtual try- on with appearance flow,.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency Tam- ing the power of diffusion models for high-quality virtual try- on with appearance flow,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:23:42.118515Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T20:23:41.428921Z digest=sha256:f73c66eed5b5f9055e43031c4e9079f2ca9c2ae261b58facfd97195563b985a1

Observation 5482fa6b-05da-47f6-8a76-a93a23448634 · outbound

This paper cites Viton-hd: High- resolution virtual try-on via misalignment-aware normaliza- tion,.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency Viton-hd: High- resolution virtual try-on via misalignment-aware normaliza- tion,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:23:42.083295Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T20:23:41.435225Z digest=sha256:b5bd2c9852101bc66e02e832b899670838a454d2534ca093df0230fe79cd77e7

Observation 647f8ce2-7440-4285-b1af-88ee7cc734b9 · outbound

This paper cites SAM 2: Segment Anything in Images and Videos.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency SAM 2: Segment Anything in Images and Videos

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-10T20:23:41.440256Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:23:41.440256Z digest=sha256:9771fdc5bccaccaeb89fee776637e0c1047caa5e6165bfadc8ff4d8b0611d357

Observation 433d9b05-bee6-4858-89fc-35278f4053e8 · outbound

This paper cites WildVidFit: Video Virtual Try-On in the Wild via Image-Based Controlled Diffusion Models.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency WildVidFit: Video Virtual Try-On in the Wild via Image-Based Controlled Diffusion Models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-10T20:23:41.445965Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:23:41.445965Z digest=sha256:9e087d8b45fb132a33c9d14cdf210f1b27e33d715abe6177587a4e8094904a49

Observation 0d4582a0-57f0-4dcf-a4bd-20ca3c36c912 · outbound

This paper cites Cat-dm: Controllable accelerated virtual try-on with diffu- sion model,.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency Cat-dm: Controllable accelerated virtual try-on with diffu- sion model,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:23:42.046765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T20:23:41.451757Z digest=sha256:dc6c3b152e9c64eebf72b8bcbfca9a36ad5d2b06c98c5cf6284b8ae8ed139301

Observation b8168b2c-a642-422c-8268-9603ad672aff · outbound

This paper cites Multimodal garment designer: Human- centric latent diffusion models for fashion image editing,.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency Multimodal garment designer: Human- centric latent diffusion models for fashion image editing,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:23:41.966118Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T20:23:41.457069Z digest=sha256:3cfbfcf8e6fcfe38f0d3201723b3b5c5779c8a053656a5bdbb3d537ae888635d

Observation ec3d2185-6b87-4195-9aff-73d7949f0f55 · outbound

This paper cites Paint by example: Exemplar-based image editing with diffusion models,.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency Paint by example: Exemplar-based image editing with diffusion models,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:23:41.934849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T20:23:41.462209Z digest=sha256:c221138dd88daaac4cd86571348e8618c953688cfa7c94e2ec8503ad19802744

Observation 814bbeae-7d9b-4220-a198-0a969c59d755 · outbound

This paper cites ViViD: Video Virtual Try-on using Diffusion Models.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency ViViD: Video Virtual Try-on using Diffusion Models

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-10T20:23:41.467754Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:23:41.467754Z digest=sha256:f76f6521f70dde408c185eb4559af3caeb8069a25547fa21afa0d738685de4a9

Observation fd6e0799-3c37-4d05-a737-8bdc91831b0b · outbound

This paper cites Fresco: Spatial- temporal correspondence for zero-shot video translation,.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency Fresco: Spatial- temporal correspondence for zero-shot video translation,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:23:41.916071Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T20:23:41.473455Z digest=sha256:fc91764e8aaca324339180d881f930b2dcaf4dd4a0e2b05de2f2ddfa51ebe380

Observation 9c4ea21c-aaf8-448e-ae38-421aefa52545 · outbound

This paper cites Lvcd: reference-based lineart video colorization with diffusion models,.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency Lvcd: reference-based lineart video colorization with diffusion models,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:23:41.819877Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T20:23:41.478486Z digest=sha256:1b9d9d9ddc1460b00f0ef5a1414e00249c9693020d452d2386eac224938b7c3d

Observation 4b5c0654-a605-497c-ab96-a0ab08d6ca5a · outbound

This paper cites Detectron2,.

RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency Detectron2,

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T20:23:41.799815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-10T20:23:41.484020Z digest=sha256:d096d6b7c9e1c5ce9fcb7b1d58c2e314f799af051d9f8647c666ab3368080666

Pith citing papers

Observation cb931202-32af-476a-97a6-f8f71ed70c7e · inbound

Eevee: Towards Close-up High-resolution Video-based Virtual Try-on cites this paper.

Eevee: Towards Close-up High-resolution Video-based Virtual Try-on RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-17T06:39:10.564830Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-17T06:34:45.045267Z digest=sha256:3c2fc90e0491454b1450d3b6b1fcf608947ba373b35f0f4df29fb9894d2637aa

Observation c504a272-756e-4e04-b20f-0bfaf4d54e58 · inbound

The devil is in the details: Enhancing Video Virtual Try-On via Keyframe-Driven Details Injection cites this paper.

The devil is in the details: Enhancing Video Virtual Try-On via Keyframe-Driven Details Injection RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-16T20:08:22.943585Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-16T20:06:30.109866Z digest=sha256:89ba1b32c9d703470a7c74bc340ed35b682efc437f1846bc05ff9610323020b9

Observation e0287f82-ae83-4d57-9d15-1cc25eb3eb4a · inbound

TripVVT: A Large-Scale Triplet Dataset and a Coarse-Mask Baseline for In-the-Wild Video Virtual Try-On cites this paper.

TripVVT: A Large-Scale Triplet Dataset and a Coarse-Mask Baseline for In-the-Wild Video Virtual Try-On RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency

Reference 27

Resolution
metadata mismatch
arxiv_id, observed 2026-05-09T05:05:12.414338Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-07T05:43:04.644875Z digest=sha256:0d2a33ed8808830a9162e586fc53e990ce36ca3e2004bc48b58dfb7ed1620922

Observation 6058e754-7fc3-4422-bf35-3cb623eac8e9 · inbound

OmniTryOn: Video Try-On Anything at Once! cites this paper.

OmniTryOn: Video Try-On Anything at Once! RealVVT: Towards Photorealistic Video Virtual Try-on via Spatio-Temporal Consistency

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-07-02T22:17:26.409824Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-27T19:00:44.243335Z digest=sha256:fe38e3a83a6c4afe9d15b4f79445d1e5bb4a8a14a9095bc3da33a2b733feb03c