Pith. sign in

Paper Citation Record · LEDGER

Edit as You See: Image-guided Video Editing via Masked Motion Modeling

As of 13 August 2026, this Paper Citation Record lists 54 of 54 outbound references and 0 inbound Pith citation observations for arXiv:2501.04325.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.04325 v1

Coverage vector

measured 54 of 54 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T21:41:23.282342Z

measured 54 of 54 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

54 of 54 outbound references displayed

  • verified exact1
  • verified fuzzy40
  • unresolved13
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 42b8a104-7d5a-4122-9737-66354e2c9150 · outbound

This paper cites Blended diffusion for text-driven editing of natural images.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Blended diffusion for text-driven editing of natural images

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.989202Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:41:23.048302Z digest=sha256:9139df8d9563b448a2387cef3dca622c2ee1ee199c81f5410ce59742365d5d99

Observation 9cb97f2d-8cf6-42a1-a1c2-d3eb9da35884 · outbound

This paper cites Text2live: Text-driven layered image and video editing.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Text2live: Text-driven layered image and video editing

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.978351Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:41:23.053362Z digest=sha256:c12c7017c9dcbd234d1cf3bd261dce13c7408be7ff966d1ffc07a209daa8389b

Observation 0fcf37a5-6e79-4c3b-99c2-d52311add416 · outbound

This paper cites In- structpix2pix: Learning to follow image editing instructions.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling In- structpix2pix: Learning to follow image editing instructions

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.967675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:41:23.057864Z digest=sha256:78e698b401eee6492202eff5254fa612a26994545a4d83498cad97cd1d86f261

Observation 3e7be476-1869-486d-9b28-9808f15909bb · outbound

This paper cites Pix2video: Video editing using image diffusion.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Pix2video: Video editing using image diffusion

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.956733Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:41:23.063303Z digest=sha256:d0c926d775f6b4cca9bd133f01c83cd5abb483ee404cd6050e0122ea7a8abfb8

Observation 3fa87c73-9e85-4344-8e4e-48911593c4de · outbound

This paper cites Stable- video: Text-driven consistency-aware diffusion video edit- ing.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Stable- video: Text-driven consistency-aware diffusion video edit- ing

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.944429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:41:23.069225Z digest=sha256:4cbdd7c22c3e63c6ac5ab57dcffab6b8746369ac51003e1c5a6fcfd28da8e12d

Observation a0f5fde4-86e9-4423-b57e-e194863c68c0 · outbound

This paper cites Zero-shot Image Editing with Reference Imitation.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Zero-shot Image Editing with Reference Imitation

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-10T21:41:23.074303Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:41:23.074303Z digest=sha256:b0215fc1ccca2af515b17aa1711987cf02cbe119811947b05d7e3a2d222fb43b

Observation 0e6f196c-371a-4f18-81ea-501dc8838046 · outbound

This paper cites Xmem: Long- term video object segmentation with an atkinson-shiffrin memory model.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Xmem: Long- term video object segmentation with an atkinson-shiffrin memory model

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.931085Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:41:23.079713Z digest=sha256:f30f95403cfcdd58fc56a4fc6516d7fdf9c3f13f5699784893fea7146be82929

Observation cc5a1e3d-7f05-402e-9b34-ebf73f9ca1d0 · outbound

This paper cites Compvis/stable-diffusion: A latent text-to- image diffusion model.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Compvis/stable-diffusion: A latent text-to- image diffusion model

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.918275Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:41:23.084951Z digest=sha256:797c6558fb3347589c9712ef449c463553c4c7a4fef2501f119dbd4fb71b53df

Observation 392b5eee-b0e5-4756-9ac6-ae9d16393a0d · outbound

This paper cites DiffEdit: Diffusion-based semantic image editing with mask guidance.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling DiffEdit: Diffusion-based semantic image editing with mask guidance

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-10T21:41:23.090284Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:41:23.090284Z digest=sha256:f4871a483b9519cfa5d453eee3697ab6fdb3340a98f82cd4038cefe755e65d94

Observation 8e3650e1-78ba-4ee5-bcdf-3b5ca474cc2a · outbound

This paper cites Videdit: Zero-shot and spatially aware text-driven video editing.IEEE Trans.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Videdit: Zero-shot and spatially aware text-driven video editing.IEEE Trans

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.906371Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:41:23.095632Z digest=sha256:962e458469617fc9b86a6b991eee4406add54948c2bef88c9026c90c116b4c4c

Observation 12e6d0ba-8e2b-4c86-b713-2e5465bd330a · outbound

This paper cites Diffusion models beat gans on image synthesis.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Diffusion models beat gans on image synthesis

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.894691Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:41:23.100475Z digest=sha256:e3c3dd5b51b894411520656fb884ef9959c964531cdd9ace5eb8f8920b39da51

Observation b9ff7a3d-bf18-4496-8175-a42f42c738b5 · outbound

This paper cites Editanything: Empower- ing unparalleled flexibility in image editing and generation.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Editanything: Empower- ing unparalleled flexibility in image editing and generation

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.882900Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:41:23.104235Z digest=sha256:724a7cd1e06e467c2a642aff057e8c8973f8c3958873987c7c1237a13d9a4769

Observation 499e42d0-3b1c-4c68-be94-31de4bc362c1 · outbound

This paper cites TokenFlow: Consistent Diffusion Features for Consistent Video Editing.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling TokenFlow: Consistent Diffusion Features for Consistent Video Editing

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-10T21:41:23.107577Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:41:23.107577Z digest=sha256:db2f07a3331fb952e0478badb9034450f8ceae52dde093186d18df0e89a7d882

Observation c61d8e86-da4b-49e6-a4c5-6cf38dbfbb06 · outbound

This paper cites AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-10T21:41:23.111674Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:41:23.111674Z digest=sha256:af1f4696889446813aa786eea95edcb416618a2c586a1a5ae52235da8b44df0f

Observation f1c5f390-96dd-4313-b408-6294732e2a52 · outbound

This paper cites Masked autoencoders are scalable vision learners.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Masked autoencoders are scalable vision learners

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.871666Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:41:23.117261Z digest=sha256:1e7a3c9ccd722cd943380f8428b01b6cb0146488784b3c273bfa642ec6049585

Observation 0bad9104-fd8f-40cb-b153-c7c5ddda50a0 · outbound

This paper cites Denoising dif- fusion probabilistic models.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Denoising dif- fusion probabilistic models

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.861625Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:41:23.121491Z digest=sha256:4d597056aa0b0d7e6946d1f9549306f8823494d90cd2822eed6e70de35979336

Observation 70106e01-8634-4d2a-a464-c6636c5f02f8 · outbound

This paper cites Gritsenko, William Chan, Mohammad Norouzi, and David J.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Gritsenko, William Chan, Mohammad Norouzi, and David J

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.850704Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:41:23.125562Z digest=sha256:ca489f6014a86654581111610854567ceaf718b2eb8005f679570d2f9b55b486

Observation 22eb0cd0-27e1-4441-a6ff-efafc57f659a · outbound

This paper cites Lite- flownet: A lightweight convolutional neural network for op- tical flow estimation.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Lite- flownet: A lightweight convolutional neural network for op- tical flow estimation

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.838500Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:41:23.130447Z digest=sha256:07de46664de38004a45ddf187ec633e3b1ff6a9e91de0ebc75c8eadf4dd49a52

Observation 303cfef8-470b-4e47-955a-e4b3ab7fedd8 · outbound

This paper cites Vmc: Video motion customization using temporal attention adap- tion for text-to-video diffusion models.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Vmc: Video motion customization using temporal attention adap- tion for text-to-video diffusion models

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.825108Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:41:23.135093Z digest=sha256:11d6fcdd6c67c4f4de0cd16986761424f4469f29f1860921fb05d3dedcf5f169

Observation 2fbb744c-715f-4aa0-a16d-64183e423ae0 · outbound

This paper cites Imagic: Text-based real image editing with diffusion models.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Imagic: Text-based real image editing with diffusion models

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.813063Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:41:23.139775Z digest=sha256:b8871f93c550b519dc2b69a59df033603f611eb5b787060f6b4c5b1967a93543

Observation ba7ee9c9-93c2-4980-818b-b5d2eec3005a · outbound

This paper cites Dif- fusionclip: Text-guided diffusion models for robust image manipulation.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Dif- fusionclip: Text-guided diffusion models for robust image manipulation

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.800802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:41:23.144193Z digest=sha256:17527826c34ac64f85837e69189dd282fdc7502eab2b369dfc21d718da64bdd7

Observation 21d1b036-6e13-49c2-9cfa-3c7771612fc2 · outbound

This paper cites Segment any- thing.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Segment any- thing

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.788808Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:41:23.148464Z digest=sha256:2c588abc6ddb3d1620cd116885a58b02804c5116c026f6d2ad8f7132bacf4617

Observation 1c400dbe-334a-4ce3-bc64-2226a555fd0a · outbound

This paper cites Open-sora-plan, 2024.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Open-sora-plan, 2024

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.776350Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:41:23.152779Z digest=sha256:af3299c4c46c0993c5354c04d7c24edb63c2779c8a06e286d39bb4c0933a9b27

Observation 28fe7f0e-30d1-4f9b-a908-377f16bc1f8a · outbound

This paper cites Learning blind video temporal consistency.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Learning blind video temporal consistency

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.764665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:41:23.157139Z digest=sha256:d0b834ea7d9da7056847e296955a13de099faa69c474fd4dd4dd89c2b5f097ad

Observation 7b8feac7-bac0-47ea-910b-a825789673d9 · outbound

This paper cites Generative image dynamics.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Generative image dynamics

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.750869Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:41:23.161189Z digest=sha256:bc678761aa598359c701483286b6285f2fea9f43544c9ed0f57cbff088d47923

Observation 6bf4ffda-2075-41f3-9724-4cd81b2aed1c · outbound

This paper cites Video-p2p: Video editing with cross-attention control.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Video-p2p: Video editing with cross-attention control

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.735746Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:41:23.165082Z digest=sha256:dbb0ea2f13c0cf1c53bc43ce496b137757db4ca05de4795afd11fd7442cb97c8

Observation 351b6968-a739-4a8d-9c88-5f125db30d89 · outbound

This paper cites Null-text inversion for editing real images using guided diffusion models.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Null-text inversion for editing real images using guided diffusion models

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.721659Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:41:23.169506Z digest=sha256:9111e2771e3debd9c038d4686dfa21a215cd5df3dca7ae1dbb4c237518ac4d54

Observation b7d3f3d1-ae61-48c9-8b46-e5f3b9286c63 · outbound

This paper cites Dreamix: Video Diffusion Models are General Video Editors.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Dreamix: Video Diffusion Models are General Video Editors

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-10T21:41:23.174350Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:41:23.174350Z digest=sha256:2a22bfdd610f921d3be296c66b689a07e221a0d950e09919ae6238917a82b881

Observation ff936bca-0c05-49c4-b8f1-d60b135f20cb · outbound

This paper cites GLIDE: Towards Photorealistic Image Generation and Editing with Text-Guided Diffusion Models.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling GLIDE: Towards Photorealistic Image Generation and Editing with Text-Guided Diffusion Models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-10T21:41:23.178964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:41:23.178964Z digest=sha256:fba9f3baace0a012e70bc48fef96d7abb5ab31edaa26a30e66ba0e82eb7b0f7d

Observation 517e7d7a-a1e1-41e9-8386-0a1cad6064ae · outbound

This paper cites The best free stock photos, royalty free images & videos shared by creators.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling The best free stock photos, royalty free images & videos shared by creators

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.707660Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:41:23.183312Z digest=sha256:718678cfffa65def4e1509f86da10244f69adffa072ca83204bb5bcf356eca01

Observation cff524ce-59b4-409a-95cb-5f344c834a7f · outbound

This paper cites Fatezero: Fus- ing attentions for zero-shot text-based video editing.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Fatezero: Fus- ing attentions for zero-shot text-based video editing

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.694231Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:41:23.186953Z digest=sha256:0031ac1ab0563410f2846dafdbeab6af0a1d7508d0e47eedf4f351707fdc8c05

Observation e44e7ee3-8a06-474e-841b-62de9cf95fcd · outbound

This paper cites High-resolution image syn- thesis with latent diffusion models.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling High-resolution image syn- thesis with latent diffusion models

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.680598Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:41:23.189983Z digest=sha256:92f4d710e5c2dc1cc0a5a7c84718456c22047e08fd10516b3d783eb2ae80104e

Observation f95c7446-3d04-448c-94d5-0893dc9c7a6a · outbound

This paper cites pytorch-fid: FID Score for PyTorch.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling pytorch-fid: FID Score for PyTorch

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-10T21:41:23.193229Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:41:23.193229Z digest=sha256:0806a2971f3ff668851e24146eb991edfaa3ddcc8ccadec997b34ad187a0a7f5

Observation 318fe1bd-0d9b-4182-a75b-4daf35332f0c · outbound

This paper cites Motion-i2v: Consistent and controllable image-to-video generation with explicit motion modeling.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Motion-i2v: Consistent and controllable image-to-video generation with explicit motion modeling

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.658961Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:41:23.197259Z digest=sha256:93e6bb9decffa21f551b70f488c72e2c9c4353cdc21feb25fa723476ed2896a6

Observation 25c25d60-4a7e-4505-91ac-52ecb2aad6a5 · outbound

This paper cites Dragdiffusion: Harnessing diffusion models for interactive point-based image editing.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Dragdiffusion: Harnessing diffusion models for interactive point-based image editing

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.647253Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:41:23.201069Z digest=sha256:19cb8d778c2809ca91cb3c4877bb45271a3215b7d14ffcf2b0a410041c8de75e

Observation 8536298f-97f5-47d7-8605-5557146a5510 · outbound

This paper cites Make-A-Video: Text-to-Video Generation without Text-Video Data.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Make-A-Video: Text-to-Video Generation without Text-Video Data

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-10T21:41:23.205165Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:41:23.205165Z digest=sha256:3d714482c0f12dfb78b1a7b7d6dafbe8d1322b061ec3744f73ccb217913ed4a0

Observation 083f91fd-af46-4d96-9ebc-9182b87255c9 · outbound

This paper cites Denoising Diffusion Implicit Models.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Denoising Diffusion Implicit Models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-10T21:41:23.210493Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:41:23.210493Z digest=sha256:065d4e8443b6577d2e516cdcc4cbafcb852ff42f59bddedbe3b5b2cdd47d33bb

Observation c352fc04-3141-449a-adb6-56113b1a92d3 · outbound

This paper cites Score-based generative modeling through stochastic differential equa- tions.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Score-based generative modeling through stochastic differential equa- tions

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.634223Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:41:23.214836Z digest=sha256:57b7d57656d17ba3b1dee70bf0fb0b5c736bb49faf2061358dc919320267b9cb

Observation 502916d0-8bbb-479d-8e24-5307fb3577bb · outbound

This paper cites Videomae: Masked autoencoders are data-efficient learners for self-supervised video pre-training.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Videomae: Masked autoencoders are data-efficient learners for self-supervised video pre-training

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.618168Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:41:23.218661Z digest=sha256:84fcd7f13c6923f861d34b281cf5a8cde6c94e7a14b327f222232c5429cc62bf

Observation ab4bff79-767a-4a6e-b913-a29a6cfb6975 · outbound

This paper cites Zero-Shot Video Editing Using Off-The-Shelf Image Diffusion Models.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Zero-Shot Video Editing Using Off-The-Shelf Image Diffusion Models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-10T21:41:23.222710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:41:23.222710Z digest=sha256:d8298cbce3143303ea8f59b47dd1c760711752f80391489368fdb1f6cd3caa98

Observation 70055fc3-79bc-432d-a0cc-067148524f7c · outbound

This paper cites Latent Image Animator: Learning to Animate Images via Latent Space Navigation.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Latent Image Animator: Learning to Animate Images via Latent Space Navigation

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-10T21:41:23.227332Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:41:23.227332Z digest=sha256:323523242e637a13464598b4082bf17732061d5b40d04fa6cc206646b537ed02

Observation 199495d0-fb7e-4021-9684-0400d74f479a · outbound

This paper cites Dynamicrafter: Animating open-domain images with video diffusion priors.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Dynamicrafter: Animating open-domain images with video diffusion priors

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.604875Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:41:23.231637Z digest=sha256:3923f65994e986ae590580f76214edecb119358aace891ed5870c3c10a601102

Observation 1d8ae152-cffd-4991-a443-6e2d8181d4d4 · outbound

This paper cites Gmflow: Learning optical flow via global matching.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Gmflow: Learning optical flow via global matching

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.591772Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:41:23.235517Z digest=sha256:b7a4420fa2313e124435eb252a2be7008cdd46450895a72073c87b2b814ec3cf

Observation 9910b07b-a008-48bb-9a9e-ae9c377efd05 · outbound

This paper cites MagicProp: Diffusion-based Video Editing via Motion-aware Appearance Propagation.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling MagicProp: Diffusion-based Video Editing via Motion-aware Appearance Propagation

Reference 44

Resolution
verified exact
local_arxiv, observed 2026-08-10T21:41:23.352369Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:41:23.239654Z digest=sha256:a4bb2071fe9aba92b3f05ba6faf595bb59538d4a1eeb6a9ecc42af70e0dd3274

Observation c0ed8248-855a-4756-8e54-53116e186248 · outbound

This paper cites Motion-Conditioned Image Animation for Video Editing.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Motion-Conditioned Image Animation for Video Editing

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-10T21:41:23.244051Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:41:23.244051Z digest=sha256:2668ca3bfac6856b33e7fd1bc5f689273f065aee7d6747cbded5394646f5e9e6

Observation b5c1d703-ac6c-4917-83d6-66ab981fce06 · outbound

This paper cites Paint by example: Exemplar-based image editing with diffusion models.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Paint by example: Exemplar-based image editing with diffusion models

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.576761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:41:23.248392Z digest=sha256:978b81e2e381860c865fc980b67d351c8dcc58f6f4882d552c5748692b65da7b

Observation 2c8b40b6-e398-46f4-a5b6-572285736344 · outbound

This paper cites Depth anything: Unleashing the power of large-scale unlabeled data.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Depth anything: Unleashing the power of large-scale unlabeled data

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.561239Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:41:23.252222Z digest=sha256:dd6c1f2beae2f3747be6005333e22c0b9f8b936de9d242bb17fe5989efb4662b

Observation d8a2c9ab-1301-4f81-9103-53f53e9856b8 · outbound

This paper cites Rerender a video: Zero-shot text-guided video-to-video translation.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Rerender a video: Zero-shot text-guided video-to-video translation

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.547123Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:41:23.256073Z digest=sha256:5611845df83ebbbcf507b43742373814a425213760968003d86800290beafb65

Observation c8588efb-512a-4e5b-b889-c58d44f1794e · outbound

This paper cites IP-Adapter: Text Compatible Image Prompt Adapter for Text-to-Image Diffusion Models.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling IP-Adapter: Text Compatible Image Prompt Adapter for Text-to-Image Diffusion Models

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-10T21:41:23.259850Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:41:23.259850Z digest=sha256:a511e10bfdbdc4a34244b289070f1016b88f490382a666e25a494dd516d83da1

Observation 0171d1bb-3607-4c92-97a5-a5bcbec45c7a · outbound

This paper cites Adding conditional control to text-to-image diffusion models.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Adding conditional control to text-to-image diffusion models

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.534380Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:41:23.264348Z digest=sha256:fe5cbdcc14ca3de427f2fc4cfa6cfd340a07f43694c8071d8ab16067b41d0d28

Observation 9dd04a1e-5fb1-45db-877a-6230a2db62ba · outbound

This paper cites Sine: Single image editing with text-to-image diffusion models.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Sine: Single image editing with text-to-image diffusion models

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.522591Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:41:23.268609Z digest=sha256:8bfa0f84971d7fc4790c0df75d3a8a9ca54b986a17f36651153552154af925a0

Observation 9c9fb6a3-1622-4fcc-9732-5645b813b3f3 · outbound

This paper cites Avid: Any-length video inpainting with dif- fusion model.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Avid: Any-length video inpainting with dif- fusion model

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.509453Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:41:23.273112Z digest=sha256:7bbd61d5c65dbd4d02569d2cae60ab9c2e0a66da9bfa9b7b61e83998527c3d32

Observation 655872eb-92ff-4783-a0db-3fd0baa50417 · outbound

This paper cites Motiondirector: Motion customization of text-to-video diffusion models.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling Motiondirector: Motion customization of text-to-video diffusion models

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.494936Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:41:23.277490Z digest=sha256:1d77106885d1804f79a6e647776260fc7f0418f713412bf37e041905da036fba

Observation b79f31ca-b2ec-4a36-9ceb-54da19575f76 · outbound

This paper cites clip-score: CLIP Score for Py- Torch.

Edit as You See: Image-guided Video Editing via Masked Motion Modeling clip-score: CLIP Score for Py- Torch

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:41:23.480291Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=pdf_text observed=2026-08-10T21:41:23.282342Z digest=sha256:47ae10c85af787e9774b3ecdefeb777b17c27a58c92f7d3fe9bdfe35f4181291

Pith citing papers

No inbound Pith citation observations are available.