Pith. sign in

Paper Citation Record · LEDGER

IDAG-Edit: Multi-Object Video Editing via Instance-Decoupled Attention and Guidance

As of 9 August 2026, this Paper Citation Record lists 28 of 28 outbound references and 1 inbound Pith citation observation for arXiv:2606.22042.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2606.22042 v1

Coverage vector

measured 28 of 28 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-06-26T12:43:34.724121Z

measured 29 of 29 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-26T12:43:34.724121Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-04T07:49:39.355337Z

Reference resolution

28 of 28 outbound references displayed

  • verified exact3
  • verified fuzzy0
  • unresolved21
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch4

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation e14e79f9-b237-406b-8572-aa4a2d6098a3 · outbound

This paper cites IDAG-Edit: Multi-Object Video Editing via Instance-Decoupled Attention and Guidance.

IDAG-Edit: Multi-Object Video Editing via Instance-Decoupled Attention and Guidance IDAG-Edit: Multi-Object Video Editing via Instance-Decoupled Attention and Guidance

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-07-04T07:49:39.356838Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-26T12:43:34.724121Z digest=sha256:afcc0d504245d958a91a1bdfe321de9cf7a9c3898bc96d254b99c0d3d4366cac

Observation 7e303af9-295d-4958-9e34-14d3e9dd1010 · outbound

This paper cites Overview Framework As shown in Fig.

IDAG-Edit: Multi-Object Video Editing via Instance-Decoupled Attention and Guidance Overview Framework As shown in Fig

Reference 2

Resolution
unresolved
no resolver link, observed 2026-06-26T12:43:34.724121Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T12:43:34.724121Z digest=sha256:833c8049f43af416adb34279cddfc0db06d2c8864c5d596532f0b582b0329917

Observation b1513b35-f65c-4117-911d-4f6141908c24 · outbound

This paper cites an unresolved cited work.

IDAG-Edit: Multi-Object Video Editing via Instance-Decoupled Attention and Guidance Unresolved cited work

Reference 3

Resolution
unresolved
no resolver link, observed 2026-06-26T12:43:34.724121Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T12:43:34.724121Z digest=sha256:51bb1ee2558d6e14c2f59a894182aad71fa983631968ac7a3b99f8e1ce113556

Observation f27b9877-8e18-43fd-ab14-b642580299e4 · outbound

This paper cites an unresolved cited work.

IDAG-Edit: Multi-Object Video Editing via Instance-Decoupled Attention and Guidance Unresolved cited work

Reference 4

Resolution
unresolved
no resolver link, observed 2026-06-26T12:43:34.724121Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T12:43:34.724121Z digest=sha256:c5779ac0378d1e595780f5f4e3318655602e3b295499e191a0c47711744e7978

Observation 019b7041-adea-415d-8af4-bbee8381394c · outbound

This paper cites an unresolved cited work.

IDAG-Edit: Multi-Object Video Editing via Instance-Decoupled Attention and Guidance Unresolved cited work

Reference 5

Resolution
unresolved
no resolver link, observed 2026-06-26T12:43:34.724121Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T12:43:34.724121Z digest=sha256:ad3afa553615409a8b18b9d44b7e7e60c2ab8499164e9fff0e55168af6db1581

Observation e591eaa2-6ed4-407a-b9a8-17583f8d6390 · outbound

This paper cites FLATTEN: optical FLow-guided ATTENtion for consistent text-to-video editing.

IDAG-Edit: Multi-Object Video Editing via Instance-Decoupled Attention and Guidance FLATTEN: optical FLow-guided ATTENtion for consistent text-to-video editing

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-07-04T07:49:39.367022Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-26T12:43:34.724121Z digest=sha256:7ab3f097739390fc11646151ee8fd69d12d975c10789152bd0a6b9687996e708

Observation 38c1c136-26eb-4eb6-9b20-171852a61ae5 · outbound

This paper cites Flowvid: Taming imperfect optical flows for consistent video-to-video synthesis,.

IDAG-Edit: Multi-Object Video Editing via Instance-Decoupled Attention and Guidance Flowvid: Taming imperfect optical flows for consistent video-to-video synthesis,

Reference 7

Resolution
unresolved
no resolver link, observed 2026-06-26T12:43:34.724121Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T12:43:34.724121Z digest=sha256:6e8818e0700df000ba86965994be2ee47e357911af774560ead0f2b0548b3307

Observation 83888145-6cc6-4366-a36d-c8d400cb7cab · outbound

This paper cites Tune-a-video: One-shot tuning of im- age diffusion models for text-to-video generation,.

IDAG-Edit: Multi-Object Video Editing via Instance-Decoupled Attention and Guidance Tune-a-video: One-shot tuning of im- age diffusion models for text-to-video generation,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-06-26T12:43:34.724121Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T12:43:34.724121Z digest=sha256:fccb854712e3aace2313010397ffca512b9c0ffb1b11ac5d057337c50948614a

Observation 0247e34f-c2f6-4e24-9e30-601382f083ee · outbound

This paper cites ControlVideo: Training-free Controllable Text-to-Video Generation.

IDAG-Edit: Multi-Object Video Editing via Instance-Decoupled Attention and Guidance ControlVideo: Training-free Controllable Text-to-Video Generation

Reference 9

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T07:49:39.347979Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-26T12:43:34.724121Z digest=sha256:a4995122474b16c3dce492c5c8a884e4b9f3cd4ce9c51b4f972e10540b74248b

Observation 48329f95-5c49-4570-aa2d-1a5111960914 · outbound

This paper cites Videograin: Modulating space-time attention for multi- grained video editing,.

IDAG-Edit: Multi-Object Video Editing via Instance-Decoupled Attention and Guidance Videograin: Modulating space-time attention for multi- grained video editing,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-06-26T12:43:34.724121Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T12:43:34.724121Z digest=sha256:8aea1c3e607cf4dff6fd09673f75453abdc4efef40114bf228104b847c07d47b

Observation 6bc0c686-a498-4e38-a11f-c10426f96f59 · outbound

This paper cites Ground-A-Video: Zero-shot Grounded Video Editing using Text-to-image Diffusion Models.

IDAG-Edit: Multi-Object Video Editing via Instance-Decoupled Attention and Guidance Ground-A-Video: Zero-shot Grounded Video Editing using Text-to-image Diffusion Models

Reference 11

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T07:49:39.350674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-26T12:43:34.724121Z digest=sha256:b3a3fb5152610fa699c646ac8cbc86756d75c7d2db5471d2f5d8cae99b2248d5

Observation c8bf65e4-a999-4dd9-a8dc-41cfae508dc7 · outbound

This paper cites Fatezero: Fusing attentions for zero-shot text-based video editing,.

IDAG-Edit: Multi-Object Video Editing via Instance-Decoupled Attention and Guidance Fatezero: Fusing attentions for zero-shot text-based video editing,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-06-26T12:43:34.724121Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T12:43:34.724121Z digest=sha256:554ace0b7a10ed7dee2afc3d9a3e7d0bef8e42c62f4b148e5e12a3a4df7b2656

Observation 3547360a-df69-46c1-9318-7514e0bfe95b · outbound

This paper cites Pix2video: Video editing using image diffusion,.

IDAG-Edit: Multi-Object Video Editing via Instance-Decoupled Attention and Guidance Pix2video: Video editing using image diffusion,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-06-26T12:43:34.724121Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T12:43:34.724121Z digest=sha256:0de74fb9a0b454630540d0bafde739c111fb67367b42c1d2ea10e4b6c9375c21

Observation a2be8264-82e6-49c5-8272-25df6ae46328 · outbound

This paper cites TokenFlow: Consistent Diffusion Features for Consistent Video Editing.

IDAG-Edit: Multi-Object Video Editing via Instance-Decoupled Attention and Guidance TokenFlow: Consistent Diffusion Features for Consistent Video Editing

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-07-04T07:49:39.349833Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-26T12:43:34.724121Z digest=sha256:07235e619489f279ae89477a4da4eb2ebb9fb93967c3b507b3c6db9609f96f3c

Observation fd87b3d6-acb2-4f88-b126-bf3850ea1af8 · outbound

This paper cites High-resolution image synthesis with latent diffusion models,.

IDAG-Edit: Multi-Object Video Editing via Instance-Decoupled Attention and Guidance High-resolution image synthesis with latent diffusion models,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-06-26T12:43:34.724121Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T12:43:34.724121Z digest=sha256:05745b482b3a0390cbcb8f1d165e4ac82a3d03a69286d62e3a4b2d92c2e2fad1

Observation 93aa88d6-b135-4de2-a947-77420f7ef6e3 · outbound

This paper cites Videodirector: Precise video editing via text-to-video models,.

IDAG-Edit: Multi-Object Video Editing via Instance-Decoupled Attention and Guidance Videodirector: Precise video editing via text-to-video models,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-06-26T12:43:34.724121Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T12:43:34.724121Z digest=sha256:bc4c2a45688177ebd4ab3d5f8293a226613d70d4ab52fc40c2686335b6db92a7

Observation 2fd1bcb0-2cac-4c97-8393-b070b931ec90 · outbound

This paper cites AnyV2V: A Tuning-Free Framework For Any Video-to-Video Editing Tasks.

IDAG-Edit: Multi-Object Video Editing via Instance-Decoupled Attention and Guidance AnyV2V: A Tuning-Free Framework For Any Video-to-Video Editing Tasks

Reference 17

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T07:49:39.360807Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-26T12:43:34.724121Z digest=sha256:831f2b491c7289773e09947a19dde0d3c58bcf34cdf755ec9eee6e23cc55fa95

Observation bcbc293f-c2a2-4c30-a384-197a6ac3a016 · outbound

This paper cites Stablev2v: Stablizing shape consistency in video-to- video editing,.

IDAG-Edit: Multi-Object Video Editing via Instance-Decoupled Attention and Guidance Stablev2v: Stablizing shape consistency in video-to- video editing,

Reference 18

Resolution
unresolved
no resolver link, observed 2026-06-26T12:43:34.724121Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T12:43:34.724121Z digest=sha256:2a2edbad5cb55df9d62ccaa92d6c395eb1fb1e769b708415e2087c64b204645f

Observation da2a343d-6e09-4831-a9fa-500ce06e9f1c · outbound

This paper cites Animatediff: Animate your personalized text-to- image diffusion models without specific tuning,.

IDAG-Edit: Multi-Object Video Editing via Instance-Decoupled Attention and Guidance Animatediff: Animate your personalized text-to- image diffusion models without specific tuning,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-06-26T12:43:34.724121Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T12:43:34.724121Z digest=sha256:994c2952c3777f5923a3ea3aba94fc57b398d574dbaa8d3d7481a75b7fb126ed

Observation 9dc34f72-fbf0-4797-aca7-141b2c5c089c · outbound

This paper cites Denoising diffusion implicit models,.

IDAG-Edit: Multi-Object Video Editing via Instance-Decoupled Attention and Guidance Denoising diffusion implicit models,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-06-26T12:43:34.724121Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T12:43:34.724121Z digest=sha256:928af01532d4ed6a2db2077919917e85bb81a401a4e075643212b797ae9c7514

Observation c7d43a11-e123-4b12-acb4-2afdbd8dc002 · outbound

This paper cites Classifier-Free Diffusion Guidance.

IDAG-Edit: Multi-Object Video Editing via Instance-Decoupled Attention and Guidance Classifier-Free Diffusion Guidance

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-07-04T07:49:39.344868Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-26T12:43:34.724121Z digest=sha256:93b6dc3951b05197a7bfead4bdd274d8dafd3fb3eb9c796c3c1a6bab1b28525b

Observation 6569343c-d998-49e6-9ea4-ea67c4deb698 · outbound

This paper cites Diffusion models beat gans on image synthesis,.

IDAG-Edit: Multi-Object Video Editing via Instance-Decoupled Attention and Guidance Diffusion models beat gans on image synthesis,

Reference 22

Resolution
unresolved
no resolver link, observed 2026-06-26T12:43:34.724121Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T12:43:34.724121Z digest=sha256:5c1dbfe2013db34f41da35f03b3c5faa3ac7fcdb9629b430b34ea8690f6e4bac

Observation 8453f64e-ee19-4da7-ab63-afb5989f7da2 · outbound

This paper cites Dense text-to-image generation with attention mod- ulation,.

IDAG-Edit: Multi-Object Video Editing via Instance-Decoupled Attention and Guidance Dense text-to-image generation with attention mod- ulation,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-06-26T12:43:34.724121Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T12:43:34.724121Z digest=sha256:a8419f20ed63b5e6cf8960a4fc44bba1a68467842448bc36e8943d33a75c24a4

Observation 97b15be1-477f-4600-85ee-2cf080f4efd5 · outbound

This paper cites A benchmark dataset and evaluation methodology for video object segmentation,.

IDAG-Edit: Multi-Object Video Editing via Instance-Decoupled Attention and Guidance A benchmark dataset and evaluation methodology for video object segmentation,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-06-26T12:43:34.724121Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T12:43:34.724121Z digest=sha256:b1733455dd52aa57bc8c4554b4350efa2eb00bf2c5c0ec26708df7e9d902944d

Observation 415ae39a-96d8-4d1d-ba1e-3f642459539e · outbound

This paper cites The 4th large-scale video object segmentation challenge - video instance segmen- tation track,.

IDAG-Edit: Multi-Object Video Editing via Instance-Decoupled Attention and Guidance The 4th large-scale video object segmentation challenge - video instance segmen- tation track,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-06-26T12:43:34.724121Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T12:43:34.724121Z digest=sha256:fcd76af3292aa72447a99bc98096cfc80dda49527c1dcaa458ab76f56cc415c1

Observation 8da24bb6-4e2b-4e7d-93c8-a060f1c11975 · outbound

This paper cites Sam 3: Segment anything with concepts,.

IDAG-Edit: Multi-Object Video Editing via Instance-Decoupled Attention and Guidance Sam 3: Segment anything with concepts,

Reference 26

Resolution
unresolved
no resolver link, observed 2026-06-26T12:43:34.724121Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T12:43:34.724121Z digest=sha256:7da00634c33f98f7c7e22c0bbf92dcc5a7a6ff016487705a69350e42b90ef3c2

Observation a96a30d6-0893-4079-8ebb-b7b17ccd8201 · outbound

This paper cites VBench++: Comprehensive and versatile benchmark suite for video gen- erative models,.

IDAG-Edit: Multi-Object Video Editing via Instance-Decoupled Attention and Guidance VBench++: Comprehensive and versatile benchmark suite for video gen- erative models,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-06-26T12:43:34.724121Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T12:43:34.724121Z digest=sha256:e318ed245e4406f0c175e11e6692d985039b4e65bf29c6ba0436c4751c77b141

Observation 9ad08ccb-ccba-44e4-a200-4dd6b4b79627 · outbound

This paper cites Learning transferable visual models from nat- ural language supervision,.

IDAG-Edit: Multi-Object Video Editing via Instance-Decoupled Attention and Guidance Learning transferable visual models from nat- ural language supervision,

Reference 28

Resolution
unresolved
no resolver link, observed 2026-06-26T12:43:34.724121Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T12:43:34.724121Z digest=sha256:1713fac23be0bf56fe8f28576328b392c83f080d5f8b07606dc73a9c4d1b2e8e

Pith citing papers

Observation e14e79f9-b237-406b-8572-aa4a2d6098a3 · inbound

IDAG-Edit: Multi-Object Video Editing via Instance-Decoupled Attention and Guidance cites this paper.

IDAG-Edit: Multi-Object Video Editing via Instance-Decoupled Attention and Guidance IDAG-Edit: Multi-Object Video Editing via Instance-Decoupled Attention and Guidance

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-07-04T07:49:39.356838Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-26T12:43:34.724121Z digest=sha256:afcc0d504245d958a91a1bdfe321de9cf7a9c3898bc96d254b99c0d3d4366cac