Pith. sign in

Paper Citation Record · LEDGER

VIVID-10M: A Dataset and Baseline for Versatile and Interactive Video Local Editing

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 7 inbound Pith citation observations for arXiv:2411.15260.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.15260 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 7 of 7 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 7 of 7 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T14:34:02.014040Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 56ceb27b-0a23-4396-b94a-19cb1c46bd9f · inbound

Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists cites this paper.

Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists VIVID-10M: A Dataset and Baseline for Versatile and Interactive Video Local Editing

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-08T14:34:02.014040Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T14:34:02.014040Z digest=sha256:79def4d9faf927b11ba18bf7917b3c709e4dd0828bea9c8b76beb27d2005ec8f

Observation 2a3962f7-d4c2-454d-ae84-c82e39445a20 · inbound

MiniMax-Remover: Taming Bad Noise Helps Video Object Removal cites this paper.

MiniMax-Remover: Taming Bad Noise Helps Video Object Removal VIVID-10M: A Dataset and Baseline for Versatile and Interactive Video Local Editing

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T12:21:38.845137Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:21:38.845137Z digest=sha256:bd321249d66670eaef7192a48d017d878dda105dc53ff30710b77080ae8e61ed

Observation 746b9c8e-e72b-4789-9a36-0f9d428cafa3 · inbound

EditVerse: Unifying Image and Video Editing and Generation with In-Context Learning cites this paper.

EditVerse: Unifying Image and Video Editing and Generation with In-Context Learning VIVID-10M: A Dataset and Baseline for Versatile and Interactive Video Local Editing

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-18T13:46:25.767860Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-18T13:46:05.547669Z digest=sha256:15e2b48d162bd4b3d7a3e21095dc0c311dd2cf08157a0d8304a2cd883f32d461

Observation 1107d8e1-f21d-4457-9c6d-3a88dec2943f · inbound

From Ideal to Real: Stable Video Object Removal under Imperfect Conditions cites this paper.

From Ideal to Real: Stable Video Object Removal under Imperfect Conditions VIVID-10M: A Dataset and Baseline for Versatile and Interactive Video Local Editing

Reference 7

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T13:35:51.545452Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-15T13:34:45.782304Z digest=sha256:afb8a1a8c4d8ff25d00e06f18370e1ba956dd9780344259060d0e82e9e34ac25

Observation 9afc401a-2586-47b4-b45c-87c8195dae2f · inbound

MiVE: Multiscale Vision-language features for reference-guided video Editing cites this paper.

MiVE: Multiscale Vision-language features for reference-guided video Editing VIVID-10M: A Dataset and Baseline for Versatile and Interactive Video Local Editing

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-15T05:25:03.032960Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-15T05:23:20.179387Z digest=sha256:7c083841f0ecf2bc97495dfcfc5d840d388785515720797194af1840fc8478ec

Observation 61cc3b24-4d13-4bc2-8044-9cd2058a8b71 · inbound

Diffusing in the Right Space: A Systematic Study of Latent Diffusability cites this paper.

Diffusing in the Right Space: A Systematic Study of Latent Diffusability VIVID-10M: A Dataset and Baseline for Versatile and Interactive Video Local Editing

Reference 52

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T02:46:27.672332Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-28T10:44:24.318786Z digest=sha256:323de11e21c949bd0e0ed38778f03bf98c2513f51708cf7970f4430ae65291cd

Observation de744b06-fdfd-4b1f-a811-e9516906f7d2 · inbound

SteerVTE: Seamless Video Text Editing with Style and Glyph Control cites this paper.

SteerVTE: Seamless Video Text Editing with Style and Glyph Control VIVID-10M: A Dataset and Baseline for Versatile and Interactive Video Local Editing

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-07-04T10:29:44.520621Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-26T08:52:23.330736Z digest=sha256:3ab18e26a13e577af516c17ea30a911f5d37684baec2332070463ce864b778b0