Pith. sign in

Paper Citation Record · LEDGER

Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists

As of 10 August 2026, this Paper Citation Record lists 35 of 35 outbound references and 18 inbound Pith citation observations for arXiv:2502.06734.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.06734 v3

Coverage vector

measured 35 of 35 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-08T14:34:02.117406Z

measured 53 of 53 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 18 of 18 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:21:41.308248Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T07:55:30.838800Z

Reference resolution

35 of 35 outbound references displayed

  • verified exact0
  • verified fuzzy12
  • unresolved23
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 3968fcc7-b0b5-41bb-b816-ec748c4a6f34 · outbound

This paper cites An Image is Worth One Word: Personalizing Text-to-Image Generation using Textual Inversion.

Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists An Image is Worth One Word: Personalizing Text-to-Image Generation using Textual Inversion

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-08T14:34:01.998242Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T14:34:01.998242Z digest=sha256:2b3b1391efec0fc3cddd458e2e6555a4f1a348b240694ac3479b1020c8e02305

Observation 3e8f7153-f15e-4aa8-a82d-36d60faf3cd5 · outbound

This paper cites TokenFlow: Consistent Diffusion Features for Consistent Video Editing.

Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists TokenFlow: Consistent Diffusion Features for Consistent Video Editing

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-08T14:34:02.002068Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T14:34:02.002068Z digest=sha256:49e645df8653f8d73c0da96a8bc67bd64be8349376d6052061a8a99dca621b2b

Observation a9ed8e8d-e873-464c-8a7c-0fb7e7e326fe · outbound

This paper cites AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning.

Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-08T14:34:02.005885Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T14:34:02.005885Z digest=sha256:759e0518fa5e7f175afe8d7086b9ce744bb995b86f0ddf1b942a248a4a9c45fc

Observation ddf45ffd-7775-4269-a74d-f6ceaeacb56e · outbound

This paper cites commas and used as input prompts for Grounded-SAM2 (Liu et al., 2023a; Ravi et al., 2024).

Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists commas and used as input prompts for Grounded-SAM2 (Liu et al., 2023a; Ravi et al., 2024)

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T14:34:02.586471Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T14:34:02.082556Z digest=sha256:440562d41cb422264b242cfa95d62d9797f384cd9285fb9c680375f3f70cf7ba

Observation 56ceb27b-0a23-4396-b94a-19cb1c46bd9f · outbound

This paper cites VIVID-10M: A Dataset and Baseline for Versatile and Interactive Video Local Editing.

Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists VIVID-10M: A Dataset and Baseline for Versatile and Interactive Video Local Editing

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-08T14:34:02.014040Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T14:34:02.014040Z digest=sha256:79def4d9faf927b11ba18bf7917b3c709e4dd0828bea9c8b76beb27d2005ec8f

Observation dc93fe88-f379-40c1-834c-e3698e829b37 · outbound

This paper cites HQ-Edit: A High-Quality Dataset for Instruction-based Image Editing.

Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists HQ-Edit: A High-Quality Dataset for Instruction-based Image Editing

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-08T14:34:02.017686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T14:34:02.017686Z digest=sha256:3d4df6b7fddf85be89c8b5de9f0c61ce7a9d6170bfb8483cff18de1ba44dc020

Observation 8e5ceaf7-b428-4401-8fea-8974b1755fe7 · outbound

This paper cites Video Diffusion Models are Strong Video Inpainter.

Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists Video Diffusion Models are Strong Video Inpainter

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-08T14:34:02.025207Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T14:34:02.025207Z digest=sha256:c781cae887e26d28c3a16a881c1825a64669b3838ae19d936c75951cdcbdcb47

Observation a110379d-fa15-4e05-bd0f-be30a42b816b · outbound

This paper cites Stablev2v: Stablizing shape consistency in video-to-video editing.

Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists Stablev2v: Stablizing shape consistency in video-to-video editing

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-08T14:34:02.029073Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T14:34:02.029073Z digest=sha256:dc010d16f172630a1ac6d350acb52f1b31cee105cf7766df26f4f2f1ed8e92a8

Observation 65508319-28ab-45a2-9db6-7b693d1e3e26 · outbound

This paper cites The video on the left depicts the original video, while the video on the right displays the edited videos.

Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists The video on the left depicts the original video, while the video on the right displays the edited videos

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T14:34:02.536489Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T14:34:02.100540Z digest=sha256:4ae8f85c26c0e1b97b903ce389fefa9492cab21d3858952f11ff62da357c7514

Observation 2943a5e1-0cbd-4409-a54b-0ca3edd5d155 · outbound

This paper cites SDEdit: Guided Image Synthesis and Editing with Stochastic Differential Equations.

Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists SDEdit: Guided Image Synthesis and Editing with Stochastic Differential Equations

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-08T14:34:02.036011Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T14:34:02.036011Z digest=sha256:b3dda54e3c82cc518f8c555e4eb0f11cd4b647be2d11a2e08da55b85ee101f0c

Observation 872c7789-ff4e-46c1-b721-5248140c7282 · outbound

This paper cites ReVideo: Remake a Video with Motion and Content Control.

Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists ReVideo: Remake a Video with Motion and Content Control

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-08T14:34:02.039641Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T14:34:02.039641Z digest=sha256:2e9ee8fc725fc121a0217678f85d95acc29aa36e664879ee8b9beb3b027f90e5

Observation be7ebc54-aa6f-48b9-b111-c69517bea17c · outbound

This paper cites Zero-shot image-to-image translation.

Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists Zero-shot image-to-image translation

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T14:34:02.604917Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T14:34:02.043064Z digest=sha256:d53f5ed368bbd1e8a96dcf4bc4af46299de1451d863505897488e768a7dd2845

Observation c20652ba-8f08-4c35-964a-287171ab34c9 · outbound

This paper cites The 2017 DAVIS Challenge on Video Object Segmentation.

Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists The 2017 DAVIS Challenge on Video Object Segmentation

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-08T14:34:02.046546Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T14:34:02.046546Z digest=sha256:2af1f0c93702711f490a19204d79d324014b7c8476b38c3a4be324bb7c4c708d

Observation 0c8114e1-09fb-4ad0-8faf-59d53d2c1a9c · outbound

This paper cites DreamFusion: Text-to-3D using 2D Diffusion.

Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists DreamFusion: Text-to-3D using 2D Diffusion

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-08T14:34:02.050432Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T14:34:02.050432Z digest=sha256:16119ed306f374dfc95901357043aee40f1ddf26af7e836050e9e4bfcb208aac

Observation bb7bec9a-b80e-4707-93b8-610c8f73af7e · outbound

This paper cites FateZero: Fusing Attentions for Zero-shot Text-based Video Editing.

Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists FateZero: Fusing Attentions for Zero-shot Text-based Video Editing

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-08T14:34:02.054247Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T14:34:02.054247Z digest=sha256:9f5b9ae5d6e2d8536fbf255210d3069362bf73932f878289f4e7932cdf6990b1

Observation 3fc7e47a-e520-4b09-be19-aa1676ca7d34 · outbound

This paper cites SAM 2: Segment Anything in Images and Videos.

Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists SAM 2: Segment Anything in Images and Videos

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-08T14:34:02.065194Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T14:34:02.065194Z digest=sha256:2e4ce0a50774fa095db59bda0ec1d8284497464b3453305d3faffbc07bfc118a

Observation 157abfa0-e1ee-43ce-992a-c0f3a5034dc2 · outbound

This paper cites Plug- and-play diffusion features for text-driven image-to-image translation.

Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists Plug- and-play diffusion features for text-driven image-to-image translation

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T14:34:02.595960Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T14:34:02.068745Z digest=sha256:9c9fc6bd7bce17ec982a5d7cb6d551906d940cd9c00c9868ecdaa219bb1b0eaa

Observation 77b33201-cf32-4aad-a065-9ea7f40b1021 · outbound

This paper cites OmniEdit: Building Image Editing Generalist Models Through Specialist Supervision.

Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists OmniEdit: Building Image Editing Generalist Models Through Specialist Supervision

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-08T14:34:02.072438Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T14:34:02.072438Z digest=sha256:cca63b6931af80a44334323bb42cbe936bb1915ceb8ac9981b97d03ce81377b6

Observation 6d671e7a-02ea-40a6-bbd1-da6aacf34d06 · outbound

This paper cites Zhang, K., Mo, L., Chen, W., Sun, H., and Su, Y.

Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists Zhang, K., Mo, L., Chen, W., Sun, H., and Su, Y

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-08T14:34:02.076020Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T14:34:02.076020Z digest=sha256:7deb277207ff6eb89289f8bce41e9a70973e3b0f476b9727512840f9a83fd3da

Observation dc54726e-b55e-448e-a843-fcd2e0f400be · outbound

This paper cites CoCoCo: Improving Text-Guided Video Inpainting for Better Consistency, Controllability and Compatibility.

Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists CoCoCo: Improving Text-Guided Video Inpainting for Better Consistency, Controllability and Compatibility

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-08T14:34:02.079224Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T14:34:02.079224Z digest=sha256:7d46aa0b897889a7bbc4b3ca2576ee082e24934ed0be81b67b69101f6cacf5bc

Observation c28dd6a3-e0f2-4b05-ba95-b98c0f9cde40 · outbound

This paper cites This limitation prevents us from applying techniques such as ControlNet to repaint a video effectively.

Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists This limitation prevents us from applying techniques such as ControlNet to repaint a video effectively

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T14:34:02.576857Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T14:34:02.086274Z digest=sha256:58f2bdce97d59626b79b0b6ef2236c31504a47b05d6689477fc3e08ecbcc5e80

Observation 19183171-3c17-41da-8f31-ae79636fa012 · outbound

This paper cites Table 5 shows that our expert model outperforms all baselines, achieving the lowest Ewarp (9.02), highest CLIPScore (0.3145), and best Temporal Consistency (0.9781).

Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists Table 5 shows that our expert model outperforms all baselines, achieving the lowest Ewarp (9.02), highest CLIPScore (0.3145), and best Temporal Consistency (0.9781)

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T14:34:02.567351Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T14:34:02.089985Z digest=sha256:79b87ae14ec1e04b88917a039bd49876c86575626dbaca4a4b537b1c7c25130f

Observation 5d8ed50c-0bc1-46d7-be75-99d457685518 · outbound

This paper cites The best results are boldfaced.

Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists The best results are boldfaced

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T14:34:02.556959Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T14:34:02.093464Z digest=sha256:8837e5c3bdf1b35e8d12f02675bca65e5b8aefa29f5d7634d4049514812c7a79

Observation 310b5d8b-27a2-45d2-8c50-e7828f4592c1 · outbound

This paper cites Bottom: The data construction pipeline for Señorita-2M using our inpainter.

Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists Bottom: The data construction pipeline for Señorita-2M using our inpainter

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T14:34:02.547526Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T14:34:02.097023Z digest=sha256:658367455c7f32a8c774f1ed4c522fc37f1514b3f4c9c00646ef242a7fedbd90

Observation 19ae2d9c-29ad-4530-a975-96bab3585997 · outbound

This paper cites an unresolved cited work.

Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-08T14:34:02.525890Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T14:34:02.103896Z digest=sha256:27a7a767cb875ada9980a9a6fd0e4715bd0e53b2da69005d1a707089df10eb8c

Observation 4a10c795-5672-4f74-8382-50f9dcd94ef0 · outbound

This paper cites For object recognition, we utilize CogVLM-video-llama3-chat (Hong et al.,.

Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists For object recognition, we utilize CogVLM-video-llama3-chat (Hong et al.,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T14:34:02.515628Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T14:34:02.107185Z digest=sha256:f097751e1f1447d965f00aad6127eb4d403b1135a4f5f4dd21a78b0f843feb3c

Observation 29e4c9da-64d4-4c24-b5ea-61148795e91a · outbound

This paper cites We set the maximum token length to 120 and use six frames per video.

Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists We set the maximum token length to 120 and use six frames per video

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T14:34:02.505628Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T14:34:02.110454Z digest=sha256:a0e06f2749906b9c24f97b0ef1703d05cc1312bba6cbc092c3d320510c20dab5

Observation f552e8c8-f3b5-4e00-a369-3fd06334b793 · outbound

This paper cites Detect" or.

Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists Detect" or

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T14:34:02.495005Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T14:34:02.114011Z digest=sha256:168d1b9b7627078be3c988c6ccd7a59e846e9cb672d3c6a10b103b9a72a812e1

Observation e6efe788-38df-4990-8af4-7998ff0f43bf · outbound

This paper cites There be.

Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists There be

Reference 592

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T14:34:02.484560Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-08T14:34:02.117406Z digest=sha256:793f1b6b496ac4baeadfaf172614a2c6387a814fcf460d37c2ef24b6b2f89e8d

Observation e9b5e5a9-be66-4812-ba75-4cafddfc9199 · outbound

This paper cites DPM-Solver++: Fast Solver for Guided Sampling of Diffusion Probabilistic Models.

Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists DPM-Solver++: Fast Solver for Guided Sampling of Diffusion Probabilistic Models

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-08-08T14:34:02.032389Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T14:34:02.032389Z digest=sha256:83a4d01f16c797154e1430a990173b8583fa62fd7ef5c047d48cf7fe8a55e968

Observation 8a8d4ad2-12ee-4734-8eb7-2b0ed1cf73de · outbound

This paper cites Hierarchical Text-Conditional Image Generation with CLIP Latents.

Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists Hierarchical Text-Conditional Image Generation with CLIP Latents

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-08T14:34:02.058260Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T14:34:02.058260Z digest=sha256:148e9f2468f94ced6d34b2e59000570aa21e8f31620c0e12b6664c558dd00471

Observation ccf3b396-b07e-4a72-b983-5689db8bcda2 · outbound

This paper cites CogVLM2: Visual Language Models for Image and Video Understanding.

Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists CogVLM2: Visual Language Models for Image and Video Understanding

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-08T14:34:02.009863Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T14:34:02.009863Z digest=sha256:9483ad74c844f2a73693bf3f378a3287b2019221b57494489e0cc0ae147e3009

Observation cda3b44a-09aa-412d-9c2f-51313623c117 · outbound

This paper cites VideoCrafter2: Overcoming Data Limitations for High-Quality Video Diffusion Models.

Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists VideoCrafter2: Overcoming Data Limitations for High-Quality Video Diffusion Models

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-08T14:34:01.989009Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T14:34:01.989009Z digest=sha256:7543fed16cd22ba91c302e5cdb76bc9cba0d19aea80d76fa1ea1dae1e2badd71

Observation 4dbe9b51-1bda-4641-b1c8-78a8ddde0fed · outbound

This paper cites The Llama 3 Herd of Models.

Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists The Llama 3 Herd of Models

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-08T14:34:01.993732Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T14:34:01.993732Z digest=sha256:a5655337d82a4166022e2c282b09db354c5d41f1549f7318ab1da13d0c533c45

Observation 430243a2-6521-4863-93bf-61ca84c74603 · outbound

This paper cites HunyuanVideo: A Systematic Framework For Large Video Generative Models.

Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists HunyuanVideo: A Systematic Framework For Large Video Generative Models

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-08T14:34:02.021500Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T14:34:02.021500Z digest=sha256:dfac238ae0c6b652eed0080706103137b43dc84d23e796f039b9f14886ebbf4e

Pith citing papers

Observation 6467c033-e05e-4c69-ac81-3574c016c6ed · inbound

Step1X-Edit: A Practical Framework for General Image Editing cites this paper.

Step1X-Edit: A Practical Framework for General Image Editing Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists

Reference 74

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T14:36:42.014639Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-11T14:36:41.467429Z digest=sha256:30e90dd0e9f23291b48c51032904d283db22979ff8da2859cdf2423f20755724

Observation 71e8ad13-ce70-4bc8-8281-4caf719b07be · inbound

MiniMax-Remover: Taming Bad Noise Helps Video Object Removal cites this paper.

MiniMax-Remover: Taming Bad Noise Helps Video Object Removal Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T12:21:41.308248Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:21:41.308248Z digest=sha256:73bb4071705978e2237828fde834bbd92f90aa7d7965803bba66798d9a176be8

Observation 6148350f-6329-420b-b9f1-cefb4e840c79 · inbound

O-DisCo-Edit: Object Distortion Control for Unified Realistic Video Editing cites this paper.

O-DisCo-Edit: Object Distortion Control for Unified Realistic Video Editing Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-05T12:27:35.562093Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:27:35.562093Z digest=sha256:0833d6279561bdc119414cd688fccfb0c42243719979f31059070e52482bb636

Observation a5dd8ca7-7353-4efb-966f-0828de6cccde · inbound

EditVerse: Unifying Image and Video Editing and Generation with In-Context Learning cites this paper.

EditVerse: Unifying Image and Video Editing and Generation with In-Context Learning Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists

Reference 35

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T13:46:25.837207Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-18T13:46:05.547669Z digest=sha256:066b89a474d742d13d8aa6b145101732e8be3d6ab9a17d052da5c568ea18127d

Observation 7253c983-e39d-4c33-a4fa-3823e6bc0878 · inbound

VideoCoF: Unified Video Editing with Temporal Reasoner cites this paper.

VideoCoF: Unified Video Editing with Temporal Reasoner Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists

Reference 51

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T00:08:43.126418Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-17T00:08:09.479706Z digest=sha256:df00f7a8c3b93e42974d0bee053873ef7de91640f0a015e8739e639a440dce8d

Observation 3fa7ad1d-f7b6-47ba-9276-1a48340ccc55 · inbound

Under One Sun: Multi-Object Generative Perception of Materials and Illumination cites this paper.

Under One Sun: Multi-Object Generative Perception of Materials and Illumination Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists

Reference 29

Resolution
unresolved
no resolver link, observed 2026-07-13T22:08:47.493022Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T22:08:47.493022Z digest=sha256:0ccd3d2854664b56e75a8575227c57aa3d152cbd80c847cb26c91c213ac3ca29

Observation 4f6801e6-67af-49a5-a181-e1107e417339 · inbound

InsEdit: Towards Instruction-based Visual Editing via Data-Efficient Video Diffusion Models Adaptation cites this paper.

InsEdit: Towards Instruction-based Visual Editing via Data-Efficient Video Diffusion Models Adaptation Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-05-11T06:25:57.579183Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T17:39:59.791758Z digest=sha256:59d22670360c578fa98bc396f8e79f8024ccb76ae6460fdd454362d74add145a

Observation e47282b7-5dac-4410-8ec7-7ffa7add2e6a · inbound

Omni-Fake: Benchmarking Unified Multimodal Social Media Deepfake Detection cites this paper.

Omni-Fake: Benchmarking Unified Multimodal Social Media Deepfake Detection Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists

Reference 104

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T17:06:04.285073Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-09T14:04:52.065878Z digest=sha256:ce761f02f122b647a26b59e52744b9ad412fbb9146defe1eeb7fa8ac3f8e6626

Observation 424be639-87c3-4e15-9fcd-e89a3e71a1be · inbound

LIVEditor-14B: Lightning Unified Video Editing via In-Context Sparse Attention cites this paper.

LIVEditor-14B: Lightning Unified Video Editing via In-Context Sparse Attention Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists

Reference 60

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T00:25:09.361619Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-07-01T00:24:30.573122Z digest=sha256:27636e1ddc502d3918b32230eb592bbe95bfa545032f6f21a195e5fbbb1c170e

Observation afcd0946-6219-4eef-ad83-30a35095024f · inbound

Sparkle: Realizing Lively Instruction-Guided Video Background Replacement via Decoupled Guidance cites this paper.

Sparkle: Realizing Lively Instruction-Guided Video Background Replacement via Decoupled Guidance Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists

Reference 29

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T19:06:11.969943Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-08T12:34:34.315135Z digest=sha256:c334dadd807a608d3decfdef7c75aedf7e799a3a5f225a30a57e9f038ccae17c

Observation 4e5b4a75-01ad-40b5-aa12-3a8d9528dd7e · inbound

InstructAV2AV: Instruction-Guided Audio-Video Joint Editing cites this paper.

InstructAV2AV: Instruction-Guided Audio-Video Joint Editing Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists

Reference 35

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T11:38:14.723253Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-20T11:34:32.558440Z digest=sha256:1fb24d2c953e4bc4ed8ded1d3b02d50a522923d307cfa0e46397f06c03a71fb8

Observation e92d4077-2ccc-4a26-bdcf-63d68177fe92 · inbound

Aurora: Unified Video Editing with a Tool-Using Agent cites this paper.

Aurora: Unified Video Editing with a Tool-Using Agent Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists

Reference 48

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T10:48:12.726167Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-20T10:47:05.038308Z digest=sha256:7dd4fa973f562039b7a8d30d1b2297fbf725d1cfba5ea738f376341241d75746

Observation b513661f-8be0-4224-aa5e-04e27f7f2e72 · inbound

StreamEdit: Training-Free Video Editing via Few-Step Streaming Video Generation cites this paper.

StreamEdit: Training-Free Video Editing via Few-Step Streaming Video Generation Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists

Reference 99

Resolution
malformed identifier
arxiv_id, observed 2026-05-21T04:53:57.681809Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:53:55.386923Z digest=sha256:847f12f2ec0f4c8758fd98b664c2d48a300e79cf141e7c241e42f01a095dd1c2

Observation c6ab60b4-a4a9-4d6d-afcc-873b1cfe21a1 · inbound

StreamEdit: Training-Free Video Editing via Few-Step Streaming Video Generation cites this paper.

StreamEdit: Training-Free Video Editing via Few-Step Streaming Video Generation Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists

Reference 99

Resolution
malformed identifier
arxiv_id, observed 2026-07-01T07:55:30.840943Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-01T07:49:06.391436Z digest=sha256:fcb15a67dc324bc3e975dbb3e59f13e2a6a11f02994862d16a085ecb71a7b949

Observation d23f9d27-f0fb-40c4-b004-66f6d2470acb · inbound

Reasoning to Align: Implicit Reasoning in Diffusion Transformers for Video Editing cites this paper.

Reasoning to Align: Implicit Reasoning in Diffusion Transformers for Video Editing Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists

Reference 61

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T13:44:41.145007Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T13:37:09.892339Z digest=sha256:33d79ea23cc7ca642141be43b1e90eafa43ad2163d1d4ac047db60bfeff942ff

Observation d61e88f4-3fe8-450f-bc2e-7d6c2322071a · inbound

SpongeBob: Sync-Aware Harmonious Audio-Visual Generative Editing cites this paper.

SpongeBob: Sync-Aware Harmonious Audio-Visual Generative Editing Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists

Reference 15

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T12:34:39.557244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T12:07:44.914337Z digest=sha256:345e41ea4d69130ae6e8702abaab1c62f74b69d5922dfce74bf954fe2fa8501d

Observation 47fa062d-9b6a-43d9-8126-df62952a8cfc · inbound

ElasticTTT: Prior-Preserving Test-Time Tuning for Video Editing cites this paper.

ElasticTTT: Prior-Preserving Test-Time Tuning for Video Editing Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-01T07:15:33.659124Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T07:15:33.659124Z digest=sha256:f39377faae4a13d6fc7ec247b974c24a132d4573268daae6f7590e1f81538941

Observation 361c599b-b00a-4c0b-8e2f-6eb2e1900059 · inbound

ElasticTTT: Prior-Preserving Test-Time Tuning for Video Editing cites this paper.

ElasticTTT: Prior-Preserving Test-Time Tuning for Video Editing Se\~norita-2M: A High-Quality Instruction-based Dataset for General Video Editing by Video Specialists

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-03T01:56:59.724051Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T01:56:59.724051Z digest=sha256:14fb6c4c8cef191ba1756155284f6f7c3309b376108e4c034beb379133998ab7