Pith. sign in

Paper Citation Record · LEDGER

I2VShield: An Efficient Proactive Defense Framework against DiT-based Image-to-Video Models

As of 8 August 2026, this Paper Citation Record lists 13 of 13 outbound references and 0 inbound Pith citation observations for arXiv:2607.25522.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.25522 v2

Coverage vector

measured 13 of 13 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-01T02:13:30.231270Z

measured 13 of 13 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

13 of 13 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved13
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation b858a6e9-b9bd-4656-ae14-1a9ec512d219 · outbound

This paper cites Towards Deep Learning Models Resistant to Adversarial Attacks.

I2VShield: An Efficient Proactive Defense Framework against DiT-based Image-to-Video Models Towards Deep Learning Models Resistant to Adversarial Attacks

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-01T02:13:29.359126Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:13:29.359126Z digest=sha256:c78472c9eb4e283eb7b1b7950bb91b6e530a77b0c89eba86129fc484de38e6b5

Observation 81c4dfa5-c9f3-4ea9-b8d0-cc597a9775ec · outbound

This paper cites Wan: Open and Advanced Large-Scale Video Generative Models.

I2VShield: An Efficient Proactive Defense Framework against DiT-based Image-to-Video Models Wan: Open and Advanced Large-Scale Video Generative Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-01T02:13:29.820660Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:13:29.820660Z digest=sha256:38572c9fffec9bd7cc179003dc72040fcacda977aae287dd8873a34d76681b4a

Observation b6dc54cf-ace0-4b21-aa70-171e202aae57 · outbound

This paper cites Q-Align: Teaching LMMs for Visual Scoring via Discrete Text-Defined Levels.

I2VShield: An Efficient Proactive Defense Framework against DiT-based Image-to-Video Models Q-Align: Teaching LMMs for Visual Scoring via Discrete Text-Defined Levels

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-01T02:13:29.918103Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:13:29.918103Z digest=sha256:5633f2d521862838ccefeeb8dfd9349a7269dca6ce6a6f7fbb62ef0c6f39a00e

Observation 7197715f-2a52-46bd-a2fb-9ba0a20af3d8 · outbound

This paper cites VideoGPT: Video Generation using VQ-VAE and Transformers.

I2VShield: An Efficient Proactive Defense Framework against DiT-based Image-to-Video Models VideoGPT: Video Generation using VQ-VAE and Transformers

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-01T02:13:30.015968Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:13:30.015968Z digest=sha256:9d1eaf43c696bc8b689233ab5382149c5564fad9f82994c6d3920d455e204fcc

Observation 69603e0f-1328-43ce-b422-2868916449f0 · outbound

This paper cites Open-Sora: Democratizing Efficient Video Production for All.

I2VShield: An Efficient Proactive Defense Framework against DiT-based Image-to-Video Models Open-Sora: Democratizing Efficient Video Production for All

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-01T02:13:30.128863Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:13:30.128863Z digest=sha256:ffa3f631e0ea839cfba41ea8e34dd23a3e649802282e1a41ed405d4adc515447

Observation f735d82a-135e-40cf-b573-9cda589f6372 · outbound

This paper cites Here, we further provide the quantitative results on each dataset separately for a more detailed comparison.

I2VShield: An Efficient Proactive Defense Framework against DiT-based Image-to-Video Models Here, we further provide the quantitative results on each dataset separately for a more detailed comparison

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-01T02:13:30.231270Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:13:30.231270Z digest=sha256:481603390942ba021b1b2a447a73c2a19ddbc1a4dbcea566e9a33fecf66f6517

Observation ab6034de-5a4c-4ae9-8e23-65e797b7c084 · outbound

This paper cites Latte: Latent Diffusion Transformer for Video Generation.

I2VShield: An Efficient Proactive Defense Framework against DiT-based Image-to-Video Models Latte: Latent Diffusion Transformer for Video Generation

Reference 2016

Resolution
unresolved
no resolver link, observed 2026-08-01T02:13:29.204402Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:13:29.204402Z digest=sha256:301465870fabee8a8544d186904eb00b4ed7f12c46c03a4aafd222e0973900bb

Observation d961e942-61da-413d-9744-790366284d06 · outbound

This paper cites Adversarial Machine Learning at Scale.

I2VShield: An Efficient Proactive Defense Framework against DiT-based Image-to-Video Models Adversarial Machine Learning at Scale

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-08-01T02:13:29.029322Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:13:29.029322Z digest=sha256:7a63b86ad95833d3744c13b0d02d9a5e72f36036d1be3106540b516a2288b7d8

Observation dfe4aeee-a307-40b9-987e-2286e9c56f9b · outbound

This paper cites Raising the Cost of Malicious AI-Powered Image Editing.

I2VShield: An Efficient Proactive Defense Framework against DiT-based Image-to-Video Models Raising the Cost of Malicious AI-Powered Image Editing

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-08-01T02:13:29.497557Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:13:29.497557Z digest=sha256:8c89019b621632b5ae64d51be09179e330a5f9f6057333d39e3f2489d51dc3de

Observation f145e9c2-1dbf-406c-9de5-42c755b165e3 · outbound

This paper cites CogVideo: Large-scale Pretraining for Text-to-Video Generation via Transformers.

I2VShield: An Efficient Proactive Defense Framework against DiT-based Image-to-Video Models CogVideo: Large-scale Pretraining for Text-to-Video Generation via Transformers

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-01T02:13:28.831727Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:13:28.831727Z digest=sha256:0b1fa8c7b8160233aedbf1446c064e33fa2c71f786460fd5382406a6e822525e

Observation c6f02ae3-f422-4e3f-8ee8-489d1e694d2f · outbound

This paper cites UCF101: A Dataset of 101 Human Actions Classes From Videos in The Wild.

I2VShield: An Efficient Proactive Defense Framework against DiT-based Image-to-Video Models UCF101: A Dataset of 101 Human Actions Classes From Videos in The Wild

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-01T02:13:29.632575Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:13:29.632575Z digest=sha256:0e5a3f91e91cd3768fe09b067bb71c242d1a998ef019aeec0268f5f67bd7d53f

Observation 81b7a5cb-1195-4b9b-811c-79832d78b517 · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

I2VShield: An Efficient Proactive Defense Framework against DiT-based Image-to-Video Models An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-01T02:13:28.652966Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:13:28.652966Z digest=sha256:9f3ef2d0c7d4cef91a4e108bf1a3fab8ed7ce1725b421cfdc3383a831ee6421c

Observation e81fb3b3-b884-4420-b1d3-0e8ea09f605a · outbound

This paper cites AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning.

I2VShield: An Efficient Proactive Defense Framework against DiT-based Image-to-Video Models AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-01T02:13:28.720596Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:13:28.720596Z digest=sha256:4cda4cd7561c4542dc8d65493d5ac81699ec046a9a8810f3cf8fbe7aa3cb5e23

Pith citing papers

No inbound Pith citation observations are available.