Pith. sign in

Paper Citation Record · LEDGER

ViDiT-Q: Efficient and Accurate Quantization of Diffusion Transformers for Image and Video Generation

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 20 inbound Pith citation observations for arXiv:2406.02540.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2406.02540 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 20 of 20 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 20 of 20 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T14:35:10.690473Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T20:48:55.686449Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation d54927b3-bdd1-42a5-85cc-fb8b08cd906d · inbound

Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity cites this paper.

Sparse VideoGen: Accelerating Video Diffusion Transformers with Spatial-Temporal Sparsity ViDiT-Q: Efficient and Accurate Quantization of Diffusion Transformers for Image and Video Generation

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-09T14:35:10.690473Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T14:35:10.690473Z digest=sha256:2b6a91e6be3a22839820985ba61c7dfc46b267443e4abea39481aface17a8ad6

Observation 7cdbc0f6-9940-4f43-a386-28a5475f1d57 · inbound

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices cites this paper.

On-device Sora: Enabling Training-Free Diffusion-based Text-to-Video Generation for Mobile Devices ViDiT-Q: Efficient and Accurate Quantization of Diffusion Transformers for Image and Video Generation

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-09T10:46:29.620727Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T10:46:29.620727Z digest=sha256:16f40f2d07e65f2115f091b05c0b652b17a4f36ee31659f6d2f5e86d22fb9ff5

Observation 2d4d10d1-c9d0-4362-90da-73bc24428fa1 · inbound

Q-VDiT: Towards Accurate Quantization and Distillation of Video-Generation Diffusion Transformers cites this paper.

Q-VDiT: Towards Accurate Quantization and Distillation of Video-Generation Diffusion Transformers ViDiT-Q: Efficient and Accurate Quantization of Diffusion Transformers for Image and Video Generation

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T13:22:19.011966Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:22:19.011966Z digest=sha256:be0d943980dc721684c9191a249f139d59da809587e81d387649a60c4cebf82a

Observation dd2ab87b-5330-4308-abda-84b02de2c8f3 · inbound

MARch\'e: Fast Masked Autoregressive Image Generation with Cache-Aware Attention cites this paper.

MARch\'e: Fast Masked Autoregressive Image Generation with Cache-Aware Attention ViDiT-Q: Efficient and Accurate Quantization of Diffusion Transformers for Image and Video Generation

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T14:51:37.905137Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:51:37.905137Z digest=sha256:0a827f49503580231e8330731b9f0f88341821e3150e9b0acfa4e5674ab68922

Observation 6f811c1c-0535-4d6c-b3a0-cc77be75548a · inbound

DMQ: Dissecting Outliers of Diffusion Models for Post-Training Quantization cites this paper.

DMQ: Dissecting Outliers of Diffusion Models for Post-Training Quantization ViDiT-Q: Efficient and Accurate Quantization of Diffusion Transformers for Image and Video Generation

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-06T16:42:12.258915Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:42:12.258915Z digest=sha256:9baa28b0bdb77c3559a8a4e1e89b47d2dbb8da7d396d13fb1db2314ed93828af

Observation 083f34a2-4834-46da-a7f6-9f66c3f7baf2 · inbound

QuantVLA: Scale-Calibrated Post-Training Quantization for Vision-Language-Action Models cites this paper.

QuantVLA: Scale-Calibrated Post-Training Quantization for Vision-Language-Action Models ViDiT-Q: Efficient and Accurate Quantization of Diffusion Transformers for Image and Video Generation

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-05-15T20:20:17.351476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T20:20:10.435886Z digest=sha256:af1e3171ce87cef0d641ed485775efaa2900c503ebdbb90f516b607b6a675132

Observation ab7851ae-b2f6-4b56-ac74-96023ac04286 · inbound

Efficient Video Diffusion Models: Advancements and Challenges cites this paper.

Efficient Video Diffusion Models: Advancements and Challenges ViDiT-Q: Efficient and Accurate Quantization of Diffusion Transformers for Image and Video Generation

Reference 193

Resolution
verified exact
arxiv_id, observed 2026-05-10T09:03:26.328251Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T08:28:29.706249Z digest=sha256:ca17a582b848bb94bc907bb9691b6514952fcc1d147bd6afc68a94a9d89f1ff3

Observation 3b211255-acf3-4967-8994-f328a5240e57 · inbound

Motion-Aware Caching for Efficient Autoregressive Video Generation cites this paper.

Motion-Aware Caching for Efficient Autoregressive Video Generation ViDiT-Q: Efficient and Accurate Quantization of Diffusion Transformers for Image and Video Generation

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:06:01.539723Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T15:40:03.605239Z digest=sha256:0e3831d42c47908c651f44d068bfdcaf3d094496a70ca61bb317c8c5b76300fd

Observation 0d2b3cef-452a-46a7-9186-d894336ea830 · inbound

Motion-Aware Caching for Efficient Autoregressive Video Generation cites this paper.

Motion-Aware Caching for Efficient Autoregressive Video Generation ViDiT-Q: Efficient and Accurate Quantization of Diffusion Transformers for Image and Video Generation

Reference 77

Resolution
verified exact
arxiv_id, observed 2026-05-15T07:19:49.890022Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T07:16:05.225105Z digest=sha256:d2d51d5a9d374aad4a571f34d48724f288308e2965da0c5fd28557f1fc068516

Observation 31882355-daf9-4383-8a9c-8c0d513798a1 · inbound

HASTE: Training-Free Video Diffusion Acceleration via Head-Wise Adaptive Sparse Attention cites this paper.

HASTE: Training-Free Video Diffusion Acceleration via Head-Wise Adaptive Sparse Attention ViDiT-Q: Efficient and Accurate Quantization of Diffusion Transformers for Image and Video Generation

Reference 52

Resolution
verified exact
arxiv_id, observed 2026-05-15T05:19:46.183911Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T05:17:48.403406Z digest=sha256:4f89fdfd8132d77c854a3c7efaa8ec8ab99225b2edf1b8d34e4ed45706abbd1d

Observation df078fb6-68b1-40aa-86e2-41cee3039018 · inbound

LongLive-2.0: An NVFP4 Parallel Infrastructure for Long Video Generation cites this paper.

LongLive-2.0: An NVFP4 Parallel Infrastructure for Long Video Generation ViDiT-Q: Efficient and Accurate Quantization of Diffusion Transformers for Image and Video Generation

Reference 78

Resolution
verified exact
arxiv_id, observed 2026-05-20T11:03:13.266687Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-20T11:02:39.465293Z digest=sha256:f68e5eecfb4988b51e67021488ef758fa87a45aa677a0313311f3d37070eb542

Observation 96f09d5e-e1d2-4e3b-8587-c23290eb137b · inbound

{\Omega}-QVLA: Robust Quantization for Vision-Language-Action Models via Composite Rotation and Per-step Scaling cites this paper.

{\Omega}-QVLA: Robust Quantization for Vision-Language-Action Models via Composite Rotation and Per-step Scaling ViDiT-Q: Efficient and Accurate Quantization of Diffusion Transformers for Image and Video Generation

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-06-29T12:53:26.935072Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-29T12:44:50.624778Z digest=sha256:52e4992206c9f85dabd4db7c5c448ad35feab2ab3880a1526bc974b3b8114e7d

Observation 6f97d15b-11b9-4c82-b1bd-decfcbb7e4b7 · inbound

ReCache: Learning Budget-Aware Caching Schedules for Diffusion Models via REINFORCE cites this paper.

ReCache: Learning Budget-Aware Caching Schedules for Diffusion Models via REINFORCE ViDiT-Q: Efficient and Accurate Quantization of Diffusion Transformers for Image and Video Generation

Reference 81

Resolution
verified exact
arxiv_id, observed 2026-07-02T12:26:56.444999Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-28T02:09:44.280357Z digest=sha256:ae9b5f242190a3ab45bbb9998109a6ed3bafba64e8a54fe683b641759f82d05a

Observation 4605c809-6c3f-4935-943c-81e7efd6604d · inbound

Holding the FP8 Quality Ceiling at 8-Bit Weights and Activations: INT8 and GGUF Post-Training Quantization of Ideogram 4.0 for Consumer GPUs cites this paper.

Holding the FP8 Quality Ceiling at 8-Bit Weights and Activations: INT8 and GGUF Post-Training Quantization of Ideogram 4.0 for Consumer GPUs ViDiT-Q: Efficient and Accurate Quantization of Diffusion Transformers for Image and Video Generation

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-07-03T10:17:58.044690Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-27T10:05:40.235610Z digest=sha256:4803e7240af2ddc7801467c3ced175bd5373e016bcb6b5ea1210124f9ea2ff34

Observation ac6aa09c-3e61-43e6-a0df-5c65c2dffeb4 · inbound

MaineCoon: Pursuing A Real-Time Audio-Visual Social World Model cites this paper.

MaineCoon: Pursuing A Real-Time Audio-Visual Social World Model ViDiT-Q: Efficient and Accurate Quantization of Diffusion Transformers for Image and Video Generation

Reference 66

Resolution
verified exact
arxiv_id, observed 2026-07-03T20:48:55.688078Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-27T01:09:02.214590Z digest=sha256:484b942fb42e701130e03513f9bd1a8fabea734e875a044c570358d380099955

Observation 3d3563f6-4eb3-42d6-b32e-7be970599bcd · inbound

OrbitQuant: Data-Agnostic Quantization for Image and Video Diffusion Transformers cites this paper.

OrbitQuant: Data-Agnostic Quantization for Image and Video Diffusion Transformers ViDiT-Q: Efficient and Accurate Quantization of Diffusion Transformers for Image and Video Generation

Reference 53

Resolution
verified exact
arxiv_id, observed 2026-07-03T14:58:32.406533Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-07-03T14:56:10.553212Z digest=sha256:b4d373477f0636989c4a4255181da6821d81469caec0dc520966d7e97f1bb19f

Observation 6f8c294e-a217-4a56-b40c-cf85130b621e · inbound

Awakening Diffusion Transformers: Eliciting Stronger Generation and Understanding via Massive Activation Modulation cites this paper.

Awakening Diffusion Transformers: Eliciting Stronger Generation and Understanding via Massive Activation Modulation ViDiT-Q: Efficient and Accurate Quantization of Diffusion Transformers for Image and Video Generation

Reference 43

Resolution
unresolved
no resolver link, observed 2026-07-12T05:46:45.293902Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T05:46:45.293902Z digest=sha256:bcfcdc90c87c1c1bf14a6bffcf006232bef61e025c2c84cb9c23339ace6dceef

Observation 835b87c2-9afa-4848-9d6f-6b4343c1cd6a · inbound

HyperVAttention: Efficient Sparse Attention with Spatio-Temporal Clustering for Video Diffusion cites this paper.

HyperVAttention: Efficient Sparse Attention with Spatio-Temporal Clustering for Video Diffusion ViDiT-Q: Efficient and Accurate Quantization of Diffusion Transformers for Image and Video Generation

Reference 66

Resolution
unresolved
no resolver link, observed 2026-07-12T05:28:53.474483Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T05:28:53.474483Z digest=sha256:f076f58751d17dd93b9addd5a9a4f1dc8bb95241ac72f443ed2477db15c54ca6

Observation 1b887995-54ba-4010-a5ba-56ef220f3120 · inbound

QuantWAMs: Calibrating at the Right Granularity for World Action Models cites this paper.

QuantWAMs: Calibrating at the Right Granularity for World Action Models ViDiT-Q: Efficient and Accurate Quantization of Diffusion Transformers for Image and Video Generation

Reference 12

Resolution
unresolved
no resolver link, observed 2026-07-31T08:25:32.533730Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T08:25:32.533730Z digest=sha256:899d722a8ff1d5b1c942e085d12f2322e1e261c0da9f9c89a89c46e64e87c2f7

Observation fe47b78e-a71b-47ca-8cdf-9d07e44ac4e9 · inbound

TASQ: Temporal-Adaptive Bit Sparsification Quantization for Diffusion Models cites this paper.

TASQ: Temporal-Adaptive Bit Sparsification Quantization for Diffusion Models ViDiT-Q: Efficient and Accurate Quantization of Diffusion Transformers for Image and Video Generation

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-08T01:04:45.889554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T01:04:45.889554Z digest=sha256:55c714bce2c5d9571028e14979a10b269d0d68a567949adda976e0911d4b2759