Pith. sign in

Paper Citation Record · LEDGER

Q-VDiT: Towards Accurate Quantization and Distillation of Video-Generation Diffusion Transformers

As of 8 August 2026, this Paper Citation Record lists 52 of 52 outbound references and 8 inbound Pith citation observations for arXiv:2505.22167.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.22167 v1

Coverage vector

measured 52 of 52 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:22:19.245131Z

measured 60 of 60 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 8 of 8 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T19:57:21.407355Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-01T09:45:40.325116Z

Reference resolution

52 of 52 outbound references displayed

  • verified exact1
  • verified fuzzy26
  • unresolved25
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 42d44208-69e9-4fcc-9441-8c13ebb60427 · outbound

This paper cites QuaRot: Outlier-Free 4-Bit Inference in Rotated LLMs.

Q-VDiT: Towards Accurate Quantization and Distillation of Video-Generation Diffusion Transformers QuaRot: Outlier-Free 4-Bit Inference in Rotated LLMs

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T13:22:14.388940Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:22:14.388940Z digest=sha256:ceb754b5539239224e3083cfcd3e8d3d24773eadb8a9571e5ea251cb319082af

Observation 40c4f2ad-234e-4d18-8350-a871dd7ea002 · outbound

This paper cites Flux.1, 2024.

Q-VDiT: Towards Accurate Quantization and Distillation of Video-Generation Diffusion Transformers Flux.1, 2024

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:22:24.533266Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:22:14.478875Z digest=sha256:2ff588b18295cb4036f54646debf34c3cf2a261856a804c255ba06ca9c81c371

Observation 225abec8-677d-4cbd-9e17-2370230b5db9 · outbound

This paper cites W., Fidler, S., and Kreis, K.

Q-VDiT: Towards Accurate Quantization and Distillation of Video-Generation Diffusion Transformers W., Fidler, S., and Kreis, K

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:22:24.371012Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:22:14.595248Z digest=sha256:21c48c54792b3fcc6e4967051722a89fd8b44340dfe6121592c5073cb0c93e4d

Observation 90373be9-13dd-4fbe-aeaf-99e779030ca8 · outbound

This paper cites Q-DiT: Accurate Post-Training Quantization for Diffusion Transformers.

Q-VDiT: Towards Accurate Quantization and Distillation of Video-Generation Diffusion Transformers Q-DiT: Accurate Post-Training Quantization for Diffusion Transformers

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T13:22:14.679938Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:22:14.679938Z digest=sha256:154a6cba58e3f9a9cce80060e54ca1aa32fae3e5aa431b6608f7bbff9d3b50e0

Observation 99e10386-786a-42a3-bc20-fa9ac6979d47 · outbound

This paper cites T., Mittal, S., Emani, M., Vishwanath, V., and Somani, A.

Q-VDiT: Towards Accurate Quantization and Distillation of Video-Generation Diffusion Transformers T., Mittal, S., Emani, M., Vishwanath, V., and Somani, A

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:22:24.195991Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:22:14.773041Z digest=sha256:8dbe28aa751723c40e600380596c8c7305fc9822cfb821dd39b171b705439b98

Observation 8b92dc1c-2ac6-4c30-98ef-1d7928aa49b8 · outbound

This paper cites and Nichol, A.

Q-VDiT: Towards Accurate Quantization and Distillation of Video-Generation Diffusion Transformers and Nichol, A

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T13:22:14.879626Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:22:14.879626Z digest=sha256:77a2c4fb98a1e3dc75404104c9cce8ddef48184f803acd1ab753cfa1457bb953

Observation 73e66e29-b3e5-4e4a-8028-ce67f7e2491f · outbound

This paper cites Reg-ptq: Regression-specialized post-training quantization for fully quantized object detector.

Q-VDiT: Towards Accurate Quantization and Distillation of Video-Generation Diffusion Transformers Reg-ptq: Regression-specialized post-training quantization for fully quantized object detector

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:22:23.986454Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:22:14.973295Z digest=sha256:1cd5f978058891a1043a2f437a3aaad5099409fd868ba763d296e8d892b1afc1

Observation 9efc581a-8040-4922-967a-61e21c4e9258 · outbound

This paper cites Perceptual quality assessment of smartphone photography.

Q-VDiT: Towards Accurate Quantization and Distillation of Video-Generation Diffusion Transformers Perceptual quality assessment of smartphone photography

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T13:22:15.169329Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:22:15.169329Z digest=sha256:4be3e6c86bccbfa5e79e4901a7732e505a34f7e9ed705b565535b1fe89bfd908

Observation f7efd17d-7b02-4aca-85f0-c890982605ed · outbound

This paper cites Relational diffusion distillation for efficient image generation.

Q-VDiT: Towards Accurate Quantization and Distillation of Video-Generation Diffusion Transformers Relational diffusion distillation for efficient image generation

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:22:23.769463Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:22:15.276248Z digest=sha256:63f549916e9aa44ddbc3ce1fb8af1b28c99d9634d622b7d2be5693a05b18ec0d

Observation 90bfd641-b19e-4967-94b9-43bd724397d4 · outbound

This paper cites Mpq-dm: Mixed precision quantization for extremely low bit diffusion models.

Q-VDiT: Towards Accurate Quantization and Distillation of Video-Generation Diffusion Transformers Mpq-dm: Mixed precision quantization for extremely low bit diffusion models

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:22:23.620119Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:22:15.443773Z digest=sha256:93ed58d91efdaee07feaeb757335921b3392e28ce835529f2dcdcbc05edabda3

Observation d3258729-ce03-4162-9886-8b38b4d43b53 · outbound

This paper cites W., and Keutzer, K.

Q-VDiT: Towards Accurate Quantization and Distillation of Video-Generation Diffusion Transformers W., and Keutzer, K

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:22:23.473188Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:22:15.615845Z digest=sha256:12f67c58b7c0dc91ecc983d26a1c22e20ecf46a65fd74ed50dac2426062961c5

Observation 6d9e08ff-bd26-4e8e-ab81-cfdec2384fd8 · outbound

This paper cites Delving deep into rectifiers: Surpassing human-level performance on imagenet classification.

Q-VDiT: Towards Accurate Quantization and Distillation of Video-Generation Diffusion Transformers Delving deep into rectifiers: Surpassing human-level performance on imagenet classification

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T13:22:15.742538Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:22:15.742538Z digest=sha256:eb6fcb7dbc384e5eef17631a4c72c410e01d7da2205015297912c71496b7d1f3

Observation 52e4babf-c224-4b40-9737-37d8f88dc76d · outbound

This paper cites EfficientDM: Efficient Quantization-Aware Fine-Tuning of Low-Bit Diffusion Models.

Q-VDiT: Towards Accurate Quantization and Distillation of Video-Generation Diffusion Transformers EfficientDM: Efficient Quantization-Aware Fine-Tuning of Low-Bit Diffusion Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T13:22:15.860417Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:22:15.860417Z digest=sha256:55b7391c8b6f7308a4af9929bc5733bc3a3c973fd4e7fd59b1449704096af748

Observation eae52490-2fda-4090-9e12-7bfd982d02ef · outbound

This paper cites Ptqd: Accurate post-training quantization for diffusion models.

Q-VDiT: Towards Accurate Quantization and Distillation of Video-Generation Diffusion Transformers Ptqd: Accurate post-training quantization for diffusion models

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:22:23.240856Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:22:15.968162Z digest=sha256:52938ae72110263b10283222722f1ec89c3547d40ed66bc41b88bb0fdba51001

Observation b3a443b6-5165-4796-855c-53475dda801e · outbound

This paper cites Denoising diffusion probabilistic models.

Q-VDiT: Towards Accurate Quantization and Distillation of Video-Generation Diffusion Transformers Denoising diffusion probabilistic models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T13:22:16.117647Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:22:16.117647Z digest=sha256:b69ce780db5ab2a820b94b1fec5818339109f59a8309787e72f73fffc9f2f843

Observation 323a40c9-3090-4006-b106-74b51a9da9b4 · outbound

This paper cites Open-sora, 2024.

Q-VDiT: Towards Accurate Quantization and Distillation of Video-Generation Diffusion Transformers Open-sora, 2024

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:22:23.083920Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:22:16.208110Z digest=sha256:853b9eb75212ddfba8ae5c309d3d87ba2944fe88bb0e1dd6f5df8e5570907b3b

Observation 2d97c5c7-ea68-43b1-be7e-0f6ed0eed961 · outbound

This paper cites LoRA: Low-Rank Adaptation of Large Language Models.

Q-VDiT: Towards Accurate Quantization and Distillation of Video-Generation Diffusion Transformers LoRA: Low-Rank Adaptation of Large Language Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T13:22:16.361887Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:22:16.361887Z digest=sha256:6b401c7853ca7a5b54972a2e02ac3404f16d20d3325b540b4c009d51c3a93efd

Observation 6d543b78-e708-4246-8000-54d87bda162f · outbound

This paper cites Tfmq-dm: Temporal feature maintenance quantization for diffusion models.

Q-VDiT: Towards Accurate Quantization and Distillation of Video-Generation Diffusion Transformers Tfmq-dm: Temporal feature maintenance quantization for diffusion models

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:22:22.952337Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:22:16.515893Z digest=sha256:8ca871ac4ed9e7c7ad92a1be426f20a10aabf8b1d24f58dc214171ce2ce1f874

Observation aab944f5-3a73-49af-b2da-dd4fbd82466a · outbound

This paper cites Vbench: Comprehensive benchmark suite for video generative models.

Q-VDiT: Towards Accurate Quantization and Distillation of Video-Generation Diffusion Transformers Vbench: Comprehensive benchmark suite for video generative models

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:22:22.742449Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:22:16.627765Z digest=sha256:344a26f60974349789765f7b7786667b1a0a11923fc9fe9fd2c653bef86e6d82

Observation a1a72d73-8869-4817-90fd-b8f191d0d3f1 · outbound

This paper cites Musiq: Multi-scale image quality transformer.

Q-VDiT: Towards Accurate Quantization and Distillation of Video-Generation Diffusion Transformers Musiq: Multi-scale image quality transformer

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T13:22:16.762681Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:22:16.762681Z digest=sha256:8622492821d4ac1adc253d0ce1ee625fae359486d2076bee642d5bcce98e65f5

Observation 7bd28cb6-65a6-43ee-8f50-7c8d2285d6c6 · outbound

This paper cites Aesthetic-predictor, 2022.

Q-VDiT: Towards Accurate Quantization and Distillation of Video-Generation Diffusion Transformers Aesthetic-predictor, 2022

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:22:22.494392Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:22:16.897187Z digest=sha256:837b43e164e145e57648a4c7f2b44a4d791646c12234e946d462b8599c74dd64

Observation 92e45033-8877-45b2-ad9a-d33be3e662cc · outbound

This paper cites Svdqunat: Absorbing outliers by low-rank components for 4-bit diffusion models.

Q-VDiT: Towards Accurate Quantization and Distillation of Video-Generation Diffusion Transformers Svdqunat: Absorbing outliers by low-rank components for 4-bit diffusion models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T13:22:17.043124Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:22:17.043124Z digest=sha256:a9ecdf6d32fe12ea81f3423192a08cdf5d22fb90d3cbdd989c01943f8ffe5ea0

Observation 408c8b91-7402-41b8-bae1-7f2d1ac272ae · outbound

This paper cites Q-diffusion: Quantizing diffusion models.

Q-VDiT: Towards Accurate Quantization and Distillation of Video-Generation Diffusion Transformers Q-diffusion: Quantizing diffusion models

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:22:22.248683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:22:17.149968Z digest=sha256:f55d5153e35c3b8916d0ecc8dd709180da77238791d45705234993d07179ee99

Observation a05a1368-8788-4dd2-b4a5-4ae7cbfc5ad1 · outbound

This paper cites Q-dm: An efficient low-bit quantized diffusion model.

Q-VDiT: Towards Accurate Quantization and Distillation of Video-Generation Diffusion Transformers Q-dm: An efficient low-bit quantized diffusion model

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:22:22.065546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:22:17.254209Z digest=sha256:2609157db7529a8a400e0aa41d66707c25f2d37b4ada2cf982367c996aa1715e

Observation 72106f91-22a9-43b2-8c18-76f86e94a4b2 · outbound

This paper cites Diffbir: Toward blind image restoration with generative diffusion prior.

Q-VDiT: Towards Accurate Quantization and Distillation of Video-Generation Diffusion Transformers Diffbir: Toward blind image restoration with generative diffusion prior

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:22:21.826663Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:22:17.404799Z digest=sha256:c312e5008385adbb0acfdb591cd1f84c98e89ced41177539b2539ca8d38a77a5

Observation d0f5f311-aa17-428b-b22d-c72ae042fd5d · outbound

This paper cites Evalcrafter: Benchmarking and evaluating large video generation models.

Q-VDiT: Towards Accurate Quantization and Distillation of Video-Generation Diffusion Transformers Evalcrafter: Benchmarking and evaluating large video generation models

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:22:21.555393Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:22:17.521124Z digest=sha256:6becee670c8131828b48dea353f3d9263217c0a3e33267f32f780b8f72bf5e8b

Observation 37d4edee-b4ad-4b64-9c50-f5b904eff26e · outbound

This paper cites Reactnet: Towards precise binary neural network with generalized activation functions.

Q-VDiT: Towards Accurate Quantization and Distillation of Video-Generation Diffusion Transformers Reactnet: Towards precise binary neural network with generalized activation functions

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:22:21.299496Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:22:17.638786Z digest=sha256:78e2dd2e734dc9d2322e1c9ad3dab3c65ccc05fc88cc7b0a6d725f084bd5a1bf

Observation 15fca04d-d514-431a-b9d1-c6a51528e5d3 · outbound

This paper cites TerDiT: Ternary Diffusion Models with Transformers.

Q-VDiT: Towards Accurate Quantization and Distillation of Video-Generation Diffusion Transformers TerDiT: Ternary Diffusion Models with Transformers

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T13:22:17.757779Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:22:17.757779Z digest=sha256:f88ccfb15ae6dc59b552edf46b8984736799e099833f3572ede5fb3ead1a7358

Observation cbdb5888-dff9-421e-be28-64ef064057a5 · outbound

This paper cites Latte: Latent Diffusion Transformer for Video Generation.

Q-VDiT: Towards Accurate Quantization and Distillation of Video-Generation Diffusion Transformers Latte: Latent Diffusion Transformer for Video Generation

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T13:22:17.824915Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:22:17.824915Z digest=sha256:e4b9c66466c54d58259953bbb5b098a5152e81917ca1adcfd9fa9173ecc917b3

Observation 4c4a7afb-cf45-486d-b404-0b2e52be06df · outbound

This paper cites and Xie, S.

Q-VDiT: Towards Accurate Quantization and Distillation of Video-Generation Diffusion Transformers and Xie, S

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T13:22:17.905892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:22:17.905892Z digest=sha256:b23a93d6c30f3656dff26a5afa6c8e7b644ec72156e354068ab6924d7761396b

Observation d142ae07-177d-4d20-be8e-4e839ce5207a · outbound

This paper cites Compression of convolutional neural networks: A short survey.

Q-VDiT: Towards Accurate Quantization and Distillation of Video-Generation Diffusion Transformers Compression of convolutional neural networks: A short survey

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:22:21.007247Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:22:17.964318Z digest=sha256:58c0b0e6e1ddbd5a6df9df46cde7a18bdb99d5549ddef91b24500ebb920e9a5e

Observation ad7da9c4-8d16-4dfc-b56b-cbed9b599890 · outbound

This paper cites Forward and backward information retention for accurate binary neural networks.

Q-VDiT: Towards Accurate Quantization and Distillation of Video-Generation Diffusion Transformers Forward and backward information retention for accurate binary neural networks

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:22:20.831361Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:22:18.014265Z digest=sha256:b1cd57516e4d3b8840535556e356d5afb8dfc2574e7f961f78dd97648977c897

Observation 20986bf5-0d0c-4d9a-acb2-97619d7785a8 · outbound

This paper cites W., Hallacy, C., Ramesh, A., Goh, G., Agarwal, S., Sastry, G., Askell, A., Mishkin, P., Clark, J., et al.

Q-VDiT: Towards Accurate Quantization and Distillation of Video-Generation Diffusion Transformers W., Hallacy, C., Ramesh, A., Goh, G., Agarwal, S., Sastry, G., Askell, A., Mishkin, P., Clark, J., et al

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T13:22:18.065263Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:22:18.065263Z digest=sha256:7c305bafc6a6623ee6ef311f97dc3a2597e44686402215fd63b2828b3edb6719

Observation 3e533fcf-643a-49b7-955c-41e150957a42 · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

Q-VDiT: Towards Accurate Quantization and Distillation of Video-Generation Diffusion Transformers High-resolution image synthesis with latent diffusion models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T13:22:18.138370Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:22:18.138370Z digest=sha256:bc9ad9e0b60e22478304e4032591b24b3dcf0c9dcd60b41057434d2ccf1de0d0

Observation 9dfc85ce-922a-48ab-aaa9-327edbf0e9ef · outbound

This paper cites Post-training quantization on diffusion models.

Q-VDiT: Towards Accurate Quantization and Distillation of Video-Generation Diffusion Transformers Post-training quantization on diffusion models

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:22:20.687033Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:22:18.222690Z digest=sha256:85380d8af3f0d61a4f021241ac0b78f2933794c5e9f589e7b76dc2d93ade0d1b

Observation f2e6ac20-41a5-46cd-87d4-a3c0bb051fe5 · outbound

This paper cites Denoising Diffusion Implicit Models.

Q-VDiT: Towards Accurate Quantization and Distillation of Video-Generation Diffusion Transformers Denoising Diffusion Implicit Models

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T13:22:18.279026Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:22:18.279026Z digest=sha256:7c36cfad6ab0a85324cd5f189209d4e31416c9adcc43baa38b175ef4df95592d

Observation 0dc5d63a-890d-4b28-815d-add6d080260c · outbound

This paper cites UCF101: A Dataset of 101 Human Actions Classes From Videos in The Wild.

Q-VDiT: Towards Accurate Quantization and Distillation of Video-Generation Diffusion Transformers UCF101: A Dataset of 101 Human Actions Classes From Videos in The Wild

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T13:22:18.342697Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:22:18.342697Z digest=sha256:a9608a349d4095298570f2834aad1c4be111a2d80bd0a4b2d40d6bc1fc6a8781

Observation dbfd2305-85e4-402a-8188-a3b9bcdc8462 · outbound

This paper cites and Deng, J.

Q-VDiT: Towards Accurate Quantization and Distillation of Video-Generation Diffusion Transformers and Deng, J

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T13:22:18.406400Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:22:18.406400Z digest=sha256:fb9d5429e7f005fa256183657489844a3c5346fd542dea892dd09f825590759f

Observation 3f693415-93eb-4e22-8899-e416ba6e79aa · outbound

This paper cites QuEST: Low-bit Diffusion Model Quantization via Efficient Selective Finetuning.

Q-VDiT: Towards Accurate Quantization and Distillation of Video-Generation Diffusion Transformers QuEST: Low-bit Diffusion Model Quantization via Efficient Selective Finetuning

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T13:22:18.467794Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:22:18.467794Z digest=sha256:30d9acde0b6b31bf602a27439a6a461d226295e50feb42a4c62a0c49aa2f7d05

Observation 5f5434b8-0cf7-4697-ba23-6297227cb4bd · outbound

This paper cites C., and Loy, C.

Q-VDiT: Towards Accurate Quantization and Distillation of Video-Generation Diffusion Transformers C., and Loy, C

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:22:20.556048Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:22:18.547599Z digest=sha256:11e9ff5c5622eb3469dd9622e5907ec9c1794b7cb0d4cc73b4f0f889144e1d3f

Observation 7d3d1fc9-29a7-43ef-ba22-2208dff8b6e8 · outbound

This paper cites InternVid: A Large-scale Video-Text Dataset for Multimodal Understanding and Generation.

Q-VDiT: Towards Accurate Quantization and Distillation of Video-Generation Diffusion Transformers InternVid: A Large-scale Video-Text Dataset for Multimodal Understanding and Generation

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T13:22:18.607441Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:22:18.607441Z digest=sha256:0893d9d4fada0e9b4d7ae8e323a03d7afc8f0b5be1ba1676362bdc71dc2aa434

Observation 4e226ceb-9057-44eb-9210-88343ad97e01 · outbound

This paper cites Exploring video quality assessment on user generated contents from aesthetic and technical perspectives.

Q-VDiT: Towards Accurate Quantization and Distillation of Video-Generation Diffusion Transformers Exploring video quality assessment on user generated contents from aesthetic and technical perspectives

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:22:20.374069Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:22:18.667166Z digest=sha256:c1fa811a0a8e3b4baf4169df28ac1f1f46af89dbb0940db56cdfdf298804ad19

Observation 67ca7b77-e310-4a14-a0ef-dda0cd4b0cf3 · outbound

This paper cites PTQ4DiT: Post-training Quantization for Diffusion Transformers.

Q-VDiT: Towards Accurate Quantization and Distillation of Video-Generation Diffusion Transformers PTQ4DiT: Post-training Quantization for Diffusion Transformers

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T13:22:18.742485Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:22:18.742485Z digest=sha256:66923a91ded41510977e93eb4bd1de2879034e7afdf449fb8002800b0dcc80e4

Observation 9be99d43-2628-4b21-87e9-ee4b8becf7bf · outbound

This paper cites Seesr: Towards semantics-aware real-world image super-resolution.

Q-VDiT: Towards Accurate Quantization and Distillation of Video-Generation Diffusion Transformers Seesr: Towards semantics-aware real-world image super-resolution

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:22:20.257269Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:22:18.828047Z digest=sha256:038b8e31c7c2425bf4da129779a15612e8918a523afedeac1e197cc565fa5f31

Observation 319346f7-cb07-4548-961e-a5ac96d4e5cc · outbound

This paper cites Smoothquant: Accurate and efficient post-training quantization for large language models.

Q-VDiT: Towards Accurate Quantization and Distillation of Video-Generation Diffusion Transformers Smoothquant: Accurate and efficient post-training quantization for large language models

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:22:20.070306Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:22:18.883254Z digest=sha256:8b198bfc043812b08bc46a458f6f5e38c902bc4299610b9289debd76a55b3177

Observation e4478665-50af-4895-a7bc-93f0e3507d7b · outbound

This paper cites Cross-image relational knowledge distillation for semantic segmentation.

Q-VDiT: Towards Accurate Quantization and Distillation of Video-Generation Diffusion Transformers Cross-image relational knowledge distillation for semantic segmentation

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:22:19.915750Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:22:18.927942Z digest=sha256:731a373e65ba67175b6e84d9fe6a8b11a4e3f68ec32bf6d24178ae098d4488aa

Observation 8647242a-2de7-427f-9024-6d18698c105e · outbound

This paper cites Multi-party Collaborative Attention Control for Image Customization.

Q-VDiT: Towards Accurate Quantization and Distillation of Video-Generation Diffusion Transformers Multi-party Collaborative Attention Control for Image Customization

Reference 47

Resolution
verified exact
local_arxiv, observed 2026-08-07T13:22:19.445271Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:22:19.008369Z digest=sha256:c64398872df32f42406f963a7e34613c0ed43f7e4de3bb65a166486f7ddad8eb

Observation 2d4d10d1-c9d0-4362-90da-73bc24428fa1 · outbound

This paper cites ViDiT-Q: Efficient and Accurate Quantization of Diffusion Transformers for Image and Video Generation.

Q-VDiT: Towards Accurate Quantization and Distillation of Video-Generation Diffusion Transformers ViDiT-Q: Efficient and Accurate Quantization of Diffusion Transformers for Image and Video Generation

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T13:22:19.011966Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:22:19.011966Z digest=sha256:7b2bf39f284667988eb4ff53957e6423c862cc905db83425092f0162461e03fd

Observation fc47f83c-143d-43d3-941e-21785bf0e418 · outbound

This paper cites Mixdq: Memory-efficient few-step text-to-image diffusion models with metric-decoupled mixed precision quantization.

Q-VDiT: Towards Accurate Quantization and Distillation of Video-Generation Diffusion Transformers Mixdq: Memory-efficient few-step text-to-image diffusion models with metric-decoupled mixed precision quantization

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:22:19.779782Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T13:22:19.016977Z digest=sha256:e2354ebc3877cb08c9800c9b5affc4db299fa5ab2e72940605a42a52fc43a47b

Observation 001445c0-a054-4550-9768-f08cd23d8106 · outbound

This paper cites BiDM: Pushing the Limit of Quantization for Diffusion Models.

Q-VDiT: Towards Accurate Quantization and Distillation of Video-Generation Diffusion Transformers BiDM: Pushing the Limit of Quantization for Diffusion Models

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T13:22:19.082312Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:22:19.082312Z digest=sha256:12285e4a696c667c201e32319032c9cad6bfa3495d5410e12969191cf4184a0f

Observation 4ab4c250-2d33-4397-8f28-0e6eaee11392 · outbound

This paper cites BinaryDM: Accurate Weight Binarization for Efficient Diffusion Models.

Q-VDiT: Towards Accurate Quantization and Distillation of Video-Generation Diffusion Transformers BinaryDM: Accurate Weight Binarization for Efficient Diffusion Models

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T13:22:19.150573Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:22:19.150573Z digest=sha256:075535eb321f93cab5c1820385f8ee69367d63c8ba9d7a924f6d6b121607b853

Observation 7fe76973-9da6-49ee-b460-20caadec97c8 · outbound

This paper cites write newline.

Q-VDiT: Towards Accurate Quantization and Distillation of Video-Generation Diffusion Transformers write newline

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T13:22:19.245131Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:22:19.245131Z digest=sha256:40dff3201e6ecd4835d4e30fd027068efe7522fe0f5926280f2139aedd3b7460

Pith citing papers

Observation 86e26931-433c-4b03-9323-d7a8c20173eb · inbound

MPQ-DMv2: Flexible Residual Mixed Precision Quantization for Low-Bit Diffusion Models with Temporal Distillation cites this paper.

MPQ-DMv2: Flexible Residual Mixed Precision Quantization for Low-Bit Diffusion Models with Temporal Distillation Q-VDiT: Towards Accurate Quantization and Distillation of Video-Generation Diffusion Transformers

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T19:57:21.407355Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:57:21.407355Z digest=sha256:421bfe1b14d9eb0a0656b3ddfb20574ec62bce0ae9e757372f66f6442a5516b2

Observation 70c3ba28-5907-4e58-81f2-9d6de20f1b70 · inbound

Charting the Future of Scholarly Knowledge with AI: A Community Perspective cites this paper.

Charting the Future of Scholarly Knowledge with AI: A Community Perspective Q-VDiT: Towards Accurate Quantization and Distillation of Video-Generation Diffusion Transformers

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-05T15:32:34.035597Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:32:34.035597Z digest=sha256:dd6db616b0b60d7c14f537cda30aed5cd81e0c73f3bf877b9eeeacf791214f84

Observation 339c94ae-8a93-4a77-bc8f-8a0a3936726e · inbound

Efficient Video Diffusion Models: Advancements and Challenges cites this paper.

Efficient Video Diffusion Models: Advancements and Challenges Q-VDiT: Towards Accurate Quantization and Distillation of Video-Generation Diffusion Transformers

Reference 268

Resolution
verified exact
arxiv_id, observed 2026-05-10T09:03:26.011053Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T08:28:29.706249Z digest=sha256:3c9bc10189fa1de7e180255657e720f008e7b3216a4a4bbd9435b8e1fe1ef9e6

Observation e50a9651-5348-45c8-8bf9-5e9b24e45328 · inbound

Motion-Aware Caching for Efficient Autoregressive Video Generation cites this paper.

Motion-Aware Caching for Efficient Autoregressive Video Generation Q-VDiT: Towards Accurate Quantization and Distillation of Video-Generation Diffusion Transformers

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:06:01.664270Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T15:40:03.605239Z digest=sha256:0363161700ea9e18a440d040a7521428b96af2479fce917aa3a320e4e419900b

Observation c2f1e0ec-12c9-459a-ac92-6bb3a2cbdbc5 · inbound

Motion-Aware Caching for Efficient Autoregressive Video Generation cites this paper.

Motion-Aware Caching for Efficient Autoregressive Video Generation Q-VDiT: Towards Accurate Quantization and Distillation of Video-Generation Diffusion Transformers

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-15T07:19:49.953112Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-15T07:16:05.225105Z digest=sha256:983fd57664c8ff8f4d6cfb0a292ae5b2b236e8b7c01b83b9526f5cd969345ec6

Observation 96c9f718-ccb3-4199-a61b-182df46185dc · inbound

GaitProtector: Impersonation-Driven Gait De-Identification via Training-Free Diffusion Latent Optimization cites this paper.

GaitProtector: Impersonation-Driven Gait De-Identification via Training-Free Diffusion Latent Optimization Q-VDiT: Towards Accurate Quantization and Distillation of Video-Generation Diffusion Transformers

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-13T06:42:26.399267Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T06:38:08.198271Z digest=sha256:fd41f442bfdcb7f02bca2edcf09322f19a4a367b0eb6828859a88e73501502f3

Observation 2a6325cd-d752-4a01-9ac2-00aa73dbf8bd · inbound

Echo-Forcing: A Scene Memory Framework for Interactive Long Video Generation cites this paper.

Echo-Forcing: A Scene Memory Framework for Interactive Long Video Generation Q-VDiT: Towards Accurate Quantization and Distillation of Video-Generation Diffusion Transformers

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-20T20:03:43.577064Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-20T20:02:45.505404Z digest=sha256:661c0b5ed42778040b934bd88f3129b809b6c781adc7ad7837d191df9777d583

Observation d33759d6-44da-41f4-aa79-f6fde75df3c1 · inbound

SyncCache: Exploiting Asymmetric Dynamics for Fast Audio-Driven Portrait Animation cites this paper.

SyncCache: Exploiting Asymmetric Dynamics for Fast Audio-Driven Portrait Animation Q-VDiT: Towards Accurate Quantization and Distillation of Video-Generation Diffusion Transformers

Reference 10

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T09:45:40.326505Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-07-01T06:13:11.504526Z digest=sha256:5f9e00bb0c28fdfac2fda99c8f15332e1254dbfe039f2369cbf33e3b299c5154