Pith. sign in

Paper Citation Record · LEDGER

Latent-Compressed Variational Autoencoder for Video Diffusion Models

As of 6 August 2026, this Paper Citation Record lists 56 of 56 outbound references and 1 inbound Pith citation observation for arXiv:2604.16479.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2604.16479 v1

Coverage vector

measured 56 of 56 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-10T15:23:19.583713Z

measured 57 of 57 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-02T05:01:24.112587Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

56 of 56 outbound references displayed

  • verified exact29
  • verified fuzzy26
  • unresolved0
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation dfed5bca-6fb6-40ee-834b-290c87435230 · outbound

This paper cites Cosmos World Foundation Model Platform for Physical AI.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Cosmos World Foundation Model Platform for Physical AI

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-11T10:41:03.061762Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:3d421a0ffb4358ebbe76b2dd6a3cc9eff01400aeddf31c1034050e8b7e80503f

Observation 299ca983-65ea-478c-a38d-a9abea0c6237 · outbound

This paper cites Frozen in time: A joint video and image encoder for end-to- end retrieval.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Frozen in time: A joint video and image encoder for end-to- end retrieval

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T02:32:20.016766Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:13c05270c6e511ff9917fc2ae12f82abd13869708a73c74a974539e13eec1900

Observation 62de98d4-5b97-4e57-a82c-1cc21d72cc0f · outbound

This paper cites Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-05-11T10:41:03.273409Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:eee0ef9572d80cbb3a47ca536a8f598b7e982f061cbf633c21b82859b66aa7a9

Observation 62664a9d-9ae2-49f5-b8a0-7593a75b9d30 · outbound

This paper cites Video generation models as world simulators.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Video generation models as world simulators

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T02:32:20.022196Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:c4ef2f191b19d534b9d99e4ed57d9879972bd98c8fbe3d8cab4c6c879593a027

Observation 20901210-6fd1-41c5-b864-9ae3cf45ed07 · outbound

This paper cites Deep Compression Autoencoder for Efficient High-Resolution Diffusion Models.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Deep Compression Autoencoder for Efficient High-Resolution Diffusion Models

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:41:03.226989Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:06d0aad53db4e805b11cb93305c152995d6b70c3cbc9609595ac9922b94aed98

Observation 9ef849ae-98c7-451d-abc4-c3364675db64 · outbound

This paper cites Dc-videogen: Efficient video gen- eration with deep compression video autoencoder.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Dc-videogen: Efficient video gen- eration with deep compression video autoencoder

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:41:03.083637Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:f626cf67eff6825837f8e035460d28fd26135b60fe2732b06a53723d1a0346db

Observation bd28493f-3edb-4436-b076-4de25ed501f4 · outbound

This paper cites Dc-ae 1.5: Accelerating dif- fusion model convergence with structured latent space.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Dc-ae 1.5: Accelerating dif- fusion model convergence with structured latent space

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T02:32:20.024781Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:00be8c18848debcee1c080c89c651312ed4ec7dd69792960c1192f8a3628369d

Observation a95f81c1-9a04-4660-8bdb-7c43a13f9e54 · outbound

This paper cites Od- vae: An omni-dimensional video compressor for improving latent video diffusion model.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Od- vae: An omni-dimensional video compressor for improving latent video diffusion model

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T02:32:20.027418Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:130ed054e4d139bcdc0b28a3c3a1e2f21e7994ac9db6743c1cfd8a428cdbc284

Observation 6d4affe2-e4f6-4c8b-9a5f-4db35d603316 · outbound

This paper cites Panda-70m: Captioning 70m videos with multiple cross- modality teachers.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Panda-70m: Captioning 70m videos with multiple cross- modality teachers

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T02:32:20.029737Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:7d9be0f951b2ceedacc63fbe73c2aa30d5c79b1d86c4baddb5a153a21b21b839

Observation 6d9acdb6-ad50-4c36-8fac-137715dc5cf9 · outbound

This paper cites Taming transformers for high-resolution image synthesis.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Taming transformers for high-resolution image synthesis

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T02:32:20.031995Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:f775028c16373327f44b53e807dce0a05575f2fcfdbe2124a9cb2cdb7e836a41

Observation b718b1df-7e54-4312-935b-527521294bf7 · outbound

This paper cites Video generation arena leader- board.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Video generation arena leader- board

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T02:32:20.064237Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:fc7b7eff9a103e152ddd86c28911ce7742e252861ba9c94a9075f7c4987b593a

Observation 5688cdf8-fe5f-463c-b86c-14fdcb0ff23d · outbound

This paper cites Seedance 1.0: Exploring the Boundaries of Video Generation Models.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Seedance 1.0: Exploring the Boundaries of Video Generation Models

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-11T12:09:57.662354Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:4de8dfc3e8d7cefd97591b1d4da6802abde95749531312afc0efd36cf52a42ca

Observation fe0fc68e-b065-4cca-bed9-cdce1f5c7ed0 · outbound

This paper cites Generative adversarial nets.Advances in Neural Information Processing Systems, 27.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Generative adversarial nets.Advances in Neural Information Processing Systems, 27

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T02:32:20.061886Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:2323149edc7f0ee79e6197df909e4ce7591e0eddf9ace670d33e10fe2ab41349

Observation 019c6939-0c20-44b8-ab20-4cf7bd6e40e6 · outbound

This paper cites An introduction to wavelets.IEEE computa- tional science and engineering, 2(2):50–61.

Latent-Compressed Variational Autoencoder for Video Diffusion Models An introduction to wavelets.IEEE computa- tional science and engineering, 2(2):50–61

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T02:32:20.066750Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:24359de4ed89f651da15c0d7ec3479d20fb799c911b6acd63bf1d08c911e356b

Observation 379c8280-c1fa-4db6-8b19-b58afae57b6f · outbound

This paper cites AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning.

Latent-Compressed Variational Autoencoder for Video Diffusion Models AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-05-11T10:41:03.030009Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:926df7271225ad10f06b9fc5b7401afd39d75ebf99317ce6d921bf19b6f891b8

Observation 1b6693c7-a193-487d-a603-9726cd377731 · outbound

This paper cites LTX-Video: Realtime Video Latent Diffusion.

Latent-Compressed Variational Autoencoder for Video Diffusion Models LTX-Video: Realtime Video Latent Diffusion

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-05-11T10:41:03.364378Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:8c858fbb356bf5dded651c088bc518132792a02942bf3ec61e8f7bd005b2eb25

Observation ac0c1624-78af-40ce-9719-35b212adf4b2 · outbound

This paper cites Learnings from Scaling Visual Tokenizers for Reconstruction and Generation.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Learnings from Scaling Visual Tokenizers for Reconstruction and Generation

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:41:03.077326Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:e31ea6c71d33cfeaa215322f77dc81300d7d138c987c4aa34579af97b0dc08b1

Observation 1f1b6ae9-bb4e-4dc1-b469-8e31eb10843c · outbound

This paper cites Latent Video Diffusion Models for High-Fidelity Long Video Generation.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Latent Video Diffusion Models for High-Fidelity Long Video Generation

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-15T04:27:43.523852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:56d3076701f871329221ad1e35d0602d7d0086c909b93149e7779eba510f53ee

Observation 2e8e02ed-5c3f-4c03-9a86-dba9533f73c5 · outbound

This paper cites simple diffusion: End-to-end diffusion for high resolution images.

Latent-Compressed Variational Autoencoder for Video Diffusion Models simple diffusion: End-to-end diffusion for high resolution images

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T02:32:20.070341Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:d34d9b4d0cbeff4a8d90836b66c40a242da4dcff2238527625b218cdd26cd4ae

Observation 2c9a178d-1bb3-46f4-add7-f34f44d93e5c · outbound

This paper cites Image quality metrics: Psnr vs.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Image quality metrics: Psnr vs

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T02:32:20.079767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:90330beac07d06c222ca88245fddbd63bb1a5db2e02702a5cb2ea459f0ddf925

Observation f3c9203b-3fe4-4318-93c9-3c3577e834cd · outbound

This paper cites The Kinetics Human Action Video Dataset.

Latent-Compressed Variational Autoencoder for Video Diffusion Models The Kinetics Human Action Video Dataset

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-05-11T10:41:03.295921Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:493f404250b1d20b7cfefa10f959c3f1d1cfae883518cc2c9340732bc4c8df48

Observation 7325a249-a7d9-48c0-a616-17cfd1c12dd2 · outbound

This paper cites Adam: A Method for Stochastic Optimization.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Adam: A Method for Stochastic Optimization

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-05-11T10:41:03.177123Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:6bd844e200459d06d469483aad0a26a9ed3a602159e7302f5796b259283757e3

Observation b8adbe17-00c9-478a-890c-41f63f0d2cc7 · outbound

This paper cites Auto-Encoding Variational Bayes.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Auto-Encoding Variational Bayes

Reference 23

Resolution
verified exact
local_arxiv, observed 2026-05-11T10:41:03.044934Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:2c86fb1505e865c89b95971650707512bf56eb5a94a25b123a546a0b0cf80e09

Observation d8789e48-97b0-4e7a-b65d-b281bf12358b · outbound

This paper cites VideoPoet: A Large Language Model for Zero-Shot Video Generation.

Latent-Compressed Variational Autoencoder for Video Diffusion Models VideoPoet: A Large Language Model for Zero-Shot Video Generation

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-15T17:51:05.830483Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:b9085cab42869bf568e1f66620fd45bad73963bfeb96ea9db905b9a0bc118fc0

Observation 2bca163d-c335-4628-b5b3-037118c5bc6a · outbound

This paper cites HunyuanVideo: A Systematic Framework For Large Video Generative Models.

Latent-Compressed Variational Autoencoder for Video Diffusion Models HunyuanVideo: A Systematic Framework For Large Video Generative Models

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-05-11T10:41:03.018392Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:7a1d27ab6e3ba38dec65109e7bbd41029d0dbc24a375dedc20305fbf16a90a85

Observation 45ba7cd5-1331-4cd5-8f01-3c90311fa8b0 · outbound

This paper cites Video autoencoder: self-supervised disentanglement of static 3d structure and motion.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Video autoencoder: self-supervised disentanglement of static 3d structure and motion

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T02:32:20.056940Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:5deb552106334dbc2889e46866d92ee4e945046df6ffea807b9699dff35616be

Observation d5a3f91a-7680-48b3-8914-5bd17281410f · outbound

This paper cites Wf-vae: Enhancing video vae by wavelet-driven energy flow for latent video diffusion model.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Wf-vae: Enhancing video vae by wavelet-driven energy flow for latent video diffusion model

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T02:32:20.044026Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:edafef910fea8eb01c114f60dc5abe2ff36868ccb0550e0dd70f1eadb267b993

Observation 8353aabe-6aa1-4242-8a43-366ddb86beb0 · outbound

This paper cites Open-Sora Plan: Open-Source Large Video Generation Model.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Open-Sora Plan: Open-Source Large Video Generation Model

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:41:03.345396Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:378f7f272fe61456ebf565e505dc0b9f14658e6ede58933e1b62d4db4ffb3852

Observation 59b3d51f-4e21-4d0a-9cba-4fba4daa9806 · outbound

This paper cites Hi-VAE: Efficient Video Autoencoding with Global and Detailed Motion.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Hi-VAE: Efficient Video Autoencoding with Global and Detailed Motion

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:41:03.093142Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:e186c85837a7b98c30945eff346cc564016002f05e267c15c62d0345d89dee53

Observation 78409137-931e-4603-9148-3cdba8c041a7 · outbound

This paper cites Decoupled Weight Decay Regularization.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Decoupled Weight Decay Regularization

Reference 30

Resolution
verified exact
local_arxiv, observed 2026-05-11T10:41:03.316506Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:fa37b4a04e5eaa56d6e7e6ae9c8b1e4010e000f54b4cfb36f1a8852ba39693b5

Observation 4c32aa39-611f-420a-95df-449cd8b044ca · outbound

This paper cites Latte: Latent diffusion transformer for video generation.Transactions on Machine Learning Research.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Latte: Latent diffusion transformer for video generation.Transactions on Machine Learning Research

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T02:32:20.051289Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:d14777e5623e59a05773720969592ce887cfab7d6bdc5c8dd3a4d626cba9971a

Observation ba2dd2a9-6785-41a6-a0ce-e476e6b03df6 · outbound

This paper cites OpenVid-1M: A Large-Scale High-Quality Dataset for Text-to-video Generation.

Latent-Compressed Variational Autoencoder for Video Diffusion Models OpenVid-1M: A Large-Scale High-Quality Dataset for Text-to-video Generation

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-14T20:34:53.267725Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:0250910c0388102f3cecece3621c7bf0b38cbc157a09e3b1cd2f0cb7704f1107

Observation 8114c507-9fcb-43bf-835d-136e00231656 · outbound

This paper cites Movie Gen: A Cast of Media Foundation Models.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Movie Gen: A Cast of Media Foundation Models

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-11T14:16:26.778393Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:66a85c7ff9a1f0004b37beb453a626cdca0df20a2bc289bba9ccc4e2ffabb755

Observation 5c6fce5f-89e4-4c23-b771-5806db933590 · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

Latent-Compressed Variational Autoencoder for Video Diffusion Models High-resolution image synthesis with latent diffusion models

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T02:32:20.074653Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:0d6f78e67ee35013b66b1b62652be5e8ace95cb46423523f428c9642b9814d93

Observation 4010fa6e-836f-40a9-a49b-d47def46bd88 · outbound

This paper cites Temporal generative adversarial nets with singular value clipping.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Temporal generative adversarial nets with singular value clipping

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T02:32:20.046459Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:444842d2c460c3eb72bd245f2b9d7c30c576ef38708261aa405ff755b7f985bf

Observation d4c993ac-391b-4521-bcc5-67f7c7a4aca3 · outbound

This paper cites The JPEG 2000 still image compression standard.

Latent-Compressed Variational Autoencoder for Video Diffusion Models The JPEG 2000 still image compression standard

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T02:32:20.054174Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:dd9d77fb2b703f65cd5bcd0d107f7a5254b06173eae3e57032c686536390773a

Observation e3fbc601-646c-47c6-a677-8c997a906602 · outbound

This paper cites Stylegan-v: A continuous video generator with the price, image quality and perks of stylegan2.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Stylegan-v: A continuous video generator with the price, image quality and perks of stylegan2

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T02:32:20.048897Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:d5590a9584c66904e8bf432a37b0bf64caaeef0fc7b6cb56f89b12cd12db57e1

Observation 22b7c1fa-f686-4732-a324-d4e83df66dfe · outbound

This paper cites Improving the Diffusability of Autoencoders.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Improving the Diffusability of Autoencoders

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:41:03.024402Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:a3d0407a727278c50158041c1f8874fe073a43574c86f08758de862e6dca643f

Observation 14da305a-1bd1-4404-9fc6-e7b623cbfa84 · outbound

This paper cites UCF101: A Dataset of 101 Human Actions Classes From Videos in The Wild.

Latent-Compressed Variational Autoencoder for Video Diffusion Models UCF101: A Dataset of 101 Human Actions Classes From Videos in The Wild

Reference 39

Resolution
verified exact
local_arxiv, observed 2026-05-11T10:41:03.251740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:cc0df5f19d4e76326ff3253c04cd11faab27e66319e08763f13d859172bd8978

Observation b5459e31-b2f8-4ed5-a603-69b9b53627d4 · outbound

This paper cites Adapting LLMs to Time Series Forecasting via Temporal Heterogeneity Modeling and Semantic Alignment.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Adapting LLMs to Time Series Forecasting via Temporal Heterogeneity Modeling and Semantic Alignment

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:41:03.307568Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:c5ea3004f3278dead417983965e9f4d60c93d775b2146e90f8ca471df24facb2

Observation e5292539-05bd-4aab-a127-72e04750685d · outbound

This paper cites Haar Wavelet Based Approach for Image Compression and Quality Assessment of Compressed Image.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Haar Wavelet Based Approach for Image Compression and Quality Assessment of Compressed Image

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-07-04T18:11:21.620230Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:6b342bc91a01e62ca08f793cef336bd8058dccaec30e1e0b27d82956773f1513

Observation 75ce7ae5-0377-4ea9-b494-9c7b4ea5a487 · outbound

This paper cites FVD: A new metric for video generation.

Latent-Compressed Variational Autoencoder for Video Diffusion Models FVD: A new metric for video generation

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T02:32:20.014436Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:1ef89e79b9f8afc5a159d2269b8f1338f0d9ffa07e3e926859e5478455917f77

Observation 7a42e7d1-a736-4ea2-b3b7-2ef1d992341c · outbound

This paper cites Attention is all you need.Advances in Neural Information Processing Systems, 30.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Attention is all you need.Advances in Neural Information Processing Systems, 30

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T02:32:20.018956Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:88e6162be08746d741401880b23503d4c030f05419e43cb50442aea9bfae86ac

Observation 1d8fd16f-127a-4aa7-89ac-9dd60d02c1db · outbound

This paper cites Wan: Open and Advanced Large-Scale Video Generative Models.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Wan: Open and Advanced Large-Scale Video Generative Models

Reference 44

Resolution
verified exact
local_arxiv, observed 2026-05-11T10:41:03.215355Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:224e1c6cb285ea2f81b35e6856e55e2ad971bf304abfe584fc1393217d5fcc41

Observation 718de5d1-f640-48e9-b5ef-ad0c8738b35f · outbound

This paper cites Image quality assessment: from error visibility to structural similarity.IEEE Transactions on Image Processing, 13(4):600–612.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Image quality assessment: from error visibility to structural similarity.IEEE Transactions on Image Processing, 13(4):600–612

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T02:32:20.036895Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:d18b16660dc75a0ea23a64dff76bcefe3ad93589355cf98b82e37d271903cd20

Observation f00a44cd-fde9-42b1-8786-25d8e2eae49f · outbound

This paper cites Improved video V AE for latent video diffusion model.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Improved video V AE for latent video diffusion model

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T02:32:20.041693Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:e14ef421f14ef35972b3fdbf7e42aa9817eae5059e5145ca3a0d1436b3194c66

Observation c279d6c0-affa-4d50-842d-6f4cecb259e9 · outbound

This paper cites H3ae: High compression, high speed, and high quality autoencoder for video diffusion models.

Latent-Compressed Variational Autoencoder for Video Diffusion Models H3ae: High compression, high speed, and high quality autoencoder for video diffusion models

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:41:03.162685Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:faf51d2e56fc69658154cb7a00b7e97fc6be64d51860b72094774d28d847b01c

Observation 3f1f18fe-4366-4385-8537-b9c8b50e29db · outbound

This paper cites Learning to generate time-lapse videos using multi-stage dy- namic generative adversarial networks.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Learning to generate time-lapse videos using multi-stage dy- namic generative adversarial networks

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T02:32:20.034424Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:a92797a0106351c73da4bde8fa655c4403fe340f606ff52c3d243c90585c3935

Observation 5636f674-a34d-423c-93c8-18822b8e2238 · outbound

This paper cites CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer.

Latent-Compressed Variational Autoencoder for Video Diffusion Models CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer

Reference 49

Resolution
verified exact
local_arxiv, observed 2026-05-11T10:41:03.053184Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:0c958740c61466c9a3e2765a5f6bed1598a0313a75cfaee5f9b1d7c6ae18cfe9

Observation 77bafbe1-84b6-488a-acff-59a4c5739655 · outbound

This paper cites Language Model Beats Diffusion -- Tokenizer is Key to Visual Generation.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Language Model Beats Diffusion -- Tokenizer is Key to Visual Generation

Reference 50

Resolution
verified exact
arxiv_id, observed 2026-05-13T20:06:45.090605Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:07d7f6ecfdf1a8013b24eb8303fb5a3e255d2f438e680240498832e5d164d4ae

Observation 1c000e58-fda9-4a7a-a9b9-0108e4c15c48 · outbound

This paper cites Efficient Video Diffusion Models via Content-Frame Motion-Latent Decomposition.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Efficient Video Diffusion Models via Content-Frame Motion-Latent Decomposition

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:41:03.353140Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:fd859fb831ff67fae0404bbd3e53a854bd629e2d7bc5ee494070996b50e2797a

Observation 8c7ce93b-2f00-4337-a11b-2cfeeed89b72 · outbound

This paper cites The unreasonable effectiveness of deep features as a perceptual metric.

Latent-Compressed Variational Autoencoder for Video Diffusion Models The unreasonable effectiveness of deep features as a perceptual metric

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T02:32:20.039061Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:4ccb221b1990af3e85117db3f9b22923b5c3ace1571eaf0e9b9590c0f67cd236

Observation b0cd959c-cd32-423a-819f-a2579642c43c · outbound

This paper cites A survey on perceptually optimized video coding.

Latent-Compressed Variational Autoencoder for Video Diffusion Models A survey on perceptually optimized video coding

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T02:32:20.059288Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:3298d053e9dad309891aeb1ecb51f271edbc83f57c0595cee2077a2f7b739f71

Observation 5a609579-f837-4fd6-9024-cd913e77c140 · outbound

This paper cites Cv- vae: A compatible video vae for latent generative video mod- els.Advances in Neural Information Processing Systems, 37: 12847–12871.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Cv- vae: A compatible video vae for latent generative video mod- els.Advances in Neural Information Processing Systems, 37: 12847–12871

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T02:32:20.077188Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:67ff9eee8588f3aa4a3c791474cac4f0baed1db9dd914b71854eda88db742e9c

Observation 192448e4-2953-4271-bbb2-8cd6066d7af5 · outbound

This paper cites Open-Sora: Democratizing Efficient Video Production for All.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Open-Sora: Democratizing Efficient Video Production for All

Reference 55

Resolution
verified exact
arxiv_id, observed 2026-05-11T12:01:52.219350Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:a95adffa57851374938de5329c3c9343de206f73c5e48b005152e087d4becc31

Observation f16c48cd-0194-4afb-9b9a-68c00bc42e9e · outbound

This paper cites Allegro: Open the Black Box of Commercial-Level Video Generation Model.

Latent-Compressed Variational Autoencoder for Video Diffusion Models Allegro: Open the Black Box of Commercial-Level Video Generation Model

Reference 56

Resolution
malformed identifier
arxiv_id, observed 2026-05-11T10:41:03.125071Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:23:19.583713Z digest=sha256:4feb64e3875ca36497f8b2a9ee88c3ad070b8c57155b0599bc5c4004d711a23e

Pith citing papers

Observation 00d907ee-09eb-41de-a468-c6ec1b203ffc · inbound

Kepler-Encoder-v0.1: Towards a Multimodal Embedding Model for Robots cites this paper.

Kepler-Encoder-v0.1: Towards a Multimodal Embedding Model for Robots Latent-Compressed Variational Autoencoder for Video Diffusion Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-02T05:01:24.112587Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T05:01:24.112587Z digest=sha256:1d4821a87ce4ed7344b57c4f70e402cb2400ca5d11f46d744adb8a4b374e92b4