Pith. sign in

Paper Citation Record · LEDGER

Diffusion Autoencoders are Scalable Image Tokenizers

As of 17 August 2026, this Paper Citation Record lists 68 of 68 outbound references and 4 inbound Pith citation observations for arXiv:2501.18593.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.18593 v1

Coverage vector

measured 68 of 68 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-09T22:56:36.235654Z

measured 72 of 72 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-11T11:23:34.104214Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T14:28:29.656526Z

Reference resolution

68 of 68 outbound references displayed

  • verified exact0
  • verified fuzzy23
  • unresolved45
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a314c0d1-094e-452f-8124-db83151f0e06 · outbound

This paper cites write newline.

Diffusion Autoencoders are Scalable Image Tokenizers write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-09T22:56:35.990857Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T22:56:35.990857Z digest=sha256:3709b6736a89aac207f0cc4ceed4a1ef80a649ddbca8fc0cd07a4a5c31ed50a7

Observation a73694ac-a45d-4ed8-bad6-f1116f4b25c3 · outbound

This paper cites Building Normalizing Flows with Stochastic Interpolants.

Diffusion Autoencoders are Scalable Image Tokenizers Building Normalizing Flows with Stochastic Interpolants

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-09T22:56:35.995742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T22:56:35.995742Z digest=sha256:035755f605a4f8f8ec50642d072851cbf69e76f08a9a977fff64dc6b22f05001

Observation 51f69fb6-8201-49ca-b14c-d800777fd0f0 · outbound

This paper cites Layer Normalization.

Diffusion Autoencoders are Scalable Image Tokenizers Layer Normalization

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-09T22:56:36.000541Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T22:56:36.000541Z digest=sha256:c21b6c34abde2d12668e27338625cab76846fbbc2fbdc83e1b5991dd61f90a47

Observation 03e7c0c6-4f34-4b09-a876-78b1235abdeb · outbound

This paper cites BEiT : B ert pre-training of image transformers.

Diffusion Autoencoders are Scalable Image Tokenizers BEiT : B ert pre-training of image transformers

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:56:37.243775Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-09T22:56:36.004819Z digest=sha256:48045a0c879d8e9f5ded3fcaf35978c229e7a8d161e5accb58e66f05d6a8dc66

Observation bd473836-3366-4c46-8d6c-e79173e98b25 · outbound

This paper cites Improving image generation with better captions.

Diffusion Autoencoders are Scalable Image Tokenizers Improving image generation with better captions

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-09T22:56:36.008688Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T22:56:36.008688Z digest=sha256:741a04bfc0a7ca0716e5c3dbf8698c5e6e62d29635516da7a6f157997fcd95dd

Observation bd147583-64f2-48fd-a6b8-39dd7732d907 · outbound

This paper cites an unresolved cited work.

Diffusion Autoencoders are Scalable Image Tokenizers Unresolved cited work

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-09T22:56:36.012692Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T22:56:36.012692Z digest=sha256:582608fe20213952cda5aebfdbf92e06b2969fc985e381026bb9abd1e563febc

Observation c1d34611-a480-4551-93b1-1c27609859e9 · outbound

This paper cites W., Fidler, S., and Kreis, K.

Diffusion Autoencoders are Scalable Image Tokenizers W., Fidler, S., and Kreis, K

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:56:37.222936Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-09T22:56:36.016572Z digest=sha256:35dcbd0265312e1a8f839623b77eb930596f85b25ccda64f470f60f06c006594

Observation 69b81172-35c2-46c2-8467-ea960fdc8e06 · outbound

This paper cites Pros and cons of gan evaluation measures.

Diffusion Autoencoders are Scalable Image Tokenizers Pros and cons of gan evaluation measures

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:56:37.209506Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-09T22:56:36.020768Z digest=sha256:7b8f030b08297d1a5d5650620d98f366403486cad1f38f1254ac4f8a07f81a06

Observation 73b084b9-e534-4c81-b9cc-2d5678e7cc4a · outbound

This paper cites Pros and cons of gan evaluation measures: New developments.

Diffusion Autoencoders are Scalable Image Tokenizers Pros and cons of gan evaluation measures: New developments

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:56:37.195969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-09T22:56:36.024342Z digest=sha256:6f93cc619e0101269aef401800590b288c0bc8141ae358198c8cada259be2c08

Observation 71b65e7b-ca2c-4f69-8a54-7bf4a4ea6b4b · outbound

This paper cites Emerging properties in self-supervised vision transformers.

Diffusion Autoencoders are Scalable Image Tokenizers Emerging properties in self-supervised vision transformers

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:56:37.182061Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-09T22:56:36.028222Z digest=sha256:bb608eebce88ff2c2df3ed267c6bff469f28fff2914b5637f67d7ad6ee1e2d12

Observation 341bcac7-b61d-4fea-ad9d-d282be6da10f · outbound

This paper cites A simple framework for contrastive learning of visual representations.

Diffusion Autoencoders are Scalable Image Tokenizers A simple framework for contrastive learning of visual representations

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:56:37.169439Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-09T22:56:36.032253Z digest=sha256:8f1bbdfa3ccbb62c4e6ab33af4b56f74ec5d4077f0e98133b96e499ba18fc9d2

Observation 70758ad5-69f3-401b-be15-d993f4ab1a4e · outbound

This paper cites Image neural field diffusion models.

Diffusion Autoencoders are Scalable Image Tokenizers Image neural field diffusion models

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:56:37.157002Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-09T22:56:36.035922Z digest=sha256:1e83a758b939e2a26583af9f240f694cf2c8095839cd3c44a4e9f578483695ed

Observation 93ebca15-40e5-4583-960e-0959d235062c · outbound

This paper cites Emu: Enhancing Image Generation Models Using Photogenic Needles in a Haystack.

Diffusion Autoencoders are Scalable Image Tokenizers Emu: Enhancing Image Generation Models Using Photogenic Needles in a Haystack

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-09T22:56:36.039705Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T22:56:36.039705Z digest=sha256:09b4dea564829984b4c4d6a9af66ca3661b511ea496a2090fa40bf7f4899cb24

Observation 26aabed6-ca60-4b88-86fa-93cfb366d1c6 · outbound

This paper cites Imagenet: A large-scale hierarchical image database.

Diffusion Autoencoders are Scalable Image Tokenizers Imagenet: A large-scale hierarchical image database

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-09T22:56:36.043727Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T22:56:36.043727Z digest=sha256:ea66a3abcf4f7a8d4f628faf6673099dd9d3b36be8244a79aeadaa164e51b188

Observation 91ac52c1-a11e-4e61-ac08-c63e585700a7 · outbound

This paper cites and Nichol, A.

Diffusion Autoencoders are Scalable Image Tokenizers and Nichol, A

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-09T22:56:36.047299Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T22:56:36.047299Z digest=sha256:f0d8259151ee0998f5265beb9d33cc4d0ca934e4f8f1c0f1114aa89576e07fee

Observation f128788f-40b9-4c8d-a4a5-d29e1ab2d499 · outbound

This paper cites Adversarial feature learning.

Diffusion Autoencoders are Scalable Image Tokenizers Adversarial feature learning

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:56:37.130822Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-09T22:56:36.052169Z digest=sha256:7d5f7dfa32b98783d65fa89ce48ad1f292960a7b3b931f5b71e5355e90b96fc8

Observation 370e115a-019e-4d3f-8fc3-e45ee9ebc0f4 · outbound

This paper cites Taming transformers for high-resolution image synthesis.

Diffusion Autoencoders are Scalable Image Tokenizers Taming transformers for high-resolution image synthesis

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-09T22:56:36.055986Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T22:56:36.055986Z digest=sha256:3222c13bff1e3bf9a1a60c1d0cd955a66222c372c34dce56d6709da9090c875b

Observation 1c0bdc77-daf9-4d5d-a6c7-979593719cff · outbound

This paper cites Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning.

Diffusion Autoencoders are Scalable Image Tokenizers Emu Video: Factorizing Text-to-Video Generation by Explicit Image Conditioning

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-09T22:56:36.059679Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T22:56:36.059679Z digest=sha256:5b221ac034488100464ad77851886e8e6b455b2a58017e10271baf6c8ba9ef26

Observation 60393605-e100-44fd-82fc-88ef34a2d397 · outbound

This paper cites Generative adversarial networks.

Diffusion Autoencoders are Scalable Image Tokenizers Generative adversarial networks

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-09T22:56:36.067740Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T22:56:36.067740Z digest=sha256:1cf406e8da69714877d3b40d95c8bcca165787c7a56e7ae2e27d80de9b3fa554

Observation 57a91fb8-b825-45af-979c-7a08be27d20b · outbound

This paper cites Bootstrap your own latent-a new approach to self-supervised learning.

Diffusion Autoencoders are Scalable Image Tokenizers Bootstrap your own latent-a new approach to self-supervised learning

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-09T22:56:36.070814Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T22:56:36.070814Z digest=sha256:598e10e9e1d27fbca163fecf13bfbbfb476ef590eccefc7e59231dc8156584cd

Observation 6b6c396b-90e0-4839-ad37-5e9f083842f5 · outbound

This paper cites Momentum contrast for unsupervised visual representation learning.

Diffusion Autoencoders are Scalable Image Tokenizers Momentum contrast for unsupervised visual representation learning

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-09T22:56:36.073560Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T22:56:36.073560Z digest=sha256:3cd8f9bb7d4ff2532d5b73dd5aa961431ac75646c0bf71f66428dd14d65d5d9f

Observation 5cadd866-2cfd-4cc9-8a48-669e560d9eba · outbound

This paper cites Masked autoencoders are scalable vision learners.

Diffusion Autoencoders are Scalable Image Tokenizers Masked autoencoders are scalable vision learners

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-09T22:56:36.076431Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T22:56:36.076431Z digest=sha256:9b64577174f41075b1f204c7bc7e47db3dc486c2cab692f9df5ad73886df1265

Observation cfb26366-2b0d-4e56-aeeb-78e701513d0a · outbound

This paper cites Gans trained by a two time-scale update rule converge to a local nash equilibrium.

Diffusion Autoencoders are Scalable Image Tokenizers Gans trained by a two time-scale update rule converge to a local nash equilibrium

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-09T22:56:36.079211Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T22:56:36.079211Z digest=sha256:c512433ce9403120571bd4d9ba34f0d4a003f88af1b012e16c6e4ae0dae1104b

Observation 7f379b02-b4c0-42a2-9f67-3799725df880 · outbound

This paper cites Denoising diffusion probabilistic models.

Diffusion Autoencoders are Scalable Image Tokenizers Denoising diffusion probabilistic models

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-09T22:56:36.081999Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T22:56:36.081999Z digest=sha256:3a5f9fe1a25c009a8af533563648081fc0df137ba25d072ce66a06ab558ddb57

Observation fcabb41b-00ba-4d56-979b-0f2336277437 · outbound

This paper cites Rethinking fid: Towards a better evaluation metric for image generation.

Diffusion Autoencoders are Scalable Image Tokenizers Rethinking fid: Towards a better evaluation metric for image generation

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:56:37.069546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-09T22:56:36.084847Z digest=sha256:b8154481486555e61e7eb203ed6867428bca58fc596124318e9037f92eb92425

Observation 98eb0240-34ff-428d-b74a-af0e0e576fba · outbound

This paper cites Elucidating the design space of diffusion-based generative models.

Diffusion Autoencoders are Scalable Image Tokenizers Elucidating the design space of diffusion-based generative models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-09T22:56:36.087805Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T22:56:36.087805Z digest=sha256:9d891c398c00e72ec5b160597b53c06f60aa5b3556bc6983b3a0f616f353748d

Observation 9a828c66-830d-438f-9587-2eb859bf1bca · outbound

This paper cites Analyzing and improving the training dynamics of diffusion models.

Diffusion Autoencoders are Scalable Image Tokenizers Analyzing and improving the training dynamics of diffusion models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-09T22:56:36.090982Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T22:56:36.090982Z digest=sha256:2780999283195285ef04b5a636ae195a4656d930e56051088dafe5efa371cdad

Observation b0d71f46-18ed-42dd-90a2-8424da663027 · outbound

This paper cites and Gao, R.

Diffusion Autoencoders are Scalable Image Tokenizers and Gao, R

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:56:37.042489Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-09T22:56:36.093981Z digest=sha256:067c5783d9d18b91a530436e32e1bcb64413f1dc5af30904d99ae5625db7f542

Observation 304f53c2-6517-486d-a8bc-e604d777c048 · outbound

This paper cites Photo-realistic single image super-resolution using a generative adversarial network.

Diffusion Autoencoders are Scalable Image Tokenizers Photo-realistic single image super-resolution using a generative adversarial network

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-09T22:56:36.097117Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T22:56:36.097117Z digest=sha256:1d56c98b0c90169292255be82b6518c77f964fbd05ea0bb399796c7ca645e28f

Observation d9a31b7c-f6bf-419a-a06b-de303ddf1f8d · outbound

This paper cites Autoregressive Image Generation without Vector Quantization.

Diffusion Autoencoders are Scalable Image Tokenizers Autoregressive Image Generation without Vector Quantization

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-09T22:56:36.099974Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T22:56:36.099974Z digest=sha256:96c427d26be51723f1d70107ef9e47db25bb96654099121b95e484b91d1a6aed

Observation f1050be2-8d99-49d2-bc2a-35b310345743 · outbound

This paper cites an unresolved cited work.

Diffusion Autoencoders are Scalable Image Tokenizers Unresolved cited work

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-09T22:56:36.103548Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T22:56:36.103548Z digest=sha256:e1e0c9008cde6ab943692dd0a23d79604234160640efc0a58849a4b32a68f455

Observation ef854193-a3a1-4fed-8259-0025f5615595 · outbound

This paper cites Flow Straight and Fast: Learning to Generate and Transfer Data with Rectified Flow.

Diffusion Autoencoders are Scalable Image Tokenizers Flow Straight and Fast: Learning to Generate and Transfer Data with Rectified Flow

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-09T22:56:36.107175Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T22:56:36.107175Z digest=sha256:2f6294af02ee6fc999e1294fffa2b0f7450319eafa703154129136f68bbb03ad

Observation 6461efa8-c9bf-48df-9ef6-f42066f6c759 · outbound

This paper cites and Hutter, F.

Diffusion Autoencoders are Scalable Image Tokenizers and Hutter, F

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-09T22:56:36.111207Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T22:56:36.111207Z digest=sha256:fecb124e46f75c3d4bbfc9266893ea4a5500b4d1b54d412aa25241096f5464bf

Observation a102542e-55fa-47a3-9481-0b6ca13e6ebd · outbound

This paper cites Simplifying, Stabilizing and Scaling Continuous-Time Consistency Models.

Diffusion Autoencoders are Scalable Image Tokenizers Simplifying, Stabilizing and Scaling Continuous-Time Consistency Models

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-09T22:56:36.114879Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T22:56:36.114879Z digest=sha256:9f8e240c2a03ff1643cf8b63535268fceef09458928ddcda7b98ff6c81dcbb39

Observation 16cb503b-4aad-468c-9ca2-86fd03a50c95 · outbound

This paper cites Dpm-solver: A fast ode solver for diffusion probabilistic model sampling in around 10 steps.

Diffusion Autoencoders are Scalable Image Tokenizers Dpm-solver: A fast ode solver for diffusion probabilistic model sampling in around 10 steps

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:56:37.009144Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-09T22:56:36.119147Z digest=sha256:5398b08d814118414ed936db84147abcfa7d9ab00f26e38798e9c887b14b18d4

Observation 73695c68-9d10-48e0-8160-b9a0a4058547 · outbound

This paper cites Stacked convolutional auto-encoders for hierarchical feature extraction.

Diffusion Autoencoders are Scalable Image Tokenizers Stacked convolutional auto-encoders for hierarchical feature extraction

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:56:36.995767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-09T22:56:36.122599Z digest=sha256:68b92be9d02281c3833d7ff7c36c42b9310dab574802d80e703516c3a8605296

Observation 9a37c3f8-4cc3-49d1-80e6-c1b03008a249 · outbound

This paper cites and Maaten, L.

Diffusion Autoencoders are Scalable Image Tokenizers and Maaten, L

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:56:36.985274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-09T22:56:36.126015Z digest=sha256:2207b43ddc83e43d69d9727dc013b32e298b90f29ae02f9bf3df44664dc3ecbc

Observation d048911a-7e17-488e-8485-5437b8136f3f · outbound

This paper cites GLIDE: Towards Photorealistic Image Generation and Editing with Text-Guided Diffusion Models.

Diffusion Autoencoders are Scalable Image Tokenizers GLIDE: Towards Photorealistic Image Generation and Editing with Text-Guided Diffusion Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-09T22:56:36.129503Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T22:56:36.129503Z digest=sha256:614cd27197b6a4b2c2934f424f1c245f36535557eb5d2db4806bad606aad9a80

Observation e652869b-bef4-4cb8-b174-44e467f135e4 · outbound

This paper cites an unresolved cited work.

Diffusion Autoencoders are Scalable Image Tokenizers Unresolved cited work

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-09T22:56:36.133073Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T22:56:36.133073Z digest=sha256:9e0e78611c7def98d81d2e4063467ba82fbc9f040380cb94a3e6b1c8685b833e

Observation 48080c0d-e3c6-45e9-ba29-845bb9f41239 · outbound

This paper cites an unresolved cited work.

Diffusion Autoencoders are Scalable Image Tokenizers Unresolved cited work

Reference 40

Resolution
unresolved
raw_fallback, observed 2026-08-09T22:56:36.969262Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-09T22:56:36.136688Z digest=sha256:df3700f8362515f7782c107d9c435c4e5809e4b2392d77b2efc766ba030f8828

Observation 4e1f994f-b339-4930-924b-b4db5644dbc8 · outbound

This paper cites Diffuse VAE : Efficient, controllable and high-fidelity generation from low-dimensional latents.

Diffusion Autoencoders are Scalable Image Tokenizers Diffuse VAE : Efficient, controllable and high-fidelity generation from low-dimensional latents

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:56:36.960143Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-09T22:56:36.140145Z digest=sha256:a1e480fa079d40b4f23d1058032d688d60f6998f4cbdb7d033512e6d6b415595

Observation f67e4eaf-2966-4250-9a3e-ad7666a9bf9e · outbound

This paper cites and Xie, S.

Diffusion Autoencoders are Scalable Image Tokenizers and Xie, S

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-09T22:56:36.143524Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T22:56:36.143524Z digest=sha256:a71cf5acce0132a111cbd38dbaf0dadd07a3a97123611695e8fee1b4686f0c93

Observation 01ea68a9-d96b-4a43-aa33-29c305f2a344 · outbound

This paper cites L., Pal, C., and Aubreville, M.

Diffusion Autoencoders are Scalable Image Tokenizers L., Pal, C., and Aubreville, M

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:56:36.943136Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-09T22:56:36.147089Z digest=sha256:b26124c66c7c6d060a2bd01743334a0021a4e4613be07f1319007771ad93e95f

Observation 22a7776a-4cb5-4829-b6b8-b7d8a1dc7469 · outbound

This paper cites SDXL : Improving latent diffusion models for high-resolution image synthesis.

Diffusion Autoencoders are Scalable Image Tokenizers SDXL : Improving latent diffusion models for high-resolution image synthesis

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:56:36.932344Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-09T22:56:36.150542Z digest=sha256:621383ffbe9be17070788a9d8a0b77919d105b0ae63658917f2fea406ce84487

Observation 6de2d389-5659-4dca-a761-eb794a9eaf20 · outbound

This paper cites Movie Gen: A Cast of Media Foundation Models.

Diffusion Autoencoders are Scalable Image Tokenizers Movie Gen: A Cast of Media Foundation Models

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-09T22:56:36.153918Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T22:56:36.153918Z digest=sha256:26477661602a15213666f9568ee155751ff6d8a80e4fe279564cb54478d91f1f

Observation 07b008da-c91a-4448-8b6b-3e6218b4167c · outbound

This paper cites Diffusion autoencoders: Toward a meaningful and decodable representation.

Diffusion Autoencoders are Scalable Image Tokenizers Diffusion autoencoders: Toward a meaningful and decodable representation

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:56:36.921496Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-09T22:56:36.157536Z digest=sha256:f02e7144dce697ba77446353c21408431737cc675db41e9a9e2c1ec9afb4fbe2

Observation b05b5223-7f12-49c5-84ad-27617c94a8f1 · outbound

This paper cites Hierarchical Text-Conditional Image Generation with CLIP Latents.

Diffusion Autoencoders are Scalable Image Tokenizers Hierarchical Text-Conditional Image Generation with CLIP Latents

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-09T22:56:36.160839Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T22:56:36.160839Z digest=sha256:a72eefaf4ce8a304b3656c0c7148c7bcb19cc6ce382644c64e4e3f41dda54af1

Observation 20f4cfb5-a412-424c-9e63-62b0f013235a · outbound

This paper cites Unsupervised learning of invariant feature hierarchies with applications to object recognition.

Diffusion Autoencoders are Scalable Image Tokenizers Unsupervised learning of invariant feature hierarchies with applications to object recognition

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:56:36.910815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-09T22:56:36.164618Z digest=sha256:0cce10bc32f136664128c836ceb199bfae049a09e30540177e5fb54c4491b0c8

Observation a740689e-1f1e-41e6-8a36-95e44c12ed49 · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

Diffusion Autoencoders are Scalable Image Tokenizers High-resolution image synthesis with latent diffusion models

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-09T22:56:36.167843Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T22:56:36.167843Z digest=sha256:a543bcf5b4363b93c24126915b82b2a375b55b02b52f6670fe7ba9c6bef7b392

Observation 719b373a-823d-4514-8e5a-a803edc8a5cb · outbound

This paper cites and Hinton, G.

Diffusion Autoencoders are Scalable Image Tokenizers and Hinton, G

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:56:36.893188Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-09T22:56:36.171377Z digest=sha256:11f6886e627eb7a1952f5335e3acfe70474a13b4ea4bbf9e20b7a67ce0301981

Observation 682fd855-219c-498d-b67f-777046f225c6 · outbound

This paper cites and Ho, J.

Diffusion Autoencoders are Scalable Image Tokenizers and Ho, J

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-09T22:56:36.175161Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T22:56:36.175161Z digest=sha256:b1138f3fd0ec10669cec92a61db04baddb9abb59ebaa4cdc3302e7a6944b7305

Observation d17e275d-414a-48c3-97bd-b2ebd8a4ced0 · outbound

This paper cites Multistep Distillation of Diffusion Models via Moment Matching.

Diffusion Autoencoders are Scalable Image Tokenizers Multistep Distillation of Diffusion Models via Moment Matching

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-09T22:56:36.178917Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T22:56:36.178917Z digest=sha256:c8ce410dfeb4f6b2c61aeca983a3310c6603f7e66ea9d92eedf99eeeac3182a7

Observation 6e733bc6-6411-4ddd-823b-015b650d5299 · outbound

This paper cites Deep unsupervised learning using nonequilibrium thermodynamics.

Diffusion Autoencoders are Scalable Image Tokenizers Deep unsupervised learning using nonequilibrium thermodynamics

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-09T22:56:36.182837Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T22:56:36.182837Z digest=sha256:3a0db71ac304f83a5483e1cfd89b0bb381311b43ef86bc6953364e4bd6e789de

Observation a454d409-f1af-4def-82e2-4a1759443042 · outbound

This paper cites Denoising diffusion implicit models.

Diffusion Autoencoders are Scalable Image Tokenizers Denoising diffusion implicit models

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-09T22:56:36.186281Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T22:56:36.186281Z digest=sha256:18d341b189240fa3055dfc08baa9bf1c99c2b2002838e459192cf0be47cfc1d4

Observation f191d790-fa97-4061-91dc-5000501e89cb · outbound

This paper cites and Dhariwal, P.

Diffusion Autoencoders are Scalable Image Tokenizers and Dhariwal, P

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-09T22:56:36.189850Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T22:56:36.189850Z digest=sha256:29e2a597f4de1b9ef67e98c78177c66bdc3ca614001dcccd7e7bbb0ebee4426f

Observation 8b9a4516-8227-4d60-be4c-2258d330effb · outbound

This paper cites and Ermon, S.

Diffusion Autoencoders are Scalable Image Tokenizers and Ermon, S

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-09T22:56:36.193379Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T22:56:36.193379Z digest=sha256:218bab36aa2020a6fed50c236dbd289c5c1d03c956a15b807a4f1bd529281648

Observation 35cb72f2-e4c3-4cae-80f6-fe051d68bf32 · outbound

This paper cites Score-Based Generative Modeling through Stochastic Differential Equations.

Diffusion Autoencoders are Scalable Image Tokenizers Score-Based Generative Modeling through Stochastic Differential Equations

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-09T22:56:36.196527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T22:56:36.196527Z digest=sha256:3f17b590e4b0fe95b54adbecf6524ba7a2eeec1d8a0acf33ffa5d516f40c6bb5

Observation 772d77e5-f8f6-47dc-a53b-b30d76865cd3 · outbound

This paper cites P., Kumar, A., Ermon, S., and Poole, B.

Diffusion Autoencoders are Scalable Image Tokenizers P., Kumar, A., Ermon, S., and Poole, B

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-09T22:56:36.199954Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T22:56:36.199954Z digest=sha256:08f65af44c51ab70650bafaa3a18fc218608ded0247061bdba9d4ed60f607458

Observation f0b51721-9f23-4365-85cc-bfec5f634cf9 · outbound

This paper cites Consistency models.

Diffusion Autoencoders are Scalable Image Tokenizers Consistency models

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:56:36.845076Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-09T22:56:36.203349Z digest=sha256:0a3bb262e6cb9fcd5974fa291f9d2dbdea6a5b848fe5bef0226b1b53da76add0

Observation 080034e6-9ff0-4c58-82d0-fb680c7da5c5 · outbound

This paper cites Extracting and composing robust features with denoising autoencoders.

Diffusion Autoencoders are Scalable Image Tokenizers Extracting and composing robust features with denoising autoencoders

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:56:36.834209Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-09T22:56:36.206301Z digest=sha256:d467ce2cddffb145b560e29315854c058f81538aab3538b7810b2761f84cc096

Observation a59741cb-79c8-4d60-8576-d39b38605987 · outbound

This paper cites Esrgan: Enhanced super-resolution generative adversarial networks.

Diffusion Autoencoders are Scalable Image Tokenizers Esrgan: Enhanced super-resolution generative adversarial networks

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:56:36.822597Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-09T22:56:36.209560Z digest=sha256:a679c1548d9dfceebb7fe1e12b521bd5c346316fd801c4de9dbfbfd0fa93773c

Observation 71413202-b732-49af-a07b-48d05e6acda8 · outbound

This paper cites Real-esrgan: Training real-world blind super-resolution with pure synthetic data.

Diffusion Autoencoders are Scalable Image Tokenizers Real-esrgan: Training real-world blind super-resolution with pure synthetic data

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-09T22:56:36.811042Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-09T22:56:36.212596Z digest=sha256:fe173a8f82ed6fad539c36d4ca7eb87c6d13475bb7a0abe54dd1d339f6cec67d

Observation 8406581b-ae79-434d-b038-dfb13a2462cb · outbound

This paper cites EM Distillation for One-step Diffusion Models.

Diffusion Autoencoders are Scalable Image Tokenizers EM Distillation for One-step Diffusion Models

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-09T22:56:36.215703Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T22:56:36.215703Z digest=sha256:03af63ae096d075ef7091461e3f92e6da91696a3ad7d305f243e1d757c02258d

Observation 3df9f2e7-9741-4d87-afe8-6caf18fa55ce · outbound

This paper cites an unresolved cited work.

Diffusion Autoencoders are Scalable Image Tokenizers Unresolved cited work

Reference 64

Resolution
unresolved
raw_fallback, observed 2026-08-09T22:56:36.798382Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-09T22:56:36.219458Z digest=sha256:2485bd97e5cad79befa73250448329b544de958772a0cb5180bb59db77d1b1fb

Observation f9f62721-9cf8-435d-9acb-a2b43d6cad23 · outbound

This paper cites T., and Park, T.

Diffusion Autoencoders are Scalable Image Tokenizers T., and Park, T

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-09T22:56:36.223355Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T22:56:36.223355Z digest=sha256:9e314b5fb9d13e47fb0d885a44ef4d0738bae61d6a182881e5b63c4104ff7bd0

Observation 10880191-3d47-4747-ab61-163f8a62be6b · outbound

This paper cites Scaling Autoregressive Models for Content-Rich Text-to-Image Generation.

Diffusion Autoencoders are Scalable Image Tokenizers Scaling Autoregressive Models for Content-Rich Text-to-Image Generation

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-09T22:56:36.227366Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T22:56:36.227366Z digest=sha256:e8089afc4b19f98020df07004e9360ef4ee10de4fb3b8b5fcbcad66e01343f82

Observation b4550f88-6475-465b-8a39-891b8ad6e24f · outbound

This paper cites A., Shechtman, E., and Wang, O.

Diffusion Autoencoders are Scalable Image Tokenizers A., Shechtman, E., and Wang, O

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-09T22:56:36.231817Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T22:56:36.231817Z digest=sha256:05f3611bf6c72ce51a5292bf2ac8111cd5b24ef6ac05b80739edebf8b5e38334

Observation 6ece528b-28f1-436c-be1c-fbf0048ec9ce · outbound

This paper cites Epsilon-VAE: Denoising as Visual Decoding.

Diffusion Autoencoders are Scalable Image Tokenizers Epsilon-VAE: Denoising as Visual Decoding

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-09T22:56:36.235654Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T22:56:36.235654Z digest=sha256:2467234fe45750df5e290c1bd0ed2d905c5399f3e1cdb8ebc440dfcf039afa94

Pith citing papers

Observation 9e403f18-b92b-4636-9bc8-c38f38308274 · inbound

3D Shape Tokenization via Latent Flow Matching cites this paper.

3D Shape Tokenization via Latent Flow Matching Diffusion Autoencoders are Scalable Image Tokenizers

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-11T11:23:34.104214Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T11:23:34.104214Z digest=sha256:1e69f25b6707bfd7173e33b5d125e05d7db2a5b1b52726cc97ae6107131f6bd9

Observation f9f6a167-692f-43c5-9e9b-258d056fdc24 · inbound

D-AR: Diffusion via Autoregressive Models cites this paper.

D-AR: Diffusion via Autoregressive Models Diffusion Autoencoders are Scalable Image Tokenizers

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T12:46:47.896319Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:46:47.896319Z digest=sha256:57a69b1510ac0155814e4473870ada657e1a0ef02cc2296bae7c4ac44e3b4307

Observation 918554f6-a36e-4b25-801b-a7567b8d98a5 · inbound

SSDD: Single-Step Diffusion Decoder for Efficient Image Tokenization cites this paper.

SSDD: Single-Step Diffusion Decoder for Efficient Image Tokenization Diffusion Autoencoders are Scalable Image Tokenizers

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-04T11:25:33.951411Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T11:25:33.951411Z digest=sha256:b02ca6cc662a18155717668b045bf103e527749a12a58dd4d98a80ccc51c5cb8

Observation c5f8e4c3-3d5f-4a0d-a217-7affa712167b · inbound

HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers cites this paper.

HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers Diffusion Autoencoders are Scalable Image Tokenizers

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-07-03T14:28:29.657972Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-06-27T07:01:07.362430Z digest=sha256:13fbac402f52ae32c49588367c57859186ba692db9a8145bd134158df486095a