Pith. sign in

Paper Citation Record · LEDGER

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer

As of 20 August 2026, this Paper Citation Record lists 100 of 126 outbound references and 5 inbound Pith citation observations for arXiv:2412.10958.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.10958 v3

Coverage vector

measured 100 of 126 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T15:33:13.575809Z

measured 105 of 105 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-09T04:47:30.290200Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-29T18:13:49.029434Z

Reference resolution

100 of 126 outbound references displayed

  • verified exact0
  • verified fuzzy17
  • unresolved83
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 9900d614-266b-49b0-9c1a-3cf57d406d9f · outbound

This paper cites Stochastic Interpolants: A Unifying Framework for Flows and Diffusions.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Stochastic Interpolants: A Unifying Framework for Flows and Diffusions

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.009735Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.009735Z digest=sha256:b17aa0db094342a0b2816f3f7d669bc63ffbdb5908c893240792cc13f82abc73

Observation 24ae505f-851c-45c7-9555-3153ca79891f · outbound

This paper cites Reverse-time diffusion equation mod- els.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Reverse-time diffusion equation mod- els

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.015948Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.015948Z digest=sha256:ebe510810d1669f297b21ec5ef30fe8313d7efbb5c03ebe5a32306f8f83944cc

Observation a47f9ab6-2a5e-4bdb-b659-f7f0be16135e · outbound

This paper cites wav2vec 2.0: A framework for self- supervised learning of speech representations.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer wav2vec 2.0: A framework for self- supervised learning of speech representations

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.020950Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.020950Z digest=sha256:26501025b6b24379153865c0881c068b5a0632d7658701e69dd9bec8d7906119

Observation bb4694c3-b57d-45a2-8cbc-01694d3b2f55 · outbound

This paper cites All are worth words: A vit backbone for diffusion models.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer All are worth words: A vit backbone for diffusion models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.025706Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.025706Z digest=sha256:f381c891e43ee0eb067487ec6cd5cde336bb9898ecad0799c042850c8f30c44f

Observation 0cff7665-d376-49fb-8422-aba97a363f2b · outbound

This paper cites BEiT: BERT Pre-Training of Image Transformers.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer BEiT: BERT Pre-Training of Image Transformers

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.030662Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.030662Z digest=sha256:62208d4c33d02328c20818735b70215518ca1c8ca26ac4c3c6271040a0082186

Observation c0446b61-46d0-40fa-9ebb-20411102d66d · outbound

This paper cites Lumiere: A Space-Time Diffusion Model for Video Generation.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Lumiere: A Space-Time Diffusion Model for Video Generation

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.035747Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.035747Z digest=sha256:a2e6e302a92d5ec059cad08c29692ba84a6fd7483c6c87f505c8c5ffbb440321

Observation d239d2be-337d-4ce8-a5a8-d0b2430809bc · outbound

This paper cites Estimating or Propagating Gradients Through Stochastic Neurons for Conditional Computation.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Estimating or Propagating Gradients Through Stochastic Neurons for Conditional Computation

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.040831Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.040831Z digest=sha256:f9afc58b20f9736ac65c74bcf8c1e6b20fd1cad99807c23de86aecd949a7d30a

Observation 8c011249-fe69-459b-b4eb-e5ff735012db · outbound

This paper cites Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.045915Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.045915Z digest=sha256:f2e54849318322904b73403cbb2a42c6cf8138b95e4242353157f81080a5c77e

Observation b1d97961-7873-40de-85d4-674af203b838 · outbound

This paper cites Emerg- ing properties in self-supervised vision transformers.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Emerg- ing properties in self-supervised vision transformers

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.051769Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.051769Z digest=sha256:ff28fe2e12b6ea2662b1b475f3793187e41475c7c0def3faf114cfd28ff10f0a

Observation ac6f7a15-4fcb-4bae-86a6-9641906ce980 · outbound

This paper cites an unresolved cited work.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Unresolved cited work

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.056446Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.056446Z digest=sha256:998b538807c22a3b66c676269a4aa57d23c83594ad5499355d20225c7d937a32

Observation a6df8b20-ece5-45ee-a4cc-a076de1ae434 · outbound

This paper cites Deep Compression Autoencoder for Efficient High-Resolution Diffusion Models.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Deep Compression Autoencoder for Efficient High-Resolution Diffusion Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.061018Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.061018Z digest=sha256:9217d5acab96edc2d0c47e5931cb347c793e0379d1734db2981f8d36d09f1945

Observation 2096eee0-1f88-45ef-b2c8-ede51a9c376a · outbound

This paper cites An Image is Worth 1/2 Tokens After Layer 2: Plug-and-Play Inference Acceleration for Large Vision-Language Models.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer An Image is Worth 1/2 Tokens After Layer 2: Plug-and-Play Inference Acceleration for Large Vision-Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.065992Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.065992Z digest=sha256:5aae4ce62d19601cd35dee2a0c3c52d04222d58a6ac5eaed87235cf543910024

Observation f70478c1-0885-4542-a63c-8248a0b35cad · outbound

This paper cites Variational Lossy Autoencoder.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Variational Lossy Autoencoder

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.070829Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.070829Z digest=sha256:d441b097dbb2b49e202669fef1f1d85db708199236e4d741e52d4bc57204cbf7

Observation c8ca871d-15cf-45a1-bae3-ea3b314ddf6c · outbound

This paper cites Deconstructing Denoising Diffusion Models for Self-Supervised Learning.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Deconstructing Denoising Diffusion Models for Self-Supervised Learning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.075744Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.075744Z digest=sha256:568d22e90df27afaaecde46408d35f103b37e4cd70bae48c16c6eefa9bdae465

Observation 743c731c-893c-4d8e-9825-f682a0a3ddb8 · outbound

This paper cites Imagenet: A large-scale hierarchical image database.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Imagenet: A large-scale hierarchical image database

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.080828Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.080828Z digest=sha256:8d026381f83db89eca6a000610824e13ecac2f239c813ad81346ba1799a6dd90

Observation 1143e83c-f0eb-464d-b2c7-ae17989a1ed8 · outbound

This paper cites Diffusion models beat gans on image synthesis, 2021.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Diffusion models beat gans on image synthesis, 2021

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.085387Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.085387Z digest=sha256:8ad085b21f24c58aac177961682044fdafb4f9707c91e56132668f0eff972972

Observation 0bf17e51-38f2-4f8d-bb23-d288308e3388 · outbound

This paper cites Deep Unsupervised Clustering with Gaussian Mixture Variational Autoencoders.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Deep Unsupervised Clustering with Gaussian Mixture Variational Autoencoders

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.090116Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.090116Z digest=sha256:b81b1c5bb2a3629fa34098338276f3935e2b3c8a956bc206850cb70022cc8573

Observation b88a35db-2660-4cc3-b414-b32fa1720ea5 · outbound

This paper cites Cogview: Mastering text-to-image generation via transformers.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Cogview: Mastering text-to-image generation via transformers

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.094730Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.094730Z digest=sha256:80c165abf735dcfb2420e5f759b7a9a20de5a2b075fccd43ab6f53cef7a053d4

Observation 92711e0a-465b-4951-a83f-8c60e2581a62 · outbound

This paper cites Peco: Perceptual codebook for bert pre-training of vision transformers.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Peco: Perceptual codebook for bert pre-training of vision transformers

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.098832Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.098832Z digest=sha256:f421e9343b92b0a97571abcc5bb967eb856a8d61a303571b9f32acdd0c38ccd5

Observation 8b22cb5f-27c3-492a-b58a-29d27824802f · outbound

This paper cites Generating images with perceptual similarity metrics based on deep networks.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Generating images with perceptual similarity metrics based on deep networks

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.103290Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.103290Z digest=sha256:7b802e4188e723a426f89805bbff551c93d9f720aa32a1fb1376fd8d0e1b82a8

Observation b382bc9e-5250-4a3f-b0bf-9cbff81601ac · outbound

This paper cites An image is worth 16x16 words: Transformers for image recognition at scale, 2021.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer An image is worth 16x16 words: Transformers for image recognition at scale, 2021

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.107469Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.107469Z digest=sha256:fbe86bb22a531b1bef4439fc056df46af7babe65fe1cbb9016497ddb7dd398e8

Observation de3dd328-38f8-4409-b8c5-f3f22d24cf55 · outbound

This paper cites A fuzzy relative of the isodata process and its use in detecting compact well-separated clusters.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer A fuzzy relative of the isodata process and its use in detecting compact well-separated clusters

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.111511Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.111511Z digest=sha256:1e909c445138da09270746ec80dbe36800cd57678451b0c83566f0f546b2cfc8

Observation e4e92bd4-5d54-4b8a-b605-8333561202de · outbound

This paper cites Taming transformers for high-resolution image synthesis.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Taming transformers for high-resolution image synthesis

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.115460Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.115460Z digest=sha256:2aafb74d9f87a94184746343f8cc7b80abc878a10f02dcbdcc501977df75190f

Observation 54ed0f68-2ac8-4c76-8713-e96c0ee2560c · outbound

This paper cites Fast Timing-Conditioned Latent Audio Diffusion.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Fast Timing-Conditioned Latent Audio Diffusion

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.119379Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.119379Z digest=sha256:87e1f36d1a5ca3880f90529594c262249af2d9e0a3ab35af6aab8f59fc0005cf

Observation ec433e16-34cc-484b-90c8-52b54acf0be1 · outbound

This paper cites Fluid: Scaling Autoregressive Text-to-image Generative Models with Continuous Tokens.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Fluid: Scaling Autoregressive Text-to-image Generative Models with Continuous Tokens

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.123915Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.123915Z digest=sha256:6add1c1f63025535bf76fa422208153aa250f9558648d9d9a0e6cb3a4f1d4e8f

Observation 1f1834e9-8542-4e38-a29d-e6bebdb4e362 · outbound

This paper cites Eva-02: A visual representation for neon genesis.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Eva-02: A visual representation for neon genesis

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.129163Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.129163Z digest=sha256:7f07b0b183f0c596d1711b5ae11465500550576fe20366e86d0dddbc048741c0

Observation 9385977d-54a2-490e-88c2-97434ceac5b2 · outbound

This paper cites Restructuring Vector Quantization with the Rotation Trick.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Restructuring Vector Quantization with the Rotation Trick

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.133672Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.133672Z digest=sha256:4e34363aa26ab41402c10ecd68efb954801bd2b9b892263a4279051571d9d8c0

Observation 46df1f4d-cd50-47f2-b942-60f948d72f81 · outbound

This paper cites Make-a-scene: Scene- based text-to-image generation with human priors.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Make-a-scene: Scene- based text-to-image generation with human priors

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.138618Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.138618Z digest=sha256:f3c381cc1fd3bbac0ae763f40ce9b8c4c06d56c71fa02b3952063efe6cda279a

Observation 3417631d-9c10-4451-96b7-39d53518d731 · outbound

This paper cites MDTv2: Masked Diffusion Transformer is a Strong Image Synthesizer.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer MDTv2: Masked Diffusion Transformer is a Strong Image Synthesizer

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.143121Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.143121Z digest=sha256:a9e090d22faf5c698c15ef53d460f2e0ff1af721f701552690b8ae120415ff3e

Observation 849e4bcd-c993-492a-a5f9-d12e86a61a47 · outbound

This paper cites Planting a seed of vision in large language model,.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Planting a seed of vision in large language model,

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.148816Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.148816Z digest=sha256:41800c10ac022ca0d9438a01654c104bc618b60935d15e161104c5efcddb9bc3

Observation 93142e31-20c6-43bc-96bf-f63d2547d81f · outbound

This paper cites Making LLaMA SEE and Draw with SEED Tokenizer.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Making LLaMA SEE and Draw with SEED Tokenizer

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.153399Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.153399Z digest=sha256:21df9e9a8d1b9eb66424c12e0d956b6df24a4e50fc9f227da2a77a476faf9e13

Observation ec8f4e62-bd3b-44be-8aa1-8030aa58c767 · outbound

This paper cites DiffuSeq: Sequence to Sequence Text Generation with Diffusion Models.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer DiffuSeq: Sequence to Sequence Text Generation with Diffusion Models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.158166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.158166Z digest=sha256:176e997a5e88a0bc3ad0dfb70d48afa0843a5c86150724f9fb63a0f7000ade7c

Observation cc01d630-7fee-40f5-afc0-eca3d3d9dca3 · outbound

This paper cites Generative adversarial networks.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Generative adversarial networks

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.163379Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.163379Z digest=sha256:19451d18044b2820e97466316e5a8ce05299be8a4e6a53d78b8d36422ddf7361

Observation fe3723f1-e282-4bf3-a68c-f17922b70194 · outbound

This paper cites Rethinking the objectives of vector-quantized tokenizers for image synthesis, 2023.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Rethinking the objectives of vector-quantized tokenizers for image synthesis, 2023

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.167952Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.167952Z digest=sha256:91a58054922375005282571823c13ea84d6cbf759769a5e55a8fad66cf2d4090

Observation a3a41adb-549a-4f40-a258-121a18145560 · outbound

This paper cites Masked autoencoders are scal- able vision learners.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Masked autoencoders are scal- able vision learners

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.172497Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.172497Z digest=sha256:04db9bf034320492b01b499f5c24d54018d65f9676b432af9ff433d6f5f8ac7a

Observation bb2da402-168d-41b0-a4db-1cea0a84597f · outbound

This paper cites Rotary position embedding for vision transformer.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Rotary position embedding for vision transformer

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.176985Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.176985Z digest=sha256:f9e6444a51433ab89f04b3b7278fd7eef5e9fbe106962216dda4ba9b44d1d503

Observation 1c1aa017-323a-43c0-98be-f8fd56d0344c · outbound

This paper cites Gans trained by a two time-scale update rule converge to a local nash equilibrium.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Gans trained by a two time-scale update rule converge to a local nash equilibrium

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.181358Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.181358Z digest=sha256:8a5ccbe175cb713110bc6fe133cae554649dca043e290f9ea131772685f2b4c1

Observation 82f625b5-b7b2-4b8a-bb97-eb70ddaab0e9 · outbound

This paper cites beta-vae: Learning basic visual concepts with a constrained variational framework.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer beta-vae: Learning basic visual concepts with a constrained variational framework

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.185786Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.185786Z digest=sha256:5726eb6049692c40afbb3c11cf7e797b3e4e53eedd2d921793239ba4a2bce977

Observation 8a34eb81-8c31-42f5-9946-c8f87659117b · outbound

This paper cites Reducing the dimensionality of data with neural networks.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Reducing the dimensionality of data with neural networks

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.190239Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.190239Z digest=sha256:676c5faa80c8ee38719095062848616830a8e86b153fdfea08fec0905134792c

Observation c1f2f2ad-369f-4d26-a36d-f33821bf8966 · outbound

This paper cites Classifier-Free Diffusion Guidance.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Classifier-Free Diffusion Guidance

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.194625Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.194625Z digest=sha256:99ce8f0f445e114d64089ce90df8af78d4d30a6a54b2808bb3239938d020236f

Observation dc383858-ad5f-462c-8882-33612857c5a6 · outbound

This paper cites Denoising diffu- sion probabilistic models.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Denoising diffu- sion probabilistic models

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.200079Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.200079Z digest=sha256:1b72c586949cdcd3e294ab9cd6730884edcb55645566371535f97c50aa6bff2f

Observation 4fd048ef-a21c-4a74-82cf-2fc837062b69 · outbound

This paper cites Video dif- fusion models.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Video dif- fusion models

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.204507Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.204507Z digest=sha256:c97608694d4ee3aa921ab2a093dbdb44b9aa54ce3247b4e4cde87c210c4bf6b9

Observation ee5901eb-78f3-4d9c-916c-2b51e1aeb6b3 · outbound

This paper cites Straightening out the straight-through estimator: Over- coming optimization challenges in vector quantized net- works.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Straightening out the straight-through estimator: Over- coming optimization challenges in vector quantized net- works

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.209089Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.209089Z digest=sha256:76fc39f35a9f4ed3ab8f38cc0c87a3cc873ebc615944a9fbdf86bc1d0093a8ae

Observation bfa48579-4528-464c-a78b-19aeee4d71cc · outbound

This paper cites an unresolved cited work.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Unresolved cited work

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.213515Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.213515Z digest=sha256:39d44582aeef9d51b52039c369b9951522c834453e6d34a1fa0718fbc84c8d5c

Observation b09be01b-fabb-4fe0-8c20-c94fb5ddfc55 · outbound

This paper cites Product quantization for nearest neighbor search.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Product quantization for nearest neighbor search

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.218002Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.218002Z digest=sha256:996633095c33697678c34bcd5dfe067cabb043875124dce43c5dbb968515e1d2

Observation 845aad99-60f4-4256-bd76-4797aa10a3ec · outbound

This paper cites Percep- tual losses for real-time style transfer and super-resolution.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Percep- tual losses for real-time style transfer and super-resolution

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.222387Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.222387Z digest=sha256:c1cfaf7a76aac1b5b54ae1359dff4e2096e0353a1a573fcbef3665b0b30f2827

Observation d97107ca-26cf-45b4-8372-3ed4378eb4fd · outbound

This paper cites Composing graphical models with neural networks for structured representations and fast inference.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Composing graphical models with neural networks for structured representations and fast inference

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.226655Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.226655Z digest=sha256:e4bf69327c8d88308f10308485fa4283bd070a9e35c6849807f781a05f942fdd

Observation d2314b4e-09e8-4e8b-a0ce-a9819092303a · outbound

This paper cites A style-based generator architecture for generative adversarial networks.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer A style-based generator architecture for generative adversarial networks

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.231037Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.231037Z digest=sha256:b69c9765c55773b52f9291833757c6aeea48aff15d12f05725a227ac4715f50b

Observation 3bb7f5d9-0e47-4e80-a3e3-c0a1f4c37cc4 · outbound

This paper cites Analyzing and improv- ing the image quality of stylegan.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Analyzing and improv- ing the image quality of stylegan

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.235169Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.235169Z digest=sha256:2be32242200ff6e7c8370b3f916908031d155ca928fd64c1bd21de5b41e35282

Observation 507122e3-93e4-4b41-9699-0aa19b6c57d2 · outbound

This paper cites Understanding diffusion objectives as the elbo with simple data augmentation.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Understanding diffusion objectives as the elbo with simple data augmentation

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.239231Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.239231Z digest=sha256:656d7a1af2a0ea6b67a830360e2a8877439c288b1a219a94aa2376b514a82510

Observation c85a29ca-c109-4c16-b6b1-ec339b0a5055 · outbound

This paper cites Auto-Encoding Variational Bayes.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Auto-Encoding Variational Bayes

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.243492Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.243492Z digest=sha256:b30eab7707ccfae6506f702b9b0fd4bb8c21f7839fa3e8fa7e06ca9d445b7db6

Observation 141ab1c4-e1e4-4c9e-afee-a89683e0367c · outbound

This paper cites Improved precision and recall metric for assessing generative models.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Improved precision and recall metric for assessing generative models

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.247514Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.247514Z digest=sha256:e16ba790de6c68c12d84364d7d0de74935249a3b054ed78673841091c962b8f9

Observation 35e731ce-1456-454f-8a95-2f8a98d62b88 · outbound

This paper cites Applying Guidance in a Limited Interval Improves Sample and Distribution Quality in Diffusion Models.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Applying Guidance in a Limited Interval Improves Sample and Distribution Quality in Diffusion Models

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.251602Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.251602Z digest=sha256:cf5a394fc5c7ab4f9781dba71590d4e22438df76c3491d5cd7005c13e015e9b6

Observation 89b309de-3202-4693-897e-11534830e673 · outbound

This paper cites Autoencoding beyond pixels using a learned similarity metric.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Autoencoding beyond pixels using a learned similarity metric

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.256119Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.256119Z digest=sha256:1772959e8584cee057e827db7ae4012e6ad6c8d2740e22e3c8096f057c9c1c8b

Observation 67582688-9b0c-44a4-96ea-43f913b4205b · outbound

This paper cites Autoregressive image generation using residual quantization.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Autoregressive image generation using residual quantization

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.260104Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.260104Z digest=sha256:4f935e7535658c0dee9c3361f8f66c351e6690acc29ac467c4a30658233b6498

Observation c5aa7035-2763-4a43-a457-12e32037db62 · outbound

This paper cites Autoregressive image generation using residual quantization, 2022.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Autoregressive image generation using residual quantization, 2022

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.264143Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.264143Z digest=sha256:a1f2be74b359c3fc04f21256b96f18033ed541371c4ce1de4eceaccbf547d7ab

Observation 57a51a0b-a6b9-4a9d-af6a-b9a6ccb24796 · outbound

This paper cites Scalable autoregressive image generation with mamba.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Scalable autoregressive image generation with mamba

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.268840Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.268840Z digest=sha256:62ba71cdbeb953e80de0b443f9d254b673fb91a589a3d3508b6583f60c5a1968

Observation 1169c724-5ac9-4147-9ce9-33e65e74ed67 · outbound

This paper cites Blip- 2: Bootstrapping language-image pre-training with frozen image encoders and large language models.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Blip- 2: Bootstrapping language-image pre-training with frozen image encoders and large language models

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.273161Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.273161Z digest=sha256:29903e3bfcf7c0aed08ab8474b97e178c1b458570d3732bc12de5f688e580650

Observation 4c1cd7f6-e3f3-4deb-b36f-e5951e305a30 · outbound

This paper cites Mage: Masked generative encoder to unify representation learning and im- age synthesis, 2023.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Mage: Masked generative encoder to unify representation learning and im- age synthesis, 2023

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.277694Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.277694Z digest=sha256:f0671b189dc35055d75da3b248690c0910fe3c3fbe2e71c71032c1ebe050eb5b

Observation bf4f4048-1066-4e46-8d0e-4c2e74dc33c2 · outbound

This paper cites Autoregressive image generation without vec- tor quantization, 2024.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Autoregressive image generation without vec- tor quantization, 2024

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.282109Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.282109Z digest=sha256:b5a7302b7f731fe41f60787345e4537cb8eab2bc8ed23a3de24b4a2529ce9d83

Observation 59e7046c-d4fb-4a3a-b6fb-0d8f77c9a4cd · outbound

This paper cites TokenPacker: Efficient Visual Projector for Multimodal LLM.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer TokenPacker: Efficient Visual Projector for Multimodal LLM

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.286584Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.286584Z digest=sha256:0b622368a23185929cd03ee2d884bbabf3a214f3f150c38c44603984e9e6ac14

Observation 499d77d8-1899-444e-8516-fa819ab0cbaa · outbound

This paper cites ImageFolder: Autoregressive Image Generation with Folded Tokens.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer ImageFolder: Autoregressive Image Generation with Folded Tokens

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.291256Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.291256Z digest=sha256:f3123dedc324fc86ae91769b2e5478939e43e999d2d872cbd28282f2dcd7d0c0

Observation 00011b92-6bd8-485b-96d9-c215a7ae257d · outbound

This paper cites Mini-Gemini: Mining the Potential of Multi-modality Vision Language Models.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Mini-Gemini: Mining the Potential of Multi-modality Vision Language Models

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.295905Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.295905Z digest=sha256:a5a1d3595a0bf6dd07a347f7d9021c2a6c5a4be0785198dcd77fbf1227e4d733

Observation 44889bf7-01ad-45ee-b1e8-db53fd2be9c9 · outbound

This paper cites SPHINX: The Joint Mixing of Weights, Tasks, and Visual Embeddings for Multi-modal Large Language Models.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer SPHINX: The Joint Mixing of Weights, Tasks, and Visual Embeddings for Multi-modal Large Language Models

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.300508Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.300508Z digest=sha256:a9101dc21583fc4f7fe61719d6ea0dab4238a69005267ab84dbcbcd85c1ba0fd

Observation c931c434-6338-4352-860f-0570feba0c8e · outbound

This paper cites Flow Matching for Generative Modeling.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Flow Matching for Generative Modeling

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.306174Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.306174Z digest=sha256:2e22925ca97586f83eb3ba9aae3d9a88f03d217cfc601d971968f735bd1303b7

Observation 9150376d-81ad-4bcb-b174-0056d66e846b · outbound

This paper cites Customize Your Visual Autoregressive Recipe with Set Autoregressive Modeling.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Customize Your Visual Autoregressive Recipe with Set Autoregressive Modeling

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.310907Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.310907Z digest=sha256:69dde53c4fe046f4ca00d8acbdf605e199c043084a427a1b94a81db0adb5de7a

Observation 726a3626-e1c0-4473-ab06-e2d8146533af · outbound

This paper cites Least squares quantization in pcm.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Least squares quantization in pcm

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.315525Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.315525Z digest=sha256:3c3d3a09048a1b62adbba6698c18ccc1f135d3266ac9f3e8f3f3a0ea7ad29532

Observation 57702906-ab4c-431e-8e14-17d2c4ae0551 · outbound

This paper cites Decoupled Weight Decay Regularization.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Decoupled Weight Decay Regularization

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.319871Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.319871Z digest=sha256:c000a31b6527b1d88b0211fca056055c7838401ccc5fa8a8b331a9fe95688fa0

Observation be5aefdb-81c7-48f6-b79b-131cbb6392c3 · outbound

This paper cites Simplifying, Stabilizing and Scaling Continuous-Time Consistency Models.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Simplifying, Stabilizing and Scaling Continuous-Time Consistency Models

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.324722Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.324722Z digest=sha256:48adfcdca03e818580a55d7f0f3926be4e4d0fcc6a459cc35de6ac81c03308cd

Observation 4f856e7e-0bec-498b-b64d-3a77b3e1661e · outbound

This paper cites Feast Your Eyes: Mixture-of-Resolution Adaptation for Multimodal Large Language Models.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Feast Your Eyes: Mixture-of-Resolution Adaptation for Multimodal Large Language Models

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.329381Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.329381Z digest=sha256:1fa5a9c47c92324f8ff56d1abd28c85377ded3a7855d264b280b96681269caf7

Observation 1865715b-0ec9-436d-8a7a-29ba1fe52158 · outbound

This paper cites Open-MAGVIT2: An Open-Source Project Toward Democratizing Auto-regressive Visual Generation.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Open-MAGVIT2: An Open-Source Project Toward Democratizing Auto-regressive Visual Generation

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.333762Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.333762Z digest=sha256:16a7475c2b0d32aae5ad3bdd373dcf9a0a025e1a70739d62ec6972cd5842462b

Observation ed533d42-d389-4536-944a-4e5090081365 · outbound

This paper cites SiT: Exploring Flow and Diffusion-based Generative Models with Scalable Interpolant Transformers.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer SiT: Exploring Flow and Diffusion-based Generative Models with Scalable Interpolant Transformers

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.338668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.338668Z digest=sha256:2c491c3539035b95d08b5d7a1468c1c390d168eb6b3fabf49c52d3f28d8d7774

Observation f497f54d-87d1-45a0-995b-47a998da0486 · outbound

This paper cites Finite scalar quantization: Vq-vae made simple, 2023.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Finite scalar quantization: Vq-vae made simple, 2023

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.453492Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.453492Z digest=sha256:969f7a796cb10320746413002054692ad702549084cea4c151af7fd4274521a9

Observation 9d681227-7bc2-4410-9fb1-bc818af2ff6e · outbound

This paper cites Midjourney.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Midjourney

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.458279Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.458279Z digest=sha256:c3a711101a87d7d65f492a3275744ea6fd9341c67a6c64271211d220c237b740

Observation ea32e042-1385-44a6-a156-62569add43f0 · outbound

This paper cites Improved denoising diffusion probabilistic models, 2021.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Improved denoising diffusion probabilistic models, 2021

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.462613Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.462613Z digest=sha256:5fd25fbe89ff41241aad6dd4b89f04448e6f750bfbad5b5b219565be47c51c5b

Observation 0ecda30f-f230-4b73-beb4-abf7a71399ca · outbound

This paper cites GLIDE: Towards Photorealistic Image Generation and Editing with Text-Guided Diffusion Models.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer GLIDE: Towards Photorealistic Image Generation and Editing with Text-Guided Diffusion Models

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.467389Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.467389Z digest=sha256:842de91bd3a37bc0e370e9e99ae6a74a1ee0aa2083a54ec6221e233aea592b45

Observation ea5e38d4-aba4-47b9-a558-11df9c14e31c · outbound

This paper cites Certain topics in telegraph transmission theory.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Certain topics in telegraph transmission theory

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.472143Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.472143Z digest=sha256:644550cc909f10ed85b84352d8821a710e2b2983d46ed69eebcc5bb35b68ac2a

Observation cb9fa985-7de4-4a16-a860-bfbd59e658a5 · outbound

This paper cites Video generation models as world simulators.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Video generation models as world simulators

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.476838Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.476838Z digest=sha256:64a7027a49f72f50d09874cbb789039a88b9a3a553f1da21609bacb161045125

Observation 6eb381dc-708f-46c7-a2af-28e66d491ef1 · outbound

This paper cites an unresolved cited work.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Unresolved cited work

Reference 79

Resolution
unresolved
raw_fallback, observed 2026-08-11T15:33:15.134776Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T15:33:13.481171Z digest=sha256:688ed12f8699e7ae3f65131a66931c3adaefcca24e3d2126733ef3e1c1f2dc19

Observation 5c2bc65b-628b-4be3-9b02-f78e63f0a83e · outbound

This paper cites Scalable diffusion models with transformers, 2023.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Scalable diffusion models with transformers, 2023

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:33:15.119464Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T15:33:13.485672Z digest=sha256:b4bfa4bf074c19d8df838fa1f0efd7723a4c089b34c36560c1d48eb3d362fd19

Observation 406bf857-fd21-48c1-8a34-16cfdd0b2451 · outbound

This paper cites Grad-tts: A diffusion probabilistic model for text-to-speech.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Grad-tts: A diffusion probabilistic model for text-to-speech

Reference 81

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:33:15.102846Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T15:33:13.490395Z digest=sha256:59e44fb23263bf9d555efe3cf726d35a2580370bf246a4ec95033dcc5577c26a

Observation 9ca122af-6f1e-4e04-ba1a-e6ed126a86fc · outbound

This paper cites Learn- ing transferable visual models from natural language super- vision.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Learn- ing transferable visual models from natural language super- vision

Reference 82

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:33:15.086872Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T15:33:13.494787Z digest=sha256:eba3419c8d8a1744faf9dac547468a5e779e03d604b9396432a645631f4f8692

Observation 14dcbe19-d435-4635-8304-b855b87e3322 · outbound

This paper cites Zero-shot text-to-image generation.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Zero-shot text-to-image generation

Reference 83

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:33:15.070087Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T15:33:13.499255Z digest=sha256:3ed950ebfa8775ee14e62e6cc9aa67cf041c0211bc30f672781ae08cfacda2e6

Observation 209686a3-ba66-49aa-8795-6f4dfe2535e3 · outbound

This paper cites Gener- ating diverse high-fidelity images with vq-vae-2.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Gener- ating diverse high-fidelity images with vq-vae-2

Reference 84

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:33:15.055563Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T15:33:13.503593Z digest=sha256:cd19174959fef80bd5f9bd5baab0e02a0f826b7973b1bbcdf8c41245331de033

Observation 1b382d1e-4e6f-49c2-8076-cd5685bc64eb · outbound

This paper cites Gen- erating diverse high-fidelity images with vq-vae-2, 2019.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Gen- erating diverse high-fidelity images with vq-vae-2, 2019

Reference 85

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:33:15.041377Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T15:33:13.507957Z digest=sha256:d73de72c2b7b7e3a1d2ad915985a11542909536a7b760e4cfa986c1c34297d77

Observation ad4194d3-5e82-4d50-b94e-c23654fba4d7 · outbound

This paper cites Gaussian mixture models.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Gaussian mixture models

Reference 86

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:33:15.026710Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T15:33:13.512417Z digest=sha256:f1b37912b43ec544b2866ba7e6b39a114fb84051f577b308f3c941f8e195afd1

Observation cdf455bc-8065-4e0f-84bf-a3e438ee94a1 · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer High-resolution image synthesis with latent diffusion models

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.516862Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.516862Z digest=sha256:61141069df412c6dad36bd12eb30d51345113bf3336d29c505a9fa9f50edd3cb

Observation 1bb16e1f-b221-43e2-8d38-2e5acfa65ad5 · outbound

This paper cites High-resolution image synthesis with latent diffusion models, 2022.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer High-resolution image synthesis with latent diffusion models, 2022

Reference 88

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:33:15.001592Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T15:33:13.520973Z digest=sha256:4e3ec211c5554d507b32c647a4e96c9d55b91f66d26cde08891f645ee76ac7f0

Observation 8fd12bbb-487a-49cb-a534-dd23aa97b56e · outbound

This paper cites Photorealistic text-to-image diffusion models with deep language understanding.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Photorealistic text-to-image diffusion models with deep language understanding

Reference 89

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:33:14.986771Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T15:33:13.525631Z digest=sha256:2adfb027a136f2c29ecffedd22cd74286ccfa6f660fdada55595f683322a0511

Observation 71269577-80cd-4ff2-8e74-cde7b10ceab7 · outbound

This paper cites Improved techniques for training gans.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Improved techniques for training gans

Reference 90

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:33:14.970885Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T15:33:13.530219Z digest=sha256:08e9bfa33e87997736969132f61ed0ab169c8501113c00b223a2aff2edfc5d8f

Observation c0b1c941-672d-44d1-8dff-4a78af647f98 · outbound

This paper cites Communication in the presence of noise.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Communication in the presence of noise

Reference 91

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:33:14.954290Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T15:33:13.534574Z digest=sha256:faf2c2cfb5929c72b2b90c78d357abe802efc37e63c6d52b1d9e752d7ff96eb9

Observation 2dd07c0d-d410-4cc2-b45c-97cb116b6729 · outbound

This paper cites Deep unsupervised learning using nonequilibrium thermodynamics.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Deep unsupervised learning using nonequilibrium thermodynamics

Reference 92

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:33:14.939417Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T15:33:13.539183Z digest=sha256:6f510e26c5ec9b908076ef8c3c2ba28b05857e08d3b132b878caa3bcfd574832

Observation cc611138-2587-411f-90ae-1614607b7534 · outbound

This paper cites Weiss, Niru Mah- eswaranathan, and Surya Ganguli.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Weiss, Niru Mah- eswaranathan, and Surya Ganguli

Reference 93

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:33:14.924652Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T15:33:13.543486Z digest=sha256:54c8be3bf888d4764b9b3dc074c088ea783aad8d867c15c5c524561ee95a4519

Observation acb165d3-b674-471a-90d0-a9a7c78b8304 · outbound

This paper cites Less is More: A Simple yet Effective Token Reduction Method for Efficient Multi-modal LLMs.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Less is More: A Simple yet Effective Token Reduction Method for Efficient Multi-modal LLMs

Reference 94

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.548179Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.548179Z digest=sha256:b18166c059c9ce9beaa889869099ab91a48562eb68fc517b51133f0da5914073

Observation f35823c9-883e-47cd-88b6-c1529839663d · outbound

This paper cites Denois- ing diffusion implicit models, 2022.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Denois- ing diffusion implicit models, 2022

Reference 95

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:33:14.909303Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T15:33:13.553024Z digest=sha256:9c7876cb6265f8821b3c23fe2160c42a9a82ffacae45563493d40ebe8034e3ff

Observation f8e0c1bc-1ad4-4908-9c88-155f982f70f9 · outbound

This paper cites Sd vae ft ema, 2023.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Sd vae ft ema, 2023

Reference 96

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:33:14.893960Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T15:33:13.557698Z digest=sha256:42e5445eca9ff2685697d959a75814c891e6c6017633028520b5a76012bc2d5f

Observation 758e142c-3232-4297-ba59-fd7f983a3cc6 · outbound

This paper cites Roformer: Enhanced transformer with rotary position embedding.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Roformer: Enhanced transformer with rotary position embedding

Reference 97

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.562297Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.562297Z digest=sha256:8148a1d9bcc032bff1bc70819ac3934cd7b0405beaea8efbcc9ebb187b541a5a

Observation 437128a1-8d6e-43fd-9d4c-098366ed7514 · outbound

This paper cites Autoregressive Model Beats Diffusion: Llama for Scalable Image Generation.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Autoregressive Model Beats Diffusion: Llama for Scalable Image Generation

Reference 98

Resolution
unresolved
no resolver link, observed 2026-08-11T15:33:13.566667Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:33:13.566667Z digest=sha256:221212a4ad837b1b1a35aac1c410cb33b42f79021d522da7b21451ba77dc0506

Observation ea6607b3-075d-441d-b325-9991145e8573 · outbound

This paper cites Visual autoregressive modeling: Scalable image generation via next-scale prediction, 2024.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Visual autoregressive modeling: Scalable image generation via next-scale prediction, 2024

Reference 99

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:33:14.869486Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T15:33:13.571484Z digest=sha256:e14290eca2d5132e904176869b76c27e2f82e68c50b0519babc0c71aa4dda548

Observation dfefb3a5-a806-49a5-82cd-c9263dd43973 · outbound

This paper cites Givt: Generative infinite-vocabulary transformers.

SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer Givt: Generative infinite-vocabulary transformers

Reference 100

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T15:33:14.854543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-11T15:33:13.575809Z digest=sha256:bad68ac0a4675456caf6bbbe1362e6ece18d24d08a7fc48e56d74eed55a2fbd1

Pith citing papers

Observation be54b425-21eb-4843-b885-592c25502b3c · inbound

Masked Autoencoders Are Effective Tokenizers for Diffusion Models cites this paper.

Masked Autoencoders Are Effective Tokenizers for Diffusion Models SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-09T04:47:30.290200Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T04:47:30.290200Z digest=sha256:4377fac55741885b615d815a2e356a591d31c779cc38868f8ef8374b08f5405a

Observation 3913c8b6-1a18-49b6-8969-9b05bb6e4b43 · inbound

Instella-T2I: Pushing the Limits of 1D Discrete Latent Space Image Generation cites this paper.

Instella-T2I: Pushing the Limits of 1D Discrete Latent Space Image Generation SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-06T22:43:02.699039Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:43:02.699039Z digest=sha256:cbeb5ce4afc6366d50d045d0faf4233a9df4ec23144e06ced340898ed7b3a601

Observation 00cb8b07-238c-497a-97e4-1ca54dcd4c11 · inbound

Rethinking Discrete Tokens: Treating Them as Conditions for Continuous Autoregressive Image Synthesis cites this paper.

Rethinking Discrete Tokens: Treating Them as Conditions for Continuous Autoregressive Image Synthesis SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T20:49:40.554531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:49:40.554531Z digest=sha256:4b50dccec128ad1d507db447bfe1b4047750346f9170f54daeec49930f990e31

Observation 62f40565-28e3-4d95-bbfe-fa5f79d53533 · inbound

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer cites this paper.

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T19:41:23.624757Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:41:23.624757Z digest=sha256:75034f230bf7ea88301574a4cb3da8c8cccba8b3472bd262f821d00807677117

Observation 7508108d-4de6-4935-9311-117ed61ec7a3 · inbound

Structure over Pixels: Learning Variable-Length Visual Programs cites this paper.

Structure over Pixels: Learning Variable-Length Visual Programs SoftVQ-VAE: Efficient 1-Dimensional Continuous Tokenizer

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-06-29T18:13:49.030820Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-29T18:06:01.684713Z digest=sha256:a92c10eba09fb7cb6778505f05355d2d0cfd571a9ba5163d5f91668a071fa3cc