Pith. sign in

Paper Citation Record · LEDGER

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation

As of 14 August 2026, this Paper Citation Record lists 52 of 52 outbound references and 0 inbound Pith citation observations for arXiv:2605.30317.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2605.30317 v1

Coverage vector

measured 52 of 52 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-06-29T08:00:16.005187Z

measured 52 of 52 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

52 of 52 outbound references displayed

  • verified exact25
  • verified fuzzy0
  • unresolved24
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch3

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation ed3b6c15-2b04-4260-bcac-4123f1d00003 · outbound

This paper cites GPT-4 Technical Report.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation GPT-4 Technical Report

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-06-29T08:03:14.053299Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:46ecda525f3b1312ca299f28569cf476e7f6e3f3ccb106b1cb03b8d426b1d01e

Observation 93b16958-4035-4284-ab1a-64a6bcaf27dc · outbound

This paper cites Self-rectifying diffusion sampling with perturbed-attention guidance.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation Self-rectifying diffusion sampling with perturbed-attention guidance

Reference 2

Resolution
unresolved
no resolver link, observed 2026-06-29T08:00:16.005187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:3321c1763f42d54cbef58ead972c8bed833fd67a262cab505771ac87cc63ea69

Observation 3f27767c-42e8-4eea-8fd4-6c305beda22c · outbound

This paper cites Findings of the Association for Computational Linguistics: ACL 2022 , publisher =.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation Findings of the Association for Computational Linguistics: ACL 2022 , publisher =

Reference 3

Resolution
metadata mismatch
doi, observed 2026-06-29T08:03:13.618766Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:0976ea15a4da5f7b9a2bb36ebf413f9c24c4c15c293d8234a8d52d1cade50d94

Observation 95dc5809-f98c-4ea4-993c-0d9e3324bd7a · outbound

This paper cites Scheduled sampling for sequence prediction with recurrent neural networks.Advances in Neural Information Processing Systems, 28, 2015.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation Scheduled sampling for sequence prediction with recurrent neural networks.Advances in Neural Information Processing Systems, 28, 2015

Reference 4

Resolution
unresolved
no resolver link, observed 2026-06-29T08:00:16.005187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:9db59d69e176f502a08ab2526acd78c0b093552a52b439e03d70da5e74a50359

Observation 98c96b3d-7a37-4884-8ca2-439d4a4de3a6 · outbound

This paper cites Generative pretraining from pixels.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation Generative pretraining from pixels

Reference 5

Resolution
unresolved
no resolver link, observed 2026-06-29T08:00:16.005187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:3369c80d695f5f734b2de1b912c74852ab1d7dceae4552287982df850677d14c

Observation 6b471dcf-9997-4c40-9171-3ea36359ac57 · outbound

This paper cites Emerging Properties in Unified Multimodal Pretraining.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation Emerging Properties in Unified Multimodal Pretraining

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-06-29T08:03:14.050583Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:9e7dcdd138c8d123e31bf9fc7203d76fc770e4e4d9224b238f603a4db25daf67

Observation a4c74d2e-963c-48bc-b7c8-d6cbfe9d16d4 · outbound

This paper cites Imagenet: A large-scale hierarchical image database.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation Imagenet: A large-scale hierarchical image database

Reference 7

Resolution
unresolved
no resolver link, observed 2026-06-29T08:00:16.005187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:8f7acc4de479043cc5dcb565aaa23be5af321fcf1578acdb1d92c009af555306

Observation 400e1312-d99b-42a3-86e5-8dd993ef4428 · outbound

This paper cites Taming transformers for high-resolution image synthesis.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation Taming transformers for high-resolution image synthesis

Reference 8

Resolution
unresolved
no resolver link, observed 2026-06-29T08:00:16.005187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:9809b2ea30286fd3d0d3309b73f287e26bd97caf050cdbebae5096a36224db9f

Observation 73b74f10-6c50-46b8-abab-0853bdd662b1 · outbound

This paper cites Geneval: An object-focused framework for evaluating text-to-image alignment.Advances in Neural Information Processing Systems, 36:52132–52152, 2023.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation Geneval: An object-focused framework for evaluating text-to-image alignment.Advances in Neural Information Processing Systems, 36:52132–52152, 2023

Reference 9

Resolution
unresolved
no resolver link, observed 2026-06-29T08:00:16.005187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:f5def73cd926328068b9ace24081feaac2f5cf3ac1b163364ce37286aee158ce

Observation 0b7a5381-863c-4e21-8f0c-aa7c8c0a3322 · outbound

This paper cites Deep autoregressive networks.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation Deep autoregressive networks

Reference 10

Resolution
unresolved
no resolver link, observed 2026-06-29T08:00:16.005187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:9771da7696305338777f58d2aa5a7475782ab76745e5ec59c9b607d9a6047cf3

Observation 0e3b39a0-77ec-48da-a31f-42593d997426 · outbound

This paper cites Suboptimal behavior of bayes and mdl in classification under misspecification.Machine Learning, 66(2):119–149, 2007.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation Suboptimal behavior of bayes and mdl in classification under misspecification.Machine Learning, 66(2):119–149, 2007

Reference 11

Resolution
unresolved
no resolver link, observed 2026-06-29T08:00:16.005187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:5479afb3a0f1cb628f202e54f02a9407cdd9e406f60b9535003a0527c29aba92

Observation 31124dfc-aa24-450a-b2dd-c8e59e2e13a3 · outbound

This paper cites Infinity: Scaling bitwise autoregressive modeling for high-resolution image synthesis.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation Infinity: Scaling bitwise autoregressive modeling for high-resolution image synthesis

Reference 12

Resolution
unresolved
no resolver link, observed 2026-06-29T08:00:16.005187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:088864a7abcfc59d8c3005dee4427457b4fefd71a6f54967bd8a659387565484

Observation a3df8ecd-32d0-4cc0-ba45-436ff05cbab3 · outbound

This paper cites Conceptrol: Concept Control of Zero-shot Personalized Image Generation.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation Conceptrol: Concept Control of Zero-shot Personalized Image Generation

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-06-29T08:03:14.006326Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:ea63a97ed8e9cf6d6ae2416d732998741c3eb73a90f194e48e75d0fcf867a44c

Observation f607d092-18b6-4404-b2c6-6e747271918c · outbound

This paper cites AID: Attention Interpolation of Text-to-Image Diffusion.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation AID: Attention Interpolation of Text-to-Image Diffusion

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-06-29T08:03:14.041712Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:4b478695cfbf25e1f71f680a59e4d46cfa1c2b330656364fe95d3c2bb10d4aaf

Observation 9083b459-91cf-404e-9352-98adfe80ea31 · outbound

This paper cites REAR: Rethinking visual autoregressive models via generator-tokenizer consistency regularization.arXiv preprint arXiv:2510.04450, 2025.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation REAR: Rethinking visual autoregressive models via generator-tokenizer consistency regularization.arXiv preprint arXiv:2510.04450, 2025

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-06-29T08:03:13.991902Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:b69cfbe66725bc017467afba924dd28d8e6eb6d9bd69fc83514c6ca2fb95e8a2

Observation 75051b49-dd15-4a8d-92f0-1338ca6f9fb7 · outbound

This paper cites Classifier-Free Diffusion Guidance.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation Classifier-Free Diffusion Guidance

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-06-29T08:03:14.003382Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:a52110d508326a08134bd724df44d7ec2d8b7df29e30d233bf24359cbe4f4235

Observation cbf2529d-be3f-4743-9b91-f1f78e4566c5 · outbound

This paper cites Smoothed energy guidance: Guiding diffusion models with reduced energy curvature of attention.Advances in Neural Information Processing Systems, 37:66743–66772, 2024.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation Smoothed energy guidance: Guiding diffusion models with reduced energy curvature of attention.Advances in Neural Information Processing Systems, 37:66743–66772, 2024

Reference 17

Resolution
unresolved
no resolver link, observed 2026-06-29T08:00:16.005187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:bde68dfbcd53ae4cedad493947497826cd95fce093c2dcb3a042fc94f228a5bb

Observation 6057887a-a83a-4bf6-be73-32b9d480e92b · outbound

This paper cites ELLA: Equip Diffusion Models with LLM for Enhanced Semantic Alignment.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation ELLA: Equip Diffusion Models with LLM for Enhanced Semantic Alignment

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-06-29T08:03:14.026465Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:57e9a7e0ae4da46308ae22172f02773d2fd58d0fb4071a8e152b439d9ac9287d

Observation 2c046fa5-3350-4f41-b1e6-556a71750e6f · outbound

This paper cites Self Forcing: Bridging the Train-Test Gap in Autoregressive Video Diffusion.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation Self Forcing: Bridging the Train-Test Gap in Autoregressive Video Diffusion

Reference 19

Resolution
verified exact
local_arxiv, observed 2026-06-29T08:03:13.986844Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:e3581d49ad56bfff273cb35490dc1ba796c7b07f27d469da857939d8214558a7

Observation 07eb17f9-7820-462b-b61b-488dff237447 · outbound

This paper cites Vbench: Comprehensive benchmark suite for video generative models.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation Vbench: Comprehensive benchmark suite for video generative models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-06-29T08:00:16.005187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:4be127dbde10108ce98196364fee522cfb7f545c7c55b43d7512ec64a1c317a4

Observation c6c66a4c-7d7d-4b70-b362-002311b3800d · outbound

This paper cites Guiding a diffusion model with a bad version of itself.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation Guiding a diffusion model with a bad version of itself

Reference 21

Resolution
unresolved
no resolver link, observed 2026-06-29T08:00:16.005187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:9553f6b49a3ba22f6804432e7f00643ddd0d78fff071257069c36a7b14f96dff

Observation bef1f905-a5e6-4dca-8756-6ba81abe32d8 · outbound

This paper cites HunyuanVideo: A Systematic Framework For Large Video Generative Models.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation HunyuanVideo: A Systematic Framework For Large Video Generative Models

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-06-29T08:03:14.000811Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:65638d143d3daac666aeb30c3ca19412d708df5b05febb24a7f26b292f7d2fa3

Observation 2c2eadf8-e9c0-40ae-b4fe-2014487d94d9 · outbound

This paper cites FLUX.2: Frontier Visual Intelligence.https://bfl.ai/blog/flux-2, 2025.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation FLUX.2: Frontier Visual Intelligence.https://bfl.ai/blog/flux-2, 2025

Reference 23

Resolution
unresolved
no resolver link, observed 2026-06-29T08:00:16.005187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:23462a009f7a0d3ed0b7c4cf4993a331c440a4761862c4278d42074491fdafe2

Observation 5d139193-acd6-4058-8e74-0999205b0ab0 · outbound

This paper cites Lamb, Anirudh Goyal, Ying Zhang, Saizheng Zhang, Aaron C.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation Lamb, Anirudh Goyal, Ying Zhang, Saizheng Zhang, Aaron C

Reference 24

Resolution
unresolved
no resolver link, observed 2026-06-29T08:00:16.005187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:a0e0768994781eb3a81fdf53b4fa252f6202863bd606166f7f8bb5c73e4d804b

Observation a4e4e4e6-d8eb-4a08-8fc3-28f786052ca0 · outbound

This paper cites Autoregressive image generation without vector quantization.Advances in Neural Information Processing Systems, 37:56424–56445, 2024.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation Autoregressive image generation without vector quantization.Advances in Neural Information Processing Systems, 37:56424–56445, 2024

Reference 25

Resolution
unresolved
no resolver link, observed 2026-06-29T08:00:16.005187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:8af36beb30d60bfbb781402529d7a5362ebbf00b4c1471fc7384c23def0b0d82

Observation 8a83ec89-a026-414e-8ce8-b7400950421a · outbound

This paper cites arXiv:2512.19680 (2025) 5.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation arXiv:2512.19680 (2025) 5

Reference 26

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T08:03:14.044788Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:933dd4751f14880718eb6487dd90a209ed50f75a3205d41fdcd7a8dba4dcbc8f

Observation e2f44700-b3bf-4fee-b294-a22df8a5e414 · outbound

This paper cites Step-level Reward for Free in RL-based T2I Diffusion Model Fine-tuning.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation Step-level Reward for Free in RL-based T2I Diffusion Model Fine-tuning

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-06-29T08:03:13.984555Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:b40d149c452bb14d3ce92b05435c1fc3c8a73fcb3e2bb4e537cb54964e38230f

Observation 3e2fce92-ecd4-4df1-ace6-002cd328c80c · outbound

This paper cites Infini- tystar: Unified spacetime autoregressive modeling for visual generation.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation Infini- tystar: Unified spacetime autoregressive modeling for visual generation

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-06-29T08:03:13.998565Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:b22c7a68487060b8ef6f8a16a0ed17b471a9a70b17bdf229c881e9ed2dafe51e

Observation 4db29d96-217c-45f7-a594-85352b5f3947 · outbound

This paper cites Interp3d: Correspondence-aware interpolation for generative textured 3d morphing.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation Interp3d: Correspondence-aware interpolation for generative textured 3d morphing

Reference 29

Resolution
unresolved
no resolver link, observed 2026-06-29T08:00:16.005187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:2655c452399426ef7e3d6db735e5529f58675d3a07494986852c26f178c2679e

Observation 90e4471b-9750-40f8-a803-89c407d567c9 · outbound

This paper cites Unitok: A unified tokenizer for visual generation and understanding.arXiv preprint arXiv:2502.20321, 2025a.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation Unitok: A unified tokenizer for visual generation and understanding.arXiv preprint arXiv:2502.20321, 2025a

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-06-29T08:03:14.038473Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:d6b3a48177a801eef51e3cd20bf216046d36d8c367dcbdaa873a271c4c552eea

Observation b4fce583-aa95-4a38-ac42-7b53996283d3 · outbound

This paper cites Finite Scalar Quantization: VQ-VAE Made Simple.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation Finite Scalar Quantization: VQ-VAE Made Simple

Reference 31

Resolution
verified exact
local_arxiv, observed 2026-06-29T08:03:14.035462Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:250b5972d9c0e25363f4e67c7ff1a5ec6fb86f799a982631a37eb90cf77f74bc

Observation 18777129-4c80-40b6-aa14-d1d680a1f58b · outbound

This paper cites Image transformer.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation Image transformer

Reference 32

Resolution
unresolved
no resolver link, observed 2026-06-29T08:00:16.005187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:fba6a83221894c70d062c14e5b220b626023519ee6724cb243d6cb507b3843a6

Observation 2ef82b06-ad04-43e7-b1e7-e8bf54ea51ba · outbound

This paper cites Beyond Next-Token: Next-X Prediction for Autoregressive Visual Generation.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation Beyond Next-Token: Next-X Prediction for Autoregressive Visual Generation

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-06-29T08:03:14.032693Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:47bfe4910e2c71522f786eda54227880314d2462647f3c3a7c5442ea72dacd15

Observation 2e13bba3-f311-43e2-b66f-5e8e7ace93e9 · outbound

This paper cites A reduction of imitation learning and structured prediction to no-regret online learning.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation A reduction of imitation learning and structured prediction to no-regret online learning

Reference 34

Resolution
unresolved
no resolver link, observed 2026-06-29T08:00:16.005187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:9469ede2dfd02e40c1c510d9895c7cb8c8c1665bffadf7db6425b68356ef98eb

Observation 1b532297-0cc7-4230-9e72-7bd436286f91 · outbound

This paper cites Generalization in generation: A closer look at exposure bias.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation Generalization in generation: A closer look at exposure bias

Reference 35

Resolution
verified exact
doi, observed 2026-06-29T08:03:13.622848Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:fd078d05913517cee5fbe069ca3d5085074e394900e489a404c81824b0ad02ea

Observation a0dab1f5-8885-4b0c-985f-dad850f4a79d · outbound

This paper cites SSG: Scaled spatial guidance for multi-scale visual autoregressive generation.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation SSG: Scaled spatial guidance for multi-scale visual autoregressive generation

Reference 36

Resolution
unresolved
no resolver link, observed 2026-06-29T08:00:16.005187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:0dba7d616ad34c107a08a9d5a66764ca4781c6645dcdab6d68a9d7d274e2359c

Observation 6fea1418-0209-443c-9c0b-52757ead38d6 · outbound

This paper cites Autoregressive Model Beats Diffusion: Llama for Scalable Image Generation.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation Autoregressive Model Beats Diffusion: Llama for Scalable Image Generation

Reference 37

Resolution
verified exact
local_arxiv, observed 2026-06-29T08:03:14.020689Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:a95d8d8837abb9d36c6ebfaa86bd2cf94251dfd3b852da25ba103569d17a7d79

Observation 4991b45d-f9ca-4900-a43a-82df66351c11 · outbound

This paper cites Chameleon: Mixed-Modal Early-Fusion Foundation Models.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation Chameleon: Mixed-Modal Early-Fusion Foundation Models

Reference 38

Resolution
verified exact
local_arxiv, observed 2026-06-29T08:03:14.011397Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:b881700900ef23af86ad1d63adaaa1d599b3a57be33f24bd19a173f759298b00

Observation fa2401a5-0ed9-4ae7-a9aa-c4162a9fbe12 · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation Gemini: A Family of Highly Capable Multimodal Models

Reference 39

Resolution
verified exact
local_arxiv, observed 2026-06-29T08:03:14.005561Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:aecd229df7ca4a3780944dac3fe1b45a22e18c77011352a8baa923130610e241

Observation c0a7076f-c789-4122-9f2e-e96ed35b0d93 · outbound

This paper cites Visual autoregressive modeling: Scalable image generation via next-scale prediction.Advances in neural information processing systems, 37:84839–84865, 2024.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation Visual autoregressive modeling: Scalable image generation via next-scale prediction.Advances in neural information processing systems, 37:84839–84865, 2024

Reference 40

Resolution
unresolved
no resolver link, observed 2026-06-29T08:00:16.005187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:ff7d64a7207530dc616817eba067822e6e298346c49822b587bb920182b283bc

Observation 6c1dd962-3c73-45f3-80bc-6ef55b7e4142 · outbound

This paper cites Neural discrete representation learning.Advances in neural information processing systems, 30, 2017.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation Neural discrete representation learning.Advances in neural information processing systems, 30, 2017

Reference 41

Resolution
unresolved
no resolver link, observed 2026-06-29T08:00:16.005187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:67984e58ab9c4909e9c4347c2433044361ed64c93ec2278f86468b28f142d3a7

Observation 1151c0a7-cda7-4d45-8b4e-cc39787496fb · outbound

This paper cites Wan: Open and Advanced Large-Scale Video Generative Models.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation Wan: Open and Advanced Large-Scale Video Generative Models

Reference 42

Resolution
verified exact
local_arxiv, observed 2026-06-29T08:03:14.029311Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:a9db8e222d7b590d23ab707eec92c823ddcd8bc8e5519aeac4948aebd8747893

Observation 8c1d0f35-8b15-4af2-8133-d6d0c427f24b · outbound

This paper cites On Exposure Bias, Hallucination and Domain Shift in Neural Machine Translation.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation On Exposure Bias, Hallucination and Domain Shift in Neural Machine Translation

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-06-29T08:03:14.047424Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:40f586b943a7e5c8aca0050d9b1bbfe0cbcc411a5f6be3b4674634d89cd60d24

Observation 0ed9c76e-7e73-46df-9462-2063178dcb34 · outbound

This paper cites Blaschko.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation Blaschko

Reference 44

Resolution
unresolved
no resolver link, observed 2026-06-29T08:00:16.005187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:0379d427329d67becb8bd84a050f5cb9d9caefe122e380f74e53bcfd93481412

Observation c7df9d30-7ee3-4de0-a9a7-7c3d1b2aa685 · outbound

This paper cites an unresolved cited work.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation Unresolved cited work

Reference 45

Resolution
unresolved
no resolver link, observed 2026-06-29T08:00:16.005187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:b07878bfdde166bb9e8edf4dc057f97279702be1f539bb00bc44d3bcee3dbf10

Observation 72d0987f-c2ca-4c5e-a5a5-0a17ef30a38b · outbound

This paper cites Infotok: Adaptive discrete video tokenizer via information- theoretic compression.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation Infotok: Adaptive discrete video tokenizer via information- theoretic compression

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-06-29T08:03:14.024593Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:34d58594a8dba07021ebdc0db8739ed3311e8ad64fa162658495b260a4282b63

Observation 601e44bc-c052-4d4b-89f9-e011dd6a7e21 · outbound

This paper cites Language Model Beats Diffusion -- Tokenizer is Key to Visual Generation.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation Language Model Beats Diffusion -- Tokenizer is Key to Visual Generation

Reference 47

Resolution
verified exact
local_arxiv, observed 2026-06-29T08:03:14.019333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:8aec4d1dc18d3d31887ef43ba39cd6d866814f275d3c032346a059bbf7603960

Observation d2b5d881-b1e2-413b-b412-6f87a6623a31 · outbound

This paper cites An image is worth 32 tokens for reconstruction and generation.Advances in Neural Information Processing Systems, 37:128940–128966, 2024.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation An image is worth 32 tokens for reconstruction and generation.Advances in Neural Information Processing Systems, 37:128940–128966, 2024

Reference 48

Resolution
unresolved
no resolver link, observed 2026-06-29T08:00:16.005187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:21aaa3c92db82bdb1f29f05118b3c89d8560a59bd6d986ab4bf7bf6c72677e22

Observation f6202b98-316e-46fa-b1a7-5f1084975064 · outbound

This paper cites Guiding a Diffusion Model by Swapping Its Tokens.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation Guiding a Diffusion Model by Swapping Its Tokens

Reference 49

Resolution
verified exact
local_arxiv, observed 2026-06-29T08:03:14.023553Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:4a85391e8612cf86c7879048f77edb88e3cfe6dd42f6074cb6b47ce5bd1f70fc

Observation 02c37d58-e6fb-466a-b1e0-f19f5d5dcb95 · outbound

This paper cites author Feng, Y.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation author Feng, Y

Reference 50

Resolution
metadata mismatch
doi, observed 2026-06-29T08:03:13.620657Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:ee4d4260787082d5892c4c4dc11560b99590b0d65ebbeb0ee5cffdef674b2294

Observation 5195f449-27e6-47ae-aa2a-7076b077e4c8 · outbound

This paper cites Image and Video Tokenization with Binary Spherical Quantization.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation Image and Video Tokenization with Binary Spherical Quantization

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-06-29T08:03:14.008832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:bd5c802cb37fba04cd4c6656fde7687caf0b93a527e040c0a3de48406bd7fcbe

Observation 14e6710d-d70c-4bd7-a6ca-16c18b370ed5 · outbound

This paper cites RelaxFlow: Text-Driven Amodal 3D Generation.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation RelaxFlow: Text-Driven Amodal 3D Generation

Reference 52

Resolution
verified exact
local_arxiv, observed 2026-06-29T08:03:14.017742Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:86f35da920cadb98b6d7b3b0dd80778a80f149b05a7d76809f2f116589e54de3

Pith citing papers

No inbound Pith citation observations are available.