Pith. sign in

Paper Citation Record · LEDGER

Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

As of 11 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 19 inbound Pith citation observations for arXiv:2503.13436.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2503.13436 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 19 of 19 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 19 of 19 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T00:30:41.596688Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 0d29f49f-1b5b-4bdc-a518-b8d1ffc5bbed · inbound

Step1X-Edit: A Practical Framework for General Image Editing cites this paper.

Step1X-Edit: A Practical Framework for General Image Editing Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-11T14:36:41.912974Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-11T14:36:41.467429Z digest=sha256:7c517cbd4af99a70652d16bd62adcb2a1ca8e5b221e22dfc5c8fe63ff90115a1

Observation 33621fbe-ebb6-43a6-a7a6-bd47ca6318c0 · inbound

Fake it till You Make it: Reward Modeling as Discriminative Prediction cites this paper.

Fake it till You Make it: Reward Modeling as Discriminative Prediction Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T00:30:41.596688Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:30:41.596688Z digest=sha256:69702aedd68b8f099c3ac88dcfdf61f02766a6b862a6fd55db507e4c119e3434

Observation ae8ad5d0-247d-4c36-9353-0225eb3b27b7 · inbound

Show-o2: Improved Native Unified Multimodal Models cites this paper.

Show-o2: Improved Native Unified Multimodal Models Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-05-12T18:51:16.091886Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-12T18:51:15.428692Z digest=sha256:5863354e6dbbca9234ef019066f98522b74c32b31070450aee390c725a15ef7f

Observation 254445f6-6f51-42b2-a357-beb673a5d945 · inbound

NeoBabel: A Multilingual Open Tower for Visual Generation cites this paper.

NeoBabel: A Multilingual Open Tower for Visual Generation Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T19:15:26.807532Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:15:26.807532Z digest=sha256:0ced1a4815537f4bb070c879cc76f6f2198f1904a2eff141ab136a0b15ff5177

Observation 3e050806-5d68-47c6-a051-591c3ae8a5e5 · inbound

Generative Distribution Distillation cites this paper.

Generative Distribution Distillation Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T16:10:22.068784Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:10:22.068784Z digest=sha256:ef39a31d397ac0805b8aa78d039be0d2eda00f46261a7c1501aba0eaaa7adaaa

Observation 4afd479f-c989-4a8c-92b3-2384f48e0880 · inbound

Skywork UniPic: Unified Autoregressive Modeling for Visual Understanding and Generation cites this paper.

Skywork UniPic: Unified Autoregressive Modeling for Visual Understanding and Generation Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T04:37:13.531723Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T04:37:13.531723Z digest=sha256:ee6d424eb7d6a3d44ace7c6c1c54c59987f008ecf46631ca9d7fa669dfb916a4

Observation 03e24383-60cf-4717-b0b4-e9cecdc8022f · inbound

Reconstruction Alignment Improves Unified Multimodal Models cites this paper.

Reconstruction Alignment Improves Unified Multimodal Models Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-04T22:36:07.931702Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T22:36:07.931702Z digest=sha256:4ef92bd29bd424960ac5e7697e65330a00fca4306fe37244bc65fe1f798d0775

Observation 9a95757f-7988-4315-9698-f93aab884bd3 · inbound

Image Diffusion Preview with Consistency Solver cites this paper.

Image Diffusion Preview with Consistency Solver Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-16T21:43:34.705010Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-16T21:42:24.098676Z digest=sha256:3193c96aeb807c75b97d462188e6ae8ccd1e68dabbdfa8144d6c11788533f73e

Observation bbaeaf67-53dd-49ed-ac6f-cdd0282a82a3 · inbound

Generation Enhances Understanding in Unified Multimodal Models via Multi-Representation Generation cites this paper.

Generation Enhances Understanding in Unified Multimodal Models via Multi-Representation Generation Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-03T07:03:12.169209Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T07:03:12.169209Z digest=sha256:214f36c0dc36247801ed42563eb181f359f3549f413d1d374fadf08ebf20655f

Observation c6306d4c-b4f4-41ef-8581-3d1ac64ef55a · inbound

FeynmanBench: Benchmarking Multimodal LLMs on Diagrammatic Physics Reasoning cites this paper.

FeynmanBench: Benchmarking Multimodal LLMs on Diagrammatic Physics Reasoning Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-13T16:52:59.401742Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-13T16:51:48.705876Z digest=sha256:08e9ab91c8cd1e7a02e5ae75d8820857faf37c3e944184dbe1783e02298c4416

Observation fccf5d1f-9135-4504-a707-6f9f71dea4b6 · inbound

FeynmanBench: Benchmarking Multimodal LLMs on Diagrammatic Physics Reasoning cites this paper.

FeynmanBench: Benchmarking Multimodal LLMs on Diagrammatic Physics Reasoning Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

Reference 14

Resolution
unresolved
no resolver link, observed 2026-07-13T12:10:53.720348Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T12:10:53.720348Z digest=sha256:6f678cfa639906bbc7f93f0628a68e005fd33e5a26d2038f66bb8d40ce88d49e

Observation d273a90e-10cc-40a3-b427-df164a0fbe4c · inbound

Pseudo-Unification: Entropy Probing Reveals Divergent Information Patterns in Unified Multimodal Models cites this paper.

Pseudo-Unification: Entropy Probing Reveals Divergent Information Patterns in Unified Multimodal Models Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-11T08:50:59.145717Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-10T16:27:14.491492Z digest=sha256:06d6de2f4116abb6cc85344faf7419dabc3431a7d912e297004c28c1049d2b19

Observation 61a52f06-dec5-4951-99fa-95a5ef2d98de · inbound

Tuna-2: Pixel Embeddings Beat Vision Encoders for Multimodal Understanding and Generation cites this paper.

Tuna-2: Pixel Embeddings Beat Vision Encoders for Multimodal Understanding and Generation Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-11T21:41:18.823854Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-08T04:31:26.325118Z digest=sha256:5ec742fd7ef1b19842038509bc644997e2015144c1df16a1493afcd3ec037e7f

Observation 1683cb3c-e9e0-4994-9a8a-409e6b24e763 · inbound

Tuna-2: Pixel Embeddings Beat Vision Encoders for Multimodal Understanding and Generation cites this paper.

Tuna-2: Pixel Embeddings Beat Vision Encoders for Multimodal Understanding and Generation Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-20T23:43:51.078460Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-20T23:41:25.275207Z digest=sha256:1d42eba760af48063a21990c37d99bbbd936121a9c1e2a26b2e33d0cbeae546a

Observation b91df5b1-f000-4f06-85b7-66cb885b08e3 · inbound

Lance: Unified Multimodal Modeling by Multi-Task Synergy cites this paper.

Lance: Unified Multimodal Modeling by Multi-Task Synergy Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-20T11:48:14.918323Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-20T11:46:52.658984Z digest=sha256:cfbe0295a5ea6a74fc9b083cd515ac34d20e2d23ec86e77d8cb932fa68bf192c

Observation d5172ce2-5713-4439-8ec6-0c928cda97c6 · inbound

Lance: Unified Multimodal Modeling by Multi-Task Synergy cites this paper.

Lance: Unified Multimodal Modeling by Multi-Task Synergy Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-21T07:59:50.752143Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-21T07:56:34.034047Z digest=sha256:a1ed9a7e779932f34d32efa72273e2326eac55cec401b8e00a714a4bec9ee848

Observation fc4b3c3a-b56b-4259-a0e3-9a079cd900d8 · inbound

HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers cites this paper.

HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

Reference 280

Resolution
verified exact
arxiv_id, observed 2026-07-03T14:28:32.117532Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-06-27T07:01:07.362430Z digest=sha256:8246b282c67e6e7afed20bbc85d1a9ea887541e894e65fa5c2b3cfd7b7db6f57

Observation cc2b5ce7-2ac5-4c70-b38c-bf713142dea9 · inbound

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation cites this paper.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

Reference 104

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.861564Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.861564Z digest=sha256:fd38d2444d4bd38fa2629d67d22aad1c2d96dfa9457feb6a8d9b5f854b7fae9c

Observation 951e616b-0053-4d9b-a2ec-a51aa180cb9d · inbound

Instruction-based Image Editing: A Survey on Data, Models, Evaluation, and Applications cites this paper.

Instruction-based Image Editing: A Survey on Data, Models, Evaluation, and Applications Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-01T01:51:57.970306Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:51:57.970306Z digest=sha256:d9bf815a2b50c8d34c5a90d5d9a41e141ba99b766853688450dae94b0b6464b8