Pith. sign in

Paper Citation Record · LEDGER

Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 19 inbound Pith citation observations for arXiv:2503.13436.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2503.13436 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 19 of 19 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 19 of 19 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T00:30:41.596688Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 0d29f49f-1b5b-4bdc-a518-b8d1ffc5bbed · inbound

Step1X-Edit: A Practical Framework for General Image Editing cites this paper.

Step1X-Edit: A Practical Framework for General Image Editing Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-11T14:36:41.912974Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-11T14:36:41.467429Z digest=sha256:1d6f44063ad4d33b0326e31a5691522227190f7e68799dcc07ba3395e519feba

Observation 33621fbe-ebb6-43a6-a7a6-bd47ca6318c0 · inbound

Fake it till You Make it: Reward Modeling as Discriminative Prediction cites this paper.

Fake it till You Make it: Reward Modeling as Discriminative Prediction Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T00:30:41.596688Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:30:41.596688Z digest=sha256:cf37944a604323f1d6fceb6ca91fe20d23a23f2715aac663673f116e0c753e99

Observation ae8ad5d0-247d-4c36-9353-0225eb3b27b7 · inbound

Show-o2: Improved Native Unified Multimodal Models cites this paper.

Show-o2: Improved Native Unified Multimodal Models Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-05-12T18:51:16.091886Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-12T18:51:15.428692Z digest=sha256:1ece7ecfd30fd7075f737468957cde37450ada83b1c52587fb11409111f95ba1

Observation 254445f6-6f51-42b2-a357-beb673a5d945 · inbound

NeoBabel: A Multilingual Open Tower for Visual Generation cites this paper.

NeoBabel: A Multilingual Open Tower for Visual Generation Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T19:15:26.807532Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:15:26.807532Z digest=sha256:807cd5e6c056649fcef0b4ffed40c3bd4f3f930dc7f5dfdbde099a49ea697c37

Observation 3e050806-5d68-47c6-a051-591c3ae8a5e5 · inbound

Generative Distribution Distillation cites this paper.

Generative Distribution Distillation Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T16:10:22.068784Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:10:22.068784Z digest=sha256:099d82f6bdaaf2f255c4c5b0075ebb5cb8bbe964f888fe7305a730e3629921b4

Observation 4afd479f-c989-4a8c-92b3-2384f48e0880 · inbound

Skywork UniPic: Unified Autoregressive Modeling for Visual Understanding and Generation cites this paper.

Skywork UniPic: Unified Autoregressive Modeling for Visual Understanding and Generation Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T04:37:13.531723Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T04:37:13.531723Z digest=sha256:2abe75be2405b8011c29232f75fd51f47dcb9e48006261351933f21b379b7052

Observation 03e24383-60cf-4717-b0b4-e9cecdc8022f · inbound

Reconstruction Alignment Improves Unified Multimodal Models cites this paper.

Reconstruction Alignment Improves Unified Multimodal Models Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-04T22:36:07.931702Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T22:36:07.931702Z digest=sha256:358e3f51f63a40f0179e372d7d4d1fbb38999ade57359661961436b5195b1b74

Observation 9a95757f-7988-4315-9698-f93aab884bd3 · inbound

Image Diffusion Preview with Consistency Solver cites this paper.

Image Diffusion Preview with Consistency Solver Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-16T21:43:34.705010Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-16T21:42:24.098676Z digest=sha256:8fe7a6af9adebdae8e87133ef613d61b5622228995fba0d989d086d9534dddfd

Observation bbaeaf67-53dd-49ed-ac6f-cdd0282a82a3 · inbound

Generation Enhances Understanding in Unified Multimodal Models via Multi-Representation Generation cites this paper.

Generation Enhances Understanding in Unified Multimodal Models via Multi-Representation Generation Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-03T07:03:12.169209Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T07:03:12.169209Z digest=sha256:18fb564eb4f1ceb84012a731b30065c2456f12514e5c6594a6d1a7a7ee1d92d2

Observation c6306d4c-b4f4-41ef-8581-3d1ac64ef55a · inbound

FeynmanBench: Benchmarking Multimodal LLMs on Diagrammatic Physics Reasoning cites this paper.

FeynmanBench: Benchmarking Multimodal LLMs on Diagrammatic Physics Reasoning Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-13T16:52:59.401742Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T16:51:48.705876Z digest=sha256:47f0124d84faf56b0d3824d670d8f10e65bc89f9731067b3808e5822607b40b4

Observation fccf5d1f-9135-4504-a707-6f9f71dea4b6 · inbound

FeynmanBench: Benchmarking Multimodal LLMs on Diagrammatic Physics Reasoning cites this paper.

FeynmanBench: Benchmarking Multimodal LLMs on Diagrammatic Physics Reasoning Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

Reference 14

Resolution
unresolved
no resolver link, observed 2026-07-13T12:10:53.720348Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T12:10:53.720348Z digest=sha256:f4ec24186a2fbd085a726cefe2668b3bc2945244de87fb9c834ae8a1be18479d

Observation d273a90e-10cc-40a3-b427-df164a0fbe4c · inbound

Pseudo-Unification: Entropy Probing Reveals Divergent Information Patterns in Unified Multimodal Models cites this paper.

Pseudo-Unification: Entropy Probing Reveals Divergent Information Patterns in Unified Multimodal Models Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-11T08:50:59.145717Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T16:27:14.491492Z digest=sha256:9df733a67312d3ba2d9a72ddf6e63e2d6eac4689c276edb31e4ebbcc3e857bce

Observation 61a52f06-dec5-4951-99fa-95a5ef2d98de · inbound

Tuna-2: Pixel Embeddings Beat Vision Encoders for Multimodal Understanding and Generation cites this paper.

Tuna-2: Pixel Embeddings Beat Vision Encoders for Multimodal Understanding and Generation Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-11T21:41:18.823854Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-08T04:31:26.325118Z digest=sha256:583c529242d8982c631f637972a042b16f1c2e640ae52401400e0d8005bb00eb

Observation 1683cb3c-e9e0-4994-9a8a-409e6b24e763 · inbound

Tuna-2: Pixel Embeddings Beat Vision Encoders for Multimodal Understanding and Generation cites this paper.

Tuna-2: Pixel Embeddings Beat Vision Encoders for Multimodal Understanding and Generation Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-20T23:43:51.078460Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-20T23:41:25.275207Z digest=sha256:8135571daa095dc5faafd58aff1035d920c1c6888eea3766a3e84841284c08fc

Observation b91df5b1-f000-4f06-85b7-66cb885b08e3 · inbound

Lance: Unified Multimodal Modeling by Multi-Task Synergy cites this paper.

Lance: Unified Multimodal Modeling by Multi-Task Synergy Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-20T11:48:14.918323Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-20T11:46:52.658984Z digest=sha256:a18a12f5f7c246e6124cd78a563e51da6a5dab6febfd60932678666440eb7851

Observation d5172ce2-5713-4439-8ec6-0c928cda97c6 · inbound

Lance: Unified Multimodal Modeling by Multi-Task Synergy cites this paper.

Lance: Unified Multimodal Modeling by Multi-Task Synergy Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-21T07:59:50.752143Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-21T07:56:34.034047Z digest=sha256:91865e8656317e646f01bb1d539d2d4419b2719d10f0efb8de6553e4f8be851d

Observation fc4b3c3a-b56b-4259-a0e3-9a079cd900d8 · inbound

HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers cites this paper.

HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

Reference 280

Resolution
verified exact
arxiv_id, observed 2026-07-03T14:28:32.117532Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-27T07:01:07.362430Z digest=sha256:fcff93f2ee1904f27ca4f17956249eb94cd00109a42f9ecba4bccc6fb0294398

Observation cc2b5ce7-2ac5-4c70-b38c-bf713142dea9 · inbound

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation cites this paper.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

Reference 104

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.861564Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.861564Z digest=sha256:e37df84b79ed8418853333d5b1b573f8e62c8395bb9937b361004236f1df27b8

Observation 951e616b-0053-4d9b-a2ec-a51aa180cb9d · inbound

Instruction-based Image Editing: A Survey on Data, Models, Evaluation, and Applications cites this paper.

Instruction-based Image Editing: A Survey on Data, Models, Evaluation, and Applications Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-01T01:51:57.970306Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:51:57.970306Z digest=sha256:b9bedc508f0c90b32aa7d8fcb940a9b2a4765267464865c1748e84e739bab5c5