Pith. sign in

Paper Citation Record · LEDGER

Fluid: Scaling Autoregressive Text-to-image Generative Models with Continuous Tokens

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 28 inbound Pith citation observations for arXiv:2410.13863.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2410.13863 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 28 of 28 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 28 of 28 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:56:34.349931Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-22T17:51:54.590192Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation dfe8eb31-0018-4818-b19a-141288966b5d · inbound

WISE: A World Knowledge-Informed Semantic Evaluation for Text-to-Image Generation cites this paper.

WISE: A World Knowledge-Informed Semantic Evaluation for Text-to-Image Generation Fluid: Scaling Autoregressive Text-to-image Generative Models with Continuous Tokens

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-15T16:24:27.572621Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T16:24:27.407376Z digest=sha256:71209d6579481989a290e3a1051a7213ba6c47f16aa2f6547d9a9a882cb4c4d0

Observation 72c0d586-95d7-4954-8201-0f1f8893fedd · inbound

Long-Context Autoregressive Video Modeling with Next-Frame Prediction cites this paper.

Long-Context Autoregressive Video Modeling with Next-Frame Prediction Fluid: Scaling Autoregressive Text-to-image Generative Models with Continuous Tokens

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-16T23:05:17.363988Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-16T23:05:17.201790Z digest=sha256:fc19120c6a61408f26570add9b0f903b383e1e6b3bd508fa1cbeadf05cc9a3ba

Observation 55a2de20-bfa7-40a1-82e5-d602c3175353 · inbound

Distilling Specialized Orders for Visual Generation cites this paper.

Distilling Specialized Orders for Visual Generation Fluid: Scaling Autoregressive Text-to-image Generative Models with Continuous Tokens

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-22T17:51:54.593401Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-22T17:51:04.337451Z digest=sha256:dcef103a8e8be5b7e717da7f581b035cf994c678621a4c61e4b46d76337c6773

Observation 63b67dff-4f12-45d7-a84c-d065bb2cc811 · inbound

CoMo: Learning Continuous Latent Motion from Internet Videos for Scalable Robot Learning cites this paper.

CoMo: Learning Continuous Latent Motion from Internet Videos for Scalable Robot Learning Fluid: Scaling Autoregressive Text-to-image Generative Models with Continuous Tokens

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T14:56:34.349931Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:56:34.349931Z digest=sha256:38977328e05c3bc6f610e303868075e489d17a20f85a2b8343f046791755bb5c

Observation d9d632a7-038c-4a2c-a2be-6af5687e0ddf · inbound

OmniGenBench: A Benchmark for Omnipotent Multimodal Generation across 50+ Tasks cites this paper.

OmniGenBench: A Benchmark for Omnipotent Multimodal Generation across 50+ Tasks Fluid: Scaling Autoregressive Text-to-image Generative Models with Continuous Tokens

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T14:28:30.474655Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:28:30.474655Z digest=sha256:23d5deecef072d3d464d995214f59221f67183d238f6bba7aa15f2831c5eb323

Observation ff9a1bf0-9ae0-41b2-bb10-e2ba1b3f28b3 · inbound

Plug-and-Play Context Feature Reuse for Efficient Masked Generation cites this paper.

Plug-and-Play Context Feature Reuse for Efficient Masked Generation Fluid: Scaling Autoregressive Text-to-image Generative Models with Continuous Tokens

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T14:22:20.617028Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:22:20.617028Z digest=sha256:414e1ca5871ebdac25158b64dc036d52929be05ca01fb32ac5547c20d3a180ff

Observation 52a0af3a-fa36-40b7-9763-9d6f62c03901 · inbound

StyleAR: Customizing Multimodal Autoregressive Model for Style-Aligned Text-to-Image Generation cites this paper.

StyleAR: Customizing Multimodal Autoregressive Model for Style-Aligned Text-to-Image Generation Fluid: Scaling Autoregressive Text-to-image Generative Models with Continuous Tokens

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:46.457415Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:09:46.457415Z digest=sha256:cc13994fc444538f3d8da1181a83f688b754bc7ccf51cc21eef5c45e0a9752b2

Observation 51c11f31-2912-494f-8d90-0423cc46167f · inbound

Hierarchical Masked Autoregressive Models with Low-Resolution Token Pivots cites this paper.

Hierarchical Masked Autoregressive Models with Low-Resolution Token Pivots Fluid: Scaling Autoregressive Text-to-image Generative Models with Continuous Tokens

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T14:00:13.347574Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:00:13.347574Z digest=sha256:ff885f0c5a9c8870e11c8afcedc3bd51e7ff05f5507f2dd41dfd30981a7dd424

Observation c3cff3eb-de30-4ddc-aa1f-7aca6106a64b · inbound

Muddit: Liberating Generation Beyond Text-to-Image with a Unified Discrete Diffusion Model cites this paper.

Muddit: Liberating Generation Beyond Text-to-Image with a Unified Discrete Diffusion Model Fluid: Scaling Autoregressive Text-to-image Generative Models with Continuous Tokens

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-19T13:02:18.369788Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-19T12:59:31.454155Z digest=sha256:2e85dac67564d90c954176b1ff7b9b8842537687c0591c49556481460f4d9d49

Observation f3d902a7-8066-4780-8d74-ea75a29bee7a · inbound

IMPACT: Iterative Mask-based Parallel Decoding for Text-to-Audio Generation with Diffusion Modeling cites this paper.

IMPACT: Iterative Mask-based Parallel Decoding for Text-to-Audio Generation with Diffusion Modeling Fluid: Scaling Autoregressive Text-to-image Generative Models with Continuous Tokens

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T12:03:24.409891Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:03:24.409891Z digest=sha256:cfe35367be46d4e4c1221264bf99493ce33bf030cf20c7347ac219e8886adbbb

Observation 70747349-f9d1-4fb1-b3cf-a2831fb2bb73 · inbound

STARFlow: Scaling Latent Normalizing Flows for High-resolution Image Synthesis cites this paper.

STARFlow: Scaling Latent Normalizing Flows for High-resolution Image Synthesis Fluid: Scaling Autoregressive Text-to-image Generative Models with Continuous Tokens

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T06:02:59.556863Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T06:02:59.556863Z digest=sha256:3517ecb236a4aded76be200671ce3023f689914d7cdcf42220e45f9bea98e0f9

Observation 497c92b9-0f40-4acc-b686-5218abdba6dc · inbound

VideoMAR: Autoregressive Video Generatio with Continuous Tokens cites this paper.

VideoMAR: Autoregressive Video Generatio with Continuous Tokens Fluid: Scaling Autoregressive Text-to-image Generative Models with Continuous Tokens

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T00:30:27.113734Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:30:27.113734Z digest=sha256:57246bd59167a426f788c6d8fb175fd0f14b06ddaa0f1426e9af25ec41513132

Observation 541dd781-ed8a-4eba-8946-5c6b248cb826 · inbound

Instella-T2I: Pushing the Limits of 1D Discrete Latent Space Image Generation cites this paper.

Instella-T2I: Pushing the Limits of 1D Discrete Latent Space Image Generation Fluid: Scaling Autoregressive Text-to-image Generative Models with Continuous Tokens

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-06T22:43:02.839383Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:43:02.839383Z digest=sha256:38f7d3789a5771976d8af4083b94be11b37bc6346c90ad305775dcd2dbcffa79

Observation 932ffd1b-bf41-46de-a8ed-deb8e07cf675 · inbound

Transition Matching: Scalable and Flexible Generative Modeling cites this paper.

Transition Matching: Scalable and Flexible Generative Modeling Fluid: Scaling Autoregressive Text-to-image Generative Models with Continuous Tokens

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T21:46:26.122915Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T21:46:26.122915Z digest=sha256:3fdd289c3bcfb230cc626ba733ba70df2b0341e52ffba3f26f3a5a82d9174b59

Observation 84ddbe34-a503-4eb4-a725-9b3fdef2f40f · inbound

Rethinking Discrete Tokens: Treating Them as Conditions for Continuous Autoregressive Image Synthesis cites this paper.

Rethinking Discrete Tokens: Treating Them as Conditions for Continuous Autoregressive Image Synthesis Fluid: Scaling Autoregressive Text-to-image Generative Models with Continuous Tokens

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T20:49:40.829577Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:49:40.829577Z digest=sha256:4f6d973e31bde91b5f64e12e5e8aa7c3b2bc4e74a9c7d5c2ab2c517629531fc1

Observation ff003569-80a0-4a66-a324-9c1a05f1c1a9 · inbound

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer cites this paper.

DC-AR: Efficient Masked Autoregressive Image Generation with Deep Compression Hybrid Tokenizer Fluid: Scaling Autoregressive Text-to-image Generative Models with Continuous Tokens

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T19:41:23.673806Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:41:23.673806Z digest=sha256:a8f310abf50efa7cae11a2375af40686fb51c4e0026ea90b6ad6c4bc21a06ec4

Observation 5c1fca8e-0408-4be2-ad98-1f1751b53db4 · inbound

PixNerd: Pixel Neural Field Diffusion cites this paper.

PixNerd: Pixel Neural Field Diffusion Fluid: Scaling Autoregressive Text-to-image Generative Models with Continuous Tokens

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-06T10:59:53.912933Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:59:53.912933Z digest=sha256:522c518f357a99875e1148757a6744eefa44e6686eafe832df63fb38c5606d1a

Observation 851a5af6-4b4e-4604-a561-73280dab8e37 · inbound

HPSv3: Towards Wide-Spectrum Human Preference Score cites this paper.

HPSv3: Towards Wide-Spectrum Human Preference Score Fluid: Scaling Autoregressive Text-to-image Generative Models with Continuous Tokens

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T04:23:05.274006Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T04:23:05.274006Z digest=sha256:99a434ae833a46efbde74da1eba1c223184ab8670b8d16b9737679d8eb5fe3f7

Observation a8946267-9d06-4a8e-b0ad-f3ab5b98770d · inbound

A Unified Low-level Foundation Model for Enhancing Pathology Image Quality cites this paper.

A Unified Low-level Foundation Model for Enhancing Pathology Image Quality Fluid: Scaling Autoregressive Text-to-image Generative Models with Continuous Tokens

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-05T13:00:48.199003Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T13:00:48.199003Z digest=sha256:3f0ac965e7a65d4cde97de5ff19fed2ab55f183afdb92fd1dfb246bd826711a8

Observation 1b3e00b6-e220-40c5-ae56-9e0de3846168 · inbound

DeCo: Frequency-Decoupled Pixel Diffusion for End-to-End Image Generation cites this paper.

DeCo: Frequency-Decoupled Pixel Diffusion for End-to-End Image Generation Fluid: Scaling Autoregressive Text-to-image Generative Models with Continuous Tokens

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-17T05:49:08.339073Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-17T05:47:24.669763Z digest=sha256:508bc9b77e38766fe39caf1872b9f73ac04e8a6e4bf53fc1f6c4a2470e8fd3ad

Observation 9bf1fc10-c7fb-4d52-9cb6-3a634dc38829 · inbound

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation cites this paper.

MoLingo: Motion-Language Alignment for Text-to-Human Motion Generation Fluid: Scaling Autoregressive Text-to-image Generative Models with Continuous Tokens

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-03T16:24:05.477812Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:24:05.477812Z digest=sha256:b4dad74405683629e8b6c3929a0f6b2a34ffe512885fe2ee876eca52029cc982

Observation ec49985c-acc0-4073-afdb-9ebdce34074e · inbound

HD-Prot: A Protein Language Model for Joint Sequence-Structure Modeling with Continuous Structure Tokens cites this paper.

HD-Prot: A Protein Language Model for Joint Sequence-Structure Modeling with Continuous Structure Tokens Fluid: Scaling Autoregressive Text-to-image Generative Models with Continuous Tokens

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-03T16:07:37.341677Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T16:07:37.341677Z digest=sha256:a29557512d12be70606df30b24cab7852c36b91ea1691d94acf37f62b6829ab9

Observation d05e4099-43ae-4b75-826a-cd860a4c3de4 · inbound

PixelGen: Improving Pixel Diffusion with Perceptual Supervision cites this paper.

PixelGen: Improving Pixel Diffusion with Perceptual Supervision Fluid: Scaling Autoregressive Text-to-image Generative Models with Continuous Tokens

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-16T07:57:33.193201Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-16T07:54:20.712620Z digest=sha256:c51066ac21b85ef0ca80b230f503b3307da49526f304cdf7dd482a5903c6afa6

Observation f57bbbab-bf3f-4535-b961-01b71b900a6f · inbound

Unifying Contrastive and Generative Objectives for Visual Understanding and Text-to-Image Generation cites this paper.

Unifying Contrastive and Generative Objectives for Visual Understanding and Text-to-Image Generation Fluid: Scaling Autoregressive Text-to-image Generative Models with Continuous Tokens

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-21T12:20:07.015956Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-21T12:15:57.019720Z digest=sha256:3c9d7a7b3845d24432910629c6f77ab6abb028ca606a9396fbf48d1730079960

Observation 4538c610-29d2-4152-8be7-cf2866c9319d · inbound

RAE-NWM: Navigation World Model in Dense Visual Representation Space cites this paper.

RAE-NWM: Navigation World Model in Dense Visual Representation Space Fluid: Scaling Autoregressive Text-to-image Generative Models with Continuous Tokens

Reference 16

Resolution
unresolved
no resolver link, observed 2026-07-15T12:07:05.640150Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-15T12:07:05.640150Z digest=sha256:f1b711d1508fc651f0ed87300288c8a10c74da4991d24dd3967aa303cb925ac9

Observation 04e45abb-1cc2-402a-8f6f-96bb122ef357 · inbound

When Models Learn to Ask Why: Adaptive Causal Reasoning for Trustworthy Medical Vision-Language Models cites this paper.

When Models Learn to Ask Why: Adaptive Causal Reasoning for Trustworthy Medical Vision-Language Models Fluid: Scaling Autoregressive Text-to-image Generative Models with Continuous Tokens

Reference 12

Resolution
unresolved
no resolver link, observed 2026-07-13T19:53:46.276873Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T19:53:46.276873Z digest=sha256:0215d670b9ad6d4446d045c798c087ffc223bf8e8d3da3c5d4ce30cbf6cfd369

Observation 2dbfb9ad-9a81-49a3-9ea7-8ae23ea5e9ef · inbound

From Broad Exploration to Stable Synthesis: Entropy-Guided Optimization for Autoregressive Image Generation cites this paper.

From Broad Exploration to Stable Synthesis: Entropy-Guided Optimization for Autoregressive Image Generation Fluid: Scaling Autoregressive Text-to-image Generative Models with Continuous Tokens

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-15T12:50:37.106776Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-15T12:50:13.764159Z digest=sha256:a96a8f96ca500b674c69e8404514ef008a56a1c36b048d477cabd2b1cd1abbb8

Observation f40448c6-b527-4eba-a856-d2b988a48e72 · inbound

MAR-GRPO: Stabilized GRPO for AR-diffusion Hybrid Image Generation cites this paper.

MAR-GRPO: Stabilized GRPO for AR-diffusion Hybrid Image Generation Fluid: Scaling Autoregressive Text-to-image Generative Models with Continuous Tokens

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-11T05:40:58.218493Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T18:00:50.105629Z digest=sha256:84447d9f6ca2213c7c47ceb28670a418a16044109b5b3ad833d284c0f96b935f