Pith. sign in

Paper Citation Record · LEDGER

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities

As of 8 August 2026, this Paper Citation Record lists 100 of 178 outbound references and 11 inbound Pith citation observations for arXiv:2505.20147.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.20147 v3

Coverage vector

measured 100 of 178 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:05:03.315466Z

measured 111 of 111 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 11 of 11 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T20:15:21.721749Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-05T11:41:02.728596Z

Reference resolution

100 of 178 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved100
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c1a2050a-50dc-4922-9ecc-560b97188b1b · outbound

This paper cites Deepseek-r1: Incentivizing reasoning capability in llms via reinforcement learning, 2025.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Deepseek-r1: Incentivizing reasoning capability in llms via reinforcement learning, 2025

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T14:04:52.894142Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:04:52.894142Z digest=sha256:e0b835362611723aeef2587724624d496a09c929c87b9427c6a99f9e49dc7e72

Observation d06f3e0a-b160-4abf-ad26-e98df6c9265a · outbound

This paper cites Qwen2.5 Technical Report.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Qwen2.5 Technical Report

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T14:04:53.001703Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:04:53.001703Z digest=sha256:b5e31d7cf21e677bbdd7725d870eac87b4f6e3391c66f6f393f9c226f378bb62

Observation 9766cfe1-7c43-40a3-adee-c7e8565673da · outbound

This paper cites The Llama 3 Herd of Models.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities The Llama 3 Herd of Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T14:04:53.087151Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:04:53.087151Z digest=sha256:ab7ce48f8fbd49b53e0981e82e36061ecb17f0646aeb5c921a0fbdf748d7c1c3

Observation 557b361b-e49e-4301-af2a-f4f445f6ada3 · outbound

This paper cites Internlm2 technical report, 2024.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Internlm2 technical report, 2024

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T14:04:53.189524Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:04:53.189524Z digest=sha256:37fdf1a8bc2678b0132eace5c4c3a2b18d081964bca580513a81017335692673

Observation 893607b1-69e4-4efa-af02-4b5edf3eb4ae · outbound

This paper cites an unresolved cited work.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Unresolved cited work

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T14:04:53.273869Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:04:53.273869Z digest=sha256:08204445ee363a92aea74d564a0c5ad326229040382506e4d13cbd07d1b65fb1

Observation 97d2b2f3-d6b7-4f33-a2c4-6fd32cfd1999 · outbound

This paper cites Visual instruction tuning.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Visual instruction tuning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T14:04:53.355130Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:04:53.355130Z digest=sha256:2dc23d09773fd4092abf9a94327bd72c5e1689d9bb33b5d8bc87cfc47f713789

Observation db23d69b-55c2-44be-abb1-0feb49777ee3 · outbound

This paper cites MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T14:04:53.432740Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:04:53.432740Z digest=sha256:832ff85f7ca8a5674abf7d2de7db742a3719dd85fb7f84ef6f71e65385cb3c03

Observation 454ddda6-a733-44f4-b480-3e1225dfb1da · outbound

This paper cites Instructblip: Towards general-purpose vision-language models with instruction tuning, 2023.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Instructblip: Towards general-purpose vision-language models with instruction tuning, 2023

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T14:04:53.479004Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:04:53.479004Z digest=sha256:1dc0c2993b9eebe126f508df39406b11ca034fd55c759a58e1937c5361b39191

Observation d01f4598-f6a0-4977-a4f7-101b2ba3b4b4 · outbound

This paper cites DeepSeek-VL: Towards Real-World Vision-Language Understanding.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities DeepSeek-VL: Towards Real-World Vision-Language Understanding

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T14:04:53.530572Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:04:53.530572Z digest=sha256:f65331690fc13aea81d31ff7d090f08908cacc4f755ef222676368ffcd685966

Observation 4683a615-9f8c-4fd8-b3f7-400d37f7e149 · outbound

This paper cites How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T14:04:53.598580Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:04:53.598580Z digest=sha256:416a3d44e952127a7be35ed4e8f57aff601d7dfd2a14eaf6a777fc04fcb55ca9

Observation 2323d415-217c-4847-bfb0-ef81d7169882 · outbound

This paper cites Diffusion models beat gans on image synthesis.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Diffusion models beat gans on image synthesis

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T14:04:53.662581Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:04:53.662581Z digest=sha256:d3ff288d8964de55c747cb10169463fa8804eae0e08f33a4822ccdbef4ffe006

Observation 636b6e21-ad95-4f6d-b3a0-a963a6adae44 · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities High-resolution image synthesis with latent diffusion models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T14:04:53.750305Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:04:53.750305Z digest=sha256:e2aec0e9bf4fa82d4bd8c275cc162e4809360d03d37b344dfba39d76bbf64cd9

Observation eb910a6c-253f-4a62-b77b-d680bc3cd313 · outbound

This paper cites PixArt-$\alpha$: Fast Training of Diffusion Transformer for Photorealistic Text-to-Image Synthesis.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities PixArt-$\alpha$: Fast Training of Diffusion Transformer for Photorealistic Text-to-Image Synthesis

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T14:04:53.892795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:04:53.892795Z digest=sha256:105957da296eff5fa7a9aa418438fc0fc025636a131309a94fb04e194bc0ad2c

Observation a5e348ca-5944-4461-adfa-7f42f8129d5e · outbound

This paper cites Scaling Rectified Flow Transformers for High-Resolution Image Synthesis.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Scaling Rectified Flow Transformers for High-Resolution Image Synthesis

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T14:04:53.994314Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:04:53.994314Z digest=sha256:e6a8ee8f94f8d2cc970ed2cf973c8e07c7a22fa860d61ba533594bf4c2dbc63b

Observation 94390c54-1090-4fc6-9be6-6c9c48ae45db · outbound

This paper cites an unresolved cited work.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Unresolved cited work

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T14:04:54.051129Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:04:54.051129Z digest=sha256:50ad85e4e6a40c3603178eafba1726c16ab2528700b6e77acac35e5a57dad06e

Observation 9121c834-8d28-4554-bafa-d53507f3dc9c · outbound

This paper cites Planting a SEED of Vision in Large Language Model.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Planting a SEED of Vision in Large Language Model

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T14:04:54.247781Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:04:54.247781Z digest=sha256:72c590fd9e3880d89a14b86d89358859541d6f1a20100bf60dba2c9de04d9aae

Observation 11016a83-596a-4a6e-bee4-aad63d187c44 · outbound

This paper cites Making LLaMA SEE and Draw with SEED Tokenizer.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Making LLaMA SEE and Draw with SEED Tokenizer

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T14:04:54.366169Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:04:54.366169Z digest=sha256:77a8a6fd531b7bc9e9ca857cd2b2566b7e6eb25e2246404940438ed85aca6379

Observation b821a411-f2e3-4a95-b6c6-3e93d6704e58 · outbound

This paper cites Emu3: Next-token prediction is all you need, 2024.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Emu3: Next-token prediction is all you need, 2024

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T14:04:54.505810Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:04:54.505810Z digest=sha256:0d0452448de03d294e0a2f61d2b1a5cdcbcf0afd144e4bbb3373bae138e3c436

Observation 7cc05029-eff6-4938-9724-845eb0cc0041 · outbound

This paper cites Transfusion: Predict the Next Token and Diffuse Images with One Multi-Modal Model.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Transfusion: Predict the Next Token and Diffuse Images with One Multi-Modal Model

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T14:04:54.628978Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:04:54.628978Z digest=sha256:ed18b11887666c5f5296d2173694e85bb932679f6b0aa5ed9deeeae3af93f94d

Observation d171d038-a94b-4373-b222-57963cb0f648 · outbound

This paper cites Janus: Decoupling Visual Encoding for Unified Multimodal Understanding and Generation.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Janus: Decoupling Visual Encoding for Unified Multimodal Understanding and Generation

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T14:04:54.764534Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:04:54.764534Z digest=sha256:45ccd05af82999cb0091bc17896317e46570d2207eed4f03b66deff47a26dc7e

Observation 3f887847-1b4b-4e0d-a8e0-b141537f80dc · outbound

This paper cites Illume: Illuminating your llms to see, draw, and self-enhance, 2024.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Illume: Illuminating your llms to see, draw, and self-enhance, 2024

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T14:04:54.940447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:04:54.940447Z digest=sha256:d393c8b4b852d5267eac9ded1e3334e0b90b98d7604468aca4d96ed1f5019afb

Observation 72b3daa5-cb80-40be-8d0d-595107375aa1 · outbound

This paper cites Janus-Pro: Unified Multimodal Understanding and Generation with Data and Model Scaling.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Janus-Pro: Unified Multimodal Understanding and Generation with Data and Model Scaling

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T14:04:55.018343Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:04:55.018343Z digest=sha256:59c3446c3fd080ccc46b35affd3307cefa0adaea0ddfff3a45a5b1d8efdd9bc8

Observation 95f3b668-cc3d-46ab-a44f-4ef7ac551063 · outbound

This paper cites Illume+: Illuminating unified mllm with dual visual tokenization and diffusion refinement, 2025.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Illume+: Illuminating unified mllm with dual visual tokenization and diffusion refinement, 2025

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T14:04:55.066665Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:04:55.066665Z digest=sha256:50f71d3dfde3b34028f4786c7d5abafe730ab79c1e6c3a140a7f375bc7bf982a

Observation 9d79ce17-4fd2-49e5-9ba8-787a0ef4d3a4 · outbound

This paper cites Sparks of artificial general intelligence: Early experiments with gpt-4, 2023.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Sparks of artificial general intelligence: Early experiments with gpt-4, 2023

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T14:04:55.161855Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:04:55.161855Z digest=sha256:56df505b6d857df77d2ee50fb870bd7fb457d33a7b3b53c8fb77d0eb62e5a4bf

Observation 118b17c7-1120-408c-bc87-39974a88512b · outbound

This paper cites Hwang, Soumya Sanyal, Sean Welleck, Xiang Ren, Allyson Ettinger, Zaid Harchaoui, and Yejin Choi.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Hwang, Soumya Sanyal, Sean Welleck, Xiang Ren, Allyson Ettinger, Zaid Harchaoui, and Yejin Choi

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T14:04:55.199306Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:04:55.199306Z digest=sha256:fa8a0f8acaadee38b49bbc39a5e0d37f55ac7ad31690e750628daccae1fc4f5f

Observation 29109999-bdad-4cb0-9c0b-2d873aa2b929 · outbound

This paper cites The pitfalls of next-token prediction.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities The pitfalls of next-token prediction

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T14:04:55.284361Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:04:55.284361Z digest=sha256:5ceb73b3896cb0b37e1d5be562900f75e951ddfa30ae9831cf6034ccf0f0dafa

Observation ee6bcc26-c9a3-42c4-8471-0480ce1f13f1 · outbound

This paper cites Beyond Autoregression: Discrete Diffusion for Complex Reasoning and Planning.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Beyond Autoregression: Discrete Diffusion for Complex Reasoning and Planning

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T14:04:55.361338Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:04:55.361338Z digest=sha256:2c41813356532fa19e24e4ff1248b6265854a65f954828057c70bc2ede4cc6a3

Observation 8f803c8a-825f-417a-a536-21852e44ef64 · outbound

This paper cites Large Language Models Cannot Self-Correct Reasoning Yet.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Large Language Models Cannot Self-Correct Reasoning Yet

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T14:04:55.454519Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:04:55.454519Z digest=sha256:5d92e7dcace0d435233dcba2e7f54b4c4fd6bbcbc00c67e93747259504beb65b

Observation 7bce1405-d86a-47cb-8cc8-faa0a7b09366 · outbound

This paper cites Structured denoising diffusion models in discrete state-spaces.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Structured denoising diffusion models in discrete state-spaces

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T14:04:55.571239Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:04:55.571239Z digest=sha256:dfbf95480f18950ab1dcfe4441d24776be32bc361efd8e0493f81dbba6caad7c

Observation d5350019-08ef-4869-9a03-a701ae7d2443 · outbound

This paper cites Discrete diffusion modeling by estimating the ratios of the data distribution.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Discrete diffusion modeling by estimating the ratios of the data distribution

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T14:04:55.654115Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:04:55.654115Z digest=sha256:3d201b8106e3aba2313b5b6b914b019473a63c9764d9de20b74a151e25c3bbd9

Observation 515895c9-5225-4d24-b22e-51667697f591 · outbound

This paper cites Simplified and generalized masked diffusion for discrete data.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Simplified and generalized masked diffusion for discrete data

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T14:04:55.747086Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:04:55.747086Z digest=sha256:0929d6a6d8d63527f2c45c1ccc945b1cf08ad60d82980c0cc99c67cf29084c10

Observation e3f3d6bc-b168-4c2d-89c7-df2311791e2e · outbound

This paper cites Simple and effective masked diffusion language models.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Simple and effective masked diffusion language models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T14:04:55.841729Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:04:55.841729Z digest=sha256:0d31af0ce32e00b0ac52a163c94efec78b6ea30e53ce1fe6816bb170776b671b

Observation b8adb77b-f844-4fa8-b17f-531e7296e865 · outbound

This paper cites Discrete flow matching.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Discrete flow matching

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T14:04:55.936733Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:04:55.936733Z digest=sha256:c65ab3087f4a73d13aa256a160db56b54ebbec61b3079b4f306692f0e3b9698d

Observation 259dca27-6153-4b04-b29e-d90de1fe47f8 · outbound

This paper cites an unresolved cited work.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Unresolved cited work

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T14:04:56.033731Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:04:56.033731Z digest=sha256:5e2e54872ed02ad816367c6dc801bce7fc3a55c25e622b91fe8098b94ca6b7c1

Observation c4a8de79-0f51-4e97-bb7c-6b56de394639 · outbound

This paper cites Generative flows on discrete state-spaces: Enabling multimodal flows with applications to protein co-design.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Generative flows on discrete state-spaces: Enabling multimodal flows with applications to protein co-design

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T14:04:56.155859Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:04:56.155859Z digest=sha256:025c8ecca6a08cd90b2b7b309b5dbe82cb8f5b97120bbf0a64dc208ae45e4396

Observation a5e545d2-a5d0-4150-ad5f-c91395878700 · outbound

This paper cites URL https://www.inceptionlabs.ai/news.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities URL https://www.inceptionlabs.ai/news

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T14:04:56.267641Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:04:56.267641Z digest=sha256:353edaf59ff16fe9227204c0a8a53c62c7951ccd8abf79c2e924a60e26bd0771

Observation d3252845-391e-4ed8-8564-25bebc7bed1b · outbound

This paper cites Flow Straight and Fast: Learning to Generate and Transfer Data with Rectified Flow.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Flow Straight and Fast: Learning to Generate and Transfer Data with Rectified Flow

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T14:04:56.402891Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:04:56.402891Z digest=sha256:b4e8956f76b8b16a56fa33f05c4b16605b4c16275e36b9101faf1aa4b00245fd

Observation 2eaeb6cb-6b90-484e-a368-ec39be947865 · outbound

This paper cites Flow Matching for Generative Modeling.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Flow Matching for Generative Modeling

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T14:04:56.567632Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:04:56.567632Z digest=sha256:24b18cdf3af2b197d456e9c1f511cdc92544d1006f37e11de1f4fe279330c4d3

Observation 516f93ed-57ba-40e8-aa40-443d0c92fff2 · outbound

This paper cites Consistency models.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Consistency models

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T14:04:56.668735Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:04:56.668735Z digest=sha256:bea0b767ad78488c854e1edb91235ee959c06ffd72add0098c4541c0bd1c3e86

Observation d3c0dfc9-9034-4a0a-8f65-5776a66aa78a · outbound

This paper cites Large Language Diffusion Models.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Large Language Diffusion Models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T14:04:56.761698Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:04:56.761698Z digest=sha256:0eab4a1d604a1775e3dc03f1153882a91aa1ea0de56f62db22b2394a6872aa57

Observation 631d0bb4-be91-4b6b-b597-8ddbb990a2b4 · outbound

This paper cites Dream 7b, 2025.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Dream 7b, 2025

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T14:04:56.821080Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:04:56.821080Z digest=sha256:7513dc649f214f4f1a663144533562278c0afb7b8f522b236be010f013e9c45c

Observation 45345e19-f584-4399-ae16-ff1d85fe254c · outbound

This paper cites Dual Diffusion for Unified Image Generation and Understanding.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Dual Diffusion for Unified Image Generation and Understanding

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T14:04:56.901976Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:04:56.901976Z digest=sha256:203bf42d57de522fb1cc71f0652a76520296d38b31d33feaa9145ffc59682e24

Observation 24d9e19b-6303-413a-bbf3-646a55d82b0a · outbound

This paper cites Unified Discrete Diffusion for Simultaneous Vision-Language Generation.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Unified Discrete Diffusion for Simultaneous Vision-Language Generation

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T14:04:56.951731Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:04:56.951731Z digest=sha256:5f2e9a703e3d64e7bf357fbd9ac74b004e414a4d2d4a39586b9ac220a81ef563

Observation 8a7ddd00-d69d-4127-bcbd-7de3b7e2b9ec · outbound

This paper cites Unified Multimodal Discrete Diffusion.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Unified Multimodal Discrete Diffusion

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T14:04:57.037941Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:04:57.037941Z digest=sha256:af8da396461d7287056a02495bc15bb85ff6581f63abb01a23cee33669f2f232

Observation 4b174cf9-7e0e-4e02-ab50-dd8c81e95aec · outbound

This paper cites Scaling Diffusion Language Models via Adaptation from Autoregressive Models.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Scaling Diffusion Language Models via Adaptation from Autoregressive Models

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T14:04:57.107795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:04:57.107795Z digest=sha256:532bd17c9d703c06e0661a2f6e558d0aded75d4709100774dae05ae0d0c7458c

Observation 5bd7432a-5bd7-4cc2-9a71-2277744ba661 · outbound

This paper cites SANA 1.5: Efficient Scaling of Training-Time and Inference-Time Compute in Linear Diffusion Transformer.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities SANA 1.5: Efficient Scaling of Training-Time and Inference-Time Compute in Linear Diffusion Transformer

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T14:04:57.190356Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:04:57.190356Z digest=sha256:677b6b7e110b81bcfdc0e56180553aedcca90a1aa220382919364be6b3db8e47

Observation 8a287536-95fb-43e8-b179-b55524e88c6c · outbound

This paper cites Dancegrpo: Unleashing grpo on visual generation, 2025.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Dancegrpo: Unleashing grpo on visual generation, 2025

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T14:04:57.273486Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:04:57.273486Z digest=sha256:a0b2716bb4373d663d58042aae7be6c97bfb3256d20c2e6356871ec607626f06

Observation 4daf39c0-2bfa-401f-8cad-40101393906b · outbound

This paper cites ILLUME+: Illuminating Unified MLLM with Dual Visual Tokenization and Diffusion Refinement.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities ILLUME+: Illuminating Unified MLLM with Dual Visual Tokenization and Diffusion Refinement

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T14:04:57.327788Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:04:57.327788Z digest=sha256:93d65523b9a0cb189242508ba90b02bdcef4321e8fa1f5423c39a17ed52bbdcb

Observation be3146dd-8cf7-459a-a608-98c4b78936f1 · outbound

This paper cites VILA-u: a unified foundation model integrating visual understanding and generation.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities VILA-u: a unified foundation model integrating visual understanding and generation

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T14:04:57.411925Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:04:57.411925Z digest=sha256:46064899e5d5a543f23f646e6041b7a9439c8714fc94c95bba839cced5127fc3

Observation be7d454e-1ea6-4802-8b50-dfcf7255c269 · outbound

This paper cites Chameleon: Mixed-Modal Early-Fusion Foundation Models.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Chameleon: Mixed-Modal Early-Fusion Foundation Models

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T14:04:57.497040Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:04:57.497040Z digest=sha256:7d710d431897cee52b618272fc2f722f03a290fbac3da0a3b02f03ec1e758c9c

Observation 7ee0f59a-7eec-43d0-87b3-82e6d30bf694 · outbound

This paper cites Unified language-vision pretraining in LLM with dynamic discrete visual tokenization.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Unified language-vision pretraining in LLM with dynamic discrete visual tokenization

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T14:04:57.555386Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:04:57.555386Z digest=sha256:e1bfeb5ae67e52aea3d30804175a8390716306ef5b4f99c2aeaee8d84a279331

Observation dbce4436-7354-453b-a11b-914505f55b45 · outbound

This paper cites Show-o: One Single Transformer to Unify Multimodal Understanding and Generation.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Show-o: One Single Transformer to Unify Multimodal Understanding and Generation

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T14:04:57.633027Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:04:57.633027Z digest=sha256:942c1b64a86e7d1198012c9c52ffd54b54bc94af26e8855cf3e4477acdc8f4ea

Observation 9e986136-9cde-44b3-bd4b-f25bb51d660f · outbound

This paper cites Denoising diffusion probabilistic models.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Denoising diffusion probabilistic models

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T14:04:57.691101Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:04:57.691101Z digest=sha256:a01de6c7402b39219710a46153ad09dcc2e44295eea44c99b24fe4b48239457a

Observation 33a9a937-9203-4e14-96c4-1b152633a63e · outbound

This paper cites Diffusionbert: Improving generative masked language models with diffusion models.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Diffusionbert: Improving generative masked language models with diffusion models

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-07T14:04:57.761658Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:04:57.761658Z digest=sha256:8c8f9097491c802519fee6917edd2823a2fcead09ecb69a64c8f04ebee40a689

Observation abfef5d2-d562-4075-8b8d-35165a4bc48c · outbound

This paper cites Sigmoid loss for language image pre-training.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Sigmoid loss for language image pre-training

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T14:04:57.843795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:04:57.843795Z digest=sha256:76aa95bd49dd845ea20f56e6425664d7dca42413459f3368a9527672894b1d55

Observation 4116490a-c833-4136-9af7-14e168160f06 · outbound

This paper cites Autoregressive Model Beats Diffusion: Llama for Scalable Image Generation.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Autoregressive Model Beats Diffusion: Llama for Scalable Image Generation

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-07T14:04:57.901384Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:04:57.901384Z digest=sha256:84d9b38ebd33a9cf4a6c3cc4bef1cb5bb7dedeea7da56535561f2579f991ce7b

Observation b6e354e6-b377-4bf1-90e4-2954bd1181c7 · outbound

This paper cites LLaVAR: Enhanced Visual Instruction Tuning for Text-Rich Image Understanding.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities LLaVAR: Enhanced Visual Instruction Tuning for Text-Rich Image Understanding

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-07T14:04:57.989894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:04:57.989894Z digest=sha256:aecf7db498d2235e4a527bd65e3fc44b3802c58fe28b0502c71746dcdb62fc49

Observation 2bf5523b-33b5-4c4f-93a7-d24eb14b9c9a · outbound

This paper cites wendlerc/renderedtext, 2023.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities wendlerc/renderedtext, 2023

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-07T14:04:57.994253Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:04:57.994253Z digest=sha256:8d61f207278d495fc157bfdadedfb213a249e3cd0842c1ae1bc8658d7fc04f97

Observation 180e9c8d-04e6-4591-93f4-45afe9049e9a · outbound

This paper cites Docvqa: A dataset for vqa on document images.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Docvqa: A dataset for vqa on document images

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-07T14:04:58.065592Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:04:58.065592Z digest=sha256:9d6387627cb4c3ba6b6568a5db967fcf61bf81bd7f097af2c7ca88dbcf56ed97

Observation c0efe589-9f20-4eb6-86dd-88ff569f833d · outbound

This paper cites Chart-to-text: Generating natural language descriptions for charts by adapting the transformer model, 2020.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Chart-to-text: Generating natural language descriptions for charts by adapting the transformer model, 2020

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-07T14:04:58.133657Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:04:58.133657Z digest=sha256:1e857eaea86c07b4842913e32b5d20ec61ce3d16e6b73030d6e242a3bbe7b63b

Observation e1206549-5bc3-4261-9251-af4a896f38c4 · outbound

This paper cites Visualmrc: Machine reading compre- hension on document images.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Visualmrc: Machine reading compre- hension on document images

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-07T14:04:58.242162Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:04:58.242162Z digest=sha256:2d3882fb44657460187108cd084267331ad4a295f2ddfd4d31d279c46077abb3

Observation 946e16cf-21bb-4a77-93bc-530a43a3a1e2 · outbound

This paper cites G-llava: Solving geomet- ric problem with multi-modal large language model, 2023.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities G-llava: Solving geomet- ric problem with multi-modal large language model, 2023

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-07T14:04:58.330589Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:04:58.330589Z digest=sha256:c6736f1e59aa119011795116a0161d3150d6533be9065c2a7a1427de4c8be7c3

Observation 9fce5abe-f0ee-483b-bdcf-af20234e15a2 · outbound

This paper cites Xing, and Liang Lin.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Xing, and Liang Lin

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-07T14:04:58.422124Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:04:58.422124Z digest=sha256:5c37b5e21efa3fff0a4c30146cb8672058cd9e3008eefcb7213fa33289f894f1

Observation 356ccaf6-4178-44f5-b634-a9a3c7375158 · outbound

This paper cites Magpie: Alignment Data Synthesis from Scratch by Prompting Aligned LLMs with Nothing.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Magpie: Alignment Data Synthesis from Scratch by Prompting Aligned LLMs with Nothing

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-07T14:04:58.508179Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:04:58.508179Z digest=sha256:1bf052b106118b5278fd6f4c68e5006048169473637f7013d657fa6bb07cd35f

Observation 608d9593-a179-42da-b8e3-f89ecada8317 · outbound

This paper cites World model on million-length video and language with blockwise ringattention.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities World model on million-length video and language with blockwise ringattention

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-07T14:04:58.598819Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:04:58.598819Z digest=sha256:06924de996978a1ecf79174b5ada540a935f54a77826e85497dc7b4812bd112d

Observation e91cac6a-9d61-4be5-8337-dcd7b674c884 · outbound

This paper cites Hierarchical Text-Conditional Image Generation with CLIP Latents.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Hierarchical Text-Conditional Image Generation with CLIP Latents

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-07T14:04:58.711136Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:04:58.711136Z digest=sha256:421c3a82fc6da60b00be9f83bce86dec3b5e2e2cb568e439ab69d40895375085

Observation d72b7cb6-759b-4c6e-8e72-dab9c8f573be · outbound

This paper cites SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-07T14:04:58.885542Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:04:58.885542Z digest=sha256:f7ba01cf6f6f5236949bfc7745f58f646e572b56ef4c8274cb199fb7535b217a

Observation 639ed477-c027-466e-843e-aecd9275f5de · outbound

This paper cites Improving image generation with better captions.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Improving image generation with better captions

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-07T14:04:59.052307Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:04:59.052307Z digest=sha256:9b56811463ecf849cb52b69de26f478537de725edb12a9446c42d632a38207ea

Observation 51ebc488-ffbf-4871-bacd-0cb52cff561e · outbound

This paper cites SEED-X: Multimodal Models with Unified Multi-granularity Comprehension and Generation.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities SEED-X: Multimodal Models with Unified Multi-granularity Comprehension and Generation

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-07T14:04:59.164314Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:04:59.164314Z digest=sha256:64200bb45f1386d56b0d272fd90dd1357e8e430e5ccf78eb9a39046fc0a29c5c

Observation d9455675-b83a-4d95-beea-188e4e763b43 · outbound

This paper cites TokenFlow: Unified Image Tokenizer for Multimodal Understanding and Generation.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities TokenFlow: Unified Image Tokenizer for Multimodal Understanding and Generation

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-07T14:04:59.258830Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:04:59.258830Z digest=sha256:7d2344e462f558d5bc7dac939b67bf9e4ef562dc6965d74b41b69523a794dd09

Observation 0ee9b758-72d2-4472-a2d9-593715f9512e · outbound

This paper cites Geneval: An object-focused framework for evaluating text-to-image alignment.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Geneval: An object-focused framework for evaluating text-to-image alignment

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-07T14:04:59.365597Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:04:59.365597Z digest=sha256:221d476332a670a0dd04d3a61b5d280f8d82c17a267fd5deae11c210220098ec

Observation 0205f82c-f161-429e-9c93-371945102ce1 · outbound

This paper cites MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-07T14:04:59.529591Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:04:59.529591Z digest=sha256:6a806dc7a4ab7b637650c185e07402d0514ee4b0b15e7e800be3b655e5f60f4a

Observation ea734a38-7f00-41dd-9e73-0c3226fbc66c · outbound

This paper cites MobileVLM V2: Faster and Stronger Baseline for Vision Language Model.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities MobileVLM V2: Faster and Stronger Baseline for Vision Language Model

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-07T14:04:59.650981Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:04:59.650981Z digest=sha256:ae3d8e2d8f8ee8483a9ef697fa9e9232dce91895d91c27a7883b10f3469f660a

Observation 62bdc963-0009-484b-8998-f6b6db471a92 · outbound

This paper cites LLaVA-Phi: Efficient Multi-Modal Assistant with Small Language Model.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities LLaVA-Phi: Efficient Multi-Modal Assistant with Small Language Model

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-07T14:04:59.785169Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:04:59.785169Z digest=sha256:da32e7b5f6ad2cb070cc164205a32fef7de2eb8908fc896a72582067d0ddd571

Observation e0b17c29-69e1-48d5-9d54-25a7d565b4bc · outbound

This paper cites Improved baselines with visual instruction tuning.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Improved baselines with visual instruction tuning

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-07T14:04:59.952940Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:04:59.952940Z digest=sha256:63da06817056e758ac4fa4c7c6825ed2e9d6ebdf33847d8a066bf92b9d07cc34

Observation 8ddb8461-d77c-4bfd-998f-58f3745bd411 · outbound

This paper cites Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-07T14:05:00.103130Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:05:00.103130Z digest=sha256:3096950ba252b3666dc44fb9877cc9088c61b9ebb3c5adefdba703e8a31d78f1

Observation 8057f8b4-a8d9-4940-b546-2217476ce832 · outbound

This paper cites Introducing idefics: An open reproduction of state-of-the-art visual language model, 2023.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Introducing idefics: An open reproduction of state-of-the-art visual language model, 2023

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-07T14:05:00.285717Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:05:00.285717Z digest=sha256:e58e5e9419ca0fe8623d754d7d4df50ce7f7efccdfe9096a8ecec4fdb348b543

Observation d7afa45c-ae24-4cd7-bc30-8283bc127c8c · outbound

This paper cites Unified language-vision pretraining with dynamic discrete visual tokenization.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Unified language-vision pretraining with dynamic discrete visual tokenization

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-07T14:05:00.396906Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:05:00.396906Z digest=sha256:ea5ca3215c6bca0c1091707e4792e2fd11755f14bad25aa2cc9cda95333a470f

Observation a288f023-8bb1-4e83-98a7-cc8f9b59bfb1 · outbound

This paper cites MetaMorph: Multimodal Understanding and Generation via Instruction Tuning.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities MetaMorph: Multimodal Understanding and Generation via Instruction Tuning

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-07T14:05:00.470954Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:05:00.470954Z digest=sha256:2b45d248e923e3590c7dd11cdc8cfc413327b5f691f99fb931a29e569debf086

Observation 27cc4a4f-dce6-46b7-a579-aead07dadab0 · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Gemini: A Family of Highly Capable Multimodal Models

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-07T14:05:00.549878Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:05:00.549878Z digest=sha256:ad5585c09a08cb77de148d8ebf13e50231c26ff210d2ece0f083ab6e9fe4ee47

Observation c70f29be-085f-4c1a-b320-9a229d9baeef · outbound

This paper cites VILA-U: a Unified Foundation Model Integrating Visual Understanding and Generation.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities VILA-U: a Unified Foundation Model Integrating Visual Understanding and Generation

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-07T14:05:00.625941Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:05:00.625941Z digest=sha256:b750e94315f3671f125c1de2324706843c25751ca02fd8bf363327cb6a992180

Observation 9cf071ae-eb3c-40a2-9a8d-13f5ba9565b5 · outbound

This paper cites Evaluating Object Hallucination in Large Vision-Language Models.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Evaluating Object Hallucination in Large Vision-Language Models

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-07T14:05:00.773460Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:05:00.773460Z digest=sha256:9229f79d6bb4baa7f2d7732666792318b8a8bbbf8000e146c67c131d3c666693

Observation e6a730c6-f49e-44b6-a039-9b4b7dd5f33c · outbound

This paper cites MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-07T14:05:00.924654Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:05:00.924654Z digest=sha256:16f8a5396c0a4aabf4046a26d4902b301495ba64c7ae68ac316cd88fc59945a1

Observation f9442855-c1f7-40af-bca1-0b4b52740018 · outbound

This paper cites SEED-Bench: Benchmarking Multimodal LLMs with Generative Comprehension.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities SEED-Bench: Benchmarking Multimodal LLMs with Generative Comprehension

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-07T14:05:01.070236Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:05:01.070236Z digest=sha256:7932a99e08043f81feb46f4aef494c690e9a8a37d9be79b2ad55f7b2ad313fe9

Observation f0db4cdd-7d8a-46fd-8aa0-322d844a3618 · outbound

This paper cites MMBench: Is Your Multi-modal Model an All-around Player?.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities MMBench: Is Your Multi-modal Model an All-around Player?

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-07T14:05:01.227959Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:05:01.227959Z digest=sha256:6e0e20868664f0aa6a467e567d914f3f34a0518fea218ffbac0d7ff9a671cd68

Observation f16fcff8-183a-48df-9fa0-0402126d77ec · outbound

This paper cites Gqa: A new dataset for real-world visual reasoning and compositional question answering.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Gqa: A new dataset for real-world visual reasoning and compositional question answering

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-07T14:05:01.400652Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:05:01.400652Z digest=sha256:0e8a6b45ee20fa2429974074c9eb271802d84a705bfab084250cd3f3dc36ab51

Observation cfb9a97b-0c24-421f-af97-1bc4efe5f708 · outbound

This paper cites Mmmu: A massive multi- discipline multimodal understanding and reasoning benchmark for expert agi.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Mmmu: A massive multi- discipline multimodal understanding and reasoning benchmark for expert agi

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-07T14:05:01.542154Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:05:01.542154Z digest=sha256:fd4e40023449c706a081adfa8e71e3f950bb05a917d1398ff61e1bffb7a9f3cc

Observation a717de63-b6ef-4029-b109-cebaf2e65ab5 · outbound

This paper cites MM-Vet: Evaluating Large Multimodal Models for Integrated Capabilities.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities MM-Vet: Evaluating Large Multimodal Models for Integrated Capabilities

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-07T14:05:01.611547Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:05:01.611547Z digest=sha256:15953260d85acfaf1bb1263ce481c9dd102ac711a951a61b10d7d32bcb8e4cec

Observation ef6700ac-f5f4-4846-b60a-0cefd3575d1f · outbound

This paper cites NVILA: Efficient Frontier Visual Language Models.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities NVILA: Efficient Frontier Visual Language Models

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-07T14:05:01.704476Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:05:01.704476Z digest=sha256:a4bf93c96160be30315f2e910160333cf966a2dedc6bb8d87da23c84d021b933

Observation c2a388f0-818e-46b1-a604-67ae8745ee15 · outbound

This paper cites Flow-GRPO: Training Flow Matching Models via Online RL.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Flow-GRPO: Training Flow Matching Models via Online RL

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-07T14:05:01.819382Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:05:01.819382Z digest=sha256:ff3beb6cce552af3d29119cbfb09576abb1fcc9a246f1353924ebc993df2d801

Observation 083b6baa-5d7f-478c-88d3-a32dcd1b8eb4 · outbound

This paper cites Unified-IO: A Unified Model for Vision, Language, and Multi-Modal Tasks.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Unified-IO: A Unified Model for Vision, Language, and Multi-Modal Tasks

Reference 91

Resolution
unresolved
no resolver link, observed 2026-08-07T14:05:01.991113Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:05:01.991113Z digest=sha256:c5d46e9d24c790d63ef4928a4b1178f00610c6ca563b4c1c90f8e61c13197118

Observation f2ee1038-e567-4fe8-8bd3-82bad39e832c · outbound

This paper cites Unified-io 2: Scaling autoregressive multimodal models with vision language audio and action.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Unified-io 2: Scaling autoregressive multimodal models with vision language audio and action

Reference 92

Resolution
unresolved
no resolver link, observed 2026-08-07T14:05:02.113485Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:05:02.113485Z digest=sha256:5f739c3b58381896500a587a8ae9c1994a5b10b72e5f0e49e7468249c02036e4

Observation ed131da7-b973-4265-9ad3-00dc48bac017 · outbound

This paper cites AnyGPT: Unified Multimodal LLM with Discrete Sequence Modeling.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities AnyGPT: Unified Multimodal LLM with Discrete Sequence Modeling

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-07T14:05:02.261823Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:05:02.261823Z digest=sha256:73f81ab9857c56b25d3affcead8409999d3887f2571c111759a8735ce811090f

Observation 73ba1ad5-474b-4b33-91bd-c6f6d0e70bbd · outbound

This paper cites Generative Multimodal Pretraining with Discrete Diffusion Timestep Tokens.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Generative Multimodal Pretraining with Discrete Diffusion Timestep Tokens

Reference 94

Resolution
unresolved
no resolver link, observed 2026-08-07T14:05:02.393418Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:05:02.393418Z digest=sha256:e57f4586e59988c268821b65bbc2c56acf5a8078d1517c4659f6c021a02d8594

Observation b913b799-b7fa-4f75-9a1e-4a153fe1ddaf · outbound

This paper cites Dreamllm: Synergistic multimodal comprehension and creation.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Dreamllm: Synergistic multimodal comprehension and creation

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-07T14:05:02.571307Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:05:02.571307Z digest=sha256:ea1cc97b430c86ef61019ed15c06286bd65f372b8157fa0d45b6afeb4f367d4f

Observation f46af32c-f3f7-41da-9a09-73517e82ff99 · outbound

This paper cites Minigpt-5: Interleaved vision-and-language generation via generative vokens.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Minigpt-5: Interleaved vision-and-language generation via generative vokens

Reference 96

Resolution
unresolved
no resolver link, observed 2026-08-07T14:05:02.723014Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:05:02.723014Z digest=sha256:32ce76b386b9da2b39f3b68f5787013fc9b38b7a4c95576ce0832fa5405d7a84

Observation de85dd05-1d00-430b-ba82-b0d600da042e · outbound

This paper cites Next-gpt: Any-to-any multimodal llm.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Next-gpt: Any-to-any multimodal llm

Reference 97

Resolution
unresolved
no resolver link, observed 2026-08-07T14:05:02.879601Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:05:02.879601Z digest=sha256:e256f926b3261bdc63dd21b37c5471d13e29d1e88158984dfb415dd74e111d92

Observation 32ce7d75-fa9f-401b-803d-6fee32024c85 · outbound

This paper cites Blip3-o: A family of fully open unified multimodal models-architecture, training and dataset, 2025.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Blip3-o: A family of fully open unified multimodal models-architecture, training and dataset, 2025

Reference 98

Resolution
unresolved
no resolver link, observed 2026-08-07T14:05:03.050083Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:05:03.050083Z digest=sha256:bb593f895413a765b3f35a503be8949c58cae82af2eae0371a0587c1ad9a0967

Observation 2066104d-ff3d-471e-993c-6f70524108d5 · outbound

This paper cites JanusFlow: Harmonizing Autoregression and Rectified Flow for Unified Multimodal Understanding and Generation.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities JanusFlow: Harmonizing Autoregression and Rectified Flow for Unified Multimodal Understanding and Generation

Reference 99

Resolution
unresolved
no resolver link, observed 2026-08-07T14:05:03.212698Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:05:03.212698Z digest=sha256:3901fc6cabd48ff2e37b76af7962672ec04527256916b7c8e5771101ce2f8bb6

Observation e033c797-86a4-46d5-971b-b01e308d1284 · outbound

This paper cites Building normalizing flows with stochastic interpolants.

FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Building normalizing flows with stochastic interpolants

Reference 100

Resolution
unresolved
no resolver link, observed 2026-08-07T14:05:03.315466Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:05:03.315466Z digest=sha256:e32a8fb142dfd0fb50eda25943236563d26be63e057a40deda7a99945431642d

Pith citing papers

Observation 7938a21f-4eeb-415c-97a7-3b94b7cfc3b4 · inbound

A Survey on Diffusion Language Models cites this paper.

A Survey on Diffusion Language Models FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities

Reference 97

Resolution
unresolved
no resolver link, observed 2026-08-05T20:15:21.721749Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:15:21.721749Z digest=sha256:1d38afc418bd5cd9d60446e89bf02757198c7d318d5e9112b208f5bbbec52813

Observation 5fab91ed-cdb0-4945-9d2e-f0487e0dbc96 · inbound

Discrete Guidance Matching: Exact Guidance for Discrete Flow Matching cites this paper.

Discrete Guidance Matching: Exact Guidance for Discrete Flow Matching FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities

Reference 78

Resolution
verified exact
arxiv_id, observed 2026-05-18T14:16:27.823170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-18T14:13:48.523955Z digest=sha256:d816aba502ddc829506b2bc18afb9d27139149d7b6aa3cc34cb0d74a431a5250

Observation 5ee8038c-463c-4d80-8abd-026fde5638ee · inbound

DFM-VLA: Iterative Action Refinement for Robot Manipulation via Discrete Flow Matching cites this paper.

DFM-VLA: Iterative Action Refinement for Robot Manipulation via Discrete Flow Matching FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-14T22:59:34.043445Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-14T22:59:21.473359Z digest=sha256:39f66643c51f63a65307dc78129bb8c1db441c332e0b1f825165fd0064cc50e9

Observation 1bbe8ea2-1a6e-4fbb-988c-df81dc05d22c · inbound

BARD: Bridging AutoRegressive and Diffusion Vision-Language Models Via Highly Efficient Progressive Block Merging and Stage-Wise Distillation cites this paper.

BARD: Bridging AutoRegressive and Diffusion Vision-Language Models Via Highly Efficient Progressive Block Merging and Stage-Wise Distillation FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-10T14:30:30.660105Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T14:28:47.505755Z digest=sha256:8ea46c3e9fc00b8a143105a254a2aece8e74a588c1025fe59033ebcb5fe41104

Observation 7d2f5983-b12d-4e73-908e-b5b3ea887e0c · inbound

UDM-GRPO: Stable and Efficient Group Relative Policy Optimization for Uniform Discrete Diffusion Models cites this paper.

UDM-GRPO: Stable and Efficient Group Relative Policy Optimization for Uniform Discrete Diffusion Models FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-10T05:36:01.889481Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T05:32:54.235578Z digest=sha256:87c573016a676876259bd1ceeef8c8e59bc1674cfdb3096bcb3ab2adf2f77fdd

Observation f7f62af5-c6db-4057-85c9-ea7f551df780 · inbound

UDM-GRPO: Stable and Efficient Group Relative Policy Optimization for Uniform Discrete Diffusion Models cites this paper.

UDM-GRPO: Stable and Efficient Group Relative Policy Optimization for Uniform Discrete Diffusion Models FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-07-05T11:41:02.730254Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-07-05T11:32:38.636335Z digest=sha256:89f310d65ede5a8fa400d86a3c6b882cc47e2d8806e1d4b422011b66ef56f6f8

Observation 6d6dab0f-be0f-49f9-8a90-ed5935a107a3 · inbound

SynerMedGen: Synergizing Medical Multimodal Understanding with Generation via Task Alignment cites this paper.

SynerMedGen: Synergizing Medical Multimodal Understanding with Generation via Task Alignment FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities

Reference 32

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T07:26:31.139664Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-12T02:52:01.518301Z digest=sha256:b2a21dc948b11a4aa6c020f3111e42bf1b0064c99384777323d264bc10a1842a

Observation 92b09563-0761-44f8-ac49-816d37abf3b3 · inbound

Discrete Flow Matching: Convergence Guarantees Under Minimal Assumptions cites this paper.

Discrete Flow Matching: Convergence Guarantees Under Minimal Assumptions FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-12T07:51:42.502998Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-12T01:47:08.976886Z digest=sha256:18991ab92f98d82d409188df0d94b5527a837c4a335c1ee73477c261a098b47d

Observation 2d614183-073a-40d1-b8dc-a1449344745d · inbound

UniCanvas: A Diffusion-base Unified Model for Text-in-Image Joint Generation cites this paper.

UniCanvas: A Diffusion-base Unified Model for Text-in-Image Joint Generation FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities

Reference 52

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T03:06:29.415805Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-28T10:23:43.501656Z digest=sha256:b3ed563f256b31675746adc745b154fc29532d05cc7c54b7bc96208db9b8c128

Observation cc12751c-e27c-4a5b-b2ed-845747d2d182 · inbound

UniTranslator: A Unified Multi-modal Framework for End-to-end In-Image Machine Translation cites this paper.

UniTranslator: A Unified Multi-modal Framework for End-to-end In-Image Machine Translation FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities

Reference 42

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T16:09:57.466387Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-26T00:50:22.839005Z digest=sha256:b9a0c01b676f22182488e962bec7b49472e599d6522b91220a8d6cc0f9e7e96d

Observation cd52193b-2f7b-45d8-89dd-1614479348c3 · inbound

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation cites this paper.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.712121Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.712121Z digest=sha256:df902fcd18e4dd0b4b17dbe55b75b0f67295b242fd6e38fdadb4094d473b1455