Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:05:03.315466Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 100 of 178 outbound references and 11 inbound Pith citation observations for arXiv:2505.20147.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:05:03.315466Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-05T20:15:21.721749Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-05T11:41:02.728596Z
100 of 178 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation c1a2050a-50dc-4922-9ecc-560b97188b1b · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Deepseek-r1: Incentivizing reasoning capability in llms via reinforcement learning, 2025
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d06f3e0a-b160-4abf-ad26-e98df6c9265a · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Qwen2.5 Technical Report
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9766cfe1-7c43-40a3-adee-c7e8565673da · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities The Llama 3 Herd of Models
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 557b361b-e49e-4301-af2a-f4f445f6ada3 · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Internlm2 technical report, 2024
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 893607b1-69e4-4efa-af02-4b5edf3eb4ae · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Unresolved cited work
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 97d2b2f3-d6b7-4f33-a2c4-6fd32cfd1999 · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Visual instruction tuning
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation db23d69b-55c2-44be-abb1-0feb49777ee3 · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 454ddda6-a733-44f4-b480-3e1225dfb1da · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Instructblip: Towards general-purpose vision-language models with instruction tuning, 2023
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d01f4598-f6a0-4977-a4f7-101b2ba3b4b4 · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities DeepSeek-VL: Towards Real-World Vision-Language Understanding
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4683a615-9f8c-4fd8-b3f7-400d37f7e149 · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2323d415-217c-4847-bfb0-ef81d7169882 · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Diffusion models beat gans on image synthesis
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 636b6e21-ad95-4f6d-b3a0-a963a6adae44 · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities High-resolution image synthesis with latent diffusion models
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eb910a6c-253f-4a62-b77b-d680bc3cd313 · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities PixArt-$\alpha$: Fast Training of Diffusion Transformer for Photorealistic Text-to-Image Synthesis
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a5e348ca-5944-4461-adfa-7f42f8129d5e · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Scaling Rectified Flow Transformers for High-Resolution Image Synthesis
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 94390c54-1090-4fc6-9be6-6c9c48ae45db · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Unresolved cited work
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9121c834-8d28-4554-bafa-d53507f3dc9c · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Planting a SEED of Vision in Large Language Model
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 11016a83-596a-4a6e-bee4-aad63d187c44 · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Making LLaMA SEE and Draw with SEED Tokenizer
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b821a411-f2e3-4a95-b6c6-3e93d6704e58 · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Emu3: Next-token prediction is all you need, 2024
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7cc05029-eff6-4938-9724-845eb0cc0041 · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Transfusion: Predict the Next Token and Diffuse Images with One Multi-Modal Model
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d171d038-a94b-4373-b222-57963cb0f648 · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Janus: Decoupling Visual Encoding for Unified Multimodal Understanding and Generation
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3f887847-1b4b-4e0d-a8e0-b141537f80dc · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Illume: Illuminating your llms to see, draw, and self-enhance, 2024
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 72b3daa5-cb80-40be-8d0d-595107375aa1 · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Janus-Pro: Unified Multimodal Understanding and Generation with Data and Model Scaling
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 95f3b668-cc3d-46ab-a44f-4ef7ac551063 · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Illume+: Illuminating unified mllm with dual visual tokenization and diffusion refinement, 2025
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9d79ce17-4fd2-49e5-9ba8-787a0ef4d3a4 · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Sparks of artificial general intelligence: Early experiments with gpt-4, 2023
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 118b17c7-1120-408c-bc87-39974a88512b · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Hwang, Soumya Sanyal, Sean Welleck, Xiang Ren, Allyson Ettinger, Zaid Harchaoui, and Yejin Choi
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 29109999-bdad-4cb0-9c0b-2d873aa2b929 · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities The pitfalls of next-token prediction
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ee6bcc26-c9a3-42c4-8471-0480ce1f13f1 · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Beyond Autoregression: Discrete Diffusion for Complex Reasoning and Planning
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8f803c8a-825f-417a-a536-21852e44ef64 · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Large Language Models Cannot Self-Correct Reasoning Yet
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7bce1405-d86a-47cb-8cc8-faa0a7b09366 · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Structured denoising diffusion models in discrete state-spaces
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d5350019-08ef-4869-9a03-a701ae7d2443 · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Discrete diffusion modeling by estimating the ratios of the data distribution
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 515895c9-5225-4d24-b22e-51667697f591 · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Simplified and generalized masked diffusion for discrete data
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e3f3d6bc-b168-4c2d-89c7-df2311791e2e · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Simple and effective masked diffusion language models
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b8adb77b-f844-4fa8-b17f-531e7296e865 · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Discrete flow matching
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 259dca27-6153-4b04-b29e-d90de1fe47f8 · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Unresolved cited work
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c4a8de79-0f51-4e97-bb7c-6b56de394639 · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Generative flows on discrete state-spaces: Enabling multimodal flows with applications to protein co-design
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a5e545d2-a5d0-4150-ad5f-c91395878700 · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities URL https://www.inceptionlabs.ai/news
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d3252845-391e-4ed8-8564-25bebc7bed1b · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Flow Straight and Fast: Learning to Generate and Transfer Data with Rectified Flow
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2eaeb6cb-6b90-484e-a368-ec39be947865 · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Flow Matching for Generative Modeling
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 516f93ed-57ba-40e8-aa40-443d0c92fff2 · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Consistency models
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d3c0dfc9-9034-4a0a-8f65-5776a66aa78a · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Large Language Diffusion Models
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 631d0bb4-be91-4b6b-b597-8ddbb990a2b4 · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Dream 7b, 2025
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 45345e19-f584-4399-ae16-ff1d85fe254c · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Dual Diffusion for Unified Image Generation and Understanding
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 24d9e19b-6303-413a-bbf3-646a55d82b0a · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Unified Discrete Diffusion for Simultaneous Vision-Language Generation
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8a7ddd00-d69d-4127-bcbd-7de3b7e2b9ec · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Unified Multimodal Discrete Diffusion
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4b174cf9-7e0e-4e02-ab50-dd8c81e95aec · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Scaling Diffusion Language Models via Adaptation from Autoregressive Models
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5bd7432a-5bd7-4cc2-9a71-2277744ba661 · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities SANA 1.5: Efficient Scaling of Training-Time and Inference-Time Compute in Linear Diffusion Transformer
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8a287536-95fb-43e8-b179-b55524e88c6c · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Dancegrpo: Unleashing grpo on visual generation, 2025
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4daf39c0-2bfa-401f-8cad-40101393906b · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities ILLUME+: Illuminating Unified MLLM with Dual Visual Tokenization and Diffusion Refinement
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation be3146dd-8cf7-459a-a608-98c4b78936f1 · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities VILA-u: a unified foundation model integrating visual understanding and generation
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation be7d454e-1ea6-4802-8b50-dfcf7255c269 · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Chameleon: Mixed-Modal Early-Fusion Foundation Models
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7ee0f59a-7eec-43d0-87b3-82e6d30bf694 · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Unified language-vision pretraining in LLM with dynamic discrete visual tokenization
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dbce4436-7354-453b-a11b-914505f55b45 · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Show-o: One Single Transformer to Unify Multimodal Understanding and Generation
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9e986136-9cde-44b3-bd4b-f25bb51d660f · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Denoising diffusion probabilistic models
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 33a9a937-9203-4e14-96c4-1b152633a63e · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Diffusionbert: Improving generative masked language models with diffusion models
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation abfef5d2-d562-4075-8b8d-35165a4bc48c · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Sigmoid loss for language image pre-training
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4116490a-c833-4136-9af7-14e168160f06 · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Autoregressive Model Beats Diffusion: Llama for Scalable Image Generation
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b6e354e6-b377-4bf1-90e4-2954bd1181c7 · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities LLaVAR: Enhanced Visual Instruction Tuning for Text-Rich Image Understanding
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2bf5523b-33b5-4c4f-93a7-d24eb14b9c9a · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities wendlerc/renderedtext, 2023
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 180e9c8d-04e6-4591-93f4-45afe9049e9a · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Docvqa: A dataset for vqa on document images
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c0efe589-9f20-4eb6-86dd-88ff569f833d · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Chart-to-text: Generating natural language descriptions for charts by adapting the transformer model, 2020
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e1206549-5bc3-4261-9251-af4a896f38c4 · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Visualmrc: Machine reading compre- hension on document images
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 946e16cf-21bb-4a77-93bc-530a43a3a1e2 · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities G-llava: Solving geomet- ric problem with multi-modal large language model, 2023
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9fce5abe-f0ee-483b-bdcf-af20234e15a2 · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Xing, and Liang Lin
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 356ccaf6-4178-44f5-b634-a9a3c7375158 · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Magpie: Alignment Data Synthesis from Scratch by Prompting Aligned LLMs with Nothing
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 608d9593-a179-42da-b8e3-f89ecada8317 · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities World model on million-length video and language with blockwise ringattention
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e91cac6a-9d61-4be5-8337-dcd7b674c884 · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Hierarchical Text-Conditional Image Generation with CLIP Latents
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d72b7cb6-759b-4c6e-8e72-dab9c8f573be · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 639ed477-c027-466e-843e-aecd9275f5de · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Improving image generation with better captions
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 51ebc488-ffbf-4871-bacd-0cb52cff561e · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities SEED-X: Multimodal Models with Unified Multi-granularity Comprehension and Generation
Reference 69
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d9455675-b83a-4d95-beea-188e4e763b43 · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities TokenFlow: Unified Image Tokenizer for Multimodal Understanding and Generation
Reference 70
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0ee9b758-72d2-4472-a2d9-593715f9512e · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Geneval: An object-focused framework for evaluating text-to-image alignment
Reference 71
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0205f82c-f161-429e-9c93-371945102ce1 · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices
Reference 72
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ea734a38-7f00-41dd-9e73-0c3226fbc66c · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities MobileVLM V2: Faster and Stronger Baseline for Vision Language Model
Reference 73
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 62bdc963-0009-484b-8998-f6b6db471a92 · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities LLaVA-Phi: Efficient Multi-Modal Assistant with Small Language Model
Reference 74
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e0b17c29-69e1-48d5-9d54-25a7d565b4bc · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Improved baselines with visual instruction tuning
Reference 75
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8ddb8461-d77c-4bfd-998f-58f3745bd411 · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond
Reference 76
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8057f8b4-a8d9-4940-b546-2217476ce832 · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Introducing idefics: An open reproduction of state-of-the-art visual language model, 2023
Reference 77
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d7afa45c-ae24-4cd7-bc30-8283bc127c8c · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Unified language-vision pretraining with dynamic discrete visual tokenization
Reference 78
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a288f023-8bb1-4e83-98a7-cc8f9b59bfb1 · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities MetaMorph: Multimodal Understanding and Generation via Instruction Tuning
Reference 79
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 27cc4a4f-dce6-46b7-a579-aead07dadab0 · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Gemini: A Family of Highly Capable Multimodal Models
Reference 80
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c70f29be-085f-4c1a-b320-9a229d9baeef · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities VILA-U: a Unified Foundation Model Integrating Visual Understanding and Generation
Reference 81
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9cf071ae-eb3c-40a2-9a8d-13f5ba9565b5 · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Evaluating Object Hallucination in Large Vision-Language Models
Reference 82
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e6a730c6-f49e-44b6-a039-9b4b7dd5f33c · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models
Reference 83
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f9442855-c1f7-40af-bca1-0b4b52740018 · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities SEED-Bench: Benchmarking Multimodal LLMs with Generative Comprehension
Reference 84
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f0db4cdd-7d8a-46fd-8aa0-322d844a3618 · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities MMBench: Is Your Multi-modal Model an All-around Player?
Reference 85
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f16fcff8-183a-48df-9fa0-0402126d77ec · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Gqa: A new dataset for real-world visual reasoning and compositional question answering
Reference 86
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cfb9a97b-0c24-421f-af97-1bc4efe5f708 · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Mmmu: A massive multi- discipline multimodal understanding and reasoning benchmark for expert agi
Reference 87
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a717de63-b6ef-4029-b109-cebaf2e65ab5 · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities MM-Vet: Evaluating Large Multimodal Models for Integrated Capabilities
Reference 88
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ef6700ac-f5f4-4846-b60a-0cefd3575d1f · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities NVILA: Efficient Frontier Visual Language Models
Reference 89
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c2a388f0-818e-46b1-a604-67ae8745ee15 · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Flow-GRPO: Training Flow Matching Models via Online RL
Reference 90
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 083b6baa-5d7f-478c-88d3-a32dcd1b8eb4 · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Unified-IO: A Unified Model for Vision, Language, and Multi-Modal Tasks
Reference 91
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f2ee1038-e567-4fe8-8bd3-82bad39e832c · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Unified-io 2: Scaling autoregressive multimodal models with vision language audio and action
Reference 92
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ed131da7-b973-4265-9ad3-00dc48bac017 · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities AnyGPT: Unified Multimodal LLM with Discrete Sequence Modeling
Reference 93
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 73ba1ad5-474b-4b33-91bd-c6f6d0e70bbd · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Generative Multimodal Pretraining with Discrete Diffusion Timestep Tokens
Reference 94
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b913b799-b7fa-4f75-9a1e-4a153fe1ddaf · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Dreamllm: Synergistic multimodal comprehension and creation
Reference 95
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f46af32c-f3f7-41da-9a09-73517e82ff99 · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Minigpt-5: Interleaved vision-and-language generation via generative vokens
Reference 96
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation de85dd05-1d00-430b-ba82-b0d600da042e · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Next-gpt: Any-to-any multimodal llm
Reference 97
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 32ce7d75-fa9f-401b-803d-6fee32024c85 · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Blip3-o: A family of fully open unified multimodal models-architecture, training and dataset, 2025
Reference 98
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2066104d-ff3d-471e-993c-6f70524108d5 · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities JanusFlow: Harmonizing Autoregression and Rectified Flow for Unified Multimodal Understanding and Generation
Reference 99
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e033c797-86a4-46d5-971b-b01e308d1284 · outbound
FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities Building normalizing flows with stochastic interpolants
Reference 100
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7938a21f-4eeb-415c-97a7-3b94b7cfc3b4 · inbound
A Survey on Diffusion Language Models FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities
Reference 97
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5fab91ed-cdb0-4945-9d2e-f0487e0dbc96 · inbound
Discrete Guidance Matching: Exact Guidance for Discrete Flow Matching FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities
Reference 78
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 5ee8038c-463c-4d80-8abd-026fde5638ee · inbound
DFM-VLA: Iterative Action Refinement for Robot Manipulation via Discrete Flow Matching FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1bbe8ea2-1a6e-4fbb-988c-df81dc05d22c · inbound
BARD: Bridging AutoRegressive and Diffusion Vision-Language Models Via Highly Efficient Progressive Block Merging and Stage-Wise Distillation FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7d2f5983-b12d-4e73-908e-b5b3ea887e0c · inbound
UDM-GRPO: Stable and Efficient Group Relative Policy Optimization for Uniform Discrete Diffusion Models FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f7f62af5-c6db-4057-85c9-ea7f551df780 · inbound
UDM-GRPO: Stable and Efficient Group Relative Policy Optimization for Uniform Discrete Diffusion Models FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6d6dab0f-be0f-49f9-8a90-ed5935a107a3 · inbound
SynerMedGen: Synergizing Medical Multimodal Understanding with Generation via Task Alignment FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 92b09563-0761-44f8-ac49-816d37abf3b3 · inbound
Discrete Flow Matching: Convergence Guarantees Under Minimal Assumptions FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 2d614183-073a-40d1-b8dc-a1449344745d · inbound
UniCanvas: A Diffusion-base Unified Model for Text-in-Image Joint Generation FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation cc12751c-e27c-4a5b-b2ed-845747d2d182 · inbound
UniTranslator: A Unified Multi-modal Framework for End-to-end In-Image Machine Translation FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation cd52193b-2f7b-45d8-89dd-1614479348c3 · inbound
Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation FUDOKI: Discrete Flow-based Unified Understanding and Generation via Kinetic-Optimal Velocities
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.