Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T04:41:58.965498Z
Paper Citation Record · LEDGER
As of 15 August 2026, this Paper Citation Record lists 60 of 60 outbound references and 7 inbound Pith citation observations for arXiv:2412.01169.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T04:41:58.965498Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:59:35.563998Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-12T05:51:25.529624Z
60 of 60 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 80172247-fb91-44be-a825-d9e64dc83ff6 · outbound
OmniFlow: Any-to-Any Generation with Multi-Modal Rectified Flows An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1ff5d55a-d177-4a51-b74d-a7f14a6feaba · outbound
OmniFlow: Any-to-Any Generation with Multi-Modal Rectified Flows Soundnet: Learning sound representations from unlabeled video
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation b7df825a-f920-4572-ba07-5776ae5e7062 · outbound
OmniFlow: Any-to-Any Generation with Multi-Modal Rectified Flows Audiosetcaps: Enriched audio captioning dataset generation using large audio language models
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 8f9c18c3-0993-478b-8dfc-def875b09431 · outbound
OmniFlow: Any-to-Any Generation with Multi-Modal Rectified Flows One transformer fits all distributions in multi-modal diffusion at scale
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation d3f5ccb4-c2b6-4989-bf17-da0a20f88049 · outbound
OmniFlow: Any-to-Any Generation with Multi-Modal Rectified Flows Coyo-700m: Image-text pair dataset
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation d7b5ff61-7cc2-4689-a578-0fe393d9e436 · outbound
OmniFlow: Any-to-Any Generation with Multi-Modal Rectified Flows Conceptual 12M: Pushing web-scale image-text pre- training to recognize long-tail visual concepts
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation e26b7d62-9c3c-4cbb-a0ce-e9f6e42f795f · outbound
OmniFlow: Any-to-Any Generation with Multi-Modal Rectified Flows Vggsound: A large-scale audio-visual dataset
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation de1471bd-58dc-41b1-a440-e190558e99dd · outbound
OmniFlow: Any-to-Any Generation with Multi-Modal Rectified Flows SLAM-AAC: Enhancing Audio Captioning with Paraphrasing Augmentation and CLAP-Refine through LLMs
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 2807c2a2-f8ae-432e-87b1-d68cd01983e2 · outbound
OmniFlow: Any-to-Any Generation with Multi-Modal Rectified Flows Scaling instruction- finetuned language models
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 8b481c2c-1a4a-4843-b2be-526c7a2da070 · outbound
OmniFlow: Any-to-Any Generation with Multi-Modal Rectified Flows Clap learning audio concepts from natural language supervision
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation b79bd482-a7fe-456b-aeed-3a34b3f078b6 · outbound
OmniFlow: Any-to-Any Generation with Multi-Modal Rectified Flows Scaling rectified flow trans- formers for high-resolution image synthesis
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation f700cbe1-38e2-4afa-a9de-ad53700f03f5 · outbound
OmniFlow: Any-to-Any Generation with Multi-Modal Rectified Flows Audio set: An ontology and human-labeled dataset for audio events
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 2bf2599f-7c1c-4ea2-bc74-109c2f134fc4 · outbound
OmniFlow: Any-to-Any Generation with Multi-Modal Rectified Flows Geneval: An object-focused framework for evaluating text-to- image alignment
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 334aee7e-57ab-48c5-aeb0-47d334d87071 · outbound
OmniFlow: Any-to-Any Generation with Multi-Modal Rectified Flows Text-to-image-2m dataset, 2024
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 320147df-2394-42f3-8bc6-94bde91ee8ad · outbound
OmniFlow: Any-to-Any Generation with Multi-Modal Rectified Flows Gans trained by a two time-scale update rule converge to a local nash equilibrium
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0c0fc8d5-a86a-4b4b-8f5a-16973d13a3e7 · outbound
OmniFlow: Any-to-Any Generation with Multi-Modal Rectified Flows Classifier-Free Diffusion Guidance
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c5f08eda-37a5-43bc-9631-bfcc9ae493da · outbound
OmniFlow: Any-to-Any Generation with Multi-Modal Rectified Flows Imagen Video: High Definition Video Generation with Diffusion Models
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b93a4ac1-4156-4ca1-b18e-ea5e3eff11dc · outbound
OmniFlow: Any-to-Any Generation with Multi-Modal Rectified Flows Make-An-Audio 2: Temporal-Enhanced Text-to-Audio Generation
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6bb9dc43-5aab-4b3e-924c-1c174ac6f688 · outbound
OmniFlow: Any-to-Any Generation with Multi-Modal Rectified Flows Make-an-audio: Text-to-audio generation with prompt-enhanced diffusion models
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 007b8214-6836-49b2-841d-1a3950d4afbd · outbound
OmniFlow: Any-to-Any Generation with Multi-Modal Rectified Flows Fr\'echet Audio Distance: A Metric for Evaluating Music Enhancement Algorithms
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a98b045d-7610-48c3-81e9-b5ca875f93a8 · outbound
OmniFlow: Any-to-Any Generation with Multi-Modal Rectified Flows Audiocaps: Generating captions for audios in the wild
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation f0204d47-2aff-49a4-ac9a-2ec73b286ddc · outbound
OmniFlow: Any-to-Any Generation with Multi-Modal Rectified Flows Understanding diffusion ob- jectives as the elbo with simple data augmentation
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 248e4d2d-532a-4b7e-9089-9afa20e068cf · outbound
OmniFlow: Any-to-Any Generation with Multi-Modal Rectified Flows Equivariant flow matching
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 04286985-c320-4d5c-b573-024e5a7f7941 · outbound
OmniFlow: Any-to-Any Generation with Multi-Modal Rectified Flows AudioGen: Textually Guided Audio Generation
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 29552902-3cb3-4d10-8f0d-7de3493cd21c · outbound
OmniFlow: Any-to-Any Generation with Multi-Modal Rectified Flows Aesthetics for open source, 2023
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 7f48e964-7d82-4ece-a0e3-d45e75b654ae · outbound
OmniFlow: Any-to-Any Generation with Multi-Modal Rectified Flows Laion coco: 600m synthetic captions from laion2b- en, 2023
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 2687f756-f306-4d73-94af-7032742ed0eb · outbound
OmniFlow: Any-to-Any Generation with Multi-Modal Rectified Flows Optimus: Organizing Sentences via Pre-trained Modeling of a Latent Space
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b96a9428-f234-4389-a5d8-93b66b7c0b1e · outbound
OmniFlow: Any-to-Any Generation with Multi-Modal Rectified Flows Blip: Bootstrapping language-image pre-training for unified vision- language understanding and generation
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 88af7c7b-04fe-4c40-b94e-b7f9cab21d57 · outbound
OmniFlow: Any-to-Any Generation with Multi-Modal Rectified Flows Blip- 2: Bootstrapping language-image pre-training with frozen image encoders and large language models
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4b3732aa-6928-41e8-a767-e8845c5c7947 · outbound
OmniFlow: Any-to-Any Generation with Multi-Modal Rectified Flows Microsoft coco: Common objects in context
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d3ac34e2-b2c3-400e-8a9b-cc29a54fdd43 · outbound
OmniFlow: Any-to-Any Generation with Multi-Modal Rectified Flows Flow Matching for Generative Modeling
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c630f42a-ffe3-4939-9baf-4c49d857662a · outbound
OmniFlow: Any-to-Any Generation with Multi-Modal Rectified Flows AudioLDM: Text-to-Audio Generation with Latent Diffusion Models
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7d3eb5f8-5726-4468-9ec8-b59e54959409 · outbound
OmniFlow: Any-to-Any Generation with Multi-Modal Rectified Flows Audioldm 2: Learning holistic audio gen- eration with self-supervised pretraining
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 3cd2133d-490a-425a-8dac-84f5b8158761 · outbound
OmniFlow: Any-to-Any Generation with Multi-Modal Rectified Flows Flow Straight and Fast: Learning to Generate and Transfer Data with Rectified Flow
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 59960bc1-3f46-448b-9ed2-7bb3ee4e58a9 · outbound
OmniFlow: Any-to-Any Generation with Multi-Modal Rectified Flows Discrete diffusion modeling by estimating the ratios of the data distri- bution
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 92dd91c5-768a-4d92-ad19-03d8a2a8f659 · outbound
OmniFlow: Any-to-Any Generation with Multi-Modal Rectified Flows Unified-io 2: Scaling autoregressive multimodal models with vision language audio and action
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 941ddad8-4831-4a07-bcd3-a71eff0aa9f2 · outbound
OmniFlow: Any-to-Any Generation with Multi-Modal Rectified Flows Wavcaps: A chatgpt-assisted weakly-labelled audio captioning dataset for audio-language multimodal re- search
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation f11ce163-9154-4ef4-8e5a-316b398398cf · outbound
OmniFlow: Any-to-Any Generation with Multi-Modal Rectified Flows Image generated using midjourney ai, 2024
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 40ef35a9-3d82-445f-b56b-c92d13e71590 · outbound
OmniFlow: Any-to-Any Generation with Multi-Modal Rectified Flows Improved denoising diffusion probabilistic models
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 11c577e3-2467-4e9f-a9e8-90ceeaa0b75b · outbound
OmniFlow: Any-to-Any Generation with Multi-Modal Rectified Flows Dall-e 3, 2023
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation a63e716a-71d1-4e97-b0ea-c41f5ba8925f · outbound
OmniFlow: Any-to-Any Generation with Multi-Modal Rectified Flows Scalable diffusion models with transformers
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cfd6255d-caec-473c-bb36-78b765dacad4 · outbound
OmniFlow: Any-to-Any Generation with Multi-Modal Rectified Flows SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aec40632-a6ea-4845-b3bf-8712a40323e6 · outbound
OmniFlow: Any-to-Any Generation with Multi-Modal Rectified Flows Learning transferable visual models from natural language supervi- sion
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4bed7f85-98df-4e04-b16c-6d7fdda8e690 · outbound
OmniFlow: Any-to-Any Generation with Multi-Modal Rectified Flows High-resolution image synthesis with latent diffusion models
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 92c621be-e835-4bdf-b4c0-88af9e8c2088 · outbound
OmniFlow: Any-to-Any Generation with Multi-Modal Rectified Flows Simple and effective masked diffusion language models
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation cf004235-c4a9-4120-8a87-d3f24a81a579 · outbound
OmniFlow: Any-to-Any Generation with Multi-Modal Rectified Flows Any-to-any generation via composable diffu- sion
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 80a4ec9c-b2c6-471a-8788-f41b94149161 · outbound
OmniFlow: Any-to-Any Generation with Multi-Modal Rectified Flows Chameleon: Mixed-Modal Early-Fusion Foundation Models
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f9f188a4-f49a-4e7a-9162-f0e03a9fa852 · outbound
OmniFlow: Any-to-Any Generation with Multi-Modal Rectified Flows Improving and generalizing flow-based generative models with minibatch optimal transport
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 86d9bee7-03d3-4c2d-a90d-114cddad3b8a · outbound
OmniFlow: Any-to-Any Generation with Multi-Modal Rectified Flows Cider: Consensus-based image description evalu- ation
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 1ac6539f-4d1f-40a5-bd7a-5ea408400e3a · outbound
OmniFlow: Any-to-Any Generation with Multi-Modal Rectified Flows Vector-quantized Image Modeling with Improved VQGAN
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0b5426c5-4765-4e1b-b98a-be32add155b4 · outbound
OmniFlow: Any-to-Any Generation with Multi-Modal Rectified Flows TinyLlama: An Open-Source Small Language Model
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7af8c7f2-fa1c-4076-8679-52e23dbc0aaa · outbound
OmniFlow: Any-to-Any Generation with Multi-Modal Rectified Flows Transfusion: Predict the Next Token and Diffuse Images with One Multi-Modal Model
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fa0f9525-67a7-420f-9ef7-1b4819366448 · outbound
OmniFlow: Any-to-Any Generation with Multi-Modal Rectified Flows noise level
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation a3430458-5820-4f57-8419-6ae832680308 · outbound
OmniFlow: Any-to-Any Generation with Multi-Modal Rectified Flows We train Model 2 for 100k steps and Model 3 for 150k steps
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 4aadeaf1-2a5a-4820-9d50-4c249a7a6c8a · outbound
OmniFlow: Any-to-Any Generation with Multi-Modal Rectified Flows The learning rate under- goes a linear warmup in the first 1000 steps and a cosine decay throughout the rest of the training
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 101711de-1266-4bde-ac93-14468eade4ea · outbound
OmniFlow: Any-to-Any Generation with Multi-Modal Rectified Flows Unresolved cited work
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 9c1ca971-3aae-47e7-9d6a-f6202be73e2e · outbound
OmniFlow: Any-to-Any Generation with Multi-Modal Rectified Flows Instead, we sample p(x0 1, x1 3|x0
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 3013141e-f663-4032-a4db-591ae65de4bc · outbound
OmniFlow: Any-to-Any Generation with Multi-Modal Rectified Flows Unresolved cited work
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 09941fc5-c0b0-487e-ae04-fccce7b390b7 · outbound
OmniFlow: Any-to-Any Generation with Multi-Modal Rectified Flows Unresolved cited work
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation eac4a76c-4909-4193-a3ca-2a2102097928 · outbound
OmniFlow: Any-to-Any Generation with Multi-Modal Rectified Flows car”, “bird
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 52608195-9992-48b1-ab74-4c2502fe686c · inbound
LaViDa: A Large Diffusion Language Model for Multimodal Understanding OmniFlow: Any-to-Any Generation with Multi-Modal Rectified Flows
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 43be3b75-d5fa-411d-a840-991ebc006b4b · inbound
UniRL: Self-Improving Unified Multimodal Models via Supervised and Reinforcement Learning OmniFlow: Any-to-Any Generation with Multi-Modal Rectified Flows
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 08756ba8-eb7c-4020-943c-5cc4f05a57f7 · inbound
Diffuse Everything: Multimodal Diffusion Models on Arbitrary State Spaces OmniFlow: Any-to-Any Generation with Multi-Modal Rectified Flows
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9ebe2114-334c-42b0-ac95-61f35730573a · inbound
Flow Diverse and Efficient: Learning Momentum Flow Matching via Stochastic Velocity Field Sampling OmniFlow: Any-to-Any Generation with Multi-Modal Rectified Flows
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e5cec857-e94b-4962-9b6d-4236679b470c · inbound
Flow Straight and Fast in Hilbert Space: Functional Rectified Flow OmniFlow: Any-to-Any Generation with Multi-Modal Rectified Flows
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4ced0174-0da9-4e56-bd04-66c6fd7506c5 · inbound
Lavida-O: Elastic Large Masked Diffusion Models for Unified Multimodal Understanding and Generation OmniFlow: Any-to-Any Generation with Multi-Modal Rectified Flows
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2bcaa1ba-a0f9-462b-8831-b78185c9ddb6 · inbound
Any2Any 3D Diffusion Models with Knowledge Transfer: A Radiotherapy Planning Study OmniFlow: Any-to-Any Generation with Multi-Modal Rectified Flows
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.