Pith. sign in

Paper Citation Record · LEDGER

MuLan: Adapting Multilingual Diffusion Models for Hundreds of Languages with Negligible Cost

As of 12 August 2026, this Paper Citation Record lists 52 of 52 outbound references and 1 inbound Pith citation observation for arXiv:2412.01271.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.01271 v2

Coverage vector

measured 52 of 52 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T04:34:52.657768Z

measured 53 of 53 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-10T15:31:35.751516Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-10T15:31:35.912902Z

Reference resolution

52 of 52 outbound references displayed

  • verified exact4
  • verified fuzzy12
  • unresolved36
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 0096384f-13f1-457c-95d8-ccbfac480b9a · outbound

This paper cites write newline.

MuLan: Adapting Multilingual Diffusion Models for Hundreds of Languages with Negligible Cost write newline

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-12T04:34:52.458486Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:34:52.458486Z digest=sha256:9202957b0ba47304da94990c3f85de807866d0a308951c9f9c59f5de3024c5e3

Observation add34e51-b1fb-48c9-bec3-3d068dfa7178 · outbound

This paper cites Kandinsky 3.0 Technical Report.

MuLan: Adapting Multilingual Diffusion Models for Hundreds of Languages with Negligible Cost Kandinsky 3.0 Technical Report

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-12T04:34:52.463687Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:34:52.463687Z digest=sha256:36b18b0ce61f1ffcfeaa7b41585574bffa74381d18b8661937f1d0c8df40409c

Observation 4ce1ee87-3286-4eb6-90a1-85670cb53e0d · outbound

This paper cites D., Dhariwal, P., Neelakantan, A., Shyam, P., Sastry, G., Askell, A., et al.

MuLan: Adapting Multilingual Diffusion Models for Hundreds of Languages with Negligible Cost D., Dhariwal, P., Neelakantan, A., Shyam, P., Sastry, G., Askell, A., et al

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-12T04:34:52.468337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:34:52.468337Z digest=sha256:1bc9b382941aa94cbf67590c206bf32877040b38e3fe8565d85275b6b6607320

Observation 490ce0d2-5c1e-44e4-8291-1e3a06f84b49 · outbound

This paper cites Cross-lingual and multilingual clip.

MuLan: Adapting Multilingual Diffusion Models for Hundreds of Languages with Negligible Cost Cross-lingual and multilingual clip

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:34:53.238344Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T04:34:52.472810Z digest=sha256:9143f773b773ae1cb6a111600b89e5d8f5414c47388c28be59b221c52bad08dc

Observation 45f589b4-d5b8-4335-8c7b-b1f41386ca43 · outbound

This paper cites Pixart- : Fast training of diffusion transformer for photorealistic text-to-image synthesis, 2023 a.

MuLan: Adapting Multilingual Diffusion Models for Hundreds of Languages with Negligible Cost Pixart- : Fast training of diffusion transformer for photorealistic text-to-image synthesis, 2023 a

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:34:53.225811Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T04:34:52.476695Z digest=sha256:7a9adb315e2f7dff3100d25617f9a48a13f37e81c8ef83afe71e112c23783cbc

Observation 439d30f4-0d50-4f3a-8563-05ecf9b4d9f4 · outbound

This paper cites AltCLIP: Altering the Language Encoder in CLIP for Extended Language Capabilities.

MuLan: Adapting Multilingual Diffusion Models for Hundreds of Languages with Negligible Cost AltCLIP: Altering the Language Encoder in CLIP for Extended Language Capabilities

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-12T04:34:52.480541Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:34:52.480541Z digest=sha256:2e16a0a8a25488ac07121f7290ac46484c7224fb905824f86ca7bb4a849eaf29

Observation f92402be-c2f8-41b5-bf40-3540ede875ce · outbound

This paper cites InternVL: Scaling up Vision Foundation Models and Aligning for Generic Visual-Linguistic Tasks.

MuLan: Adapting Multilingual Diffusion Models for Hundreds of Languages with Negligible Cost InternVL: Scaling up Vision Foundation Models and Aligning for Generic Visual-Linguistic Tasks

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-12T04:34:52.484601Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:34:52.484601Z digest=sha256:4b3d364b4a06b5878a110eda251ee5e32d08a6fd322e159a1342caa275abb89f

Observation f1d1a06d-0a54-4f45-aa57-e1adda4681e9 · outbound

This paper cites Unsupervised Cross-lingual Representation Learning at Scale.

MuLan: Adapting Multilingual Diffusion Models for Hundreds of Languages with Negligible Cost Unsupervised Cross-lingual Representation Learning at Scale

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-12T04:34:52.488909Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:34:52.488909Z digest=sha256:a12c7ba5c699cf53a555fb11138e28efa6043deb25cbabb2ba44fbdc53ba5720

Observation a5962030-b7f1-48e9-9f94-073e31f6f871 · outbound

This paper cites BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding.

MuLan: Adapting Multilingual Diffusion Models for Hundreds of Languages with Negligible Cost BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-12T04:34:52.493264Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:34:52.493264Z digest=sha256:2439d771a59c1478e13b2307ae91fafa24331e4d0b7c26c3e9306d6847514271

Observation 6b792da0-1994-4d3f-a7db-2853e9e0b4ab · outbound

This paper cites The Llama 3 Herd of Models.

MuLan: Adapting Multilingual Diffusion Models for Hundreds of Languages with Negligible Cost The Llama 3 Herd of Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-12T04:34:52.497257Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:34:52.497257Z digest=sha256:cd7ff95d36fc6a3360a86782769ebf17b891b74b7fd846f5982bc770dd1c576a

Observation 9523ebfe-afd0-490f-9d59-d0e67e1d2abe · outbound

This paper cites Scaling Rectified Flow Transformers for High-Resolution Image Synthesis.

MuLan: Adapting Multilingual Diffusion Models for Hundreds of Languages with Negligible Cost Scaling Rectified Flow Transformers for High-Resolution Image Synthesis

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-12T04:34:52.501185Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:34:52.501185Z digest=sha256:2e313073ac28d421c4a8323fd434f1716e83e8adef083e3091ecdd94b90adcf1

Observation f35225be-af6c-477b-b1a9-b8bde9da0ee5 · outbound

This paper cites Efficient diffusion training via min-snr weighting strategy.

MuLan: Adapting Multilingual Diffusion Models for Hundreds of Languages with Negligible Cost Efficient diffusion training via min-snr weighting strategy

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:34:53.213055Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T04:34:52.505032Z digest=sha256:2da277bf9be8d62e449486b6c59a043e7d4bd827f1ca3cc76074ef5fb4513649

Observation a9dbf3c5-b087-464f-837f-f3f3d94b5597 · outbound

This paper cites LoRA: Low-Rank Adaptation of Large Language Models.

MuLan: Adapting Multilingual Diffusion Models for Hundreds of Languages with Negligible Cost LoRA: Low-Rank Adaptation of Large Language Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-12T04:34:52.508739Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:34:52.508739Z digest=sha256:1e6712ba656252a7abd3148ad1dc7a80aa8a281eb15af1c2e5192eae5c073900

Observation f58c1641-f862-48e3-b6c6-3f1f598e086f · outbound

This paper cites Scaling Up Visual and Vision-Language Representation Learning With Noisy Text Supervision.

MuLan: Adapting Multilingual Diffusion Models for Hundreds of Languages with Negligible Cost Scaling Up Visual and Vision-Language Representation Learning With Noisy Text Supervision

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-12T04:34:52.512550Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:34:52.512550Z digest=sha256:5273ad5c32e817346a5aab6415eea49725e8bee169badaf411ba74677b305970

Observation 7dec1bfe-41a3-4fa7-bc4f-cf3292995dfe · outbound

This paper cites Clip-vit-h-14-frozen-xlm-roberta-large-laion5b-s13b-b90k, 2023.

MuLan: Adapting Multilingual Diffusion Models for Hundreds of Languages with Negligible Cost Clip-vit-h-14-frozen-xlm-roberta-large-laion5b-s13b-b90k, 2023

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:34:53.201186Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T04:34:52.516355Z digest=sha256:0dad211662a45399f89ec92167c892180a75f2d9d7502b958aaa8f67f4038fc0

Observation 02d9cd97-3ddc-48f7-a891-95c2a6f6a130 · outbound

This paper cites BART: Denoising Sequence-to-Sequence Pre-training for Natural Language Generation, Translation, and Comprehension.

MuLan: Adapting Multilingual Diffusion Models for Hundreds of Languages with Negligible Cost BART: Denoising Sequence-to-Sequence Pre-training for Natural Language Generation, Translation, and Comprehension

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-12T04:34:52.519894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:34:52.519894Z digest=sha256:bbeac587b197e3b19a0a7a06d66ae6aa208120dc8967856d288ed793e81988b5

Observation 2d52467a-6d46-4f9c-9e48-760c56a34b8d · outbound

This paper cites Playground v2.5: Three Insights towards Enhancing Aesthetic Quality in Text-to-Image Generation.

MuLan: Adapting Multilingual Diffusion Models for Hundreds of Languages with Negligible Cost Playground v2.5: Three Insights towards Enhancing Aesthetic Quality in Text-to-Image Generation

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-12T04:34:52.524028Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:34:52.524028Z digest=sha256:96c99f2bd20a97a21612aec9a53a47ada423bf7953d1611478e0f5636b5be835

Observation 44c70204-fe22-4ab5-b652-290ec56af837 · outbound

This paper cites Hunyuan-dit: A powerful multi-resolution diffusion transformer with fine-grained chinese understanding, 2024 b.

MuLan: Adapting Multilingual Diffusion Models for Hundreds of Languages with Negligible Cost Hunyuan-dit: A powerful multi-resolution diffusion transformer with fine-grained chinese understanding, 2024 b

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:34:53.189160Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T04:34:52.527937Z digest=sha256:682e5a0bc8be6a704d1b6a9c512345e60e84f93551f27184aa0adbc82465c24c

Observation 440d936b-de3a-4925-87c8-aa88cf87804d · outbound

This paper cites Microsoft COCO: Common Objects in Context.

MuLan: Adapting Multilingual Diffusion Models for Hundreds of Languages with Negligible Cost Microsoft COCO: Common Objects in Context

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-12T04:34:52.531519Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:34:52.531519Z digest=sha256:192b9eb82a2f24a94300bf761c703d5ea1271459dbac05f507a0f96f1944a1ce

Observation 7a65eff2-a7a7-4ba9-90db-90fd96dcc535 · outbound

This paper cites Evaluating text-to-visual generation with image-to-text generation.

MuLan: Adapting Multilingual Diffusion Models for Hundreds of Languages with Negligible Cost Evaluating text-to-visual generation with image-to-text generation

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:34:53.177223Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T04:34:52.535284Z digest=sha256:79b60f57f8bbf651d40ff57786743d43ad3023fd4355b5cf5b5782ce773c55b1

Observation 5fe1171b-f4eb-4219-ac24-1ba135e0e56e · outbound

This paper cites Decoupled Weight Decay Regularization.

MuLan: Adapting Multilingual Diffusion Models for Hundreds of Languages with Negligible Cost Decoupled Weight Decay Regularization

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-12T04:34:52.538691Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:34:52.538691Z digest=sha256:2d50e36179949b5149dcfad80d71b709a4a106af8de4e0a992a595a3283c0fa8

Observation 2caef497-6f5c-4451-beb6-70e565d999c0 · outbound

This paper cites PanGu-Draw: Advancing Resource-Efficient Text-to-Image Synthesis with Time-Decoupled Training and Reusable Coop-Diffusion.

MuLan: Adapting Multilingual Diffusion Models for Hundreds of Languages with Negligible Cost PanGu-Draw: Advancing Resource-Efficient Text-to-Image Synthesis with Time-Decoupled Training and Reusable Coop-Diffusion

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-08-12T04:34:52.902845Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T04:34:52.542300Z digest=sha256:5c0206ff5eb1fe08563dada8ebae4b900d2033f1e1e07b230d1ff99cf63a7aff

Observation 072101b4-285f-4887-86c7-8b5f16602807 · outbound

This paper cites Latent Consistency Models: Synthesizing High-Resolution Images with Few-Step Inference.

MuLan: Adapting Multilingual Diffusion Models for Hundreds of Languages with Negligible Cost Latent Consistency Models: Synthesizing High-Resolution Images with Few-Step Inference

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-12T04:34:52.546272Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:34:52.546272Z digest=sha256:1e9c4de026ee56367027ef1198df95ca0e5fbcb0c21b0d1bb408cbe8aa0a1446

Observation 79a89b17-af5f-40da-961a-a84aec41a555 · outbound

This paper cites PEA-Diffusion: Parameter-Efficient Adapter with Knowledge Distillation in non-English Text-to-Image Generation.

MuLan: Adapting Multilingual Diffusion Models for Hundreds of Languages with Negligible Cost PEA-Diffusion: Parameter-Efficient Adapter with Knowledge Distillation in non-English Text-to-Image Generation

Reference 24

Resolution
verified exact
local_arxiv, observed 2026-08-12T04:34:52.872696Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T04:34:52.550158Z digest=sha256:0d40b2ce3f81dbdd611eb063603f4596eb2bd0e8096665adfafa88eca1f407d3

Observation f3a303fc-1027-4c8a-ba00-9514c34866f6 · outbound

This paper cites GLIDE: Towards Photorealistic Image Generation and Editing with Text-Guided Diffusion Models.

MuLan: Adapting Multilingual Diffusion Models for Hundreds of Languages with Negligible Cost GLIDE: Towards Photorealistic Image Generation and Editing with Text-Guided Diffusion Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-12T04:34:52.554495Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:34:52.554495Z digest=sha256:00b59de1ae25f71934da9a8b7e809f23c911f9d2b8be3d33b4ee86872ea0c056

Observation 7af2321a-c9e1-4711-a482-fd4390bf6092 · outbound

This paper cites an unresolved cited work.

MuLan: Adapting Multilingual Diffusion Models for Hundreds of Languages with Negligible Cost Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-12T04:34:53.165651Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T04:34:52.558199Z digest=sha256:21931cc3aa77810fae5cf9434881fc0840cf8240d1b231e785579e934e6a5b85

Observation 5049546a-b893-4096-9069-bb6ebe432f21 · outbound

This paper cites Gpt-4 technical report, 2023.

MuLan: Adapting Multilingual Diffusion Models for Hundreds of Languages with Negligible Cost Gpt-4 technical report, 2023

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-12T04:34:52.561727Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:34:52.561727Z digest=sha256:1b2f2a5e38e08c10fe51f9592a50980b48d26c5b58dac55d97dfd373252a3c88

Observation 256dcfc7-e606-4692-aae4-d67b5f40292b · outbound

This paper cites Journeydb: A benchmark for generative image understanding, 2023.

MuLan: Adapting Multilingual Diffusion Models for Hundreds of Languages with Negligible Cost Journeydb: A benchmark for generative image understanding, 2023

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:34:53.147374Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T04:34:52.565203Z digest=sha256:ed3694d28049e36077499e82a7cccb4fcb4a3ba45f7efd5ea6629745fbf174a6

Observation e42f1503-c048-4994-9232-9280e19d8618 · outbound

This paper cites SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis.

MuLan: Adapting Multilingual Diffusion Models for Hundreds of Languages with Negligible Cost SDXL: Improving Latent Diffusion Models for High-Resolution Image Synthesis

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-12T04:34:52.568696Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:34:52.568696Z digest=sha256:fc2a819455b02e412e69c5bbc2e8f539479a3dda826cedfc286361d67f4b497d

Observation fa05039f-d798-47e6-b5d5-9b492e17126c · outbound

This paper cites Gluegen: Plug and play multi-modal encoders for x-to-image generation.

MuLan: Adapting Multilingual Diffusion Models for Hundreds of Languages with Negligible Cost Gluegen: Plug and play multi-modal encoders for x-to-image generation

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:34:53.135969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T04:34:52.572401Z digest=sha256:521a2ef8503ce110a5bce8bb73f0c672a4cd468739ba4478196378f79a2ad63e

Observation 57b983e2-d39e-4044-8e37-ff5ae3244428 · outbound

This paper cites W., Hallacy, C., Ramesh, A., Goh, G., Agarwal, S., Sastry, G., Askell, A., Mishkin, P., Clark, J., et al.

MuLan: Adapting Multilingual Diffusion Models for Hundreds of Languages with Negligible Cost W., Hallacy, C., Ramesh, A., Goh, G., Agarwal, S., Sastry, G., Askell, A., Mishkin, P., Clark, J., et al

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-12T04:34:52.575738Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:34:52.575738Z digest=sha256:896327c385313cb3cb5756358a5aa4afd282757802ab9cfee13480dd6b359b8d

Observation 455c5a27-234c-4d3d-b1b0-8239afdf7aff · outbound

This paper cites an unresolved cited work.

MuLan: Adapting Multilingual Diffusion Models for Hundreds of Languages with Negligible Cost Unresolved cited work

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-12T04:34:52.579213Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:34:52.579213Z digest=sha256:5b8f56a192afaa1954d6df57b13ea9c14524f3ffd48fc45ce6f1b28486b31a22

Observation 5b539808-6bc9-46d3-bd61-3d6b6179bd18 · outbound

This paper cites Zero-shot text-to-image generation.

MuLan: Adapting Multilingual Diffusion Models for Hundreds of Languages with Negligible Cost Zero-shot text-to-image generation

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-12T04:34:52.582784Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:34:52.582784Z digest=sha256:aa2d7653241d811a0147bfd65b3ec40ba884a6d97fe9ee4fc637e913ef98fd69

Observation 5d82bcc6-9bfd-4af9-ac38-c98ca71f7f42 · outbound

This paper cites Hierarchical Text-Conditional Image Generation with CLIP Latents.

MuLan: Adapting Multilingual Diffusion Models for Hundreds of Languages with Negligible Cost Hierarchical Text-Conditional Image Generation with CLIP Latents

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-12T04:34:52.586197Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:34:52.586197Z digest=sha256:14e03468a675b92130a2c3e84867d95f5a4605e13be22b29585cdcd934b73321

Observation 9875caab-4e9d-491e-bf10-b8d4f114399e · outbound

This paper cites High-resolution image synthesis with latent diffusion models.

MuLan: Adapting Multilingual Diffusion Models for Hundreds of Languages with Negligible Cost High-resolution image synthesis with latent diffusion models

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-12T04:34:52.589684Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:34:52.589684Z digest=sha256:191d0bdab05f8e9f10f40ad9890fa08dc5692a112c758f8ae627b809bd0f25a1

Observation 8a77dc0b-746e-4fa6-b990-547d15dcb405 · outbound

This paper cites Photorealistic Text-to-Image Diffusion Models with Deep Language Understanding.

MuLan: Adapting Multilingual Diffusion Models for Hundreds of Languages with Negligible Cost Photorealistic Text-to-Image Diffusion Models with Deep Language Understanding

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-12T04:34:52.593189Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:34:52.593189Z digest=sha256:2cec808f334ddfa758acb431361b70a59f75864c37e9b1e824f003b33e149ea1

Observation 0729244f-b06b-4311-a8b1-d83c5d583342 · outbound

This paper cites LAION-5B: An open large-scale dataset for training next generation image-text models.

MuLan: Adapting Multilingual Diffusion Models for Hundreds of Languages with Negligible Cost LAION-5B: An open large-scale dataset for training next generation image-text models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-12T04:34:52.597013Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:34:52.597013Z digest=sha256:c531521451635e2e8469e7639a433023aedb77361df8e5dbe2372a7118cc7144

Observation 16ea769d-1537-4adc-991d-3f561cf8fbf5 · outbound

This paper cites CCMatrix: Mining Billions of High-Quality Parallel Sentences on the WEB.

MuLan: Adapting Multilingual Diffusion Models for Hundreds of Languages with Negligible Cost CCMatrix: Mining Billions of High-Quality Parallel Sentences on the WEB

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-12T04:34:52.600901Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:34:52.600901Z digest=sha256:68b50d7a0221bd11474f628d41e9753ddf348e327d3622bec8c5133a4fc6cbb3

Observation 775c2cd2-0160-4a56-b9e6-736dd176ed34 · outbound

This paper cites MVDream: Multi-view Diffusion for 3D Generation.

MuLan: Adapting Multilingual Diffusion Models for Hundreds of Languages with Negligible Cost MVDream: Multi-view Diffusion for 3D Generation

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-12T04:34:52.604664Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:34:52.604664Z digest=sha256:bde21a4eb0d9330b0331309c36c2dc2e06014cf2cc994598e39e0d9138b9d731

Observation 57803d5b-b49d-4c62-a1ee-8d82635dc7df · outbound

This paper cites and Akiba, T.

MuLan: Adapting Multilingual Diffusion Models for Hundreds of Languages with Negligible Cost and Akiba, T

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:34:53.097370Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T04:34:52.608279Z digest=sha256:38407f0ced5d51726dac6e2a7912967febbc67344ec4ebc3a154397aff00f80b

Observation da5e9098-5c60-4da1-a0bd-35de5b3d1f96 · outbound

This paper cites Recent advances in implicit representation-based 3d shape generation.

MuLan: Adapting Multilingual Diffusion Models for Hundreds of Languages with Negligible Cost Recent advances in implicit representation-based 3d shape generation

Reference 41

Resolution
verified exact
doi, observed 2026-08-12T04:34:52.703561Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T04:34:52.615680Z digest=sha256:2fb15f963b507cbb4a329bc3073490db7d7923bfe2c3059f0c9a5dcef464b57e

Observation 6e3a6694-cd28-474b-abd7-eb653d0d66a8 · outbound

This paper cites Kolors: Effective training of diffusion model for photorealistic text-to-image synthesis.

MuLan: Adapting Multilingual Diffusion Models for Hundreds of Languages with Negligible Cost Kolors: Effective training of diffusion model for photorealistic text-to-image synthesis

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:34:53.086055Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T04:34:52.619693Z digest=sha256:199545047d4f3bd0470132dcb0355fb4ecff59f76ce90c986824839dfbdd6c6e

Observation f6886cba-c109-4b34-bfc5-7780d8c37c13 · outbound

This paper cites Crossmodal-3600: A Massively Multilingual Multimodal Evaluation Dataset.

MuLan: Adapting Multilingual Diffusion Models for Hundreds of Languages with Negligible Cost Crossmodal-3600: A Massively Multilingual Multimodal Evaluation Dataset

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-12T04:34:52.623249Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:34:52.623249Z digest=sha256:ff72f84f3d6c5c4531e3b142afa6b4cd893d2d479640d5341bd5dc1e71d2d2a6

Observation 0ee658da-171d-49de-902a-5c16e0da6f7e · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

MuLan: Adapting Multilingual Diffusion Models for Hundreds of Languages with Negligible Cost Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-12T04:34:52.626969Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:34:52.626969Z digest=sha256:ffb0a4e5ca53b09d9a5359024b4ace3b360deeb6733e325906dcec31f27586d2

Observation f54334f3-3391-4cc4-82c8-d4163daf6f6d · outbound

This paper cites and Hinton, G.

MuLan: Adapting Multilingual Diffusion Models for Hundreds of Languages with Negligible Cost and Hinton, G

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:34:53.074109Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T04:34:52.630533Z digest=sha256:e4779921cf92f6c195ee20b23c8ea185f4fc0cad43dd6a27052d0882ca888fe9

Observation 3b1f2fc2-9e3a-40d2-8ff8-73f51c9c8ca9 · outbound

This paper cites Taiyi-Diffusion-XL: Advancing Bilingual Text-to-Image Generation with Large Vision-Language Model Support.

MuLan: Adapting Multilingual Diffusion Models for Hundreds of Languages with Negligible Cost Taiyi-Diffusion-XL: Advancing Bilingual Text-to-Image Generation with Large Vision-Language Model Support

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-12T04:34:52.634309Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:34:52.634309Z digest=sha256:ef96e7b05fef47af368ee58b14d160c656f67c91ae13a673b1b8143a1d85e1fa

Observation cd3e66fb-ef2a-404f-ac23-a383f17a84c6 · outbound

This paper cites mT5: A massively multilingual pre-trained text-to-text transformer.

MuLan: Adapting Multilingual Diffusion Models for Hundreds of Languages with Negligible Cost mT5: A massively multilingual pre-trained text-to-text transformer

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-12T04:34:52.638245Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:34:52.638245Z digest=sha256:97cbdc6f6b6d622f3fff5bfc7cbbaa66f175f2a4e1b806ec3b66fb5e09f720a6

Observation e7564aa5-5d81-4f78-8d26-266640dbd62d · outbound

This paper cites Dialoguenerf: towards realistic avatar face-to-face conversation video generation.

MuLan: Adapting Multilingual Diffusion Models for Hundreds of Languages with Negligible Cost Dialoguenerf: towards realistic avatar face-to-face conversation video generation

Reference 48

Resolution
verified exact
doi, observed 2026-08-12T04:34:52.691656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T04:34:52.642166Z digest=sha256:39f0a901ce01de5ab686442b37d81761c374323254d7fce7f9d27c39df1b7f40

Observation 359b6738-1e45-4e29-9079-5fe99e6c0873 · outbound

This paper cites AltDiffusion: A Multilingual Text-to-Image Diffusion Model.

MuLan: Adapting Multilingual Diffusion Models for Hundreds of Languages with Negligible Cost AltDiffusion: A Multilingual Text-to-Image Diffusion Model

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-12T04:34:52.646040Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:34:52.646040Z digest=sha256:e975d60d0505bcd968096d823c7dd8edd6ab1408201595f74517d2371b18b165

Observation 5a4ded44-f864-4432-8b86-8df506542cf5 · outbound

This paper cites Ip-adapter: Text compatible image prompt adapter for text-to-image diffusion models.

MuLan: Adapting Multilingual Diffusion Models for Hundreds of Languages with Negligible Cost Ip-adapter: Text compatible image prompt adapter for text-to-image diffusion models

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T04:34:53.062083Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=arxiv_source observed=2026-08-12T04:34:52.650156Z digest=sha256:710fa51132812ddd6233ac3ac48f4f5fcc8e6e2781c734e261381f93bd32fe85

Observation 64d6d7f1-557b-48c4-b287-0ba79fa3c8f3 · outbound

This paper cites Fengshenbang 1.0: Being the Foundation of Chinese Cognitive Intelligence.

MuLan: Adapting Multilingual Diffusion Models for Hundreds of Languages with Negligible Cost Fengshenbang 1.0: Being the Foundation of Chinese Cognitive Intelligence

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-12T04:34:52.653845Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:34:52.653845Z digest=sha256:7a630f827ee27f2b0a7d5cba9616a710cdcd7886f95c3559cf6f6e6602002214

Observation dfc0ca5d-8e6d-40c0-bbfc-00013c900db6 · outbound

This paper cites Adding conditional control to text-to-image diffusion models.

MuLan: Adapting Multilingual Diffusion Models for Hundreds of Languages with Negligible Cost Adding conditional control to text-to-image diffusion models

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-12T04:34:52.657768Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T04:34:52.657768Z digest=sha256:283bec54153a01173b7cacfcdd632db104921a8c862e8423a9a0b532c45481a8

Pith citing papers

Observation 89843210-b96e-4734-9fc1-afe37ee1c9b6 · inbound

IMAGINE-E: Image Generation Intelligence Evaluation of State-of-the-art Text-to-Image Models cites this paper.

IMAGINE-E: Image Generation Intelligence Evaluation of State-of-the-art Text-to-Image Models MuLan: Adapting Multilingual Diffusion Models for Hundreds of Languages with Negligible Cost

Reference 78

Resolution
verified exact
local_arxiv, observed 2026-08-10T15:31:35.919905Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T15:31:35.751516Z digest=sha256:c60a4db56d4662483cc689504d69b053f48cb3a830b36187d6f1487b5c93c9f1