Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T05:45:54.168593Z
Paper Citation Record · LEDGER
As of 13 August 2026, this Paper Citation Record lists 39 of 39 outbound references and 4 inbound Pith citation observations for arXiv:2412.17225.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T05:45:54.168593Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-01T05:51:12.656627Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-03T21:28:58.246299Z
39 of 39 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 32dcee00-c070-47fe-a50a-c395288bfe3c · outbound
CharGen: High Accurate Character-Level Visual Text Generation Model with MultiModal Encoder Improving Image Gen- eration with Better Captions
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation d644f7f7-218a-4241-8348-8e52a477c2d7 · outbound
CharGen: High Accurate Character-Level Visual Text Generation Model with MultiModal Encoder Freeman, Michael Rubinstein, Yuanzhen Li, and Dilip Krishnan
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation aa59fca9-478b-42bf-ad77-66a533438e79 · outbound
CharGen: High Accurate Character-Level Visual Text Generation Model with MultiModal Encoder Diffute: Universal text editing diffusion model
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation af6c4dbe-363e-4cd0-a785-3fd33ceaf91e · outbound
CharGen: High Accurate Character-Level Visual Text Generation Model with MultiModal Encoder TextDiffuser-2: Unleashing the Power of Language Models for Text Rendering
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5be04580-7bc3-41bf-a3b5-123cfd3eeea4 · outbound
CharGen: High Accurate Character-Level Visual Text Generation Model with MultiModal Encoder Textdiffuser: Diffusion models as text painters
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 0f6a3fba-73d8-4b6f-9915-8b934c5132e1 · outbound
CharGen: High Accurate Character-Level Visual Text Generation Model with MultiModal Encoder Diffusion Models Beat GANs on Image Synthesis
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a105cabf-43bc-426b-b935-3f12b69ce34c · outbound
CharGen: High Accurate Character-Level Visual Text Generation Model with MultiModal Encoder Odm: A text-image further alignment pre-training approach for scene text detection and spotting
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation a33befb2-80d5-47d5-98fd-700c80dbde74 · outbound
CharGen: High Accurate Character-Level Visual Text Generation Model with MultiModal Encoder Duguangocr
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 4fdb2fe9-d5c1-46df-a71d-66b3a84b97ba · outbound
CharGen: High Accurate Character-Level Visual Text Generation Model with MultiModal Encoder Scaling recti- fied flow transformers for high-resolution image synthesis
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f6730221-9216-40c4-bccd-9301fa904acb · outbound
CharGen: High Accurate Character-Level Visual Text Generation Model with MultiModal Encoder Wukong: 100 million large-scale chinese cross-modal pre-training dataset and a foundation frame- work, 2022
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 78774e74-34d5-4000-bd64-99259f3d4ee9 · outbound
CharGen: High Accurate Character-Level Visual Text Generation Model with MultiModal Encoder Denoising Dif- fusion Probabilistic Models, 2020
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 87be2020-c36d-43de-ad5c-6df12c40f426 · outbound
CharGen: High Accurate Character-Level Visual Text Generation Model with MultiModal Encoder Auto-encoding varia- tional bayes, 2022
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 11d9325a-7517-4838-9095-db005ac98a92 · outbound
CharGen: High Accurate Character-Level Visual Text Generation Model with MultiModal Encoder Pp-ocrv3: More attempts for the improvement of ultra lightweight ocr system,
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 4a50ce3e-f731-474e-8afe-a05b6511c937 · outbound
CharGen: High Accurate Character-Level Visual Text Generation Model with MultiModal Encoder Character-aware models improve visual text rendering, 2023
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation ba0c7ce2-e223-4f68-b8c1-413f868882c1 · outbound
CharGen: High Accurate Character-Level Visual Text Generation Model with MultiModal Encoder Character-Aware Models Improve Visual Text Rendering
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5708ea8d-5278-4cc3-8bc1-46d4e43f1593 · outbound
CharGen: High Accurate Character-Level Visual Text Generation Model with MultiModal Encoder Glyph-ByT5: A Cus- tomized Text Encoder for Accurate Visual Text Rendering,
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 0c45c0d0-fc72-4ab9-b27b-a285b2cdd64f · outbound
CharGen: High Accurate Character-Level Visual Text Generation Model with MultiModal Encoder Glyph-ByT5-v2: A Strong Aesthetic Baseline for Accurate Multilingual Visual Text Rendering
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2ecb5dfd-4802-4cc8-b0dc-60425446d0bd · outbound
CharGen: High Accurate Character-Level Visual Text Generation Model with MultiModal Encoder GlyphDraw: Seamlessly Rendering Text with Intricate Spatial Structures in Text-to-Image Generation
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ad03fc6a-a382-4c15-9bc4-a7aba992e08e · outbound
CharGen: High Accurate Character-Level Visual Text Generation Model with MultiModal Encoder Glyphdraw2: Automatic generation of complex glyph posters with diffusion models and large language models,
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 5e344d80-5f47-487b-9baf-fec339db194f · outbound
CharGen: High Accurate Character-Level Visual Text Generation Model with MultiModal Encoder Midjourney
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 37c395c1-3718-43a0-91ce-d99a458aecaf · outbound
CharGen: High Accurate Character-Level Visual Text Generation Model with MultiModal Encoder Improved denoising dif- fusion probabilistic models, 2021
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation fe117036-b7ff-4752-85c2-25ea340eb4ee · outbound
CharGen: High Accurate Character-Level Visual Text Generation Model with MultiModal Encoder Learning transferable visual models from natural language supervi- sion
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 098b4e94-2f61-4228-a1b5-6d8025581694 · outbound
CharGen: High Accurate Character-Level Visual Text Generation Model with MultiModal Encoder Exploring the limits of transfer learning with a unified text-to-text transformer
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 59b35edd-5e0c-46be-8542-5ed4add82c13 · outbound
CharGen: High Accurate Character-Level Visual Text Generation Model with MultiModal Encoder Zero-shot text-to-image generation, 2021
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f74ce017-246f-456e-85b6-1d9d3748c3ed · outbound
CharGen: High Accurate Character-Level Visual Text Generation Model with MultiModal Encoder Hierarchical text-conditional image gener- ation with clip latents, 2022
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8f16315b-4474-4658-a26c-796c1c0801fe · outbound
CharGen: High Accurate Character-Level Visual Text Generation Model with MultiModal Encoder Ocr-vqgan: Taming text- within-image generation
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 18160161-c409-440f-88f1-a2b150a5e58c · outbound
CharGen: High Accurate Character-Level Visual Text Generation Model with MultiModal Encoder High-resolution image synthesis with latent diffusion models
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 2b65121d-7aa4-46a7-8d28-d0f51b1b9144 · outbound
CharGen: High Accurate Character-Level Visual Text Generation Model with MultiModal Encoder Photorealistic Text-to-Image Diffusion Models with Deep Language Understanding
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a1522277-adbe-44d6-8b8e-e58ff6b58c12 · outbound
CharGen: High Accurate Character-Level Visual Text Generation Model with MultiModal Encoder LAION-400M: Open Dataset of CLIP-Filtered 400 Million Image-Text Pairs
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3dfd3bf8-f8d6-4cdb-a245-7165c583f528 · outbound
CharGen: High Accurate Character-Level Visual Text Generation Model with MultiModal Encoder Kingma, Ab- hishek Kumar, Stefano Ermon, and Ben Poole
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0d02f40a-f4d3-4215-9539-a03046be485f · outbound
CharGen: High Accurate Character-Level Visual Text Generation Model with MultiModal Encoder AnyText: Multilingual Visual Text Generation And Editing
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5c17193d-3251-4244-af5a-c4db9ddaa9a3 · outbound
CharGen: High Accurate Character-Level Visual Text Generation Model with MultiModal Encoder Glyphcontrol: Glyph conditional control for visual text generation
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 2d8c04b4-8d3e-46c3-a1c3-eee0ee9c063d · outbound
CharGen: High Accurate Character-Level Visual Text Generation Model with MultiModal Encoder Long-clip: Unlocking the long-text capability of clip, 2024
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation f03aad97-a15e-49a2-ad58-e27632021fe1 · outbound
CharGen: High Accurate Character-Level Visual Text Generation Model with MultiModal Encoder Adding conditional control to text-to-image diffusion models
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation c02d6877-585f-4d82-ab71-d409ae8d0b2b · outbound
CharGen: High Accurate Character-Level Visual Text Generation Model with MultiModal Encoder Adding Conditional Control to Text-to-Image Diffusion Models,
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 40ef395d-b24b-4269-8b3e-ad783d5a255e · outbound
CharGen: High Accurate Character-Level Visual Text Generation Model with MultiModal Encoder Brush your text: Synthesize any scene text on im- ages via diffusion model
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 13b981e2-4f9b-4bc2-9c69-9a6572dca6a3 · outbound
CharGen: High Accurate Character-Level Visual Text Generation Model with MultiModal Encoder UDiffText: A Unified Framework for High-quality Text Synthesis in Arbitrary Images via Character-aware Diffusion Models
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation bfd22a16-6f0b-4425-9200-70da3b6ed929 · outbound
CharGen: High Accurate Character-Level Visual Text Generation Model with MultiModal Encoder Adding Conditional Control to Text-to-Image Diffusion Models
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fa5c5579-92f7-4b53-9ea1-4408977906f0 · outbound
CharGen: High Accurate Character-Level Visual Text Generation Model with MultiModal Encoder Glyph-ByT5: A Customized Text Encoder for Accurate Visual Text Rendering
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 800ea31d-774f-4249-835e-68c4aa3970ec · inbound
Holding the FP8 Quality Ceiling at 8-Bit Weights and Activations: INT8 and GGUF Post-Training Quantization of Ideogram 4.0 for Consumer GPUs CharGen: High Accurate Character-Level Visual Text Generation Model with MultiModal Encoder
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 8daa24de-0235-4b42-a24e-0ddd934bb10e · inbound
GMO-E$^2$DIT: Grounded Multi-Operation Editing for E-Commerce Images CharGen: High Accurate Character-Level Visual Text Generation Model with MultiModal Encoder
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation f9fc143e-5a29-4ec9-8905-b2e272ea7998 · inbound
GMO-E$^2$DIT: Grounded Multi-Operation Editing for E-Commerce Images CharGen: High Accurate Character-Level Visual Text Generation Model with MultiModal Encoder
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation df547916-2f60-4f9e-a17d-0164114a8eb5 · inbound
InnoText: A Unified Model for Visual Text Generation and Editing CharGen: High Accurate Character-Level Visual Text Generation Model with MultiModal Encoder
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.