Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T15:39:32.833572Z
Paper Citation Record · LEDGER
As of 19 August 2026, this Paper Citation Record lists 96 of 96 outbound references and 1 inbound Pith citation observation for arXiv:2411.14062.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T15:39:32.833572Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-06T18:50:43.512312Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-06T18:50:46.335683Z
96 of 96 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation f6c7de02-b8f7-4689-8b48-9fcb1a5c0988 · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective Phi-3 technical report: A highly capable language model locally on your phone, 2024
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b48fb7d6-b12b-4d2b-a6e1-c079d9f10e26 · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective Lawrence Zitnick, Dhruv Batra, and Devi Parikh
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6b13b67a-2826-4d42-a277-939c243e4e08 · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective Pixtral 12b, 2024
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 88503e0e-eb56-4065-834c-333684f824a2 · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective Flamingo: a visual language model for few-shot learning
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 620f42f1-e921-4387-b5af-0bea0aa64520 · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective Unicom: Universal and compact representation learning for image re- trieval, 2023
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 392d676f-d717-43b3-a963-6a4e784f0ad4 · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5a3cb3c3-c7b5-4b63-8759-2e2c9483f6ca · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective Benchmarking foundation models with language- model-as-an-examiner
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3a569c53-eca4-41cc-aed1-e0e13b528f6a · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective AutoBench-V: Can Large Vision-Language Models Benchmark Themselves?
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 34af97cc-c653-4fc4-9f3e-53580ba164c6 · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective Are We on the Right Way for Evaluating Large Vision-Language Models?
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4d426660-b006-4629-9d50-508e9e74dc1b · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective InternVL: Scaling up Vision Foundation Models and Aligning for Generic Visual-Linguistic Tasks
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 01337695-0e8e-4add-bfb0-1d19d00479f6 · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e25963b0-5c21-4047-95e0-4209caa29209 · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective MobileVLM V2: Faster and Stronger Baseline for Vision Language Model
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 89cf1586-6f20-4a4f-a2c2-8eafcca0e571 · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective Opencompass: A universal evaluation platform for foundation models
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1761d266-6cdf-430a-90a2-27df1ace2ead · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective Molmo and pixmo: Open weights and open data for state-of-the-art vision-language models, 2024
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d8c43b71-31f3-4880-a221-e0fbd37bd55e · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective Vlmevalkit: An open- source toolkit for evaluating large multi-modality models,
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1b4b3530-0bfb-4f77-ae21-051394a1d927 · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective Scaling recti- fied flow transformers for high-resolution image synthesis
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e2144318-a87c-4cf8-af47-a01a5c975885 · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 043b6139-149a-47eb-acba-cb8576dc294f · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective Ocrbench v2: An improved benchmark for evaluating large multimodal models on visual text localization and reasoning, 2024
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4a7cc438-a291-4348-a1f4-bfaaa58aa2c2 · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective Smith, Wei-Chiu Ma, and Ranjay Krishna
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bf3e3257-aaa7-48c2-a71c-89d72ecb73bd · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective Lumina-t2x: Transforming text into any modality, resolution, and dura- tion via flow-based large diffusion transformers, 2024
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c5ece020-f293-4eae-85a2-43dcce0cec50 · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective Mini-InternVL: A Flexible-Transfer Pocket Multimodal Model with 5% Parameters and 90% Performance
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c78fc9af-9c35-49fb-9379-e274a6031276 · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective Chatglm: A family of large language models from glm-130b to glm-4 all tools, 2024
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 24d9c418-a3c1-4770-8901-12ff2e6cc96d · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective Making the v in vqa matter: Elevating the role of image understanding in visual question answer- ing
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e87e963a-4527-4e4c-807f-9faec4e66514 · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective Olympiadbench: A challenging benchmark for promoting agi with olympiad-level bilingual multimodal scientific problems, 2024
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c52fba38-7d99-4a5c-a8d5-d72edc5df881 · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective Denoising dif- fusion probabilistic models
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 65865f0e-ac7c-437a-ad06-8ba98191d52c · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective Cogvlm2: Visual language models for image and video understanding, 2024
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 928cb6f8-4a69-42c3-beb4-245f80d879fc · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective Chatgpt for shaping the future of 9 dentistry: the potential of multi-modal large language model
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation b98100ee-c7ff-44be-b245-1c10baff7f79 · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective Genmac: Compositional text-to-video generation with multi-agent collaboration, 2024
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 998299a5-2518-461b-bb68-19aa7336bed3 · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective Mini-monkey: Alleviating the semantic saw- tooth effect for lightweight mllms via complementary image pyramid, 2024
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 994807de-2d76-408d-b19b-83b065a3f60f · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective Hudson and Christopher D
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ded9f1a8-8f81-4d6a-955b-a09e68a7221a · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective Ku, Qian Liu, and Wenhu Chen
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 499efcef-457d-4012-b8ca-bba28fb504b1 · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective Chatgpt for good? on opportuni- ties and challenges of large language models for education
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 4b9286f8-31cc-40c9-842e-31262017dad3 · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective Reflective decoding network for image captioning
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 93d3472c-f64a-45aa-9faf-dba61c362afb · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective Unresolved cited work
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation d4e2799c-b6c2-4394-915b-400f3e2f62d7 · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective Building and better understanding vision- language models: insights and future directions., 2024
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation b745b877-e7ee-44cc-9cd5-20d6df8ef4ca · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective What matters when building vision-language models?,
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation abfeeb7c-77c2-49e9-9b15-1d47c90208bb · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective Llava-onevision: Easy visual task transfer, 2024
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation e4660870-ddae-4088-a8f4-b2bcb5c4e527 · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective AutoBencher: Towards Declarative Benchmark Construction
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 603b06ca-39c5-4251-94fd-94203ff1d0cb · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective Llm-grounded video diffusion models, 2024
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 894a6aa1-2230-4cdf-a99e-f594e9c768f1 · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective Vila: On pre-training for visual language models, 2023
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 97984e33-8afc-4b0f-b019-ec6c0f30fff3 · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective Microsoft coco: Common objects in context
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation d8f55a43-6f18-48cd-87bc-5e13ab9587a1 · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective Improved baselines with visual instruction tuning
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation efbb7fb3-2d82-47db-beac-e355f76519c8 · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective Llava-next: Im- proved reasoning, ocr, and world knowledge, 2024
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 6ec08b7a-98d7-4311-a264-2f5311eaa607 · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective Tempcom- pass: Do video llms really understand videos?, 2024
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 72df9ccf-b402-4686-831d-2b7fa24dd0f9 · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective Ocrbench: On the hidden mystery of ocr in large multimodal models, 2024
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a0b82f83-739c-4368-befa-a19ba47778dc · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective Mmbench: Is your multi-modal model an all-around player? In European Conference on Computer Vision, pages 216–233
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation de84a044-9868-4e8b-85c4-d245ff68e623 · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective Mmdu: A multi-turn multi-image dia- log understanding benchmark and instruction-tuning dataset for lvlms, 2024
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 6b2a2f7e-e0dc-4e9b-830f-9fc7de4a394c · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective Mmalaya2
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 19f4c7ed-d0e8-4f7f-a8f7-e97484ad952e · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective Mathvista: Evaluating mathematical reasoning of foundation models in visual contexts, 2024
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 1fefffe4-841f-4476-ad1b-aab72a973087 · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective Ovis: Structural Embedding Alignment for Multimodal Large Language Model
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dad6b72c-ab48-4ee2-b150-6c4992ebac98 · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective Mmlongbench-doc: Bench- marking long-context document understanding with visual- izations, 2024
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 99adf939-6ebf-4b71-abd2-739a3ea8e914 · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective The llama 3 herd of models, 2024
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation dc7f92d1-c0ea-4920-86d2-78b33d333f96 · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective Gpt-4o system card, 2024
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation f6c63cb7-7aef-443b-a138-e5014b289879 · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective GPT-4 Technical Report
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eec231ab-3931-4a8f-8ee7-b49b1d93bb09 · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective Omnidocbench: Benchmarking diverse pdf document parsing with comprehensive annota- tions, 2024
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 451e32c6-30bd-46f2-af7d-313d7bfb8822 · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective Sowing information: Cultivating con- textual coherence with mllms in image generation, 2024
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation af72decb-b80c-483c-93b6-e71f39f199fd · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective Rbdash-v1.2-72b
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 0e3ad64d-18be-49fe-8188-374e0702de0b · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective High-resolution image synthesis with latent diffusion models
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 2e14a600-424d-46e1-ae8d-9f53737c730d · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective Eagle: Exploring the design space for multimodal llms with mixture of encoders, 2025
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation c4f332b9-96ce-415e-8db0-7f0afaeb2b0d · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective Journeydb: A benchmark for generative im- age understanding
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation b7d4203c-012c-4d6f-a4ce-96b627f69fb4 · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective Gemini: A Family of Highly Capable Multimodal Models
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7912cc58-07eb-4614-aa15-2a5fd20f5c43 · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective Kolors: Effective training of diffusion model for photorealistic text-to-image synthesis
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ca0b0a84-17aa-422e-9259-f029bb20aeaa · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective Cambrian-1: A Fully Open, Vision-Centric Exploration of Multimodal LLMs
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e0680d9d-1c65-4540-8eff-861fabde1c63 · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective Llama-3-mixsensev1 1
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation b9e32ecc-4031-4d02-b48e-45a77f86cbea · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 77e1f016-8d2a-4f80-bf98-81ea93970981 · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective Large-scale multi-modal pre-trained models: A comprehensive survey
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation d911fd21-c8b9-428b-b48d-0bd46701a678 · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective Unresolved cited work
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 7530f209-eea4-4070-a986-39e837a3f2e5 · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective BloombergGPT: A Large Language Model for Finance
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 10ba1df2-2978-4223-94fa-ad2afe866761 · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective Unigen: A unified framework for textual dataset generation using large language models
Reference 69
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ec0b6ef1-fb73-4fd3-8e0b-71d15839f9ec · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective Self-correcting llm-controlled diffu- sion models
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 90e9bae8-6ace-4529-8971-259e9c382a40 · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective Unresolved cited work
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation b007837a-dc2f-44b1-b02d-f0988bbd5bd7 · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective LVLM-eHub: A Comprehensive Evaluation Benchmark for Large Vision-Language Models
Reference 72
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1017213b-7fe3-4803-b0f4-1117cc6fb4bb · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective xgen-mm (blip-3): A family of open large multimodal models, 2024
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 2eff9fe7-f24f-4fac-a20e-ae515ea03c3c · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective Cc-ocr: A comprehensive and challenging ocr benchmark for evalu- ating large multimodal models in literacy, 2024
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 5a146a39-7414-4757-a6de-acab5baa365d · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective MiniCPM-V: A GPT-4V Level MLLM on Your Phone
Reference 75
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7df529b8-8387-446e-9b09-be541edef8e3 · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective Lamm: Language-assisted multi-modal instruction-tuning dataset, framework, and benchmark
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 6aa2bdd9-6a97-435c-b870-3ba4ce4dd882 · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective Benchmarking chinese text recognition: Datasets, baselines, and an empirical study, 2022
Reference 77
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 695ee012-2c3b-46ae-bbba-e70926fb0b97 · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective MM-Vet: Evaluating Large Multimodal Models for Integrated Capabilities
Reference 78
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ec4128b2-cdd5-4f5d-b7a5-01df68b31685 · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective Mm-vet: Evaluating large multimodal models for integrated capabilities, 2023
Reference 79
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation d00654b8-e5bc-4326-a09e-6b903bbc1853 · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective Task Me Anything
Reference 80
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3e43a91c-b0d8-44d8-b76c-198276ae641d · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective LLaMA-Adapter: Efficient Fine-tuning of Language Models with Zero-init Attention
Reference 81
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b12178ac-ae59-4a29-a9c0-45ea07ced1ab · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective Omchat: A recipe to train multimodal language models with strong long context and video under- standing, 2024
Reference 82
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 254e35a6-d337-45af-8368-ab56ccc548e9 · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective Dyval: Dynamic evalua- tion of large language models for reasoning tasks
Reference 83
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 5a8b3be0-124b-4f00-980f-eceea2f9ad56 · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective role”, “definition
Reference 84
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 2a295e12-c4d6-4c4a-aec9-35beabeab30d · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective image pattern
Reference 86
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 5f766ab4-efdc-4edb-b508-1d5d5eea9f56 · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective Unresolved cited work
Reference 87
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 9fe9aa89-99dd-4282-9ab5-ad91dd43825c · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective Surreal”: 2262, “Lighting
Reference 88
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 165296a6-f543-4b96-9ec8-fe604d315a75 · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective Unresolved cited work
Reference 89
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation e18433df-7137-4cd0-9fe9-44f9097ecb03 · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective # Key Points
Reference 90
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 7d7f9315-0a39-49e0-91fc-20d079ce331d · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective Unresolved cited work
Reference 91
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation ff20a192-012f-46d2-9461-884fdf8d3c61 · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective You may annotate multiple patterns as appropriate
Reference 92
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation cb493ce5-771e-4a9d-a2a8-cd789f8efacb · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective Surreal”: “This pattern is characterized by its prevalence in depicting scenes that mix elements of fantasy with reality, often creating imaginative or dream-like visuals
Reference 93
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 8a05b0ac-0152-48d7-9186-ce600dc46680 · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective Unresolved cited work
Reference 94
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 35228db3-7b17-4e64-868b-e644dd2c0fb4 · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective Unresolved cited work
Reference 95
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 5da42579-5ce4-4599-b205-0e82eb4f5245 · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective Unresolved cited work
Reference 96
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation a416d5f0-2a35-4b0e-946e-751778cb5ece · outbound
MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective BHNORAK TOP
Reference 97
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 17068afa-0c71-4cec-9951-a0b0225303ec · inbound
Towards Evaluating Robustness of Prompt Adherence in Text to Image Models MMGenBench: Fully Automatically Evaluating LMMs from the Text-to-Image Generation Perspective
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.