Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T11:26:29.251213Z
Paper Citation Record · LEDGER
As of 11 August 2026, this Paper Citation Record lists 36 of 36 outbound references and 2 inbound Pith citation observations for arXiv:2501.16688.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T11:26:29.251213Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T00:40:53.903009Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-10T23:00:50.182088Z
36 of 36 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 56dee464-c6dc-4e45-94ae-38a77f8ab241 · outbound
MME-Industry: A Cross-Industry Multimodal Evaluation Benchmark GPT-4 Technical Report
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0ee1f056-df51-40de-a25d-1334ec1bf855 · outbound
MME-Industry: A Cross-Industry Multimodal Evaluation Benchmark Qwen Technical Report
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7541bcbd-747b-4a1e-949a-f766a4eb5ec7 · outbound
MME-Industry: A Cross-Industry Multimodal Evaluation Benchmark Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6a476e83-9698-48be-bf23-0ffde00e9e27 · outbound
MME-Industry: A Cross-Industry Multimodal Evaluation Benchmark Are We on the Right Way for Evaluating Large Vision-Language Models?
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 082f44b9-beeb-4f8b-adc0-9a808a21cd22 · outbound
MME-Industry: A Cross-Industry Multimodal Evaluation Benchmark InternLM-XComposer2: Mastering Free-form Text-Image Composition and Comprehension in Vision-Language Large Model
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7b7b239c-115f-466d-b19a-0da33d7cbb05 · outbound
MME-Industry: A Cross-Industry Multimodal Evaluation Benchmark PaLM-E: An Embodied Multimodal Language Model
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e58c7726-ab28-4d01-867d-e240faf4511c · outbound
MME-Industry: A Cross-Industry Multimodal Evaluation Benchmark Vlmevalkit: An open-source toolkit for evaluating large multi-modality models,
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation f59c7468-ccd1-43b9-b979-c4bd9fb569cb · outbound
MME-Industry: A Cross-Industry Multimodal Evaluation Benchmark MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 71da6fb4-77e6-47e2-9fa3-fe8291ac4b4e · outbound
MME-Industry: A Cross-Industry Multimodal Evaluation Benchmark ChatGLM: A Family of Large Language Models from GLM-130B to GLM-4 All Tools
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 82b174a3-5d82-4411-a330-5a0c2363df82 · outbound
MME-Industry: A Cross-Industry Multimodal Evaluation Benchmark LoRA: Low-Rank Adaptation of Large Language Models
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation da99e05f-54cd-4157-bac8-f596a5559318 · outbound
MME-Industry: A Cross-Industry Multimodal Evaluation Benchmark Blip-2: Bootstrapping language-image pre- training with frozen image encoders and large language models
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation a9ae5d96-87d5-4fa8-915e-3fcf61f6979f · outbound
MME-Industry: A Cross-Industry Multimodal Evaluation Benchmark Monkey: Image resolution and text label are important things for large multi-modal models
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 945af6d2-102f-47e0-84a2-4a0a090dafa8 · outbound
MME-Industry: A Cross-Industry Multimodal Evaluation Benchmark MMBench: Is Your Multi-modal Model an All-around Player?
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e83c63da-d2a5-43fd-98e7-9449af5fd5e3 · outbound
MME-Industry: A Cross-Industry Multimodal Evaluation Benchmark OCRBench: On the Hidden Mystery of OCR in Large Multimodal Models
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7b675fb5-e226-4e93-853a-ee05c37b0b8b · outbound
MME-Industry: A Cross-Industry Multimodal Evaluation Benchmark SPHINX-X: Scaling Data and Parameters for a Family of Multi-modal Large Language Models
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 07d84a8d-8c82-4c1e-9f4e-114d04f26442 · outbound
MME-Industry: A Cross-Industry Multimodal Evaluation Benchmark ChartQA: A Benchmark for Question Answering about Charts with Visual and Logical Reasoning
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 40f43e90-910f-4e14-881c-eb6eff3996f9 · outbound
MME-Industry: A Cross-Industry Multimodal Evaluation Benchmark Training language models to follow instruc- tions with human feedback
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6b632a4b-1418-4641-be56-30d1e43b18a5 · outbound
MME-Industry: A Cross-Industry Multimodal Evaluation Benchmark Multilayer perceptron (mlp)
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 519b8997-851e-41ea-8111-42621960c03b · outbound
MME-Industry: A Cross-Industry Multimodal Evaluation Benchmark Internlm: A multilingual language model with progressively enhanced capabilities,
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 38df5ea7-3fb6-450b-ad8c-f086bdfc7e49 · outbound
MME-Industry: A Cross-Industry Multimodal Evaluation Benchmark Llama 2: Open Foundation and Fine-Tuned Chat Models
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e5d65e6f-ce52-4940-894f-d073327309c7 · outbound
MME-Industry: A Cross-Industry Multimodal Evaluation Benchmark To See is to Believe: Prompting GPT-4V for Better Visual Instruction Tuning
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 50448d07-38c6-45dc-997b-dde94c40c5b2 · outbound
MME-Industry: A Cross-Industry Multimodal Evaluation Benchmark CogVLM: Visual Expert for Pretrained Language Models
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b88beee7-c818-4108-82b2-687df36d5a08 · outbound
MME-Industry: A Cross-Industry Multimodal Evaluation Benchmark Baichuan 2: Open Large-scale Language Models
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 37d07d4e-13e8-45f8-ab6c-0509483a77fb · outbound
MME-Industry: A Cross-Industry Multimodal Evaluation Benchmark mPLUG-Owl: Modularization Empowers Large Language Models with Multimodality
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3093fe24-26ce-4dd0-8863-13feed24e6dc · outbound
MME-Industry: A Cross-Industry Multimodal Evaluation Benchmark mplug-owl3: Towards long image-sequence under- standing in multi-modal large language models,
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 0941ea37-996e-46b0-a446-d7a3510215e6 · outbound
MME-Industry: A Cross-Industry Multimodal Evaluation Benchmark MM-Vet: Evaluating Large Multimodal Models for Integrated Capabilities
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 05f96119-e74e-4678-ae06-f8e452a0beba · outbound
MME-Industry: A Cross-Industry Multimodal Evaluation Benchmark Mmmu: A massive multi-discipline multimodal understanding and reasoning benchmark for expert agi
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 8e1a6271-5d9e-4084-b77c-f2a4a9a756f8 · outbound
MME-Industry: A Cross-Industry Multimodal Evaluation Benchmark MM-LLMs: Recent Advances in MultiModal Large Language Models
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4b5ce515-2087-4987-9dc4-0acbe47afee9 · outbound
MME-Industry: A Cross-Industry Multimodal Evaluation Benchmark MME-RealWorld: Could Your Multimodal LLM Challenge High-Resolution Real-World Scenarios that are Difficult for Humans?
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e6d42da3-5a4f-4899-80b3-07d14b2c7270 · outbound
MME-Industry: A Cross-Industry Multimodal Evaluation Benchmark MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 361c4051-056e-45f1-919e-07e7e8cbebaf · outbound
MME-Industry: A Cross-Industry Multimodal Evaluation Benchmark Gemini: A Family of Highly Capable Multimodal Models
Reference 2018
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation de929faf-3875-4ebd-bf05-fa60ce8b9334 · outbound
MME-Industry: A Cross-Industry Multimodal Evaluation Benchmark Scaling up visual and vision-language representation learning with noisy text su- pervision
Reference 2021
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 4b5406b8-2c45-4456-89e2-507d466cc627 · outbound
MME-Industry: A Cross-Industry Multimodal Evaluation Benchmark Docvqa: A dataset for vqa on document images
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9064c011-934f-4443-b298-8b06eea20d22 · outbound
MME-Industry: A Cross-Industry Multimodal Evaluation Benchmark OpenFlamingo: An Open-Source Framework for Training Large Autoregressive Vision-Language Models
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bfdcad41-56da-4654-9e62-6125effff869 · outbound
MME-Industry: A Cross-Industry Multimodal Evaluation Benchmark Sharegpt4v: Improving large multi-modal models with better captions
Reference 2024
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 1a8370a7-1b44-4ba9-80bb-363b55e9d0bd · outbound
MME-Industry: A Cross-Industry Multimodal Evaluation Benchmark Palm: Scaling lan- guage modeling with pathways
Reference 2025
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 54de9421-cea9-4749-b7f6-ae489d8dfa0a · inbound
CFBenchmark-MM: Chinese Financial Assistant Benchmark for Multimodal Large Language Model MME-Industry: A Cross-Industry Multimodal Evaluation Benchmark
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e28b1841-9f53-4e79-9ad8-b828921e8db5 · inbound
FORGE: Fine-grained Multimodal Evaluation for Manufacturing Scenarios MME-Industry: A Cross-Industry Multimodal Evaluation Benchmark
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.