Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T18:43:22.479627Z
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 44 of 44 outbound references and 0 inbound Pith citation observations for arXiv:2507.07572.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T18:43:22.479627Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
44 of 44 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 6dc2b321-d1db-4656-a462-9264db353127 · outbound
Single-to-mix Modality Alignment with Multimodal Large Language Model for Document Image Machine Translation Unresolved cited work
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation bdac6906-0ab9-460a-9edc-43e17dc20641 · outbound
Single-to-mix Modality Alignment with Multimodal Large Language Model for Document Image Machine Translation Unresolved cited work
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6b6fbd1a-17e0-4ea4-b980-901eb30e7fd4 · outbound
Single-to-mix Modality Alignment with Multimodal Large Language Model for Document Image Machine Translation Unresolved cited work
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 310cdb34-bc96-4d48-ab61-5808c1d8a2b8 · outbound
Single-to-mix Modality Alignment with Multimodal Large Language Model for Document Image Machine Translation mPLUG-DocOwl 1.5: Unified Structure Learning for OCR-free Document Understanding
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2296abb2-4e8f-4392-8606-0a28df6bc9e0 · outbound
Single-to-mix Modality Alignment with Multimodal Large Language Model for Document Image Machine Translation Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen - Zhu, Yuanzhi Li, Shean Wang, Lu Wang, and Weizhu Chen
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1bc7f4d3-af09-475e-99da-44cf5e6753ad · outbound
Single-to-mix Modality Alignment with Multimodal Large Language Model for Document Image Machine Translation GPT-4o System Card
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a6de0343-f519-4e49-9730-79ccf97b7b90 · outbound
Single-to-mix Modality Alignment with Multimodal Large Language Model for Document Image Machine Translation Unresolved cited work
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 57d06870-3a37-42ac-92e3-6f8c3c6c4e6b · outbound
Single-to-mix Modality Alignment with Multimodal Large Language Model for Document Image Machine Translation Unresolved cited work
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c926d771-bc43-425c-b240-5862a28ee999 · outbound
Single-to-mix Modality Alignment with Multimodal Large Language Model for Document Image Machine Translation Translatotron-V(ison): An End-to-End Model for In-Image Machine Translation
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 09771ffe-fd32-4820-bab3-4bfadea98e0c · outbound
Single-to-mix Modality Alignment with Multimodal Large Language Model for Document Image Machine Translation Unresolved cited work
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0979d8c0-a23b-4506-9f84-d6c4531577fa · outbound
Single-to-mix Modality Alignment with Multimodal Large Language Model for Document Image Machine Translation Unresolved cited work
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d9374080-267b-459d-bd30-5ce2bfe06085 · outbound
Single-to-mix Modality Alignment with Multimodal Large Language Model for Document Image Machine Translation Unresolved cited work
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 621e9192-ad72-48e2-b972-2c2d8ac597bf · outbound
Single-to-mix Modality Alignment with Multimodal Large Language Model for Document Image Machine Translation Unresolved cited work
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 836f1d85-7cb8-4389-ad5a-7a2d6669fde8 · outbound
Single-to-mix Modality Alignment with Multimodal Large Language Model for Document Image Machine Translation TextMonkey: An OCR-Free Large Multimodal Model for Understanding Document
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 08f09daf-7f9c-4d6a-a3e0-a912c63669aa · outbound
Single-to-mix Modality Alignment with Multimodal Large Language Model for Document Image Machine Translation Unresolved cited work
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 47b5cf15-cab5-4aea-8a21-edace5787eef · outbound
Single-to-mix Modality Alignment with Multimodal Large Language Model for Document Image Machine Translation Unresolved cited work
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation cc494ea0-9e09-496f-a697-4a3da22fc6d6 · outbound
Single-to-mix Modality Alignment with Multimodal Large Language Model for Document Image Machine Translation Unresolved cited work
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 44ea4592-9e4c-4f04-aaa2-1f04274590ca · outbound
Single-to-mix Modality Alignment with Multimodal Large Language Model for Document Image Machine Translation Unresolved cited work
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation d2fe24c6-c3e2-4f60-8151-cb5b13f1c0f6 · outbound
Single-to-mix Modality Alignment with Multimodal Large Language Model for Document Image Machine Translation Unresolved cited work
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 3b055af7-2b40-4165-8088-e074e9d6d084 · outbound
Single-to-mix Modality Alignment with Multimodal Large Language Model for Document Image Machine Translation Unresolved cited work
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation eb7ea040-f597-4173-a005-0e4d8149a5b0 · outbound
Single-to-mix Modality Alignment with Multimodal Large Language Model for Document Image Machine Translation Unresolved cited work
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 87267277-c07d-4b20-82c3-e49eaf6d0c0b · outbound
Single-to-mix Modality Alignment with Multimodal Large Language Model for Document Image Machine Translation AnyTrans: Translate AnyText in the Image with Large Scale Models
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 6fc84834-7f6a-4097-a237-80af945b0015 · outbound
Single-to-mix Modality Alignment with Multimodal Large Language Model for Document Image Machine Translation Unresolved cited work
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 661ee238-238a-4928-9db6-02a159c7b0e6 · outbound
Single-to-mix Modality Alignment with Multimodal Large Language Model for Document Image Machine Translation Unresolved cited work
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f6083f23-d158-4533-91f5-b857296d4947 · outbound
Single-to-mix Modality Alignment with Multimodal Large Language Model for Document Image Machine Translation Unresolved cited work
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e7b7882e-a38a-49d7-a955-15442eb3ed0a · outbound
Single-to-mix Modality Alignment with Multimodal Large Language Model for Document Image Machine Translation Gemini 1.5: Unlocking multimodal understanding across millions of tokens of context
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e2d68b63-7cb2-41a2-8ee1-57cc5bf0ac6a · outbound
Single-to-mix Modality Alignment with Multimodal Large Language Model for Document Image Machine Translation Unresolved cited work
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 47265098-39ea-48f1-a157-bd79d7ede202 · outbound
Single-to-mix Modality Alignment with Multimodal Large Language Model for Document Image Machine Translation Unresolved cited work
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 367341b0-89e6-42f2-b246-9aedb38f19ec · outbound
Single-to-mix Modality Alignment with Multimodal Large Language Model for Document Image Machine Translation Vary: Scaling up the Vision Vocabulary for Large Vision-Language Models
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b8c558e9-a200-4c3d-994c-2ab58b37d8ff · outbound
Single-to-mix Modality Alignment with Multimodal Large Language Model for Document Image Machine Translation Small Language Model Meets with Reinforced Vision Vocabulary
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b3366063-80bd-4a49-a4b4-65b52d237c8e · outbound
Single-to-mix Modality Alignment with Multimodal Large Language Model for Document Image Machine Translation The Dawn of LMMs: Preliminary Explorations with GPT-4V(ision)
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b54d8299-b8f4-4942-8f0c-e52353333fd5 · outbound
Single-to-mix Modality Alignment with Multimodal Large Language Model for Document Image Machine Translation DocXChain: A Powerful Open-Source Toolchain for Document Parsing and Beyond
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4b668490-1497-4a9e-ba9e-36a94dd2cd0e · outbound
Single-to-mix Modality Alignment with Multimodal Large Language Model for Document Image Machine Translation mPLUG-DocOwl: Modularized Multimodal Large Language Model for Document Understanding
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation da945551-70eb-4509-a555-6964945401cb · outbound
Single-to-mix Modality Alignment with Multimodal Large Language Model for Document Image Machine Translation TextHawk: Exploring Efficient Fine-Grained Perception of Multimodal Large Language Models
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 568769e4-9114-4bd8-9e2e-fe8d9b7c8914 · outbound
Single-to-mix Modality Alignment with Multimodal Large Language Model for Document Image Machine Translation LLaVAR: Enhanced Visual Instruction Tuning for Text-Rich Image Understanding
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 058aca8d-a209-44e2-9cf3-80997d90c1c4 · outbound
Single-to-mix Modality Alignment with Multimodal Large Language Model for Document Image Machine Translation Unresolved cited work
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 927e280a-4c4d-4dbc-b4d8-860dd826acea · outbound
Single-to-mix Modality Alignment with Multimodal Large Language Model for Document Image Machine Translation Unresolved cited work
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b14c7078-59e3-485b-a49c-2be46bcfabc7 · outbound
Single-to-mix Modality Alignment with Multimodal Large Language Model for Document Image Machine Translation Unresolved cited work
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 23791114-4098-49e3-a17c-314ea538591d · outbound
Single-to-mix Modality Alignment with Multimodal Large Language Model for Document Image Machine Translation Unresolved cited work
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5a7ff8e7-e621-4895-8db6-0127cb835b59 · outbound
Single-to-mix Modality Alignment with Multimodal Large Language Model for Document Image Machine Translation Unresolved cited work
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 31d7ac5b-3258-4d07-a4fd-5570c973ddad · outbound
Single-to-mix Modality Alignment with Multimodal Large Language Model for Document Image Machine Translation Unresolved cited work
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 4f85e258-13dc-4932-9cd9-69d55e67015d · outbound
Single-to-mix Modality Alignment with Multimodal Large Language Model for Document Image Machine Translation Unresolved cited work
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 3c75423e-e542-4227-9811-01443ea5758c · outbound
Single-to-mix Modality Alignment with Multimodal Large Language Model for Document Image Machine Translation online" 'onlinestring :=
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9db02041-8bfc-4c1a-8e66-e438ae89158e · outbound
Single-to-mix Modality Alignment with Multimodal Large Language Model for Document Image Machine Translation write newline
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.