Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T04:57:29.583850Z
Paper Citation Record · LEDGER
As of 14 August 2026, this Paper Citation Record lists 87 of 87 outbound references and 0 inbound Pith citation observations for arXiv:2608.11167.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T04:57:29.583850Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
87 of 87 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 565c71e8-44ee-40a7-bc9c-ca5e5cff1448 · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment Visual Instruction Tuning , booktitle =
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 885c6a52-3474-490d-95a4-3e54da950b4a · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment LLaVA-NeXT: Improved reasoning, OCR, and world knowledge , url=
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4b2b103e-19f2-4ef5-8ddd-af34039fd017 · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment Unresolved cited work
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0bc2528a-21ed-49d6-afd1-97a4894d6f61 · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 754d5abb-09e1-49bb-a305-b31610cae958 · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment Qwen2.5-VL Technical Report
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a3f5ab58-210a-4f26-be6b-b9e5e7637237 · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment 2025 , eprint=
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ec6fb014-cc3f-411c-97ab-3317bc25fe51 · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment MiMo-VL Technical Report
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c498bd22-1e36-48e9-8295-3b69342f6ca9 · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment 2025 , eprint=
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 56cd082d-a781-48ff-a7b9-25ea335c561a · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9700f8ad-d9d6-4c9c-9ab2-e256ead06748 · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment InternVL3: Exploring Advanced Training and Test-Time Recipes for Open-Source Multimodal Models
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a98493a4-1f8a-4c78-8c33-d5e957b981c8 · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment Scalable Vision Language Model Training via High Quality Data Curation , booktitle =
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f4742661-3eac-49a7-96af-d507aed679e0 · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment Flamingo: a Visual Language Model for Few-Shot Learning , booktitle =
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5e997eb9-1d21-4376-aba1-c8cb9b751f88 · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment mPLUG-Owl2: Revolutionizing Multi-modal Large Language Model with Modality Collaboration
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a1ca3212-96f5-4120-8be9-32b9fd38b029 · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment Seed1.5-VL Technical Report
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 606d72bd-1521-41af-86f8-ed8e16e64fee · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment CoRR , volume =
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation ea48c056-da6f-48a1-ab50-01bb588aa137 · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment Computer Vision -
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a47e8c63-bfea-4588-8ce7-f1fe15e5d248 · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment Qwen2.5 Technical Report
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation de5b4bd7-3978-4b42-83fc-81c5c01d00f2 · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment Qwen3 Technical Report
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8679e8f7-be7a-4ec3-9f82-ee1df8f835b9 · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment The Llama 3 Herd of Models
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cd8b26b5-f973-4828-8103-9c4e6686aec1 · outbound
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7843997f-c70e-4364-835e-31a78e905f08 · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment LMMs-Eval: Reality Check on the Evaluation of Large Multimodal Models , booktitle =
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8fb8e444-3744-443c-9fa4-582fcb27a03e · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment MMBench: Is Your Multi-modal Model an All-Around Player? , booktitle =
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f478145a-65b0-4b2f-92cb-dec2857d45c9 · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cc75e00e-fa7a-443b-9502-7c12dd2e591b · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment Are We on the Right Way for Evaluating Large Vision-Language Models? , booktitle =
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b6f4399b-c5c1-438f-832f-2d47f8d26dc6 · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment Forty-first International Conference on Machine Learning,
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 92af91a1-ab07-4a2f-bbd3-d324f400ce42 · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment Hudson and Christopher D
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b2889eba-9548-4e16-b9c6-e9ec1d9db9fb · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment A Diagram is Worth a Dozen Images , booktitle =
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5d21991e-be69-43e5-b0e1-c1de80f27ee0 · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment Joty and Enamul Hoque , editor =
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 966b00db-cd54-4d65-85b5-0de2840a07ed · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment Cambrian-1:
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation e8eb2ba9-63e7-4ac4-aa06-97e6ee4e2dfd · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment OCRBench: on the hidden mystery of
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0bb53f02-c088-4c53-9237-85ab45e99862 · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment 2019 , url =
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8d61bc53-0d30-4d5b-8979-cb784e26e15d · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment Berg , editor =
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 438c1bf1-98ec-4d3d-8c16-65f9f746a306 · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment Yuille and Kevin Murphy , title =
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e74cd858-eb1d-44c7-9060-f68719cf461e · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment Unresolved cited work
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 205dfa8c-497f-424e-a738-adee591aca4e · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment Molmo and PixMo: Open Weights and Open Data for State-of-the-Art Vision-Language Models , booktitle =
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 45e93531-70cd-4e47-bbed-9e0b49aca790 · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment ShareGPT4V: Improving Large Multi-modal Models with Better Captions , booktitle =
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3cec23b2-d15e-479f-8f48-d39036e9083a · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cef80f05-23ca-4fd2-b9f9-f0cdad2e25d2 · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment DenseFusion-1M: Merging Vision Experts for Comprehensive Multimodal Perception , booktitle =
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 89047adb-f0dd-4db6-8626-d5c982287beb · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment DenseWorld-1M: Towards Detailed Dense Grounded Caption in the Real World
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 21da1ec2-b606-4317-af1b-4c006199bc6b · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment ImageInWords: Unlocking Hyper-Detailed Image Descriptions , booktitle =
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4e2ded75-9897-483d-bdaa-3244b3cc8133 · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment Computer Vision -
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 59fce482-b53e-4896-a479-f67ead3f281a · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment CoRR , volume =
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation b417e9be-da9b-45fe-ade6-3cc9286b023f · outbound
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c5ddacd0-7702-4c31-95f8-6ca162d3a25d · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment Unresolved cited work
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ce3f8042-9f48-4740-8fa3-884b0755b6c6 · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment Unresolved cited work
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9d10c119-6fdb-454c-b553-625a839350ea · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment 2024 , url =
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation abbab471-c565-464f-aea3-0451272a2df6 · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment Segment Anything , booktitle =
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bb5db8b4-3b52-4168-a186-825279c19c33 · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment Unresolved cited work
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 70fb8413-9ed0-4208-8925-81225e0307ff · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment 2019 International Conference on Document Analysis and Recognition,
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0be0bab2-6ab8-41c5-b7ff-f0acd559e871 · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment Visual Genome: Connecting Language and Vision Using Crowdsourced Dense Image Annotations , journal =
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 99e3d290-888c-4409-8749-adef6a03714e · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment Price and Scott Cohen and Christopher Kanan , title =
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e793620b-01d4-4fc2-8d3d-0b07548dbd58 · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment OCR-Free Document Understanding Transformer , booktitle =
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d7e7e17a-16cd-40c2-9c9c-6fce36402e1b · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment Unresolved cited work
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 2ff79465-f090-4877-939c-03b535a9919f · outbound
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d872d31f-f3b5-4831-a098-3ecb84aa09a2 · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment The Thirteenth International Conference on Learning Representations,
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5f8c0848-552d-4c53-b33b-54ac20f4bec2 · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment Forty-first International Conference on Machine Learning,
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation bdaf6862-6507-4e7c-9463-79024530986e · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment Hinton , editor =
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 9e0ab83a-63a3-4426-b20c-d62751cdb8e8 · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment SEA : Supervised Embedding Alignment for Token-Level Visual-Textual Integration in MLLM s
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation d6ac425d-f669-4b50-ae94-43c4517bc5f7 · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment Analyzing Fine-Grained Alignment and Enhancing Vision Understanding in Multimodal Language Models
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 13a8c10a-4f3b-4043-88ba-1996568f87b1 · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment GroundingGPT: Language Enhanced Multi-modal Grounding Model , booktitle =
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation e4d0ef44-8a33-4824-a5f3-a1f312028ad2 · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment Advancing Fine-Grained Visual Understanding with Multi-Scale Alignment in Multi-Modal Models
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation d34a3720-006b-476d-9616-28a8ab2b83c5 · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment Kosmos-2: Grounding Multimodal Large Language Models to the World
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 771488ca-ee58-44b5-878e-c446e0a4dfe8 · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment ParGo: Bridging Vision-Language with Partial and Global Views , booktitle =
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 73485882-dd2d-4a8e-930b-2214a5f87f1f · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment Shikra: Unleashing Multimodal LLM's Referential Dialogue Magic
Reference 69
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b0b25b03-1f68-4821-9d4a-7fad0d63e5f6 · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment Ferret: Refer and Ground Anything Anywhere at Any Granularity , booktitle =
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 2d1c01d8-6b9d-4668-901f-e007c922e92b · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment ICCV , year=
Reference 71
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 2046d4a3-1d46-4e01-8fb0-c5a24677d9fc · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment Unresolved cited work
Reference 72
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 9aa41628-5252-4051-b715-16d7d06c68d3 · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment LLaVA-OneVision-1.5: Fully Open Framework for Democratized Multimodal Training
Reference 73
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation baf50106-d779-424c-acb3-61de245ec292 · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment Learning Transferable Visual Models From Natural Language Supervision , booktitle =
Reference 74
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1b9d230d-3562-416b-80aa-8a8b8503e51a · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment SigLIP 2: Multilingual Vision-Language Encoders with Improved Semantic Understanding, Localization, and Dense Features
Reference 75
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2d5521f7-924d-4bba-b023-f15988c62f71 · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment GPT4RoI: Instruction Tuning Large Language Model on Region-of-Interest , booktitle =
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 02cc1978-67db-47a9-bbb7-fcad49a87637 · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment Shaker and Salman H
Reference 77
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 70b2912e-eb61-455e-8ada-155eaddd0cbc · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment NExT-Chat: An
Reference 78
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation e9d00afa-6941-4047-8a5c-b6247aaab4f7 · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment LLaVA-Grounding: Grounded Visual Chat with Large Multimodal Models , booktitle =
Reference 79
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d660a6f6-af95-4a67-b46f-6defdd0d40d6 · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment The All-Seeing Project
Reference 80
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 474a49c6-8f61-4d83-9ac5-b84e8e4d9867 · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment Investigating and Scaling up Code-Switching for Multilingual Language Model Pre-Training , booktitle =
Reference 81
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation e313d1cf-c9e1-4d48-985b-13a69e1b5aac · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment PreAlign: Boosting Cross-Lingual Transfer by Early Establishment of Multilingual Alignment , booktitle =
Reference 82
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3f70b69b-cdfc-4972-b6fc-dbdce5df0209 · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment Code-Switching Curriculum Learning for Multilingual Transfer in LLMs , booktitle =
Reference 83
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation cd13adba-e00f-494f-af79-67378847e00b · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment DICT-MLM: Improved Multilingual Pre-Training using Bilingual Dictionaries
Reference 84
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d323adac-6712-4058-92e6-6a814d721a39 · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment CoSDA-ML: Multi-Lingual Code-Switching Data Augmentation for Zero-Shot Cross-Lingual
Reference 85
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation c4518eaa-3692-4674-9be4-5fb1b2a65984 · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment Bilingual Alignment Pre-Training for Zero-Shot Cross-Lingual Transfer
Reference 86
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation ed807199-313f-400c-b98a-76c76d9c7e41 · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment Latino Language and Communicative Behavior , editor =
Reference 87
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 3626a8b2-33ad-484e-9e91-26d287c368d9 · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment Thara and Prabaharan Poornachandran , title =
Reference 88
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 0f9fbaa1-9520-482c-b3a1-b903594a379a · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment Alabdulmohsin and Avital Oliver and Piotr Padlewski and Alexey A
Reference 89
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 0fb93f31-adc7-44a4-9fd3-175496eb9a09 · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment Hu and Yelong Shen and Phillip Wallis and Zeyuan Allen
Reference 90
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aa44995f-2e8f-4629-81e1-0efa883a7051 · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment 2025 , howpublished =
Reference 91
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e8ec4d31-14bf-4d50-ae36-e862ddc6b724 · outbound
MultiModal Code-Switching: Interleaving Visual Objects into Language for Explicit Object-Level Alignment 18th International Conference on Pattern Recognition
Reference 92
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.