Pith. sign in

Paper Citation Record · LEDGER

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models

As of 14 August 2026, this Paper Citation Record lists 53 of 53 outbound references and 2 inbound Pith citation observations for arXiv:2411.19103.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.19103 v1

Coverage vector

measured 53 of 53 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T10:36:36.109735Z

measured 55 of 55 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T15:23:53.069460Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-11T08:50:57.754355Z

Reference resolution

53 of 53 outbound references displayed

  • verified exact0
  • verified fuzzy11
  • unresolved42
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c3ec4983-b98b-4b07-bdfe-49a8bae6b79d · outbound

This paper cites Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-12T10:36:35.852542Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:36:35.852542Z digest=sha256:8c789ee5fa5a2ecc252f089f4bb29024ea85d0bd088175c673ff1f18c48da6b0

Observation c7f4e65e-14c8-4085-9674-9268d6349542 · outbound

This paper cites GPT-4 Technical Report.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models GPT-4 Technical Report

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-12T10:36:35.858504Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:36:35.858504Z digest=sha256:8f921d008c9c344254259c5a108cb1188077956c01e392c41dc4e91b3bc49e72

Observation c1e5ee41-0493-4459-a3d8-4db1ff6bf164 · outbound

This paper cites Pixtral 12B.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models Pixtral 12B

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-12T10:36:35.864070Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:36:35.864070Z digest=sha256:d49ba095c075ef40b15b262a631eda165019d48fe131e52a72ad95c0e53fa3cc

Observation b93e74b5-54da-4b69-8c7c-7cd75a6814ba · outbound

This paper cites Claude 3.5 sonnet, 2024.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models Claude 3.5 sonnet, 2024

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-12T10:36:35.869702Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:36:35.869702Z digest=sha256:97e99d7f981feada0515b3be40860a0f9d701cfc9761ba6ca7b8b181744c9ca9

Observation 3b116a9b-aaf5-4d76-9988-cf27460567a7 · outbound

This paper cites Are we on the right way for evaluating large vision-language models? In The Thirty-eighth Annual Conference on Neural Information Processing Systems, 2024.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models Are we on the right way for evaluating large vision-language models? In The Thirty-eighth Annual Conference on Neural Information Processing Systems, 2024

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:36:36.927658Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T10:36:35.873893Z digest=sha256:e56c5adbf9ac365b431109ec6cfa077dbe2ca70ac900e7de28704bccc8cc570d

Observation b16a62da-450e-4651-b93b-b9fed078a453 · outbound

This paper cites How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-12T10:36:35.879674Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:36:35.879674Z digest=sha256:31cf34d5593c97e33b7257dd745f8656fa1e3db342011178ebe2d75fb7268f4e

Observation 246d330c-3898-44fc-b5c1-fd9b8351ba9f · outbound

This paper cites Molmo and PixMo: Open Weights and Open Data for State-of-the-Art Vision-Language Models.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models Molmo and PixMo: Open Weights and Open Data for State-of-the-Art Vision-Language Models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-12T10:36:35.885630Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:36:35.885630Z digest=sha256:2685460e15bf5f529462db413ddf94d2e008ca7417bd340e5b0dde9b06e5cdc9

Observation 5714edc0-0fb7-401a-aa5b-ba5e6dfac8b5 · outbound

This paper cites VLMEvalKit: An Open-Source Toolkit for Evaluating Large Multi-Modality Models.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models VLMEvalKit: An Open-Source Toolkit for Evaluating Large Multi-Modality Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-12T10:36:35.890601Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:36:35.890601Z digest=sha256:d235b44cfb23ac2981096ab63611e7626f4140affe7272d94b774937f7999868

Observation 51c57779-af72-41e5-8fee-5046260785c0 · outbound

This paper cites Pororo: Platform of neural models for natural language processing.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models Pororo: Platform of neural models for natural language processing

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:36:36.912470Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T10:36:35.895328Z digest=sha256:18d7f5d332b02a03614897183d1780c734c7987f2e53d5b07e00fbb85c7b0450

Observation c0b004ca-f92b-4cb4-a871-3eadd68f546f · outbound

This paper cites Icdar 2013 robust reading competition.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models Icdar 2013 robust reading competition

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:36:36.897869Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T10:36:35.899817Z digest=sha256:8e9f57682cc1b777fe4be43a142b400f7850bb106448aa5dba0b57d1b6ac6861

Observation d1372440-1e8f-4eac-8b8a-cbe10f98cbec · outbound

This paper cites Icdar 2015 competition on robust reading.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models Icdar 2015 competition on robust reading

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:36:36.884238Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T10:36:35.904048Z digest=sha256:e7f82902a2293374cdffa4b204c00bfc82dbb3371b3799becf4662c38f418ed9

Observation 51682c5f-09ce-42e3-ac5a-701740ba3fe5 · outbound

This paper cites Ocr-free document understanding transformer.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models Ocr-free document understanding transformer

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:36:36.870270Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T10:36:35.908719Z digest=sha256:f47f65219f3999c65273a32b54ed4f91ba37bb796e3473f0d86c0cc08edf454c

Observation ef31c4a8-ef05-48a1-ab75-b25a673efd42 · outbound

This paper cites Korean Localization of Visual Question Answering for Blind People.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models Korean Localization of Visual Question Answering for Blind People

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:36:36.856356Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T10:36:35.913040Z digest=sha256:8fdf3165ac2a20125d4e8b0b45c69c75b7502c943a5cb165676aad95945b7987

Observation e481b500-b1fb-47f6-a568-685b8dc5b950 · outbound

This paper cites Building and better understanding vision-language models: insights and future directions.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models Building and better understanding vision-language models: insights and future directions

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-12T10:36:35.917708Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:36:35.917708Z digest=sha256:709303c0d218253217c165192341d1217a8b3efb91d9173a3b8f223e2be9d0c1

Observation eb303d3c-5e48-4e5f-9f4b-e34c74bfd650 · outbound

This paper cites Popeval: A character-level approach to end-to-end evaluation compatible with word-level benchmark dataset.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models Popeval: A character-level approach to end-to-end evaluation compatible with word-level benchmark dataset

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:36:36.843146Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T10:36:35.922716Z digest=sha256:bd30304b0b6565b7f28ff16535da65184e442e3b0bae7456f60a7717cfa2a688

Observation eb77d461-959f-4b92-896c-53fd31b8bfbf · outbound

This paper cites LLaVA-OneVision: Easy Visual Task Transfer.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models LLaVA-OneVision: Easy Visual Task Transfer

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-12T10:36:35.927460Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:36:35.927460Z digest=sha256:6bbc00c3fbb720daa5de341e69bec938998c7c18b170817bb949b1f0e4face2e

Observation 17dbb9ee-fe15-4080-aa48-7c3b080ca89e · outbound

This paper cites Seed-bench: Benchmarking multimodal large language models.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models Seed-bench: Benchmarking multimodal large language models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-12T10:36:35.932104Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:36:35.932104Z digest=sha256:8ce96f24686a41523f95378a7d7a5c88f5dbf5c70ae8c6a0e9fb690e7f7e14dd

Observation 67f14028-ed47-4e08-9371-ad03019cda90 · outbound

This paper cites LLaVA-NeXT-Interleave: Tackling Multi-image, Video, and 3D in Large Multimodal Models.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models LLaVA-NeXT-Interleave: Tackling Multi-image, Video, and 3D in Large Multimodal Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-12T10:36:35.936853Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:36:35.936853Z digest=sha256:49a14585ee7ed74a0f4ae7e84e3002347d94d7c0429a6f4d1077ea5c19d174c6

Observation 8fbd28c9-3f04-4a69-a00d-4c74e4a7748e · outbound

This paper cites A Survey on Benchmarks of Multimodal Large Language Models.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models A Survey on Benchmarks of Multimodal Large Language Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-12T10:36:35.943050Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:36:35.943050Z digest=sha256:c7779d470dfc8ad767801c2cd6253fe6ddfbea0805388d148cf8ed2627ea013f

Observation 16a8dd29-5225-48db-94d6-791e9927e454 · outbound

This paper cites Visual instruction tuning.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models Visual instruction tuning

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-12T10:36:35.948507Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:36:35.948507Z digest=sha256:56e8dce0036c5f3f0c40dc862cfe35596d005c86a57120b118966d16e093b761

Observation e9b90bd9-b359-4b8f-b787-b79dee9b7014 · outbound

This paper cites Visual instruction tuning.Advances in neural information processing systems , 36, 2024.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models Visual instruction tuning.Advances in neural information processing systems , 36, 2024

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-12T10:36:35.952362Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:36:35.952362Z digest=sha256:99cd520489245ce61c3e2d86e786685c25056a8c71e04c11e79105b2ddf7e2ef

Observation 9a7b7c2d-3b38-464d-a53e-41b2259a319a · outbound

This paper cites Mmbench: Is your multi-modal model an all-around player? In European Conference on Computer Vision, pages 216–233.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models Mmbench: Is your multi-modal model an all-around player? In European Conference on Computer Vision, pages 216–233

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-12T10:36:35.956783Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:36:35.956783Z digest=sha256:3eeabbdfb55dc565a664508c5d982963cdbd53657d2e475a12a4708680051ca4

Observation 644ef2d0-4c4f-46d9-9072-4f75b800ab02 · outbound

This paper cites OCRBench: On the Hidden Mystery of OCR in Large Multimodal Models.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models OCRBench: On the Hidden Mystery of OCR in Large Multimodal Models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-12T10:36:35.961256Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:36:35.961256Z digest=sha256:d4db0f5c60cb075b983051b4716ffbeebf465069d02a716a8ac9b2202c576923

Observation 62577d7d-b611-4861-a10a-b66b17ba5381 · outbound

This paper cites DeepSeek-VL: Towards Real-World Vision-Language Understanding.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models DeepSeek-VL: Towards Real-World Vision-Language Understanding

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-12T10:36:35.966053Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:36:35.966053Z digest=sha256:e5f01cae3159b5dd2c564c70a8baf80a3ab5fa0828c0b24f567074673853675e

Observation a482db7a-3c49-4dcf-a98d-b55608b817c6 · outbound

This paper cites Feast Your Eyes: Mixture-of-Resolution Adaptation for Multimodal Large Language Models.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models Feast Your Eyes: Mixture-of-Resolution Adaptation for Multimodal Large Language Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-12T10:36:35.970570Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:36:35.970570Z digest=sha256:fec50a91628f4d61102f84dbba02482a26deef4e5235d07b9a40e0b6a3146d5b

Observation c061e2c0-08c5-49dc-b3f4-024b4b23566d · outbound

This paper cites Cord: A consolidated receipt dataset for post-ocr parsing.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models Cord: A consolidated receipt dataset for post-ocr parsing

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:36:36.793542Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T10:36:35.974921Z digest=sha256:c5690b16f7aa46cbb8dd6070cbd535b46983dd83c9443921561248c673916d1f

Observation da1c4e4a-caa6-43f2-8733-0d84693ad957 · outbound

This paper cites Kosmos-2: Grounding Multimodal Large Language Models to the World.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models Kosmos-2: Grounding Multimodal Large Language Models to the World

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-12T10:36:35.979973Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:36:35.979973Z digest=sha256:cf178b8dec7d52f5fe30af3cfa1dba1ad6aa0d351c30fc4d14e23632ffc4391d

Observation f7e29a92-e734-4ce7-a87b-d385a2a9644f · outbound

This paper cites Direct preference optimization: Your language model is secretly a reward model.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models Direct preference optimization: Your language model is secretly a reward model

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-12T10:36:35.985077Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:36:35.985077Z digest=sha256:0bb70462499ee3d0f296665fd2f41f5672dde09f91f86f86b5d4acb5fe3c7a96

Observation 604c59d3-a8fe-4a00-b43c-68248d20f7d1 · outbound

This paper cites Exaone 3.0 7.8 b instruction tuned language model.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models Exaone 3.0 7.8 b instruction tuned language model

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-12T10:36:35.990115Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:36:35.990115Z digest=sha256:ddc70dbf97ae79cff16e86c24eea331b975d760fc38ccd4a0b1cc87c74edb7f1

Observation b23c4022-859d-42bf-b0fa-541cbf73ec0c · outbound

This paper cites X-LLaV A: Optimizing bilingual large vision-language alignment.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models X-LLaV A: Optimizing bilingual large vision-language alignment

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-12T10:36:35.995460Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:36:35.995460Z digest=sha256:5b14ba974e01b54f2726fe4e590c648d951414fa23e2001330eb2f3ed57bc803

Observation 19ec7602-0137-4c88-a501-a41b2eee73c8 · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models Gemini: A Family of Highly Capable Multimodal Models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-12T10:36:36.001932Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:36:36.001932Z digest=sha256:341f5e555370d17d425f1da8eba228188d74ad1e5639d4a4895d26dc17527ab4

Observation 452c9471-7122-4496-a1f6-8cb1ff36ffa1 · outbound

This paper cites Qwen2.5: A party of foundation models, September 2024.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models Qwen2.5: A party of foundation models, September 2024

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-12T10:36:36.006796Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:36:36.006796Z digest=sha256:12c146db5e77ecfd3eb052ed2b2609166aa76784b634a88838b80b32730e9ff2

Observation 8db81b3a-4c41-42d0-a111-7d3770ee2e35 · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-12T10:36:36.010906Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:36:36.010906Z digest=sha256:bd0ac17e419041de3b8771b905beb08f3c872a009dc7e6b5b9545809347a792c

Observation 3545269b-b6bc-4f4c-9c61-60214502c22f · outbound

This paper cites MiniCPM-V: A GPT-4V Level MLLM on Your Phone.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models MiniCPM-V: A GPT-4V Level MLLM on Your Phone

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-12T10:36:36.015486Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:36:36.015486Z digest=sha256:a69fff523641f90ff69693c1f15c54df717ae0b6f6ad0a3e3038ef4aa72b685a

Observation 8960d159-d36d-4147-9bed-865e5c3489f0 · outbound

This paper cites Mmmu: A massive multi-discipline multimodal understanding and reasoning benchmark for expert agi.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models Mmmu: A massive multi-discipline multimodal understanding and reasoning benchmark for expert agi

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-12T10:36:36.020089Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:36:36.020089Z digest=sha256:f47f1ce035122e3a2ebae321f343f59252dffedd91e18409a3e040a3e08a335e

Observation 5befc4bb-fea9-4a40-bb2a-3c5767bb97a3 · outbound

This paper cites Pangea: A Fully Open Multilingual Multimodal LLM for 39 Languages.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models Pangea: A Fully Open Multilingual Multimodal LLM for 39 Languages

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-12T10:36:36.025075Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:36:36.025075Z digest=sha256:1e2ac8db8d88ff1b12bfa10c9066634e4fce9859d406fd856f3468c2a7f7985b

Observation 2f56c486-75f2-45ef-b34e-0104b5948ad0 · outbound

This paper cites Sigmoid loss for language image pre-training.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models Sigmoid loss for language image pre-training

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-12T10:36:36.029653Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:36:36.029653Z digest=sha256:53b84422dfcda4f5635ca44db77c8105e3243eb487b2f14c9c51f0a046fcbc83

Observation adab2f40-cc40-44f4-8c01-717e2f6e1b60 · outbound

This paper cites LMMs-Eval: Reality Check on the Evaluation of Large Multimodal Models.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models LMMs-Eval: Reality Check on the Evaluation of Large Multimodal Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-12T10:36:36.033894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:36:36.033894Z digest=sha256:729dbc2a1a0f460fe6f96b9a01e98098bbd0ef2f397315ab7db060e1ceb096ab

Observation 8f6afc56-4097-454a-9150-3fc39708ad0f · outbound

This paper cites Swift:a scal- able lightweight infrastructure for fine-tuning, 2024.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models Swift:a scal- able lightweight infrastructure for fine-tuning, 2024

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:36:36.748316Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T10:36:36.039180Z digest=sha256:a3ba3041603a76439740bd7d0df8eca4fd7e298baef9f29ca4ef0eee5094b568

Observation 80e390e2-b100-4c33-a098-68045181a8b3 · outbound

This paper cites Judging llm-as-a-judge with mt-bench and chatbot arena.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models Judging llm-as-a-judge with mt-bench and chatbot arena

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-12T10:36:36.044493Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:36:36.044493Z digest=sha256:7905a7af271ae4fa9aa4313d1723b54873fa0dee0e104e40ed19c70202f18407

Observation 139285f2-244a-45bd-973c-d4c7a734478c · outbound

This paper cites LlamaFactory: Unified Efficient Fine-Tuning of 100+ Language Models.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models LlamaFactory: Unified Efficient Fine-Tuning of 100+ Language Models

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-12T10:36:36.049341Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:36:36.049341Z digest=sha256:7721809fc69b7bc459545321bfbc8667faec81075a8adae89c9e96f4cb55988a

Observation 809f3286-352b-42bf-9f2c-9a5b7840e845 · outbound

This paper cites an unresolved cited work.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models Unresolved cited work

Reference 42

Resolution
unresolved
raw_fallback, observed 2026-08-12T10:36:36.719589Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T10:36:36.054202Z digest=sha256:a0f9093d5b965cdaeac2c65f08fc2fe94882b57d33f4a74dc586a88d1efaa271

Observation fcd469a4-f479-4bf2-a266-66771f85addc · outbound

This paper cites an unresolved cited work.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models Unresolved cited work

Reference 43

Resolution
unresolved
raw_fallback, observed 2026-08-12T10:36:36.705504Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T10:36:36.059447Z digest=sha256:6226546843d2df9094a80fb726f5bc34acedee0f92bded8dd32f260d2b4f68f0

Observation 4af0a18f-ab65-4411-9732-09c638dc520c · outbound

This paper cites an unresolved cited work.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models Unresolved cited work

Reference 44

Resolution
unresolved
raw_fallback, observed 2026-08-12T10:36:36.691332Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T10:36:36.065506Z digest=sha256:afe18df3780bdd7820ea5b736efbe061ea33d97c73d6764525aa93ca653b7379

Observation 3a37a793-b66e-49fa-b975-b201fe16a6bd · outbound

This paper cites # 출력 형식 - 첫 번째 줄: `어시스턴트1_점수 어시스턴트2_점수` (예: `8 9`) - 두 번째 줄: `유용성`, `관련성`, `정확성`, `세부 수준`, `한국어 생성능력` 기준으로 점수를 설명 하는 자세한 문단을 제공합니다.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models # 출력 형식 - 첫 번째 줄: `어시스턴트1_점수 어시스턴트2_점수` (예: `8 9`) - 두 번째 줄: `유용성`, `관련성`, `정확성`, `세부 수준`, `한국어 생성능력` 기준으로 점수를 설명 하는 자세한 문단을 제공합니다

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:36:36.676764Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T10:36:36.069989Z digest=sha256:26515801bc0ba6f7e75a05e00d00bdf671e03b0d35a9912c77054c915ddf6df9

Observation 9e90c8ed-c6c5-420f-97fe-23b5148bdec5 · outbound

This paper cites an unresolved cited work.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models Unresolved cited work

Reference 46

Resolution
unresolved
raw_fallback, observed 2026-08-12T10:36:36.663376Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T10:36:36.076612Z digest=sha256:02840d0c56b77b014c32601d9d564d31aef504e3ce72c97a782b110be5e1a1c7

Observation 8c43a167-2fae-459b-9913-4bb7fe9235a8 · outbound

This paper cites an unresolved cited work.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models Unresolved cited work

Reference 47

Resolution
unresolved
raw_fallback, observed 2026-08-12T10:36:36.650092Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T10:36:36.080778Z digest=sha256:60638be671ec7032b093a4164355a491574a80034938ffdc9c8776cf1df5177e

Observation 5c19e8b7-c018-4eb3-bb3b-fc4f50a87db1 · outbound

This paper cites an unresolved cited work.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models Unresolved cited work

Reference 48

Resolution
unresolved
raw_fallback, observed 2026-08-12T10:36:36.631921Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T10:36:36.085491Z digest=sha256:72b1db56644055ad1712c86c5e1a421858dedf578599c151d2797f936d923b56

Observation bc7c32ca-e971-4795-8d4d-38402f0f4955 · outbound

This paper cites an unresolved cited work.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models Unresolved cited work

Reference 49

Resolution
unresolved
raw_fallback, observed 2026-08-12T10:36:36.615033Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T10:36:36.089879Z digest=sha256:e0a2a84bb07da7bb7c44d1ae563012bc444ed322de0035a34c1ad387e1577960

Observation 36ebed22-f097-4b91-93c6-d375de28c098 · outbound

This paper cites an unresolved cited work.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models Unresolved cited work

Reference 50

Resolution
unresolved
raw_fallback, observed 2026-08-12T10:36:36.601697Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T10:36:36.094275Z digest=sha256:8b483f12b94dac86e3f8bd0aba16ac9bd4c7c6376076859cedf176facbe36e69

Observation 51633135-2074-408c-9ef9-6e88c3558468 · outbound

This paper cites an unresolved cited work.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models Unresolved cited work

Reference 51

Resolution
unresolved
raw_fallback, observed 2026-08-12T10:36:36.588969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T10:36:36.099332Z digest=sha256:679937395fe9b3e2186585e253e64bf82d0ce1e90e6aa726c57c0024507664f5

Observation 21d48145-90f6-48b4-8569-dd08c1d97009 · outbound

This paper cites an unresolved cited work.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models Unresolved cited work

Reference 52

Resolution
unresolved
raw_fallback, observed 2026-08-12T10:36:36.576150Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T10:36:36.104063Z digest=sha256:2afdf39b5f0e24a50643e49a776e9df50b22be0467fb7bfffc5d2eb01b78b388

Observation a2bd0708-bf91-4878-b984-d136dd20162c · outbound

This paper cites STARBUCKS.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models STARBUCKS

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:36:36.563962Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-12T10:36:36.109735Z digest=sha256:30d76934ab0e33eb11990ea419320b163c1a2790b121f3a37c5ce77637f0d71c

Pith citing papers

Observation 42f4af09-c14f-415e-a1e6-575f581e615f · inbound

KRETA: A Benchmark for Korean Reading and Reasoning in Text-Rich VQA Attuned to Diverse Visual Contexts cites this paper.

KRETA: A Benchmark for Korean Reading and Reasoning in Text-Rich VQA Attuned to Diverse Visual Contexts VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-05T15:23:53.069460Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:23:53.069460Z digest=sha256:04dff7b2ae65db7b02e8f27f747f12538c6690993b615618daa13e4c19e9c24e

Observation 579ef28e-9638-4b05-bd0d-dbb1829f3b4d · inbound

Anthropogenic Regional Adaptation in Multimodal Vision-Language Model cites this paper.

Anthropogenic Regional Adaptation in Multimodal Vision-Language Model VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-11T08:50:57.757159Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-10T16:29:12.064221Z digest=sha256:fdc1cf789c12d06894eec5eb26d0aa6032567ec442d6362928c00cffafa0fdf0