Pith. sign in

Paper Citation Record · LEDGER

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models

As of 15 August 2026, this Paper Citation Record lists 53 of 53 outbound references and 2 inbound Pith citation observations for arXiv:2411.19103.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.19103 v1

Coverage vector

measured 53 of 53 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T10:36:36.109735Z

measured 55 of 55 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T15:23:53.069460Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-11T08:50:57.754355Z

Reference resolution

53 of 53 outbound references displayed

  • verified exact0
  • verified fuzzy11
  • unresolved42
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c3ec4983-b98b-4b07-bdfe-49a8bae6b79d · outbound

This paper cites Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-12T10:36:35.852542Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:36:35.852542Z digest=sha256:8c789ee5fa5a2ecc252f089f4bb29024ea85d0bd088175c673ff1f18c48da6b0

Observation c7f4e65e-14c8-4085-9674-9268d6349542 · outbound

This paper cites GPT-4 Technical Report.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models GPT-4 Technical Report

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-12T10:36:35.858504Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:36:35.858504Z digest=sha256:8f921d008c9c344254259c5a108cb1188077956c01e392c41dc4e91b3bc49e72

Observation c1e5ee41-0493-4459-a3d8-4db1ff6bf164 · outbound

This paper cites Pixtral 12B.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models Pixtral 12B

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-12T10:36:35.864070Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:36:35.864070Z digest=sha256:d49ba095c075ef40b15b262a631eda165019d48fe131e52a72ad95c0e53fa3cc

Observation b93e74b5-54da-4b69-8c7c-7cd75a6814ba · outbound

This paper cites Claude 3.5 sonnet, 2024.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models Claude 3.5 sonnet, 2024

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-12T10:36:35.869702Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:36:35.869702Z digest=sha256:97e99d7f981feada0515b3be40860a0f9d701cfc9761ba6ca7b8b181744c9ca9

Observation 3b116a9b-aaf5-4d76-9988-cf27460567a7 · outbound

This paper cites Are we on the right way for evaluating large vision-language models? In The Thirty-eighth Annual Conference on Neural Information Processing Systems, 2024.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models Are we on the right way for evaluating large vision-language models? In The Thirty-eighth Annual Conference on Neural Information Processing Systems, 2024

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:36:36.927658Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T10:36:35.873893Z digest=sha256:52722d44a6a03814655025a0dc50a411beefc0b4f00195f4b15e4244fa16d371

Observation b16a62da-450e-4651-b93b-b9fed078a453 · outbound

This paper cites How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models How Far Are We to GPT-4V? Closing the Gap to Commercial Multimodal Models with Open-Source Suites

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-12T10:36:35.879674Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:36:35.879674Z digest=sha256:31cf34d5593c97e33b7257dd745f8656fa1e3db342011178ebe2d75fb7268f4e

Observation 246d330c-3898-44fc-b5c1-fd9b8351ba9f · outbound

This paper cites Molmo and PixMo: Open Weights and Open Data for State-of-the-Art Vision-Language Models.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models Molmo and PixMo: Open Weights and Open Data for State-of-the-Art Vision-Language Models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-12T10:36:35.885630Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:36:35.885630Z digest=sha256:2685460e15bf5f529462db413ddf94d2e008ca7417bd340e5b0dde9b06e5cdc9

Observation 5714edc0-0fb7-401a-aa5b-ba5e6dfac8b5 · outbound

This paper cites VLMEvalKit: An Open-Source Toolkit for Evaluating Large Multi-Modality Models.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models VLMEvalKit: An Open-Source Toolkit for Evaluating Large Multi-Modality Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-12T10:36:35.890601Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:36:35.890601Z digest=sha256:d235b44cfb23ac2981096ab63611e7626f4140affe7272d94b774937f7999868

Observation 51c57779-af72-41e5-8fee-5046260785c0 · outbound

This paper cites Pororo: Platform of neural models for natural language processing.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models Pororo: Platform of neural models for natural language processing

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:36:36.912470Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T10:36:35.895328Z digest=sha256:56b95976cce02b8d258c1a6d1f85a439064b81381cb89c4005d848d6e70a79aa

Observation c0b004ca-f92b-4cb4-a871-3eadd68f546f · outbound

This paper cites Icdar 2013 robust reading competition.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models Icdar 2013 robust reading competition

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:36:36.897869Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T10:36:35.899817Z digest=sha256:93178698ef524a95883ac0129f716817ad446e846a777576efb2742995a7f03c

Observation d1372440-1e8f-4eac-8b8a-cbe10f98cbec · outbound

This paper cites Icdar 2015 competition on robust reading.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models Icdar 2015 competition on robust reading

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:36:36.884238Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T10:36:35.904048Z digest=sha256:006c467ef8344f6ff2b81edff7be171402cabc4cb7face8edc2ef0c7f70142ab

Observation 51682c5f-09ce-42e3-ac5a-701740ba3fe5 · outbound

This paper cites Ocr-free document understanding transformer.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models Ocr-free document understanding transformer

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:36:36.870270Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T10:36:35.908719Z digest=sha256:e9ad43219064f7b2e84953658f2cf61e2d494ff528692258b3914c4b575245d1

Observation ef31c4a8-ef05-48a1-ab75-b25a673efd42 · outbound

This paper cites Korean Localization of Visual Question Answering for Blind People.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models Korean Localization of Visual Question Answering for Blind People

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:36:36.856356Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T10:36:35.913040Z digest=sha256:85c766a843ff2401ae887869fa5380242a130f77086858b1112e4fd2776c2fdb

Observation e481b500-b1fb-47f6-a568-685b8dc5b950 · outbound

This paper cites Building and better understanding vision-language models: insights and future directions.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models Building and better understanding vision-language models: insights and future directions

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-12T10:36:35.917708Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:36:35.917708Z digest=sha256:709303c0d218253217c165192341d1217a8b3efb91d9173a3b8f223e2be9d0c1

Observation eb303d3c-5e48-4e5f-9f4b-e34c74bfd650 · outbound

This paper cites Popeval: A character-level approach to end-to-end evaluation compatible with word-level benchmark dataset.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models Popeval: A character-level approach to end-to-end evaluation compatible with word-level benchmark dataset

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:36:36.843146Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T10:36:35.922716Z digest=sha256:b77420f911cce448d1ce0a8f5d98c5edd84a13c50da4489fc5426c49cd82d6d6

Observation eb77d461-959f-4b92-896c-53fd31b8bfbf · outbound

This paper cites LLaVA-OneVision: Easy Visual Task Transfer.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models LLaVA-OneVision: Easy Visual Task Transfer

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-12T10:36:35.927460Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:36:35.927460Z digest=sha256:6bbc00c3fbb720daa5de341e69bec938998c7c18b170817bb949b1f0e4face2e

Observation 17dbb9ee-fe15-4080-aa48-7c3b080ca89e · outbound

This paper cites Seed-bench: Benchmarking multimodal large language models.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models Seed-bench: Benchmarking multimodal large language models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-12T10:36:35.932104Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:36:35.932104Z digest=sha256:8ce96f24686a41523f95378a7d7a5c88f5dbf5c70ae8c6a0e9fb690e7f7e14dd

Observation 67f14028-ed47-4e08-9371-ad03019cda90 · outbound

This paper cites LLaVA-NeXT-Interleave: Tackling Multi-image, Video, and 3D in Large Multimodal Models.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models LLaVA-NeXT-Interleave: Tackling Multi-image, Video, and 3D in Large Multimodal Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-12T10:36:35.936853Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:36:35.936853Z digest=sha256:49a14585ee7ed74a0f4ae7e84e3002347d94d7c0429a6f4d1077ea5c19d174c6

Observation 8fbd28c9-3f04-4a69-a00d-4c74e4a7748e · outbound

This paper cites A Survey on Benchmarks of Multimodal Large Language Models.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models A Survey on Benchmarks of Multimodal Large Language Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-12T10:36:35.943050Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:36:35.943050Z digest=sha256:c7779d470dfc8ad767801c2cd6253fe6ddfbea0805388d148cf8ed2627ea013f

Observation 16a8dd29-5225-48db-94d6-791e9927e454 · outbound

This paper cites Visual instruction tuning.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models Visual instruction tuning

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-12T10:36:35.948507Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:36:35.948507Z digest=sha256:56e8dce0036c5f3f0c40dc862cfe35596d005c86a57120b118966d16e093b761

Observation e9b90bd9-b359-4b8f-b787-b79dee9b7014 · outbound

This paper cites Visual instruction tuning.Advances in neural information processing systems , 36, 2024.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models Visual instruction tuning.Advances in neural information processing systems , 36, 2024

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-12T10:36:35.952362Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:36:35.952362Z digest=sha256:99cd520489245ce61c3e2d86e786685c25056a8c71e04c11e79105b2ddf7e2ef

Observation 9a7b7c2d-3b38-464d-a53e-41b2259a319a · outbound

This paper cites Mmbench: Is your multi-modal model an all-around player? In European Conference on Computer Vision, pages 216–233.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models Mmbench: Is your multi-modal model an all-around player? In European Conference on Computer Vision, pages 216–233

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-12T10:36:35.956783Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:36:35.956783Z digest=sha256:3eeabbdfb55dc565a664508c5d982963cdbd53657d2e475a12a4708680051ca4

Observation 644ef2d0-4c4f-46d9-9072-4f75b800ab02 · outbound

This paper cites OCRBench: On the Hidden Mystery of OCR in Large Multimodal Models.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models OCRBench: On the Hidden Mystery of OCR in Large Multimodal Models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-12T10:36:35.961256Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:36:35.961256Z digest=sha256:d4db0f5c60cb075b983051b4716ffbeebf465069d02a716a8ac9b2202c576923

Observation 62577d7d-b611-4861-a10a-b66b17ba5381 · outbound

This paper cites DeepSeek-VL: Towards Real-World Vision-Language Understanding.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models DeepSeek-VL: Towards Real-World Vision-Language Understanding

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-12T10:36:35.966053Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:36:35.966053Z digest=sha256:e5f01cae3159b5dd2c564c70a8baf80a3ab5fa0828c0b24f567074673853675e

Observation a482db7a-3c49-4dcf-a98d-b55608b817c6 · outbound

This paper cites Feast Your Eyes: Mixture-of-Resolution Adaptation for Multimodal Large Language Models.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models Feast Your Eyes: Mixture-of-Resolution Adaptation for Multimodal Large Language Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-12T10:36:35.970570Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:36:35.970570Z digest=sha256:dda7b11f3e25cdce06beae7b9fa0437388bb4bf6a7d04fa742aa048d889d0b4f

Observation c061e2c0-08c5-49dc-b3f4-024b4b23566d · outbound

This paper cites Cord: A consolidated receipt dataset for post-ocr parsing.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models Cord: A consolidated receipt dataset for post-ocr parsing

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:36:36.793542Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T10:36:35.974921Z digest=sha256:eef21ada746b62aa507d4db7b71f275849227bdc852c5216fcd8fa0f28e199ba

Observation da1c4e4a-caa6-43f2-8733-0d84693ad957 · outbound

This paper cites Kosmos-2: Grounding Multimodal Large Language Models to the World.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models Kosmos-2: Grounding Multimodal Large Language Models to the World

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-12T10:36:35.979973Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:36:35.979973Z digest=sha256:cf178b8dec7d52f5fe30af3cfa1dba1ad6aa0d351c30fc4d14e23632ffc4391d

Observation f7e29a92-e734-4ce7-a87b-d385a2a9644f · outbound

This paper cites Direct preference optimization: Your language model is secretly a reward model.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models Direct preference optimization: Your language model is secretly a reward model

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-12T10:36:35.985077Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:36:35.985077Z digest=sha256:0bb70462499ee3d0f296665fd2f41f5672dde09f91f86f86b5d4acb5fe3c7a96

Observation 604c59d3-a8fe-4a00-b43c-68248d20f7d1 · outbound

This paper cites Exaone 3.0 7.8 b instruction tuned language model.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models Exaone 3.0 7.8 b instruction tuned language model

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-12T10:36:35.990115Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:36:35.990115Z digest=sha256:ddc70dbf97ae79cff16e86c24eea331b975d760fc38ccd4a0b1cc87c74edb7f1

Observation b23c4022-859d-42bf-b0fa-541cbf73ec0c · outbound

This paper cites X-LLaV A: Optimizing bilingual large vision-language alignment.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models X-LLaV A: Optimizing bilingual large vision-language alignment

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-12T10:36:35.995460Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:36:35.995460Z digest=sha256:5b14ba974e01b54f2726fe4e590c648d951414fa23e2001330eb2f3ed57bc803

Observation 19ec7602-0137-4c88-a501-a41b2eee73c8 · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models Gemini: A Family of Highly Capable Multimodal Models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-12T10:36:36.001932Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:36:36.001932Z digest=sha256:341f5e555370d17d425f1da8eba228188d74ad1e5639d4a4895d26dc17527ab4

Observation 452c9471-7122-4496-a1f6-8cb1ff36ffa1 · outbound

This paper cites Qwen2.5: A party of foundation models, September 2024.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models Qwen2.5: A party of foundation models, September 2024

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-12T10:36:36.006796Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:36:36.006796Z digest=sha256:12c146db5e77ecfd3eb052ed2b2609166aa76784b634a88838b80b32730e9ff2

Observation 8db81b3a-4c41-42d0-a111-7d3770ee2e35 · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-12T10:36:36.010906Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:36:36.010906Z digest=sha256:bd0ac17e419041de3b8771b905beb08f3c872a009dc7e6b5b9545809347a792c

Observation 3545269b-b6bc-4f4c-9c61-60214502c22f · outbound

This paper cites MiniCPM-V: A GPT-4V Level MLLM on Your Phone.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models MiniCPM-V: A GPT-4V Level MLLM on Your Phone

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-12T10:36:36.015486Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:36:36.015486Z digest=sha256:a69fff523641f90ff69693c1f15c54df717ae0b6f6ad0a3e3038ef4aa72b685a

Observation 8960d159-d36d-4147-9bed-865e5c3489f0 · outbound

This paper cites Mmmu: A massive multi-discipline multimodal understanding and reasoning benchmark for expert agi.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models Mmmu: A massive multi-discipline multimodal understanding and reasoning benchmark for expert agi

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-12T10:36:36.020089Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:36:36.020089Z digest=sha256:f47f1ce035122e3a2ebae321f343f59252dffedd91e18409a3e040a3e08a335e

Observation 5befc4bb-fea9-4a40-bb2a-3c5767bb97a3 · outbound

This paper cites Pangea: A Fully Open Multilingual Multimodal LLM for 39 Languages.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models Pangea: A Fully Open Multilingual Multimodal LLM for 39 Languages

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-12T10:36:36.025075Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:36:36.025075Z digest=sha256:1e2ac8db8d88ff1b12bfa10c9066634e4fce9859d406fd856f3468c2a7f7985b

Observation 2f56c486-75f2-45ef-b34e-0104b5948ad0 · outbound

This paper cites Sigmoid loss for language image pre-training.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models Sigmoid loss for language image pre-training

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-12T10:36:36.029653Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:36:36.029653Z digest=sha256:53b84422dfcda4f5635ca44db77c8105e3243eb487b2f14c9c51f0a046fcbc83

Observation adab2f40-cc40-44f4-8c01-717e2f6e1b60 · outbound

This paper cites LMMs-Eval: Reality Check on the Evaluation of Large Multimodal Models.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models LMMs-Eval: Reality Check on the Evaluation of Large Multimodal Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-12T10:36:36.033894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:36:36.033894Z digest=sha256:729dbc2a1a0f460fe6f96b9a01e98098bbd0ef2f397315ab7db060e1ceb096ab

Observation 8f6afc56-4097-454a-9150-3fc39708ad0f · outbound

This paper cites Swift:a scal- able lightweight infrastructure for fine-tuning, 2024.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models Swift:a scal- able lightweight infrastructure for fine-tuning, 2024

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:36:36.748316Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T10:36:36.039180Z digest=sha256:228f4dc975a03e519ebfb52f017e1d0fd97c723b15f8d92d921bfa95a032c9b7

Observation 80e390e2-b100-4c33-a098-68045181a8b3 · outbound

This paper cites Judging llm-as-a-judge with mt-bench and chatbot arena.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models Judging llm-as-a-judge with mt-bench and chatbot arena

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-12T10:36:36.044493Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:36:36.044493Z digest=sha256:7905a7af271ae4fa9aa4313d1723b54873fa0dee0e104e40ed19c70202f18407

Observation 139285f2-244a-45bd-973c-d4c7a734478c · outbound

This paper cites LlamaFactory: Unified Efficient Fine-Tuning of 100+ Language Models.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models LlamaFactory: Unified Efficient Fine-Tuning of 100+ Language Models

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-12T10:36:36.049341Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:36:36.049341Z digest=sha256:7721809fc69b7bc459545321bfbc8667faec81075a8adae89c9e96f4cb55988a

Observation 809f3286-352b-42bf-9f2c-9a5b7840e845 · outbound

This paper cites an unresolved cited work.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models Unresolved cited work

Reference 42

Resolution
unresolved
raw_fallback, observed 2026-08-12T10:36:36.719589Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T10:36:36.054202Z digest=sha256:13b590260244b2af684feb8526f61366cd7fdfed156f312b420654e806553083

Observation fcd469a4-f479-4bf2-a266-66771f85addc · outbound

This paper cites an unresolved cited work.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models Unresolved cited work

Reference 43

Resolution
unresolved
raw_fallback, observed 2026-08-12T10:36:36.705504Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T10:36:36.059447Z digest=sha256:349b80e26bbf8357740f8e0b01ce7d57cd9757a41e6bbc1bb527f29d0c06ecc5

Observation 4af0a18f-ab65-4411-9732-09c638dc520c · outbound

This paper cites an unresolved cited work.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models Unresolved cited work

Reference 44

Resolution
unresolved
raw_fallback, observed 2026-08-12T10:36:36.691332Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T10:36:36.065506Z digest=sha256:8aae88f66d096c8af1063e961a481abd3cb4c05e8bcabf88147fa2984c7e10a3

Observation 3a37a793-b66e-49fa-b975-b201fe16a6bd · outbound

This paper cites # 출력 형식 - 첫 번째 줄: `어시스턴트1_점수 어시스턴트2_점수` (예: `8 9`) - 두 번째 줄: `유용성`, `관련성`, `정확성`, `세부 수준`, `한국어 생성능력` 기준으로 점수를 설명 하는 자세한 문단을 제공합니다.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models # 출력 형식 - 첫 번째 줄: `어시스턴트1_점수 어시스턴트2_점수` (예: `8 9`) - 두 번째 줄: `유용성`, `관련성`, `정확성`, `세부 수준`, `한국어 생성능력` 기준으로 점수를 설명 하는 자세한 문단을 제공합니다

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:36:36.676764Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T10:36:36.069989Z digest=sha256:0602d0eb843cadcd4f5eac5687f7e0137b832b34220f5bdd84306b302e584c1c

Observation 9e90c8ed-c6c5-420f-97fe-23b5148bdec5 · outbound

This paper cites an unresolved cited work.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models Unresolved cited work

Reference 46

Resolution
unresolved
raw_fallback, observed 2026-08-12T10:36:36.663376Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T10:36:36.076612Z digest=sha256:cd8a6efc1ae76c369752bc082cc14d3ffe8fc8790c0fbb64f8935ba6ae6642d3

Observation 8c43a167-2fae-459b-9913-4bb7fe9235a8 · outbound

This paper cites an unresolved cited work.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models Unresolved cited work

Reference 47

Resolution
unresolved
raw_fallback, observed 2026-08-12T10:36:36.650092Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T10:36:36.080778Z digest=sha256:8945d142795ea59dc08423f27b4f75cc539fc3920e83cc9de842421048b4414e

Observation 5c19e8b7-c018-4eb3-bb3b-fc4f50a87db1 · outbound

This paper cites an unresolved cited work.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models Unresolved cited work

Reference 48

Resolution
unresolved
raw_fallback, observed 2026-08-12T10:36:36.631921Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T10:36:36.085491Z digest=sha256:38c461726d6703fb810eed5a2e4f7ab54b2967c3b36ca732124b682f45a12408

Observation bc7c32ca-e971-4795-8d4d-38402f0f4955 · outbound

This paper cites an unresolved cited work.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models Unresolved cited work

Reference 49

Resolution
unresolved
raw_fallback, observed 2026-08-12T10:36:36.615033Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T10:36:36.089879Z digest=sha256:5b0445de26c804f79cfd1422cebb22b5d879b58cad666056c24d396166b3bfdd

Observation 36ebed22-f097-4b91-93c6-d375de28c098 · outbound

This paper cites an unresolved cited work.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models Unresolved cited work

Reference 50

Resolution
unresolved
raw_fallback, observed 2026-08-12T10:36:36.601697Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T10:36:36.094275Z digest=sha256:4aeaaa3c440f4be865a647fa6238e5d2453a7924b7cbc0c76baf2bbfa6eb75e5

Observation 51633135-2074-408c-9ef9-6e88c3558468 · outbound

This paper cites an unresolved cited work.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models Unresolved cited work

Reference 51

Resolution
unresolved
raw_fallback, observed 2026-08-12T10:36:36.588969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T10:36:36.099332Z digest=sha256:9b64f7566e1d41158efac4fb9f0b04877d08a2844671f28e0836d890f863b130

Observation 21d48145-90f6-48b4-8569-dd08c1d97009 · outbound

This paper cites an unresolved cited work.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models Unresolved cited work

Reference 52

Resolution
unresolved
raw_fallback, observed 2026-08-12T10:36:36.576150Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T10:36:36.104063Z digest=sha256:1c32953aded1f02c78a4807a388fbf5ae6610e42c5d8452fa66bdea5ea1ed194

Observation a2bd0708-bf91-4878-b984-d136dd20162c · outbound

This paper cites STARBUCKS.

VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models STARBUCKS

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T10:36:36.563962Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-12T10:36:36.109735Z digest=sha256:27ddf042b2772b1dd9fd91bc889c887bac876b89e5e6d596eba00bb14a9b0829

Pith citing papers

Observation 42f4af09-c14f-415e-a1e6-575f581e615f · inbound

KRETA: A Benchmark for Korean Reading and Reasoning in Text-Rich VQA Attuned to Diverse Visual Contexts cites this paper.

KRETA: A Benchmark for Korean Reading and Reasoning in Text-Rich VQA Attuned to Diverse Visual Contexts VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-05T15:23:53.069460Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:23:53.069460Z digest=sha256:04dff7b2ae65db7b02e8f27f747f12538c6690993b615618daa13e4c19e9c24e

Observation 579ef28e-9638-4b05-bd0d-dbb1829f3b4d · inbound

Anthropogenic Regional Adaptation in Multimodal Vision-Language Model cites this paper.

Anthropogenic Regional Adaptation in Multimodal Vision-Language Model VARCO-VISION: Expanding Frontiers in Korean Vision-Language Models

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-11T08:50:57.757159Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-05-10T16:29:12.064221Z digest=sha256:66378664ea4a0f8855a950da4d8554e77d3beb30efec45c6b40145e57fdcfeb8