Pith. sign in

Paper Citation Record · LEDGER

Rethinking Multilingual Vision-Language Translation: Dataset, Evaluation, and Adaptation

As of 8 August 2026, this Paper Citation Record lists 42 of 42 outbound references and 2 inbound Pith citation observations for arXiv:2506.11820.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.11820 v1

Coverage vector

measured 42 of 42 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T04:09:38.557854Z

measured 44 of 44 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-26T00:50:22.839005Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T16:09:57.560157Z

Reference resolution

42 of 42 outbound references displayed

  • verified exact2
  • verified fuzzy0
  • unresolved37
  • parse uncertain0
  • malformed identifier2
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a07cb292-aab3-4548-b423-965549ca451a · outbound

This paper cites an unresolved cited work.

Rethinking Multilingual Vision-Language Translation: Dataset, Evaluation, and Adaptation Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:09:42.473118Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:09:34.190461Z digest=sha256:dbad6a36fcd7c9838af88e5c98189ba3e411791fde3261d2bfbf8825a7436c3d

Observation a902591a-d0ed-43fa-82af-d9f9dacb2089 · outbound

This paper cites an unresolved cited work.

Rethinking Multilingual Vision-Language Translation: Dataset, Evaluation, and Adaptation Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:09:42.364665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:09:34.274201Z digest=sha256:d373a31f15e626ed6cd7506dd9668f8345a06c6eda2923ea5979f3cd6914ffad

Observation 6465ca04-a3de-47c9-8cea-a9f255bb5518 · outbound

This paper cites an unresolved cited work.

Rethinking Multilingual Vision-Language Translation: Dataset, Evaluation, and Adaptation Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:09:42.252074Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:09:34.368938Z digest=sha256:e0e2426768f7b3c4d3aae739b30a1aba0c294555d68553e9e4452541ef1aaed9

Observation 7b8dc7da-69e9-4c9d-b52a-a27c1c825e7b · outbound

This paper cites an unresolved cited work.

Rethinking Multilingual Vision-Language Translation: Dataset, Evaluation, and Adaptation Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:09:42.147588Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:09:34.456818Z digest=sha256:2ab9d5b2bfa0768741d07485e3bedd73c46f0793687c68db0e9776c8378cf35b

Observation da00d981-6080-4a81-afd3-fd9404a8f77f · outbound

This paper cites Qwen Technical Report.

Rethinking Multilingual Vision-Language Translation: Dataset, Evaluation, and Adaptation Qwen Technical Report

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T04:09:34.536244Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:09:34.536244Z digest=sha256:f57a33c201499d280649be7f3b46cebe03e6b79334ee164fa1eec2d9cc978b81

Observation 9ce7deb8-c1b8-41d5-83cd-fa0285ca6a55 · outbound

This paper cites Qwen2.5-VL Technical Report.

Rethinking Multilingual Vision-Language Translation: Dataset, Evaluation, and Adaptation Qwen2.5-VL Technical Report

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T04:09:34.620101Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:09:34.620101Z digest=sha256:ea9bd40acb4000401214ff4566a46ce3f72830ae5bbfb5c233931122ad8f4c52

Observation 54ad94a7-f6da-43a5-9a6e-d0fbeafe9033 · outbound

This paper cites Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling.

Rethinking Multilingual Vision-Language Translation: Dataset, Evaluation, and Adaptation Expanding Performance Boundaries of Open-Source Multimodal Models with Model, Data, and Test-Time Scaling

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T04:09:34.693712Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:09:34.693712Z digest=sha256:5228e8d2bbc6c5c77585fe3284a4a7243bb1a4cbeaf9edddaaa9156e30deef32

Observation c8217aed-9829-4158-ad9c-3f61f4b3c1d8 · outbound

This paper cites an unresolved cited work.

Rethinking Multilingual Vision-Language Translation: Dataset, Evaluation, and Adaptation Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:09:41.938815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:09:34.802169Z digest=sha256:f3ac91eab0e2f98c96ea9e284974a69ab695d5ced716402b1ed5302e2f624211

Observation beae5aa9-804d-47ce-85fd-bbc5d049e92e · outbound

This paper cites InstructBLIP: Towards General-purpose Vision-Language Models with Instruction Tuning.

Rethinking Multilingual Vision-Language Translation: Dataset, Evaluation, and Adaptation InstructBLIP: Towards General-purpose Vision-Language Models with Instruction Tuning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T04:09:34.888015Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:09:34.888015Z digest=sha256:7708c154cc684158a9625fa25e42cfc90419a1032b52fa1bfc89adc22a8a65c0

Observation a5ccf9eb-4b48-42f9-b2d2-c439bca80c13 · outbound

This paper cites an unresolved cited work.

Rethinking Multilingual Vision-Language Translation: Dataset, Evaluation, and Adaptation Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:09:41.704019Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:09:34.981820Z digest=sha256:ccea9ef254796c8d96459b7c6d0378f2b42a0db335c8c5199e666ad550d304fe

Observation a489ee83-af0a-4b92-92a4-dc0270172f2d · outbound

This paper cites InternLM-XComposer2: Mastering Free-form Text-Image Composition and Comprehension in Vision-Language Large Model.

Rethinking Multilingual Vision-Language Translation: Dataset, Evaluation, and Adaptation InternLM-XComposer2: Mastering Free-form Text-Image Composition and Comprehension in Vision-Language Large Model

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T04:09:35.068668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:09:35.068668Z digest=sha256:646005dc85f2319b810d7aad8cba7b9001995ec186e9c992cb1e4d31723d118d

Observation c3a40e3e-c8df-4cce-9e1d-0f4f7080c97f · outbound

This paper cites OCRBench v2: An Improved Benchmark for Evaluating Large Multimodal Models on Visual Text Localization and Reasoning.

Rethinking Multilingual Vision-Language Translation: Dataset, Evaluation, and Adaptation OCRBench v2: An Improved Benchmark for Evaluating Large Multimodal Models on Visual Text Localization and Reasoning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T04:09:35.164573Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:09:35.164573Z digest=sha256:a01764ada532f911cb11b76ebd8b5fa3347ad602c0b67e0db339fbe0cae9a905

Observation 5c7fcaa6-3b2c-45da-a5e6-c1a0f2802610 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Rethinking Multilingual Vision-Language Translation: Dataset, Evaluation, and Adaptation DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T04:09:35.268346Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:09:35.268346Z digest=sha256:c72919dc501bd2bb9aa0d08edbee52cc2e22a4913e5f40f96dd3071637c645d0

Observation 5a50f67f-5712-4359-b53e-c6e6834f46ea · outbound

This paper cites Why do LLaVA Vision-Language Models Reply to Images in English?.

Rethinking Multilingual Vision-Language Translation: Dataset, Evaluation, and Adaptation Why do LLaVA Vision-Language Models Reply to Images in English?

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-08-07T04:09:39.218598Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:09:35.375611Z digest=sha256:a53c10de073dbc6d2835ed68201dc3b558fb80d19a6ecd699a092d75305261fa

Observation fdb6b414-0240-4ad5-b963-8bc1245d4bea · outbound

This paper cites an unresolved cited work.

Rethinking Multilingual Vision-Language Translation: Dataset, Evaluation, and Adaptation Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:09:41.519045Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:09:35.482280Z digest=sha256:7231d9eae8b9fd21ba56843a70450083946c7d24e31dfd153b8d4fab7adc7405

Observation 021f66d7-dbf3-40cf-962b-d4020be13807 · outbound

This paper cites an unresolved cited work.

Rethinking Multilingual Vision-Language Translation: Dataset, Evaluation, and Adaptation Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:09:41.265749Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:09:35.712901Z digest=sha256:e66eec2dd18049a4f96c2c3fe61bccb18d2a0a9dd213632a93f5e82fe29a6b4a

Observation 85313ac1-7066-42a3-851f-1d431dd08420 · outbound

This paper cites an unresolved cited work.

Rethinking Multilingual Vision-Language Translation: Dataset, Evaluation, and Adaptation Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:09:41.033037Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:09:35.896344Z digest=sha256:09db3122bca35fffeffa2acfcb89e9ccdfd25ab9b12816c9e9cd9c8d017e98a8

Observation 9bf2ad5b-2a7a-4c85-9c6b-1ca2c814fe70 · outbound

This paper cites Exploring Better Text Image Translation with Multimodal Codebook.

Rethinking Multilingual Vision-Language Translation: Dataset, Evaluation, and Adaptation Exploring Better Text Image Translation with Multimodal Codebook

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T04:09:35.989987Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:09:35.989987Z digest=sha256:5de5d8e6a997e764e48ebe23233faee1ee9fe9537a941472c9558dab25690d2c

Observation 1d7ee12b-bb86-405d-a86f-87fd3b061468 · outbound

This paper cites LLaVA-OneVision: Easy Visual Task Transfer.

Rethinking Multilingual Vision-Language Translation: Dataset, Evaluation, and Adaptation LLaVA-OneVision: Easy Visual Task Transfer

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T04:09:36.118268Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:09:36.118268Z digest=sha256:a53c9fbdb9dd8854366d3d070f117ff598a672a9a3fc83e78bf6801cce161d85

Observation 7bc98be5-2f6d-41db-a796-33cbe7da28d4 · outbound

This paper cites MIT-10M: A Large Scale Parallel Corpus of Multilingual Image Translation.

Rethinking Multilingual Vision-Language Translation: Dataset, Evaluation, and Adaptation MIT-10M: A Large Scale Parallel Corpus of Multilingual Image Translation

Reference 20

Resolution
verified exact
local_arxiv, observed 2026-08-07T04:09:38.856996Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:09:36.200053Z digest=sha256:f326080b1a8ed240e87a97b274cc7777ad1d6dbc3c7cd7437fafd6dfaffbf42c

Observation c6021510-60ba-4509-a864-9b08541391ce · outbound

This paper cites an unresolved cited work.

Rethinking Multilingual Vision-Language Translation: Dataset, Evaluation, and Adaptation Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:09:40.859384Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:09:36.309154Z digest=sha256:f5752102c37d0ff7011e0cad53f36bf65f8e1f7e9bbaeb09677b87152788aa04

Observation 0557a70b-1a8b-4421-ac70-5e653519dacc · outbound

This paper cites an unresolved cited work.

Rethinking Multilingual Vision-Language Translation: Dataset, Evaluation, and Adaptation Unresolved cited work

Reference 22

Resolution
malformed identifier
no resolver link, observed 2026-08-07T04:09:36.415839Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:09:36.415839Z digest=sha256:b9636f2dab02e67d75d8519a1b7d9bc7ba07d1a3c67f3206e5ef4e7a4a104513

Observation a68d240e-463e-4c1a-bc8b-b2d08de45c9c · outbound

This paper cites DeepSeek-V3 Technical Report.

Rethinking Multilingual Vision-Language Translation: Dataset, Evaluation, and Adaptation DeepSeek-V3 Technical Report

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T04:09:36.524679Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:09:36.524679Z digest=sha256:97cf091ceddc6777f0c15b7698dc919f6df58b217188cf313593079a88fb4273

Observation 0fb004d3-12dd-46a7-a6ed-78a076c9ae56 · outbound

This paper cites Visual Instruction Tuning.

Rethinking Multilingual Vision-Language Translation: Dataset, Evaluation, and Adaptation Visual Instruction Tuning

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T04:09:36.610381Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:09:36.610381Z digest=sha256:adc6f8e67155ea61e59efbe24067c2dfa9e32066a04a9f5b8cf90ad09ea71034

Observation 091fc455-d1bf-4298-bc20-200ef325884a · outbound

This paper cites an unresolved cited work.

Rethinking Multilingual Vision-Language Translation: Dataset, Evaluation, and Adaptation Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:09:40.660349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:09:36.753882Z digest=sha256:6541da8a751f06f691c59485385e9cf7fd8ede39c4dc13a0a9eea6a51900046b

Observation 16f63c40-6054-47f8-8093-4e9e5391bc84 · outbound

This paper cites an unresolved cited work.

Rethinking Multilingual Vision-Language Translation: Dataset, Evaluation, and Adaptation Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:09:40.470286Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:09:36.828092Z digest=sha256:b0a6d28668f4b6a2dc94c176b4d3d7efedeb66d49533d46bb24ded4e82cd579d

Observation 93aa56c0-dd99-47e5-bd64-6cd2a0fa9677 · outbound

This paper cites an unresolved cited work.

Rethinking Multilingual Vision-Language Translation: Dataset, Evaluation, and Adaptation Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:09:40.347547Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:09:36.958398Z digest=sha256:2f854bf31964fd641f26666c1fc9f189a4646a24c12c926b4c31f781cb752b52

Observation af145a36-0bd4-4174-8123-b1f8c71a6dd9 · outbound

This paper cites an unresolved cited work.

Rethinking Multilingual Vision-Language Translation: Dataset, Evaluation, and Adaptation Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:09:40.184971Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:09:37.105600Z digest=sha256:329fce9d939a6543bc183e0dee92b8379da1e54873fa7c8c8dccf28ce46ab141

Observation c436a5a2-c01f-4921-a1a4-93e42c68ecb2 · outbound

This paper cites GPT-4 Technical Report.

Rethinking Multilingual Vision-Language Translation: Dataset, Evaluation, and Adaptation GPT-4 Technical Report

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T04:09:37.234227Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:09:37.234227Z digest=sha256:c81d6f86b1ffc7249130e945261be035ab29a34e51f112548bbe3aa369dbb695

Observation 27a7eb21-c56a-44ab-8bee-7ce0b70a4792 · outbound

This paper cites an unresolved cited work.

Rethinking Multilingual Vision-Language Translation: Dataset, Evaluation, and Adaptation Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:09:40.038387Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:09:37.356553Z digest=sha256:470fa545a90f34b7a1f3e853e6213552d015945e0af5cf3f3d9a4e5099488264

Observation f6425688-e5e3-4c1d-ac57-b5a1b924ea92 · outbound

This paper cites AnyTrans: Translate AnyText in the Image with Large Scale Models.

Rethinking Multilingual Vision-Language Translation: Dataset, Evaluation, and Adaptation AnyTrans: Translate AnyText in the Image with Large Scale Models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T04:09:37.454148Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:09:37.454148Z digest=sha256:b54b68dbfc9b6abe9d0f89df48f8328b71cf221a41456a2c6f1b4bb2226e756b

Observation e49542db-372d-439b-b646-7c932d159184 · outbound

This paper cites an unresolved cited work.

Rethinking Multilingual Vision-Language Translation: Dataset, Evaluation, and Adaptation Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:09:39.890062Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:09:37.579741Z digest=sha256:18915a1cea426136c089da7c3d676dc6c30e0c2b2c1421b6f8d7d37a884af57d

Observation 352fe44c-dc6f-4756-849e-989a9caf09d1 · outbound

This paper cites an unresolved cited work.

Rethinking Multilingual Vision-Language Translation: Dataset, Evaluation, and Adaptation Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:09:39.763487Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:09:37.753785Z digest=sha256:339764ff696f805f17eb5e1c4d1cadea4f59740b65b0e60cb22eb8ee39f51cc4

Observation 3a25f781-e255-471e-b826-4df209a5197e · outbound

This paper cites an unresolved cited work.

Rethinking Multilingual Vision-Language Translation: Dataset, Evaluation, and Adaptation Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:09:39.625251Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:09:37.875638Z digest=sha256:0839248040b891a6428db4605e81d14bfc0749995b63e1a06e97b03e55315515

Observation 8eaa5ac8-6fba-48bd-877f-4f8836538ceb · outbound

This paper cites an unresolved cited work.

Rethinking Multilingual Vision-Language Translation: Dataset, Evaluation, and Adaptation Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:09:39.468213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:09:38.008146Z digest=sha256:a774ca5556bfb112434396d8c9bdcb4301fb6c7c96a4ae5659427bbe0592589d

Observation 6c3d87d5-6e49-4bb3-b4f6-44b90f62adb2 · outbound

This paper cites an unresolved cited work.

Rethinking Multilingual Vision-Language Translation: Dataset, Evaluation, and Adaptation Unresolved cited work

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T04:09:38.151001Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:09:38.151001Z digest=sha256:d848de769e0a4e34a6063a8d9f541d4791150dd7fe3a541aa579e068796603ad

Observation 00cb081d-4089-4d71-9413-23de75b09525 · outbound

This paper cites DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding.

Rethinking Multilingual Vision-Language Translation: Dataset, Evaluation, and Adaptation DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T04:09:38.305643Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:09:38.305643Z digest=sha256:fd214aca033b7afd80fae8dcb9be60751fb7169ca215c74e03ee35e5d75460c6

Observation 666eba29-f95c-49a0-9801-dba863f9e9b1 · outbound

This paper cites Qwen2.5 Technical Report.

Rethinking Multilingual Vision-Language Translation: Dataset, Evaluation, and Adaptation Qwen2.5 Technical Report

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T04:09:38.380048Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:09:38.380048Z digest=sha256:dda0e61ee220c3572749c5aaa91adac74e18516b77b80ec9b6fb8b29b0a9f370

Observation c4292ea1-2568-4750-ae9c-1109fad454ab · outbound

This paper cites CC-OCR: A Comprehensive and Challenging OCR Benchmark for Evaluating Large Multimodal Models in Literacy.

Rethinking Multilingual Vision-Language Translation: Dataset, Evaluation, and Adaptation CC-OCR: A Comprehensive and Challenging OCR Benchmark for Evaluating Large Multimodal Models in Literacy

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T04:09:38.470496Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:09:38.470496Z digest=sha256:7935ae2140ce0c76341ae1840c13bcad5495d83ec433446117a554ad713e5d9d

Observation f36073cc-ae6b-47bb-a1c1-cd5d021c379b · outbound

This paper cites MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models.

Rethinking Multilingual Vision-Language Translation: Dataset, Evaluation, and Adaptation MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models

Reference 40

Resolution
malformed identifier
no resolver link, observed 2026-08-07T04:09:38.557854Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:09:38.557854Z digest=sha256:b76958d145c0098107595b7caf641e5971d6d7069499933106ba447e00ac6644

Observation e2b7ff45-9a2c-4c19-9edb-b4fa4e9b5391 · outbound

This paper cites The State and Fate of Linguistic Diversity and Inclusion in the NLP World.

Rethinking Multilingual Vision-Language Translation: Dataset, Evaluation, and Adaptation The State and Fate of Linguistic Diversity and Inclusion in the NLP World

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-07T04:09:35.593856Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:09:35.593856Z digest=sha256:6177badb135374da05cd7e5f5cc3b8fb04e737181d8ccb324b799cd3a87c4d5a

Observation 5b8334f9-b787-4344-8067-57e8e8af5110 · outbound

This paper cites Chitranuvad: Adapting Multi-Lingual LLMs for Multimodal Translation.

Rethinking Multilingual Vision-Language Translation: Dataset, Evaluation, and Adaptation Chitranuvad: Adapting Multi-Lingual LLMs for Multimodal Translation

Reference 2025

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T04:09:39.064229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:09:35.787476Z digest=sha256:924526ae129557ca80751083e04786eb4841e4023be1e6f371a6de660052842b

Pith citing papers

Observation 959eaf23-72d1-44a7-b8c7-c507cf6f1cbe · inbound

VIDA: A dataset for Visually Dependent Ambiguity in Multimodal Machine Translation cites this paper.

VIDA: A dataset for Visually Dependent Ambiguity in Multimodal Machine Translation Rethinking Multilingual Vision-Language Translation: Dataset, Evaluation, and Adaptation

Reference 103

Resolution
verified exact
arxiv_id, observed 2026-05-09T05:50:25.683678Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-08T19:34:02.860270Z digest=sha256:e285f4b20a342abd930b232d881285e30c5aaeec2cdc87de7a2bdce7aa7b2465

Observation 29de0f1f-ec21-48f3-bb35-39b29add18b7 · inbound

UniTranslator: A Unified Multi-modal Framework for End-to-end In-Image Machine Translation cites this paper.

UniTranslator: A Unified Multi-modal Framework for End-to-end In-Image Machine Translation Rethinking Multilingual Vision-Language Translation: Dataset, Evaluation, and Adaptation

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-07-04T16:09:57.561630Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-26T00:50:22.839005Z digest=sha256:210823321be4ad0cffd9a5f1889d7064ca796f36e7dc9f7e41891e19f94cb9c7