Pith. sign in

Paper Citation Record · LEDGER

ViDove: A Translation Agent System with Multimodal Context and Memory-Augmented Reasoning

As of 21 August 2026, this Paper Citation Record lists 12 of 12 outbound references and 1 inbound Pith citation observation for arXiv:2507.07306.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.07306 v1

Coverage vector

measured 12 of 12 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T18:48:51.453332Z

measured 13 of 13 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-08T19:34:02.860270Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-09T05:50:25.837175Z

Reference resolution

12 of 12 outbound references displayed

  • verified exact0
  • verified fuzzy3
  • unresolved8
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a9bd8f55-93ce-4bbc-87d3-99ed0d237ec9 · outbound

This paper cites 0:00:01,229.

ViDove: A Translation Agent System with Multimodal Context and Memory-Augmented Reasoning 0:00:01,229

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:48:52.564849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T18:48:51.266705Z digest=sha256:e3f9d8bbcd8c398adcfb9427d328a34223999bd36f668afd15380bd616fc3c96

Observation fbc4129c-cc86-4fa1-8c96-27f2c8d7ac72 · outbound

This paper cites - Translate into natural, fluent Simplified Chinese.

ViDove: A Translation Agent System with Multimodal Context and Memory-Augmented Reasoning - Translate into natural, fluent Simplified Chinese

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:48:52.301134Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T18:48:51.354932Z digest=sha256:ed254451640f4ab80fd55cbdcbd7f88d35acb25c9396a78ec3e3626a06be789c

Observation 49dbd9a3-455b-4399-bfc5-c3124132ee7f · outbound

This paper cites Qwen2-Audio Technical Report.

ViDove: A Translation Agent System with Multimodal Context and Memory-Augmented Reasoning Qwen2-Audio Technical Report

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:50.400665Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:50.400665Z digest=sha256:0617cebb82a6ab1064026c245c790cd86aa6675a8a24f942be64c3aa8e398eaf

Observation eade5100-c891-44b2-bbc9-5497669b11b7 · outbound

This paper cites Generative Multi-Modal Knowledge Retrieval with Large Language Models.

ViDove: A Translation Agent System with Multimodal Context and Memory-Augmented Reasoning Generative Multi-Modal Knowledge Retrieval with Large Language Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:50.667682Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:50.667682Z digest=sha256:8ebf5a32b7c9402d66a0a392c43e6249cacc01da747e7729804524de4ebafbc4

Observation 40b6a302-24ba-4964-af30-91310563ccd1 · outbound

This paper cites Low-Resource Machine Translation through Retrieval-Augmented LLM Prompting: A Study on the Mambai Language.

ViDove: A Translation Agent System with Multimodal Context and Memory-Augmented Reasoning Low-Resource Machine Translation through Retrieval-Augmented LLM Prompting: A Study on the Mambai Language

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:50.814311Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:50.814311Z digest=sha256:205900634470c4988ea68034648f80f4656d27071c86f5a2d877c046f0076618

Observation 507589dc-73e6-4233-8273-186ec0158572 · outbound

This paper cites A Survey on Multi-modal Machine Translation: Tasks, Methods and Challenges.

ViDove: A Translation Agent System with Multimodal Context and Memory-Augmented Reasoning A Survey on Multi-modal Machine Translation: Tasks, Methods and Challenges

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:51.081327Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:51.081327Z digest=sha256:ae84977c93fba999676935280dfe82a65b76a4addd8a26e8473ca4e43d9bb74b

Observation 756f39ab-2379-4120-8206-c9aba4205673 · outbound

This paper cites 翻译提供的视频中的说话内容到中文。只需要输出翻译内容原文,不要输出任何解释。.

ViDove: A Translation Agent System with Multimodal Context and Memory-Augmented Reasoning 翻译提供的视频中的说话内容到中文。只需要输出翻译内容原文,不要输出任何解释。

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:48:52.090711Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T18:48:51.453332Z digest=sha256:7ea1b71a95fc49430f3ac573944013b8a864cd503da9dc3e96c95d2fa2483d54

Observation 6c36e146-7be7-4520-bc03-e7fef0e2f1da · outbound

This paper cites Google's Neural Machine Translation System: Bridging the Gap between Human and Machine Translation.

ViDove: A Translation Agent System with Multimodal Context and Memory-Augmented Reasoning Google's Neural Machine Translation System: Bridging the Gap between Human and Machine Translation

Reference 2016

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:51.181622Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:51.181622Z digest=sha256:87b42f23577ba563abb273205503dd7a4950ba377d41a5b38fee1c70c45562b4

Observation 60b8b508-fee2-4c63-a1d5-756ac76760a8 · outbound

This paper cites How to Design Translation Prompts for ChatGPT: An Empirical Study.

ViDove: A Translation Agent System with Multimodal Context and Memory-Augmented Reasoning How to Design Translation Prompts for ChatGPT: An Empirical Study

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:50.559294Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:50.559294Z digest=sha256:c3f5301f9598c436e35ec3b49a21ca1d6b595277f0885663123c7c3b44a7ffa8

Observation a052cd3b-a9f7-4471-9a7d-561997f55ad8 · outbound

This paper cites Thibault Sellam, Dipanjan Das, and Ankur P Parikh.

ViDove: A Translation Agent System with Multimodal Context and Memory-Augmented Reasoning Thibault Sellam, Dipanjan Das, and Ankur P Parikh

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:50.931379Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:50.931379Z digest=sha256:4c1dab5cdb98b9fe3265e1a2a76b1dad164f66c2f1e5209f2256858a8bab0668

Observation 29858307-8c49-4cdb-a116-dd6566757533 · outbound

This paper cites Retrieving Examples from Memory for Retrieval Augmented Neural Machine Translation: A Systematic Comparison.

ViDove: A Translation Agent System with Multimodal Context and Memory-Augmented Reasoning Retrieving Examples from Memory for Retrieval Augmented Neural Machine Translation: A Systematic Comparison

Reference 2024

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T18:48:51.849051Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T18:48:50.286314Z digest=sha256:c97732ba1cbb1a3fec59c2874d1955f4ecf9a73c0d8ab8edcd4862a5523bd165

Observation 5bd26649-41f8-44b1-a26e-b69870a8c3f1 · outbound

This paper cites Qwen2.5-VL Technical Report.

ViDove: A Translation Agent System with Multimodal Context and Memory-Augmented Reasoning Qwen2.5-VL Technical Report

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-06T18:48:50.189858Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:48:50.189858Z digest=sha256:aeb02adeb48e3f8642ff7ff0f256dd35769c2d3e2648a42ecd285d025057ee63

Pith citing papers

Observation 59257bbb-5a73-44df-9dbc-4678817bf33b · inbound

VIDA: A dataset for Visually Dependent Ambiguity in Multimodal Machine Translation cites this paper.

VIDA: A dataset for Visually Dependent Ambiguity in Multimodal Machine Translation ViDove: A Translation Agent System with Multimodal Context and Memory-Augmented Reasoning

Reference 78

Resolution
verified exact
arxiv_id, observed 2026-05-09T05:50:25.838847Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-05-08T19:34:02.860270Z digest=sha256:30e624a5ab480c5c6fd721aa04ce727d7fa63af60efdf51935f2a41bd89c7bf4