Pith. sign in

Paper Citation Record · LEDGER

EchoSight: Advancing Visual-Language Models with Wiki Knowledge

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 14 inbound Pith citation observations for arXiv:2407.12735.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2407.12735 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 14 of 14 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 14 of 14 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:46:44.260422Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T10:59:46.886481Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 8efb0f53-c2e0-43c8-9e89-beba86e99e88 · inbound

Towards General Continuous Memory for Vision-Language Models cites this paper.

Towards General Continuous Memory for Vision-Language Models EchoSight: Advancing Visual-Language Models with Wiki Knowledge

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T14:46:44.260422Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:46:44.260422Z digest=sha256:251190a0cd632a4847d4ab775900e71e863b676710e31f62ebd4e9912fd14d8e

Observation 6585cc9f-7072-4fe3-a5a7-a27ac79cb5cc · inbound

Spa-VLM: Stealthy Poisoning Attacks on RAG-based VLM cites this paper.

Spa-VLM: Stealthy Poisoning Attacks on RAG-based VLM EchoSight: Advancing Visual-Language Models with Wiki Knowledge

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T13:21:39.948489Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:21:39.948489Z digest=sha256:df1ad7a221646bf4fe79bdd7ddc14d4fb7868c446376886f7e78ea806809d2be

Observation d3ef9275-f502-49c9-ab57-ffb0919c85a8 · inbound

Augmented Vision-Language Models: A Systematic Review cites this paper.

Augmented Vision-Language Models: A Systematic Review EchoSight: Advancing Visual-Language Models with Wiki Knowledge

Reference 121

Resolution
unresolved
no resolver link, observed 2026-08-06T14:33:43.532433Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:33:43.532433Z digest=sha256:f4bebd613f1040986043831adddcd9e4acbd4be57cf05ccc5597c267a02545eb

Observation f7871e68-32ff-4d91-81aa-d9d2307a3a28 · inbound

Empowering Multimodal LLMs with External Tools: A Comprehensive Survey cites this paper.

Empowering Multimodal LLMs with External Tools: A Comprehensive Survey EchoSight: Advancing Visual-Language Models with Wiki Knowledge

Reference 94

Resolution
unresolved
no resolver link, observed 2026-08-05T20:28:51.739526Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T20:28:51.739526Z digest=sha256:969e2f3983ab170d71db2274fef5ab51abf273ab6e99a76a57dae119eb3d3190

Observation 909a70fa-6c91-4148-aec9-d57a7798f601 · inbound

Wiki-R1: Incentivizing Multimodal Reasoning for Knowledge-based VQA via Data and Sampling Curriculum cites this paper.

Wiki-R1: Incentivizing Multimodal Reasoning for Knowledge-based VQA via Data and Sampling Curriculum EchoSight: Advancing Visual-Language Models with Wiki Knowledge

Reference 35

Resolution
unresolved
no resolver link, observed 2026-07-15T14:43:21.866055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-15T14:43:21.866055Z digest=sha256:dc8d7325679d0eb10cbef1580d2dde18bb62e593920cda298356eadb1e15ca2e

Observation b105937d-eadc-428e-bd23-551f0953ac4b · inbound

WikiCLIP: An Efficient Contrastive Baseline for Open-domain Visual Entity Recognition cites this paper.

WikiCLIP: An Efficient Contrastive Baseline for Open-domain Visual Entity Recognition EchoSight: Advancing Visual-Language Models with Wiki Knowledge

Reference 43

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T13:15:50.573740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T13:11:54.384284Z digest=sha256:74cd8b5730d5e38b5b2415a428092d35cbef5b4491cd7ead35bb28404f6c60c9

Observation e30d155d-b208-49da-850f-831c8dcfda72 · inbound

WikiCLIP: An Efficient Contrastive Baseline for Open-domain Visual Entity Recognition cites this paper.

WikiCLIP: An Efficient Contrastive Baseline for Open-domain Visual Entity Recognition EchoSight: Advancing Visual-Language Models with Wiki Knowledge

Reference 43

Resolution
unresolved
no resolver link, observed 2026-07-14T23:55:24.006436Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T23:55:24.006436Z digest=sha256:755c3fc6f761ff06ba99dfc1838cdbb109490640a9d5772c3ae9bd614a171a64

Observation 2b7e3776-9f5f-4795-8230-af053ee6119e · inbound

MCERF: Advancing Multimodal LLM Evaluation of Engineering Documentation with Enhanced Retrieval cites this paper.

MCERF: Advancing Multimodal LLM Evaluation of Engineering Documentation with Enhanced Retrieval EchoSight: Advancing Visual-Language Models with Wiki Knowledge

Reference 33

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T09:30:48.227529Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-16T09:29:32.250418Z digest=sha256:457aa1a0351215dc68053ba8b88038e4e87a0466e2c9072468b0e6eca17e1415

Observation be4e461e-93f0-4222-a7a9-83751cd35bc2 · inbound

WikiVQABench: A Knowledge-Grounded Visual Question Answering Benchmark from Wikipedia and Wikidata cites this paper.

WikiVQABench: A Knowledge-Grounded Visual Question Answering Benchmark from Wikipedia and Wikidata EchoSight: Advancing Visual-Language Models with Wiki Knowledge

Reference 100

Resolution
verified exact
arxiv_id, observed 2026-05-21T04:43:58.727343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-21T04:41:06.779529Z digest=sha256:afe68e20d3a4f56f104a3cdbc43d05b589df24cd9013b31a968b74ad26b427f9

Observation e42bb48c-b815-47bc-a373-7db31c3079ff · inbound

MathVis-Fine: Aligning Visual Supervision with Necessity via Progressive Dependency-Guided Training for Multimodal Mathematical Reasoning cites this paper.

MathVis-Fine: Aligning Visual Supervision with Necessity via Progressive Dependency-Guided Training for Multimodal Mathematical Reasoning EchoSight: Advancing Visual-Language Models with Wiki Knowledge

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-07-03T20:18:57.617210Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-27T01:23:40.564561Z digest=sha256:44667dbbb1a92ef5b16b74d6654e161b3d7997054ff389fb8dbde7cb1a8780cc

Observation 61ffe059-d1ae-4d04-950f-47add1214d65 · inbound

Ground Then Rank: Revisiting Knowledge-Based VQA with Training-Free Entity Identification cites this paper.

Ground Then Rank: Revisiting Knowledge-Based VQA with Training-Free Entity Identification EchoSight: Advancing Visual-Language Models with Wiki Knowledge

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-07-04T10:59:46.888382Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-26T08:12:14.829556Z digest=sha256:d131342ac0a06d9700091e98694a7a2fe9e522bad7705e6ee49dc64deb568f54

Observation 27be4eae-a4ff-43a8-8cbe-3dece275ac9f · inbound

Reason Before You Retrieve: Agentic Planning for Multi-modal RAG cites this paper.

Reason Before You Retrieve: Agentic Planning for Multi-modal RAG EchoSight: Advancing Visual-Language Models with Wiki Knowledge

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-02T10:20:50.309084Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T10:20:50.309084Z digest=sha256:a9c0cf7c205223a3a4bf3ea57b911bc8eb1c228b4edf2c389656ac1d41295938

Observation d3e5af33-0afa-4b71-ba0c-f91d25a5e579 · inbound

UniHEAR: Unified Heterogeneous-Source Attentive Retrieval for Knowledge-Based Visual Question Answering cites this paper.

UniHEAR: Unified Heterogeneous-Source Attentive Retrieval for Knowledge-Based Visual Question Answering EchoSight: Advancing Visual-Language Models with Wiki Knowledge

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T00:31:17.396355Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:31:17.396355Z digest=sha256:2df19e5ab4ff946d20d22542660add654c4147571705ffa302eb674e732023df

Observation bd2f5435-1540-49da-8431-66d651fd345d · inbound

UniHEAR: Unified Heterogeneous-Source Attentive Retrieval for Knowledge-Based Visual Question Answering cites this paper.

UniHEAR: Unified Heterogeneous-Source Attentive Retrieval for Knowledge-Based Visual Question Answering EchoSight: Advancing Visual-Language Models with Wiki Knowledge

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T04:19:53.199025Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T04:19:53.199025Z digest=sha256:734e38153723ad656feecfd7f5814cc2a9108487ca6475c076275d33b64e21ee