Pith. sign in

Paper Citation Record · LEDGER

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model

As of 15 August 2026, this Paper Citation Record lists 67 of 67 outbound references and 0 inbound Pith citation observations for arXiv:2505.23358.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.23358 v1

Coverage vector

measured 67 of 67 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:52:46.278075Z

measured 67 of 67 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

67 of 67 outbound references displayed

  • verified exact8
  • verified fuzzy21
  • unresolved37
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation e70c888e-fa71-4752-9764-9110e71fe033 · outbound

This paper cites Anderson, X.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Anderson, X

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:38.225153Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:52:38.225153Z digest=sha256:2638fbdbde4dc67c2b9cec4454c47bbd8592e8aa8bc54bd63a250601b4a79fd1

Observation 4888885c-d95a-475b-9753-6f939a93b7e2 · outbound

This paper cites Cornia, M.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Cornia, M

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:58.633849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T12:52:38.329252Z digest=sha256:9df60254ed67676858b5f0a9d2e12852b9b6ab3714a07b453e5517bcb817ddd6

Observation 1e9e79fd-4936-415a-bfc2-6d29ce0e0484 · outbound

This paper cites Vinyals, A.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Vinyals, A

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:58.403872Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T12:52:38.400669Z digest=sha256:deb977cd66416ca4072cae7f63d5eef834a0bab8fb1b63dd3ee3da3abbdffaf2

Observation e7a65481-e849-446c-9b33-c414de9c311c · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:58.054236Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T12:52:38.489602Z digest=sha256:b6137495645c94b1504dd81177bfcf372880cefdfb0961fc2d4c6bbe47a150d2

Observation f9435839-2957-4c63-9835-47d5448e576c · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:57.704900Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T12:52:38.569375Z digest=sha256:e0cea953da1efe53ab3c6db5df16e844a88b90f78a9540cbfaaabb2255a7f4a4

Observation f6af9ec9-d65a-4cc5-a68e-deaa09cf7908 · outbound

This paper cites Stefanini, M.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Stefanini, M

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:57.397582Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T12:52:38.647327Z digest=sha256:e7dad1190e800de041387c98db621fb7df22df2ec65d6a0ca5182cab7052d031

Observation fef9e153-5601-4eaf-968a-29db34242249 · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 7

Resolution
verified exact
doi, observed 2026-08-07T12:52:46.868814Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T12:52:38.727922Z digest=sha256:b32a009567c2eab9a7111cd52df10acdb48e90c53b7002918140edf767d92e46

Observation 785ad6f3-02cf-40b6-ae2b-bfaed1889868 · outbound

This paper cites Multi-Modal Image Captioning for the Visually Impaired.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Multi-Modal Image Captioning for the Visually Impaired

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:52:48.959440Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T12:52:38.796139Z digest=sha256:7285a034211c43430f329beb8a5849c75f8e5c96e0e6e6523afb3cbade842b71

Observation 029b2ba3-8538-4874-a84c-66bbcea5383d · outbound

This paper cites Image Captioning as an Assistive Technology: Lessons Learned from VizWiz 2020 Challenge.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Image Captioning as an Assistive Technology: Lessons Learned from VizWiz 2020 Challenge

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:52:48.712342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T12:52:38.871685Z digest=sha256:fba7dc5feb88a102bfad9cee5600042514be2bd59816caaa14566fdc83e202ef

Observation 75d3e471-50d6-41aa-b64f-e410ff65e797 · outbound

This paper cites Captioning Images Taken by People Who Are Blind.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Captioning Images Taken by People Who Are Blind

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:38.976502Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:52:38.976502Z digest=sha256:6a0d930a1ecf54d16e9b86fae61ad98cff7ef8ad836345e255f5caeaa563c711

Observation f159e2fb-340f-4bf6-957e-dfbea85de435 · outbound

This paper cites Faurina, A.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Faurina, A

Reference 11

Resolution
verified exact
doi, observed 2026-08-07T12:52:46.561064Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T12:52:39.078368Z digest=sha256:1837b01ae2af21d3085bf9d59b709cd4403d33a2ea8c1e8ec8bc499e7526aee6

Observation ac40d6e8-8e45-4d00-b667-8e002468660b · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:57.078394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T12:52:39.151464Z digest=sha256:899ffb11b898728440eb7c52e36b674e18b0d412adebe52e81c8ded3a834e040

Observation be2580ea-5407-4093-a27d-0affbbdd62ab · outbound

This paper cites Nikiforova, T.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Nikiforova, T

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:56.773631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T12:52:39.232914Z digest=sha256:4b2707ee4a1d830ac72c29334b23cc175f12f315e4e62f7f659b1271f0e7209d

Observation dd4c3f21-c962-4f7f-b8a0-c02c1272136c · outbound

This paper cites Informative Image Captioning with External Sources of Information.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Informative Image Captioning with External Sources of Information

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:39.321919Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:52:39.321919Z digest=sha256:73ff217850da8de03c808b5e701592d573444ee282545da8d55c43c11a8a4ea3

Observation 5d799102-8f3a-42d4-bca7-8b91f9706c54 · outbound

This paper cites NOC-REK: Novel Object Captioning with Retrieved Vocabulary from External Knowledge.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model NOC-REK: Novel Object Captioning with Retrieved Vocabulary from External Knowledge

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:39.390400Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:52:39.390400Z digest=sha256:8d8199ffc0ad67cdf04884933ed05f04df386b23268e2732fc6bcbf8741fe5fb

Observation 5fd0f2c5-6556-4c46-9c11-51f0167ff03a · outbound

This paper cites Boosting Entity-aware Image Captioning with Multi-modal Knowledge Graph.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Boosting Entity-aware Image Captioning with Multi-modal Knowledge Graph

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:52:48.458488Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T12:52:39.476170Z digest=sha256:d9777cf6a337a9d59a40f2325f128c1ca968e25baa4d019655d4fa2ef688d8d3

Observation 60d63913-eb1e-4af9-abd2-d5caaf9684ea · outbound

This paper cites Entity-aware Image Caption Generation.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Entity-aware Image Caption Generation

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:39.553652Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:52:39.553652Z digest=sha256:87fc4c973a847dc3bf7ee5f82b65a92131f24eab3dd78517e3533e279333d7a2

Observation bbb191e8-7261-4676-8baf-c2bd74b465ac · outbound

This paper cites EVCap: Retrieval-Augmented Image Captioning with External Visual-Name Memory for Open-World Comprehension.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model EVCap: Retrieval-Augmented Image Captioning with External Visual-Name Memory for Open-World Comprehension

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:52:48.254957Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T12:52:39.641589Z digest=sha256:5a10cee14cb1f9432f0ce425ffd50329bb89cc1f576029492eeecab511537ebc

Observation 87ec3b0d-016b-4d07-a39a-87c165e74893 · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:56.473669Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T12:52:39.752067Z digest=sha256:5c31d9a75b8894436f21cfc881629db758dafc09720e71f9cb9a8f2ed9daa2e7

Observation 0b82bc37-5fdc-46fd-8889-aa1c7f230e41 · outbound

This paper cites Zhang, X.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Zhang, X

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:56.216551Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T12:52:39.866012Z digest=sha256:d61a594e08826ca6358d6dd9c6edf09db282cf531f8662d6ca00008d296745e3

Observation 903d26c9-c6ec-40ce-8b66-e5489d9accc2 · outbound

This paper cites Ayesha, J.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Ayesha, J

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:55.875785Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T12:52:40.022083Z digest=sha256:de3396219dc8e8d08160b1329b8eea10e6fd637e9ba1e88ecf09cc6ea9175da0

Observation f931b5b8-6cc6-44b8-9613-0f4090991634 · outbound

This paper cites Joint Image Captioning and Question Answering.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Joint Image Captioning and Question Answering

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:52:48.040856Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T12:52:40.177749Z digest=sha256:97108dab0cd968d81ab44335c06b2af32855823b308962a702d7d1569071a914

Observation 554effaf-be40-4e95-9c68-f3535c39c000 · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:55.602175Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T12:52:40.237589Z digest=sha256:49968c719a4902ce4ca76815c6a6e3de353fabbdc718095b85688d95003f54f4

Observation 54812251-928e-4475-90ea-ce0238a75fe7 · outbound

This paper cites Salaberria, G.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Salaberria, G

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:55.352987Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T12:52:40.385666Z digest=sha256:9bbe0f7b0390734b2e2a34ee19f11fb88f67fdb39cc66c76c2ccaa2b0e00ba9e

Observation 7c0678fa-8fff-4c8a-b8d4-380ad916ee57 · outbound

This paper cites Enhancing Visual Question Answering through Question-Driven Image Captions as Prompts.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Enhancing Visual Question Answering through Question-Driven Image Captions as Prompts

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:52:47.869346Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T12:52:40.470416Z digest=sha256:87d99cafcabd20b0fbe88c9e1fa450dce3ea73f534625a3a97b4160c6395bec7

Observation 62d4644b-8acb-453a-ba90-f045a5abbfc8 · outbound

This paper cites Image Captioning and Visual Question Answering Based on Attributes and External Knowledge.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Image Captioning and Visual Question Answering Based on Attributes and External Knowledge

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:40.597487Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:52:40.597487Z digest=sha256:30e00f76e2e23d1a1a0299f95090fa765c78464a71fd86d7a33e6c4fe3969eda

Observation 5c33a43f-692b-48c9-a3a5-187a92206d95 · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:55.015213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T12:52:40.746781Z digest=sha256:995ca1527c36b6d942c31da6f058c0b803488e2344a7a1707103ec0d27e4852f

Observation bdc05ffb-3341-4d4f-a6b6-b089287f8993 · outbound

This paper cites Whitehead, H.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Whitehead, H

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:54.736213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T12:52:40.912671Z digest=sha256:08ad99d0cb27e06463f3c82785576a4ae7949c0f058be40ead18b9f52add30ab

Observation f3755080-71f4-45c2-9779-3bd814c6b63d · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:54.551892Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T12:52:41.044280Z digest=sha256:a61602382db0ef47ef30ad6c857baeeabb0c03314d74c97458952d82f7cb3259

Observation 785807f6-025b-422a-9fd9-ac259b927aaa · outbound

This paper cites Radford, J.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Radford, J

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:54.357001Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T12:52:41.156367Z digest=sha256:7954a8b6faa2709cb410dd9fffe7e6887b1eeb4d9279705776377f5cf88d2a8f

Observation 9ca17ee5-e151-4c6a-83e3-e9bf56c682a4 · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:54.188963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T12:52:41.242756Z digest=sha256:38038c87b52fa24e35201715b37858bfdb57ee0e6eaf1e21508d8d1d4d338f3e

Observation 9bbf83b5-4d06-42c1-a951-73890c52ff38 · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:53.939757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T12:52:41.347860Z digest=sha256:8cee5786de11c0f2bf1fc8b3f0055c5cc97b4c8232cbad5a51976546543d7f19

Observation 77bd1637-6379-4ef3-8d25-928d3b36ff97 · outbound

This paper cites SimVLM: Simple Visual Language Model Pretraining with Weak Supervision.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model SimVLM: Simple Visual Language Model Pretraining with Weak Supervision

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:41.477507Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:52:41.477507Z digest=sha256:a7ccc3679fffb7189bfd3f6f269102669e8979786b7ab1117656305244f45970

Observation c5dbae6e-3d28-445d-a8a7-7d60e9ac6351 · outbound

This paper cites Zhang, X.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Zhang, X

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:53.718800Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T12:52:41.620139Z digest=sha256:c28badf0dcd3d3e92ac42b1e08b4c58ef8ca4bb63016a9bea3e2560f22d5207c

Observation e21a4c6f-2011-4bb3-83f2-fca4edc42986 · outbound

This paper cites Cheng, W.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Cheng, W

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:41.737519Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:52:41.737519Z digest=sha256:9e727b41d2fa0fde85937327a1b98bdac54669e6411c093bdc84fc8f75d6c200

Observation 63038f23-68e0-4cf7-b811-4a6d07d8714b · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:41.863756Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:52:41.863756Z digest=sha256:9b03ee5042d73924c9c57524c70550a43c36fd066f1a4ce4bbf7749aaa8271a2

Observation 9bf901f3-42d4-4796-b3f9-57dca0f93082 · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 37

Resolution
metadata mismatch
raw_fallback, observed 2026-08-07T12:52:47.355468Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T12:52:41.935467Z digest=sha256:ba5484e7723068e3a1a5e1a2e760fb7eb3547bd735dd65582392f55743b5c828

Observation aa067ddc-360f-49ec-ac90-0e63bab9460a · outbound

This paper cites Cheng, W.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Cheng, W

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:42.077219Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:52:42.077219Z digest=sha256:2aa7eb4d002d725e57def9ee5a0753f497d7602e6d63a50da17c9376d3794d70

Observation e94ec38d-69ec-4738-a967-b5faadbeec74 · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 39

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:53.500630Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T12:52:42.260304Z digest=sha256:3a2845b56dcbf0cdd7f0a16fec9920ae838c2cd472d848988bfc47418de22325

Observation bb68489d-5d11-41d8-9982-ea5f287d7e21 · outbound

This paper cites Zhang, Z.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Zhang, Z

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:53.308178Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T12:52:42.414916Z digest=sha256:876fa88c31f48ffee4c517fec1194a7afea633e8f6371367457c296b8583fa7b

Observation 317e716f-d6f1-4b0d-b7e0-5c577690f75f · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:53.118429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T12:52:42.564442Z digest=sha256:9df0bc73efd0844a8620ba2f844f1d4fd0512d4e65e0df86c6167a68d78de5d0

Observation ead8fbd5-99c9-4163-9a60-66e4898bc5d6 · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 42

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:52.872374Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T12:52:42.746129Z digest=sha256:f2c6402f6274a00d2c42ac5d220c01ef266b43bb91c16cc1e8a0f57557a61691

Observation 2ad7f8fa-a2e5-4653-9fc5-a25fb201a9c9 · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 43

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:52.726239Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T12:52:42.902878Z digest=sha256:5492ea8f0cb65e3fa0979332c9516fc14569115f0af48d5d1b586a21293b504c

Observation a1497428-a809-49e1-ba46-d42f066e3c26 · outbound

This paper cites Huang, W.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Huang, W

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:52.516230Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T12:52:43.080542Z digest=sha256:68a757b501032c61c4f8354fd6ea06df91a60c4839fdafbc8f5b6a169b7c85a7

Observation 7eb8d239-1849-4376-b212-2f30817b08e1 · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:43.291761Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:52:43.291761Z digest=sha256:835d16b8e86632988cb07fe568725ac281c7068780492152d39b5df3bfa2ab3b

Observation 939a91d4-e703-4647-8e57-dbec012f5be4 · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 46

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:52.328504Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T12:52:43.376452Z digest=sha256:5c034487fc78d8b4fbd85664110393d51e8ea4378433b5d3d0b3a6d3f8e4d2c0

Observation 83882436-fe97-4f28-b95c-1947616b957c · outbound

This paper cites Devlin, M.-W.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Devlin, M.-W

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:52.113059Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T12:52:43.502994Z digest=sha256:e079b50010ca0444e1d60127acc78bcff907f34fbcf0618cbb87307a537f5601

Observation 0c5efa97-c4d9-4a2a-a5ba-41674afba9bc · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 48

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:51.880237Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T12:52:43.665941Z digest=sha256:1b3d3b2a00dc20e545463392faac433eabd02693b38a4ab7a84e1b394797263d

Observation 68447c61-88fc-49cb-8413-4300f61ddb0a · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 49

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:51.717387Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T12:52:43.799953Z digest=sha256:e0e696faaaf5189bb163a63fb070bebe274a003f027e250e3793362d03749e91

Observation e97e983c-750b-499f-b0fe-c605f3f115da · outbound

This paper cites Radford, J.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Radford, J

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:51.485927Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T12:52:43.943310Z digest=sha256:014e68f69e748020528f57e65bb54e3fa868325cb33d6c1fa0161e99dd6710de

Observation ab95742f-af91-4837-a8be-418c7492f131 · outbound

This paper cites ClipCap: CLIP Prefix for Image Captioning.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model ClipCap: CLIP Prefix for Image Captioning

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:44.090056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:52:44.090056Z digest=sha256:d0b4e09d41bb575082e20e0ffdb0ad9a45dce29e9ce2561c9f3e80321c2c0d0b

Observation 583f908b-8132-4b1d-a22e-58bdab4df765 · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 52

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:51.296324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T12:52:44.217497Z digest=sha256:771a5495c3004ad558d40ad69e2f7102edfc927b3fe08b7a17f50d427cec11b5

Observation d6987114-ce4b-4866-b9c4-b3af430e460a · outbound

This paper cites GIT: A Generative Image-to-text Transformer for Vision and Language.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model GIT: A Generative Image-to-text Transformer for Vision and Language

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:44.360480Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:52:44.360480Z digest=sha256:4d5747c4086e0f581e30231f3f469007dfd8a7bd26ccff694d5ad329038a54f8

Observation f33e1d90-fa99-4a87-ba0d-ebb038698a5e · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 54

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:51.143213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T12:52:44.524819Z digest=sha256:cd29af12098eb371fc8a374d0eea24159180dfa280c9ec7981ad7a1d3b66c36d

Observation 619ce196-557d-4234-a37b-8be37228eae5 · outbound

This paper cites 13035–13045.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model 13035–13045

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:50.963362Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T12:52:44.675623Z digest=sha256:e615bcb67df3cc4f86d551992296d1527bdaf0c6011ae42c1a8f680a81ac7d33

Observation 771492b5-9dd1-4e5a-908c-30499c358417 · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 56

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:50.768170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T12:52:44.829181Z digest=sha256:ac6bae0f27bb64381e1e52560a5485aede4e7502af21938c4748c5ed5417b902

Observation eaa26ee6-5c18-4998-ad25-3464b4440290 · outbound

This paper cites Changpinyo, P.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Changpinyo, P

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:44.950358Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:52:44.950358Z digest=sha256:b1d623448a9d220fe3ff5e9c2485a78d8d7522e29f66870dd85ac0204dea210b

Observation 1faceee6-58c1-42c1-a747-6b0968ee5f17 · outbound

This paper cites Kullback, R.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Kullback, R

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:50.556760Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T12:52:45.091862Z digest=sha256:aa7e6686e05e144d2a04eafa7285e0755932836a9bdd3872b18709aff2e04f51

Observation a9da7bcf-96d6-4e97-aa27-d303f3179f4c · outbound

This paper cites Microsoft COCO Captions: Data Collection and Evaluation Server.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Microsoft COCO Captions: Data Collection and Evaluation Server

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:45.252289Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:52:45.252289Z digest=sha256:5f577e2f98baa29bc67a44fea8825de88c5dab4fdbab85e663c7a4ae5c82faf2

Observation 7632d89f-de66-4896-af43-b06f96587e43 · outbound

This paper cites Papineni, S.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Papineni, S

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:50.405269Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T12:52:45.322179Z digest=sha256:c9faf3b4dfc6d93a16bd5a1988cf33a1624f64795d09a648439263b7365e2a54

Observation 3504e205-b35a-4e11-a653-83c1f2407f11 · outbound

This paper cites Banerjee, A.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Banerjee, A

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:50.178964Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T12:52:45.478971Z digest=sha256:018df76070981c7f74524b5862ce33bf2d46ab39459a776692a2e1f4d0ded240

Observation 1aea2639-4e4d-4a37-ba4e-f76a1944a962 · outbound

This paper cites Lin, Rouge: A package for automatic evaluation of summaries, in: Text summarization branches out: Proceedings of the ACL-04 workshop, 2004, pp.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Lin, Rouge: A package for automatic evaluation of summaries, in: Text summarization branches out: Proceedings of the ACL-04 workshop, 2004, pp

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:50.009788Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T12:52:45.626619Z digest=sha256:d1d541906994da5fe341895dd7cca08c0ac982b98bc702578f3416a76e1e50cb

Observation 9225c830-2bce-45e3-a697-43c60fdde696 · outbound

This paper cites Vedantam, C.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Vedantam, C

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:49.793273Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T12:52:45.757986Z digest=sha256:03bf5e8dd0e571df996c932100df27333430aac6ea6fe91aa38ee56c2cfa12c9

Observation 301c5f23-0b3f-4bc1-8635-bf8b1b7805f3 · outbound

This paper cites Kirkpatrick, R.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Kirkpatrick, R

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:49.605423Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T12:52:45.859012Z digest=sha256:3035db082e364f690f3107ab1b7b392aa4ed51daa97214f79bf776d90a57afea

Observation 210021e2-202f-45f3-968e-251efdbd32d1 · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 65

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:49.376017Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T12:52:45.945542Z digest=sha256:719921df4fc88e42093ae40a6f3f17a432c3f2c38828d9daa2085ae917f48c43

Observation 57f827d0-533d-40e0-9a1d-80b3884f9dbc · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 66

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:49.146830Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T12:52:46.100120Z digest=sha256:5e88127e6dc8db6ef4ea4cd690ae57f74404e58b06d883747cb102fca7d8c1cf

Observation e0bfc597-ddae-4ac7-9bf9-2bf7e9f83ac8 · outbound

This paper cites CLIP-Adapter: Better Vision-Language Models with Feature Adapters.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model CLIP-Adapter: Better Vision-Language Models with Feature Adapters

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:46.278075Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:52:46.278075Z digest=sha256:bb99ff866d4e87f2e19500e87b3211b5fd86e7f1d4f003cb8a06961bc5879726

Pith citing papers

No inbound Pith citation observations are available.