Pith. sign in

Paper Citation Record · LEDGER

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model

As of 7 August 2026, this Paper Citation Record lists 67 of 67 outbound references and 0 inbound Pith citation observations for arXiv:2505.23358.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.23358 v1

Coverage vector

measured 67 of 67 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:52:46.278075Z

measured 67 of 67 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

67 of 67 outbound references displayed

  • verified exact8
  • verified fuzzy21
  • unresolved37
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation e70c888e-fa71-4752-9764-9110e71fe033 · outbound

This paper cites Anderson, X.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Anderson, X

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:38.225153Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:52:38.225153Z digest=sha256:2cc7a479e3e817c3203fd2fcc168833691255751d8ca1cbfa660b2130a3e5f87

Observation 4888885c-d95a-475b-9753-6f939a93b7e2 · outbound

This paper cites Cornia, M.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Cornia, M

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:58.633849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:38.329252Z digest=sha256:f7c4adc883d3259bbdf36b54ce8375cd7c34edf71873d7317a0ee046aaf5bcf9

Observation 1e9e79fd-4936-415a-bfc2-6d29ce0e0484 · outbound

This paper cites Vinyals, A.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Vinyals, A

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:58.403872Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:38.400669Z digest=sha256:7092acf45352a6a84366e79b5209c9f35c726d59f4b99260385d78169b2ddf95

Observation e7a65481-e849-446c-9b33-c414de9c311c · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:58.054236Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:38.489602Z digest=sha256:44cce68ba9ca073f19bdf3a6306f034e0b1a9def562a58832debe1195684a591

Observation f9435839-2957-4c63-9835-47d5448e576c · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:57.704900Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:38.569375Z digest=sha256:1f74940437e85181264da7e0f213944922062a6316a3785198202fdaf57a4273

Observation f6af9ec9-d65a-4cc5-a68e-deaa09cf7908 · outbound

This paper cites Stefanini, M.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Stefanini, M

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:57.397582Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:38.647327Z digest=sha256:d25ec5908b1b8a145717f5b42c847bf2efdb811cfc9c79a3465c06197102214b

Observation fef9e153-5601-4eaf-968a-29db34242249 · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 7

Resolution
verified exact
doi, observed 2026-08-07T12:52:46.868814Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:38.727922Z digest=sha256:7b61c21c7b6b554e68c3ea315daf406cada19e63f48ea2c1398b1b1dcedefb1c

Observation 785ad6f3-02cf-40b6-ae2b-bfaed1889868 · outbound

This paper cites Multi-Modal Image Captioning for the Visually Impaired.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Multi-Modal Image Captioning for the Visually Impaired

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:52:48.959440Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:38.796139Z digest=sha256:ddd2f65a1fe387d1357aa92ac7cedebe1a666bfe01a2762a8ffecaca2e4c4127

Observation 029b2ba3-8538-4874-a84c-66bbcea5383d · outbound

This paper cites Image Captioning as an Assistive Technology: Lessons Learned from VizWiz 2020 Challenge.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Image Captioning as an Assistive Technology: Lessons Learned from VizWiz 2020 Challenge

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:52:48.712342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:38.871685Z digest=sha256:27fd2dfb4ffbe72c516782f0952177c7f6f34b4076a67ef510a699ae144fa1eb

Observation 75d3e471-50d6-41aa-b64f-e410ff65e797 · outbound

This paper cites Captioning Images Taken by People Who Are Blind.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Captioning Images Taken by People Who Are Blind

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:38.976502Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:52:38.976502Z digest=sha256:6bd3724df995adfed0fb17f7ac9cd0dd4aeca7fc830c8ddbb04dfe30b1f5a48c

Observation f159e2fb-340f-4bf6-957e-dfbea85de435 · outbound

This paper cites Faurina, A.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Faurina, A

Reference 11

Resolution
verified exact
doi, observed 2026-08-07T12:52:46.561064Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:39.078368Z digest=sha256:e592e19c57287bf45a2e8dbea7992db36ce93da1a9b69de72b2fbce5db7a22e5

Observation ac40d6e8-8e45-4d00-b667-8e002468660b · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:57.078394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:39.151464Z digest=sha256:a6394d7815cd49eed65b7a06621e1400de45b67345bf107011ce16380eaf8369

Observation be2580ea-5407-4093-a27d-0affbbdd62ab · outbound

This paper cites Nikiforova, T.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Nikiforova, T

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:56.773631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:39.232914Z digest=sha256:774195a385d846692bf8858f294135c85506d6388017af28f2ea9ff412c1867e

Observation dd4c3f21-c962-4f7f-b8a0-c02c1272136c · outbound

This paper cites Informative Image Captioning with External Sources of Information.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Informative Image Captioning with External Sources of Information

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:39.321919Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:52:39.321919Z digest=sha256:b7fd8565ac1d63acf3453628f0f6cfcbc6139371fc79cf799baf6f05d0031253

Observation 5d799102-8f3a-42d4-bca7-8b91f9706c54 · outbound

This paper cites NOC-REK: Novel Object Captioning with Retrieved Vocabulary from External Knowledge.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model NOC-REK: Novel Object Captioning with Retrieved Vocabulary from External Knowledge

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:39.390400Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:52:39.390400Z digest=sha256:35247fd19afa785afbef35113f4f2b5a4d51e6f3a3d1e05763699ae59814d995

Observation 5fd0f2c5-6556-4c46-9c11-51f0167ff03a · outbound

This paper cites Boosting Entity-aware Image Captioning with Multi-modal Knowledge Graph.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Boosting Entity-aware Image Captioning with Multi-modal Knowledge Graph

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:52:48.458488Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:39.476170Z digest=sha256:d4ba85dff30401a16a46498a4c43fdf0cfaf06c51da2b2b0217d44dbbce4a3f2

Observation 60d63913-eb1e-4af9-abd2-d5caaf9684ea · outbound

This paper cites Entity-aware Image Caption Generation.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Entity-aware Image Caption Generation

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:39.553652Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:52:39.553652Z digest=sha256:f70961fb7c4953f74f59124771980940a9f4f8c7d772ad318d738888e3e9c8ea

Observation bbb191e8-7261-4676-8baf-c2bd74b465ac · outbound

This paper cites EVCap: Retrieval-Augmented Image Captioning with External Visual-Name Memory for Open-World Comprehension.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model EVCap: Retrieval-Augmented Image Captioning with External Visual-Name Memory for Open-World Comprehension

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:52:48.254957Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:39.641589Z digest=sha256:22ae31a44faab74cc7e6ce5c728b8a0c525777fb24b739ac045886f58884db0c

Observation 87ec3b0d-016b-4d07-a39a-87c165e74893 · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:56.473669Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:39.752067Z digest=sha256:9571b8d11d1f7c5c79cd5d3b45401f176382d9123fdf64593cfa9ae4ab20365a

Observation 0b82bc37-5fdc-46fd-8889-aa1c7f230e41 · outbound

This paper cites Zhang, X.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Zhang, X

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:56.216551Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:39.866012Z digest=sha256:e5879ec68bd2308e07b256abbba11d86612734a3c521e69bbcb515613b611970

Observation 903d26c9-c6ec-40ce-8b66-e5489d9accc2 · outbound

This paper cites Ayesha, J.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Ayesha, J

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:55.875785Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:40.022083Z digest=sha256:d2f81335fcb0eee722e169e9fe3dc8be88c4fd08d9ee1b02b869d848b3e15bed

Observation f931b5b8-6cc6-44b8-9613-0f4090991634 · outbound

This paper cites Joint Image Captioning and Question Answering.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Joint Image Captioning and Question Answering

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:52:48.040856Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:40.177749Z digest=sha256:a3d142980b5a924b99f190ebd18bc90285da9f92919d2c56cbae50c851c75734

Observation 554effaf-be40-4e95-9c68-f3535c39c000 · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:55.602175Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:40.237589Z digest=sha256:abf625245d70fc75c0f4708cdf4331359a73362958990cb04c4382be40ebfb60

Observation 54812251-928e-4475-90ea-ce0238a75fe7 · outbound

This paper cites Salaberria, G.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Salaberria, G

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:55.352987Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:40.385666Z digest=sha256:ce4fe3bd10085193a87ff36d2d752b1960f1ad1ebce9c64687aea274ce5aa95e

Observation 7c0678fa-8fff-4c8a-b8d4-380ad916ee57 · outbound

This paper cites Enhancing Visual Question Answering through Question-Driven Image Captions as Prompts.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Enhancing Visual Question Answering through Question-Driven Image Captions as Prompts

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:52:47.869346Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:40.470416Z digest=sha256:18b699b04bd048ebfe81b4ebbc3ba1373e708a08371f882cc3ad7bad587ac3fb

Observation 62d4644b-8acb-453a-ba90-f045a5abbfc8 · outbound

This paper cites Image Captioning and Visual Question Answering Based on Attributes and External Knowledge.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Image Captioning and Visual Question Answering Based on Attributes and External Knowledge

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:40.597487Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:52:40.597487Z digest=sha256:7bd121e8ba741d366bc3d70642d3d6648c1f18b85e175e04a4e1a114d6befb9c

Observation 5c33a43f-692b-48c9-a3a5-187a92206d95 · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:55.015213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:40.746781Z digest=sha256:37bc24c8adfbdb32b2ba2e2d7332584d9336c88d92c319a4a00b47f08523e838

Observation bdc05ffb-3341-4d4f-a6b6-b089287f8993 · outbound

This paper cites Whitehead, H.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Whitehead, H

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:54.736213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:40.912671Z digest=sha256:3e2c9b20761b95a24cf4a2e7cdd0f5c01f87974bb4f6e6628af1327c36086c6e

Observation f3755080-71f4-45c2-9779-3bd814c6b63d · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:54.551892Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:41.044280Z digest=sha256:d00fe0dce44995939bf0b02499444565bf9afba31917e0f373fa616ee004bb2b

Observation 785807f6-025b-422a-9fd9-ac259b927aaa · outbound

This paper cites Radford, J.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Radford, J

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:54.357001Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:41.156367Z digest=sha256:8826382c2bea2e472fce21a6415049ee60bb72c96cbae3d920e7fea56917131d

Observation 9ca17ee5-e151-4c6a-83e3-e9bf56c682a4 · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:54.188963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:41.242756Z digest=sha256:8ab71aae2fce864d301d9ab164d4c8e89ca7159a103a99ecd3f32f6a83e35eae

Observation 9bbf83b5-4d06-42c1-a951-73890c52ff38 · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:53.939757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:41.347860Z digest=sha256:72796dd2edde19633c8dc406c3043e6552f43f644ac5e46cd4e533df8e71269e

Observation 77bd1637-6379-4ef3-8d25-928d3b36ff97 · outbound

This paper cites SimVLM: Simple Visual Language Model Pretraining with Weak Supervision.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model SimVLM: Simple Visual Language Model Pretraining with Weak Supervision

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:41.477507Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:52:41.477507Z digest=sha256:3961cf069aca9a20c3d35cc9cd9849fe16a74c54c497403c9cf2bbb1d6b7ef92

Observation c5dbae6e-3d28-445d-a8a7-7d60e9ac6351 · outbound

This paper cites Zhang, X.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Zhang, X

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:53.718800Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:41.620139Z digest=sha256:a4bd4cc677c21dc0d19001ed407bf34d21ccc048a02c81d18f08e3d1932f62c8

Observation e21a4c6f-2011-4bb3-83f2-fca4edc42986 · outbound

This paper cites Cheng, W.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Cheng, W

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:41.737519Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:52:41.737519Z digest=sha256:1e979da2a0546aefaf47e53fe9a0380c7945e7b85d1425297de8cc0aaec47aed

Observation 63038f23-68e0-4cf7-b811-4a6d07d8714b · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:41.863756Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:52:41.863756Z digest=sha256:08033de3acad529bc99443a95bf4efbc49f23b935916d81cb42cc8ce06b9004c

Observation 9bf901f3-42d4-4796-b3f9-57dca0f93082 · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 37

Resolution
metadata mismatch
raw_fallback, observed 2026-08-07T12:52:47.355468Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:41.935467Z digest=sha256:b3a77077e6e15b9180ccf90b5129a1757afdfca7b4cf63d22faf26e4bb3648f6

Observation aa067ddc-360f-49ec-ac90-0e63bab9460a · outbound

This paper cites Cheng, W.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Cheng, W

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:42.077219Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:52:42.077219Z digest=sha256:67df9347815d78e87d27c875aec6f4cb381d1482bb7de052d20db919d4a0ebab

Observation e94ec38d-69ec-4738-a967-b5faadbeec74 · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 39

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:53.500630Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:42.260304Z digest=sha256:7f1cd82bf67851b5235700434e6ceae37f07695a6f5ae66892bac4dc3f1389a2

Observation bb68489d-5d11-41d8-9982-ea5f287d7e21 · outbound

This paper cites Zhang, Z.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Zhang, Z

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:53.308178Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:42.414916Z digest=sha256:dbd2dbd66e75bbfd464168f62fad27c58eb960128aa0c9030e85d6ecab75971c

Observation 317e716f-d6f1-4b0d-b7e0-5c577690f75f · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:53.118429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:42.564442Z digest=sha256:0835d634f61d25634a223b46a6f07dae6c3145c0445e0fb7c35607f529fa8db7

Observation ead8fbd5-99c9-4163-9a60-66e4898bc5d6 · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 42

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:52.872374Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:42.746129Z digest=sha256:cdf3741127861a26ab295c79e322c94d929783411403e7535d88978ff6070365

Observation 2ad7f8fa-a2e5-4653-9fc5-a25fb201a9c9 · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 43

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:52.726239Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:42.902878Z digest=sha256:73fe8afef1fd1c017107349d5cbc5fe3f82172e9fa02d942ef3c362e9b9dba9d

Observation a1497428-a809-49e1-ba46-d42f066e3c26 · outbound

This paper cites Huang, W.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Huang, W

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:52.516230Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:43.080542Z digest=sha256:441cbffac394a02ca6ae0349b1a1b52a75eb3db285f7c5ae028490ff365769c6

Observation 7eb8d239-1849-4376-b212-2f30817b08e1 · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:43.291761Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:52:43.291761Z digest=sha256:89faf1de660620867834118b65ddadda8fcf48f5212527ccaa07f65590eeca53

Observation 939a91d4-e703-4647-8e57-dbec012f5be4 · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 46

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:52.328504Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:43.376452Z digest=sha256:f39abf5e3e42f79047773610260bc9a8b8d193a8d75e921284ba47641a4e0d53

Observation 83882436-fe97-4f28-b95c-1947616b957c · outbound

This paper cites Devlin, M.-W.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Devlin, M.-W

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:52.113059Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:43.502994Z digest=sha256:179e0c4eb1aef158b1f6367a7b9d4b170e34eeafeeccc98330fb79582bd88b84

Observation 0c5efa97-c4d9-4a2a-a5ba-41674afba9bc · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 48

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:51.880237Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:43.665941Z digest=sha256:93b71d224c1dfec79e7b94ac9470f8ef8dbf644cce191d86a7f04421c6afba1f

Observation 68447c61-88fc-49cb-8413-4300f61ddb0a · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 49

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:51.717387Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:43.799953Z digest=sha256:0f94e2f6b8f17328dc5871b2eafcb845a064d6d646368c97d5129b10154aefe1

Observation e97e983c-750b-499f-b0fe-c605f3f115da · outbound

This paper cites Radford, J.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Radford, J

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:51.485927Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:43.943310Z digest=sha256:180301e60a587ae34e388c47aadf5af9dfa539140af75370462f5ea7b7564132

Observation ab95742f-af91-4837-a8be-418c7492f131 · outbound

This paper cites ClipCap: CLIP Prefix for Image Captioning.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model ClipCap: CLIP Prefix for Image Captioning

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:44.090056Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:52:44.090056Z digest=sha256:df0e7213e971f63bfe89448a856030fc436016587fe8ee2a38179a18c3d43373

Observation 583f908b-8132-4b1d-a22e-58bdab4df765 · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 52

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:51.296324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:44.217497Z digest=sha256:6d6ed973185a43ba0b232589d9f9c024f35e4ef7bddc7ac0f4982a74a28922ce

Observation d6987114-ce4b-4866-b9c4-b3af430e460a · outbound

This paper cites GIT: A Generative Image-to-text Transformer for Vision and Language.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model GIT: A Generative Image-to-text Transformer for Vision and Language

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:44.360480Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:52:44.360480Z digest=sha256:3511d58b0f604322a7c06c4c59b0ab629d3ebe6ba047e1b1892217dffeea95e9

Observation f33e1d90-fa99-4a87-ba0d-ebb038698a5e · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 54

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:51.143213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:44.524819Z digest=sha256:edee01f91fb73dd0e1b207f3883effcbff2779bf7fe15cdce5ee5a84987cf053

Observation 619ce196-557d-4234-a37b-8be37228eae5 · outbound

This paper cites 13035–13045.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model 13035–13045

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:50.963362Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:44.675623Z digest=sha256:4e0a08759edb558aa41f12d8fd0119451001fcd234a0eae49a66082058cb211e

Observation 771492b5-9dd1-4e5a-908c-30499c358417 · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 56

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:50.768170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:44.829181Z digest=sha256:dfa3ce6c24535ddd2a9918a63190365adbbca19881c2131877988bcc04262802

Observation eaa26ee6-5c18-4998-ad25-3464b4440290 · outbound

This paper cites Changpinyo, P.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Changpinyo, P

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:44.950358Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:52:44.950358Z digest=sha256:a93a8ad92d38c348e0ab640b9d3abbff48856c9eab32308ceefdc098ac33c1f8

Observation 1faceee6-58c1-42c1-a747-6b0968ee5f17 · outbound

This paper cites Kullback, R.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Kullback, R

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:50.556760Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:45.091862Z digest=sha256:f512aa0f8b55fc750f17b557878d57504053ecc17322a51789ff6b6ff13cfb68

Observation a9da7bcf-96d6-4e97-aa27-d303f3179f4c · outbound

This paper cites Microsoft COCO Captions: Data Collection and Evaluation Server.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Microsoft COCO Captions: Data Collection and Evaluation Server

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:45.252289Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:52:45.252289Z digest=sha256:f820142f05d9bda2c5a95d4998b8c13946e7146b1be23e87f97728ffb68e1757

Observation 7632d89f-de66-4896-af43-b06f96587e43 · outbound

This paper cites Papineni, S.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Papineni, S

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:50.405269Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:45.322179Z digest=sha256:a6332e0a284ecc828fb87f8a73e0db1cd99c7f051a5401d54b6fdd344955e651

Observation 3504e205-b35a-4e11-a653-83c1f2407f11 · outbound

This paper cites Banerjee, A.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Banerjee, A

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:50.178964Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:45.478971Z digest=sha256:e769d035a0c92aa78f92349b29eac0080ecac1a5dec18017baa9b0f04a59a463

Observation 1aea2639-4e4d-4a37-ba4e-f76a1944a962 · outbound

This paper cites Lin, Rouge: A package for automatic evaluation of summaries, in: Text summarization branches out: Proceedings of the ACL-04 workshop, 2004, pp.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Lin, Rouge: A package for automatic evaluation of summaries, in: Text summarization branches out: Proceedings of the ACL-04 workshop, 2004, pp

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:50.009788Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:45.626619Z digest=sha256:6c7229417485e285b0ae410480b20e7b9b61a69be14d5793465e046a2a21d404

Observation 9225c830-2bce-45e3-a697-43c60fdde696 · outbound

This paper cites Vedantam, C.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Vedantam, C

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:49.793273Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:45.757986Z digest=sha256:4069909e3615950d5afe786de258b1cfdd923632fc8e99e8a5c02677a7d88d5e

Observation 301c5f23-0b3f-4bc1-8635-bf8b1b7805f3 · outbound

This paper cites Kirkpatrick, R.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Kirkpatrick, R

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:52:49.605423Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:45.859012Z digest=sha256:0993c2f957746cf9c8e4e4abeb60d0e0970d1e678085353d4ee4bcc1ec1106e0

Observation 210021e2-202f-45f3-968e-251efdbd32d1 · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 65

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:49.376017Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:45.945542Z digest=sha256:07704fc9d8a2780104ec9d410160ae6c9c37e2afca7608872849393f695953ac

Observation 57f827d0-533d-40e0-9a1d-80b3884f9dbc · outbound

This paper cites an unresolved cited work.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model Unresolved cited work

Reference 66

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:52:49.146830Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:52:46.100120Z digest=sha256:7e84084eabb8b23fb78683b013d17e435c1b161e4b91f92f1ff27744b7992937

Observation e0bfc597-ddae-4ac7-9bf9-2bf7e9f83ac8 · outbound

This paper cites CLIP-Adapter: Better Vision-Language Models with Feature Adapters.

Beam-Guided Knowledge Replay for Knowledge-Rich Image Captioning using Vision-Language Model CLIP-Adapter: Better Vision-Language Models with Feature Adapters

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:46.278075Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:52:46.278075Z digest=sha256:9cc5afc2be19271669b604bc2a4ecf2f8da24fd9dfd4b1d9aaf6bf2ed7bc30a7

Pith citing papers

No inbound Pith citation observations are available.