Pith. sign in

Paper Citation Record · LEDGER

From Image Captioning to Visual Storytelling

As of 16 August 2026, this Paper Citation Record lists 93 of 93 outbound references and 0 inbound Pith citation observations for arXiv:2508.14045.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.14045 v1

Coverage vector

measured 93 of 93 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T10:27:39.491216Z

measured 93 of 93 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

93 of 93 outbound references displayed

  • verified exact34
  • verified fuzzy0
  • unresolved56
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch3

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 460933a4-410f-404f-a920-cb7a238ba56a · outbound

This paper cites URL: " 'urlintro :=.

From Image Captioning to Visual Storytelling URL: " 'urlintro :=

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.142990Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.142990Z digest=sha256:440802f2c940f28258e1269b757cef468f3594658a38616692d39071dbd95778

Observation 68db7014-2d91-40f8-9766-8128851954a9 · outbound

This paper cites write newline.

From Image Captioning to Visual Storytelling write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.147551Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.147551Z digest=sha256:f705a8a28892e9faeb5a7caa35f12efb85f0cba1fac78d9faba3bf399c396b2d

Observation 67917612-ccc8-49d8-a5d3-50d109e71a17 · outbound

This paper cites SPICE: Semantic Propositional Image Caption Evaluation.

From Image Captioning to Visual Storytelling SPICE: Semantic Propositional Image Caption Evaluation

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.151762Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.151762Z digest=sha256:4b9596ddcd051cc015a7f863d20fac0b213c5b2736b12f9c6946bda85e0206b3

Observation 8922e562-aa55-4287-95b4-a4f6c09712ef · outbound

This paper cites Bottom-Up and Top-Down Attention for Image Captioning and Visual Question Answering.

From Image Captioning to Visual Storytelling Bottom-Up and Top-Down Attention for Image Captioning and Visual Question Answering

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.156320Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.156320Z digest=sha256:3ecc6a47e3eb71b62baf38485d0676fe4729481a6eefb5ce7c648fe07bf92cac

Observation 8208d6aa-a0b3-4ca8-8279-123b8ef4fb3c · outbound

This paper cites TouchStone: Evaluating Vision-Language Models by Language Models.

From Image Captioning to Visual Storytelling TouchStone: Evaluating Vision-Language Models by Language Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.160499Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.160499Z digest=sha256:acc8e1299c84f27f01611b1cb5ad0f1e30ebc4310f878522aaedd839cc7d92db

Observation 19d49c2c-5034-483f-885d-61a32c798eae · outbound

This paper cites an unresolved cited work.

From Image Captioning to Visual Storytelling Unresolved cited work

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.164465Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.164465Z digest=sha256:3359594e9070432c2d6ff482185fac504bb99198062fb40a4d461e0a6c492611

Observation 1136e2c7-61fc-4591-b176-f46da3e3f86a · outbound

This paper cites Commonsense Knowledge Aware Concept Selection For Diverse and Informative Visual Storytelling.

From Image Captioning to Visual Storytelling Commonsense Knowledge Aware Concept Selection For Diverse and Informative Visual Storytelling

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-08-06T10:27:40.555602Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.168379Z digest=sha256:0b209050571de1400cdf9a53a6dc83508346e9df934cde245212d9aeed7f735e

Observation fc08cee6-7010-4466-8aef-9b7c28a56ec1 · outbound

This paper cites an unresolved cited work.

From Image Captioning to Visual Storytelling Unresolved cited work

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.172413Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.172413Z digest=sha256:339916049cb34333834fb0b263024f14787a706bdcb2a5136dfd251c86ac370c

Observation ee8efa95-40e1-4047-9f60-72b97afb0630 · outbound

This paper cites TARN-VIST: Topic Aware Reinforcement Network for Visual Storytelling.

From Image Captioning to Visual Storytelling TARN-VIST: Topic Aware Reinforcement Network for Visual Storytelling

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-08-06T10:27:40.457057Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.176728Z digest=sha256:70bf6769cb264e8ef87c1f99445a4c469c95694e8eb42945bbb5093875adbf4f

Observation 1e2fad49-b672-4775-bce9-b8e9a9f4ad87 · outbound

This paper cites an unresolved cited work.

From Image Captioning to Visual Storytelling Unresolved cited work

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.180421Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.180421Z digest=sha256:6f1aabbadcb51998c53ac116165cff729b57c1c7507ca09d0ed28fc85dfe5a99

Observation 09d7a993-ac83-4163-a3f9-7861dc75b291 · outbound

This paper cites an unresolved cited work.

From Image Captioning to Visual Storytelling Unresolved cited work

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.183894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.183894Z digest=sha256:6f597c990efc79894c4a845a3baa4deb8d159fd20f39f5342f41168c36b4da86

Observation 6662c5ec-dd13-476d-8be4-0d40515305e6 · outbound

This paper cites Exploring Nearest Neighbor Approaches for Image Captioning.

From Image Captioning to Visual Storytelling Exploring Nearest Neighbor Approaches for Image Captioning

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-08-06T10:27:40.442624Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.187962Z digest=sha256:a8346516606e19e4f5efce3d0f90a2f9f3a8104c210d48d0c29a30194006e0ab

Observation 01b27af5-ae41-49e7-a259-165e2d91c103 · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

From Image Captioning to Visual Storytelling An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.191648Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.191648Z digest=sha256:6ba191923dabf8fb2d3012c2a9fd9a55e494e04950787831225c691469c8ee50

Observation fbed2d2d-eadb-49c3-b8e3-090c433f12ad · outbound

This paper cites an unresolved cited work.

From Image Captioning to Visual Storytelling Unresolved cited work

Reference 14

Resolution
metadata mismatch
raw_fallback, observed 2026-08-06T10:27:40.417408Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.195374Z digest=sha256:b72cd707633844730d0aa77c166855d3cf70eeec15259481bae60f2026032af2

Observation fb83ea58-8d96-44da-96d7-3690c62f155d · outbound

This paper cites Transformer-based Conditional Variational Autoencoder for Controllable Story Generation.

From Image Captioning to Visual Storytelling Transformer-based Conditional Variational Autoencoder for Controllable Story Generation

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-08-06T10:27:40.343040Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.199030Z digest=sha256:f896b1f4ca40d3dee0fbc16fdb987d1009391f0cfc19dabc131093f2d3986e6b

Observation 78cb959e-289a-4eec-927b-980ea2f91043 · outbound

This paper cites an unresolved cited work.

From Image Captioning to Visual Storytelling Unresolved cited work

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.202850Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.202850Z digest=sha256:6e57bd564228e4bd865c12091b47bd9f091cdea27098e2c807d83a725bcecfed

Observation cfb859f3-90cd-4b9c-90a8-93cf080365af · outbound

This paper cites an unresolved cited work.

From Image Captioning to Visual Storytelling Unresolved cited work

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.206292Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.206292Z digest=sha256:04077b8b276548a8e52971badc45b736e95e49a00e0791a6920ecc8da63f0a7b

Observation 65897f70-5776-44ba-b7e1-953f6a2545d6 · outbound

This paper cites Contextualize, Show and Tell: A Neural Visual Storyteller.

From Image Captioning to Visual Storytelling Contextualize, Show and Tell: A Neural Visual Storyteller

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-08-06T10:27:40.328324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.209675Z digest=sha256:a891c230447c4225649609648186938978cdd2a0dfb50e47fd45d15ab1083d50

Observation 5319b260-92f2-4cf1-85d6-529914bd1dbd · outbound

This paper cites A Knowledge-Enhanced Pretraining Model for Commonsense Story Generation.

From Image Captioning to Visual Storytelling A Knowledge-Enhanced Pretraining Model for Commonsense Story Generation

Reference 19

Resolution
verified exact
local_arxiv, observed 2026-08-06T10:27:40.314505Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.213265Z digest=sha256:1ee96a624a7bd1f34406f3a139b5d036572bf678e2befc3e50c74427ca15eff3

Observation 6919e57d-be18-41dc-8699-d038eb3be25d · outbound

This paper cites VICTR: Visual Information Captured Text Representation for Text-to-Image Multimodal Tasks.

From Image Captioning to Visual Storytelling VICTR: Visual Information Captured Text Representation for Text-to-Image Multimodal Tasks

Reference 20

Resolution
verified exact
local_arxiv, observed 2026-08-06T10:27:40.300254Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.216713Z digest=sha256:55ac6011ceeb50f6062657beafa74c07a7b72dac36662c0bae225d28065d654f

Observation 77020388-c12f-4cc3-8929-a18ee2cd8ff7 · outbound

This paper cites Deep Residual Learning for Image Recognition.

From Image Captioning to Visual Storytelling Deep Residual Learning for Image Recognition

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.220359Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.220359Z digest=sha256:592c7ba9ae29162d1fdab8ce9ddd25187be8c012419fd0d8c1990ca00604f060

Observation 0578e642-1779-4e5f-9d47-9397ae743d20 · outbound

This paper cites Image Captioning through Image Transformer.

From Image Captioning to Visual Storytelling Image Captioning through Image Transformer

Reference 22

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T10:27:40.276721Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.223799Z digest=sha256:d4b0ddd4b3ce90a7f1e01d389bf3338682cfe017afbcc3ca524123a935305c82

Observation 2c6c43eb-392d-4435-bb10-1af549fdd19f · outbound

This paper cites Image Captioning: Transforming Objects into Words.

From Image Captioning to Visual Storytelling Image Captioning: Transforming Objects into Words

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.227854Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.227854Z digest=sha256:65ee91e2eab84362dcf20864c48ea08fd80a8dc98f1d93577f7ea4e7f71d8c30

Observation 7efec995-1c58-4e65-a53d-5a0a5e6d3878 · outbound

This paper cites Visual Writing Prompts: Character-Grounded Story Generation with Curated Image Sequences.

From Image Captioning to Visual Storytelling Visual Writing Prompts: Character-Grounded Story Generation with Curated Image Sequences

Reference 24

Resolution
verified exact
local_arxiv, observed 2026-08-06T10:27:40.252870Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.231546Z digest=sha256:2d2a567b18799d2d57ae5fe827e893722166813e147f199026bfadb9c82bd915

Observation 427b46ef-5ae3-46dc-a62b-30d54a780374 · outbound

This paper cites an unresolved cited work.

From Image Captioning to Visual Storytelling Unresolved cited work

Reference 25

Resolution
verified exact
doi, observed 2026-08-06T10:27:39.600141Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.235186Z digest=sha256:941c8b8bac23b6ca588fab726815580343f25e49f7c7458389c8dd166a064005

Observation 8e834285-e94b-4497-8653-3365b1d32738 · outbound

This paper cites Knowledge-Enriched Visual Storytelling.

From Image Captioning to Visual Storytelling Knowledge-Enriched Visual Storytelling

Reference 26

Resolution
verified exact
local_arxiv, observed 2026-08-06T10:27:40.237734Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.238764Z digest=sha256:7fb019abe0fbfeb46cd2b97fce62db08c36eb22ebec268ebeacb18e282ef6832

Observation 3e16745c-30a3-48a8-a26b-a174808360ba · outbound

This paper cites Plot and Rework: Modeling Storylines for Visual Storytelling.

From Image Captioning to Visual Storytelling Plot and Rework: Modeling Storylines for Visual Storytelling

Reference 27

Resolution
verified exact
local_arxiv, observed 2026-08-06T10:27:40.223637Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.242438Z digest=sha256:4cbfe14497de765fd6a3ea7f2642d0e83c34b20efe70d2602dec0c784a0abee9

Observation 4c1f7389-1b57-4464-b4c2-bdee5178d99b · outbound

This paper cites Visual Story Post-Editing.

From Image Captioning to Visual Storytelling Visual Story Post-Editing

Reference 28

Resolution
verified exact
local_arxiv, observed 2026-08-06T10:27:40.209796Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.246217Z digest=sha256:e51870d98baafeaa7b8d7a6ce181e006d249105a28f0c8bc307f410d70437b05

Observation 105aa1fc-90dc-4512-b17e-786916095155 · outbound

This paper cites What Makes A Good Story? Designing Composite Rewards for Visual Storytelling.

From Image Captioning to Visual Storytelling What Makes A Good Story? Designing Composite Rewards for Visual Storytelling

Reference 29

Resolution
verified exact
local_arxiv, observed 2026-08-06T10:27:40.194184Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.250137Z digest=sha256:764c99a8bb95b7e6c765c0c743f2fe84395e48ebaa8901699549ee5d5f23b1c8

Observation c67cc21f-a03e-440f-ae93-4237f8229808 · outbound

This paper cites Attention on Attention for Image Captioning.

From Image Captioning to Visual Storytelling Attention on Attention for Image Captioning

Reference 30

Resolution
verified exact
local_arxiv, observed 2026-08-06T10:27:40.179323Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.254800Z digest=sha256:662b7a37a1eb453cfc70b2a2f1f57f9f91ea3c239cd1991290290c6d6630082e

Observation 2ad04455-f1e3-4249-86cb-1b4716457b81 · outbound

This paper cites Lawrence Zitnick, Devi Parikh, Lucy Vanderwende, Galley Michel, and Mitchell Margaret.

From Image Captioning to Visual Storytelling Lawrence Zitnick, Devi Parikh, Lucy Vanderwende, Galley Michel, and Mitchell Margaret

Reference 31

Resolution
verified exact
doi, observed 2026-08-06T10:27:39.589103Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.258864Z digest=sha256:c0cc46a8527f4310eb7b049ef80c2c8c957b257d34cfac3e006bf540def8bb93

Observation a2cdddcc-5ee7-4e64-9a08-7c793b365491 · outbound

This paper cites Story Generation from Sequence of Independent Short Descriptions.

From Image Captioning to Visual Storytelling Story Generation from Sequence of Independent Short Descriptions

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.262362Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.262362Z digest=sha256:e5d49dff8e40d3f4c65860f538db60062f689ec9979e14affa8347203f795190

Observation cbf1c71d-8fe6-485f-8532-77576a34ca9f · outbound

This paper cites an unresolved cited work.

From Image Captioning to Visual Storytelling Unresolved cited work

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.266799Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.266799Z digest=sha256:e5560a27ec9b80920d171baefa38425aa8585b609c3f1c6ad6cd684dad879936

Observation 51554768-4817-4571-ac6e-fc292c53417a · outbound

This paper cites an unresolved cited work.

From Image Captioning to Visual Storytelling Unresolved cited work

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.270245Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.270245Z digest=sha256:d151a08600fe7d4a59187b04f34f43f51cc2cae0225059115470cab8bb1b56d7

Observation 419e3971-28c0-48b6-a6b4-af2f9f896912 · outbound

This paper cites GLAC Net: GLocal Attention Cascading Networks for Multi-image Cued Story Generation.

From Image Captioning to Visual Storytelling GLAC Net: GLocal Attention Cascading Networks for Multi-image Cued Story Generation

Reference 35

Resolution
verified exact
local_arxiv, observed 2026-08-06T10:27:40.154163Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.274281Z digest=sha256:85cd58265dfb3b4ea58b9c3511362649d83d87267ec0311791121a606c6f9e2f

Observation c8910c74-efb6-4850-a6c5-12a63ed028af · outbound

This paper cites Adam: A Method for Stochastic Optimization.

From Image Captioning to Visual Storytelling Adam: A Method for Stochastic Optimization

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.278137Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.278137Z digest=sha256:e4c873d67f96c970c5c7e45bcfc5a1a166a5d187b2d1bea8b88f4664963e8dd3

Observation 989190ee-ceb5-4ac2-ad7f-c3b8d92420f6 · outbound

This paper cites Visual Genome: Connecting Language and Vision Using Crowdsourced Dense Image Annotations.

From Image Captioning to Visual Storytelling Visual Genome: Connecting Language and Vision Using Crowdsourced Dense Image Annotations

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.281717Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.281717Z digest=sha256:1788adb39d601c6d9f0c98cb5fc53eba8d6360fb3bc8914e7f89ec4588744fea

Observation 2e146646-9bd6-4472-a3fd-c05ec8112898 · outbound

This paper cites an unresolved cited work.

From Image Captioning to Visual Storytelling Unresolved cited work

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.285646Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.285646Z digest=sha256:465686abe080cbe065a79bbd4f5c834d5e59decff98abc8a67e9d05a66d88785

Observation 903672af-7ab0-4424-88f5-2d33bebc4e7e · outbound

This paper cites Can large language models provide useful feedback on research papers? A large-scale empirical analysis.

From Image Captioning to Visual Storytelling Can large language models provide useful feedback on research papers? A large-scale empirical analysis

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.290156Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.290156Z digest=sha256:23d7ea4e2a3681e61101be4ed9cbf0569bc4bf08cc7e525805cf565c8921f2c6

Observation 3da6f782-8b0d-4469-a45b-045a2cc88fde · outbound

This paper cites an unresolved cited work.

From Image Captioning to Visual Storytelling Unresolved cited work

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.295020Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.295020Z digest=sha256:87e939adf0ea3db29fd2877888e5ac3f0b8f0fdbd60cb20b03b2010f3385d131

Observation 73029131-2d63-47bd-8138-b5e332d5cf80 · outbound

This paper cites Microsoft COCO: Common Objects in Context.

From Image Captioning to Visual Storytelling Microsoft COCO: Common Objects in Context

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.299512Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.299512Z digest=sha256:acf635712686f99b788567a7635817ddbff400eb16f535b15d6aed00a6448ad7

Observation 058e1eda-5b59-432e-b044-8063073aa986 · outbound

This paper cites Detecting and Grounding Important Characters in Visual Stories.

From Image Captioning to Visual Storytelling Detecting and Grounding Important Characters in Visual Stories

Reference 42

Resolution
verified exact
local_arxiv, observed 2026-08-06T10:27:40.101851Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.303627Z digest=sha256:278340f334370d714070eba94517b6884052a4d4bc35ab4314fc46fb4fa38cd4

Observation 6e3ab560-f4dc-4e6e-8d2c-973f8c351b55 · outbound

This paper cites Generating Visual Stories with Grounded and Coreferent Characters.

From Image Captioning to Visual Storytelling Generating Visual Stories with Grounded and Coreferent Characters

Reference 43

Resolution
verified exact
local_arxiv, observed 2026-08-06T10:27:40.087549Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.307440Z digest=sha256:5011019c996d8cc881c97ad1286e47dea170fd0d58ea8878b369eabe7d5c750b

Observation 57653fa1-3860-4cc6-a222-ce63084026dd · outbound

This paper cites an unresolved cited work.

From Image Captioning to Visual Storytelling Unresolved cited work

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.311506Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.311506Z digest=sha256:508a511048b3fe0ae711d17629d9a13f21a0c9ac67794386c701f66e731aa6f4

Observation da206540-3816-4e31-b2b8-531f42f8f6d2 · outbound

This paper cites Visual Instruction Tuning.

From Image Captioning to Visual Storytelling Visual Instruction Tuning

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.315063Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.315063Z digest=sha256:97079a5cafa2755b3efc58e0ccc5b5f989a8de0144c72a7409e5253a462b863f

Observation fdee897e-d3b5-402d-8946-069e6a5a8903 · outbound

This paper cites an unresolved cited work.

From Image Captioning to Visual Storytelling Unresolved cited work

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.318888Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.318888Z digest=sha256:c0b824f05d414c5cdfc24de5cc957c51865a0bc4b06395052b849f7e503acf17

Observation fffd135f-d63d-4b72-8ed5-9e1082cc1057 · outbound

This paper cites an unresolved cited work.

From Image Captioning to Visual Storytelling Unresolved cited work

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.322346Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.322346Z digest=sha256:afa0af5fe9b95909b03addd66413244b5393c72b3bafe71cc0a76806fed38a8f

Observation 43d4022d-768c-44cc-8564-49fa5ed5d58c · outbound

This paper cites Decoupled Weight Decay Regularization.

From Image Captioning to Visual Storytelling Decoupled Weight Decay Regularization

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.325723Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.325723Z digest=sha256:991ca23483f865d099256841b77465dfc3fd2fbe9315f7658790cf2dbef57b36

Observation e585f74d-cd12-4a2c-a89d-fea58f28007f · outbound

This paper cites ViLBERT: Pretraining Task-Agnostic Visiolinguistic Representations for Vision-and-Language Tasks.

From Image Captioning to Visual Storytelling ViLBERT: Pretraining Task-Agnostic Visiolinguistic Representations for Vision-and-Language Tasks

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.329615Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.329615Z digest=sha256:73bf920ef23b562d67f7ac2209472598f70633a11ff8ef89587e365b6d61d375

Observation 6c4c66f5-985b-494a-9897-bab94d50dfca · outbound

This paper cites Knowing When to Look: Adaptive Attention via A Visual Sentinel for Image Captioning.

From Image Captioning to Visual Storytelling Knowing When to Look: Adaptive Attention via A Visual Sentinel for Image Captioning

Reference 50

Resolution
verified exact
local_arxiv, observed 2026-08-06T10:27:40.043109Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.334282Z digest=sha256:d8e0a2085310325299450826578de245c5a5885f731e805fc2d1c59caad2e1b1

Observation 37b79eda-e591-486e-aa15-4fabec394308 · outbound

This paper cites ClipCap: CLIP Prefix for Image Captioning.

From Image Captioning to Visual Storytelling ClipCap: CLIP Prefix for Image Captioning

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.338214Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.338214Z digest=sha256:952ed8cfad55361a2ece965238a20fab22d3c5418b436812f46b2fba29bd3f7b

Observation a0104fdd-bca1-41ec-ab2c-d9117d4ee5fe · outbound

This paper cites Album Storytelling with Iterative Story-aware Captioning and Large Language Models.

From Image Captioning to Visual Storytelling Album Storytelling with Iterative Story-aware Captioning and Large Language Models

Reference 52

Resolution
verified exact
local_arxiv, observed 2026-08-06T10:27:40.017233Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.341873Z digest=sha256:34d771d1a3d42b4330953ff73335dd4ee67cea62c458dce3b276ba5afd1f3470

Observation ec36a8b8-11dd-4cfd-b7ad-2a5f9824ceee · outbound

This paper cites GPT-4 Technical Report.

From Image Captioning to Visual Storytelling GPT-4 Technical Report

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.345535Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.345535Z digest=sha256:8ca24b9269828915b31fd98bde70a99b92d8e4ef9f99f8117060d435427fc965

Observation 0f4f6d66-e639-4762-b23a-42658458b030 · outbound

This paper cites an unresolved cited work.

From Image Captioning to Visual Storytelling Unresolved cited work

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.349238Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.349238Z digest=sha256:cb6aeb6274dbd9e45f867557af6cea8c8c20f6f19de7fd4f34f76c1b7980f87a

Observation 687713ed-6311-4c42-892c-7f49a7474639 · outbound

This paper cites an unresolved cited work.

From Image Captioning to Visual Storytelling Unresolved cited work

Reference 55

Resolution
unresolved
raw_fallback, observed 2026-08-06T10:27:40.699746Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.353116Z digest=sha256:0575c52ab461eb6b89211ef75875ca1693dff23c185b249e8d61b28d5f2f9eae

Observation 554ae120-bd17-4590-901f-7cb3c32814d2 · outbound

This paper cites "My Way of Telling a Story": Persona based Grounded Story Generation.

From Image Captioning to Visual Storytelling "My Way of Telling a Story": Persona based Grounded Story Generation

Reference 56

Resolution
verified exact
local_arxiv, observed 2026-08-06T10:27:39.993159Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.356460Z digest=sha256:b037f65ecb4b7a3304b8cd584fac0d00790b97abe7a7b4015e3ab0642444231e

Observation 18cf3879-f740-46b8-a74e-c2dd89070af7 · outbound

This paper cites an unresolved cited work.

From Image Captioning to Visual Storytelling Unresolved cited work

Reference 57

Resolution
metadata mismatch
raw_fallback, observed 2026-08-06T10:27:39.978523Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.360161Z digest=sha256:35cef2170074c6e28d3926dbfe17df2366c2258867291fc8ffef4f75cad45016

Observation 043bfdfd-ae57-491c-b34e-a2cf088ec02b · outbound

This paper cites an unresolved cited work.

From Image Captioning to Visual Storytelling Unresolved cited work

Reference 58

Resolution
unresolved
raw_fallback, observed 2026-08-06T10:27:40.689280Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.363586Z digest=sha256:140d1e38615eb48430d143091e8c6480524075cf208db08bb96a491897ea4e0c

Observation e60aa79e-f929-44e9-aadd-fe864c611ff1 · outbound

This paper cites an unresolved cited work.

From Image Captioning to Visual Storytelling Unresolved cited work

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.367121Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.367121Z digest=sha256:475a7df915172083e921649bfef03d8a2b67c32a33a40d18242d0ac572cdfd25

Observation 952d01be-fd97-4645-b3dd-1138179d7cd9 · outbound

This paper cites an unresolved cited work.

From Image Captioning to Visual Storytelling Unresolved cited work

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.370515Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.370515Z digest=sha256:33bdfb956c0dc803080aebc4da266c61d525977b88c4b7f0682182d3412d6ad0

Observation 13985ff9-03a2-4143-9bed-a6cec49e6f69 · outbound

This paper cites an unresolved cited work.

From Image Captioning to Visual Storytelling Unresolved cited work

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.373867Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.373867Z digest=sha256:8bb95898071c16add22585b526ba564c871ec393d1906945f22596eb2875483d

Observation edead90b-c421-48a6-b33d-d747ba762ea9 · outbound

This paper cites an unresolved cited work.

From Image Captioning to Visual Storytelling Unresolved cited work

Reference 62

Resolution
unresolved
raw_fallback, observed 2026-08-06T10:27:40.660493Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.377562Z digest=sha256:8064845462a4367b445b70d83d688aaac55c0055a4cc87a66c6c22ebb2b7933d

Observation 94a30da2-4380-4351-9805-5a4ed67bfcc8 · outbound

This paper cites Self-critical Sequence Training for Image Captioning.

From Image Captioning to Visual Storytelling Self-critical Sequence Training for Image Captioning

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.381077Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.381077Z digest=sha256:d4446953e7df7a3a5abe952014a482249089338d5b21a787ffda327f3f0319ae

Observation 557f3364-c092-41db-96b4-739057dc3d61 · outbound

This paper cites an unresolved cited work.

From Image Captioning to Visual Storytelling Unresolved cited work

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.384650Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.384650Z digest=sha256:47a9472ea789bcefb250a1a472d59991e8d982848b94ff46bafec162c2919f33

Observation 217a277e-110f-4fa1-a021-7f4e7464098c · outbound

This paper cites an unresolved cited work.

From Image Captioning to Visual Storytelling Unresolved cited work

Reference 65

Resolution
unresolved
raw_fallback, observed 2026-08-06T10:27:40.649191Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.387909Z digest=sha256:19b6309089ffb08ea6999718d78f6d336a15bb09b39ebfe386ca9f68ce53a75e

Observation 231da63b-840f-4809-bb7e-7ac72fb0b435 · outbound

This paper cites an unresolved cited work.

From Image Captioning to Visual Storytelling Unresolved cited work

Reference 66

Resolution
verified exact
doi, observed 2026-08-06T10:27:39.565514Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.391257Z digest=sha256:9f9d047e9ea93abd918a22212eca2c1d5f5cf50a8c14cca7c54755a3f0d10aae

Observation 33fc1b98-a558-4f14-aa29-ef4449227539 · outbound

This paper cites Context-aware Visual Storytelling with Visual Prefix Tuning and Contrastive Learning.

From Image Captioning to Visual Storytelling Context-aware Visual Storytelling with Visual Prefix Tuning and Contrastive Learning

Reference 67

Resolution
verified exact
local_arxiv, observed 2026-08-06T10:27:39.895991Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.394862Z digest=sha256:6cee828e36fcb0c37e49ad12a72685a8763af813347faf1d4e2009d4c7edfdc2

Observation d456efa1-f29f-4fc5-a19c-783add29eff0 · outbound

This paper cites BERT-hLSTMs: BERT and Hierarchical LSTMs for Visual Storytelling.

From Image Captioning to Visual Storytelling BERT-hLSTMs: BERT and Hierarchical LSTMs for Visual Storytelling

Reference 68

Resolution
verified exact
local_arxiv, observed 2026-08-06T10:27:39.880542Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.398532Z digest=sha256:fbec4284e2526110cb3abf9f29e1e7c67b8b5a32e3feecc2ffd0106638b4815d

Observation 896b20bb-827b-4a89-a7f0-d91bfc9c3f1f · outbound

This paper cites A Contrastive Framework for Neural Text Generation.

From Image Captioning to Visual Storytelling A Contrastive Framework for Neural Text Generation

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.402209Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.402209Z digest=sha256:edbb136e393a7fab9edc56bc084383a63990da8f7912047a9154047832a875dc

Observation 3cf0b8ea-ae3c-4ba7-b452-e0a68339f0c8 · outbound

This paper cites GROOViST: A Metric for Grounding Objects in Visual Storytelling.

From Image Captioning to Visual Storytelling GROOViST: A Metric for Grounding Objects in Visual Storytelling

Reference 70

Resolution
verified exact
local_arxiv, observed 2026-08-06T10:27:39.856031Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.406036Z digest=sha256:26e233786a2cf35028a11b6f155ca102f6d4cbf70ca79421e200ed8157eb6adc

Observation dda0013f-0dc2-4353-87e1-1a443e757f60 · outbound

This paper cites an unresolved cited work.

From Image Captioning to Visual Storytelling Unresolved cited work

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.410173Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.410173Z digest=sha256:56dae534a265a826bb8746cb82f648c47b3cf005fa6e95b7ca1ecfaff088814d

Observation f8082d88-3674-49a4-ada3-2f159cb73630 · outbound

This paper cites Attention Is All You Need.

From Image Captioning to Visual Storytelling Attention Is All You Need

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.413633Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.413633Z digest=sha256:724f413faa3ce5ac6b888a0eaa887dc01055051d802c8ef5ffe09721f5f2a9e9

Observation fb173263-4328-4989-b022-42b86e67887d · outbound

This paper cites CIDEr: Consensus-based Image Description Evaluation.

From Image Captioning to Visual Storytelling CIDEr: Consensus-based Image Description Evaluation

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.417323Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.417323Z digest=sha256:954f8d741283bddc808d46fd0ebc57bbbc7728fbe2c73a13f644929a4721d1a4

Observation b32d3152-5ffd-4c64-bbae-1b51149e54a4 · outbound

This paper cites Show and Tell: A Neural Image Caption Generator.

From Image Captioning to Visual Storytelling Show and Tell: A Neural Image Caption Generator

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.421035Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.421035Z digest=sha256:ae62d4191c15357228f9d8bd197d58aa70ea695b2660c84bdb1bd4c36bd199ad

Observation 5c2aa0b5-1d93-4188-be11-52d9e6166858 · outbound

This paper cites RoViST:Learning Robust Metrics for Visual Storytelling.

From Image Captioning to Visual Storytelling RoViST:Learning Robust Metrics for Visual Storytelling

Reference 75

Resolution
verified exact
local_arxiv, observed 2026-08-06T10:27:39.812385Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.424827Z digest=sha256:5691782712f3310dc5acc654376a9490803705a396425741d2468a8e40003916

Observation 1ce7e186-bbf4-4ec8-a59a-7397098b1477 · outbound

This paper cites SCO-VIST: Social Interaction Commonsense Knowledge-based Visual Storytelling.

From Image Captioning to Visual Storytelling SCO-VIST: Social Interaction Commonsense Knowledge-based Visual Storytelling

Reference 76

Resolution
verified exact
local_arxiv, observed 2026-08-06T10:27:39.797114Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.428595Z digest=sha256:84b6ffa8c44e852fd727a185f26fba07d56aabf9c238690c59c0571c5bc13602

Observation 3aa95786-b222-4c36-93f7-c096dbe5113a · outbound

This paper cites an unresolved cited work.

From Image Captioning to Visual Storytelling Unresolved cited work

Reference 77

Resolution
unresolved
raw_fallback, observed 2026-08-06T10:27:40.636867Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.432373Z digest=sha256:3f69e34b7fa9dbd732ce2e4709f5ae4414c61e66a0e7f0442cf82b46b67262bd

Observation b091807c-d599-4f50-915b-7ddf03bc2b1f · outbound

This paper cites an unresolved cited work.

From Image Captioning to Visual Storytelling Unresolved cited work

Reference 78

Resolution
verified exact
doi, observed 2026-08-06T10:27:39.547730Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.435892Z digest=sha256:749eb9872e4ceab9824e6794be79b0fa83189eeb1fda029f3342cbb21ca7158f

Observation 02c3b8ec-8ce0-4881-b390-f8d6d5f5dab1 · outbound

This paper cites No Metrics Are Perfect: Adversarial Reward Learning for Visual Storytelling.

From Image Captioning to Visual Storytelling No Metrics Are Perfect: Adversarial Reward Learning for Visual Storytelling

Reference 79

Resolution
verified exact
local_arxiv, observed 2026-08-06T10:27:39.781778Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.439808Z digest=sha256:e6a94bccd826aef2a96e03ce1bab8da8fff5ae9f589cc8e78fe5c45d8b6202e3

Observation a5e90f3d-6977-4b6f-978e-86667d9f41a6 · outbound

This paper cites an unresolved cited work.

From Image Captioning to Visual Storytelling Unresolved cited work

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.443605Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.443605Z digest=sha256:e2a1b0e7a05f46b737f1628a825a6f4fb0b8d3aa81b17a0536d8b51e66cf247a

Observation 2e7bb518-1332-4a8b-9932-b8d85a7200b3 · outbound

This paper cites A Comparative Study of Open-Source Large Language Models, GPT-4 and Claude 2: Multiple-Choice Test Taking in Nephrology.

From Image Captioning to Visual Storytelling A Comparative Study of Open-Source Large Language Models, GPT-4 and Claude 2: Multiple-Choice Test Taking in Nephrology

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.447385Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.447385Z digest=sha256:df556d0e2d0334803c66f74d8d69364959adaa2e7aa669d42b68ee96027038f7

Observation b90fe462-ddb0-4e97-b1c7-6b6410a24976 · outbound

This paper cites an unresolved cited work.

From Image Captioning to Visual Storytelling Unresolved cited work

Reference 82

Resolution
unresolved
raw_fallback, observed 2026-08-06T10:27:40.625992Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.450741Z digest=sha256:188ced00d3535354a7b2daef032a2f7fe9222a79e6337ce09aebe2a3075878f4

Observation e22c170b-500e-49d1-a96d-3afe984aa08d · outbound

This paper cites Show, Attend and Tell: Neural Image Caption Generation with Visual Attention.

From Image Captioning to Visual Storytelling Show, Attend and Tell: Neural Image Caption Generation with Visual Attention

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.454437Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.454437Z digest=sha256:a0391fb13a8ea72e68c7b54120e739fbf632c2f737c907dabbfd6546a5a5380f

Observation e6ee732c-a3a7-49e1-a83b-bf8ca9c895cc · outbound

This paper cites an unresolved cited work.

From Image Captioning to Visual Storytelling Unresolved cited work

Reference 84

Resolution
verified exact
doi, observed 2026-08-06T10:27:39.536839Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.458015Z digest=sha256:d926b2baf7edc28a2121148a01ecf285f1806487977d0fcb5693e61c8824339e

Observation 60ad4f21-f899-4498-8fa3-30c6c4fbf7fa · outbound

This paper cites an unresolved cited work.

From Image Captioning to Visual Storytelling Unresolved cited work

Reference 85

Resolution
verified exact
doi, observed 2026-08-06T10:27:39.525086Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.461591Z digest=sha256:01d099f43c81fe17258e8abcac64f1c5295afd86f59dec4b37998b7ab81a8c11

Observation 77d70103-ca16-4fff-aed7-5b5f4e1da0aa · outbound

This paper cites an unresolved cited work.

From Image Captioning to Visual Storytelling Unresolved cited work

Reference 86

Resolution
unresolved
raw_fallback, observed 2026-08-06T10:27:40.615483Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.465764Z digest=sha256:3843358230067822898337c23c3383d6f3ed4aca5e707d28043dcb0b24abbf3f

Observation b5e64da4-7ca6-48b4-b0dc-63abf58a9530 · outbound

This paper cites an unresolved cited work.

From Image Captioning to Visual Storytelling Unresolved cited work

Reference 87

Resolution
unresolved
raw_fallback, observed 2026-08-06T10:27:40.605045Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.469781Z digest=sha256:dcd5f0472c4eb681a6f5ec11db704ecd76cf6f39a4929c75702afe02dc628066

Observation 97cfa074-f42b-4275-8a3a-3693350a57af · outbound

This paper cites Auto-Encoding Scene Graphs for Image Captioning.

From Image Captioning to Visual Storytelling Auto-Encoding Scene Graphs for Image Captioning

Reference 88

Resolution
verified exact
local_arxiv, observed 2026-08-06T10:27:39.669306Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.473297Z digest=sha256:b28ed43d735276c7afb3f14536f95bb4a4d17580d923695b71e8c7504ac32895

Observation ad058722-43bc-4b39-ab1e-06cda13bbe28 · outbound

This paper cites Plan-And-Write: Towards Better Automatic Storytelling.

From Image Captioning to Visual Storytelling Plan-And-Write: Towards Better Automatic Storytelling

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.477116Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.477116Z digest=sha256:04f12bf902b142333201e3f9e49ab3fbf6809d7ec9660a0b17afc1f8663e0021

Observation c6a4ef8e-df43-43c9-9b35-1b91c819723a · outbound

This paper cites Exploring Visual Relationship for Image Captioning.

From Image Captioning to Visual Storytelling Exploring Visual Relationship for Image Captioning

Reference 90

Resolution
verified exact
local_arxiv, observed 2026-08-06T10:27:39.645511Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.480734Z digest=sha256:c24a18c3f1018c93b16e17d0d429f656459593ebb4fa6d382145ef97c535e526

Observation 312199dd-8e2a-4cf3-9e5c-4e18d5bd2259 · outbound

This paper cites Hierarchically-Attentive RNN for Album Summarization and Storytelling.

From Image Captioning to Visual Storytelling Hierarchically-Attentive RNN for Album Summarization and Storytelling

Reference 91

Resolution
verified exact
local_arxiv, observed 2026-08-06T10:27:39.631169Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.484397Z digest=sha256:f67bc0543be4c54f0dc44447deaa6a37bcc5b724d1b5b1523bfeca8bb4c732f8

Observation d3d47cb8-d634-4809-bf01-5d2ea0aafcab · outbound

This paper cites an unresolved cited work.

From Image Captioning to Visual Storytelling Unresolved cited work

Reference 92

Resolution
unresolved
raw_fallback, observed 2026-08-06T10:27:40.594457Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.487868Z digest=sha256:b4fd58cbd578337010b563ae9edb2852c0d745da2cd17732f2bda89318e23fa1

Observation 29a3e9b4-a2d3-46c1-80dc-c3f24e36a74e · outbound

This paper cites Vision-Language Models for Vision Tasks: A Survey.

From Image Captioning to Visual Storytelling Vision-Language Models for Vision Tasks: A Survey

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.491216Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.491216Z digest=sha256:6762a42877d610c04ddba09b4d03ee8c921b88e46bf2589dbb6bdb8c478a4685

Pith citing papers

No inbound Pith citation observations are available.