Pith. sign in

Paper Citation Record · LEDGER

From Image Captioning to Visual Storytelling

As of 10 August 2026, this Paper Citation Record lists 93 of 93 outbound references and 0 inbound Pith citation observations for arXiv:2508.14045.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.14045 v1

Coverage vector

measured 93 of 93 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T10:27:39.491216Z

measured 93 of 93 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

93 of 93 outbound references displayed

  • verified exact34
  • verified fuzzy0
  • unresolved56
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch3

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 460933a4-410f-404f-a920-cb7a238ba56a · outbound

This paper cites URL: " 'urlintro :=.

From Image Captioning to Visual Storytelling URL: " 'urlintro :=

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.142990Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.142990Z digest=sha256:2fdd701d95685552f5a831a4acaab2a12607c07cd83973287fd8c565efc882a2

Observation 68db7014-2d91-40f8-9766-8128851954a9 · outbound

This paper cites write newline.

From Image Captioning to Visual Storytelling write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.147551Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.147551Z digest=sha256:174c67bd0efdd7cf67d9edadbf713dad4cdbb29beca21e925f99d1f312349f8e

Observation 67917612-ccc8-49d8-a5d3-50d109e71a17 · outbound

This paper cites SPICE: Semantic Propositional Image Caption Evaluation.

From Image Captioning to Visual Storytelling SPICE: Semantic Propositional Image Caption Evaluation

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.151762Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.151762Z digest=sha256:63e3f39d77ef968bcfc57b355f279845f3b72629c712b477eaf5c9241886cdf8

Observation 8922e562-aa55-4287-95b4-a4f6c09712ef · outbound

This paper cites Bottom-Up and Top-Down Attention for Image Captioning and Visual Question Answering.

From Image Captioning to Visual Storytelling Bottom-Up and Top-Down Attention for Image Captioning and Visual Question Answering

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.156320Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.156320Z digest=sha256:6ecdd451fbee8cfbcc7f8a954077dbf2052dd4a497b1fee037ece051edca0012

Observation 8208d6aa-a0b3-4ca8-8279-123b8ef4fb3c · outbound

This paper cites TouchStone: Evaluating Vision-Language Models by Language Models.

From Image Captioning to Visual Storytelling TouchStone: Evaluating Vision-Language Models by Language Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.160499Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.160499Z digest=sha256:63b464d3c527020ee379fe53f2a7e32473824f9597c36c0b5266d1b81e09d856

Observation 19d49c2c-5034-483f-885d-61a32c798eae · outbound

This paper cites an unresolved cited work.

From Image Captioning to Visual Storytelling Unresolved cited work

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.164465Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.164465Z digest=sha256:7411169cc074597e6cf5bdc5c8a544ef6ecc9687e2c3b0d4f7061aabb39e4ecd

Observation 1136e2c7-61fc-4591-b176-f46da3e3f86a · outbound

This paper cites Commonsense Knowledge Aware Concept Selection For Diverse and Informative Visual Storytelling.

From Image Captioning to Visual Storytelling Commonsense Knowledge Aware Concept Selection For Diverse and Informative Visual Storytelling

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-08-06T10:27:40.555602Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.168379Z digest=sha256:e66994dfbe36bda115dd9d9fe5408c5d8b60ee259309b3e4864cb08a7239cbc5

Observation fc08cee6-7010-4466-8aef-9b7c28a56ec1 · outbound

This paper cites an unresolved cited work.

From Image Captioning to Visual Storytelling Unresolved cited work

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.172413Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.172413Z digest=sha256:976cc6160308144d585d1c64396f98dbcbd3a3328c3e96ff6d8166c88398798a

Observation ee8efa95-40e1-4047-9f60-72b97afb0630 · outbound

This paper cites TARN-VIST: Topic Aware Reinforcement Network for Visual Storytelling.

From Image Captioning to Visual Storytelling TARN-VIST: Topic Aware Reinforcement Network for Visual Storytelling

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-08-06T10:27:40.457057Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.176728Z digest=sha256:ffb5adb153cb2b05a334387947b4fdffb70962ea8471f66fbb6e96f5426dc639

Observation 1e2fad49-b672-4775-bce9-b8e9a9f4ad87 · outbound

This paper cites an unresolved cited work.

From Image Captioning to Visual Storytelling Unresolved cited work

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.180421Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.180421Z digest=sha256:c2b42ad3b68d0e8e343f5b883d260e062a6575a63c5b90fcb8e029af9c02d338

Observation 09d7a993-ac83-4163-a3f9-7861dc75b291 · outbound

This paper cites an unresolved cited work.

From Image Captioning to Visual Storytelling Unresolved cited work

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.183894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.183894Z digest=sha256:6753a1e57de3d845a333ea1000c814b775d7e8b83d6410ab8d541c1c912c4b00

Observation 6662c5ec-dd13-476d-8be4-0d40515305e6 · outbound

This paper cites Exploring Nearest Neighbor Approaches for Image Captioning.

From Image Captioning to Visual Storytelling Exploring Nearest Neighbor Approaches for Image Captioning

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-08-06T10:27:40.442624Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.187962Z digest=sha256:79aeb523b8d60b7cabc893dee30fb68f3532e48ede6bea26103405f724096955

Observation 01b27af5-ae41-49e7-a259-165e2d91c103 · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

From Image Captioning to Visual Storytelling An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.191648Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.191648Z digest=sha256:6433c11a681b0785fd97d577f930fb41cccedc6b1095abe0584b798b65166177

Observation fbed2d2d-eadb-49c3-b8e3-090c433f12ad · outbound

This paper cites an unresolved cited work.

From Image Captioning to Visual Storytelling Unresolved cited work

Reference 14

Resolution
metadata mismatch
raw_fallback, observed 2026-08-06T10:27:40.417408Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.195374Z digest=sha256:13e0f7fd86757eb63361ac33465c2ab5fac416e79c23d6daedd7fdb0d81c828e

Observation fb83ea58-8d96-44da-96d7-3690c62f155d · outbound

This paper cites Transformer-based Conditional Variational Autoencoder for Controllable Story Generation.

From Image Captioning to Visual Storytelling Transformer-based Conditional Variational Autoencoder for Controllable Story Generation

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-08-06T10:27:40.343040Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.199030Z digest=sha256:9908719fdfc7c329167f8a2439770981f4aa24dc18034cbd44321ac12b306b52

Observation 78cb959e-289a-4eec-927b-980ea2f91043 · outbound

This paper cites an unresolved cited work.

From Image Captioning to Visual Storytelling Unresolved cited work

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.202850Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.202850Z digest=sha256:ad3b35faa9006d905ad5a17b78afbf129ecc7607d9702c674635106847bcfa46

Observation cfb859f3-90cd-4b9c-90a8-93cf080365af · outbound

This paper cites an unresolved cited work.

From Image Captioning to Visual Storytelling Unresolved cited work

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.206292Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.206292Z digest=sha256:4e17d8de536e407c83ab5c3d5081a907d2b264c95025ee90a23622646ce1f652

Observation 65897f70-5776-44ba-b7e1-953f6a2545d6 · outbound

This paper cites Contextualize, Show and Tell: A Neural Visual Storyteller.

From Image Captioning to Visual Storytelling Contextualize, Show and Tell: A Neural Visual Storyteller

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-08-06T10:27:40.328324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.209675Z digest=sha256:cee30c7e00ff025bfb6c41e51241f6cb4af50aa7bf82010d9f9067d4fb93b077

Observation 5319b260-92f2-4cf1-85d6-529914bd1dbd · outbound

This paper cites A Knowledge-Enhanced Pretraining Model for Commonsense Story Generation.

From Image Captioning to Visual Storytelling A Knowledge-Enhanced Pretraining Model for Commonsense Story Generation

Reference 19

Resolution
verified exact
local_arxiv, observed 2026-08-06T10:27:40.314505Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.213265Z digest=sha256:706b63e9fbc52e9e0c9ca6cbfbd246742ae6aa4adcc5109e3c523dde8adef999

Observation 6919e57d-be18-41dc-8699-d038eb3be25d · outbound

This paper cites VICTR: Visual Information Captured Text Representation for Text-to-Image Multimodal Tasks.

From Image Captioning to Visual Storytelling VICTR: Visual Information Captured Text Representation for Text-to-Image Multimodal Tasks

Reference 20

Resolution
verified exact
local_arxiv, observed 2026-08-06T10:27:40.300254Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.216713Z digest=sha256:ef9b33f32de75a6b02fc17d67c756f9cf3ec242fdac1ad6bd50f8aabe2d2ebd3

Observation 77020388-c12f-4cc3-8929-a18ee2cd8ff7 · outbound

This paper cites Deep Residual Learning for Image Recognition.

From Image Captioning to Visual Storytelling Deep Residual Learning for Image Recognition

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.220359Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.220359Z digest=sha256:f8a5c3224984be781b145999aaa6912160ccce281fe275fcdd0010d00a58ae05

Observation 0578e642-1779-4e5f-9d47-9397ae743d20 · outbound

This paper cites Image Captioning through Image Transformer.

From Image Captioning to Visual Storytelling Image Captioning through Image Transformer

Reference 22

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T10:27:40.276721Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.223799Z digest=sha256:88129b77dbc147e67c29a0811de924c2e0c2d9ae682f6bf1b0f6f0b54b7770f6

Observation 2c6c43eb-392d-4435-bb10-1af549fdd19f · outbound

This paper cites Image Captioning: Transforming Objects into Words.

From Image Captioning to Visual Storytelling Image Captioning: Transforming Objects into Words

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.227854Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.227854Z digest=sha256:952ee7e663fe2a802a9605eb5cbbcd92e82288ef4bd465034066bda2661d33ef

Observation 7efec995-1c58-4e65-a53d-5a0a5e6d3878 · outbound

This paper cites Visual Writing Prompts: Character-Grounded Story Generation with Curated Image Sequences.

From Image Captioning to Visual Storytelling Visual Writing Prompts: Character-Grounded Story Generation with Curated Image Sequences

Reference 24

Resolution
verified exact
local_arxiv, observed 2026-08-06T10:27:40.252870Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.231546Z digest=sha256:2539cb132dd05cd0ba85c89b34866f3215f15bf7642f6ca1d8c7f29fe2fc0e81

Observation 427b46ef-5ae3-46dc-a62b-30d54a780374 · outbound

This paper cites an unresolved cited work.

From Image Captioning to Visual Storytelling Unresolved cited work

Reference 25

Resolution
verified exact
doi, observed 2026-08-06T10:27:39.600141Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.235186Z digest=sha256:70407bcb1eba2afb5440dedd58402719ce1390ed0686954e6a51d34dd5b320f4

Observation 8e834285-e94b-4497-8653-3365b1d32738 · outbound

This paper cites Knowledge-Enriched Visual Storytelling.

From Image Captioning to Visual Storytelling Knowledge-Enriched Visual Storytelling

Reference 26

Resolution
verified exact
local_arxiv, observed 2026-08-06T10:27:40.237734Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.238764Z digest=sha256:683be7fa4464c6044563343df822ed89be52ea776e3ee431ff29e8169447bfd0

Observation 3e16745c-30a3-48a8-a26b-a174808360ba · outbound

This paper cites Plot and Rework: Modeling Storylines for Visual Storytelling.

From Image Captioning to Visual Storytelling Plot and Rework: Modeling Storylines for Visual Storytelling

Reference 27

Resolution
verified exact
local_arxiv, observed 2026-08-06T10:27:40.223637Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.242438Z digest=sha256:1612958628f382f002f40b8eadadbab7b1e636b1559fd01e3195211a5f85efd3

Observation 4c1f7389-1b57-4464-b4c2-bdee5178d99b · outbound

This paper cites Visual Story Post-Editing.

From Image Captioning to Visual Storytelling Visual Story Post-Editing

Reference 28

Resolution
verified exact
local_arxiv, observed 2026-08-06T10:27:40.209796Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.246217Z digest=sha256:81bcfd401d32072b90e138279d9599ecab394352ff63199543359aa9b2cf2852

Observation 105aa1fc-90dc-4512-b17e-786916095155 · outbound

This paper cites What Makes A Good Story? Designing Composite Rewards for Visual Storytelling.

From Image Captioning to Visual Storytelling What Makes A Good Story? Designing Composite Rewards for Visual Storytelling

Reference 29

Resolution
verified exact
local_arxiv, observed 2026-08-06T10:27:40.194184Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.250137Z digest=sha256:dfad7161a042f6b62f34ff3e91c02e9064853f2999ecb739ba917a822e481e9b

Observation c67cc21f-a03e-440f-ae93-4237f8229808 · outbound

This paper cites Attention on Attention for Image Captioning.

From Image Captioning to Visual Storytelling Attention on Attention for Image Captioning

Reference 30

Resolution
verified exact
local_arxiv, observed 2026-08-06T10:27:40.179323Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.254800Z digest=sha256:a8414e97ac217d12a951730b2a8b21ebf99b4667c6d45c5981c5834722822b89

Observation 2ad04455-f1e3-4249-86cb-1b4716457b81 · outbound

This paper cites Lawrence Zitnick, Devi Parikh, Lucy Vanderwende, Galley Michel, and Mitchell Margaret.

From Image Captioning to Visual Storytelling Lawrence Zitnick, Devi Parikh, Lucy Vanderwende, Galley Michel, and Mitchell Margaret

Reference 31

Resolution
verified exact
doi, observed 2026-08-06T10:27:39.589103Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.258864Z digest=sha256:a286e492a473b6c6bd9208507be13fd7b1f34bcece88f1603fac6c5da853d091

Observation a2cdddcc-5ee7-4e64-9a08-7c793b365491 · outbound

This paper cites Story Generation from Sequence of Independent Short Descriptions.

From Image Captioning to Visual Storytelling Story Generation from Sequence of Independent Short Descriptions

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.262362Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.262362Z digest=sha256:4bbe2f4e0681392dafc1a13d464fc848382929293d1f760521ac14829f30a8a8

Observation cbf1c71d-8fe6-485f-8532-77576a34ca9f · outbound

This paper cites an unresolved cited work.

From Image Captioning to Visual Storytelling Unresolved cited work

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.266799Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.266799Z digest=sha256:3dda554bd54942dbe1cd7ea4333e0aeebf4af46a70b3c75b3dfb6886cee1752e

Observation 51554768-4817-4571-ac6e-fc292c53417a · outbound

This paper cites an unresolved cited work.

From Image Captioning to Visual Storytelling Unresolved cited work

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.270245Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.270245Z digest=sha256:545f07413894033b9235f11a342547c7127fe576800f6071d08148aaf9e398cf

Observation 419e3971-28c0-48b6-a6b4-af2f9f896912 · outbound

This paper cites GLAC Net: GLocal Attention Cascading Networks for Multi-image Cued Story Generation.

From Image Captioning to Visual Storytelling GLAC Net: GLocal Attention Cascading Networks for Multi-image Cued Story Generation

Reference 35

Resolution
verified exact
local_arxiv, observed 2026-08-06T10:27:40.154163Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.274281Z digest=sha256:deae1b6430d335e6a6b68e913f205fd59afbfc14bcab26eed101b90be4db485f

Observation c8910c74-efb6-4850-a6c5-12a63ed028af · outbound

This paper cites Adam: A Method for Stochastic Optimization.

From Image Captioning to Visual Storytelling Adam: A Method for Stochastic Optimization

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.278137Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.278137Z digest=sha256:33d32e596aabe2900fa30a5bb2a413b16ef70554f59aa6afb78f9a6f3dadc117

Observation 989190ee-ceb5-4ac2-ad7f-c3b8d92420f6 · outbound

This paper cites Visual Genome: Connecting Language and Vision Using Crowdsourced Dense Image Annotations.

From Image Captioning to Visual Storytelling Visual Genome: Connecting Language and Vision Using Crowdsourced Dense Image Annotations

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.281717Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.281717Z digest=sha256:f3ad78d59944ef5740950be9c126b172764aec51563686ed9fc109cdae5d005e

Observation 2e146646-9bd6-4472-a3fd-c05ec8112898 · outbound

This paper cites an unresolved cited work.

From Image Captioning to Visual Storytelling Unresolved cited work

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.285646Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.285646Z digest=sha256:ca899ce062270f6e042ac414e07e4f09a5b61576aedaf37e72887fa7ee6ca806

Observation 903672af-7ab0-4424-88f5-2d33bebc4e7e · outbound

This paper cites Can large language models provide useful feedback on research papers? A large-scale empirical analysis.

From Image Captioning to Visual Storytelling Can large language models provide useful feedback on research papers? A large-scale empirical analysis

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.290156Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.290156Z digest=sha256:2a92b102a09756795fdb3d8c83d6a11905a64a0f7447ac9a1658f65950523ab3

Observation 3da6f782-8b0d-4469-a45b-045a2cc88fde · outbound

This paper cites an unresolved cited work.

From Image Captioning to Visual Storytelling Unresolved cited work

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.295020Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.295020Z digest=sha256:73080d84b97a45ba6684480dd3fbbca4cb612de3beb2af3fa7f1e36258ca5695

Observation 73029131-2d63-47bd-8138-b5e332d5cf80 · outbound

This paper cites Microsoft COCO: Common Objects in Context.

From Image Captioning to Visual Storytelling Microsoft COCO: Common Objects in Context

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.299512Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.299512Z digest=sha256:7edb130a1bc696617ba63ecd511db40c59f8dc387e2139d6ebe3ba4fe81f56a4

Observation 058e1eda-5b59-432e-b044-8063073aa986 · outbound

This paper cites Detecting and Grounding Important Characters in Visual Stories.

From Image Captioning to Visual Storytelling Detecting and Grounding Important Characters in Visual Stories

Reference 42

Resolution
verified exact
local_arxiv, observed 2026-08-06T10:27:40.101851Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.303627Z digest=sha256:902e36b770df035f75fd88c5376a612d458dd928a62d41317a65423aecf3d15a

Observation 6e3ab560-f4dc-4e6e-8d2c-973f8c351b55 · outbound

This paper cites Generating Visual Stories with Grounded and Coreferent Characters.

From Image Captioning to Visual Storytelling Generating Visual Stories with Grounded and Coreferent Characters

Reference 43

Resolution
verified exact
local_arxiv, observed 2026-08-06T10:27:40.087549Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.307440Z digest=sha256:681118c57b153f363bba256a56481decde74e2566eef39d7b29320337f760727

Observation 57653fa1-3860-4cc6-a222-ce63084026dd · outbound

This paper cites an unresolved cited work.

From Image Captioning to Visual Storytelling Unresolved cited work

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.311506Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.311506Z digest=sha256:b0a668e9c9ce513554019bf5f85f54274cce58181ef925c1f379e3e0fdc2c567

Observation da206540-3816-4e31-b2b8-531f42f8f6d2 · outbound

This paper cites Visual Instruction Tuning.

From Image Captioning to Visual Storytelling Visual Instruction Tuning

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.315063Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.315063Z digest=sha256:1155bfc71cd56a259ef9fad76ee881bbd7096d0c467ecb42d5a6db07366837ab

Observation fdee897e-d3b5-402d-8946-069e6a5a8903 · outbound

This paper cites an unresolved cited work.

From Image Captioning to Visual Storytelling Unresolved cited work

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.318888Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.318888Z digest=sha256:8c3d5fdbf20d017296afd1c59078d29b39a7e3e55a7581a3b82c6ddec84deed9

Observation fffd135f-d63d-4b72-8ed5-9e1082cc1057 · outbound

This paper cites an unresolved cited work.

From Image Captioning to Visual Storytelling Unresolved cited work

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.322346Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.322346Z digest=sha256:174414983bcce029a3771712ffcd51934ed627bb12c0a1ab40c06f7667ba58e3

Observation 43d4022d-768c-44cc-8564-49fa5ed5d58c · outbound

This paper cites Decoupled Weight Decay Regularization.

From Image Captioning to Visual Storytelling Decoupled Weight Decay Regularization

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.325723Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.325723Z digest=sha256:dc499b57f945f5dbb0e8d215569114600a30fee9a134d8c5d6e0c237d28be06b

Observation e585f74d-cd12-4a2c-a89d-fea58f28007f · outbound

This paper cites ViLBERT: Pretraining Task-Agnostic Visiolinguistic Representations for Vision-and-Language Tasks.

From Image Captioning to Visual Storytelling ViLBERT: Pretraining Task-Agnostic Visiolinguistic Representations for Vision-and-Language Tasks

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.329615Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.329615Z digest=sha256:215f7a52bc993f350d410ac36e6539873125dfcf947f666307f7604402e47c90

Observation 6c4c66f5-985b-494a-9897-bab94d50dfca · outbound

This paper cites Knowing When to Look: Adaptive Attention via A Visual Sentinel for Image Captioning.

From Image Captioning to Visual Storytelling Knowing When to Look: Adaptive Attention via A Visual Sentinel for Image Captioning

Reference 50

Resolution
verified exact
local_arxiv, observed 2026-08-06T10:27:40.043109Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.334282Z digest=sha256:05953c7f479891f15540e51273e3b497905f1936ac8b11f9646187a8bd1279a9

Observation 37b79eda-e591-486e-aa15-4fabec394308 · outbound

This paper cites ClipCap: CLIP Prefix for Image Captioning.

From Image Captioning to Visual Storytelling ClipCap: CLIP Prefix for Image Captioning

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.338214Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.338214Z digest=sha256:ebcf05311ba5fc8b5508ab9ab0f1f4e7932e0223e4857ec5b19098d1f5c103b2

Observation a0104fdd-bca1-41ec-ab2c-d9117d4ee5fe · outbound

This paper cites Album Storytelling with Iterative Story-aware Captioning and Large Language Models.

From Image Captioning to Visual Storytelling Album Storytelling with Iterative Story-aware Captioning and Large Language Models

Reference 52

Resolution
verified exact
local_arxiv, observed 2026-08-06T10:27:40.017233Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.341873Z digest=sha256:47627e179bed4b3db3417c489e1c4b011c618eb81f19108ea84608c0c0228b5c

Observation ec36a8b8-11dd-4cfd-b7ad-2a5f9824ceee · outbound

This paper cites GPT-4 Technical Report.

From Image Captioning to Visual Storytelling GPT-4 Technical Report

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.345535Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.345535Z digest=sha256:88a613ee859501f79c8b065d7100be97f66a39e0e64062ddf1ec58ac7ea4610d

Observation 0f4f6d66-e639-4762-b23a-42658458b030 · outbound

This paper cites an unresolved cited work.

From Image Captioning to Visual Storytelling Unresolved cited work

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.349238Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.349238Z digest=sha256:8c3c5d69a08540e945d74ef6290af62cae95b17b2d9322c7981cc2236670fbcc

Observation 687713ed-6311-4c42-892c-7f49a7474639 · outbound

This paper cites an unresolved cited work.

From Image Captioning to Visual Storytelling Unresolved cited work

Reference 55

Resolution
unresolved
raw_fallback, observed 2026-08-06T10:27:40.699746Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.353116Z digest=sha256:a8d213a72429ef53344a654d48295e68ca47a44b4431d44e21909fbe360d9f2d

Observation 554ae120-bd17-4590-901f-7cb3c32814d2 · outbound

This paper cites "My Way of Telling a Story": Persona based Grounded Story Generation.

From Image Captioning to Visual Storytelling "My Way of Telling a Story": Persona based Grounded Story Generation

Reference 56

Resolution
verified exact
local_arxiv, observed 2026-08-06T10:27:39.993159Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.356460Z digest=sha256:71b4dd223cc4c2a6fdd7a1d52461b8c9198dd3e1f7e70d1c7e0ce8f3d20b3926

Observation 18cf3879-f740-46b8-a74e-c2dd89070af7 · outbound

This paper cites an unresolved cited work.

From Image Captioning to Visual Storytelling Unresolved cited work

Reference 57

Resolution
metadata mismatch
raw_fallback, observed 2026-08-06T10:27:39.978523Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.360161Z digest=sha256:a796d42aed4f26f10ddfa8c52b28536dfa926d311a558e65a41026d1bbf273c7

Observation 043bfdfd-ae57-491c-b34e-a2cf088ec02b · outbound

This paper cites an unresolved cited work.

From Image Captioning to Visual Storytelling Unresolved cited work

Reference 58

Resolution
unresolved
raw_fallback, observed 2026-08-06T10:27:40.689280Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.363586Z digest=sha256:627d16a442e184e69a5198f30948a1a5287abf44979af84635bdbea9ed640711

Observation e60aa79e-f929-44e9-aadd-fe864c611ff1 · outbound

This paper cites an unresolved cited work.

From Image Captioning to Visual Storytelling Unresolved cited work

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.367121Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.367121Z digest=sha256:6a26dc5cfc59fbad64fcfeddd2d9573db9aae5e52c90a36540115b7ea29068ab

Observation 952d01be-fd97-4645-b3dd-1138179d7cd9 · outbound

This paper cites an unresolved cited work.

From Image Captioning to Visual Storytelling Unresolved cited work

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.370515Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.370515Z digest=sha256:456161bd310dfe2d8c27c9d90db97a8d0b349b57cf66b13eb2358fc7b055edb6

Observation 13985ff9-03a2-4143-9bed-a6cec49e6f69 · outbound

This paper cites an unresolved cited work.

From Image Captioning to Visual Storytelling Unresolved cited work

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.373867Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.373867Z digest=sha256:b7cda55c2c6d6978355ed62145c5427697b3de0dc360ac3b83a7b88925640929

Observation edead90b-c421-48a6-b33d-d747ba762ea9 · outbound

This paper cites an unresolved cited work.

From Image Captioning to Visual Storytelling Unresolved cited work

Reference 62

Resolution
unresolved
raw_fallback, observed 2026-08-06T10:27:40.660493Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.377562Z digest=sha256:5ef5bfd58de6752d0819bcb11f6f69c27cbeb3d05b798d0451b3348353185f6d

Observation 94a30da2-4380-4351-9805-5a4ed67bfcc8 · outbound

This paper cites Self-critical Sequence Training for Image Captioning.

From Image Captioning to Visual Storytelling Self-critical Sequence Training for Image Captioning

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.381077Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.381077Z digest=sha256:c464c33070b6635cbb744c65226dd41d4b7a973bd9c8a16bc737dcfd76ed7344

Observation 557f3364-c092-41db-96b4-739057dc3d61 · outbound

This paper cites an unresolved cited work.

From Image Captioning to Visual Storytelling Unresolved cited work

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.384650Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.384650Z digest=sha256:b6949533c5bb133430303242df501335068c7acc87f2f1ed23be512261487c87

Observation 217a277e-110f-4fa1-a021-7f4e7464098c · outbound

This paper cites an unresolved cited work.

From Image Captioning to Visual Storytelling Unresolved cited work

Reference 65

Resolution
unresolved
raw_fallback, observed 2026-08-06T10:27:40.649191Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.387909Z digest=sha256:369740e0eeb0049c5d7167c42a17f70dada822318fff28f03197312fd98289e5

Observation 231da63b-840f-4809-bb7e-7ac72fb0b435 · outbound

This paper cites an unresolved cited work.

From Image Captioning to Visual Storytelling Unresolved cited work

Reference 66

Resolution
verified exact
doi, observed 2026-08-06T10:27:39.565514Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.391257Z digest=sha256:cf749312e691eff66deb86dd6147f3b87e210618a82c644e116be393c4e9b944

Observation 33fc1b98-a558-4f14-aa29-ef4449227539 · outbound

This paper cites Context-aware Visual Storytelling with Visual Prefix Tuning and Contrastive Learning.

From Image Captioning to Visual Storytelling Context-aware Visual Storytelling with Visual Prefix Tuning and Contrastive Learning

Reference 67

Resolution
verified exact
local_arxiv, observed 2026-08-06T10:27:39.895991Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.394862Z digest=sha256:4d1b4d6d31cd9b6c4cad8d7ac701f842859b7d3fb623d6e19d1186c5e10184de

Observation d456efa1-f29f-4fc5-a19c-783add29eff0 · outbound

This paper cites BERT-hLSTMs: BERT and Hierarchical LSTMs for Visual Storytelling.

From Image Captioning to Visual Storytelling BERT-hLSTMs: BERT and Hierarchical LSTMs for Visual Storytelling

Reference 68

Resolution
verified exact
local_arxiv, observed 2026-08-06T10:27:39.880542Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.398532Z digest=sha256:4bfcc9d7c7f8570aeaa4794c076851af725ebbcde1ddf38c9815963e1c6665b0

Observation 896b20bb-827b-4a89-a7f0-d91bfc9c3f1f · outbound

This paper cites A Contrastive Framework for Neural Text Generation.

From Image Captioning to Visual Storytelling A Contrastive Framework for Neural Text Generation

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.402209Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.402209Z digest=sha256:0ed7a7c3f48bf09b8b9e6dc91fbe6f61a503dc437a6eafef60bd93d1e7b1d72f

Observation 3cf0b8ea-ae3c-4ba7-b452-e0a68339f0c8 · outbound

This paper cites GROOViST: A Metric for Grounding Objects in Visual Storytelling.

From Image Captioning to Visual Storytelling GROOViST: A Metric for Grounding Objects in Visual Storytelling

Reference 70

Resolution
verified exact
local_arxiv, observed 2026-08-06T10:27:39.856031Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.406036Z digest=sha256:d956201b0d7070872b18dbfbb59d254ffd22a6834f0d1ce4d2a23e97dd5d5f6d

Observation dda0013f-0dc2-4353-87e1-1a443e757f60 · outbound

This paper cites an unresolved cited work.

From Image Captioning to Visual Storytelling Unresolved cited work

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.410173Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.410173Z digest=sha256:14c19e3a14a0a09a7e250826f06c66c93adfaea13e3c58cec0dba1f432079ad2

Observation f8082d88-3674-49a4-ada3-2f159cb73630 · outbound

This paper cites Attention Is All You Need.

From Image Captioning to Visual Storytelling Attention Is All You Need

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.413633Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.413633Z digest=sha256:28b42758108a462ec1927e9904267339fe193cda03bd450d308104bba313a557

Observation fb173263-4328-4989-b022-42b86e67887d · outbound

This paper cites CIDEr: Consensus-based Image Description Evaluation.

From Image Captioning to Visual Storytelling CIDEr: Consensus-based Image Description Evaluation

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.417323Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.417323Z digest=sha256:2627555f209068c095f4031b5ef0b153f89eaca25f34c5b4afee97e7329616a0

Observation b32d3152-5ffd-4c64-bbae-1b51149e54a4 · outbound

This paper cites Show and Tell: A Neural Image Caption Generator.

From Image Captioning to Visual Storytelling Show and Tell: A Neural Image Caption Generator

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.421035Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.421035Z digest=sha256:0114e28024587cbd246f73b31550a75019f1f67909de56bde394e6cd70e49592

Observation 5c2aa0b5-1d93-4188-be11-52d9e6166858 · outbound

This paper cites RoViST:Learning Robust Metrics for Visual Storytelling.

From Image Captioning to Visual Storytelling RoViST:Learning Robust Metrics for Visual Storytelling

Reference 75

Resolution
verified exact
local_arxiv, observed 2026-08-06T10:27:39.812385Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.424827Z digest=sha256:cdd128d2e33201b20ab0163368a69a283a859919af40f6e371e369f26593fbfe

Observation 1ce7e186-bbf4-4ec8-a59a-7397098b1477 · outbound

This paper cites SCO-VIST: Social Interaction Commonsense Knowledge-based Visual Storytelling.

From Image Captioning to Visual Storytelling SCO-VIST: Social Interaction Commonsense Knowledge-based Visual Storytelling

Reference 76

Resolution
verified exact
local_arxiv, observed 2026-08-06T10:27:39.797114Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.428595Z digest=sha256:3fe0170c2b3c3ab2f1c86606339b7154c29e4cd07ff325dd14f277953c0a7d16

Observation 3aa95786-b222-4c36-93f7-c096dbe5113a · outbound

This paper cites an unresolved cited work.

From Image Captioning to Visual Storytelling Unresolved cited work

Reference 77

Resolution
unresolved
raw_fallback, observed 2026-08-06T10:27:40.636867Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.432373Z digest=sha256:c2c0d24d131ea193f77b3732ba35dc8e9c86864d787e5ed266fcb92ebbf99fbf

Observation b091807c-d599-4f50-915b-7ddf03bc2b1f · outbound

This paper cites an unresolved cited work.

From Image Captioning to Visual Storytelling Unresolved cited work

Reference 78

Resolution
verified exact
doi, observed 2026-08-06T10:27:39.547730Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.435892Z digest=sha256:b56c92acf3f7fb26f38dff999effdf1ac524fbad088a404306a712b7bee7cc16

Observation 02c3b8ec-8ce0-4881-b390-f8d6d5f5dab1 · outbound

This paper cites No Metrics Are Perfect: Adversarial Reward Learning for Visual Storytelling.

From Image Captioning to Visual Storytelling No Metrics Are Perfect: Adversarial Reward Learning for Visual Storytelling

Reference 79

Resolution
verified exact
local_arxiv, observed 2026-08-06T10:27:39.781778Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.439808Z digest=sha256:b553dda4169f63aba7e2b2c0669c7fc07964fc9d3fe2dcf08e749ab317f21de8

Observation a5e90f3d-6977-4b6f-978e-86667d9f41a6 · outbound

This paper cites an unresolved cited work.

From Image Captioning to Visual Storytelling Unresolved cited work

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.443605Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.443605Z digest=sha256:260924837c1470c5880906897b44a4da4f17e2e96c19a0889b25bc9267a3655a

Observation 2e7bb518-1332-4a8b-9932-b8d85a7200b3 · outbound

This paper cites A Comparative Study of Open-Source Large Language Models, GPT-4 and Claude 2: Multiple-Choice Test Taking in Nephrology.

From Image Captioning to Visual Storytelling A Comparative Study of Open-Source Large Language Models, GPT-4 and Claude 2: Multiple-Choice Test Taking in Nephrology

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.447385Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.447385Z digest=sha256:66627faf4eb117a6c4857be39ebf7ec51228b3e3b6c7ab619893dcdd5c72f271

Observation b90fe462-ddb0-4e97-b1c7-6b6410a24976 · outbound

This paper cites an unresolved cited work.

From Image Captioning to Visual Storytelling Unresolved cited work

Reference 82

Resolution
unresolved
raw_fallback, observed 2026-08-06T10:27:40.625992Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.450741Z digest=sha256:c9f0de6765fe19c9f4f21cb792ca9c0de45eaa37ac19c8f75df09d46f0ce7bc2

Observation e22c170b-500e-49d1-a96d-3afe984aa08d · outbound

This paper cites Show, Attend and Tell: Neural Image Caption Generation with Visual Attention.

From Image Captioning to Visual Storytelling Show, Attend and Tell: Neural Image Caption Generation with Visual Attention

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.454437Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.454437Z digest=sha256:ecac8c14e7d084f33f5d8a3bf66b1f32f098df61e6545a838dc5db14a6d36437

Observation e6ee732c-a3a7-49e1-a83b-bf8ca9c895cc · outbound

This paper cites an unresolved cited work.

From Image Captioning to Visual Storytelling Unresolved cited work

Reference 84

Resolution
verified exact
doi, observed 2026-08-06T10:27:39.536839Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.458015Z digest=sha256:10b4817b33f153c341b72debf8a523501c2d3b084b150027683cb2c7cce33a6e

Observation 60ad4f21-f899-4498-8fa3-30c6c4fbf7fa · outbound

This paper cites an unresolved cited work.

From Image Captioning to Visual Storytelling Unresolved cited work

Reference 85

Resolution
verified exact
doi, observed 2026-08-06T10:27:39.525086Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.461591Z digest=sha256:9ffd8771bf06556ddb5dc0543c55b8d570285cbde02f1ee7447ed31c9f483ed7

Observation 77d70103-ca16-4fff-aed7-5b5f4e1da0aa · outbound

This paper cites an unresolved cited work.

From Image Captioning to Visual Storytelling Unresolved cited work

Reference 86

Resolution
unresolved
raw_fallback, observed 2026-08-06T10:27:40.615483Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.465764Z digest=sha256:f61122bd645921045ec04ce0ad01e01bbedb1025cac17861b9c68b31c43294fe

Observation b5e64da4-7ca6-48b4-b0dc-63abf58a9530 · outbound

This paper cites an unresolved cited work.

From Image Captioning to Visual Storytelling Unresolved cited work

Reference 87

Resolution
unresolved
raw_fallback, observed 2026-08-06T10:27:40.605045Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.469781Z digest=sha256:074440be07ee71adf0b9190acb77fd18e571301baf293bc7c147797108647770

Observation 97cfa074-f42b-4275-8a3a-3693350a57af · outbound

This paper cites Auto-Encoding Scene Graphs for Image Captioning.

From Image Captioning to Visual Storytelling Auto-Encoding Scene Graphs for Image Captioning

Reference 88

Resolution
verified exact
local_arxiv, observed 2026-08-06T10:27:39.669306Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.473297Z digest=sha256:bf638fd15dc11e4f9871cac4d21e0e16042a064231073904e9cefa47d995ce92

Observation ad058722-43bc-4b39-ab1e-06cda13bbe28 · outbound

This paper cites Plan-And-Write: Towards Better Automatic Storytelling.

From Image Captioning to Visual Storytelling Plan-And-Write: Towards Better Automatic Storytelling

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.477116Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.477116Z digest=sha256:3dda91d7e267b2326ae8a85c1396beee0af490c40427b0481803c0d21b03bb07

Observation c6a4ef8e-df43-43c9-9b35-1b91c819723a · outbound

This paper cites Exploring Visual Relationship for Image Captioning.

From Image Captioning to Visual Storytelling Exploring Visual Relationship for Image Captioning

Reference 90

Resolution
verified exact
local_arxiv, observed 2026-08-06T10:27:39.645511Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.480734Z digest=sha256:8c078820a39abc179d0a178205e8184e837ca56762a9d351e03a74c8ae723aed

Observation 312199dd-8e2a-4cf3-9e5c-4e18d5bd2259 · outbound

This paper cites Hierarchically-Attentive RNN for Album Summarization and Storytelling.

From Image Captioning to Visual Storytelling Hierarchically-Attentive RNN for Album Summarization and Storytelling

Reference 91

Resolution
verified exact
local_arxiv, observed 2026-08-06T10:27:39.631169Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.484397Z digest=sha256:33077a6434c50983717ff9ec021bb4595bdc826adbee4eed64c07b9abb8eac70

Observation d3d47cb8-d634-4809-bf01-5d2ea0aafcab · outbound

This paper cites an unresolved cited work.

From Image Captioning to Visual Storytelling Unresolved cited work

Reference 92

Resolution
unresolved
raw_fallback, observed 2026-08-06T10:27:40.594457Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-06T10:27:39.487868Z digest=sha256:5736c146aae0707e83383502126442ea6a8d479bfea87505d238d40ec5384dce

Observation 29a3e9b4-a2d3-46c1-80dc-c3f24e36a74e · outbound

This paper cites Vision-Language Models for Vision Tasks: A Survey.

From Image Captioning to Visual Storytelling Vision-Language Models for Vision Tasks: A Survey

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-06T10:27:39.491216Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T10:27:39.491216Z digest=sha256:ea590f3167a41606f743fa64aaef37f1f4eda6b748ed42ca6558a3a1793753a3

Pith citing papers

No inbound Pith citation observations are available.