Pith. sign in

Paper Citation Record · LEDGER

Show, Tell and Summarize: Dense Video Captioning Using Visual Cue Aided Sentence Summarization

As of 16 August 2026, this Paper Citation Record lists 45 of 45 outbound references and 0 inbound Pith citation observations for arXiv:2506.20567.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.20567 v1

Coverage vector

measured 45 of 45 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T22:50:19.144468Z

measured 45 of 45 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

45 of 45 outbound references displayed

  • verified exact1
  • verified fuzzy34
  • unresolved10
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation be343b00-0a56-4c27-848f-04d828ab018d · outbound

This paper cites Dense- captioning events in videos.

Show, Tell and Summarize: Dense Video Captioning Using Visual Cue Aided Sentence Summarization Dense- captioning events in videos

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:50:21.836616Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T22:50:14.947835Z digest=sha256:c76f415088624e855afd31254f5b4f2d7f149d622d0279b64e091b53c316e40f

Observation 33a67473-8759-42cf-8a2a-607f6a88c040 · outbound

This paper cites Bidirectional attentive fusion with context gating for dense video captioning,.

Show, Tell and Summarize: Dense Video Captioning Using Visual Cue Aided Sentence Summarization Bidirectional attentive fusion with context gating for dense video captioning,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:50:21.824027Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T22:50:15.138319Z digest=sha256:e6b8d2d2afc979ad9a86ec4e83ac43d3ff8939a4bc9877887c7610a861dc3ecf

Observation 2dbc9973-8cec-4bd0-b42f-f5b0aa8e81e4 · outbound

This paper cites Jointly localizing and describing events for dense video captioning,.

Show, Tell and Summarize: Dense Video Captioning Using Visual Cue Aided Sentence Summarization Jointly localizing and describing events for dense video captioning,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:50:21.810753Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T22:50:15.506523Z digest=sha256:cb6c070514e99beabed9a6bf8e19747ea3c7880aae44558a25d97071270a0277

Observation ece2b424-bb17-4a27-8843-0b7c3806da3a · outbound

This paper cites End-to-end dense video captioning with masked transformer,.

Show, Tell and Summarize: Dense Video Captioning Using Visual Cue Aided Sentence Summarization End-to-end dense video captioning with masked transformer,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:50:21.798297Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T22:50:15.751479Z digest=sha256:b69d47d4380515426439dfd04d5163cb3d8e5a9cae919d9b670d7d00eb595967

Observation 5286d440-f748-4688-a186-0688100c1150 · outbound

This paper cites Show, attend and tell: Neural image caption generation with visual attention,.

Show, Tell and Summarize: Dense Video Captioning Using Visual Cue Aided Sentence Summarization Show, attend and tell: Neural image caption generation with visual attention,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:50:21.785815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T22:50:15.821613Z digest=sha256:07c0d11b738de54f00e4264542f31f7fb2eaca7457987ac778fe74fa28ebf5f9

Observation 51087b00-fc3d-4d69-8aed-a39609d4b09a · outbound

This paper cites Automatic annotation of human actions in video,.

Show, Tell and Summarize: Dense Video Captioning Using Visual Cue Aided Sentence Summarization Automatic annotation of human actions in video,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:50:21.773304Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T22:50:15.928863Z digest=sha256:b08aa501f29d87de099ef791890207e4c0159c7398ad86e2ba0bdd2a3490f9b6

Observation 1f492d7f-d6a3-4b07-96ff-394272b47d28 · outbound

This paper cites Fast temporal activity proposals for efficient detection of human actions in untrimmed videos,.

Show, Tell and Summarize: Dense Video Captioning Using Visual Cue Aided Sentence Summarization Fast temporal activity proposals for efficient detection of human actions in untrimmed videos,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:50:21.760433Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T22:50:16.027777Z digest=sha256:479373738205070efb524ff397fcbcec5cda0f0536485c70a782adf105e19d99

Observation 4e2e2220-8af5-4048-bcd9-f6e07f71945a · outbound

This paper cites Daps: Deep action proposals for action understanding,.

Show, Tell and Summarize: Dense Video Captioning Using Visual Cue Aided Sentence Summarization Daps: Deep action proposals for action understanding,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:50:21.747757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T22:50:16.131829Z digest=sha256:34da34de606c1aafcdca7cca1f037b23eed86a96f76d8f5d7c60a14ac1acd60f

Observation 1cb9ebd6-f373-4ab0-9544-2181ac28f2b5 · outbound

This paper cites Sst: Single-stream temporal action proposals,.

Show, Tell and Summarize: Dense Video Captioning Using Visual Cue Aided Sentence Summarization Sst: Single-stream temporal action proposals,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:50:21.734681Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T22:50:16.205197Z digest=sha256:e23b41263489df6818084a3813b348e4f2c87224b8ac09397bdb2d3ad64664b4

Observation 80efb3de-abc5-4b9f-aeeb-46e29a54dec9 · outbound

This paper cites Weakly supervised dense video captioning,.

Show, Tell and Summarize: Dense Video Captioning Using Visual Cue Aided Sentence Summarization Weakly supervised dense video captioning,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:50:21.721776Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T22:50:16.272610Z digest=sha256:ce5187533b227f3e630fa410bc55da4808288c28a9b5fd21772314472b817584

Observation b111f691-401d-42d8-b482-7d930212da28 · outbound

This paper cites Jointly modeling embedding and translation to bridge video and language,.

Show, Tell and Summarize: Dense Video Captioning Using Visual Cue Aided Sentence Summarization Jointly modeling embedding and translation to bridge video and language,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:50:21.708393Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T22:50:16.351104Z digest=sha256:15065c7bc206fe6081ddbff174a12f0c8fa6ab9194139aeb44f5078e749809e7

Observation 34691900-1b4e-4d53-b0ff-4884cf7f6abc · outbound

This paper cites Less Is More: Picking Informative Frames for Video Captioning.

Show, Tell and Summarize: Dense Video Captioning Using Visual Cue Aided Sentence Summarization Less Is More: Picking Informative Frames for Video Captioning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T22:50:16.420489Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:50:16.420489Z digest=sha256:ee0aeac739097e5d28adf62a62123b47a186860449cb0dbc73050d82432d4b37

Observation 64a7f426-94ce-4bee-bcdf-c206b3dcb5bc · outbound

This paper cites Video captioning via hierarchical reinforcement learning,.

Show, Tell and Summarize: Dense Video Captioning Using Visual Cue Aided Sentence Summarization Video captioning via hierarchical reinforcement learning,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:50:21.694298Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T22:50:16.495187Z digest=sha256:6a53429e75635d563860ba28b7f270ad956ae190276332ae6c4a67a46523ebc1

Observation 26fb4082-e17d-4b5c-baa5-69bac34be641 · outbound

This paper cites Reconstruction network for video captioning,.

Show, Tell and Summarize: Dense Video Captioning Using Visual Cue Aided Sentence Summarization Reconstruction network for video captioning,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:50:21.680005Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T22:50:16.571296Z digest=sha256:84ade2c32f149c78208b1a7062c35b84cdd9e15db3d821fdb85f7cc0c04a5547

Observation 9659b044-96d4-4b68-9875-b5f77204ba98 · outbound

This paper cites Interpretable video captioning via trajectory structured localization,.

Show, Tell and Summarize: Dense Video Captioning Using Visual Cue Aided Sentence Summarization Interpretable video captioning via trajectory structured localization,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:50:21.664698Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T22:50:16.644115Z digest=sha256:7a2d49430a7681c7f3058e6f674b64e5a753c11fd08e6591b1b8cb87b26d89a7

Observation f29e687c-2cd5-42ab-889c-d2f3c6092f24 · outbound

This paper cites Regularizing RNNs for Caption Generation by Reconstructing The Past with The Present.

Show, Tell and Summarize: Dense Video Captioning Using Visual Cue Aided Sentence Summarization Regularizing RNNs for Caption Generation by Reconstructing The Past with The Present

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-08-06T22:50:19.418898Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T22:50:16.710262Z digest=sha256:cf2e33751abfa84d1f14b3e53eb2a49313c126595f01e190cb30343a5016f543

Observation ce9bf090-d342-4cfb-a08a-a40191860b7b · outbound

This paper cites Improving action localization by progressive cross-stream cooperation,.

Show, Tell and Summarize: Dense Video Captioning Using Visual Cue Aided Sentence Summarization Improving action localization by progressive cross-stream cooperation,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:50:21.651024Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T22:50:16.774200Z digest=sha256:6a2d40a74020a9415acc6b28b90205bef4c8d4156750b14e71404750f0bdfd6c

Observation 5a5566f1-892b-441a-8b57-612754d6b74f · outbound

This paper cites Translating Videos to Natural Language Using Deep Recurrent Neural Networks.

Show, Tell and Summarize: Dense Video Captioning Using Visual Cue Aided Sentence Summarization Translating Videos to Natural Language Using Deep Recurrent Neural Networks

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T22:50:16.850678Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:50:16.850678Z digest=sha256:4d87fd5f8be62e057913cbf54916be04e02f8ee34ada627ac578521bc37741e5

Observation b6da2887-c6d6-46f5-8f61-df3d60d97517 · outbound

This paper cites Semantic compositional networks for visual captioning,.

Show, Tell and Summarize: Dense Video Captioning Using Visual Cue Aided Sentence Summarization Semantic compositional networks for visual captioning,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:50:21.638022Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T22:50:16.928322Z digest=sha256:52a5842b5fe52b31db930723e6d4da2ff8e55fc390f14c3fdf253992f2fe7c97

Observation cba3db13-0d13-4e52-acdb-799695ca3fc8 · outbound

This paper cites Describing videos by exploiting temporal structure,.

Show, Tell and Summarize: Dense Video Captioning Using Visual Cue Aided Sentence Summarization Describing videos by exploiting temporal structure,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:50:21.624379Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T22:50:16.993545Z digest=sha256:6102abf52b801ccc56145634ebe284f31ff114bd911b60e1f676ea3f74968955

Observation eecebb60-f696-4b9d-8081-c57a3fb3b305 · outbound

This paper cites Video captioning with transferred semantic attributes,.

Show, Tell and Summarize: Dense Video Captioning Using Visual Cue Aided Sentence Summarization Video captioning with transferred semantic attributes,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:50:21.611453Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T22:50:17.057222Z digest=sha256:2ddbef00f250280927a610df53b790a4b8cd60020dbca82dd23f1ffbec12bca0

Observation 47670dd5-f172-4f4a-8a76-f47d67139d96 · outbound

This paper cites Sequence to sequence learning with neural networks,.

Show, Tell and Summarize: Dense Video Captioning Using Visual Cue Aided Sentence Summarization Sequence to sequence learning with neural networks,

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T22:50:17.136369Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:50:17.136369Z digest=sha256:ca00fe0680774b7fb97bfccf93bf5f164b0b6da595de3e8d4bcdcf8bd6ca711f

Observation 3fca801c-b22a-427b-9c4a-81463bd9f9ab · outbound

This paper cites Long-term recurrent convolutional IEEE TRANSACTIONS ON XXXX, VOL.XX, NO.XX, XXXX,XXXX 10 networks for visual recognition and description,.

Show, Tell and Summarize: Dense Video Captioning Using Visual Cue Aided Sentence Summarization Long-term recurrent convolutional IEEE TRANSACTIONS ON XXXX, VOL.XX, NO.XX, XXXX,XXXX 10 networks for visual recognition and description,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:50:21.588755Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T22:50:17.171573Z digest=sha256:c97e33e41ce9e2046db3ee120138e24001357782bab1d0e8f3beb13ba47dcf69

Observation 9e959d1f-aeee-4df8-a6bb-b2575e7027fd · outbound

This paper cites Hierarchical recurrent neural encoder for video representation with application to captioning,.

Show, Tell and Summarize: Dense Video Captioning Using Visual Cue Aided Sentence Summarization Hierarchical recurrent neural encoder for video representation with application to captioning,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:50:21.574208Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T22:50:17.278138Z digest=sha256:bec692996bc2da26c59cdf2960ee7cb5c5826ef38b0573db9af0c25fe49e4fa8

Observation afdcac9b-cd09-4ae5-b00b-969b5c1147f1 · outbound

This paper cites Summarization-based video caption via deep neural networks,.

Show, Tell and Summarize: Dense Video Captioning Using Visual Cue Aided Sentence Summarization Summarization-based video caption via deep neural networks,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:50:21.558868Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T22:50:17.341956Z digest=sha256:6a4de327c3cc94dace304a2efa1ec0bf4a5da5d13523f57ff922ae80d8a55c62

Observation bcfd8f4c-8c67-4854-8a7f-cc145ae2ae7a · outbound

This paper cites Boosting video description generation by explicitly translating from frame-level captions,.

Show, Tell and Summarize: Dense Video Captioning Using Visual Cue Aided Sentence Summarization Boosting video description generation by explicitly translating from frame-level captions,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:50:21.546119Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T22:50:17.443160Z digest=sha256:5c9fc0c6bad0913bc512c20d7f55b510403bab831ea89f71e60bf9d8b5279910

Observation 153e326f-a6cc-4833-b38a-90984ed248d2 · outbound

This paper cites Lexrank: Graph-based lexical centrality as salience in text summarization,.

Show, Tell and Summarize: Dense Video Captioning Using Visual Cue Aided Sentence Summarization Lexrank: Graph-based lexical centrality as salience in text summarization,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:50:21.533177Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T22:50:17.493254Z digest=sha256:ea131f4a66eadfdfa31dbb3233559a7a9cfde6f50e4025678ac08fea1ee2cefa

Observation 5096925b-ce27-4472-80d4-02e4aed4e32e · outbound

This paper cites Ask, attend and answer: Exploring question- guided spatial attention for visual question answering,.

Show, Tell and Summarize: Dense Video Captioning Using Visual Cue Aided Sentence Summarization Ask, attend and answer: Exploring question- guided spatial attention for visual question answering,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:50:21.519131Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T22:50:17.632919Z digest=sha256:a27646d7837b72e488ba5c1a5767a7d79f2b6f068fe4c4de952ec73c7f5bf5e2

Observation 9142f11f-c8c5-4c42-bcb5-f0fefa19ec27 · outbound

This paper cites An end-to-end spatio- temporal attention model for human action recognition from skeleton data.

Show, Tell and Summarize: Dense Video Captioning Using Visual Cue Aided Sentence Summarization An end-to-end spatio- temporal attention model for human action recognition from skeleton data

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:50:21.431310Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T22:50:17.727526Z digest=sha256:fb8959799884350201e532b7473d9f30ba2882e570bbe9e9ef5261492eaf407b

Observation c238bbcc-909f-4504-95d2-52234236109a · outbound

This paper cites Image captioning with semantic attention,.

Show, Tell and Summarize: Dense Video Captioning Using Visual Cue Aided Sentence Summarization Image captioning with semantic attention,

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T22:50:17.810265Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:50:17.810265Z digest=sha256:f4267883d1a4a90d33e83d988e45c0565925d12a5dc1ae5e88afbd8ce2496231

Observation 67b1e6b7-a537-4a0e-ad74-280516af2d9b · outbound

This paper cites Stacked attention networks for image question answering,.

Show, Tell and Summarize: Dense Video Captioning Using Visual Cue Aided Sentence Summarization Stacked attention networks for image question answering,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:50:21.206292Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T22:50:17.888817Z digest=sha256:1fee959b6dc90aa45fbb65ad1432fa33ac33bedb6fd07b2b34fc5af0c59e961c

Observation d2a57b90-233d-4438-ba2c-26a090fb6fb9 · outbound

This paper cites Residual Attention Network for Image Classification.

Show, Tell and Summarize: Dense Video Captioning Using Visual Cue Aided Sentence Summarization Residual Attention Network for Image Classification

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T22:50:17.952445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:50:17.952445Z digest=sha256:c8339e8c57aba15139e77d6220dc8a7f5f9995be7e36c0e7940d93312d51b352

Observation f185ee68-2d35-4448-99b7-2ed1a5f56e17 · outbound

This paper cites Multiple Object Recognition with Visual Attention.

Show, Tell and Summarize: Dense Video Captioning Using Visual Cue Aided Sentence Summarization Multiple Object Recognition with Visual Attention

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T22:50:18.031129Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:50:18.031129Z digest=sha256:a15cf2637e3351db2f2566f7f9151d801323334d09cf13ab2967b9a320419eb2

Observation 64f901d5-1fc5-4219-bbb6-5cc4ab005fdc · outbound

This paper cites Show and tell: A neural image caption generator,.

Show, Tell and Summarize: Dense Video Captioning Using Visual Cue Aided Sentence Summarization Show and tell: A neural image caption generator,

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T22:50:18.148720Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:50:18.148720Z digest=sha256:52139db0c3f35566bcf5b3db008d17e416fa25c91038788c453207b2535c5ff1

Observation 7acad870-ab6f-4d61-ab92-81cc6614a751 · outbound

This paper cites Long short-term memory,.

Show, Tell and Summarize: Dense Video Captioning Using Visual Cue Aided Sentence Summarization Long short-term memory,

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T22:50:18.251741Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:50:18.251741Z digest=sha256:c8cc0018215ca1d026f27c5ab4a5684dc8f9f3614e300a5a33c7bd376c627dfb

Observation 776b54e1-8256-4075-a76c-48aaaea7ddc3 · outbound

This paper cites Review networks for caption generation,.

Show, Tell and Summarize: Dense Video Captioning Using Visual Cue Aided Sentence Summarization Review networks for caption generation,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:50:21.098629Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T22:50:18.310504Z digest=sha256:b210f0e0293784e095507d3dbf14d9f449e8c5205df0c0150a29b37eac915fbf

Observation a64cea73-d1ab-4008-88cd-dd6170dd4276 · outbound

This paper cites Self- critical sequence training for image captioning,.

Show, Tell and Summarize: Dense Video Captioning Using Visual Cue Aided Sentence Summarization Self- critical sequence training for image captioning,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:50:20.810377Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T22:50:18.416926Z digest=sha256:3b575deb3d3089d84681b5aa08d04e4ab7455cc89de8ebed788d22cfb2ae29e3

Observation 34981dca-27dd-40ef-b3f7-81a90eee49c4 · outbound

This paper cites Sequence to sequence-video to text,.

Show, Tell and Summarize: Dense Video Captioning Using Visual Cue Aided Sentence Summarization Sequence to sequence-video to text,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:50:20.599387Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T22:50:18.523498Z digest=sha256:90bed6f4e1e1ca5459559cd755015877250574d80115f214b2186ef9dec29c48

Observation 7b410b05-f45d-4306-931c-1e325669fc92 · outbound

This paper cites Video paragraph captioning using hierarchical recurrent neural networks,.

Show, Tell and Summarize: Dense Video Captioning Using Visual Cue Aided Sentence Summarization Video paragraph captioning using hierarchical recurrent neural networks,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:50:20.332965Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T22:50:18.609179Z digest=sha256:be6c506d2c75556ecef29124a3cd402c8e5274ba9d36653f32b6f55e1d7ebddd

Observation 6f361231-9711-47cb-9312-7f862d8e5375 · outbound

This paper cites Activitynet: A large-scale video benchmark for human activity under- standing,.

Show, Tell and Summarize: Dense Video Captioning Using Visual Cue Aided Sentence Summarization Activitynet: A large-scale video benchmark for human activity under- standing,

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T22:50:18.680565Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:50:18.680565Z digest=sha256:6cb97619407acca29841e35e88f5a7378ad5016687e3de7292539ad54dcca584

Observation 70c9825e-c934-431e-8d64-f54b8d32d00f · outbound

This paper cites Adam: A Method for Stochastic Optimization.

Show, Tell and Summarize: Dense Video Captioning Using Visual Cue Aided Sentence Summarization Adam: A Method for Stochastic Optimization

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-06T22:50:18.810031Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:50:18.810031Z digest=sha256:7d3100fc14a8f9be2716d095050a128be971db002cc228351a404e4a28408b7a

Observation 32085a35-0253-4df7-8cd5-ad8bcac640db · outbound

This paper cites Bleu: a method for automatic evaluation of machine translation,.

Show, Tell and Summarize: Dense Video Captioning Using Visual Cue Aided Sentence Summarization Bleu: a method for automatic evaluation of machine translation,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:50:20.185852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T22:50:18.910040Z digest=sha256:b82276f3cd0afe98a1c436dd1a94f1acbdaa84253beaad5667a9fa8d7e3de648

Observation 01d30384-c551-4554-8164-c849f04af813 · outbound

This paper cites Rouge: A package for automatic evaluation of summaries,.

Show, Tell and Summarize: Dense Video Captioning Using Visual Cue Aided Sentence Summarization Rouge: A package for automatic evaluation of summaries,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:50:20.023610Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T22:50:18.986504Z digest=sha256:bda269cd06db36c0e6441068cd532026631d3d35bd22befd82b69e5a213746ae

Observation e26829a0-8529-4f03-830a-b4ef52a28203 · outbound

This paper cites Meteor universal: Language specific translation evaluation for any target language,.

Show, Tell and Summarize: Dense Video Captioning Using Visual Cue Aided Sentence Summarization Meteor universal: Language specific translation evaluation for any target language,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:50:19.837312Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T22:50:19.064066Z digest=sha256:d492199c52659ecbee4a2bed6d35496d266757025ef48b99da23883446c0eed9

Observation 74b25128-b4d9-47a4-83a4-de850d387842 · outbound

This paper cites Cider: Consensus- based image description evaluation,.

Show, Tell and Summarize: Dense Video Captioning Using Visual Cue Aided Sentence Summarization Cider: Consensus- based image description evaluation,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T22:50:19.640902Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-06T22:50:19.144468Z digest=sha256:a2dac795108f4fd95a7533c9742698a0a4a3ac01681aa3a510b3f60a344e1978

Pith citing papers

No inbound Pith citation observations are available.