Pith. sign in

Paper Citation Record · LEDGER

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer

As of 15 August 2026, this Paper Citation Record lists 63 of 63 outbound references and 0 inbound Pith citation observations for arXiv:2412.11836.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.11836 v1

Coverage vector

measured 63 of 63 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T14:36:11.326499Z

measured 63 of 63 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

63 of 63 outbound references displayed

  • verified exact3
  • verified fuzzy51
  • unresolved9
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 5dcf5570-694a-4036-b6d5-5615c9cdee6d · outbound

This paper cites Bottom -up and top-down attention for image captioning and visual question answering,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Bottom -up and top-down attention for image captioning and visual question answering,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:12.675214Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:10.830605Z digest=sha256:17d71e10b70fd93afb5be11117b66163c3310ea7a9d362ab766a1c36ddf7245a

Observation 33740153-13c8-4e31-874f-f34eac01de6c · outbound

This paper cites Entangled transformer for image captioning,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Entangled transformer for image captioning,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:12.657153Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:10.836912Z digest=sha256:21526b7c3cc3db9e54314c8b507d3d75c3538c0ad2906045f29707d0394f3ae0

Observation 73756992-8194-44f4-a512-386375f61be3 · outbound

This paper cites Automatic alt -text: Computer-generated image descriptions for blind users on a social network service,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Automatic alt -text: Computer-generated image descriptions for blind users on a social network service,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:12.641821Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:10.842187Z digest=sha256:331ca1f473529b9be8989c30c85df2a7f877206f2c58aa87f2a9accd37ee436b

Observation ced5a1ab-4368-433e-aec3-ac86a62ceae0 · outbound

This paper cites Vizwiz grand challenge: Answering visual questions from blind people,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Vizwiz grand challenge: Answering visual questions from blind people,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:12.625214Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:10.847943Z digest=sha256:915e2f765149ca9f0b9dd9391b63ac5494a2b17d0dea820b4f69dba73d4a49b7

Observation e84212a4-f758-4a50-8066-a480211d8b1a · outbound

This paper cites VQA: Visual question answering,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer VQA: Visual question answering,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:12.608995Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:10.853726Z digest=sha256:d854aa93671f23973a4776479048a83c42bacec45324a845a3010185a7b7b666

Observation 6ef1c2ca-6847-4cfb-aab5-110a52965764 · outbound

This paper cites SCA-CNN: Spatial and Channel-wise Attention in Convolutional Networks for Image Captioning.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer SCA-CNN: Spatial and Channel-wise Attention in Convolutional Networks for Image Captioning

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-08-11T14:36:11.693828Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:10.859627Z digest=sha256:5cdd2c2b7b05e729631d7d539d5e569c2c12bb37503b36bfbdb996083ef3f9af

Observation 7b8fdcae-c1bc-4df8-bd3d-62c7cd817a29 · outbound

This paper cites Image captioning using DenseNet network and adaptive attention,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Image captioning using DenseNet network and adaptive attention,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:12.593011Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:10.866084Z digest=sha256:db8055cca00db68372a1e0cdff513d6d3023b0ce7fe37a57c3c2116ff52e9e72

Observation 406f6488-6303-497d-ba87-9d154d94274a · outbound

This paper cites High-Order Interaction Learning for Image Captioning,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer High-Order Interaction Learning for Image Captioning,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:12.575319Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:10.871021Z digest=sha256:9a22b5f19c107a43858759666dabf2165ed886f0b286b433d0c791ae7544d4ef

Observation 73381eec-3e93-40bb-b050-0e0da7c19983 · outbound

This paper cites Stylenet: Generating attractive visual captions with styles,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Stylenet: Generating attractive visual captions with styles,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:12.559263Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:10.876203Z digest=sha256:cc6acb4d44a20cc215c7c26fee124ec5dad4e9e71bffb5c2bf2c650c8a9549ff

Observation 76aeddb8-c4bb-4af0-9c11-cefcc9a24299 · outbound

This paper cites Similar Scenes arouse Similar Emotions: Parallel Data Augmentation for Stylized Image Captioning,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Similar Scenes arouse Similar Emotions: Parallel Data Augmentation for Stylized Image Captioning,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:12.542168Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:10.881597Z digest=sha256:bb7b970f9081d202e7c59240a90e99c2e5a0de4d015c0c5dec7d1eba9fd91d82

Observation b933ddea-0d31-4893-a4ea-7d8c0d89f4ef · outbound

This paper cites “Factual.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer “Factual

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:12.525606Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:10.886556Z digest=sha256:e01596bfb2839f8b7f44c38171ef69b9a837322d0d2cd83e89d421320e97c12e

Observation 0475a42e-e94d-42c9-a58e-4b4f914bae1b · outbound

This paper cites MSap: Multi -Style Image Captioning With Unpaired Stylized Text,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer MSap: Multi -Style Image Captioning With Unpaired Stylized Text,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:12.509646Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:10.891808Z digest=sha256:0dad1b797ded9d42f1a4485b6f552bed8acbc5db65a7fdd31f366ccd8baa45d9

Observation 754592c0-fa9b-45cb-a784-480c6f4b1ead · outbound

This paper cites SummaRuNNer: A Recurrent Neural Network based Sequence Model for Extractive Summarization of Documents,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer SummaRuNNer: A Recurrent Neural Network based Sequence Model for Extractive Summarization of Documents,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:12.493489Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:10.896661Z digest=sha256:f03216d9677bc4a49563629ab940afb41b94c88642b566129cf3a99c13bd5b31

Observation 0559f682-77f9-45f5-8379-ff86f584968e · outbound

This paper cites Abstractive Text Summarization Using Sequence-to-Sequence RNNs and Beyond.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Abstractive Text Summarization Using Sequence-to-Sequence RNNs and Beyond

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-11T14:36:10.901271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:36:10.901271Z digest=sha256:5ff5d82b27bbd3a7eaea789a9a6d8b87d2a9d59fc53b64a74a36f75041e40f13

Observation 41e1617a-e91b-45ba-be05-de0e16934956 · outbound

This paper cites From image descriptions to visual denotations: New similarity metrics for semantic inference over event descriptions,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer From image descriptions to visual denotations: New similarity metrics for semantic inference over event descriptions,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:12.474347Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:10.906245Z digest=sha256:f2ebabae24f67398e233a86fc88c8be01d2a9779483f4cf00ab745374174d6ad

Observation ab859f51-3b72-4e25-954d-5c04434c4879 · outbound

This paper cites Corpus -guided sentence generation of natural images,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Corpus -guided sentence generation of natural images,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:12.458659Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:10.910978Z digest=sha256:0f5bda477e775b008b54e27f894cb32d712183db15e0f933d3d8421e1116f340

Observation 1fd48cea-baf8-45c0-ad6d-90b16d55aef4 · outbound

This paper cites Generating image descriptions from computer vision detections,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Generating image descriptions from computer vision detections,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:12.437163Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:10.915569Z digest=sha256:2fe590fe3fed00b97542c47aa15b42905315e5392e259c98a304eae9129b1490

Observation 5ea2c62a-ba3b-4d27-8f7a-483bdaacc863 · outbound

This paper cites Every picture tells a story: Generating sentences from images,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Every picture tells a story: Generating sentences from images,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:12.418966Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:10.920115Z digest=sha256:56f31689adf2f4ada1889ae6d337a6a82af4279627944e275083d8e8d9d04f3f

Observation d9800e96-5bc3-463b-81fa-6af29f13f755 · outbound

This paper cites Framing image description as a ranking task: Data, models and evaluation metrics,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Framing image description as a ranking task: Data, models and evaluation metrics,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:12.401366Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:10.924698Z digest=sha256:58c745a4b9abc730c402d5b4dcb88f4417d1e9a266ea68c52ed87e2f0e3c62ad

Observation b7a817f2-0297-4ab8-9bf8-0519ec504459 · outbound

This paper cites Deep Captioning with Multimodal Recurrent Neural Networks (m-RNN).

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Deep Captioning with Multimodal Recurrent Neural Networks (m-RNN)

Reference 20

Resolution
verified exact
local_arxiv, observed 2026-08-11T14:36:11.637073Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:10.929525Z digest=sha256:4f8fe38bfa5632ef1f1a789bb7c6a084087f598cbacb7b74d594bce91a9f2a9a

Observation 9c432f49-8056-4a10-a98f-115817353b86 · outbound

This paper cites Show and Tell: A Neural Image Caption Generator,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Show and Tell: A Neural Image Caption Generator,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:12.383317Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:10.935288Z digest=sha256:ecf82c266f4d3e270c5482e0adf3e179022709dab126f2658642d3cd2844a5bb

Observation 7e7d3542-c492-42d8-ba3a-2c7bd26e8ca3 · outbound

This paper cites Guiding the long -short term memory model for image caption generation,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Guiding the long -short term memory model for image caption generation,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:12.365444Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:10.941162Z digest=sha256:30f4415d1abfa86c7d5872434460198cc9a90b5a43b965bda1f379f6804e12c4

Observation 8b686616-17d0-4ce0-9fe2-603bd47c2106 · outbound

This paper cites Long-Term Recurrent Convolutional Networks for Visual Recognition and Description,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Long-Term Recurrent Convolutional Networks for Visual Recognition and Description,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:12.347317Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:10.945752Z digest=sha256:4d791be1735c5542c88243e4d9c100126db90335e750183386c5be649191e778

Observation 82b70d48-3428-46fb-b820-62c3ac98ef5a · outbound

This paper cites Image Captioning with Deep Bidirectional LSTMs,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Image Captioning with Deep Bidirectional LSTMs,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:12.330872Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:10.950396Z digest=sha256:0fec1ecd0e255792b97fe8b2b323882acdf95ed449a7863f86982fd726213e72

Observation 5385956a-552c-4749-8034-d4e30a327ef9 · outbound

This paper cites Knowing When to Look: Adaptive Attention via A Visual Sentinel for Image Captioning.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Knowing When to Look: Adaptive Attention via A Visual Sentinel for Image Captioning

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-11T14:36:10.955397Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:36:10.955397Z digest=sha256:5d4f8b366f195aefb2a372ad1f26665477ddf5f24404843f3215afd8c107c8de

Observation 345def87-d7ed-485f-9fa0-666af389ba29 · outbound

This paper cites Automated Image Caption Generation Framework using Adaptive Attention and Bi-LSTM,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Automated Image Caption Generation Framework using Adaptive Attention and Bi-LSTM,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:12.308873Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:10.961891Z digest=sha256:f6c42c460b2b7f704fcb0c304c65e260a09139673e7943b0328353ef2463026a

Observation d09275ac-15c7-4c24-aae9-1dad69abcfa9 · outbound

This paper cites Exploring region relationships implicitly: Image captioning with visual relationship attention,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Exploring region relationships implicitly: Image captioning with visual relationship attention,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:12.288513Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:10.966713Z digest=sha256:a0ec29378be514ecac06c3d033d33aab213028b3003a130420d8931ee3b06324

Observation 22b7618a-a8a4-4f26-afe7-4e2ad326f445 · outbound

This paper cites Image captioning with semantic attention,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Image captioning with semantic attention,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:12.271829Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:10.971450Z digest=sha256:fd0282964c3749429a7c237231e8fd95f2a6bf89f2b799d1834047de7876ab4e

Observation d77e436a-9321-40e2-a9cc-cc86901e7698 · outbound

This paper cites DAA: Dual LSTMs with adaptive attention for image captioning,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer DAA: Dual LSTMs with adaptive attention for image captioning,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:12.253096Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:10.976138Z digest=sha256:9576344c800aca3cdfd53e532e15d66f10fc0d2d3ee19ae2b94d8c58884912fd

Observation 03224d93-cc52-4926-a6a4-692e450a64ad · outbound

This paper cites Task -Adaptive Attention for Image Captioning,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Task -Adaptive Attention for Image Captioning,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:12.236033Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:10.980764Z digest=sha256:4c74a5ca204722c8b1dab60cb5798982ded9ec8fb68f9e6351e4d55c020b51f7

Observation 108922a8-c154-4c2c-9afa-ac7d4715a6e9 · outbound

This paper cites A New Attention -Based LSTM for Image Captioning,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer A New Attention -Based LSTM for Image Captioning,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:12.217603Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:10.985439Z digest=sha256:6595e7e0a425730525db58453a0f39a104b2ae15d8b7fa48fedb081b94e07f74

Observation e5de7fc4-29a8-46e5-961b-37513e79eb8f · outbound

This paper cites Semstyle: Learning to generate stylised image captions using unaligned text,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Semstyle: Learning to generate stylised image captions using unaligned text,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:12.199066Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:10.990398Z digest=sha256:b21a63a905aeda1d13b5afd210a062804d7c804dadaab5c54fe9c2829fc864b3

Observation c8162142-c764-4cb3-aa13-41d35d8b6ea3 · outbound

This paper cites MemCap: Memorizing Style Knowledge for Image Captioning,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer MemCap: Memorizing Style Knowledge for Image Captioning,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:12.181637Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:10.995823Z digest=sha256:42a8f8edec3310b51064433b88afb120a9de68070650cbb2e87dc24076ae4d7f

Observation 0da9e6ae-c127-46eb-9d2b-83efda56a66e · outbound

This paper cites SentiCap: Generating Image Descriptions with Sentiments,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer SentiCap: Generating Image Descriptions with Sentiments,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:12.163361Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:11.000464Z digest=sha256:34452f8e67adb4be00037fa17fad7da3013b7a18e58a11c4cfd280727ff28e8f

Observation a3c0135e-fea8-4499-9037-7684d658dcbd · outbound

This paper cites Engaging Image Captioning Via Personality.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Engaging Image Captioning Via Personality

Reference 35

Resolution
verified exact
local_arxiv, observed 2026-08-11T14:36:11.570178Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:11.005630Z digest=sha256:cb6573a69e75eb1d67a8d420e6ee66e42cab7b07672d34b5eded6facb1e1437d

Observation 1874a31d-42f9-4070-9c14-2930a52b7351 · outbound

This paper cites Image Captioning with Inherent Sentiment,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Image Captioning with Inherent Sentiment,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:12.146160Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:11.011308Z digest=sha256:10ddbb736c6a4c8a1af1df8a96d89f90173f3d3451077b1b31c695817e315bdb

Observation a537017d-5d0b-4776-bbd1-e5fc6ccc864e · outbound

This paper cites Assessing shallow sentence scoring techniques and combinations for single and multi -document summarization,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Assessing shallow sentence scoring techniques and combinations for single and multi -document summarization,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:12.130452Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:11.016846Z digest=sha256:826797fe75d6d4731473d11eda173dc1fd1d431690f7af2375cccf59b748b52a

Observation 56e78d23-1f88-44ad-bda1-fd540323820b · outbound

This paper cites Summarization of changes in dynamic text collections using latent dirichlet allocation model,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Summarization of changes in dynamic text collections using latent dirichlet allocation model,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:12.111973Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:11.021446Z digest=sha256:beb63baf6b13d06e36319009952805a93f26131a2633a494bdc5ca6d30fb653e

Observation 1aa3fc4a-b0db-43d3-9c23-af939f9a694c · outbound

This paper cites Fully abstractive approach to guided summarization,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Fully abstractive approach to guided summarization,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:12.088961Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:11.026083Z digest=sha256:dc956329829d7e60b211711556289b4796d5892d0c9d62cc08241d3c2682cadf

Observation da0cebe7-cb25-43e9-85fa-fbd112c39daa · outbound

This paper cites A bayesian method to incorporate background knowledge during automatic text summarization,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer A bayesian method to incorporate background knowledge during automatic text summarization,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:12.071321Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:11.031064Z digest=sha256:1a0ce43cce8509317271dc30e16e4f0679acc78f07865e47be522430d0103574

Observation 20c4642e-cead-4692-aaf1-361652a98d13 · outbound

This paper cites Multi-document abstractive summarization using ilp based multi- sentence compression,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Multi-document abstractive summarization using ilp based multi- sentence compression,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:12.039793Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:11.036472Z digest=sha256:9cb9021eca3dc0f268b3549eea48618f240c5ee8aa6e6875228edca112e5352d

Observation ea938668-5ea7-4883-afb0-bfd7a3a8a533 · outbound

This paper cites A Neural Attention Model for Abstractive Sentence Summarization,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer A Neural Attention Model for Abstractive Sentence Summarization,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:12.021356Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:11.042058Z digest=sha256:3881a148f239fb95d3ff580697244d6312ff130b7f241b6d394abfe45bf30075

Observation 12042d4a-1d45-4e75-9e77-d38f42b9c74d · outbound

This paper cites Get To The Point: Summarization with Pointer-Generator Networks.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Get To The Point: Summarization with Pointer-Generator Networks

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-11T14:36:11.047117Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:36:11.047117Z digest=sha256:1d476a36bfb77dc097ea05ee34e354651b7d0d0a462e8c831133aad91fa68eaf

Observation e2b8f66f-c7e6-41cc-8af8-9d954a6c6ee9 · outbound

This paper cites A Unified Model for Extractive and Abstractive Summarization using Inconsistency Loss,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer A Unified Model for Extractive and Abstractive Summarization using Inconsistency Loss,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:12.002684Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:11.052323Z digest=sha256:00673b450f99ff35c4945944cb5b790ec2dd197bff05c468aa6700b563f9f6d3

Observation 83847568-d1a2-4862-8d34-5c73e2586f0b · outbound

This paper cites Abstractive Text Summarization with Multi-Head Attention,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Abstractive Text Summarization with Multi-Head Attention,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:11.980276Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:11.057544Z digest=sha256:867cb17fb0905e40295ff81d2c8a7a01758e40058e9df1daf0bbe7d42f8be32c

Observation f31f7f4a-d289-49f6-a958-b87adf3572fe · outbound

This paper cites Dual Encoding for Abstractive Text Summarization,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Dual Encoding for Abstractive Text Summarization,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:11.962325Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:11.062722Z digest=sha256:4cb0b74b8774f67fbd59a7c670b2e3b23bccc92ddf2ec7b60475f32e0e001b91

Observation 9811558c-4640-4523-93b5-50752940949e · outbound

This paper cites Transformers and Pointer -Generator Networks for Abstractive Summarization,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Transformers and Pointer -Generator Networks for Abstractive Summarization,

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:11.941952Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:11.068759Z digest=sha256:c47e08fe8af6bef10acf0f137fb68750bff8836af3b08f1931b844130855871e

Observation 089aa2d2-9225-44d1-96b8-dcbd8ca95a6b · outbound

This paper cites Attention is all you need,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Attention is all you need,

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:11.918274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:11.074358Z digest=sha256:eb77d6c9610a3fbeb454a886a8b8b5bce1e6e3478cb4e4e2e6c62d49458a8db3

Observation 8e325db5-da09-4cfb-92b1-85dcefa2b212 · outbound

This paper cites Faster R -CNN: Towards Real -Time Object Detection with Region Proposal Networks,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Faster R -CNN: Towards Real -Time Object Detection with Region Proposal Networks,

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:11.896013Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:11.080758Z digest=sha256:3a1650d65b26333e9498e49f9e6bd9cc3d538a1d4c6d1c37299bebaba68e5a08

Observation e104252a-d3f6-4ed9-86c8-bd1b1914b898 · outbound

This paper cites Rethinking the Inception Architecture for Computer Vision,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Rethinking the Inception Architecture for Computer Vision,

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:11.870867Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:11.085932Z digest=sha256:f7889c9feb5c90a04e11c045f5b266e6a565254985a2e939a667b07e1467821b

Observation 5e309434-f6e0-4d60-8c99-cbddadbca44d · outbound

This paper cites Multi-task Sequence to Sequence Learning.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Multi-task Sequence to Sequence Learning

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-11T14:36:11.097117Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:36:11.097117Z digest=sha256:5111f48c80f76e2573516f40d6767d578b364339564ddee27d59f08c1d12c1c4

Observation 6cae5062-235b-4406-b893-5946ea0e7a14 · outbound

This paper cites Enriching Word Vectors with Subword Information.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Enriching Word Vectors with Subword Information

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-11T14:36:11.103396Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:36:11.103396Z digest=sha256:8913f13f574c44f0f836c712ce6ed4a9619909c173395975ac99512120cec240

Observation 3108421a-ec0f-4d30-bdf3-93b674cf8a3a · outbound

This paper cites Neural Machine Translation by Jointly Learning to Align and Translate.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Neural Machine Translation by Jointly Learning to Align and Translate

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-11T14:36:11.109445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:36:11.109445Z digest=sha256:67d13d1c94b470e104dc8e25a12eb2b08923a888c96b93c8aed5b3e77068ca66

Observation 8b0095c1-55e1-48c6-b54c-c2a77eecdb8c · outbound

This paper cites Frustratingly Short Attention Spans in Neural Language Modeling.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Frustratingly Short Attention Spans in Neural Language Modeling

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-11T14:36:11.115085Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:36:11.115085Z digest=sha256:4eecd4413690615a62bcf121b34bb772009e56591d68872a76a748c78bea6ea7

Observation 8881e44a-09cd-42bd-88c7-04f9f25b7727 · outbound

This paper cites "Bleu: a method for automatic evaluation of machine,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer "Bleu: a method for automatic evaluation of machine,

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:11.851559Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:11.121019Z digest=sha256:f91e0f5a54a175d5e981a3a8df66b8715b4b2b41d303e58e751342f49e0138ec

Observation 8655df51-9866-471b-917d-9e0ad80521d6 · outbound

This paper cites Meteor: An automatic Metric for MT Evaluation with Improved Correlation with Human Judgments,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Meteor: An automatic Metric for MT Evaluation with Improved Correlation with Human Judgments,

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:11.829860Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:11.126492Z digest=sha256:6a3690b8e0f27638650e52f999c9be56ac1f12b32639a7c5b1727e0ae4b25124

Observation daf48de2-ca23-4cf0-b01c-5c43d8937fd1 · outbound

This paper cites Adam: A Method for Stochastic Optimization.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Adam: A Method for Stochastic Optimization

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-11T14:36:11.131790Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:36:11.131790Z digest=sha256:74d8755a76f0ae2bd973069d63992f534ace01bf2dc50fd80494f8c1dce73c8c

Observation d136edbe-af6d-4c0e-ad07-036ae801d9f5 · outbound

This paper cites Guiding Long-Short Term Memory for Image Caption Generation.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Guiding Long-Short Term Memory for Image Caption Generation

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-11T14:36:11.137805Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:36:11.137805Z digest=sha256:24c8f8ca4d3e9cd672ee59c52cab86ccd77408d11faf72df3e30390225d92117

Observation edd66cae-c832-4555-8578-81dfcc8d6827 · outbound

This paper cites Multimodal Neural Language Models,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Multimodal Neural Language Models,

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:11.809439Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:11.143935Z digest=sha256:65cbdf9ee0dbafcd2bb02fdfc76c63b5d6c967bb178512eedb888a5fa38029cf

Observation edec3636-92a0-4326-8948-9a33998384ae · outbound

This paper cites Stimulus -driven and concept -driven analysis for image caption generation,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Stimulus -driven and concept -driven analysis for image caption generation,

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:11.790431Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:11.300760Z digest=sha256:be8f1358af235348160dc671f3b9104ce3aa693520405da21fa548b08954b380

Observation e96a5079-addf-4e18-a5f1-4d1694cab461 · outbound

This paper cites Learning joint relationship attention network for image captioning,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Learning joint relationship attention network for image captioning,

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:11.767496Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:11.308224Z digest=sha256:a411a6cad799bdff20fc72d28c74cfbe692bb200a10dc5020895a9d9bb5edd23

Observation 17f63919-fb5b-4918-8536-1997f245ab32 · outbound

This paper cites Detach and Attach: Stylized Image Captioning without Paired Stylized Dataset,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Detach and Attach: Stylized Image Captioning without Paired Stylized Dataset,

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:11.742491Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:11.318914Z digest=sha256:4f5ad62e3a0856d48d25692b2fea209c9edb49f0c258ba05ec6f2eb1efb9bd90

Observation 52339b4d-69e0-4388-93a6-d780ad4cc7fc · outbound

This paper cites Learning Cooperative Neural Modules for Stylized Image Captioning,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Learning Cooperative Neural Modules for Stylized Image Captioning,

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:11.723077Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:11.326499Z digest=sha256:016d235c861a6d0699a58c879f0b8b43db9e845a2620edb3a90e63d69960e4d8

Pith citing papers

No inbound Pith citation observations are available.