Pith. sign in

Paper Citation Record · LEDGER

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer

As of 15 August 2026, this Paper Citation Record lists 63 of 63 outbound references and 0 inbound Pith citation observations for arXiv:2412.11836.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.11836 v1

Coverage vector

measured 63 of 63 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T14:36:11.326499Z

measured 63 of 63 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

63 of 63 outbound references displayed

  • verified exact3
  • verified fuzzy51
  • unresolved9
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 5dcf5570-694a-4036-b6d5-5615c9cdee6d · outbound

This paper cites Bottom -up and top-down attention for image captioning and visual question answering,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Bottom -up and top-down attention for image captioning and visual question answering,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:12.675214Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:10.830605Z digest=sha256:2640a79a7b62a24f13d4d38d15d1de0e29ca3c573ccd7f121c079f2db71c2821

Observation 33740153-13c8-4e31-874f-f34eac01de6c · outbound

This paper cites Entangled transformer for image captioning,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Entangled transformer for image captioning,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:12.657153Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:10.836912Z digest=sha256:4150d60a24e4e771672baa6c336e1f1ee0cd63692243ea4d1e74a9f9ef05a3ca

Observation 73756992-8194-44f4-a512-386375f61be3 · outbound

This paper cites Automatic alt -text: Computer-generated image descriptions for blind users on a social network service,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Automatic alt -text: Computer-generated image descriptions for blind users on a social network service,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:12.641821Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:10.842187Z digest=sha256:a0545002915b935c2f4ddf710bf5b449d9290d9a2c6c582991c46c04744de78a

Observation ced5a1ab-4368-433e-aec3-ac86a62ceae0 · outbound

This paper cites Vizwiz grand challenge: Answering visual questions from blind people,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Vizwiz grand challenge: Answering visual questions from blind people,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:12.625214Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:10.847943Z digest=sha256:db09949f9e3cbe7c99e1c49ac18744fc94297812cb28b157a57398d425855842

Observation e84212a4-f758-4a50-8066-a480211d8b1a · outbound

This paper cites VQA: Visual question answering,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer VQA: Visual question answering,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:12.608995Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:10.853726Z digest=sha256:664ec27407f5c0f47d0e5f2b182adbe7575cbce99caed089bc68a29361884b3e

Observation 6ef1c2ca-6847-4cfb-aab5-110a52965764 · outbound

This paper cites SCA-CNN: Spatial and Channel-wise Attention in Convolutional Networks for Image Captioning.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer SCA-CNN: Spatial and Channel-wise Attention in Convolutional Networks for Image Captioning

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-08-11T14:36:11.693828Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:10.859627Z digest=sha256:52869493a9be32863552bfc03b798fc677f317e4cb1be55ff46b8fba29d2b96c

Observation 7b8fdcae-c1bc-4df8-bd3d-62c7cd817a29 · outbound

This paper cites Image captioning using DenseNet network and adaptive attention,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Image captioning using DenseNet network and adaptive attention,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:12.593011Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:10.866084Z digest=sha256:5b96e9fb98c433284952d8da0a83b4cf49afe0a193da096ca427eebef83d8444

Observation 406f6488-6303-497d-ba87-9d154d94274a · outbound

This paper cites High-Order Interaction Learning for Image Captioning,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer High-Order Interaction Learning for Image Captioning,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:12.575319Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:10.871021Z digest=sha256:1ba860e386949eac256d337596172b66f042fa8d87964db95992fa939c45a291

Observation 73381eec-3e93-40bb-b050-0e0da7c19983 · outbound

This paper cites Stylenet: Generating attractive visual captions with styles,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Stylenet: Generating attractive visual captions with styles,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:12.559263Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:10.876203Z digest=sha256:b29f307888a621319eb82f9e41391cb180873bba1b67519b1b7ae94c81b10bbc

Observation 76aeddb8-c4bb-4af0-9c11-cefcc9a24299 · outbound

This paper cites Similar Scenes arouse Similar Emotions: Parallel Data Augmentation for Stylized Image Captioning,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Similar Scenes arouse Similar Emotions: Parallel Data Augmentation for Stylized Image Captioning,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:12.542168Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:10.881597Z digest=sha256:be3c3458d6e2e0e2d952465b7f9add2daf41cf6941e21d00103a41bd531a656b

Observation b933ddea-0d31-4893-a4ea-7d8c0d89f4ef · outbound

This paper cites “Factual.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer “Factual

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:12.525606Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:10.886556Z digest=sha256:dd769c1fad1a835e1b11efac48dd8783087741d98cd97d259cde831694d8ac69

Observation 0475a42e-e94d-42c9-a58e-4b4f914bae1b · outbound

This paper cites MSap: Multi -Style Image Captioning With Unpaired Stylized Text,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer MSap: Multi -Style Image Captioning With Unpaired Stylized Text,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:12.509646Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:10.891808Z digest=sha256:684fe017d3578750049ce64423d68e0ee96d4aa1c2673ab17ac623739f2792cd

Observation 754592c0-fa9b-45cb-a784-480c6f4b1ead · outbound

This paper cites SummaRuNNer: A Recurrent Neural Network based Sequence Model for Extractive Summarization of Documents,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer SummaRuNNer: A Recurrent Neural Network based Sequence Model for Extractive Summarization of Documents,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:12.493489Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:10.896661Z digest=sha256:1e0e3cd0df6eb073452b0961a38d0ab005bc33166f0b0091126554eec0b62537

Observation 0559f682-77f9-45f5-8379-ff86f584968e · outbound

This paper cites Abstractive Text Summarization Using Sequence-to-Sequence RNNs and Beyond.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Abstractive Text Summarization Using Sequence-to-Sequence RNNs and Beyond

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-11T14:36:10.901271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:36:10.901271Z digest=sha256:07e0bfd48c1875802ce8398a21ec08db33d4c9a2ec3e89afbffae96fe7f013e7

Observation 41e1617a-e91b-45ba-be05-de0e16934956 · outbound

This paper cites From image descriptions to visual denotations: New similarity metrics for semantic inference over event descriptions,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer From image descriptions to visual denotations: New similarity metrics for semantic inference over event descriptions,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:12.474347Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:10.906245Z digest=sha256:5ae17bf6646f0633b46debcb06f12de15cb5eb932ae47137595e78aeef6c4b06

Observation ab859f51-3b72-4e25-954d-5c04434c4879 · outbound

This paper cites Corpus -guided sentence generation of natural images,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Corpus -guided sentence generation of natural images,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:12.458659Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:10.910978Z digest=sha256:84dcf3185e4643b07f1b67d660cdb2b272dcaeb2e8b772bd29c80ebba4511518

Observation 1fd48cea-baf8-45c0-ad6d-90b16d55aef4 · outbound

This paper cites Generating image descriptions from computer vision detections,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Generating image descriptions from computer vision detections,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:12.437163Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:10.915569Z digest=sha256:798ec6872c821fffa4cac7eb9d514a2a40388ee7aa05468db6ef9aeb585108aa

Observation 5ea2c62a-ba3b-4d27-8f7a-483bdaacc863 · outbound

This paper cites Every picture tells a story: Generating sentences from images,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Every picture tells a story: Generating sentences from images,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:12.418966Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:10.920115Z digest=sha256:c9005dff0150d09ebd96594bd82ac8dd5deb0c79cc0f2d16139872cf153a5db5

Observation d9800e96-5bc3-463b-81fa-6af29f13f755 · outbound

This paper cites Framing image description as a ranking task: Data, models and evaluation metrics,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Framing image description as a ranking task: Data, models and evaluation metrics,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:12.401366Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:10.924698Z digest=sha256:10ff80bba0493cd81a8870150a163df930cdf8c7b6dca669d3d88202d586f528

Observation b7a817f2-0297-4ab8-9bf8-0519ec504459 · outbound

This paper cites Deep Captioning with Multimodal Recurrent Neural Networks (m-RNN).

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Deep Captioning with Multimodal Recurrent Neural Networks (m-RNN)

Reference 20

Resolution
verified exact
local_arxiv, observed 2026-08-11T14:36:11.637073Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:10.929525Z digest=sha256:15330f6308cf6e6fba3edfb72bbbe2066ed7069594a1e38e32e772f7c00dd0c6

Observation 9c432f49-8056-4a10-a98f-115817353b86 · outbound

This paper cites Show and Tell: A Neural Image Caption Generator,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Show and Tell: A Neural Image Caption Generator,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:12.383317Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:10.935288Z digest=sha256:86d287c3c9713fc5af0abd35b2892eeaee60282875a5fa7ba885a56102d9aa81

Observation 7e7d3542-c492-42d8-ba3a-2c7bd26e8ca3 · outbound

This paper cites Guiding the long -short term memory model for image caption generation,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Guiding the long -short term memory model for image caption generation,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:12.365444Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:10.941162Z digest=sha256:57eac97aaa925ba719e16fd8a43986413356ce2a5fdae537099241495fd1527d

Observation 8b686616-17d0-4ce0-9fe2-603bd47c2106 · outbound

This paper cites Long-Term Recurrent Convolutional Networks for Visual Recognition and Description,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Long-Term Recurrent Convolutional Networks for Visual Recognition and Description,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:12.347317Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:10.945752Z digest=sha256:1b2b060dc07b6dd9884a4cdf83498247b1e85252930de298ab725427692b672e

Observation 82b70d48-3428-46fb-b820-62c3ac98ef5a · outbound

This paper cites Image Captioning with Deep Bidirectional LSTMs,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Image Captioning with Deep Bidirectional LSTMs,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:12.330872Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:10.950396Z digest=sha256:4a778618ab6f5b070dc4618501feccc258426ee08543d16d9320e2d044cc19a6

Observation 5385956a-552c-4749-8034-d4e30a327ef9 · outbound

This paper cites Knowing When to Look: Adaptive Attention via A Visual Sentinel for Image Captioning.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Knowing When to Look: Adaptive Attention via A Visual Sentinel for Image Captioning

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-11T14:36:10.955397Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:36:10.955397Z digest=sha256:a729f721a01c7d6700fd883bddfab89726beaba02297531b7f231ab99a583771

Observation 345def87-d7ed-485f-9fa0-666af389ba29 · outbound

This paper cites Automated Image Caption Generation Framework using Adaptive Attention and Bi-LSTM,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Automated Image Caption Generation Framework using Adaptive Attention and Bi-LSTM,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:12.308873Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:10.961891Z digest=sha256:547236dddaf34fc7902a049afe744215bc2019556d8d5cca37966cb97860fa2e

Observation d09275ac-15c7-4c24-aae9-1dad69abcfa9 · outbound

This paper cites Exploring region relationships implicitly: Image captioning with visual relationship attention,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Exploring region relationships implicitly: Image captioning with visual relationship attention,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:12.288513Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:10.966713Z digest=sha256:4d7c6fe1a1798970baba8f025b363e83d8cd3ea7d25786abf9c8417502217041

Observation 22b7618a-a8a4-4f26-afe7-4e2ad326f445 · outbound

This paper cites Image captioning with semantic attention,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Image captioning with semantic attention,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:12.271829Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:10.971450Z digest=sha256:5e90f506f9213ca5ea070a53ab0b682283216ff5ce087e31223253fedbf3e66b

Observation d77e436a-9321-40e2-a9cc-cc86901e7698 · outbound

This paper cites DAA: Dual LSTMs with adaptive attention for image captioning,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer DAA: Dual LSTMs with adaptive attention for image captioning,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:12.253096Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:10.976138Z digest=sha256:8876a6d2e1c60d612dad875b739609ed99161a961370fae5d6a5812b255110e2

Observation 03224d93-cc52-4926-a6a4-692e450a64ad · outbound

This paper cites Task -Adaptive Attention for Image Captioning,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Task -Adaptive Attention for Image Captioning,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:12.236033Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:10.980764Z digest=sha256:acd561795ceb027171daf7d672651000d9f702d96d00125c6eb4db11d2daffa1

Observation 108922a8-c154-4c2c-9afa-ac7d4715a6e9 · outbound

This paper cites A New Attention -Based LSTM for Image Captioning,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer A New Attention -Based LSTM for Image Captioning,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:12.217603Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:10.985439Z digest=sha256:989118f0e50120496a765eb6cd4e514b704b3d1f7a2aaa62426b556bcd18d6eb

Observation e5de7fc4-29a8-46e5-961b-37513e79eb8f · outbound

This paper cites Semstyle: Learning to generate stylised image captions using unaligned text,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Semstyle: Learning to generate stylised image captions using unaligned text,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:12.199066Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:10.990398Z digest=sha256:1e3c3d13966c2684bc189ffe59e26939e7b15e36296c63bfb5ceae99280d55cb

Observation c8162142-c764-4cb3-aa13-41d35d8b6ea3 · outbound

This paper cites MemCap: Memorizing Style Knowledge for Image Captioning,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer MemCap: Memorizing Style Knowledge for Image Captioning,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:12.181637Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:10.995823Z digest=sha256:e2aaf467f2299e6102d7f79bc9da606bb48710f17eb56c3c17ef969c2247281a

Observation 0da9e6ae-c127-46eb-9d2b-83efda56a66e · outbound

This paper cites SentiCap: Generating Image Descriptions with Sentiments,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer SentiCap: Generating Image Descriptions with Sentiments,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:12.163361Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:11.000464Z digest=sha256:7aca10735935e3ac95938a2874173e2fadc4f58271834826381e8f9dac6c037f

Observation a3c0135e-fea8-4499-9037-7684d658dcbd · outbound

This paper cites Engaging Image Captioning Via Personality.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Engaging Image Captioning Via Personality

Reference 35

Resolution
verified exact
local_arxiv, observed 2026-08-11T14:36:11.570178Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:11.005630Z digest=sha256:187a4ebc2a258f59952acfcfde3cce30170c9cb491c7c33f6f9eb70c31b06e57

Observation 1874a31d-42f9-4070-9c14-2930a52b7351 · outbound

This paper cites Image Captioning with Inherent Sentiment,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Image Captioning with Inherent Sentiment,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:12.146160Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:11.011308Z digest=sha256:a49b5165a4a6647b7bf6cb9ac7e78b5db08c4b4f271ced6ff12f5883901aae08

Observation a537017d-5d0b-4776-bbd1-e5fc6ccc864e · outbound

This paper cites Assessing shallow sentence scoring techniques and combinations for single and multi -document summarization,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Assessing shallow sentence scoring techniques and combinations for single and multi -document summarization,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:12.130452Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:11.016846Z digest=sha256:4795f13218ab897b5101b8b9b6cc47fb12da99f6f4186f4ee5367a0ae3a85afe

Observation 56e78d23-1f88-44ad-bda1-fd540323820b · outbound

This paper cites Summarization of changes in dynamic text collections using latent dirichlet allocation model,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Summarization of changes in dynamic text collections using latent dirichlet allocation model,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:12.111973Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:11.021446Z digest=sha256:2709102f7609ef28565cae5114af82906154188b22333786048984612095e3de

Observation 1aa3fc4a-b0db-43d3-9c23-af939f9a694c · outbound

This paper cites Fully abstractive approach to guided summarization,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Fully abstractive approach to guided summarization,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:12.088961Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:11.026083Z digest=sha256:c15ea3ad20c4c9451a814f54595fae889864616b38284d0177b2043553583892

Observation da0cebe7-cb25-43e9-85fa-fbd112c39daa · outbound

This paper cites A bayesian method to incorporate background knowledge during automatic text summarization,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer A bayesian method to incorporate background knowledge during automatic text summarization,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:12.071321Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:11.031064Z digest=sha256:fcd15392af73ca4936592b8ff979cd2a7912defbaea87d431019c645e2a17f63

Observation 20c4642e-cead-4692-aaf1-361652a98d13 · outbound

This paper cites Multi-document abstractive summarization using ilp based multi- sentence compression,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Multi-document abstractive summarization using ilp based multi- sentence compression,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:12.039793Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:11.036472Z digest=sha256:618656ac1b916730d2a648a42256f05efd553c109da072402d7fa35b7752288f

Observation ea938668-5ea7-4883-afb0-bfd7a3a8a533 · outbound

This paper cites A Neural Attention Model for Abstractive Sentence Summarization,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer A Neural Attention Model for Abstractive Sentence Summarization,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:12.021356Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:11.042058Z digest=sha256:4dc2d6f739a1220717eab3fc10898cba1bb91ebe7bf7c14234c6ba8bf120b2a1

Observation 12042d4a-1d45-4e75-9e77-d38f42b9c74d · outbound

This paper cites Get To The Point: Summarization with Pointer-Generator Networks.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Get To The Point: Summarization with Pointer-Generator Networks

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-11T14:36:11.047117Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:36:11.047117Z digest=sha256:f89251b37ce1493d91737dbe5cb25f9af5e4e5d332266bcd68ff627830a16527

Observation e2b8f66f-c7e6-41cc-8af8-9d954a6c6ee9 · outbound

This paper cites A Unified Model for Extractive and Abstractive Summarization using Inconsistency Loss,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer A Unified Model for Extractive and Abstractive Summarization using Inconsistency Loss,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:12.002684Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:11.052323Z digest=sha256:bf20a311e81c775608f1515d9751e6c5eb80ff1b01b1240e75ff839b59687001

Observation 83847568-d1a2-4862-8d34-5c73e2586f0b · outbound

This paper cites Abstractive Text Summarization with Multi-Head Attention,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Abstractive Text Summarization with Multi-Head Attention,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:11.980276Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:11.057544Z digest=sha256:c0f50df1a792f481bccb553b2e8ee9c5d2726ef0d5a3d159e9de6f07e48b98a1

Observation f31f7f4a-d289-49f6-a958-b87adf3572fe · outbound

This paper cites Dual Encoding for Abstractive Text Summarization,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Dual Encoding for Abstractive Text Summarization,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:11.962325Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:11.062722Z digest=sha256:c352146e80c11456db48897e44c5f17613ae22982d0771c8993eee4154286064

Observation 9811558c-4640-4523-93b5-50752940949e · outbound

This paper cites Transformers and Pointer -Generator Networks for Abstractive Summarization,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Transformers and Pointer -Generator Networks for Abstractive Summarization,

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:11.941952Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:11.068759Z digest=sha256:ab75900f1116c1a2123b914947da9d9ab452e6c5efafdec5cebb6d83f91c0215

Observation 089aa2d2-9225-44d1-96b8-dcbd8ca95a6b · outbound

This paper cites Attention is all you need,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Attention is all you need,

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:11.918274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:11.074358Z digest=sha256:a13f8937814512b29e70b78bf9e0485498c348dc9161bea3d21c7647c9a346a9

Observation 8e325db5-da09-4cfb-92b1-85dcefa2b212 · outbound

This paper cites Faster R -CNN: Towards Real -Time Object Detection with Region Proposal Networks,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Faster R -CNN: Towards Real -Time Object Detection with Region Proposal Networks,

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:11.896013Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:11.080758Z digest=sha256:1c6f2f6706b3ffa767ffcc6af4d75675eb2b808862c82da04c0ffc4d09c75278

Observation e104252a-d3f6-4ed9-86c8-bd1b1914b898 · outbound

This paper cites Rethinking the Inception Architecture for Computer Vision,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Rethinking the Inception Architecture for Computer Vision,

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:11.870867Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:11.085932Z digest=sha256:b05bb936e1da56d32d6f3e89eaf70763a2148641314685e6640552133e9b5f1d

Observation 5e309434-f6e0-4d60-8c99-cbddadbca44d · outbound

This paper cites Multi-task Sequence to Sequence Learning.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Multi-task Sequence to Sequence Learning

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-11T14:36:11.097117Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:36:11.097117Z digest=sha256:d69e6ddfeda91c4ba6ad549c15579684893aa66558d62b416928d342aa145b10

Observation 6cae5062-235b-4406-b893-5946ea0e7a14 · outbound

This paper cites Enriching Word Vectors with Subword Information.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Enriching Word Vectors with Subword Information

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-11T14:36:11.103396Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:36:11.103396Z digest=sha256:cd0f0b318dd5e17bf00326d4843adc901a32f953ee3d144ba3e52fce34665f69

Observation 3108421a-ec0f-4d30-bdf3-93b674cf8a3a · outbound

This paper cites Neural Machine Translation by Jointly Learning to Align and Translate.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Neural Machine Translation by Jointly Learning to Align and Translate

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-11T14:36:11.109445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:36:11.109445Z digest=sha256:2530f569c8ece58e44a89e9dbafeb2748765605bf876c2b0bfc89e1271187cfa

Observation 8b0095c1-55e1-48c6-b54c-c2a77eecdb8c · outbound

This paper cites Frustratingly Short Attention Spans in Neural Language Modeling.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Frustratingly Short Attention Spans in Neural Language Modeling

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-11T14:36:11.115085Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:36:11.115085Z digest=sha256:ea246b4c74f0cdb4c3602be221060589ab0b43d3a3d99c121937341afb2fc83b

Observation 8881e44a-09cd-42bd-88c7-04f9f25b7727 · outbound

This paper cites "Bleu: a method for automatic evaluation of machine,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer "Bleu: a method for automatic evaluation of machine,

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:11.851559Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:11.121019Z digest=sha256:48dff15cc9178c21e6af7ac9f6d973108c8b3988ead989e1eb281700fffd8d35

Observation 8655df51-9866-471b-917d-9e0ad80521d6 · outbound

This paper cites Meteor: An automatic Metric for MT Evaluation with Improved Correlation with Human Judgments,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Meteor: An automatic Metric for MT Evaluation with Improved Correlation with Human Judgments,

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:11.829860Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:11.126492Z digest=sha256:cd71bd7c6372089e44a03ae59693b25cec2300a70d6d95b8ea2bb2087424f4fb

Observation daf48de2-ca23-4cf0-b01c-5c43d8937fd1 · outbound

This paper cites Adam: A Method for Stochastic Optimization.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Adam: A Method for Stochastic Optimization

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-11T14:36:11.131790Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:36:11.131790Z digest=sha256:94def19afc6c1b146e522d884ccc89fc43f30dbb3c4de17aa9620f7da002af40

Observation d136edbe-af6d-4c0e-ad07-036ae801d9f5 · outbound

This paper cites Guiding Long-Short Term Memory for Image Caption Generation.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Guiding Long-Short Term Memory for Image Caption Generation

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-11T14:36:11.137805Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:36:11.137805Z digest=sha256:92948108820f7b2d0dac7694b18db9e43e5d8f085ea63415901238cd05df4357

Observation edd66cae-c832-4555-8578-81dfcc8d6827 · outbound

This paper cites Multimodal Neural Language Models,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Multimodal Neural Language Models,

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:11.809439Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:11.143935Z digest=sha256:3a648dc6fbb385443b0de306569bc46366090211472b542ea26259a811ab15b8

Observation edec3636-92a0-4326-8948-9a33998384ae · outbound

This paper cites Stimulus -driven and concept -driven analysis for image caption generation,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Stimulus -driven and concept -driven analysis for image caption generation,

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:11.790431Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:11.300760Z digest=sha256:8743ec4a53e5ad1ae7eb26f96aaa30ea8bc138e7344cff82c45357e2edcb8498

Observation e96a5079-addf-4e18-a5f1-4d1694cab461 · outbound

This paper cites Learning joint relationship attention network for image captioning,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Learning joint relationship attention network for image captioning,

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:11.767496Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:11.308224Z digest=sha256:8b5b7d27bb198cf77e93a3a2768584580b8ac0b62c5a1606213062d7315da3d3

Observation 17f63919-fb5b-4918-8536-1997f245ab32 · outbound

This paper cites Detach and Attach: Stylized Image Captioning without Paired Stylized Dataset,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Detach and Attach: Stylized Image Captioning without Paired Stylized Dataset,

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:11.742491Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:11.318914Z digest=sha256:656b43e7d69cb9397d90a071de1866a4b70cef69f41fc47b5148409a45839551

Observation 52339b4d-69e0-4388-93a6-d780ad4cc7fc · outbound

This paper cites Learning Cooperative Neural Modules for Stylized Image Captioning,.

UnMA-CapSumT: Unified and Multi-Head Attention-driven Caption Summarization Transformer Learning Cooperative Neural Modules for Stylized Image Captioning,

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T14:36:11.723077Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-11T14:36:11.326499Z digest=sha256:1d05c938805c9cd7ebfb8e2eaa167f8a1ac3a9e5ef77a0978aa868e1253eb727

Pith citing papers

No inbound Pith citation observations are available.