Pith. sign in

Paper Citation Record · LEDGER

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation

As of 9 August 2026, this Paper Citation Record lists 43 of 43 outbound references and 1 inbound Pith citation observation for arXiv:2507.07568.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.07568 v1

Coverage vector

measured 43 of 43 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T18:40:59.075135Z

measured 44 of 44 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-10T14:09:06.139625Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-10T14:10:28.670574Z

Reference resolution

43 of 43 outbound references displayed

  • verified exact1
  • verified fuzzy31
  • unresolved11
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 78479a4d-5929-44c4-a896-406487b8c21b · outbound

This paper cites Bottom-up and top-down attention for image captioning and visual question answering.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation Bottom-up and top-down attention for image captioning and visual question answering

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T18:40:54.634276Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:40:54.634276Z digest=sha256:be5f827aeab4fedc38da9ce0d8ff327affbce631b2894e41f1098adf6f08656e

Observation 66c350e7-409e-48e6-b3ab-d8610a0fc357 · outbound

This paper cites Instance-level expert knowledge and aggregate discrimina- tive attention for radiology report generation.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation Instance-level expert knowledge and aggregate discrimina- tive attention for radiology report generation

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:41:05.610400Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T18:40:54.698556Z digest=sha256:7e25256256c6b15164086d119f5a0fe22388294f608a1159a669b6f1d511df50

Observation a3ec215e-a41d-40fe-81a3-32556a435ba6 · outbound

This paper cites Fine-Grained Image-Text Alignment in Medical Imaging Enables Explainable Cyclic Image-Report Generation.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation Fine-Grained Image-Text Alignment in Medical Imaging Enables Explainable Cyclic Image-Report Generation

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T18:40:54.786069Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:40:54.786069Z digest=sha256:15001147b9fd3632697cb1803e6479603bb1fa9eecb5ce8953f34468ef368a63

Observation d504dc60-88db-48ec-b52f-f26993d8d51c · outbound

This paper cites Generating Radiology Reports via Memory-driven Transformer.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation Generating Radiology Reports via Memory-driven Transformer

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T18:40:54.869300Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:40:54.869300Z digest=sha256:5cf21c78d12bbc70449c99048a8e8c116d37aab9a6f73f3df902c279aa39431c

Observation 9d6a7184-3743-452e-b37b-29994f9f3119 · outbound

This paper cites Cross-modal Memory Networks for Radiology Report Generation.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation Cross-modal Memory Networks for Radiology Report Generation

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T18:40:54.940751Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:40:54.940751Z digest=sha256:b2b08df743aba98c2996a08dd75836fe616b238585019545c8edb51da18d685d

Observation 52a97cc1-afeb-4dec-95d9-80a01c7c7f76 · outbound

This paper cites Meshed-memory transformer for image cap- tioning.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation Meshed-memory transformer for image cap- tioning

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T18:40:55.064495Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:40:55.064495Z digest=sha256:fa097ebe501c2cd8ff9b4eff7523f23e5989b90e70621f3be4a0b27c32cf8761

Observation e2f3b51a-843d-4293-9f9d-eec123549feb · outbound

This paper cites To- wards diverse and natural image descriptions via a condi- tional gan.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation To- wards diverse and natural image descriptions via a condi- tional gan

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:41:05.462231Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T18:40:55.148019Z digest=sha256:de2cc8816fcfcd7b1beab354f28cfeb865b254251879e2a673910b104958aae9

Observation e41aad48-9aba-4b14-8f8b-079ca4ba5034 · outbound

This paper cites Meteor 1.3: Automatic metric for reliable optimization and evaluation of machine translation systems.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation Meteor 1.3: Automatic metric for reliable optimization and evaluation of machine translation systems

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:41:05.278577Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T18:40:55.209257Z digest=sha256:6cad6ca5d73e6d095e66f391e9383d0562ee56e293ebd7b7f9d8d3267f062b7f

Observation 444720c9-34f8-4b33-ba88-03685b39b303 · outbound

This paper cites Long short-term memory.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation Long short-term memory

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:41:05.109519Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T18:40:55.323303Z digest=sha256:27e7b029cab2297a7a4ef9a08a51f4abb376660aba4124f9e525663bc047694c

Observation 69c3d97e-2961-4c1f-a040-89f0b82b9409 · outbound

This paper cites Scaling up vision-language pre-training for image captioning.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation Scaling up vision-language pre-training for image captioning

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:41:04.969520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T18:40:55.433036Z digest=sha256:0c20cb34bd7bb0185475b3985456b532fc70b6e9a95610a188e01e133ebffdad

Observation fd4f6c4e-9dfe-4597-9e56-c1c3019e36e6 · outbound

This paper cites Kiut: Knowledge-injected u-transformer for radiology report generation.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation Kiut: Knowledge-injected u-transformer for radiology report generation

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:41:04.768171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T18:40:55.534153Z digest=sha256:1c6b965db1e8fefef02280ba1cec78b25c83a5f4b9b26379abea9cb086bc5dc2

Observation b3f8e06e-977d-4568-84cd-db6808602b1d · outbound

This paper cites Chexpert: A large chest radiograph dataset with uncertainty labels and expert comparison.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation Chexpert: A large chest radiograph dataset with uncertainty labels and expert comparison

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:41:04.578931Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T18:40:55.655311Z digest=sha256:cc1617b0cb7960e923beeec12d469e34346a2d00caa3da345c8924b621f6a46d

Observation 66150774-7d3f-42a1-96e6-d9944ecf8717 · outbound

This paper cites Promptmrg: Diagnosis-driven prompts for medical report generation.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation Promptmrg: Diagnosis-driven prompts for medical report generation

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:41:04.442761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T18:40:55.760521Z digest=sha256:c72cafd590d79b70d1fc7d021d61840913693f69abc299402ff79add6fed3c3f

Observation 5b4ef1b4-7b98-4a38-993f-91d441f63684 · outbound

This paper cites Zero-shot camouflaged object detection.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation Zero-shot camouflaged object detection

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:41:04.318825Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T18:40:55.898674Z digest=sha256:b148632c60831e14e5f5d0f84db2432dda3b07b16854dfbfa0ecaccebf8351d4

Observation baffaa0e-0a8a-4797-b2e5-e7060b176690 · outbound

This paper cites Dynamic graph enhanced contrastive learning for chest x-ray report generation.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation Dynamic graph enhanced contrastive learning for chest x-ray report generation

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:41:03.750014Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T18:40:56.011195Z digest=sha256:b4233d73842144ed04143befb49d70678e62a1d70f6badad45baacac2db6719e

Observation db6feba9-9b44-4b77-87e1-a16bb8ee3b6e · outbound

This paper cites Unify, align and refine: Multi- level semantic alignment for radiology report generation.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation Unify, align and refine: Multi- level semantic alignment for radiology report generation

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:41:03.303434Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T18:40:56.141380Z digest=sha256:90992946d30d423c2c849f7f87631068abbfd609cc1aa2605bbe131be48e9b42

Observation 445c98da-8d75-4fc3-9628-def98b1f73b0 · outbound

This paper cites Rouge: A package for automatic evaluation of summaries.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation Rouge: A package for automatic evaluation of summaries

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T18:40:56.262544Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:40:56.262544Z digest=sha256:8acebad7062d0f72fc4fc6b29ea1534a1164b0eb6d22572a6fb74b800ad1246b

Observation 3e1d39db-f330-47b2-b6be-84e91f921b60 · outbound

This paper cites Exploring and distilling posterior and prior knowl- edge for radiology report generation.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation Exploring and distilling posterior and prior knowl- edge for radiology report generation

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:41:03.132064Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T18:40:56.361215Z digest=sha256:7db77d9abecbf4531546b36c549123a7d4a23d2411968b51af8507e9f15e1cf4

Observation 26596a90-efe7-46db-97ae-9bbbeb7ec035 · outbound

This paper cites Grounding dino: Marry- ing dino with grounded pre-training for open-set object de- tection, 2024.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation Grounding dino: Marry- ing dino with grounded pre-training for open-set object de- tection, 2024

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:41:02.954214Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T18:40:56.507414Z digest=sha256:1e434c34c767cca9f1da289e08757552069ef28a9da774c8a62999c7295884f9

Observation c30f39ce-c50a-4523-a79e-52880908d818 · outbound

This paper cites Knowing when to look: Adaptive attention via a visual sen- tinel for image captioning.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation Knowing when to look: Adaptive attention via a visual sen- tinel for image captioning

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:41:02.739319Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T18:40:56.639657Z digest=sha256:86dd7813c513a6e2af68c5e1d78665720726715583c82adcf05e74460e9fcd60

Observation f1115d3c-2adf-4404-beba-d14d20bf12da · outbound

This paper cites Im- proving chest x-ray report generation by leveraging warm starting.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation Im- proving chest x-ray report generation by leveraging warm starting

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:41:02.508737Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T18:40:56.748965Z digest=sha256:9df5cd5de391743f84e0bd93c7ea123b07e5f6b90165985220cdb242464b3b3f

Observation 7600f648-b485-4dec-a247-a68a9f9f67cb · outbound

This paper cites Progressive Transformer-Based Generation of Radiology Reports.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation Progressive Transformer-Based Generation of Radiology Reports

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-08-06T18:40:59.265156Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T18:40:56.851674Z digest=sha256:d0098fcf4c7b0ee40cc881ad7d467d57e5a09d076a1c7bde8dcbbc8b6c60bfd0

Observation 481197d6-1da4-4f23-a372-1101e123d7bb · outbound

This paper cites An Introduction to Convolutional Neural Networks.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation An Introduction to Convolutional Neural Networks

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T18:40:56.975156Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:40:56.975156Z digest=sha256:9210355be76c4adbc2972ba594b188ba8a4492d131fd1ca530618ffe8fe9f404

Observation 8a47ff2b-b8ab-4ff7-8152-aeba2f888650 · outbound

This paper cites X-linear attention networks for image captioning.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation X-linear attention networks for image captioning

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:41:02.280231Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T18:40:57.092731Z digest=sha256:1d7b9143964ae72829932968586756601c945895edb1404bbf50a93b94adeb92

Observation 399b2de2-7ed2-4d59-b049-64f5531a590b · outbound

This paper cites Bleu: a method for automatic evaluation of machine translation.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation Bleu: a method for automatic evaluation of machine translation

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T18:40:57.192845Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:40:57.192845Z digest=sha256:da1352813604a2e52bbff482716600760f9c974c65215709e197df6fe00d42d9

Observation 2279dcf1-cd36-4f63-90c1-8bcec903b8a5 · outbound

This paper cites Self-critical sequence training for image captioning.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation Self-critical sequence training for image captioning

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:41:02.077441Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T18:40:57.300223Z digest=sha256:c786f80a6248c9549f9478dc4adbcc94226a986871057b5b78f60b86eb56072c

Observation 6cbdd847-1624-48ba-bea7-e050a2f5ce2f · outbound

This paper cites CheXbert: Combining Automatic Labelers and Expert Annotations for Accurate Radiology Report Labeling Using BERT.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation CheXbert: Combining Automatic Labelers and Expert Annotations for Accurate Radiology Report Labeling Using BERT

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T18:40:57.387664Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:40:57.387664Z digest=sha256:9fb7bce96f2a8e920eabe269df88c721fb91d1c580224b5318390f9573e466c9

Observation bb509744-ab22-429f-9980-1e8dc68f5de1 · outbound

This paper cites Interactive and explainable region-guided radiol- ogy report generation.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation Interactive and explainable region-guided radiol- ogy report generation

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:41:01.885758Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T18:40:57.521412Z digest=sha256:fee3514b17851f9f7c8dea877f713aa8fb262057a9d61f9dfaf42d683388b596

Observation 9f787231-8d18-4ab0-b334-5b4bb60b2e60 · outbound

This paper cites Attention is all you need.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation Attention is all you need

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T18:40:57.653944Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:40:57.653944Z digest=sha256:302fc2a86614b72cefed0f99ae16898a7d11e88fb219ecbf9b07cef9a1d5ef77

Observation 726f4580-4d9c-4fb2-a79d-9b8813b37521 · outbound

This paper cites Hergen: El- evating radiology report generation with longitudinal data,.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation Hergen: El- evating radiology report generation with longitudinal data,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:41:01.679143Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T18:40:57.808493Z digest=sha256:a1f615b7bd4ca5056541837d48028efcd0946b70cf90c301d16fa816dd685526

Observation dc82586b-a8f1-4b1b-a1e6-cee1f9b0c2b3 · outbound

This paper cites Cross-modal pro- totype driven network for radiology report generation.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation Cross-modal pro- totype driven network for radiology report generation

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:41:01.372332Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T18:40:57.876827Z digest=sha256:4b6de26e1314c00e8c9ce294203be3467bc4ad0d7195b4fba81cba976a6899a5

Observation b743310e-a002-4e63-b0ea-cbe3e98c47af · outbound

This paper cites Multi-view feature fusion and visual prompt for remote sensing image captioning.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation Multi-view feature fusion and visual prompt for remote sensing image captioning

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:41:01.242470Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T18:40:57.968629Z digest=sha256:3663fe28771917ce7ae0e404d16c9c7101fa219e401a521a0480acc14c43d353

Observation 777f3ed9-0a61-4858-a7bf-c2a69703f466 · outbound

This paper cites Medclip: Contrastive learning from unpaired medical images and text, 2022.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation Medclip: Contrastive learning from unpaired medical images and text, 2022

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:41:01.050486Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T18:40:58.066074Z digest=sha256:da2f6b81437777f04997e735f06a1db3d24e28f152dbab42240fe6ac4c8f7531

Observation 81bd850c-94a1-4f1e-9f70-de01d224ba5b · outbound

This paper cites Metransformer: Radiology report generation by transformer with multiple learnable expert tokens.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation Metransformer: Radiology report generation by transformer with multiple learnable expert tokens

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:41:00.765205Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T18:40:58.159836Z digest=sha256:56214e8e1f52d884bdf7a72b8365043ce390f2ecfe0fe97d414a15d8bf35f244

Observation 00e48e2c-8dbf-4bc5-a4ec-39c6a2f85fcf · outbound

This paper cites Medklip: Medical knowledge enhanced language-image pre-training in radiology, 2023.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation Medklip: Medical knowledge enhanced language-image pre-training in radiology, 2023

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:41:00.516888Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T18:40:58.283944Z digest=sha256:b7efcdb064ccdc2760d382a2eb6fba4a2a4760992799c702fe34fc30a6ffefb0

Observation ca36f8da-a3e4-4651-88e2-b907d02e5427 · outbound

This paper cites Clinical-bert: Vision-language pre-training for radiograph diagnosis and reports generation.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation Clinical-bert: Vision-language pre-training for radiograph diagnosis and reports generation

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:41:00.236712Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T18:40:58.395579Z digest=sha256:3b34ac8e6a31c693e7acfbaf946f25d86f40ace939349c4f745af7c65c35b540

Observation a1c43eb3-35d8-4341-8c96-d4285d96a71d · outbound

This paper cites Knowledge matters: Chest radiology report genera- tion with general and specific knowledge.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation Knowledge matters: Chest radiology report genera- tion with general and specific knowledge

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:41:00.110219Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T18:40:58.484139Z digest=sha256:785690194cc4e54495fa88493e45d54f9a896a4c2bd7ccbfc81c717190ed8a66

Observation 3a4ab974-bc06-472f-a825-8bb259196c28 · outbound

This paper cites Radiology report generation with a learned knowledge base and multi-modal alignment.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation Radiology report generation with a learned knowledge base and multi-modal alignment

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:40:59.972046Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T18:40:58.584779Z digest=sha256:6f6bb26e208156ced73f77f14182914c36379f493d5ae2f11626a47fa873f5bf

Observation fc0a99fe-33e2-4751-b791-a7d51508a97f · outbound

This paper cites Improving hyperbolic representations via gromov- wasserstein regularization.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation Improving hyperbolic representations via gromov- wasserstein regularization

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:40:59.865065Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T18:40:58.684853Z digest=sha256:446cd072a6a224baa6e7df126859be9bc8a77d79ee106986bf7aee04c73712f3

Observation 61a4302a-1b9b-4391-b4d0-5c6efedcccd2 · outbound

This paper cites Otseg: Multi-prompt sinkhorn attention for zero-shot semantic segmentation.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation Otseg: Multi-prompt sinkhorn attention for zero-shot semantic segmentation

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:40:59.719371Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T18:40:58.767838Z digest=sha256:b73f20850d6de41d01c231026ede424b162b05b86cda0cefbf0b9b9d65b66947

Observation 6f210f77-71be-4632-afa1-6bd23d153e4a · outbound

This paper cites CoCa: Contrastive Captioners are Image-Text Foundation Models.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation CoCa: Contrastive Captioners are Image-Text Foundation Models

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-06T18:40:58.884213Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:40:58.884213Z digest=sha256:7be11640a967595e87b2dbbb636531c1311a2a9d01732d5ba1a1c9b643c2500c

Observation 59a90ce7-959b-4231-9389-27b178fc4b20 · outbound

This paper cites Anatomy-guided weakly- supervised abnormality localization in chest x-rays.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation Anatomy-guided weakly- supervised abnormality localization in chest x-rays

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:40:59.576735Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T18:40:58.967913Z digest=sha256:03d493b4a53a78a713d2d57776ae9422b4c7a12ee2dfd4d46ceb24412a932662

Observation 21525339-621f-46bf-876d-c7d8cecf6915 · outbound

This paper cites Sam-guided enhanced fine-grained encoding with mixed semantic learning for medical image captioning.

Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation Sam-guided enhanced fine-grained encoding with mixed semantic learning for medical image captioning

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:40:59.428509Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T18:40:59.075135Z digest=sha256:734f6725ac1a127aae2a0e0de1e49c857907ae810c8c96eec053b6b33c269fbe

Pith citing papers

Observation 5c7c268a-d65a-4457-a4fa-2afbffd5e22a · inbound

Enhancing Reinforcement Learning for Radiology Report Generation with Evidence-aware Rewards and Self-correcting Preference Learning cites this paper.

Enhancing Reinforcement Learning for Radiology Report Generation with Evidence-aware Rewards and Self-correcting Preference Learning Learnable Retrieval Enhanced Visual-Text Alignment and Fusion for Radiology Report Generation

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-10T14:10:28.672612Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-10T14:09:06.139625Z digest=sha256:0fc967085fb1fdc5e7364ca24c184097e64029629e0bea1d797ea9a9b0c1a4ef