Pith. sign in

Paper Citation Record · LEDGER

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding

As of 18 August 2026, this Paper Citation Record lists 52 of 52 outbound references and 0 inbound Pith citation observations for arXiv:2506.09634.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.09634 v1

Coverage vector

measured 52 of 52 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T04:47:08.060954Z

measured 52 of 52 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

52 of 52 outbound references displayed

  • verified exact1
  • verified fuzzy24
  • unresolved26
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 59d1df71-1389-4d58-8a2c-ee09b32c4a33 · outbound

This paper cites Phi-4-Mini Technical Report: Compact yet Powerful Multimodal Language Models via Mixture-of-LoRAs.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding Phi-4-Mini Technical Report: Compact yet Powerful Multimodal Language Models via Mixture-of-LoRAs

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:07.620921Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:47:07.620921Z digest=sha256:ae7a1d6ca1373d1ea9ce020e8cbc55a9369932fa90d1993dd7556f4440134502

Observation 656efc58-2110-4998-a288-a64b06889b1a · outbound

This paper cites DCFormer: Efficient 3D Vision-Language Modeling with Decomposed Convolutions.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding DCFormer: Efficient 3D Vision-Language Modeling with Decomposed Convolutions

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:07.624830Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:47:07.624830Z digest=sha256:db630c1aa7f552f820408d562959846374e8a613f109de8ba20c51cee3b3a64f

Observation 7a954a7c-d5a8-4442-ba12-c67648decd15 · outbound

This paper cites M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding M3D: Advancing 3D Medical Image Analysis with Multi-Modal Large Language Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:07.628602Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:47:07.628602Z digest=sha256:949bd2bf8df410035c19e3121a53a17930bb50a8be09655fba13ffd91176bb4e

Observation 1c5b0f29-76a7-4997-9239-afd5e0597e8c · outbound

This paper cites Shah, Andrew Johnston, Robert D.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding Shah, Andrew Johnston, Robert D

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:07.631850Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:47:07.631850Z digest=sha256:da2fc5033bd351a26d8f995529c6e4c0181da3fbf80edb98867516958c9b1fc1

Observation ddf966f4-e600-4c6d-aa4b-73395c2e83df · outbound

This paper cites Bruno, Eric A.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding Bruno, Eric A

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:47:14.534106Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T04:47:07.634653Z digest=sha256:9108fbcd215d378fd379ab84bfab1329a45fdc0f1bbfa5e06678204a406a25b1

Observation 4b26adec-b0b4-4955-ab61-8013d1320908 · outbound

This paper cites 3D-CT-GPT: Generating 3D Radiology Reports through Integration of Large Vision-Language Models.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding 3D-CT-GPT: Generating 3D Radiology Reports through Integration of Large Vision-Language Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:07.637441Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:47:07.637441Z digest=sha256:62d37c0c7d0270f0bfdf1663a7b37ccdeb6114b02c7d48dc6f13c9dd3b065afa

Observation 48dcebf1-4bf3-43a9-822f-079ce5e67f00 · outbound

This paper cites HuatuoGPT-Vision, Towards Injecting Medical Visual Knowledge into Multimodal LLMs at Scale.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding HuatuoGPT-Vision, Towards Injecting Medical Visual Knowledge into Multimodal LLMs at Scale

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:07.641054Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:47:07.641054Z digest=sha256:7390cc64e16d269dfcaf8d1e36e3d9893a7d54618e65411e3acd77a4d79353c8

Observation 95305907-eca2-414a-9d4f-20ead735841f · outbound

This paper cites Dragonfly: Multi-Resolution Zoom-In Encoding Enhances Vision-Language Models.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding Dragonfly: Multi-Resolution Zoom-In Encoding Enhances Vision-Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:07.644114Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:47:07.644114Z digest=sha256:ee037981c7d2e2820d3b0f7372aa7b1238a03517861dc089e1bbd714d018fb12

Observation a76d515b-13ab-4ff0-8da7-1aa827f86195 · outbound

This paper cites MedBLIP: Bootstrapping Language-Image Pre-training from 3D Medical Images and Texts.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding MedBLIP: Bootstrapping Language-Image Pre-training from 3D Medical Images and Texts

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:07.647075Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:47:07.647075Z digest=sha256:928a8683fbf0d6a76b0d9591a667040243664e1e0badcfa64ccab3c4c72c1442

Observation 37b70fec-5257-4c93-851d-29b00c830b61 · outbound

This paper cites BIMCV-R: A Landmark Dataset for 3D CT Text-Image Retrieval.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding BIMCV-R: A Landmark Dataset for 3D CT Text-Image Retrieval

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:47:14.321988Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T04:47:07.650006Z digest=sha256:72263caaeb5a8932d3ee410544723b59de090cfb05ca63a9e2e1d1be665df82e

Observation 3a2baa62-f388-436a-95e3-6391794fb94d · outbound

This paper cites Dia-LLaMA: Towards Large Language Model-driven CT Report Generation.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding Dia-LLaMA: Towards Large Language Model-driven CT Report Generation

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:07.652832Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:47:07.652832Z digest=sha256:bdece1aa144a993742d0cf5cf0b320ccc5f954c5a7f0dc3deb4927e45e324da1

Observation 594844b1-255a-40dc-ad4d-807471d347b9 · outbound

This paper cites BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:47:14.100340Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T04:47:07.655589Z digest=sha256:79ea2418871139ca572a73457238a1bbcbda218b9b64b20fec9d3390dceefd6c

Observation 124937a7-342c-49b7-8b75-0f59990804b1 · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:47:13.968497Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T04:47:07.658248Z digest=sha256:905bc2905d25e41cba1de7412a609cb715dd4afeff6e62a32edc07dd2f91fc8f

Observation 2b779508-9103-42b8-a0cb-631d79a054b6 · outbound

This paper cites Developing Generalist Foundation Models from a Multimodal Dataset for 3D Computed Tomography.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding Developing Generalist Foundation Models from a Multimodal Dataset for 3D Computed Tomography

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:07.660922Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:47:07.660922Z digest=sha256:f713a1a10648a0e810c9c586ccaee1953e562b873971a42ce17ce9359f10e234

Observation 4feccbb4-ac79-4e5d-ad6f-7d192f85fe18 · outbound

This paper cites LoRA: Low-Rank Adaptation of Large Language Models.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding LoRA: Low-Rank Adaptation of Large Language Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:07.663591Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:47:07.663591Z digest=sha256:28d0f726db21ada1abc9d80f2b77114a52a59429089138d92108cf96369e7d0d

Observation f5fff607-8b6c-4f1d-923e-6379ef9deebd · outbound

This paper cites Ball, Norah Borus, Andrew Huang, Bhavik N.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding Ball, Norah Borus, Andrew Huang, Bhavik N

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:47:13.708650Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T04:47:07.666325Z digest=sha256:ff363a5c13b3eb65392c48aca5792944b0044ae8cce6d4911b42a6677a4e8d77

Observation b356897e-0e06-49ab-90ae-43215cb14aa4 · outbound

This paper cites Lungren, und Serena Yeung.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding Lungren, und Serena Yeung

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:47:13.488370Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T04:47:07.669216Z digest=sha256:ab7a9f7c52f4343b8100961d742a97823df1ec60263ab262b1a90a7b608750a1

Observation 556ea59f-01cf-47ae-9322-af38c4d2464b · outbound

This paper cites an unresolved cited work.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:47:13.321335Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T04:47:07.672081Z digest=sha256:879950c45831bf499dc279dd95f2d6ca4cb5fbd0313eb377e7cf23c79d2303db

Observation 79a40d5c-1666-4444-a0c4-39568d4a1496 · outbound

This paper cites MDAgents: An Adaptive Collaboration of LLMs for Medical Decision-Making.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding MDAgents: An Adaptive Collaboration of LLMs for Medical Decision-Making

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:47:13.159421Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T04:47:07.675122Z digest=sha256:abd91e466a7941c9d96a445be5da83d28ae6bb2ae855403cd5f0dcc43191914a

Observation 3cbc623f-fd10-44ae-abf3-1f67173aa724 · outbound

This paper cites E3D-GPT: Enhanced 3D Visual Foundation for Medical Vision-Language Model.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding E3D-GPT: Enhanced 3D Visual Foundation for Medical Vision-Language Model

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:07.677788Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:47:07.677788Z digest=sha256:ef0a0a89825d5bd448d396143b78d397ce7d2393afbade6578ac567b7768e174

Observation f398a14c-3cd1-4287-a0f3-2c58b6662927 · outbound

This paper cites Kevin Zhou.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding Kevin Zhou

Reference 21

Resolution
verified exact
raw_fallback, observed 2026-08-07T04:47:08.567433Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T04:47:07.680780Z digest=sha256:b5fb7487ea7d1f8da1b83fd21b6f5cd91b51b6e0c4f037bf45e5eeec715e8f73

Observation a5794760-a898-4bd0-beee-2b7d456c3087 · outbound

This paper cites METEOR: An Automatic Metric for MT Evaluation with High Levels of Correlation with Human Judgments.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding METEOR: An Automatic Metric for MT Evaluation with High Levels of Correlation with Human Judgments

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:47:12.937482Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T04:47:07.683384Z digest=sha256:b7d54759af96e19b79034b75b7956de4e7b9c0b777497d6e017ce2e9d02aaf4f

Observation e444ef61-6a4c-40c4-b933-5116bbda84a8 · outbound

This paper cites Towards a holistic framework for multimodal LLM in 3D brain CT radiology report generation.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding Towards a holistic framework for multimodal LLM in 3D brain CT radiology report generation

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:47:12.755672Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T04:47:07.686018Z digest=sha256:6c91c250ea9422b9103d029a935603e9f46097b7dc3aeab25ac445a3c233c2a4

Observation 7e10a83b-decf-4894-b774-0bef881cfbcb · outbound

This paper cites LLaV A-Med: Training a Large Language-and- Vision Assistant for Biomedicine in One Day.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding LLaV A-Med: Training a Large Language-and- Vision Assistant for Biomedicine in One Day

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:47:12.584478Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T04:47:07.688760Z digest=sha256:410e2f870845de8b3ae1cd2377b40dfcdee6173f13b4bc3d2396df479184ac8d

Observation ff9013ea-85a7-4230-833d-a2291a85dc49 · outbound

This paper cites an unresolved cited work.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:47:12.340999Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T04:47:07.691469Z digest=sha256:ee0b4ca036b7aa50b9c4c8d4b590a51fa3c06713711778425d9cd23a2e9178fa

Observation 1e1d1347-ba9f-44c9-99d4-540f68fcd32a · outbound

This paper cites Dynamic Graph Enhanced Contrastive Learning for Chest X-ray Report Generation.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding Dynamic Graph Enhanced Contrastive Learning for Chest X-ray Report Generation

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:47:12.149866Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T04:47:07.694107Z digest=sha256:b7d3cb00bcedf53c54787385b655fd321a3dd3d57493ef600e80ae29c910b92b

Observation 4adaa931-5207-495e-bbd1-7f0ca55cb3db · outbound

This paper cites TokenPacker: Efficient Visual Projector for Multimodal LLM.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding TokenPacker: Efficient Visual Projector for Multimodal LLM

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:07.697677Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:47:07.697677Z digest=sha256:8a844ea9b1fb3e49d48e348e2fb18e9d7d63c9d2d14226e93c33cf9de67d5d4a

Observation d752036f-f3da-41d7-bf24-79bda8a8573c · outbound

This paper cites Mini-Gemini: Mining the Potential of Multi-modality Vision Language Models.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding Mini-Gemini: Mining the Potential of Multi-modality Vision Language Models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:07.700642Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:47:07.700642Z digest=sha256:5fe145ebdd5c2d1f9168796d34cfee414a2d86d80a99ba991320cd5c0f766146

Observation 3f76dae1-5a53-4bd6-b0d0-5d850f82c64a · outbound

This paper cites Macro-and micro- anatomical, histological and computed tomography scan characterization of the nasopalatine canal.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding Macro-and micro- anatomical, histological and computed tomography scan characterization of the nasopalatine canal

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:47:12.020007Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T04:47:07.703490Z digest=sha256:01e3bd3853c0b6f40991f0284d4c50c4145cefa15067cbd8e6ee4bbb668cb4c0

Observation 910fd7f9-d15c-495d-9346-30f4a4f64400 · outbound

This paper cites Rouge: A package for automatic evaluation of summaries.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding Rouge: A package for automatic evaluation of summaries

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:47:11.842518Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T04:47:07.706327Z digest=sha256:7f0fd519552d2b4eb3263d49744d444f750485748e59ee4734b0ec2f48bb83df

Observation 40c0597a-cc1a-444e-999d-9166c23816f1 · outbound

This paper cites MG-3D: Multi-Grained Knowledge-Enhanced 3D Medical Vision-Language Pre-training.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding MG-3D: Multi-Grained Knowledge-Enhanced 3D Medical Vision-Language Pre-training

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:07.708862Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:47:07.708862Z digest=sha256:722ff1a9c5f6067c32f6824dbe450a188b867fe355a4ea3d93467eaafb7cf4bb

Observation 2a42e4d0-3fa2-499c-a749-a9664e804bd4 · outbound

This paper cites Bleu: a Method for Auto- matic Evaluation of Machine Translation.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding Bleu: a Method for Auto- matic Evaluation of Machine Translation

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:47:11.676720Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T04:47:07.711866Z digest=sha256:dd2c9c0dbc3a69c98753dc02f64f487bc15661368ffbe29d2ec733320f55b508

Observation d16d19a9-7886-4572-a920-c0f099914bdc · outbound

This paper cites Learning Transferable Visual Models From Natural Language Supervision.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding Learning Transferable Visual Models From Natural Language Supervision

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:47:11.474165Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T04:47:07.714448Z digest=sha256:4c420fa5c2d03364dfe573198b2c75310c8ae2cd66d39c349c9bb7f1048881aa

Observation e89def40-1a80-42a2-b779-0d5302249e3b · outbound

This paper cites Salvolini, E.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding Salvolini, E

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:47:11.349751Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T04:47:07.717125Z digest=sha256:a680638eb1885411fd6595975ecb801fda214a9f8727a0f1b7db51e3b94c7d99

Observation d7c55625-fda1-40c0-b6bb-b382ad468434 · outbound

This paper cites Time Is Money: Considerations for Measuring the Radiological Reading Time.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding Time Is Money: Considerations for Measuring the Radiological Reading Time

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:47:11.173127Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T04:47:07.719912Z digest=sha256:21b13ec717595267c762b48a98c4267e82cda98364d3dfa8c7d20145fa4d2470

Observation 5b7dceb2-6ae9-4f9d-9767-12fea05b9215 · outbound

This paper cites Med-2E3: A 2D-Enhanced 3D Medical Multimodal Large Language Model.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding Med-2E3: A 2D-Enhanced 3D Medical Multimodal Large Language Model

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:07.722774Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:47:07.722774Z digest=sha256:35793f95fdd91f9d67183c542eb5059828ac829589bcfb042b30b2f6f3099e9c

Observation 7480eb91-463e-4c1b-b889-80e3340626b9 · outbound

This paper cites an unresolved cited work.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:47:11.001306Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T04:47:07.725473Z digest=sha256:05783af51fe54de6aaa096443212ac52b9436ababb97e95e0330236c26b48207

Observation 7e9939ad-ea6f-4414-a99a-344e21c0a209 · outbound

This paper cites an unresolved cited work.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:47:10.861999Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T04:47:07.728638Z digest=sha256:76a7160b7c9787dd5aa428366fe34c0f3320fb1fa68a31b7d8a4c521eef6d712

Observation d3d44794-c411-4187-8c28-c7d1cc90b1b5 · outbound

This paper cites Towards Generalist Biomedical AI.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding Towards Generalist Biomedical AI

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:07.731105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:47:07.731105Z digest=sha256:fc9b9509c103bd84cf5174d24dc02e20e22c36f5a149725246130491fd41abe3

Observation 1f57a25c-8f75-4339-8799-4a0e942798c7 · outbound

This paper cites Representation Learning with Contrastive Predictive Coding.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding Representation Learning with Contrastive Predictive Coding

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:07.733934Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:47:07.733934Z digest=sha256:c5230f999d8cc597075f4e8233b63210ec999787ec4a2d67f2261f03b230404e

Observation fa98463d-1155-416d-867b-0f2e6da219b4 · outbound

This paper cites Cross-modal prototype driven network for radiology report generation.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding Cross-modal prototype driven network for radiology report generation

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:47:10.670325Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T04:47:07.736992Z digest=sha256:6d8a7e99bc10bda628b0bda76fee7ca8e4e59abc49b445935fbdd87183b23170

Observation fd024453-c5ee-4cd6-b23a-8214a5f63df9 · outbound

This paper cites MMCLIP: Cross-modal Attention Masked Modelling for Medical Language-Image Pre-Training.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding MMCLIP: Cross-modal Attention Masked Modelling for Medical Language-Image Pre-Training

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:07.741096Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:47:07.741096Z digest=sha256:09c04ae603dd937b37c6bf2262bb3cfe06f37c06f1e5073ddb7a1d12c5fb052f

Observation 98a3ec93-657c-4a92-bf45-7c1b99b0c292 · outbound

This paper cites Towards Generalist Foundation Model for Radiology by Leveraging Web-scale 2D&3D Medical Data.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding Towards Generalist Foundation Model for Radiology by Leveraging Web-scale 2D&3D Medical Data

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:07.745722Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:47:07.745722Z digest=sha256:1cfadadbedb36f52d57cf073fb5c4da5910033ead0710bdb83cc9f47440c5443

Observation 883a7492-0f61-419c-868a-29c40e50901f · outbound

This paper cites Zou, und Huaxiu Yao.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding Zou, und Huaxiu Yao

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:47:10.399927Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T04:47:07.752474Z digest=sha256:59d91436638bc0f55b96acd15b7771125c21850dfd2f4426d2deddc27ec57b11

Observation 0d50dc99-07f8-4d3a-a0b8-dbf602292b1d · outbound

This paper cites Med3DVLM: An Efficient Vision- Language Model for 3D Medical Image Analysis.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding Med3DVLM: An Efficient Vision- Language Model for 3D Medical Image Analysis

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:07.756129Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:47:07.756129Z digest=sha256:a729193dbc43945f7b00a65705b9f761b2bb1f86f6142f86c8a6f0564e177d98

Observation f3bc874f-1ec7-4a1a-a361-98eefc53a0bb · outbound

This paper cites Sigmoid Loss for Language Image Pre-Training.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding Sigmoid Loss for Language Image Pre-Training

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:47:10.209432Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T04:47:07.760736Z digest=sha256:44cd912eaa4fc939b991f10b9ff3e3a81f507669beb2c2aa9325be354ed533fc

Observation 3eb5ce39-5de0-4804-8dca-ccda4f254aa2 · outbound

This paper cites Lungren, Tristan Naumann, Sheng Wang, und Hoifung Poon.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding Lungren, Tristan Naumann, Sheng Wang, und Hoifung Poon

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:47:09.979845Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T04:47:07.770261Z digest=sha256:73e6e5e7c1fa5a13db861009fec7ff8cba21f6ae8b145869367efff87ba7c195

Observation a8424c63-97ce-46e9-8a0a-4604076e458d · outbound

This paper cites Weinberger, und Yoav Artzi.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding Weinberger, und Yoav Artzi

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:47:09.799164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T04:47:07.784587Z digest=sha256:6058e1146fd4280cb7fedfb548cbc3af1f7f48017eafb50e5ba7dea0a8eb5ac1

Observation 12f782b5-3cf3-4c2f-9a47-c9d7f591e4ac · outbound

This paper cites MEPNet: Medical Entity-Balanced Prompting Network for Brain CT Report Generation.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding MEPNet: Medical Entity-Balanced Prompting Network for Brain CT Report Generation

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:47:09.557804Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T04:47:07.805586Z digest=sha256:c005647e83787756748c76cbb3f40f2a70ae4eeddc7e0131be84c827ba2a0f98

Observation 3f99ebed-d651-47b5-87b0-1f2f8c00f6b6 · outbound

This paper cites RadGenome-Chest CT: A Grounded Vision-Language Dataset for Chest CT Analysis.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding RadGenome-Chest CT: A Grounded Vision-Language Dataset for Chest CT Analysis

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T04:47:07.828501Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:47:07.828501Z digest=sha256:5bcdf984bbfa66870a484b7f1862b678246db28375d0ee4d40f34e2640a9700a

Observation dd49e689-3feb-418a-a7b8-e50e15fce348 · outbound

This paper cites There are emphysematous changes in both lungs.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding There are emphysematous changes in both lungs

Reference 51

Resolution
malformed identifier
raw_fallback, observed 2026-08-07T04:47:09.276205Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T04:47:07.941763Z digest=sha256:f0a8855d1b7e6d888ea2ca94870117d10c6945bc35d88cedcf1c229a6d038402

Observation 16f3d46d-e57f-4dec-a54b-23b8802df80d · outbound

This paper cites Guidelines: • The answer NA means that the paper does not involve crowdsourcing nor research with human subjects.

HSENet: Hybrid Spatial Encoding Network for 3D Medical Vision-Language Understanding Guidelines: • The answer NA means that the paper does not involve crowdsourcing nor research with human subjects

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:47:08.987645Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T04:47:08.060954Z digest=sha256:c2474c3beb52c5a1a6b88ba6140c804c656ca97c32365675bc5d9d8adb945397

Pith citing papers

No inbound Pith citation observations are available.