Pith. sign in

Paper Citation Record · LEDGER

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images

As of 18 August 2026, this Paper Citation Record lists 49 of 49 outbound references and 0 inbound Pith citation observations for arXiv:2608.03322.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.03322 v1

Coverage vector

measured 49 of 49 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T21:07:09.893269Z

measured 49 of 49 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

49 of 49 outbound references displayed

  • verified exact2
  • verified fuzzy23
  • unresolved24
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 8139d1e2-3c3a-4309-a652-b50b58240e49 · outbound

This paper cites International conference on medical image computing and computer-assisted intervention , pages=.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images International conference on medical image computing and computer-assisted intervention , pages=

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:07:14.440464Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T21:07:07.291303Z digest=sha256:cc411b9b34e3c95baf7acdb50b9e4c3e304e01ce8e511d46aff8768eef0ddd8e

Observation 798da038-d359-4dd3-bc2b-29e39aaafa7f · outbound

This paper cites International Conference on Medical Image Computing and Computer-Assisted Intervention , pages=.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images International Conference on Medical Image Computing and Computer-Assisted Intervention , pages=

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:07:14.235802Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T21:07:07.323838Z digest=sha256:3fda968ff622188ca8d22f67b7e034cad21f5a123834253767db02f3ffe5f008

Observation 34ec7a43-e77e-4a12-864b-b65bc4ffd6fd · outbound

This paper cites IEEE Transactions on Pattern Analysis and Machine Intelligence , year=.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images IEEE Transactions on Pattern Analysis and Machine Intelligence , year=

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:07:14.039872Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T21:07:07.354778Z digest=sha256:c8844f981699b0813711e5737e2e5db977801e5c6d11220728cdeaef2c931766

Observation f3ceb8d4-476c-4986-b790-39f15fb39ec7 · outbound

This paper cites LocateAnything: Fast and High-Quality Vision-Language Grounding with Parallel Box Decoding.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images LocateAnything: Fast and High-Quality Vision-Language Grounding with Parallel Box Decoding

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-05T21:07:07.387536Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:07:07.387536Z digest=sha256:3f568ee2b637737207e05e6cc4aa1245a7837ac367c5faf6b3ef7266b3e04640

Observation 6ec4edd4-a18a-4843-b5a9-59152b7a7b83 · outbound

This paper cites Nature communications , volume=.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images Nature communications , volume=

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:07:13.872592Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T21:07:07.433529Z digest=sha256:92ab752cdb0a7ff891898c335ca09d6196b9a651c483810d811c715fd0ee0e39

Observation 4f2736d8-74e2-4d5b-8d6c-91256a884cde · outbound

This paper cites Nature methods , volume=.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images Nature methods , volume=

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:07:13.738062Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T21:07:07.468330Z digest=sha256:c72b6c7920accce3cfda6bf3f5c2dfdd9608eb8e6d18ed32b0d50697eeefc79f

Observation 59b2ecb1-120e-491d-b221-857adb44898f · outbound

This paper cites international conference on medical image computing and computer-assisted intervention , pages=.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images international conference on medical image computing and computer-assisted intervention , pages=

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:07:13.575409Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T21:07:07.508242Z digest=sha256:c321f7634c941a97aa08462d1b2cd71bbdec685d680f41916f8a5deb06077647

Observation 57e88c4d-9892-4f82-997b-395e7c3a7058 · outbound

This paper cites MedMoE: Modality-Specialized Mixture of Experts for Medical Vision-Language Understanding.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images MedMoE: Modality-Specialized Mixture of Experts for Medical Vision-Language Understanding

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-05T21:07:07.540354Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:07:07.540354Z digest=sha256:6f92e28a74409598dd113a5e63689a10602145554303ebe1aee04df4ec9bb79e

Observation f388ae50-5da8-4c7b-828c-16459a5034e2 · outbound

This paper cites Rex-Thinker: Grounded Object Referring via Chain-of-Thought Reasoning.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images Rex-Thinker: Grounded Object Referring via Chain-of-Thought Reasoning

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-05T21:07:07.567191Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:07:07.567191Z digest=sha256:9c84112d6a8de4c2d81958226bfb0d90dafa9aa84c4fd74f5d03db5315f0fcd8

Observation a2fcd220-4af4-4ced-9811-a1969e258422 · outbound

This paper cites arXiv preprint arXiv:2510.04477 , year=.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images arXiv preprint arXiv:2510.04477 , year=

Reference 10

Resolution
verified exact
raw_fallback, observed 2026-08-05T21:07:10.819578Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T21:07:07.598807Z digest=sha256:f9ee218dc8dd6f318e9d0c917f34dafd25e549da57c3936c49a15739471fdcc3

Observation 2f6eb89d-c272-4b9b-87fc-d06fe82f5a8f · outbound

This paper cites Journal of medical imaging , volume=.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images Journal of medical imaging , volume=

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:07:13.464628Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T21:07:07.657756Z digest=sha256:3295222763a907ce5dcaf241c1aa8784fe00d511ec016f4731e9dd7303c56b6d

Observation 544d608b-642c-49db-8be5-36ac71182503 · outbound

This paper cites Proceedings of the IEEE/CVF international conference on computer vision , pages=.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images Proceedings of the IEEE/CVF international conference on computer vision , pages=

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T21:07:07.693581Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:07:07.693581Z digest=sha256:9fa6efae6d9947360ca517f373c0e81c49d037a2ceac4e5618a3ba7cdfdfbe0b

Observation 7f3499f1-877e-4ae3-bcf6-7a721c8371a0 · outbound

This paper cites Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:07:13.334089Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T21:07:07.735063Z digest=sha256:5e41bd16e4c47df99d638cd31b7cb8e001d4f3b125a5b84d4b3a13727b176526

Observation 2bd2698e-bf62-4d58-94e7-124c85a2665d · outbound

This paper cites European conference on computer vision , pages=.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images European conference on computer vision , pages=

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:07:13.172153Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T21:07:07.778837Z digest=sha256:2326cf5422007a4a9e9e11f959bb3eed200997160dc2c3fc9a482e3190c6c53d

Observation 6a4eae3b-817b-4988-a3fd-315e68536dc9 · outbound

This paper cites European conference on computer vision , pages=.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images European conference on computer vision , pages=

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-05T21:07:07.802293Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:07:07.802293Z digest=sha256:e633a21d124aeead46532793c8855620308ab13b7a79f50921190d78c741b58b

Observation c8c46346-8a68-4dcd-8c56-1a320ba32416 · outbound

This paper cites International Conference on Learning Representations , volume=.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images International Conference on Learning Representations , volume=

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-05T21:07:07.838697Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:07:07.838697Z digest=sha256:daf956249a6e7fd94d05a205cdff621b62d61d64dc520be7ad9cd984b10f8e05

Observation a378364e-1e5e-42da-acba-05ca18c631c6 · outbound

This paper cites arXiv preprint arXiv:2510.12798 , year=.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images arXiv preprint arXiv:2510.12798 , year=

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-05T21:07:07.882591Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:07:07.882591Z digest=sha256:18e3e5b8e7d23fbc788a63eed9cde944545fa945fce9b325e1927c96a5517f38

Observation 085a2332-2867-4624-a043-da1bf3de5719 · outbound

This paper cites 2026 IEEE/CVF Winter Conference on Applications of Computer Vision (WACV) , pages=.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images 2026 IEEE/CVF Winter Conference on Applications of Computer Vision (WACV) , pages=

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:07:12.986260Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T21:07:07.931795Z digest=sha256:42845f6a7140d236bc49c4788178c0c56666993426b51978e632e52f6f225a14

Observation 110f6385-b605-4c39-94a6-bc57dde51b2b · outbound

This paper cites VividMed: Vision Language Model with Versatile Visual Grounding for Medicine.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images VividMed: Vision Language Model with Versatile Visual Grounding for Medicine

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-05T21:07:08.007254Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:07:08.007254Z digest=sha256:a3cc6f9601b65ea0e77a1498719e4c84ca2b678becf4fdc3007e33b3201d9162

Observation 49ad6160-661a-4c88-b080-3bc11726618b · outbound

This paper cites Proceedings of the Computer Vision and Pattern Recognition Conference , pages=.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images Proceedings of the Computer Vision and Pattern Recognition Conference , pages=

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:07:12.882338Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T21:07:08.069248Z digest=sha256:483546b58afa9c25fa9f4577cadd71385999315239ea24fe8ec2be4888fc46f1

Observation 92b64d50-bfb0-4007-8452-c2d84cb062ac · outbound

This paper cites Proceedings of the AAAI Conference on Artificial Intelligence , volume=.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images Proceedings of the AAAI Conference on Artificial Intelligence , volume=

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:07:12.776827Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T21:07:08.158591Z digest=sha256:781568e843a69c27979911e8456d117b0a0b5944498b8e9060fd1c2f595e4814

Observation 539ca9fe-d91c-4f3f-8426-223cdba068c4 · outbound

This paper cites Nature Communications , year=.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images Nature Communications , year=

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:07:12.672341Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T21:07:08.218339Z digest=sha256:08a8f39345fb9b181d8148b44b8b4392e21a21540afe2314df62ed12d645ed6f

Observation 3652678f-b974-46c2-a861-824a124d0c48 · outbound

This paper cites European conference on computer vision , pages=.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images European conference on computer vision , pages=

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-05T21:07:08.287875Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:07:08.287875Z digest=sha256:877eaf9e49114ced3d306499be086488c3691ab779a35727daed03807afc9576

Observation d0f5a677-cf76-4f15-b909-147757ebb773 · outbound

This paper cites arXiv preprint arXiv:2601.06847 , year=.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images arXiv preprint arXiv:2601.06847 , year=

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-05T21:07:08.385821Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:07:08.385821Z digest=sha256:dee00b3d646b71217831cac3d6a96d9f9b571946c8121299ab18ac88ad213816

Observation bfc923ed-8082-4118-8eb6-30463cbad083 · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images Advances in Neural Information Processing Systems , volume=

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:07:12.461937Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T21:07:08.421805Z digest=sha256:5c8bb0ec4bc6fa0df03567193bca662c59189760da9539d3d03e3f01e8fa0caa

Observation c88710ae-06f9-4c58-bc6f-b9a7e8c8a542 · outbound

This paper cites UniVG-R1: Reasoning Guided Universal Visual Grounding with Reinforcement Learning.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images UniVG-R1: Reasoning Guided Universal Visual Grounding with Reinforcement Learning

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-05T21:07:08.465735Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:07:08.465735Z digest=sha256:9c9a279940bee665c7ca3f6fe54097529b7c74579ebdfca05cf7aa774349b4bb

Observation 0021fea5-eb31-41b4-b5db-4ee7239a1c2d · outbound

This paper cites Proceedings of the IEEE/CVF International Conference on Computer Vision , pages=.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images Proceedings of the IEEE/CVF International Conference on Computer Vision , pages=

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:07:12.310068Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T21:07:08.521244Z digest=sha256:cc9c8620f6e2b6121b1f7eef397d991f388383a24f27ec1fcee9d094e3b3ec29

Observation 7baefc8b-2614-44e4-84bb-7fa50f48fa78 · outbound

This paper cites Better Eyes, Better Thoughts: Why Vision Chain-of-Thought Fails in Medicine.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images Better Eyes, Better Thoughts: Why Vision Chain-of-Thought Fails in Medicine

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-05T21:07:08.563485Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:07:08.563485Z digest=sha256:7c4636697d109821436d1590c2880904cf2bb22bd13885467ca72269844c0ec8

Observation 0cdb2e9e-7bca-4770-9a48-2ba6af965c1b · outbound

This paper cites Nature communications , volume=.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images Nature communications , volume=

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-05T21:07:08.634558Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:07:08.634558Z digest=sha256:4ecb219868e2205d81678386a40eeb1e9449c86a22e431b5fcaee9e947d0f207

Observation 98196b98-4ce3-4d47-94f3-30ac11d264a5 · outbound

This paper cites Radiology: Artificial Intelligence , volume=.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images Radiology: Artificial Intelligence , volume=

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-05T21:07:08.695541Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:07:08.695541Z digest=sha256:2cd2245b10bb31533d78fbd804cedda4a8cad3cbcae14bf6cc43a99b0a3a2ce1

Observation b9446222-422b-47a8-bf0e-47cd25906429 · outbound

This paper cites Medical image analysis , volume=.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images Medical image analysis , volume=

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-05T21:07:08.748748Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:07:08.748748Z digest=sha256:1f495aa682de0c498281057a5d425a5022d19cb3605ff092d7ff7bbd49d0b26a

Observation e28eec6a-38e8-4e3f-b8bf-f709c7521296 · outbound

This paper cites International Conference on Medical Image Computing and Computer-Assisted Intervention , pages=.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images International Conference on Medical Image Computing and Computer-Assisted Intervention , pages=

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:07:12.099143Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T21:07:08.786746Z digest=sha256:bf7c369ee7f0135ecd2493c0c33824ffea7fbc12008853b0c0fe12d3587dba25

Observation 3728f14f-26b9-4c6b-b8b5-5b7838536633 · outbound

This paper cites Medical physics , volume=.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images Medical physics , volume=

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:07:11.934478Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T21:07:08.815007Z digest=sha256:6d79b7de38d97c7a505aa7309c43eb259291582eeabaca69e0a1df62d7146de4

Observation 886c2177-2804-4377-a1c5-8790a793c558 · outbound

This paper cites Scientific data , volume=.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images Scientific data , volume=

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:07:11.770374Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T21:07:08.851524Z digest=sha256:0ca1a7483efabcf515cc03a2a7f9d2f96751b2a6af149790c20b249826a3c83e

Observation 160b1582-f6be-48de-a60d-5ff7ccd42fd2 · outbound

This paper cites Scientific data , volume=.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images Scientific data , volume=

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:07:11.537733Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T21:07:08.877954Z digest=sha256:176cd24036b47d5a0bfec58ef06d3a62f17282e57f95ef2fefe9ecd71d497ec4

Observation 53f19895-266c-44f3-b1cd-03a41cd2adbe · outbound

This paper cites arXiv preprint arXiv:2305.19112 , year=.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images arXiv preprint arXiv:2305.19112 , year=

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-05T21:07:08.919218Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:07:08.919218Z digest=sha256:b374da450f45623e19c3652236b9ff88b3bdc26ec4ebf9b9af137fcae61a09b2

Observation c8535cf8-1809-4a3e-8a35-262a22cdbf41 · outbound

This paper cites International conference on multimedia modeling , pages=.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images International conference on multimedia modeling , pages=

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-05T21:07:08.948707Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:07:08.948707Z digest=sha256:c41d6c53318834bd7c1e15b984d8dbcd1c2e6c8a448fa2e9a648417d8df48b97

Observation ee625d8d-b7fe-43bd-bcf2-062ab53d7f6f · outbound

This paper cites saliency maps from physicians , author=.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images saliency maps from physicians , author=

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:07:11.392202Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T21:07:08.996716Z digest=sha256:e3a3ae2b38a975cd0e83684ada418c8ad083d94299bfdf07a83beb2a5f1ae42c

Observation 7cade63a-ecd3-4fc1-9294-c121ba9e450b · outbound

This paper cites Information Sciences , volume=.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images Information Sciences , volume=

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:07:11.308346Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T21:07:09.028296Z digest=sha256:50117355e9d8f717835bd2f908a8128a1e7bfa7c3210d6babed7a914a4929d01

Observation c60c39ba-7f7b-4009-bbea-c50fe9b3cc2b · outbound

This paper cites IEEE transactions on medical imaging , volume=.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images IEEE transactions on medical imaging , volume=

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:07:11.227731Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T21:07:09.065811Z digest=sha256:5f451f3d989fc5f13af9ab541dbaeac7e6cb9b477afd7cec77cdc702aee42126

Observation 5dfb962e-91c8-4480-b4ce-152c824890a8 · outbound

This paper cites IEEE Transactions on Medical imaging , volume=.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images IEEE Transactions on Medical imaging , volume=

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T21:07:11.062611Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T21:07:09.098336Z digest=sha256:85dd4affd3c5cf80edb4a5d79a5c093e461c0258b161d8ed8117aa69f70b2618

Observation efe2f606-036d-4080-978e-e841ff6e7743 · outbound

This paper cites European conference on computer vision , pages=.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images European conference on computer vision , pages=

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-05T21:07:09.170037Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:07:09.170037Z digest=sha256:89bdfbfe77a9db3ef230ce5ae645c275fc5dde6e511330a2fe83ec200844ce81

Observation 9d5ab858-d88d-4c63-bfa5-96bcbb3c514d · outbound

This paper cites Ultralytics YOLO26: Unified Real-Time End-to-End Vision Models.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images Ultralytics YOLO26: Unified Real-Time End-to-End Vision Models

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-05T21:07:09.213500Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:07:09.213500Z digest=sha256:55c4ac7e423f97946c1211e640e3baa54afb9a2c42ad4427215639511db6f4f6

Observation a9cd5b55-8b47-41bd-be3f-a25d42f77648 · outbound

This paper cites LW-DETR: A Transformer Replacement to YOLO for Real-Time Detection.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images LW-DETR: A Transformer Replacement to YOLO for Real-Time Detection

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-05T21:07:09.302031Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:07:09.302031Z digest=sha256:071ad81991684818f7096230a5c08949462c50dd55bf236e14a2cf364ff8c6e5

Observation 6c6cbf1d-23fb-4ccb-85cf-f5d563015e11 · outbound

This paper cites EdgeCrafter: Compact ViTs for Edge Dense Prediction via Task-Specialized Distillation.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images EdgeCrafter: Compact ViTs for Edge Dense Prediction via Task-Specialized Distillation

Reference 45

Resolution
verified exact
raw_fallback, observed 2026-08-18T02:13:47.526664Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T21:07:09.384132Z digest=sha256:29ca7cbc0aaa985e7762d84e94d01f9b61c3f1e243e19b8a289b36bc17d4bf06

Observation cb508499-4c1a-41a4-a672-9f366aa90408 · outbound

This paper cites DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-05T21:07:09.449564Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:07:09.449564Z digest=sha256:8317effa8ee26d812f0233877c9c79f57d519700c724c437e065be59becd4dd0

Observation db45a968-61e3-45d5-b5c7-a97ac6f009bd · outbound

This paper cites arXiv preprint arXiv:2512.17436 , year=.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images arXiv preprint arXiv:2512.17436 , year=

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-05T21:07:09.599566Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:07:09.599566Z digest=sha256:be2bfec957b288695be8cda437c070f3f0b4b023ed53915c27a9547111c26807

Observation 07a1f85b-7386-431d-b18e-40ed5284516e · outbound

This paper cites Ovis2.5 Technical Report.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images Ovis2.5 Technical Report

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-05T21:07:09.717152Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:07:09.717152Z digest=sha256:017670127b1cce5329f5db3e8dddce69225d233570adbb2b2415f0467c8e0eaf

Observation 6ebcfbae-2f97-459c-9f74-b717d95ccad2 · outbound

This paper cites Qwen3-VL Technical Report.

LocAnyMed: Vision-Language Grounding for Multimodal Medical Images Qwen3-VL Technical Report

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-05T21:07:09.893269Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:07:09.893269Z digest=sha256:caa82dd533bb5be5e58ee853b626ac535e5c7e14fbb69f8af7cfce1286ae2fed

Pith citing papers

No inbound Pith citation observations are available.