Pith. sign in

Paper Citation Record · LEDGER

Can Multimodal Large Language Models Understand Spatial Relations?

As of 8 August 2026, this Paper Citation Record lists 37 of 37 outbound references and 2 inbound Pith citation observations for arXiv:2505.19015.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.19015 v2

Coverage vector

measured 37 of 37 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:24:07.489391Z

measured 39 of 39 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-03T13:26:06.713886Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

37 of 37 outbound references displayed

  • verified exact1
  • verified fuzzy0
  • unresolved36
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 94638c2e-606d-4dbb-9e17-61e2c405cf06 · outbound

This paper cites online" 'onlinestring :=.

Can Multimodal Large Language Models Understand Spatial Relations? online" 'onlinestring :=

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T14:24:04.016067Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:24:04.016067Z digest=sha256:31474aee24874363f01df4e7e41788865e33b8c81a9aac2dd52bf5b47c239d4f

Observation 9ea1da04-f617-4fe9-8880-a1f75defe4b7 · outbound

This paper cites write newline.

Can Multimodal Large Language Models Understand Spatial Relations? write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T14:24:04.199896Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:24:04.199896Z digest=sha256:3430fb998a337cb91535bb9875210d7a88c19fd3f54962adf6db79090fe9c747

Observation 85c76fd1-5255-4319-9c6b-f442c04bdd16 · outbound

This paper cites GPT-4 Technical Report.

Can Multimodal Large Language Models Understand Spatial Relations? GPT-4 Technical Report

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T14:24:04.305365Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:24:04.305365Z digest=sha256:4265680af5493a858c5f6854fd4375f97730372a7056221f45949f1277a6f595

Observation 04321268-c26f-4c1a-8693-d555a8a2fd45 · outbound

This paper cites an unresolved cited work.

Can Multimodal Large Language Models Understand Spatial Relations? Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:24:10.227755Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:24:04.434248Z digest=sha256:df3907933473516a081c3ef616eb3e128a8c09123f5a154193b770c70ec89c17

Observation 7b406e20-d445-4681-a415-e8c4673aa066 · outbound

This paper cites SpatialBot: Precise Spatial Understanding with Vision Language Models.

Can Multimodal Large Language Models Understand Spatial Relations? SpatialBot: Precise Spatial Understanding with Vision Language Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T14:24:04.556539Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:24:04.556539Z digest=sha256:cf3cda99fc898bbcfec7a33d688a29702ce29c4ac076f16208ce33571748f835

Observation 71a025e6-50ec-4938-aa1a-ffaeb4266c67 · outbound

This paper cites an unresolved cited work.

Can Multimodal Large Language Models Understand Spatial Relations? Unresolved cited work

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T14:24:04.683366Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:24:04.683366Z digest=sha256:5c3cae1878158f8beb24d6fcc5fc61755adc56e2e0aebae88625d521ed737174

Observation 6ee4ee5c-9a39-4345-90d7-00518300fbf6 · outbound

This paper cites SpatialRGPT: Grounded Spatial Reasoning in Vision Language Models.

Can Multimodal Large Language Models Understand Spatial Relations? SpatialRGPT: Grounded Spatial Reasoning in Vision Language Models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T14:24:04.777438Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:24:04.777438Z digest=sha256:953084cbe4a52f633959061a2c513121e0dfd3795de6a85ad7d7c31c467975a2

Observation 5910c5c7-c0b6-417f-b3f8-26090edfbe5c · outbound

This paper cites EgoThink: Evaluating First-Person Perspective Thinking Capability of Vision-Language Models.

Can Multimodal Large Language Models Understand Spatial Relations? EgoThink: Evaluating First-Person Perspective Thinking Capability of Vision-Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T14:24:04.861523Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:24:04.861523Z digest=sha256:224015c38c8d5cd3b10bbae8753db04ae933195c0641502ad48ce767e6b53f28

Observation e5b6b00a-112f-4fb0-b7bb-d1d5ba9207a0 · outbound

This paper cites an unresolved cited work.

Can Multimodal Large Language Models Understand Spatial Relations? Unresolved cited work

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T14:24:04.990570Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:24:04.990570Z digest=sha256:1a1a8135058632ce579fcc20c9e34eebe3fd137e5d44f9810e4d1270eb82eca2

Observation 208b9e5c-4569-46e9-bf2a-e06398c05537 · outbound

This paper cites EmbSpatial-Bench: Benchmarking Spatial Understanding for Embodied Tasks with Large Vision-Language Models.

Can Multimodal Large Language Models Understand Spatial Relations? EmbSpatial-Bench: Benchmarking Spatial Understanding for Embodied Tasks with Large Vision-Language Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T14:24:05.093854Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:24:05.093854Z digest=sha256:8eaaf73d79d27b2bce420126f5cb7ed97b0377782ff071413eca21d8afa0695b

Observation 24c57652-fcc8-46e0-ae5e-2bb6da3ed56b · outbound

This paper cites MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models.

Can Multimodal Large Language Models Understand Spatial Relations? MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T14:24:05.175871Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:24:05.175871Z digest=sha256:2b52a1306b674ce0f3cdcbc7760ccff10ce9bcfeb806d449eff988934d7ee8d3

Observation cd04fc05-1c0b-432d-b845-c8eabd87d837 · outbound

This paper cites A Survey for Foundation Models in Autonomous Driving.

Can Multimodal Large Language Models Understand Spatial Relations? A Survey for Foundation Models in Autonomous Driving

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T14:24:05.264871Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:24:05.264871Z digest=sha256:77b12d0ecdf359599c22303725a55368584bb4c3833118c45a7db9721ebdf213

Observation 652cca04-34ff-4909-9b8d-335a74b4ecfa · outbound

This paper cites an unresolved cited work.

Can Multimodal Large Language Models Understand Spatial Relations? Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:24:09.988251Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:24:05.387147Z digest=sha256:9fefae60576ab050958e70b6801fd65101f345f41f703c347e4dfec315cc02ff

Observation cbed56f3-c34d-4f67-965c-dc247beb41c2 · outbound

This paper cites an unresolved cited work.

Can Multimodal Large Language Models Understand Spatial Relations? Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:24:09.741763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:24:05.520208Z digest=sha256:9bf13aae81230a4a800cfd93578a42317a7cb84be4c7281e4d8de4589e6949c7

Observation fec98b1a-2dcc-4e76-8fe1-e6d477ccd908 · outbound

This paper cites an unresolved cited work.

Can Multimodal Large Language Models Understand Spatial Relations? Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:24:09.579034Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:24:05.606345Z digest=sha256:b106af5075669182650937b13c329172bdb7225e904c8116c62953fe51f446e1

Observation 1773f94f-d438-4a2a-bb74-d5b2755fb262 · outbound

This paper cites LoRA: Low-Rank Adaptation of Large Language Models.

Can Multimodal Large Language Models Understand Spatial Relations? LoRA: Low-Rank Adaptation of Large Language Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T14:24:05.718317Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:24:05.718317Z digest=sha256:c071e67925eef650581b958c7e6f48f85c719cd7b1f9ee2a0693c35f7cd7ec11

Observation a0d41e7b-d23b-4c9d-bd9c-4de51431891a · outbound

This paper cites an unresolved cited work.

Can Multimodal Large Language Models Understand Spatial Relations? Unresolved cited work

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T14:24:05.842198Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:24:05.842198Z digest=sha256:cf888f712dd98b78e715a3d70508ae9aaff9ccb331fcc6e93adfee6283f46946

Observation 9eace458-3a77-43b0-9f48-a1c7af94ac54 · outbound

This paper cites an unresolved cited work.

Can Multimodal Large Language Models Understand Spatial Relations? Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:24:09.427448Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:24:05.957479Z digest=sha256:213a1172e5f10d2c85e5990934359b2e82fd4733bddb5ad0af3e854877cc9394

Observation 55cdf097-1c50-46a1-bfae-b9cfa6a8fcd9 · outbound

This paper cites an unresolved cited work.

Can Multimodal Large Language Models Understand Spatial Relations? Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:24:09.241890Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:24:06.041764Z digest=sha256:b3fb61da5d359068eb30e3246b03f3572c3fe280013e8a198273a4a221f69877

Observation ed5c6e44-bc67-476c-8aab-6c311db9b9c8 · outbound

This paper cites an unresolved cited work.

Can Multimodal Large Language Models Understand Spatial Relations? Unresolved cited work

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T14:24:06.159421Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:24:06.159421Z digest=sha256:f9a447b459a8aa7af83064d20e810813ab3dae91d3dbb4704d3edfd4528c76d5

Observation 412c00e5-dc90-46ec-b90f-1bc987e3c628 · outbound

This paper cites an unresolved cited work.

Can Multimodal Large Language Models Understand Spatial Relations? Unresolved cited work

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T14:24:06.250116Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:24:06.250116Z digest=sha256:ef12cc9af7a71bd34c3c80546e53ecac26d29557f4bdda6de5c44b3cb91dc3a5

Observation 159aa6df-2de7-4683-a052-955e6dcee207 · outbound

This paper cites an unresolved cited work.

Can Multimodal Large Language Models Understand Spatial Relations? Unresolved cited work

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T14:24:06.318633Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:24:06.318633Z digest=sha256:462589d84c150cf71694b7f17144f51c2fd9ab56b5c5e42555e6624ac8d4b983

Observation 93aa9eee-a76a-4107-9616-0622e863650c · outbound

This paper cites an unresolved cited work.

Can Multimodal Large Language Models Understand Spatial Relations? Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:24:09.026833Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:24:06.384825Z digest=sha256:b0fd3358c263692116ca1231e70e52bfe56faa102af0e01b7b248ea0798a4fef

Observation 0b6a188c-c912-4935-9f2f-77a8edfa9679 · outbound

This paper cites an unresolved cited work.

Can Multimodal Large Language Models Understand Spatial Relations? Unresolved cited work

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T14:24:06.425516Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:24:06.425516Z digest=sha256:3f8db02aa48ffec505aea1aab817d1193536b551cee24e91a01031ea6ec21033

Observation 1d0ad457-f85f-435c-a485-8344616daf1e · outbound

This paper cites MMHQA-ICL: Multimodal In-context Learning for Hybrid Question Answering over Text, Tables and Images.

Can Multimodal Large Language Models Understand Spatial Relations? MMHQA-ICL: Multimodal In-context Learning for Hybrid Question Answering over Text, Tables and Images

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T14:24:06.523784Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:24:06.523784Z digest=sha256:a327143dfd83cb6c4fa3877aa9b70264db4af2c9b74eb709661338f8ce3c9063

Observation 8fbdc9b1-e232-4dd3-807e-8d2afa261895 · outbound

This paper cites P-Tuning v2: Prompt Tuning Can Be Comparable to Fine-tuning Universally Across Scales and Tasks.

Can Multimodal Large Language Models Understand Spatial Relations? P-Tuning v2: Prompt Tuning Can Be Comparable to Fine-tuning Universally Across Scales and Tasks

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T14:24:06.627542Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:24:06.627542Z digest=sha256:8de881931864a8297f6ff8dea59d4d891fa7a1f8f0167473b95c5485e05ba21a

Observation 7f272d00-8eeb-4138-8f64-c735d3004cb6 · outbound

This paper cites an unresolved cited work.

Can Multimodal Large Language Models Understand Spatial Relations? Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:24:08.828915Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:24:06.716472Z digest=sha256:728748c2701e574260bb2dbda5b9c9c4edc0936d89d9ac6bf2012510bd765895

Observation 37e2bfd1-b1d1-4f4b-be8b-4daff3631f8c · outbound

This paper cites an unresolved cited work.

Can Multimodal Large Language Models Understand Spatial Relations? Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:24:08.657963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:24:06.800214Z digest=sha256:83c78ad7dc32eb4acbefd52b40e9bce85e022389f099310274951e4096905f38

Observation ca18be29-2106-47db-b102-f380972139fe · outbound

This paper cites an unresolved cited work.

Can Multimodal Large Language Models Understand Spatial Relations? Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:24:08.500757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:24:06.908390Z digest=sha256:ec5ffd612a7b24e5b855eded4f0f08008247607bb7880f9553b5785154aabaad

Observation 15815e6a-a2f6-437f-b5aa-4257a1c7f842 · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

Can Multimodal Large Language Models Understand Spatial Relations? Gemini: A Family of Highly Capable Multimodal Models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T14:24:06.999213Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:24:06.999213Z digest=sha256:fc0241243e522bfb6205f29c69ff768d12f6d62400331536e6a099aa37a47ea7

Observation 1bfa95ca-ef8e-4085-91ad-ba81db48bbc0 · outbound

This paper cites an unresolved cited work.

Can Multimodal Large Language Models Understand Spatial Relations? Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:24:08.369030Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:24:07.055340Z digest=sha256:699c00f7f29039a62766d9fad3c9f09694a8a2b3ab00599e570e25b09fad1342

Observation e4ad624f-e49c-4527-8863-40e463aac19b · outbound

This paper cites an unresolved cited work.

Can Multimodal Large Language Models Understand Spatial Relations? Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:24:08.228301Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:24:07.117272Z digest=sha256:021c30608a4b34a41c0b9c09c292ec6ae117df2c61da2a26181169aa89919063

Observation 6d4bca06-c6b1-4636-aa82-31f910ef2921 · outbound

This paper cites Can Transformers Capture Spatial Relations between Objects?.

Can Multimodal Large Language Models Understand Spatial Relations? Can Transformers Capture Spatial Relations between Objects?

Reference 33

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:24:07.735657Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:24:07.210177Z digest=sha256:d8c7d3ed168014f4e842c811769baf99eeef77b282bd5dcc24895560805411eb

Observation 31e9da66-8395-45cb-ace7-d3729ca6b859 · outbound

This paper cites an unresolved cited work.

Can Multimodal Large Language Models Understand Spatial Relations? Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:24:08.054120Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T14:24:07.281480Z digest=sha256:ccbde6e8a5d80237739ea4642a20c957541615854613d64aadd52f0260638e4f

Observation 9bef279a-abe8-4ce1-a3fc-342b17aa70ff · outbound

This paper cites mPLUG-Owl: Modularization Empowers Large Language Models with Multimodality.

Can Multimodal Large Language Models Understand Spatial Relations? mPLUG-Owl: Modularization Empowers Large Language Models with Multimodality

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T14:24:07.339557Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:24:07.339557Z digest=sha256:9c66908e638b14d617cb2951b1079fa125fabc930d3f63d288f40364e621eb20

Observation 8670db17-2935-4c00-b800-55c2984d1be3 · outbound

This paper cites CoCoT: Contrastive Chain-of-Thought Prompting for Large Multimodal Models with Multiple Image Inputs.

Can Multimodal Large Language Models Understand Spatial Relations? CoCoT: Contrastive Chain-of-Thought Prompting for Large Multimodal Models with Multiple Image Inputs

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T14:24:07.427455Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:24:07.427455Z digest=sha256:2ca38a1937bfa3f3e636c1d7aeec55de1a804fcbd57be88e44cef0d9d1bcd5bf

Observation 52c7551b-2261-42e7-b9d5-ad0aadc87c5d · outbound

This paper cites an unresolved cited work.

Can Multimodal Large Language Models Understand Spatial Relations? Unresolved cited work

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T14:24:07.489391Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:24:07.489391Z digest=sha256:317d31f9fa94796b8941695df96f19177f8faba0db65327428e8ece96b138904

Pith citing papers

Observation 09564187-549e-4c2d-84e4-0eef8ac830b1 · inbound

Spatial-aware Vision Language Model for Autonomous Driving cites this paper.

Spatial-aware Vision Language Model for Autonomous Driving Can Multimodal Large Language Models Understand Spatial Relations?

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-03T13:26:06.713886Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:26:06.713886Z digest=sha256:aba29e8e04038e6dbc002e7794b260c7d9b5e8cda83dd972326d4f807ca331c5

Observation e795f2a0-0f3f-4d36-9c85-8229c6859310 · inbound

AnatomiX, an Anatomy-Aware Grounded Multimodal Large Language Model for Chest X-Ray Interpretation cites this paper.

AnatomiX, an Anatomy-Aware Grounded Multimodal Large Language Model for Chest X-Ray Interpretation Can Multimodal Large Language Models Understand Spatial Relations?

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-03T12:25:50.226593Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:25:50.226593Z digest=sha256:aeeb678fa2213c869dc6e8785e17f2a24785d5e785396651902dd61631e5eb30