Pith. sign in

Paper Citation Record · LEDGER

Can Multimodal Large Language Models Understand Spatial Relations?

As of 13 August 2026, this Paper Citation Record lists 37 of 37 outbound references and 2 inbound Pith citation observations for arXiv:2505.19015.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.19015 v2

Coverage vector

measured 37 of 37 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:24:07.489391Z

measured 39 of 39 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-03T13:26:06.713886Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

37 of 37 outbound references displayed

  • verified exact1
  • verified fuzzy0
  • unresolved36
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 94638c2e-606d-4dbb-9e17-61e2c405cf06 · outbound

This paper cites online" 'onlinestring :=.

Can Multimodal Large Language Models Understand Spatial Relations? online" 'onlinestring :=

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T14:24:04.016067Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:24:04.016067Z digest=sha256:657e299d83013cce817841b8f4809813674218d09679a855224cd0e8f4c827b0

Observation 9ea1da04-f617-4fe9-8880-a1f75defe4b7 · outbound

This paper cites write newline.

Can Multimodal Large Language Models Understand Spatial Relations? write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T14:24:04.199896Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:24:04.199896Z digest=sha256:ea5324bcbd0bbcc1fececea8cae6d27882f12434af2ffc0f52207817457ce2d8

Observation 85c76fd1-5255-4319-9c6b-f442c04bdd16 · outbound

This paper cites GPT-4 Technical Report.

Can Multimodal Large Language Models Understand Spatial Relations? GPT-4 Technical Report

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T14:24:04.305365Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:24:04.305365Z digest=sha256:0c2eb5f8547fea30d7bb7a282f61d78e8884f94158dc4de49489ad8511025059

Observation 04321268-c26f-4c1a-8693-d555a8a2fd45 · outbound

This paper cites an unresolved cited work.

Can Multimodal Large Language Models Understand Spatial Relations? Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:24:10.227755Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-07T14:24:04.434248Z digest=sha256:a198391b3fa5933e61299bc7f7ae0ecaab50a5fd9e4e5eebcb5953ea813590e4

Observation 7b406e20-d445-4681-a415-e8c4673aa066 · outbound

This paper cites SpatialBot: Precise Spatial Understanding with Vision Language Models.

Can Multimodal Large Language Models Understand Spatial Relations? SpatialBot: Precise Spatial Understanding with Vision Language Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T14:24:04.556539Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:24:04.556539Z digest=sha256:ee092178bac76034e5dff917f102289adc898080d04aa2ce4a24ebea1d7bad5f

Observation 71a025e6-50ec-4938-aa1a-ffaeb4266c67 · outbound

This paper cites an unresolved cited work.

Can Multimodal Large Language Models Understand Spatial Relations? Unresolved cited work

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T14:24:04.683366Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:24:04.683366Z digest=sha256:a02ee9edaa10416b9de54b7167a57837409a90260638e8a41b3ad43fb764721d

Observation 6ee4ee5c-9a39-4345-90d7-00518300fbf6 · outbound

This paper cites SpatialRGPT: Grounded Spatial Reasoning in Vision Language Models.

Can Multimodal Large Language Models Understand Spatial Relations? SpatialRGPT: Grounded Spatial Reasoning in Vision Language Models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T14:24:04.777438Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:24:04.777438Z digest=sha256:05d96d44d9df96e6c545a0cec7f96ca6b66359f315a24cd119d1c3c8766a0d70

Observation 5910c5c7-c0b6-417f-b3f8-26090edfbe5c · outbound

This paper cites EgoThink: Evaluating First-Person Perspective Thinking Capability of Vision-Language Models.

Can Multimodal Large Language Models Understand Spatial Relations? EgoThink: Evaluating First-Person Perspective Thinking Capability of Vision-Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T14:24:04.861523Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:24:04.861523Z digest=sha256:7619eda9451d975c339af71dce2fb115665a23b3a3d161744693c3fc101c42b9

Observation e5b6b00a-112f-4fb0-b7bb-d1d5ba9207a0 · outbound

This paper cites an unresolved cited work.

Can Multimodal Large Language Models Understand Spatial Relations? Unresolved cited work

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T14:24:04.990570Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:24:04.990570Z digest=sha256:b3d3b8a48bc4ce69f9befa1fe437eac3164fdfbbba28b451e7d28fc41bb20004

Observation 208b9e5c-4569-46e9-bf2a-e06398c05537 · outbound

This paper cites EmbSpatial-Bench: Benchmarking Spatial Understanding for Embodied Tasks with Large Vision-Language Models.

Can Multimodal Large Language Models Understand Spatial Relations? EmbSpatial-Bench: Benchmarking Spatial Understanding for Embodied Tasks with Large Vision-Language Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T14:24:05.093854Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:24:05.093854Z digest=sha256:5ce44d836e2a67636d36bbbfa91018f099a8d8e32738aab34a2c854c61fcd778

Observation 24c57652-fcc8-46e0-ae5e-2bb6da3ed56b · outbound

This paper cites MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models.

Can Multimodal Large Language Models Understand Spatial Relations? MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T14:24:05.175871Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:24:05.175871Z digest=sha256:02d9d3385e7ba43116565b7ac4e14d063c56967501991265aed2eca64675f963

Observation cd04fc05-1c0b-432d-b845-c8eabd87d837 · outbound

This paper cites A Survey for Foundation Models in Autonomous Driving.

Can Multimodal Large Language Models Understand Spatial Relations? A Survey for Foundation Models in Autonomous Driving

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T14:24:05.264871Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:24:05.264871Z digest=sha256:a74e97aa734cd4faaa7d299e6f4d899c79258458612f1ee90140465dd449553f

Observation 652cca04-34ff-4909-9b8d-335a74b4ecfa · outbound

This paper cites an unresolved cited work.

Can Multimodal Large Language Models Understand Spatial Relations? Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:24:09.988251Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-07T14:24:05.387147Z digest=sha256:989325469b01cf0b9f0ed07891c2cd8196d334e3f5e15d949d6f7b5bf35a13a9

Observation cbed56f3-c34d-4f67-965c-dc247beb41c2 · outbound

This paper cites an unresolved cited work.

Can Multimodal Large Language Models Understand Spatial Relations? Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:24:09.741763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-07T14:24:05.520208Z digest=sha256:c3da63aa9d78005843ec7beddcbcb1bb087d3cf0a98dae363f723ae29e3fedd7

Observation fec98b1a-2dcc-4e76-8fe1-e6d477ccd908 · outbound

This paper cites an unresolved cited work.

Can Multimodal Large Language Models Understand Spatial Relations? Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:24:09.579034Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-07T14:24:05.606345Z digest=sha256:c7ccf0cb4ae8262487e67861a68d1199cbc41e0010883ade8c8869f4281d83ff

Observation 1773f94f-d438-4a2a-bb74-d5b2755fb262 · outbound

This paper cites LoRA: Low-Rank Adaptation of Large Language Models.

Can Multimodal Large Language Models Understand Spatial Relations? LoRA: Low-Rank Adaptation of Large Language Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T14:24:05.718317Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:24:05.718317Z digest=sha256:ad59d8698036cb8d4f22f7259deaa2bed59888aed7150a2e900b7b0940fff5ae

Observation a0d41e7b-d23b-4c9d-bd9c-4de51431891a · outbound

This paper cites an unresolved cited work.

Can Multimodal Large Language Models Understand Spatial Relations? Unresolved cited work

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T14:24:05.842198Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:24:05.842198Z digest=sha256:ac5ec674e75a2aa036fc580965f706f650c23511a90285203adb0658adbb3ae9

Observation 9eace458-3a77-43b0-9f48-a1c7af94ac54 · outbound

This paper cites an unresolved cited work.

Can Multimodal Large Language Models Understand Spatial Relations? Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:24:09.427448Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-07T14:24:05.957479Z digest=sha256:5fc5a57b3264aaf76a582cd8eb84e91682ca5f6d8d1fb519127d769768c4d4b1

Observation 55cdf097-1c50-46a1-bfae-b9cfa6a8fcd9 · outbound

This paper cites an unresolved cited work.

Can Multimodal Large Language Models Understand Spatial Relations? Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:24:09.241890Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-07T14:24:06.041764Z digest=sha256:5732ee2b51a18ac2382439890096cfee003aecdf31277f3fc7ae6660d23c168d

Observation ed5c6e44-bc67-476c-8aab-6c311db9b9c8 · outbound

This paper cites an unresolved cited work.

Can Multimodal Large Language Models Understand Spatial Relations? Unresolved cited work

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T14:24:06.159421Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:24:06.159421Z digest=sha256:572a91949f5fcea9bd1da032b3eda3cad814949ca51015d3b8a7dd20202a5f91

Observation 412c00e5-dc90-46ec-b90f-1bc987e3c628 · outbound

This paper cites an unresolved cited work.

Can Multimodal Large Language Models Understand Spatial Relations? Unresolved cited work

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T14:24:06.250116Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:24:06.250116Z digest=sha256:dca53bf144a18b56a2ef62caef725c0854020bafe496d9d5dd0dd420f8faa792

Observation 159aa6df-2de7-4683-a052-955e6dcee207 · outbound

This paper cites an unresolved cited work.

Can Multimodal Large Language Models Understand Spatial Relations? Unresolved cited work

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T14:24:06.318633Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:24:06.318633Z digest=sha256:c6858007345c81c9b480332c693a079ef49b28e82ee2a109eb99fcc791a3c5ea

Observation 93aa9eee-a76a-4107-9616-0622e863650c · outbound

This paper cites an unresolved cited work.

Can Multimodal Large Language Models Understand Spatial Relations? Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:24:09.026833Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-07T14:24:06.384825Z digest=sha256:8c6162b5cf29f1a81b4bb972676650fbdd996dd9950f0941ebe1e68267917160

Observation 0b6a188c-c912-4935-9f2f-77a8edfa9679 · outbound

This paper cites an unresolved cited work.

Can Multimodal Large Language Models Understand Spatial Relations? Unresolved cited work

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T14:24:06.425516Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:24:06.425516Z digest=sha256:57c0f6bf52fde1faca88a90d8440e902ef8aeeda7964243297bfdf2dacb5c755

Observation 1d0ad457-f85f-435c-a485-8344616daf1e · outbound

This paper cites MMHQA-ICL: Multimodal In-context Learning for Hybrid Question Answering over Text, Tables and Images.

Can Multimodal Large Language Models Understand Spatial Relations? MMHQA-ICL: Multimodal In-context Learning for Hybrid Question Answering over Text, Tables and Images

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T14:24:06.523784Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:24:06.523784Z digest=sha256:74d8f6483473a244b7276533b84d326917b182bc12580586102b55600d95b6af

Observation 8fbdc9b1-e232-4dd3-807e-8d2afa261895 · outbound

This paper cites P-Tuning v2: Prompt Tuning Can Be Comparable to Fine-tuning Universally Across Scales and Tasks.

Can Multimodal Large Language Models Understand Spatial Relations? P-Tuning v2: Prompt Tuning Can Be Comparable to Fine-tuning Universally Across Scales and Tasks

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T14:24:06.627542Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:24:06.627542Z digest=sha256:3447b9580b261cfd039c7e4a51b75bcf1e411929a90d6ad995f7edbeb2a99104

Observation 7f272d00-8eeb-4138-8f64-c735d3004cb6 · outbound

This paper cites an unresolved cited work.

Can Multimodal Large Language Models Understand Spatial Relations? Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:24:08.828915Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-07T14:24:06.716472Z digest=sha256:7284992aa5e4dff5a3f3d50466602f8d26ca6131cc624299713727adfd801ccf

Observation 37e2bfd1-b1d1-4f4b-be8b-4daff3631f8c · outbound

This paper cites an unresolved cited work.

Can Multimodal Large Language Models Understand Spatial Relations? Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:24:08.657963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-07T14:24:06.800214Z digest=sha256:4625415fb165c4518a67ec4c6168ca237ded6c6ff980d8cf18fdab951fadcb8d

Observation ca18be29-2106-47db-b102-f380972139fe · outbound

This paper cites an unresolved cited work.

Can Multimodal Large Language Models Understand Spatial Relations? Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:24:08.500757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-07T14:24:06.908390Z digest=sha256:03930a0c0533843b9b3c6ccd5f3162d12b495568b3d7a3309002122bf0087124

Observation 15815e6a-a2f6-437f-b5aa-4257a1c7f842 · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

Can Multimodal Large Language Models Understand Spatial Relations? Gemini: A Family of Highly Capable Multimodal Models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T14:24:06.999213Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:24:06.999213Z digest=sha256:b0b97d7a10d3615a5d82bee770c6a0306130aa3f0c4de601482656ead06db12f

Observation 1bfa95ca-ef8e-4085-91ad-ba81db48bbc0 · outbound

This paper cites an unresolved cited work.

Can Multimodal Large Language Models Understand Spatial Relations? Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:24:08.369030Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-07T14:24:07.055340Z digest=sha256:b2d18ed5a76329fbcf48d373a055f853465453690ebcbfc8f6df2ff152791661

Observation e4ad624f-e49c-4527-8863-40e463aac19b · outbound

This paper cites an unresolved cited work.

Can Multimodal Large Language Models Understand Spatial Relations? Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:24:08.228301Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-07T14:24:07.117272Z digest=sha256:33dedc0c9b388eb581559befcfd728121e1fbac68ea183522175f85b81983920

Observation 6d4bca06-c6b1-4636-aa82-31f910ef2921 · outbound

This paper cites Can Transformers Capture Spatial Relations between Objects?.

Can Multimodal Large Language Models Understand Spatial Relations? Can Transformers Capture Spatial Relations between Objects?

Reference 33

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:24:07.735657Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-07T14:24:07.210177Z digest=sha256:37255932fdde061c8d322e5c67d70aa36827edc21cf27cd371586a4d13660c8f

Observation 31e9da66-8395-45cb-ace7-d3729ca6b859 · outbound

This paper cites an unresolved cited work.

Can Multimodal Large Language Models Understand Spatial Relations? Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:24:08.054120Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.

source=arxiv_source observed=2026-08-07T14:24:07.281480Z digest=sha256:ed6ea40b3116506e8245fca646ee7b925a7b647d0061bdb0ac32dfb59d94978a

Observation 9bef279a-abe8-4ce1-a3fc-342b17aa70ff · outbound

This paper cites mPLUG-Owl: Modularization Empowers Large Language Models with Multimodality.

Can Multimodal Large Language Models Understand Spatial Relations? mPLUG-Owl: Modularization Empowers Large Language Models with Multimodality

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T14:24:07.339557Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:24:07.339557Z digest=sha256:f0b2247abd91c3b763d67d8b9f4bdaef295d00aad307747bb4b93e41be34ea2d

Observation 8670db17-2935-4c00-b800-55c2984d1be3 · outbound

This paper cites CoCoT: Contrastive Chain-of-Thought Prompting for Large Multimodal Models with Multiple Image Inputs.

Can Multimodal Large Language Models Understand Spatial Relations? CoCoT: Contrastive Chain-of-Thought Prompting for Large Multimodal Models with Multiple Image Inputs

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T14:24:07.427455Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:24:07.427455Z digest=sha256:8ef42f355d8f5ce2ecad64c917ad8a3c0417dad4b282bad9645fbbe465c635b2

Observation 52c7551b-2261-42e7-b9d5-ad0aadc87c5d · outbound

This paper cites an unresolved cited work.

Can Multimodal Large Language Models Understand Spatial Relations? Unresolved cited work

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T14:24:07.489391Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:24:07.489391Z digest=sha256:fff62a3ca3eede93d99ab8b9518094fddbdb19574b5461eb142f23fb0a2afeb7

Pith citing papers

Observation 09564187-549e-4c2d-84e4-0eef8ac830b1 · inbound

Spatial-aware Vision Language Model for Autonomous Driving cites this paper.

Spatial-aware Vision Language Model for Autonomous Driving Can Multimodal Large Language Models Understand Spatial Relations?

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-03T13:26:06.713886Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T13:26:06.713886Z digest=sha256:9c98ea65e84ac11c1e5fa1f13cfb404c74921071f83a5bb947f7cbec17650413

Observation e795f2a0-0f3f-4d36-9c85-8229c6859310 · inbound

AnatomiX, an Anatomy-Aware Grounded Multimodal Large Language Model for Chest X-Ray Interpretation cites this paper.

AnatomiX, an Anatomy-Aware Grounded Multimodal Large Language Model for Chest X-Ray Interpretation Can Multimodal Large Language Models Understand Spatial Relations?

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-03T12:25:50.226593Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:25:50.226593Z digest=sha256:f09bc1f299f5c8ed4f665f101d38c8099bad1bbb044036708196afe32811a902