Pith. sign in

Paper Citation Record · LEDGER

Chain of Spatial Thoughts: Modality-Agnostic Spatial Grounding for Vision Language Models

As of 24 August 2026, this Paper Citation Record lists 44 of 44 outbound references and 0 inbound Pith citation observations for arXiv:2608.10278.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.10278 v1

Coverage vector

measured 44 of 44 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-14T04:16:06.304920Z

measured 44 of 44 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

44 of 44 outbound references displayed

  • verified exact1
  • verified fuzzy20
  • unresolved23
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a048de66-c3c5-4f1b-a8f8-a17296a9532e · outbound

This paper cites and Han, Rilyn and Fei-Fei, Li and Xie, Saining , title =.

Chain of Spatial Thoughts: Modality-Agnostic Spatial Grounding for Vision Language Models and Han, Rilyn and Fei-Fei, Li and Xie, Saining , title =

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-14T04:16:06.098155Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:16:06.098155Z digest=sha256:2a54b8e21e625559a368eec83d5c5fa1c501f32d3ff961b2a270195c70df5580

Observation 89999f83-87a5-4da9-b401-41dafaa82bed · outbound

This paper cites 2025 , eprint=.

Chain of Spatial Thoughts: Modality-Agnostic Spatial Grounding for Vision Language Models 2025 , eprint=

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:16:07.712507Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-14T04:16:06.105261Z digest=sha256:a3a4cb3fec007e0bdad06562f6de674697a1992e399c42f7aa2083575dac75ae

Observation aa9e10aa-50c4-48cc-9c1a-9284b0affd2e · outbound

This paper cites ARKitScenes: A Diverse Real-World Dataset For 3D Indoor Scene Understanding Using Mobile RGB-D Data , url =.

Chain of Spatial Thoughts: Modality-Agnostic Spatial Grounding for Vision Language Models ARKitScenes: A Diverse Real-World Dataset For 3D Indoor Scene Understanding Using Mobile RGB-D Data , url =

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:16:07.699483Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-14T04:16:06.109356Z digest=sha256:23aedaafce1a2114577ae64f53867de32ecb76514daeefdd7d4d836baff82e69

Observation bcaf662b-6c9f-4dec-b564-7e86b4b05cf8 · outbound

This paper cites an unresolved cited work.

Chain of Spatial Thoughts: Modality-Agnostic Spatial Grounding for Vision Language Models Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-14T04:16:07.685314Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-14T04:16:06.113547Z digest=sha256:7d023e05a6fb580cf6bafa50745d77dc1d6902f730500243137350c33fa8e379

Observation 77531e96-2392-449e-aace-12227ee86910 · outbound

This paper cites Proceedings of the IEEE/CVF International Conference on Computer Vision , pages=.

Chain of Spatial Thoughts: Modality-Agnostic Spatial Grounding for Vision Language Models Proceedings of the IEEE/CVF International Conference on Computer Vision , pages=

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-14T04:16:06.119929Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:16:06.119929Z digest=sha256:5570b00fd94649fcb14d8e16c45c37d33eff68578b38ea700cad3e53286fd1f8

Observation 0dd7f347-cbf8-4132-b9d5-125184b65256 · outbound

This paper cites 2026 , eprint=.

Chain of Spatial Thoughts: Modality-Agnostic Spatial Grounding for Vision Language Models 2026 , eprint=

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:16:07.662456Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-14T04:16:06.125646Z digest=sha256:5e26ba4c364aead527cad24e2d9df8a937d3dbe7d395e744802c9954f17e1b71

Observation 842b44ff-c10d-47af-a083-3e98cc92e8c9 · outbound

This paper cites 2024 , eprint=.

Chain of Spatial Thoughts: Modality-Agnostic Spatial Grounding for Vision Language Models 2024 , eprint=

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-14T04:16:06.131624Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:16:06.131624Z digest=sha256:5407ad5af1c2549ac9aa2f9d805d814a07aee3cb70858ac7b267493c1f903d14

Observation 5068f620-fae6-4c6a-a281-fa5a19d51b4c · outbound

This paper cites 2021 , eprint=.

Chain of Spatial Thoughts: Modality-Agnostic Spatial Grounding for Vision Language Models 2021 , eprint=

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-14T04:16:06.135857Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:16:06.135857Z digest=sha256:d380497a3107d66fcd82306814fa8dd3c76f1b9235d9f21272cd2b46a5359b5b

Observation 56c5f222-19a9-4817-874e-2b5263b67153 · outbound

This paper cites 2024 , eprint=.

Chain of Spatial Thoughts: Modality-Agnostic Spatial Grounding for Vision Language Models 2024 , eprint=

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-14T04:16:06.148875Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:16:06.148875Z digest=sha256:35c7a060d99b2a8574115df2387da0992d1d097936418e3d04cd3696565f3963

Observation 81f8cced-5111-477d-8a6d-f46e45e074a7 · outbound

This paper cites SpaceR: Reinforcing MLLMs in Video Spatial Reasoning.

Chain of Spatial Thoughts: Modality-Agnostic Spatial Grounding for Vision Language Models SpaceR: Reinforcing MLLMs in Video Spatial Reasoning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-14T04:16:06.153749Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:16:06.153749Z digest=sha256:899a6dceef7bb97e98cfa8c8c98855ec456d3e5a40c16f48987e4094ac2aa2c5

Observation 0a4d1178-3f1b-4f98-bfd0-9c0c5eff08ef · outbound

This paper cites 2025 , eprint=.

Chain of Spatial Thoughts: Modality-Agnostic Spatial Grounding for Vision Language Models 2025 , eprint=

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:16:07.594502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-14T04:16:06.158133Z digest=sha256:488249e953d32187a6d38a771219b752a2bf83484e26c0742374a8d980026dbb

Observation 1f8f3b33-1211-48ca-8edc-72981be6719b · outbound

This paper cites Spatial-MLLM: Boosting MLLM Capabilities in Visual-based Spatial Intelligence , url =.

Chain of Spatial Thoughts: Modality-Agnostic Spatial Grounding for Vision Language Models Spatial-MLLM: Boosting MLLM Capabilities in Visual-based Spatial Intelligence , url =

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:16:07.581548Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-14T04:16:06.164256Z digest=sha256:047da1b8abd0821a315ff87adf8688459a2ef374386b4c685a5d643c46f0992f

Observation a21c4ee0-4a45-4b18-8446-3375ad32ca0e · outbound

This paper cites 2025 , eprint=.

Chain of Spatial Thoughts: Modality-Agnostic Spatial Grounding for Vision Language Models 2025 , eprint=

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:16:07.555514Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-14T04:16:06.168944Z digest=sha256:6510b23558512ee05a856953cb5dc40563e9d47572545152d6d4b7e4466ece06

Observation e5544cca-8d0c-4bf2-ad6b-6ba3e7bdf40d · outbound

This paper cites 2026 , eprint=.

Chain of Spatial Thoughts: Modality-Agnostic Spatial Grounding for Vision Language Models 2026 , eprint=

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:16:07.530586Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-14T04:16:06.173679Z digest=sha256:1120a8bdcc057df31eb4d11aa9e7158953c21be5de4ee84ce50c87368d6dc884

Observation 5054642a-2078-4cb6-8af7-964696b9bf08 · outbound

This paper cites 2025 , eprint=.

Chain of Spatial Thoughts: Modality-Agnostic Spatial Grounding for Vision Language Models 2025 , eprint=

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:16:07.505341Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-14T04:16:06.178280Z digest=sha256:5f2700beca37c471fdc2cf2be600c478b9d7d0b2851a77e19141a9836a3b23b0

Observation 0c4c09ec-57fd-4b59-9cb4-0a7f295fda9f · outbound

This paper cites Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , month =.

Chain of Spatial Thoughts: Modality-Agnostic Spatial Grounding for Vision Language Models Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) , month =

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:16:07.490857Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-14T04:16:06.182292Z digest=sha256:c889b6e5a6508790e72ebf39e2d91887cf410ef24512d53c744a480bc014c7be

Observation 8fc276b4-a6d9-4866-bacf-058391c57cbc · outbound

This paper cites and Saenko, Kate and Krishna, Ranjay and Guibas, Leonidas and Chu, Wen-Sheng , title =.

Chain of Spatial Thoughts: Modality-Agnostic Spatial Grounding for Vision Language Models and Saenko, Kate and Krishna, Ranjay and Guibas, Leonidas and Chu, Wen-Sheng , title =

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:16:07.478356Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-14T04:16:06.185853Z digest=sha256:d130396b60e2033c2bb1247deecf7af9fe3e5ae9ca5210fec413772bfc880ba9

Observation d1c4922b-b9c9-4a04-b667-3d21bc76b694 · outbound

This paper cites Chain-of-Visual-Thought: Teaching VLMs to See and Think Better with Continuous Visual Tokens.

Chain of Spatial Thoughts: Modality-Agnostic Spatial Grounding for Vision Language Models Chain-of-Visual-Thought: Teaching VLMs to See and Think Better with Continuous Visual Tokens

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-14T04:16:06.189513Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:16:06.189513Z digest=sha256:07cd08b22572cdb2edf72fab34e453abcd60d429dd0e50c3fcbeb5aa2579c382

Observation 80c287ae-7989-4b66-809b-b47517fafa22 · outbound

This paper cites 2026 , eprint=.

Chain of Spatial Thoughts: Modality-Agnostic Spatial Grounding for Vision Language Models 2026 , eprint=

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:16:07.464611Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-14T04:16:06.193944Z digest=sha256:c0918edb6360e65580fa619de490c9d42d5ee3b4dd1799740056404489e0d14c

Observation 79d57768-cb70-43ea-8934-894b865776b9 · outbound

This paper cites 2026 , eprint=.

Chain of Spatial Thoughts: Modality-Agnostic Spatial Grounding for Vision Language Models 2026 , eprint=

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:16:07.451597Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-14T04:16:06.207574Z digest=sha256:e4b7a5c3a903212c3db00abb3536851a8870110c7531f96350a4a121af9dc913

Observation df96d802-e41c-49d1-bb56-7eaf8bbcec90 · outbound

This paper cites 2025 , eprint=.

Chain of Spatial Thoughts: Modality-Agnostic Spatial Grounding for Vision Language Models 2025 , eprint=

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:16:07.438470Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-14T04:16:06.211416Z digest=sha256:f20f59890a81f63b96f25151ba6474f48bf02a9244c0c9e982c1f97df22ce187

Observation 7901008e-b669-48a8-b0a6-a464788b3a4b · outbound

This paper cites 2026 , eprint=.

Chain of Spatial Thoughts: Modality-Agnostic Spatial Grounding for Vision Language Models 2026 , eprint=

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:16:07.401393Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-14T04:16:06.217737Z digest=sha256:22134f3db0ded8b2a123bd6a1b034bb8ea87fc8ac64d81e31ba308c08668cbed

Observation 40fbdaea-d985-40e5-bd40-23edb8cd8d22 · outbound

This paper cites 2023 , eprint=.

Chain of Spatial Thoughts: Modality-Agnostic Spatial Grounding for Vision Language Models 2023 , eprint=

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-14T04:16:06.221915Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:16:06.221915Z digest=sha256:e0cc99cc4293076658608e6b48747269815f8854d23962fcf50fee25226e986e

Observation 4ed25394-c91c-4307-a56f-eec7d7724674 · outbound

This paper cites DUSt3R: Geometric 3D Vision Made Easy , year=.

Chain of Spatial Thoughts: Modality-Agnostic Spatial Grounding for Vision Language Models DUSt3R: Geometric 3D Vision Made Easy , year=

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-14T04:16:06.225825Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:16:06.225825Z digest=sha256:4a4296b04417ba3d7ea6b0b06159d72600f77564a7173eac40c25f9b03a9f94a

Observation ff9da692-a220-49af-9a59-947ef4d496a6 · outbound

This paper cites VGGT: Visual Geometry Grounded Transformer , year=.

Chain of Spatial Thoughts: Modality-Agnostic Spatial Grounding for Vision Language Models VGGT: Visual Geometry Grounded Transformer , year=

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-14T04:16:06.230458Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:16:06.230458Z digest=sha256:a5a5d94c2bd9093c969a976fdac5aa5ff48eb28faed6e04f6d1048d50fc13406

Observation 48ce547a-ba16-426a-9594-bc100ccc0cb2 · outbound

This paper cites 2026 , eprint=.

Chain of Spatial Thoughts: Modality-Agnostic Spatial Grounding for Vision Language Models 2026 , eprint=

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:16:07.360021Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-14T04:16:06.234573Z digest=sha256:4bcfbed8aeda71d6c0fe1a8fdb78b1463a0c9cb593ec7943d94df2f03ce55b08

Observation 4ef36d43-f8b2-4c9e-b809-7961530eec03 · outbound

This paper cites 2025 , eprint=.

Chain of Spatial Thoughts: Modality-Agnostic Spatial Grounding for Vision Language Models 2025 , eprint=

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:16:07.319594Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-14T04:16:06.238191Z digest=sha256:3416d3e08af9502f991987822c5f8a62302c307790816095fbb62687a0db145a

Observation 65b64441-4b1b-44b4-998b-a3d0b4f3b6cd · outbound

This paper cites 2025 , eprint=.

Chain of Spatial Thoughts: Modality-Agnostic Spatial Grounding for Vision Language Models 2025 , eprint=

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-14T04:16:06.241882Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:16:06.241882Z digest=sha256:595f6a92bb2b6b822714e102e88ef98f4b8e28ce15db47b40c177f1b7e78127f

Observation d393e972-2053-47dc-b658-906965d3641b · outbound

This paper cites 2023 , editor =.

Chain of Spatial Thoughts: Modality-Agnostic Spatial Grounding for Vision Language Models 2023 , editor =

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-14T04:16:06.245541Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:16:06.245541Z digest=sha256:e5bf714d556e20ce1dee9dfc9828b8c13772214fbc805b7015df34ed52356da5

Observation 5b8be4d3-5cf3-44d4-9520-60cd72021c46 · outbound

This paper cites 2025 , eprint=.

Chain of Spatial Thoughts: Modality-Agnostic Spatial Grounding for Vision Language Models 2025 , eprint=

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-14T04:16:06.249252Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:16:06.249252Z digest=sha256:32a82ac62c601a8136ab5697150c6d3cceb96a98a4ae082ec494153184a95848

Observation 12ce40ad-eff6-4611-bfb6-7e44a38465cb · outbound

This paper cites 2024 , eprint=.

Chain of Spatial Thoughts: Modality-Agnostic Spatial Grounding for Vision Language Models 2024 , eprint=

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-14T04:16:06.252750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:16:06.252750Z digest=sha256:34a356f96f3e6bba62497ffcbda27ff555c780bb624646d918720552aa58a714

Observation 7084bdeb-3698-4f52-98ee-bf46996ba7a0 · outbound

This paper cites 2025 , eprint=.

Chain of Spatial Thoughts: Modality-Agnostic Spatial Grounding for Vision Language Models 2025 , eprint=

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-14T04:16:06.256910Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:16:06.256910Z digest=sha256:ec0044d1f32c81eb638269cf878413eff5da1b61f961c7478c08acfda49e88c7

Observation 82a5ab25-45f0-4b36-bff7-330e0d835fe8 · outbound

This paper cites 2025 , eprint=.

Chain of Spatial Thoughts: Modality-Agnostic Spatial Grounding for Vision Language Models 2025 , eprint=

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-14T04:16:06.260581Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:16:06.260581Z digest=sha256:ed3f24426100c6b545061ec4a4302f947b26e5045060a721a3572ccd8247e6ae

Observation e4856399-78a1-4000-b370-d19c29453c6c · outbound

This paper cites 2025 , eprint=.

Chain of Spatial Thoughts: Modality-Agnostic Spatial Grounding for Vision Language Models 2025 , eprint=

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-14T04:16:06.264998Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:16:06.264998Z digest=sha256:36a90db874b6fd87fa1f411ba769a7bae4dcca5b4f1ce757594560b1f1fa30ea

Observation e382c2bf-45a9-4486-a5fe-728f85e4e734 · outbound

This paper cites arXiv preprint arXiv:2512.16561 , year=.

Chain of Spatial Thoughts: Modality-Agnostic Spatial Grounding for Vision Language Models arXiv preprint arXiv:2512.16561 , year=

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-14T04:16:06.268674Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:16:06.268674Z digest=sha256:35129d6475fd80a99d37034a2f6a8add29764637bc069099dddfa21cc20df120

Observation 5b56c2c9-ad05-4a87-bc8f-c6ae39cffa23 · outbound

This paper cites arXiv preprint arXiv:2511.20648 , year=.

Chain of Spatial Thoughts: Modality-Agnostic Spatial Grounding for Vision Language Models arXiv preprint arXiv:2511.20648 , year=

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-14T04:16:06.272360Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:16:06.272360Z digest=sha256:ec46822a336b3b2c4a46112885041063f761ad4f18ed31a8c5f5c1cd2aa83456

Observation e0ed8563-706a-4356-8dc2-90244b2d4e36 · outbound

This paper cites Abstract 3D Perception for Spatial Intelligence in Vision-Language Models.

Chain of Spatial Thoughts: Modality-Agnostic Spatial Grounding for Vision Language Models Abstract 3D Perception for Spatial Intelligence in Vision-Language Models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-14T04:16:06.276162Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:16:06.276162Z digest=sha256:233cca7f467928fcd58316946c636c8e7e19eb2f8931666cf92f4686adf68e09

Observation 0399d557-9803-45a7-911c-56396225da80 · outbound

This paper cites arXiv preprint arXiv:2603.00409 , year=.

Chain of Spatial Thoughts: Modality-Agnostic Spatial Grounding for Vision Language Models arXiv preprint arXiv:2603.00409 , year=

Reference 38

Resolution
verified exact
raw_fallback, observed 2026-08-14T04:16:06.520498Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-14T04:16:06.280152Z digest=sha256:8cadaab2d03571af9f70980ab94e8c7c9fb3f2f6c2067cc879bf92a25137454f

Observation b7d45d47-2d83-48f4-9b2a-b4435f195216 · outbound

This paper cites arXiv preprint arXiv:2603.05591 , year=.

Chain of Spatial Thoughts: Modality-Agnostic Spatial Grounding for Vision Language Models arXiv preprint arXiv:2603.05591 , year=

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-14T04:16:06.284460Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:16:06.284460Z digest=sha256:517cac964ea3d0a298b2665029bcc9d4afca45e46435be0867f5fb749dbb5ed2

Observation 6618bb78-667c-4abc-81f7-0ae82328503d · outbound

This paper cites European Conference on Computer Vision , pages=.

Chain of Spatial Thoughts: Modality-Agnostic Spatial Grounding for Vision Language Models European Conference on Computer Vision , pages=

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:16:07.100599Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-14T04:16:06.288452Z digest=sha256:dd5852a95b91d1ca652cbd36e110d16489c19b8e7862bf29018953985fab67ec

Observation de951d94-bbea-4c11-aa9e-1b575b7a308d · outbound

This paper cites Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=.

Chain of Spatial Thoughts: Modality-Agnostic Spatial Grounding for Vision Language Models Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-14T04:16:06.292308Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:16:06.292308Z digest=sha256:b1d5d0a363e9983b9f8b570c4ec374128f2c75360f1cff1ff1c399842d21423f

Observation 090025c8-624c-4946-95a9-c39af4443988 · outbound

This paper cites The Thirty-eighth Annual Conference on Neural Information Processing Systems , year=.

Chain of Spatial Thoughts: Modality-Agnostic Spatial Grounding for Vision Language Models The Thirty-eighth Annual Conference on Neural Information Processing Systems , year=

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:16:06.978567Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-14T04:16:06.297394Z digest=sha256:507ce0d71c73264c58910ddb49b7f553006e9bc182d5a3b98003964fd759b42c

Observation 34e5db0e-208c-40d9-92a4-925aa8c461d8 · outbound

This paper cites 2025 , eprint=.

Chain of Spatial Thoughts: Modality-Agnostic Spatial Grounding for Vision Language Models 2025 , eprint=

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:16:06.965836Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-14T04:16:06.301321Z digest=sha256:b3a1ec38345bfdc63c911198a3cb8240c9748fe288041687661ddd18dc7675e1

Observation dea5c195-f580-48aa-b7bd-dc1335f1c166 · outbound

This paper cites 2025 , eprint=.

Chain of Spatial Thoughts: Modality-Agnostic Spatial Grounding for Vision Language Models 2025 , eprint=

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:16:06.916028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-23T06:30:58.430688+00:00.

source=arxiv_source observed=2026-08-14T04:16:06.304920Z digest=sha256:6b53d44fddb844b4a6461b33e04d68ecd531361fa30c3b36f2d7ebe62c2f9ced

Pith citing papers

No inbound Pith citation observations are available.