Pith. sign in

Paper Citation Record · LEDGER

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models

As of 16 August 2026, this Paper Citation Record lists 79 of 79 outbound references and 0 inbound Pith citation observations for arXiv:2608.01899.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.01899 v1

Coverage vector

measured 79 of 79 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T15:08:10.299262Z

measured 79 of 79 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

79 of 79 outbound references displayed

  • verified exact1
  • verified fuzzy7
  • unresolved70
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation b8848310-dcb0-4718-9285-b4e2314cce9c · outbound

This paper cites International conference on machine learning , pages=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models International conference on machine learning , pages=

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:09.994202Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:09.994202Z digest=sha256:6a20acf9a75696c91396543741bcf29a7dd50b516855ccdc5b3fc398a2d0799e

Observation 6f62e481-92ce-444e-944b-4e8c5553c78b · outbound

This paper cites Advances in neural information processing systems , volume=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models Advances in neural information processing systems , volume=

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:09.998452Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:09.998452Z digest=sha256:397cb1530aef6d2169ba8a86bf5ee8044984378a3b3a119a2d7bc536c01063e5

Observation e62aa995-b48d-44d8-863f-fee87493cd0a · outbound

This paper cites IEEE transactions on pattern analysis and machine intelligence , volume=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models IEEE transactions on pattern analysis and machine intelligence , volume=

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.002778Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.002778Z digest=sha256:77df111753e01cebb85a57eca1bfa118285b592b4b579c7054391f2548994e49

Observation 0db17a0d-c4ab-4d8c-882d-16631c3bbf13 · outbound

This paper cites arXiv preprint arXiv:2503.01773 , year=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models arXiv preprint arXiv:2503.01773 , year=

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.007021Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.007021Z digest=sha256:a154808aa0b32164a52b3bd1472b3f136f36a44ab26a4f85eca9e93bc3097451

Observation 3155d4cd-15a0-4ee0-b01d-d95d3c24b4f5 · outbound

This paper cites arXiv preprint arXiv:2503.17349 , year=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models arXiv preprint arXiv:2503.17349 , year=

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.011077Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.011077Z digest=sha256:95cbf3b66c66b2e8beca9afe3ef3c4cffbbefa16bc96a51ded74431aa0bd41ee

Observation 773a7ea6-a1a1-455f-95f3-027f0ee76c11 · outbound

This paper cites arXiv preprint arXiv:2509.18905 , year=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models arXiv preprint arXiv:2509.18905 , year=

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.014809Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.014809Z digest=sha256:d67122bb86314989299ae2160838dc50441810bfe2f3da450a0fe6b6e245f581

Observation 1244430b-6002-4e75-a6b4-b9a56f73e0ba · outbound

This paper cites arXiv preprint arXiv:2509.09332 , year=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models arXiv preprint arXiv:2509.09332 , year=

Reference 7

Resolution
verified exact
raw_fallback, observed 2026-08-15T15:08:11.541670Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T15:08:10.018527Z digest=sha256:2897c4d4fd6578eaf4dc46de921e8dd4f7485dffa297ccb663d3bdf2dd89841c

Observation 38fe9fd8-b4f3-4faa-9345-022aca713ce9 · outbound

This paper cites arXiv preprint arXiv:2506.01946 , year=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models arXiv preprint arXiv:2506.01946 , year=

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.022372Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.022372Z digest=sha256:8b9c1c7f0df06788f2291d92f5ff857b6ac67bf2b705e1f13bfc680d42849f12

Observation 3e4aff90-dad5-4416-b2f4-c9d8184cacdd · outbound

This paper cites arXiv preprint arXiv:2506.04308 , year=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models arXiv preprint arXiv:2506.04308 , year=

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.026194Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.026194Z digest=sha256:5a18b1308c8c9ae686898d80d1431a928d2d633c17c18fab93e434fb868539d0

Observation e2f449e7-3da7-4c1c-b817-317100323fa0 · outbound

This paper cites Proceedings of the IEEE/CVF International Conference on Computer Vision , pages=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models Proceedings of the IEEE/CVF International Conference on Computer Vision , pages=

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.029967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.029967Z digest=sha256:e0f2a5254bfdb7b14596c38ef197117e9e7ed29fe2675d3fc82342800b4f1c84

Observation 3e339f89-744a-4ad4-b61f-fd3cf427ceb1 · outbound

This paper cites arXiv preprint arXiv:2510.13800 , year=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models arXiv preprint arXiv:2510.13800 , year=

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.033848Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.033848Z digest=sha256:3c107ed866ede8ab59c72f5912bdb22a702c4721c0f367fd7d515030069c16e8

Observation 34013dd4-eb19-40aa-aac2-9ca0c4ae3d88 · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models Advances in Neural Information Processing Systems , volume=

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.037808Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.037808Z digest=sha256:8f8a8e95d6768f56ebb0a6709593e274587bf853f909626a9065c1f74003e7ca

Observation 5f047b01-a2f6-4fee-b446-6a8a3089d695 · outbound

This paper cites arXiv preprint arXiv:2505.12448 , year=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models arXiv preprint arXiv:2505.12448 , year=

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.041514Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.041514Z digest=sha256:d25bad326087bbbe29ed31eaf38457e2773686682ceff06ff48b25288e9d1f55

Observation cfbe3632-8440-452b-bec6-eadbb9a5a180 · outbound

This paper cites SpatialVLA: Exploring Spatial Representations for Visual-Language-Action Model.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models SpatialVLA: Exploring Spatial Representations for Visual-Language-Action Model

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.044667Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.044667Z digest=sha256:b73d7969620a0b1a039d47de0726fff13e78a000807508cf2abd74b4ff027b7a

Observation f950af22-c19e-4641-964b-a6bad81814a1 · outbound

This paper cites Proceedings of the Computer Vision and Pattern Recognition Conference , pages=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models Proceedings of the Computer Vision and Pattern Recognition Conference , pages=

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.048406Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.048406Z digest=sha256:0148a1b618a03a817d87fdfe0afc0193fad5b2a61ee7c85b42df79d7cfacb040

Observation 24210930-bfa1-4daa-ad6d-a4f5423e8fcc · outbound

This paper cites VLM-3R: Vision-Language Models Augmented with Instruction-Aligned 3D Reconstruction.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models VLM-3R: Vision-Language Models Augmented with Instruction-Aligned 3D Reconstruction

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.052392Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.052392Z digest=sha256:b6f1833e81b44cddd560a840fa8d824293d7d670533725681b803cda0c0c57a1

Observation e0f52b50-6e78-4608-b6ce-0fa06998de0a · outbound

This paper cites arXiv preprint arXiv:2505.24625 , year=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models arXiv preprint arXiv:2505.24625 , year=

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.056496Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.056496Z digest=sha256:89b51a70f08bba04e80c2c6af4d15aae819a75121aa5d82593626327c6c07a2c

Observation a999f830-30c9-4849-941a-0d633ea2419f · outbound

This paper cites Spatial-MLLM: Boosting MLLM Capabilities in Visual-based Spatial Intelligence.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models Spatial-MLLM: Boosting MLLM Capabilities in Visual-based Spatial Intelligence

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.060200Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.060200Z digest=sha256:9c7ca31dde5a0fa479004ba35e3e0b4737c8735f11105af0ecedbdecefe7505f

Observation 335fa677-d79d-4fe3-9c1c-36bfb7967581 · outbound

This paper cites arXiv preprint arXiv:2510.13375 , year=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models arXiv preprint arXiv:2510.13375 , year=

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.063880Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.063880Z digest=sha256:52c754ffa20822ea861f208dca49eeac7925c0bff479b9f81dae58d6cb56b8f5

Observation 81dfeac8-49f2-44c0-9a32-225ba20cccf7 · outbound

This paper cites arXiv preprint arXiv:2509.25413 , year=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models arXiv preprint arXiv:2509.25413 , year=

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.067422Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.067422Z digest=sha256:a480d4b4a87003031fb2b05189506eca35419be24bcf856d829e28f5033d4204

Observation 929f4769-f1ee-4734-b302-2cf289500171 · outbound

This paper cites Cambrian-S: Towards Spatial Supersensing in Video.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models Cambrian-S: Towards Spatial Supersensing in Video

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.070923Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.070923Z digest=sha256:f7ed4aa65c38f518a160fdde448a93051f8af985eab50d35b2603c693ce240bc

Observation d9b4a329-1382-4eff-a447-0bd6d489075b · outbound

This paper cites Depth Pro: Sharp Monocular Metric Depth in Less Than a Second.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models Depth Pro: Sharp Monocular Metric Depth in Less Than a Second

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.074520Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.074520Z digest=sha256:42b9f55df4de86aec616c52e4d725fc1196fff42eca294671c6043209b90773a

Observation 92017cab-6a06-4fdf-975b-4476c640dd4f · outbound

This paper cites IEEE Transactions on Pattern Analysis and Machine Intelligence , year=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models IEEE Transactions on Pattern Analysis and Machine Intelligence , year=

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.078255Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.078255Z digest=sha256:e3fffffd2107871895279ba0dd51f881f1afc23e95e757fea8ea5a5d3fa408ec

Observation f020c111-2719-4b60-a421-b5466240724e · outbound

This paper cites Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , pages=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , pages=

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.082156Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.082156Z digest=sha256:433b80925eb9f185caa1704486e8ed226ec7b5b3aa06476fd291b4da43dc11e5

Observation fae56157-e09c-46e3-a597-0934b0b9ade6 · outbound

This paper cites 2025 IEEE International Conference on Robotics and Automation (ICRA) , pages=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models 2025 IEEE International Conference on Robotics and Automation (ICRA) , pages=

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.085693Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.085693Z digest=sha256:88c312d92b57597d8250e6a2998c1f9610e2e641df7ead426a8c621dd8831fdb

Observation 25be8de6-2c5d-4fae-8189-9e946d0e4e43 · outbound

This paper cites arXiv preprint arXiv:2511.05491 , year=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models arXiv preprint arXiv:2511.05491 , year=

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.090072Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.090072Z digest=sha256:c1d7fe13058aaa566934233d03bd61392907e84744b45f560905109dbaacf72c

Observation 4978c184-5cfb-43da-8e50-576c13c3bb7f · outbound

This paper cites arXiv preprint arXiv:2511.13719 , year=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models arXiv preprint arXiv:2511.13719 , year=

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.093794Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.093794Z digest=sha256:84092f3b48c0737dfd24feab74d56469bdc10452b79f3883976a9beaff38c594

Observation 33f88e37-dfaa-4b87-b331-0bee34a0b124 · outbound

This paper cites OpenVLA: An Open-Source Vision-Language-Action Model.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models OpenVLA: An Open-Source Vision-Language-Action Model

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.098045Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.098045Z digest=sha256:8406726b7e6c1ac96a1a5c96bbaddeed6b75dd64a583cbe4ffd7eb19239a069f

Observation 52e51b28-043c-48f3-b668-4fd39d6a13b8 · outbound

This paper cites $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control

Reference 29

Resolution
malformed identifier
no resolver link, observed 2026-08-15T15:08:10.102107Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.102107Z digest=sha256:0226a6c074ed7e6711cfa4e24b2ca9062721eee714512d2ad9cd1cf40b469b8f

Observation c7877a9f-6149-4cf2-a13c-15229f3b6ec7 · outbound

This paper cites $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.106096Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.106096Z digest=sha256:4af59e30d21bc0a641e4ce1c283de2e45548fd0b47b7c2794441e7a2301549b8

Observation 9a74ffe6-be85-4d75-93c7-31f500e6c3f1 · outbound

This paper cites GR00T N1: An Open Foundation Model for Generalist Humanoid Robots.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models GR00T N1: An Open Foundation Model for Generalist Humanoid Robots

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.110198Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.110198Z digest=sha256:561a3b1f8e04c6897c58610493682e41417c1d58727d625f2a20121e5e55e63b

Observation 7182c33b-5725-4bcf-a650-ddc8ef32db45 · outbound

This paper cites ChatVLA-2: Vision-Language-Action Model with Open-World Embodied Reasoning from Pretrained Knowledge.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models ChatVLA-2: Vision-Language-Action Model with Open-World Embodied Reasoning from Pretrained Knowledge

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.115509Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.115509Z digest=sha256:7f5ca3ea84b9decb4a72d8774d45c93fa20f0643de45d7cdb50efb8038d2a0ea

Observation f00df1b7-6bd7-4984-bada-03c0ecdbad78 · outbound

This paper cites Proceedings of the Computer Vision and Pattern Recognition Conference , pages=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models Proceedings of the Computer Vision and Pattern Recognition Conference , pages=

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.119439Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.119439Z digest=sha256:fb0e69f4162750afa817f5acb47341cb5bfa8b9a586de359a80b80f474b39d23

Observation 5d4ba3d4-208c-4c33-b0dc-1a591bf7f3dc · outbound

This paper cites Depth Anything 3: Recovering the Visual Space from Any Views.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models Depth Anything 3: Recovering the Visual Space from Any Views

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.122723Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.122723Z digest=sha256:f46fef3c0dbfec0b3bdbaa9589589d4f6d136107e82905699f03cad4a7a6b12a

Observation 76472e7d-95ca-4762-85ee-032562b1cfff · outbound

This paper cites Proceedings of the Computer Vision and Pattern Recognition Conference , pages=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models Proceedings of the Computer Vision and Pattern Recognition Conference , pages=

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.126229Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.126229Z digest=sha256:941c66d2f07211be7c7d95e106b7b02c9a70e1a4ec4cca7cc62b934a800f24bc

Observation 004c95d5-15d1-4e36-97e1-9c5c5c9319ca · outbound

This paper cites arXiv preprint arXiv:2510.13054 , year=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models arXiv preprint arXiv:2510.13054 , year=

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.129455Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.129455Z digest=sha256:f873e76a8299aba64b384120f190c8d6a1b4b1da0ec8a6ff285fef31b9885d6f

Observation d0971804-4448-4522-8f26-f03a8c6f2937 · outbound

This paper cites Fine-Tuning Vision-Language-Action Models: Optimizing Speed and Success.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models Fine-Tuning Vision-Language-Action Models: Optimizing Speed and Success

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.132964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.132964Z digest=sha256:32f8379880d235d39591e5612237daa7d9b1d1c6d003191485b0bf741a429cfe

Observation 3693f6d4-daa7-4dc8-b3e2-2ba944310fb6 · outbound

This paper cites Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:08:11.940877Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T15:08:10.136567Z digest=sha256:8108be2f703e3d6482b9b4412859c7ebd63cf7384f1e1ff423387894b12656a2

Observation d0a7176f-7016-4010-8273-648b330d6392 · outbound

This paper cites Proceedings of the Computer Vision and Pattern Recognition Conference , pages=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models Proceedings of the Computer Vision and Pattern Recognition Conference , pages=

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.139878Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.139878Z digest=sha256:97a1c0ed2f18b245918e28d6fed90515d0fd563bdb469a1b1b43833a175e4f85

Observation 383aa0f8-3988-4f04-bd1e-e965e6b4c36e · outbound

This paper cites Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , pages=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , pages=

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.143408Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.143408Z digest=sha256:0db5adfcabcce1d80bf9e7dcfb79991768459c3f2d7b35d7a15e69e2d2abdf64

Observation a49079e5-35d1-4bf5-8624-5341038c98fe · outbound

This paper cites Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , pages=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , pages=

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.146948Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.146948Z digest=sha256:036db1937bef40a6a4ddc17382175c163a253d41d16b0ab1444a883555bef005

Observation 06e5e7e9-ab29-454e-9fd2-728c23d8783f · outbound

This paper cites Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.151163Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.151163Z digest=sha256:42a09db8f852f83aef74951cd7c5d31a5338c91caba2be62059ceb3e21314113

Observation 9fda7044-ab8f-41f5-b434-b7431ef78512 · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models Advances in Neural Information Processing Systems , volume=

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.154858Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.154858Z digest=sha256:4aede392f4a6e276165b8a55e01005d959d9ada770f186f306beadec027dc959

Observation 86c022b2-abff-4f3c-96bc-b3ec76b18060 · outbound

This paper cites UniVLA: Learning to Act Anywhere with Task-centric Latent Actions.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models UniVLA: Learning to Act Anywhere with Task-centric Latent Actions

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.158321Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.158321Z digest=sha256:6f196db8b965fc59843410fd72cb3e1eca5274930016fa50b50164da1d6c7037

Observation efa3b36c-893b-4835-a838-3ae598fb5470 · outbound

This paper cites SmolVLA: A Vision-Language-Action Model for Affordable and Efficient Robotics.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models SmolVLA: A Vision-Language-Action Model for Affordable and Efficient Robotics

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.161892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.161892Z digest=sha256:65bf88b0e200e7eeb8b339f2d2cbc7d21c75b3bbac36655771995ba36610b13d

Observation 82f7e42a-beb1-4f3c-9716-fd1472c7d891 · outbound

This paper cites arXiv preprint arXiv:2508.18269 , year=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models arXiv preprint arXiv:2508.18269 , year=

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.165360Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.165360Z digest=sha256:f84e2f66d6a91bc9d4fc52cb64335aa5a3d65f9cb3846dd0ee961e14ade1a8c5

Observation 073605da-fa7d-4423-9bc0-590f142dbc87 · outbound

This paper cites Cosmos-Reason1: From Physical Common Sense To Embodied Reasoning.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models Cosmos-Reason1: From Physical Common Sense To Embodied Reasoning

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.169529Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.169529Z digest=sha256:5a7f26ecd16659702be04d4361b65b2c6f7885c5cd514da288a242d0b99c6e99

Observation 32085fe6-0ef5-4750-930a-518898088f9d · outbound

This paper cites Qwen2.5-VL Technical Report.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models Qwen2.5-VL Technical Report

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.173432Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.173432Z digest=sha256:94c4cd561bafb9fc69ca2e813a90e672ffcd972e82a69a8c1f3d098dfbd9adab

Observation 5d0ea939-7a64-4072-9dc8-aaf2c4727e86 · outbound

This paper cites 2025 , eprint=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models 2025 , eprint=

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.176743Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.176743Z digest=sha256:aaccc11e800d3850a16966a28b13929ffccb1283ca9f9ee3e14f5fd779218047

Observation 6081fc44-87e2-4eed-a87b-3473cc8adba3 · outbound

This paper cites InternVL3.5: Advancing Open-Source Multimodal Models in Versatility, Reasoning, and Efficiency.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models InternVL3.5: Advancing Open-Source Multimodal Models in Versatility, Reasoning, and Efficiency

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.180305Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.180305Z digest=sha256:286eef40b6c1dc5ad50af064f50965ba1fec05308095d55ff5bd02265cf97f1c

Observation 242df485-fd2b-4be4-81eb-7d97fb2e1940 · outbound

This paper cites LLaVA-NeXT: A Strong Zero-shot Video Understanding Model , url=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models LLaVA-NeXT: A Strong Zero-shot Video Understanding Model , url=

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.184135Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.184135Z digest=sha256:05ebc8a51ba0949a26528a3025991024f1d24048daa2bc1f9d08ce48bda6659c

Observation 989e9969-7e7c-4081-9049-498ef2b9efb1 · outbound

This paper cites 2025 , howpublished =.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models 2025 , howpublished =

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.187614Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.187614Z digest=sha256:e5c8ca764e83e5b01a65f50c07c988a49c74398a8fb3ffd7b0aded83af3d2fb2

Observation 782bbbab-77f7-4c03-a458-bf49910d53c5 · outbound

This paper cites Seed1.5-VL Technical Report.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models Seed1.5-VL Technical Report

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.191584Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.191584Z digest=sha256:7b3396c4bedc0f170f9983aaaf60bf4d0d98dab5bd549f45c197b41165ad412b

Observation 28dad9aa-0c89-4f2d-bb70-bd1bdd1007ab · outbound

This paper cites Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.195853Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.195853Z digest=sha256:25ac4d8d12d48158971f8ce747d31b8c764f3ee8fbc2003dbe0f84c8a4d6bbc6

Observation e37a52df-66ab-499a-93fe-2c4c00a5651c · outbound

This paper cites MiMo-Embodied: X-Embodied Foundation Model Technical Report.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models MiMo-Embodied: X-Embodied Foundation Model Technical Report

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.200174Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.200174Z digest=sha256:74edf681a2c06028830bb6114b3c5bebd0e4e5a5705f0590b7823dc38e177c57

Observation 81b17312-3105-4e6f-a28b-c51687e73922 · outbound

This paper cites 2025 , eprint=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models 2025 , eprint=

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.204796Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.204796Z digest=sha256:60c84da22fa72401e7756fab476cca8b04d93be2eb85e935d9ad3c19521e2f85

Observation e072c118-4f45-4919-af2d-8716e6292995 · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models Advances in Neural Information Processing Systems , volume=

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.208183Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.208183Z digest=sha256:1f6303c75aa7300d246c2eca7b83c3a341b37642ca3038ace7618996a40c5d4b

Observation 6721cb0d-8e0d-434c-ac64-2fc0cdbbf464 · outbound

This paper cites Proceedings of the IEEE conference on computer vision and pattern recognition , pages=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models Proceedings of the IEEE conference on computer vision and pattern recognition , pages=

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.211361Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.211361Z digest=sha256:c59b7cfd6375dab62a17bbb1c77fd7aa335a6e91e023c8105dd96701da3620da

Observation 0b7e0b1b-2c81-489c-9cc7-fb096fb74d41 · outbound

This paper cites Proceedings of the IEEE/CVF International Conference on Computer Vision , pages=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models Proceedings of the IEEE/CVF International Conference on Computer Vision , pages=

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.214609Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.214609Z digest=sha256:67a60439ab85f4c7a6b86a504472dfee732e7cd3057185b72053470c7d2bfe24

Observation 45ba04f5-5a83-40c6-9ae9-93ddeadf835f · outbound

This paper cites ARKitScenes: A Diverse Real-World Dataset For 3D Indoor Scene Understanding Using Mobile RGB-D Data.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models ARKitScenes: A Diverse Real-World Dataset For 3D Indoor Scene Understanding Using Mobile RGB-D Data

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.218090Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.218090Z digest=sha256:0686edfdc56a247a8cdd1fe77d58c6f68b18cd6c5e7d4f46f8e821d3eeb0c436

Observation b63cb3c8-5df9-40c9-b0ad-ca7cd23ad943 · outbound

This paper cites Proceedings of the IEEE/CVF international conference on computer vision , pages=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models Proceedings of the IEEE/CVF international conference on computer vision , pages=

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.221464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.221464Z digest=sha256:5fb651f935acadfc500df19a8f29a97827a25e6bb012721d8c030cc833f4e44c

Observation 39bf2a94-bfbe-4563-bbf9-516092352eac · outbound

This paper cites Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.224769Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.224769Z digest=sha256:c2f91713411738b9f1aa82404a042c7e8802aaae2185a130e8b4b53add8f3984

Observation f25691ba-d002-4b08-b370-48ffeaa1d0cc · outbound

This paper cites Proceedings of the IEEE conference on computer vision and pattern recognition , pages=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models Proceedings of the IEEE conference on computer vision and pattern recognition , pages=

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.228209Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.228209Z digest=sha256:b18d7c19c7ca195d779cbb98c5363e520df1ab4b88cd4c80b3daf060ec1298d2

Observation 2b2607a5-c18a-45a8-a910-c61efb6508ac · outbound

This paper cites Virtual KITTI 2.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models Virtual KITTI 2

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.232197Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.232197Z digest=sha256:cd559410df796e6a82b3b07eeda335bb812fe63d17e8a7d061df709fcf3b8114

Observation 1c97c294-4d7f-4147-b2f8-df8ad2e07620 · outbound

This paper cites SQA3D: Situated Question Answering in 3D Scenes.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models SQA3D: Situated Question Answering in 3D Scenes

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.235780Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.235780Z digest=sha256:416b11c036aef53a92888d619323aa37f95ae4321611d38837b81dcbe948490a

Observation 4dd5da3b-bee4-41f6-9735-66bf3dabc150 · outbound

This paper cites proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.239710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.239710Z digest=sha256:076c2a94294a041f09832b870f43c5efb76878778761191d7d049e29b52e2940

Observation b11e7624-cf49-4813-8575-fd2ee1887efa · outbound

This paper cites Proceedings of the IEEE conference on computer vision and pattern recognition , pages=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models Proceedings of the IEEE conference on computer vision and pattern recognition , pages=

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.242927Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.242927Z digest=sha256:6825b451156df9c271dd237648d52741ec979c6475be7459a4f34a05b7024ec6

Observation 767ad612-a545-47d4-b4c1-ec951e3eafdd · outbound

This paper cites European conference on computer vision , pages=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models European conference on computer vision , pages=

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.246223Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.246223Z digest=sha256:a168bfe8db6f3f5fa3bb2769baaf9f93aa5b759be13a159f8f4eb0b4449d82b3

Observation 433131b0-99ba-4a91-bf5f-cf0eeba3d844 · outbound

This paper cites Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , pages=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , pages=

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:08:11.825372Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T15:08:10.249673Z digest=sha256:09ea64a418b277a950cae22a4a65979b60d81305820baf976b79528e21ea32ea

Observation 5c4b15ae-f0af-4320-a5bf-a57c4b1eb904 · outbound

This paper cites The international journal of robotics research , volume=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models The international journal of robotics research , volume=

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.252980Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.252980Z digest=sha256:4a2cc165b1e5beadb0a9a0cc37cfcf7710aeff02a4eea46bc0b73df1d9b0ce64

Observation 502c1cb8-27c7-4b61-9345-9ae5154a4e49 · outbound

This paper cites Proceedings of the IEEE conference on computer vision and pattern recognition , pages=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models Proceedings of the IEEE conference on computer vision and pattern recognition , pages=

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.256437Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.256437Z digest=sha256:9e4f2adc2bbc559d3b488a23799f90052f06eab9bbc97b24b65ddc3a5da7c3ec

Observation 3d1a9da0-79e6-43c9-bdb6-66e5e438e00a · outbound

This paper cites GPT-4o System Card.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models GPT-4o System Card

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.260290Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.260290Z digest=sha256:d3e092a3d9a711bb9349692a6d50ece2f6b533e834ca53f6e12d76e4d625798d

Observation 03b360a9-2ceb-418e-81d8-bf2f27c60d57 · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models Advances in Neural Information Processing Systems , volume=

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:08:11.802754Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T15:08:10.263361Z digest=sha256:e8e2ad0a0aed1db50a409413a9c2ccd7d21c4bd63716742381a923dff749ef2f

Observation 7b36932b-0625-4e4b-91b5-50ae042171d8 · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models Advances in Neural Information Processing Systems , volume=

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:08:11.791941Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T15:08:10.267499Z digest=sha256:a4849a6e03152a683ffb1ffe2e31b2478fe5c95b352622cac055e88cef91ec75

Observation 7b776672-c1e3-4c52-9f86-ff88c6b9befe · outbound

This paper cites Proceedings of the IEEE/CVF international conference on computer vision , pages=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models Proceedings of the IEEE/CVF international conference on computer vision , pages=

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:08:11.781165Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T15:08:10.271044Z digest=sha256:f8a70d5baf93896f4ec36f521b0674da93c4b6fe3b74bad20ea4f22930d15755

Observation 6a23ec32-603d-4c59-b7b7-070b9206706b · outbound

This paper cites Proceedings of the Computer Vision and Pattern Recognition Conference , pages=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models Proceedings of the Computer Vision and Pattern Recognition Conference , pages=

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:08:11.770220Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T15:08:10.286451Z digest=sha256:e3f03d72491a086992a93af258bac138fd85b400e71730c244e4d9251096dd99

Observation 64c875bf-ea2f-40e5-b7c3-77e3ad16fc07 · outbound

This paper cites Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.290790Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.290790Z digest=sha256:16cce175ab47f61b7e5930a65f007e49456a0635fb153b4548b469b8e017ca62

Observation 47d221e8-4b8a-404b-a578-206db2051555 · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models Advances in Neural Information Processing Systems , volume=

Reference 78

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:08:11.753395Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=arxiv_source observed=2026-08-15T15:08:10.295112Z digest=sha256:d458fb34cd28de38c2c2202614453aacd4749c96c6753484c423cb776f291183

Observation 32449b10-0cda-41e4-b507-e6c0d2b2ca1b · outbound

This paper cites arXiv preprint arXiv:2512.08889 , year=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models arXiv preprint arXiv:2512.08889 , year=

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.299262Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.299262Z digest=sha256:55e6969d3d43878160d84c9a6d9adee0aa36d5bc89d0349050e7e597e9cec016

Pith citing papers

No inbound Pith citation observations are available.