Pith. sign in

Paper Citation Record · LEDGER

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models

As of 15 August 2026, this Paper Citation Record lists 79 of 79 outbound references and 0 inbound Pith citation observations for arXiv:2608.01899.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.01899 v1

Coverage vector

measured 79 of 79 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T15:08:10.299262Z

measured 79 of 79 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

79 of 79 outbound references displayed

  • verified exact1
  • verified fuzzy7
  • unresolved70
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation b8848310-dcb0-4718-9285-b4e2314cce9c · outbound

This paper cites International conference on machine learning , pages=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models International conference on machine learning , pages=

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:09.994202Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:09.994202Z digest=sha256:e933abbd2d37eede9d58557faf4d0c9e62d0890cbd4ea0dd37035a5b205408c4

Observation 6f62e481-92ce-444e-944b-4e8c5553c78b · outbound

This paper cites Advances in neural information processing systems , volume=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models Advances in neural information processing systems , volume=

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:09.998452Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:09.998452Z digest=sha256:6bfc3ce74f89e4133d9b08d124bc251db13745f782fd89fba24be5301cf106d3

Observation e62aa995-b48d-44d8-863f-fee87493cd0a · outbound

This paper cites IEEE transactions on pattern analysis and machine intelligence , volume=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models IEEE transactions on pattern analysis and machine intelligence , volume=

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.002778Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.002778Z digest=sha256:4090a9477a6ae3080f6f1313cb08a531d7ae2ecde4767a292fe3ca459f9bc7a5

Observation 0db17a0d-c4ab-4d8c-882d-16631c3bbf13 · outbound

This paper cites arXiv preprint arXiv:2503.01773 , year=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models arXiv preprint arXiv:2503.01773 , year=

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.007021Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.007021Z digest=sha256:39521c840487ab31e63147cfd828f7f60e845589a194a5819c5d5f097225aaa7

Observation 3155d4cd-15a0-4ee0-b01d-d95d3c24b4f5 · outbound

This paper cites arXiv preprint arXiv:2503.17349 , year=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models arXiv preprint arXiv:2503.17349 , year=

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.011077Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.011077Z digest=sha256:df0efdf2ea7a3bbe2a745540f4f98e793ee114035ac7df9c6f7e75357977e1ca

Observation 773a7ea6-a1a1-455f-95f3-027f0ee76c11 · outbound

This paper cites arXiv preprint arXiv:2509.18905 , year=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models arXiv preprint arXiv:2509.18905 , year=

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.014809Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.014809Z digest=sha256:38fa5a0db2ca3fb3498cbedc107730ef914208a00b49b93227ba8278d1e13e88

Observation 1244430b-6002-4e75-a6b4-b9a56f73e0ba · outbound

This paper cites arXiv preprint arXiv:2509.09332 , year=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models arXiv preprint arXiv:2509.09332 , year=

Reference 7

Resolution
verified exact
raw_fallback, observed 2026-08-15T15:08:11.541670Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-15T15:08:10.018527Z digest=sha256:f5601af4c734f84c1dc748f639f079575d76277ad8b2322c5dc74e3eee94579a

Observation 38fe9fd8-b4f3-4faa-9345-022aca713ce9 · outbound

This paper cites arXiv preprint arXiv:2506.01946 , year=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models arXiv preprint arXiv:2506.01946 , year=

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.022372Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.022372Z digest=sha256:6265e51f179895f14c96c5336c14b46824b937fdd5c9978a65c09c51e6bc4083

Observation 3e4aff90-dad5-4416-b2f4-c9d8184cacdd · outbound

This paper cites arXiv preprint arXiv:2506.04308 , year=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models arXiv preprint arXiv:2506.04308 , year=

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.026194Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.026194Z digest=sha256:f4069bbdfd97e2d7f47edd1a20b1a59b8d36fbb4a716d15fdc3aae9a8efb9a88

Observation e2f449e7-3da7-4c1c-b817-317100323fa0 · outbound

This paper cites Proceedings of the IEEE/CVF International Conference on Computer Vision , pages=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models Proceedings of the IEEE/CVF International Conference on Computer Vision , pages=

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.029967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.029967Z digest=sha256:62a4a3bffad8b52b0f26bc47637f3c7c5519340e1cb915d015efe090e1101f07

Observation 3e339f89-744a-4ad4-b61f-fd3cf427ceb1 · outbound

This paper cites arXiv preprint arXiv:2510.13800 , year=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models arXiv preprint arXiv:2510.13800 , year=

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.033848Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.033848Z digest=sha256:642b0f06684ea5ab378da85aeec1d61923ae04d87021866318c23a252d3ad520

Observation 34013dd4-eb19-40aa-aac2-9ca0c4ae3d88 · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models Advances in Neural Information Processing Systems , volume=

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.037808Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.037808Z digest=sha256:cbfce78b725e4dac730bc99acf94b896d9bc3f02e461232ae242671956727b21

Observation 5f047b01-a2f6-4fee-b446-6a8a3089d695 · outbound

This paper cites arXiv preprint arXiv:2505.12448 , year=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models arXiv preprint arXiv:2505.12448 , year=

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.041514Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.041514Z digest=sha256:7fdd79452c9977d9f5ca110c7c42833fb8d363831ae677aa05ba3df7f830e8cd

Observation cfbe3632-8440-452b-bec6-eadbb9a5a180 · outbound

This paper cites SpatialVLA: Exploring Spatial Representations for Visual-Language-Action Model.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models SpatialVLA: Exploring Spatial Representations for Visual-Language-Action Model

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.044667Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.044667Z digest=sha256:62335c0f23ec510c39e7820160b49e58e4a7171317bdfd7e5f3ab61e838bb3bd

Observation f950af22-c19e-4641-964b-a6bad81814a1 · outbound

This paper cites Proceedings of the Computer Vision and Pattern Recognition Conference , pages=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models Proceedings of the Computer Vision and Pattern Recognition Conference , pages=

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.048406Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.048406Z digest=sha256:ed9674a6d1e2f848ee887640dc5bae642366f47c36d7aeace181ea69a36c6651

Observation 24210930-bfa1-4daa-ad6d-a4f5423e8fcc · outbound

This paper cites VLM-3R: Vision-Language Models Augmented with Instruction-Aligned 3D Reconstruction.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models VLM-3R: Vision-Language Models Augmented with Instruction-Aligned 3D Reconstruction

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.052392Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.052392Z digest=sha256:f1435c9b7b57324218a0320bea22acf3b876b1f77dda53b08deda96e00c0521b

Observation e0f52b50-6e78-4608-b6ce-0fa06998de0a · outbound

This paper cites arXiv preprint arXiv:2505.24625 , year=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models arXiv preprint arXiv:2505.24625 , year=

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.056496Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.056496Z digest=sha256:9c15d6f72c6340b77de08a2cf83c933122d19777d3fac60680b3f968ec2d1682

Observation a999f830-30c9-4849-941a-0d633ea2419f · outbound

This paper cites Spatial-MLLM: Boosting MLLM Capabilities in Visual-based Spatial Intelligence.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models Spatial-MLLM: Boosting MLLM Capabilities in Visual-based Spatial Intelligence

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.060200Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.060200Z digest=sha256:773b68ee31020bcae549c81abbdeb5ef228d5ec54586bcfb06cfbe6b428c7491

Observation 335fa677-d79d-4fe3-9c1c-36bfb7967581 · outbound

This paper cites arXiv preprint arXiv:2510.13375 , year=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models arXiv preprint arXiv:2510.13375 , year=

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.063880Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.063880Z digest=sha256:ca99d5f20063dd24070f90d85370caeabfe678f6c309dd8ac8325a39ab8caf11

Observation 81dfeac8-49f2-44c0-9a32-225ba20cccf7 · outbound

This paper cites arXiv preprint arXiv:2509.25413 , year=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models arXiv preprint arXiv:2509.25413 , year=

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.067422Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.067422Z digest=sha256:02a161cccf16adf1ab9344086d48b5b26dc119816b6a4b16cd68dd324b9517e4

Observation 929f4769-f1ee-4734-b302-2cf289500171 · outbound

This paper cites Cambrian-S: Towards Spatial Supersensing in Video.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models Cambrian-S: Towards Spatial Supersensing in Video

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.070923Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.070923Z digest=sha256:1a45ecbb69b19bd6c54384c27d88d74f41c3081c5d97ddee7b0b7a473848534a

Observation d9b4a329-1382-4eff-a447-0bd6d489075b · outbound

This paper cites Depth Pro: Sharp Monocular Metric Depth in Less Than a Second.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models Depth Pro: Sharp Monocular Metric Depth in Less Than a Second

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.074520Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.074520Z digest=sha256:437a0ab32e80fc03ab8bc4928829cc2da900c68e8509761cb1987b94ff4800b4

Observation 92017cab-6a06-4fdf-975b-4476c640dd4f · outbound

This paper cites IEEE Transactions on Pattern Analysis and Machine Intelligence , year=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models IEEE Transactions on Pattern Analysis and Machine Intelligence , year=

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.078255Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.078255Z digest=sha256:a604c9cc0f3294e3bb91209f7ed241b5b55056f3b16b893c464dd922b7433a2e

Observation f020c111-2719-4b60-a421-b5466240724e · outbound

This paper cites Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , pages=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , pages=

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.082156Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.082156Z digest=sha256:46ae923fdf970c3e21f32e368595237c22beeb0724e9706f5f5b6d9bc49f2af9

Observation fae56157-e09c-46e3-a597-0934b0b9ade6 · outbound

This paper cites 2025 IEEE International Conference on Robotics and Automation (ICRA) , pages=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models 2025 IEEE International Conference on Robotics and Automation (ICRA) , pages=

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.085693Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.085693Z digest=sha256:7479cfbc167e68d8bf7cccc40c1b828545c3bda887f0f43219e946bea77c832e

Observation 25be8de6-2c5d-4fae-8189-9e946d0e4e43 · outbound

This paper cites arXiv preprint arXiv:2511.05491 , year=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models arXiv preprint arXiv:2511.05491 , year=

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.090072Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.090072Z digest=sha256:9dffcca17289d56759c9fd794b688fc90819f8c097c23aa818fef35f2ab6051e

Observation 4978c184-5cfb-43da-8e50-576c13c3bb7f · outbound

This paper cites arXiv preprint arXiv:2511.13719 , year=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models arXiv preprint arXiv:2511.13719 , year=

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.093794Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.093794Z digest=sha256:69b287ac9f8e7daaaa2cab0700fc5decdcff3b8f9e9944a6f4ebe77b748392e3

Observation 33f88e37-dfaa-4b87-b331-0bee34a0b124 · outbound

This paper cites OpenVLA: An Open-Source Vision-Language-Action Model.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models OpenVLA: An Open-Source Vision-Language-Action Model

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.098045Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.098045Z digest=sha256:31f42ecaa5878ce3a6aef21efc586b978d14f73b8717d63e44acadd530aad88f

Observation 52e51b28-043c-48f3-b668-4fd39d6a13b8 · outbound

This paper cites $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models $\pi_0$: A Vision-Language-Action Flow Model for General Robot Control

Reference 29

Resolution
malformed identifier
no resolver link, observed 2026-08-15T15:08:10.102107Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.102107Z digest=sha256:3f241f5853aa1b978643d33a79da13e3395b290c54cb3f1d4dc742238a121ebe

Observation c7877a9f-6149-4cf2-a13c-15229f3b6ec7 · outbound

This paper cites $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models $\pi_{0.5}$: a Vision-Language-Action Model with Open-World Generalization

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.106096Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.106096Z digest=sha256:5f07c907fa33216ea583c80505258d39d6642064497c145d5e3f4b169512e390

Observation 9a74ffe6-be85-4d75-93c7-31f500e6c3f1 · outbound

This paper cites GR00T N1: An Open Foundation Model for Generalist Humanoid Robots.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models GR00T N1: An Open Foundation Model for Generalist Humanoid Robots

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.110198Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.110198Z digest=sha256:994f0edf7a9625afab9ab793fb493c8dda98f7009228dac993cf6e89402ab794

Observation 7182c33b-5725-4bcf-a650-ddc8ef32db45 · outbound

This paper cites ChatVLA-2: Vision-Language-Action Model with Open-World Embodied Reasoning from Pretrained Knowledge.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models ChatVLA-2: Vision-Language-Action Model with Open-World Embodied Reasoning from Pretrained Knowledge

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.115509Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.115509Z digest=sha256:87462e82ee36cbd6c82afc3506b7b19ceb406facec171e0f731388c75a4aa50b

Observation f00df1b7-6bd7-4984-bada-03c0ecdbad78 · outbound

This paper cites Proceedings of the Computer Vision and Pattern Recognition Conference , pages=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models Proceedings of the Computer Vision and Pattern Recognition Conference , pages=

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.119439Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.119439Z digest=sha256:23d2dc1d46595f76f894423e2021940569a0649c5922693729616d97557282f9

Observation 5d4ba3d4-208c-4c33-b0dc-1a591bf7f3dc · outbound

This paper cites Depth Anything 3: Recovering the Visual Space from Any Views.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models Depth Anything 3: Recovering the Visual Space from Any Views

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.122723Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.122723Z digest=sha256:deefacd017bca17bf58d10d2ec5828ce4a5d9cf3ccca4f12f9ce3fbfd2c048cf

Observation 76472e7d-95ca-4762-85ee-032562b1cfff · outbound

This paper cites Proceedings of the Computer Vision and Pattern Recognition Conference , pages=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models Proceedings of the Computer Vision and Pattern Recognition Conference , pages=

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.126229Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.126229Z digest=sha256:d8706ca7952967582b6e15fa5a9176c2538a133d9706d8ae0be73b3feb1bed42

Observation 004c95d5-15d1-4e36-97e1-9c5c5c9319ca · outbound

This paper cites arXiv preprint arXiv:2510.13054 , year=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models arXiv preprint arXiv:2510.13054 , year=

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.129455Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.129455Z digest=sha256:6ab7ccbad517909eede1a8c5f0d39e7394df9125e37323e602214d2ba90d3d5e

Observation d0971804-4448-4522-8f26-f03a8c6f2937 · outbound

This paper cites Fine-Tuning Vision-Language-Action Models: Optimizing Speed and Success.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models Fine-Tuning Vision-Language-Action Models: Optimizing Speed and Success

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.132964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.132964Z digest=sha256:5104baba25ba709d5805ba8c9e7e9e7941cb56b0fda98b7d8ee2dc525626e413

Observation 3693f6d4-daa7-4dc8-b3e2-2ba944310fb6 · outbound

This paper cites Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:08:11.940877Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-15T15:08:10.136567Z digest=sha256:aaed9d2d551d974dc5ee215271cbb3e084c7ae3f3534575d7c3e5198d25bd5a3

Observation d0a7176f-7016-4010-8273-648b330d6392 · outbound

This paper cites Proceedings of the Computer Vision and Pattern Recognition Conference , pages=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models Proceedings of the Computer Vision and Pattern Recognition Conference , pages=

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.139878Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.139878Z digest=sha256:d963d261cf8ebdf054017bfcac4af1a82d9ea9d77bc666ebfe68b923881f7cd6

Observation 383aa0f8-3988-4f04-bd1e-e965e6b4c36e · outbound

This paper cites Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , pages=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , pages=

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.143408Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.143408Z digest=sha256:f00903b9ab3729027dee65c6ecb914a66e6066e010e73956fd868a2bcb4ef9ef

Observation a49079e5-35d1-4bf5-8624-5341038c98fe · outbound

This paper cites Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , pages=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , pages=

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.146948Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.146948Z digest=sha256:2f65a462daf38c89eacae5215bd692744c38054cdb13fd76e0fac26e2777ef0c

Observation 06e5e7e9-ab29-454e-9fd2-728c23d8783f · outbound

This paper cites Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.151163Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.151163Z digest=sha256:e124968e56d27e3a572c3c92ab618773026de2457b4577416d507c6f7b9435f3

Observation 9fda7044-ab8f-41f5-b434-b7431ef78512 · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models Advances in Neural Information Processing Systems , volume=

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.154858Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.154858Z digest=sha256:e175ef8ecaffad9d69941441901fef085402bd4d3917f8e046f1d1ed86820c4a

Observation 86c022b2-abff-4f3c-96bc-b3ec76b18060 · outbound

This paper cites UniVLA: Learning to Act Anywhere with Task-centric Latent Actions.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models UniVLA: Learning to Act Anywhere with Task-centric Latent Actions

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.158321Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.158321Z digest=sha256:9a95198994e2fde1606fa82961f24f88fa7ea5176d7f4b8cf0d3b1367a6728ec

Observation efa3b36c-893b-4835-a838-3ae598fb5470 · outbound

This paper cites SmolVLA: A Vision-Language-Action Model for Affordable and Efficient Robotics.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models SmolVLA: A Vision-Language-Action Model for Affordable and Efficient Robotics

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.161892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.161892Z digest=sha256:fe1353c58b0e88ebc1b52d7b98b26688f18b70a89c49b76395439686f9906ec0

Observation 82f7e42a-beb1-4f3c-9716-fd1472c7d891 · outbound

This paper cites arXiv preprint arXiv:2508.18269 , year=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models arXiv preprint arXiv:2508.18269 , year=

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.165360Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.165360Z digest=sha256:b011b6a9b21b14e39e79fe7e5314ab8f07ad0cac609316253b68254242f92e6a

Observation 073605da-fa7d-4423-9bc0-590f142dbc87 · outbound

This paper cites Cosmos-Reason1: From Physical Common Sense To Embodied Reasoning.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models Cosmos-Reason1: From Physical Common Sense To Embodied Reasoning

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.169529Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.169529Z digest=sha256:40e6ed6516932c9931f42948a25278de0a26d4519d71cb11039a2d3dd8312c5e

Observation 32085fe6-0ef5-4750-930a-518898088f9d · outbound

This paper cites Qwen2.5-VL Technical Report.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models Qwen2.5-VL Technical Report

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.173432Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.173432Z digest=sha256:75655c99d1df0530542a82d0a3acb6c058e2b9fc54ceaa89019420768e829c07

Observation 5d0ea939-7a64-4072-9dc8-aaf2c4727e86 · outbound

This paper cites 2025 , eprint=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models 2025 , eprint=

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.176743Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.176743Z digest=sha256:2876835612a7878ecec392d9b429192c50295054735699bd32a23353336ff94e

Observation 6081fc44-87e2-4eed-a87b-3473cc8adba3 · outbound

This paper cites InternVL3.5: Advancing Open-Source Multimodal Models in Versatility, Reasoning, and Efficiency.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models InternVL3.5: Advancing Open-Source Multimodal Models in Versatility, Reasoning, and Efficiency

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.180305Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.180305Z digest=sha256:833f0a3df8aaa748275050b3eaabfb69be335a2d5ca93b3b700d477e227f420f

Observation 242df485-fd2b-4be4-81eb-7d97fb2e1940 · outbound

This paper cites LLaVA-NeXT: A Strong Zero-shot Video Understanding Model , url=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models LLaVA-NeXT: A Strong Zero-shot Video Understanding Model , url=

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.184135Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.184135Z digest=sha256:27afc9b113330dacbaf3a7ebe77081d21d82cbbefd566e3b4943ad2bd9b981e4

Observation 989e9969-7e7c-4081-9049-498ef2b9efb1 · outbound

This paper cites 2025 , howpublished =.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models 2025 , howpublished =

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.187614Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.187614Z digest=sha256:20e95892b918fde24456b605c3bac19ec7f99db278c6a5e9eb06f1cc84abb6e0

Observation 782bbbab-77f7-4c03-a458-bf49910d53c5 · outbound

This paper cites Seed1.5-VL Technical Report.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models Seed1.5-VL Technical Report

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.191584Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.191584Z digest=sha256:db6f4eb1888852f21d47bec15f7a338d1022e727918b8ffb5561e1e341955f12

Observation 28dad9aa-0c89-4f2d-bb70-bd1bdd1007ab · outbound

This paper cites Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.195853Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.195853Z digest=sha256:5ce35a66e922233f149172366ce125d3c3d12885ed093b05a2676413773bc8ab

Observation e37a52df-66ab-499a-93fe-2c4c00a5651c · outbound

This paper cites MiMo-Embodied: X-Embodied Foundation Model Technical Report.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models MiMo-Embodied: X-Embodied Foundation Model Technical Report

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.200174Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.200174Z digest=sha256:570217e097280cc36a1a140dda652742b9512d1538429ae0a39f7284da8f2588

Observation 81b17312-3105-4e6f-a28b-c51687e73922 · outbound

This paper cites 2025 , eprint=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models 2025 , eprint=

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.204796Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.204796Z digest=sha256:442b32bcebeb53349159fbbc5777b450baa14b15328c67d244bdb6ed1ce30d66

Observation e072c118-4f45-4919-af2d-8716e6292995 · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models Advances in Neural Information Processing Systems , volume=

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.208183Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.208183Z digest=sha256:5349969a107c4ec2c3b18dcc9b61028decfc25b0b426f37e7bca07c5944a3c91

Observation 6721cb0d-8e0d-434c-ac64-2fc0cdbbf464 · outbound

This paper cites Proceedings of the IEEE conference on computer vision and pattern recognition , pages=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models Proceedings of the IEEE conference on computer vision and pattern recognition , pages=

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.211361Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.211361Z digest=sha256:55d4b659315f82e1c6c4b620069d6e5d17d3ff04dc4ca288d2310e13af1229be

Observation 0b7e0b1b-2c81-489c-9cc7-fb096fb74d41 · outbound

This paper cites Proceedings of the IEEE/CVF International Conference on Computer Vision , pages=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models Proceedings of the IEEE/CVF International Conference on Computer Vision , pages=

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.214609Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.214609Z digest=sha256:ac11b14e5867a28391a5bb485ca9932ada05f214e4e910e752f211dd11c685c4

Observation 45ba04f5-5a83-40c6-9ae9-93ddeadf835f · outbound

This paper cites ARKitScenes: A Diverse Real-World Dataset For 3D Indoor Scene Understanding Using Mobile RGB-D Data.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models ARKitScenes: A Diverse Real-World Dataset For 3D Indoor Scene Understanding Using Mobile RGB-D Data

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.218090Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.218090Z digest=sha256:dfc87c1f2aca5a49287fc6b87b32ecd6c3481a244b450b2d0735be361d17b4bb

Observation b63cb3c8-5df9-40c9-b0ad-ca7cd23ad943 · outbound

This paper cites Proceedings of the IEEE/CVF international conference on computer vision , pages=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models Proceedings of the IEEE/CVF international conference on computer vision , pages=

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.221464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.221464Z digest=sha256:27395fdc800703021ce2f7048f9567b3b380cdc7e09ff2a2b9b24095bd9dff7d

Observation 39bf2a94-bfbe-4563-bbf9-516092352eac · outbound

This paper cites Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.224769Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.224769Z digest=sha256:f04db069f17df5b53798cd1f5009c9a16fab367405ec5af92b1c41cf85cd8e8d

Observation f25691ba-d002-4b08-b370-48ffeaa1d0cc · outbound

This paper cites Proceedings of the IEEE conference on computer vision and pattern recognition , pages=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models Proceedings of the IEEE conference on computer vision and pattern recognition , pages=

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.228209Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.228209Z digest=sha256:ff64244382803aaec1d0766e0597b1b331d0e57e8e01d46c87ee1281e37f1dd5

Observation 2b2607a5-c18a-45a8-a910-c61efb6508ac · outbound

This paper cites Virtual KITTI 2.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models Virtual KITTI 2

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.232197Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.232197Z digest=sha256:66262630543a04cf74d2799b0888280ae46dc73de91e94d7b770bbf6b330a66b

Observation 1c97c294-4d7f-4147-b2f8-df8ad2e07620 · outbound

This paper cites SQA3D: Situated Question Answering in 3D Scenes.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models SQA3D: Situated Question Answering in 3D Scenes

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.235780Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.235780Z digest=sha256:38a8d2463fbbcfd6e71d0df807c7c4436fd0f94b4ad0092a92e9dda8b530e38b

Observation 4dd5da3b-bee4-41f6-9735-66bf3dabc150 · outbound

This paper cites proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.239710Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.239710Z digest=sha256:fd11a24f65db9e018422868f8c1762801b31e2f345fe850126c63787180c98ab

Observation b11e7624-cf49-4813-8575-fd2ee1887efa · outbound

This paper cites Proceedings of the IEEE conference on computer vision and pattern recognition , pages=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models Proceedings of the IEEE conference on computer vision and pattern recognition , pages=

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.242927Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.242927Z digest=sha256:cf5efab8aa5e0acf04e767ad692422bef5d08fa9778ec1b47e19346f11a1ed94

Observation 767ad612-a545-47d4-b4c1-ec951e3eafdd · outbound

This paper cites European conference on computer vision , pages=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models European conference on computer vision , pages=

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.246223Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.246223Z digest=sha256:f19640386630f50bfe7342823a6763de28e8452688ab4f8dc8e7160bbfa68aec

Observation 433131b0-99ba-4a91-bf5f-cf0eeba3d844 · outbound

This paper cites Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , pages=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , pages=

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:08:11.825372Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-15T15:08:10.249673Z digest=sha256:1305171284d330fce7561807f80fe96344003f652152259127ab621ae885cde7

Observation 5c4b15ae-f0af-4320-a5bf-a57c4b1eb904 · outbound

This paper cites The international journal of robotics research , volume=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models The international journal of robotics research , volume=

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.252980Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.252980Z digest=sha256:ffbbfd41a054c63e7d508eb4e1126c0658b82ca17f0e09fe0c9d664380b9ec36

Observation 502c1cb8-27c7-4b61-9345-9ae5154a4e49 · outbound

This paper cites Proceedings of the IEEE conference on computer vision and pattern recognition , pages=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models Proceedings of the IEEE conference on computer vision and pattern recognition , pages=

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.256437Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.256437Z digest=sha256:87a5aa6fad2cf9c261d8ca80d0cb7544eaeaff624a8c8a6dfe95dc7fa2975af6

Observation 3d1a9da0-79e6-43c9-bdb6-66e5e438e00a · outbound

This paper cites GPT-4o System Card.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models GPT-4o System Card

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.260290Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.260290Z digest=sha256:47016a93510cbaf8b9c0e42d563c53d9b2eff8ea9758f2634d14079132a55aee

Observation 03b360a9-2ceb-418e-81d8-bf2f27c60d57 · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models Advances in Neural Information Processing Systems , volume=

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:08:11.802754Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-15T15:08:10.263361Z digest=sha256:e6b1e933fe19ff04194f2ba974eae07cf702a733ca0b6643ed047f9bba6cda85

Observation 7b36932b-0625-4e4b-91b5-50ae042171d8 · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models Advances in Neural Information Processing Systems , volume=

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:08:11.791941Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-15T15:08:10.267499Z digest=sha256:5151d03d3dbe08520f664de3bbe58b96cb8cca2c191a5b2de775f7f28c9dd7a2

Observation 7b776672-c1e3-4c52-9f86-ff88c6b9befe · outbound

This paper cites Proceedings of the IEEE/CVF international conference on computer vision , pages=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models Proceedings of the IEEE/CVF international conference on computer vision , pages=

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:08:11.781165Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-15T15:08:10.271044Z digest=sha256:adb44fd3c2226f1e7b51b86460b5b5a1e10f873ee79f14a9eb6bcf1738e2e84d

Observation 6a23ec32-603d-4c59-b7b7-070b9206706b · outbound

This paper cites Proceedings of the Computer Vision and Pattern Recognition Conference , pages=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models Proceedings of the Computer Vision and Pattern Recognition Conference , pages=

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:08:11.770220Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-15T15:08:10.286451Z digest=sha256:fe55c3477ecdc7a6eadee29663fd4c9f99d1f6a184eb564c58d51f7cf96f117c

Observation 64c875bf-ea2f-40e5-b7c3-77e3ad16fc07 · outbound

This paper cites Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models Proceedings of the IEEE/CVF conference on computer vision and pattern recognition , pages=

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.290790Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.290790Z digest=sha256:880e79e20002a3d1dc4a4ac2fdc608c7fa444a2b3dd2d4ec28cbe6cace67ea3b

Observation 47d221e8-4b8a-404b-a578-206db2051555 · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models Advances in Neural Information Processing Systems , volume=

Reference 78

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T15:08:11.753395Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=arxiv_source observed=2026-08-15T15:08:10.295112Z digest=sha256:6f0b7a1af9f0aa2dfa93e43f1232e48d39829a1e7c5cdbbfc69967f8a4da008e

Observation 32449b10-0cda-41e4-b507-e6c0d2b2ca1b · outbound

This paper cites arXiv preprint arXiv:2512.08889 , year=.

SpatioLM: Towards General Physical Spatial Intelligence in Vision-Language Models arXiv preprint arXiv:2512.08889 , year=

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-15T15:08:10.299262Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T15:08:10.299262Z digest=sha256:252ffa98cd37532f647090cd7946486875cc51637c840c689853b873ff27b95f

Pith citing papers

No inbound Pith citation observations are available.