Pith. sign in

Paper Citation Record · LEDGER

SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 29 inbound Pith citation observations for arXiv:2501.10074.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.10074 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 29 of 29 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 29 of 29 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T21:22:27.158065Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T22:28:59.692534Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 71090c80-2619-4d3f-9187-59ef4726c803 · inbound

Multimodal Chain-of-Thought Reasoning: A Comprehensive Survey cites this paper.

Multimodal Chain-of-Thought Reasoning: A Comprehensive Survey SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 141

Resolution
verified exact
arxiv_id, observed 2026-05-15T17:18:53.227502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T17:18:52.996467Z digest=sha256:58693a533da7bea924c0893d3b4b92067aa989f362ffc053a807f22ba7298b32

Observation b022d3f3-a60b-435e-99dc-d710ad8ea032 · inbound

Ascending the Infinite Ladder: Benchmarking Spatial Deformation Reasoning in Vision-Language Models cites this paper.

Ascending the Infinite Ladder: Benchmarking Spatial Deformation Reasoning in Vision-Language Models SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T21:22:27.158065Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:22:27.158065Z digest=sha256:86cf4dc5671b8bb691d4f5287744d4df2dc0e9d896d440bf1c78880666103050

Observation 6b9a3432-c1f9-4b28-913c-24ea6a71de4d · inbound

AutoLayout: Closed-Loop Layout Synthesis via Slow-Fast Collaborative Reasoning cites this paper.

AutoLayout: Closed-Loop Layout Synthesis via Slow-Fast Collaborative Reasoning SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T19:56:14.640225Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:56:14.640225Z digest=sha256:bd90917114b553a4cdeb40101f0cd95ee5b7dd4f7a5d5ee6b2a8ea113fe71c8b

Observation 474139ab-99b5-46cc-801e-99b140d37b0b · inbound

Enhancing Spatial Reasoning in Vision-Language Models via Chain-of-Thought Prompting and Reinforcement Learning cites this paper.

Enhancing Spatial Reasoning in Vision-Language Models via Chain-of-Thought Prompting and Reinforcement Learning SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T19:54:48.113277Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:54:48.113277Z digest=sha256:cc0c02366a1a196c1b0043cdffc04d72535a29725cd7129c0281765022125288

Observation fb2d0e45-7ae2-4b1d-b81b-91325487b923 · inbound

Pseudo Depth Meets Gaussian: A Feed-forward RGB SLAM Baseline cites this paper.

Pseudo Depth Meets Gaussian: A Feed-forward RGB SLAM Baseline SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-05T23:56:04.037430Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:56:04.037430Z digest=sha256:ef9972a3112965d63b3b65756adedb9766999982bb58ca4f2ccaaeb09fb4484c

Observation d7e59e26-d885-4a83-9c37-ab0c44ded2ff · inbound

PySeizure: A single machine learning classifier framework to detect seizures in diverse datasets cites this paper.

PySeizure: A single machine learning classifier framework to detect seizures in diverse datasets SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-05T22:18:05.094176Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:18:05.094176Z digest=sha256:1f91be250a991716f4306bed8869e705bdf428b64104c6e59aed92d805a3482d

Observation 7ee2fe4f-6421-460a-b40b-9b26b89109b7 · inbound

Why Do MLLMs Struggle with Spatial Understanding? A Systematic Analysis from Data to Architecture cites this paper.

Why Do MLLMs Struggle with Spatial Understanding? A Systematic Analysis from Data to Architecture SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-05T11:40:23.101934Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T11:40:23.101934Z digest=sha256:e8dc59643ce340b6273e50cd5cdb8e5dfe128f804d06d6ad854214e1a46365dc

Observation 5dad4a7b-f9b2-4a70-b6c6-d4b53cb38a56 · inbound

DeepThinkVLA: Enhancing Reasoning Capability of Vision-Language-Action Models cites this paper.

DeepThinkVLA: Enhancing Reasoning Capability of Vision-Language-Action Models SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-18T03:10:48.842964Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-18T03:09:09.713822Z digest=sha256:b980bb907f8aee4fba92529752a1deb2f131f2199ac55bbbe605bfed48c19614

Observation 634d84b9-9dcc-413a-b90b-19e7c5a5b5b3 · inbound

SpaceTools: Tool-Augmented Spatial Reasoning via Double Interactive RL cites this paper.

SpaceTools: Tool-Augmented Spatial Reasoning via Double Interactive RL SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-03T18:42:12.111148Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:42:12.111148Z digest=sha256:354ffca295d433b70d01c2b2c4cc95e5604364f78c287627a000a89b4b98c97f

Observation b30b3fe8-a60c-4285-afbf-d40364a3d9eb · inbound

SCP: Spatial Causal Prediction in Video cites this paper.

SCP: Spatial Causal Prediction in Video SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-15T16:50:11.146675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T16:47:44.523606Z digest=sha256:307a8d711eb4f908389a2216c54dd308ae350df0a7d46e93ebb82741b312a4a3

Observation 4de075f5-7088-49b8-9dab-da2a933f0d85 · inbound

Token Warping Helps MLLMs Look from Nearby Viewpoints cites this paper.

Token Warping Helps MLLMs Look from Nearby Viewpoints SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 57

Resolution
verified exact
arxiv_id, observed 2026-05-13T21:08:17.361639Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-13T21:07:55.062113Z digest=sha256:86a01b33d6c7d9881104b6efbbcd1e5ca37f5d939688449530166ae20acb144e

Observation 5779c4ba-d08c-4616-b328-18e48d2f23df · inbound

Spatio-Temporal Grounding of Large Language Models from Perception Streams cites this paper.

Spatio-Temporal Grounding of Large Language Models from Perception Streams SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-11T07:26:02.684607Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-10T17:10:45.837684Z digest=sha256:e0ab1c59e4c18e2fe99a8f539d7a1014503616c845ca64d7c8cf0081df3a0b4d

Observation b6d757a7-7c85-4524-a156-a2e9aadef4ea · inbound

4DThinker: Thinking with 4D Imagery for Dynamic Spatial Understanding cites this paper.

4DThinker: Thinking with 4D Imagery for Dynamic Spatial Understanding SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-11T18:41:11.747234Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-08T14:20:08.404090Z digest=sha256:23d028e8dce5327e9ce3d79f741871b048a18ba7ce1bbbe6bc560d91d2ab1cdd

Observation 735994be-fcfe-4f8e-b488-ed43f427ee0b · inbound

4DThinker: Thinking with 4D Imagery for Dynamic Spatial Understanding cites this paper.

4DThinker: Thinking with 4D Imagery for Dynamic Spatial Understanding SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-25T06:16:39.949775Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-25T06:15:33.062980Z digest=sha256:b3c9e3e1ebb31d628cd5c9153aac0cf2814d2489367d21487558ed76991916ad

Observation 7d5b6bdc-d4c0-4847-b0dc-db113d0d9d6d · inbound

Proxy3D: Efficient 3D Representations for Vision-Language Models via Semantic Clustering and Alignment cites this paper.

Proxy3D: Efficient 3D Representations for Vision-Language Models via Semantic Clustering and Alignment SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:05:59.892965Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-11T01:55:39.721560Z digest=sha256:26645dad6cddd809c5b8c1d559883c31696cf749a338848ad89652aebabc0020

Observation 6f9e0018-747c-4b0d-b457-958542a9d008 · inbound

SpatialForge: Bootstrapping 3D-Aware Spatial Reasoning from Open-World 2D Images cites this paper.

SpatialForge: Bootstrapping 3D-Aware Spatial Reasoning from Open-World 2D Images SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-13T01:27:01.622856Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-13T01:26:47.052025Z digest=sha256:9ceb2a21dc82c8a7a0f2b8c16b3472b6ffbc64f4c3d8c26bdf0690754f7c1b3f

Observation e69a6f8e-d685-4842-8d40-009b644c3d13 · inbound

ESI-Bench: Towards Embodied Spatial Intelligence that Closes the Perception-Action Loop cites this paper.

ESI-Bench: Towards Embodied Spatial Intelligence that Closes the Perception-Action Loop SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 10

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T10:53:13.184532Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-20T10:52:22.778489Z digest=sha256:6d7fb4a58fd33db8a878459640caa7d13241758ba33b62524242d3fd4107c2ac

Observation 89df9b66-4673-47fa-82f8-04cce5bdfade · inbound

ESI-Bench: Towards Embodied Spatial Intelligence that Closes the Perception-Action Loop cites this paper.

ESI-Bench: Towards Embodied Spatial Intelligence that Closes the Perception-Action Loop SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 10

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T15:05:47.169697Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T18:25:17.831116Z digest=sha256:86828032a6a70aa728115e3af5552c715fae3f621a129717e689cfc7309fc885

Observation 30082d11-5f43-49ce-a533-2931f22751ee · inbound

Sketch2MinSurf: Vision-Language Guided Generation of Editable Minimal Surfaces from Hand-Drawn Sketches cites this paper.

Sketch2MinSurf: Vision-Language Guided Generation of Editable Minimal Surfaces from Hand-Drawn Sketches SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-21T05:03:58.061704Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T05:00:11.848526Z digest=sha256:1f2fa236df544d7fb52bb78f8bc23fa8b549b6a20244be591b4b1affc3ff1d40

Observation f0d63128-776f-4dc8-92fb-3379e8744763 · inbound

DRIVESPATIAL: A Benchmark for Spatiotemporal Intelligence in VLMs for Autonomous Driving cites this paper.

DRIVESPATIAL: A Benchmark for Spatiotemporal Intelligence in VLMs for Autonomous Driving SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 58

Resolution
verified exact
arxiv_id, observed 2026-05-25T05:15:22.732097Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-25T05:10:32.522453Z digest=sha256:36355ca4e9bfa865e8d9c71db60f095d65730ea3c53abfbb240b5cd3992427cd

Observation c178bf64-6215-4858-8932-b616a2f27228 · inbound

DRIVESPATIAL: A Benchmark for Spatiotemporal Intelligence in VLMs for Autonomous Driving cites this paper.

DRIVESPATIAL: A Benchmark for Spatiotemporal Intelligence in VLMs for Autonomous Driving SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 58

Resolution
verified exact
arxiv_id, observed 2026-06-30T16:44:56.062222Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-30T16:40:22.441025Z digest=sha256:1c4304a1c6ff4c23a3ee8e851e083bcb6721883fa0bc20aff28dd14a34a44d6a

Observation f2fb586f-e6fd-49cc-a50b-6a6d5ebf95c2 · inbound

Grounded 3D-Aware Spatial Vision-Language Modeling cites this paper.

Grounded 3D-Aware Spatial Vision-Language Modeling SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 74

Resolution
verified exact
arxiv_id, observed 2026-06-29T08:13:15.342690Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-29T08:08:36.012761Z digest=sha256:38cd29537ee662be5a2b6544ce073d81ac7e95968554f3f42ab978c80dcf16fb

Observation 8cc528be-b179-4200-b688-23ea88a620e4 · inbound

SpatialWorld: Benchmarking Interactive Spatial Reasoning of Multimodal Agents in Real-World Tasks cites this paper.

SpatialWorld: Benchmarking Interactive Spatial Reasoning of Multimodal Agents in Real-World Tasks SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-07-03T01:27:30.607933Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-27T16:35:14.099586Z digest=sha256:d7eacb151624a95a4ed4ccc459a18eca1d52f4c2043579f6ded1fe7be49924ad

Observation 32232175-98d9-4877-ba72-d02745cd166f · inbound

Reinforcing Dual-Path Reasoning in Spatial Vision Language Models cites this paper.

Reinforcing Dual-Path Reasoning in Spatial Vision Language Models SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-07-03T20:08:55.601454Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-06-27T01:42:30.005911Z digest=sha256:545700a7b6c1839322cea8d7d4bbeda37a760485ddb96db7b246acb4c8c417bb

Observation 4344ecef-9e55-4a5e-9e19-5f42073828f2 · inbound

Spatial Reasoning via Modality Switching Between Language and Symbolic Representation cites this paper.

Spatial Reasoning via Modality Switching Between Language and Symbolic Representation SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 68

Resolution
verified exact
arxiv_id, observed 2026-07-01T10:15:44.479409Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-07-01T05:41:46.057986Z digest=sha256:edc4cfd27902331fb0041dc2334c85aff45853e60989c103073e11a3faefcabb

Observation f481d8e6-21b1-4613-9024-233ef425327c · inbound

Spatial Reasoning via Modality Switching Between Language and Symbolic Representation cites this paper.

Spatial Reasoning via Modality Switching Between Language and Symbolic Representation SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 68

Resolution
verified exact
arxiv_id, observed 2026-07-03T22:28:59.694996Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-07-03T22:28:24.036122Z digest=sha256:b5bdbca1c0cee89b548de6d4d508da1452f7f6491391438230a0baeb6fb019de

Observation 5317129b-a50e-4e53-882f-79221021b8ed · inbound

OmniView-Space: Reinforcing Spatial Reasoning via Multi-Perspective Spatial Mapping cites this paper.

OmniView-Space: Reinforcing Spatial Reasoning via Multi-Perspective Spatial Mapping SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 52

Resolution
verified exact
arxiv_id, observed 2026-07-02T14:17:02.420247Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-07-02T14:16:39.649823Z digest=sha256:2923b3ec12ca64ce987c2cc3dcf8481730f7a9e3d0b3a61d72228b15210d7583

Observation 490f06ad-ce8e-40d3-a95f-7ae53961e6b1 · inbound

SpaceEra++: A Unified Framework Towards 3D Spatial Reasoning in Video cites this paper.

SpaceEra++: A Unified Framework Towards 3D Spatial Reasoning in Video SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-07-03T16:18:37.311510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-07-03T16:16:41.412451Z digest=sha256:94f140959ed52f578c8adef1de5f29348535b90ba8b690fe6e7329c231c6bc97

Observation ddc2f67e-6117-4bc4-8b82-5fb77958e9ed · inbound

ViSTR-Bench: Can MLLMs Reason from Continuous Visual Cues in Dynamic Scenes? cites this paper.

ViSTR-Bench: Can MLLMs Reason from Continuous Visual Cues in Dynamic Scenes? SpatialCoT: Advancing Spatial Reasoning through Coordinate Alignment and Chain-of-Thought for Embodied Task Planning

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-01T09:09:24.863958Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T09:09:24.863958Z digest=sha256:053a3bc3ce1bf03c8ba9fc3cf15a786373659a2c2e5406b8deffbc0a93c67df4