Pith. sign in

Paper Citation Record · LEDGER

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning

As of 13 August 2026, this Paper Citation Record lists 100 of 129 outbound references and 0 inbound Pith citation observations for arXiv:2608.09682.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.09682 v1

Coverage vector

measured 100 of 129 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T12:53:14.503030Z

measured 100 of 100 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

100 of 129 outbound references displayed

  • verified exact3
  • verified fuzzy0
  • unresolved97
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 21d126e9-bc08-49be-ac7a-28c7872f3451 · outbound

This paper cites Latent reasoning with supervised thinking states.arXiv preprint arXiv:2602.08332, 2026.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Latent reasoning with supervised thinking states.arXiv preprint arXiv:2602.08332, 2026

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:13.922601Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:13.922601Z digest=sha256:2ff931e57ee58e8edbfe514a75441b26f332e8075c286e41e91d3f606971be74

Observation 3b40b537-9d4c-4dae-87c7-72e0cbaead1d · outbound

This paper cites Acloserlookatbiasandchain-of-thought faithfulness of large (vision) language models.arXiv preprint arXiv:2505.23945, 2025.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Acloserlookatbiasandchain-of-thought faithfulness of large (vision) language models.arXiv preprint arXiv:2505.23945, 2025

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:13.930040Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:13.930040Z digest=sha256:70b14abb09022e2c3f625e56fdfc29612a1cc992e8db5789c862fb93b1a3d8c4

Observation 54c530f6-4381-45b3-8282-d55da1520d61 · outbound

This paper cites An Image is Worth 1/2 Tokens After Layer 2: Plug-and-Play Inference Acceleration for Large Vision-Language Models.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning An Image is Worth 1/2 Tokens After Layer 2: Plug-and-Play Inference Acceleration for Large Vision-Language Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:13.935715Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:13.935715Z digest=sha256:f4201fb555243e333aa43233063589dbab80140a94e1d05064337cc2b59a8194

Observation 4127ec8f-cb29-49ef-bf11-6a0e20ae4c03 · outbound

This paper cites MMStar: An evaluator-centric benchmark for multi-modal large language models, 2024.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning MMStar: An evaluator-centric benchmark for multi-modal large language models, 2024

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:13.941754Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:13.941754Z digest=sha256:ed86e129968b81dcd9520d03cfc4a0428907a3960f80983ea71c7a47307c517a

Observation e3a06ba7-1a84-4d28-9062-55bd6db2c470 · outbound

This paper cites Perception before reasoning: Two-stage reinforcement learning for visual reasoning.arXiv preprint arXiv:2509.13031, 2025.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Perception before reasoning: Two-stage reinforcement learning for visual reasoning.arXiv preprint arXiv:2509.13031, 2025

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:13.947142Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:13.947142Z digest=sha256:e8e58ed006195f714cbe0fa765835fab737ad4387caedb070542038f1dc3ede7

Observation c1d6048d-a793-41d2-8a99-295b545a062e · outbound

This paper cites Reasoning Models Don't Always Say What They Think.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Reasoning Models Don't Always Say What They Think

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:13.952698Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:13.952698Z digest=sha256:f382f61af0da123972cb5fc51187cb8648e514df4155b11a69ff08ae395d4767

Observation 527f98d2-c979-4acc-b913-fbf7eec197af · outbound

This paper cites v1: Learning to Point Visual Tokens for Multimodal Grounded Reasoning.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning v1: Learning to Point Visual Tokens for Multimodal Grounded Reasoning

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:13.959442Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:13.959442Z digest=sha256:b5cf355a163e7deceab43e5f557435c4615cf0dc8873ffe96a10ebf00f45d540

Observation 4b2e1502-e09d-4db2-af0d-9f9e3250c7e9 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:13.964796Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:13.964796Z digest=sha256:85f1ffc624302b4ac3522e79d2124c2fa7f3c1216b472bbff979d0b538747b6b

Observation 509a382a-9523-41ba-98f9-4c79aa25a0ef · outbound

This paper cites Molmo and PixMo: Open Weights and Open Data for State-of-the-Art Vision-Language Models.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Molmo and PixMo: Open Weights and Open Data for State-of-the-Art Vision-Language Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:13.970205Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:13.970205Z digest=sha256:ccb5bef3fba1b0e218986a4c467b8c8e0d5ff16a6ac941e9cc11aa9a003fff94

Observation 93107529-ff1e-4d89-8ff6-a0d9f5579f93 · outbound

This paper cites Virgo: A Preliminary Exploration on Reproducing o1-like MLLM.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Virgo: A Preliminary Exploration on Reproducing o1-like MLLM

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:13.978385Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:13.978385Z digest=sha256:e1600f5f85c42b774db4151a4cbadc69ad331cc839b37ac14fb55ba1c25069c6

Observation 65d3ed44-1f15-4843-b40c-8316c5ac469c · outbound

This paper cites Revisiting the necessity of lengthy chain-of-thought in vision-centric reasoning.arXiv preprint arXiv:2511.22586, 2025.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Revisiting the necessity of lengthy chain-of-thought in vision-centric reasoning.arXiv preprint arXiv:2511.22586, 2025

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:13.984603Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:13.984603Z digest=sha256:50fc4fecfebd7b2653635164cad15a920fd0eb7d07caea9b38ccac36c6d5b345

Observation be44b9db-1f68-4022-8868-c22a5ce9043b · outbound

This paper cites VLMEvalKit: An open-source toolkit for evaluating large multi-modality models.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning VLMEvalKit: An open-source toolkit for evaluating large multi-modality models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:13.990131Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:13.990131Z digest=sha256:885812c4dc04fefbdbd401d632a65106cc95ff041cf0ef8f3de96e99cac7b982

Observation 7931aa8e-29c5-4a9f-9949-92b81a812407 · outbound

This paper cites GRIT: Teaching MLLMs to Think with Images.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning GRIT: Teaching MLLMs to Think with Images

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:13.996266Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:13.996266Z digest=sha256:f05559c56f9501addf19e30be94bec018c5e38b9bfc271249ac34951aaa50ffa

Observation 9056dd3f-3141-49d4-8937-2bf9ac597373 · outbound

This paper cites Reward Shaping to Mitigate Reward Hacking in RLHF.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Reward Shaping to Mitigate Reward Hacking in RLHF

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.002043Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.002043Z digest=sha256:a614715bd174819b6f0f5cf00c6038ab64fd1cb58f290ff0422e76c5f8d9de13

Observation 195d31a5-11c3-492e-af18-b14ebf648ec6 · outbound

This paper cites BLINK: Multimodal Large Language Models Can See but Not Perceive.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning BLINK: Multimodal Large Language Models Can See but Not Perceive

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.008737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.008737Z digest=sha256:bed5ea99feb278c592debc7c2935b8c8775f4ebac0fe35b3a8d9c8b83efad8f6

Observation 64d7c8d5-a9f5-446f-ab58-b427e8915449 · outbound

This paper cites Thinking with deltas: Incen- tivizingreinforcementlearningviadifferentialvisualreasoningpolicy.arXivpreprintarXiv:2601.06801, 2026.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Thinking with deltas: Incen- tivizingreinforcementlearningviadifferentialvisualreasoningpolicy.arXivpreprintarXiv:2601.06801, 2026

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.016201Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.016201Z digest=sha256:a7d014581b08c55f29232629836af90a68d5aa3434563ee050cf7915bbacb747

Observation 068e1c91-7dab-46d4-be3a-3a5502407ccd · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Gemini: A Family of Highly Capable Multimodal Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.021564Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.021564Z digest=sha256:8990cdbe3ba97f9a044c97f49fa600ed90d973ce503433d2791ba4df0b1fee5b

Observation a332fe5b-83bd-4f84-a92e-1559c14595f7 · outbound

This paper cites GLM-5V-Turbo: Toward a Native Foundation Model for Multimodal Agents.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning GLM-5V-Turbo: Toward a Native Foundation Model for Multimodal Agents

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.026958Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.026958Z digest=sha256:333407769c5eebd17ff40d3fed6aaa91d4930b12a617649f9944982466683b76

Observation 0456dc97-cbd6-472e-9331-b839f796af49 · outbound

This paper cites Visual programming: Compositional visual reasoning without training.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Visual programming: Compositional visual reasoning without training

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.032344Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.032344Z digest=sha256:3c0b5c84dc6c0241ed00eb0166d55876ebfbb3432bf518494187aa820c21f706

Observation e15c4967-d4e1-4555-9a7d-b39829d8404d · outbound

This paper cites Training Large Language Models to Reason in a Continuous Latent Space.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Training Large Language Models to Reason in a Continuous Latent Space

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.037446Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.037446Z digest=sha256:93b2d76e9435f68f59285d24b7c90adaba156d075e97dfedd4c8e0b2a6d4c81e

Observation 661bcc98-115d-42d1-af70-47b2e98f38f0 · outbound

This paper cites DeepEyesV2: Toward Agentic Multimodal Model.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning DeepEyesV2: Toward Agentic Multimodal Model

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.046256Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.046256Z digest=sha256:ac35fa03f570df3d5443427bea0edb19108daacf25416f0550eb1e82c1afa404

Observation caeb8699-53c6-4525-ae64-b3f61ae401e4 · outbound

This paper cites Hollon, and Bryan Wang.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Hollon, and Bryan Wang

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.052496Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.052496Z digest=sha256:5409c19a8e70f37bb1fe8383e09debd34ded275da525c9c941a58538f2faf24a

Observation 24edadf1-b96c-40e0-b450-1d1055e80bc6 · outbound

This paper cites Visual Sketchpad: Sketching as a Visual Chain of Thought for Multimodal Language Models.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Visual Sketchpad: Sketching as a Visual Chain of Thought for Multimodal Language Models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.059008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.059008Z digest=sha256:fe938864beec438321f38964f79e5c312a124101126fc167689704046daaa2a5

Observation 2688c6b2-61ba-4f3a-b127-b6d76bfd8826 · outbound

This paper cites VerlTool: Towards holistic agentic reinforcement learning with tool use.arXiv preprint arXiv:2509.01055, 2025.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning VerlTool: Towards holistic agentic reinforcement learning with tool use.arXiv preprint arXiv:2509.01055, 2025

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.065378Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.065378Z digest=sha256:0e3e89552444bfb078b6b9c44a0907e9ecae9cdf987d431a10993ea4a86dc4ad

Observation 59f00b1d-851b-48ac-9c49-cf03c313ff85 · outbound

This paper cites Kimi K2.5: Visual Agentic Intelligence.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Kimi K2.5: Visual Agentic Intelligence

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.072026Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.072026Z digest=sha256:5cfc2b0cd4762156816b647368cd1908c0c7869214f1675ec3f87a55db7577b6

Observation b105647b-87f3-4fb7-82bf-6f0446626b1f · outbound

This paper cites Mini-o3: Scaling Up Reasoning Patterns and Interaction Turns for Visual Search.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Mini-o3: Scaling Up Reasoning Patterns and Interaction Turns for Visual Search

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.077712Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.077712Z digest=sha256:eed87fe1c00663542b7f3ec220b40feb6cc2dd6dc205b71afef355d8aeb6ce04

Observation 2897de1e-e8b0-45dd-9850-acec5ccdf6bc · outbound

This paper cites Measuring Faithfulness in Chain-of-Thought Reasoning.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Measuring Faithfulness in Chain-of-Thought Reasoning

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.083137Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.083137Z digest=sha256:3d563288fb7bcc71bd8f890d40556e087cc66cedac31c5dc992d4094329e5f60

Observation 35cf9fd5-68e4-40d2-972e-a5679640d9ca · outbound

This paper cites Zebra-cot: A dataset for interleaved vision language reasoning.arXiv preprint arXiv:2507.16746, 2025.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Zebra-cot: A dataset for interleaved vision language reasoning.arXiv preprint arXiv:2507.16746, 2025

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.088153Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.088153Z digest=sha256:c91b3456608675aecd929440f062e52a70dcfa3eca89cc5c2db199477d06fcd6

Observation 76d54afa-1dd1-45c7-bd61-5613d7a3532b · outbound

This paper cites Tir-bench: A comprehensive benchmark for agentic thinking-with-images reasoning.arXiv preprint arXiv:2511.01833, 2025.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Tir-bench: A comprehensive benchmark for agentic thinking-with-images reasoning.arXiv preprint arXiv:2511.01833, 2025

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.092965Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.092965Z digest=sha256:d0bae8742db8514b079b4844751499d92c6d4037e0937cc317610f965be388bf

Observation 6155944d-2c98-4ebf-ac10-b83d52418bbb · outbound

This paper cites Seg-Zero: Reasoning-Chain Guided Segmentation via Cognitive Reinforcement.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Seg-Zero: Reasoning-Chain Guided Segmentation via Cognitive Reinforcement

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.097596Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.097596Z digest=sha256:33896f0fff79b0d1e6459efec382711abc27407abab0705ba6a55dfbab1483cb

Observation 8c4c7087-0bf3-4c1c-ba1f-f214947de793 · outbound

This paper cites On the faithfulness of visual thinking: Measurement and enhancement.arXiv preprint arXiv:2510.23482, 2025.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning On the faithfulness of visual thinking: Measurement and enhancement.arXiv preprint arXiv:2510.23482, 2025

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.102763Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.102763Z digest=sha256:0c5154410a26955497621d1aef35107673e76d0f6bcd4fa2a2753ce191d871a5

Observation fed3ace4-7f7a-4068-a766-df90f1678d4d · outbound

This paper cites Chain-of-Spot: Interactive Reasoning Improves Large Vision-Language Models.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Chain-of-Spot: Interactive Reasoning Improves Large Vision-Language Models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.107674Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.107674Z digest=sha256:40c3d86449565739b84a18b3648261166827c4b3d90b8605105f29ead4223878

Observation 62e485d8-49de-4c3a-9605-ae3b6734b188 · outbound

This paper cites Chameleon: Plug-and-play compositional reasoning with large language models.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Chameleon: Plug-and-play compositional reasoning with large language models

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.113186Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.113186Z digest=sha256:a6945d3f5a055a9fe270da03da1bd93fae47431bbb36bf1e9622b6dd1a6ca777

Observation 9004a401-c7d2-49cc-ba01-b07b1b464f24 · outbound

This paper cites Reasoning Models Can Be Effective Without Thinking.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Reasoning Models Can Be Effective Without Thinking

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.118396Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.118396Z digest=sha256:20bfe37be785480da686fbf1097d3ffaead0361dd755649e2dea7ec0fca3f20c

Observation 87e7c981-a900-4a4f-9443-7f23e41f7676 · outbound

This paper cites What Does Vision Tool-Use Reinforcement Learning Really Learn? Disentangling Tool-Induced and Intrinsic Effects for Crop-and-Zoom.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning What Does Vision Tool-Use Reinforcement Learning Really Learn? Disentangling Tool-Induced and Intrinsic Effects for Crop-and-Zoom

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.125300Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.125300Z digest=sha256:5aaa4fc7521a620fc4c0d9808dee4d1560e34829dc04c22d140f0a43241ecd4f

Observation 8c9b6ca7-8c46-4542-b286-5c2210ec5976 · outbound

This paper cites ChartQA: A Benchmark for Question Answering about Charts with Visual and Logical Reasoning.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning ChartQA: A Benchmark for Question Answering about Charts with Visual and Logical Reasoning

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.131378Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.131378Z digest=sha256:379f71d9796cf2a722d9f5b99b455a57a5b29ae23ca911be44068a32eb7737e5

Observation d915d33d-2552-490f-b362-54b7e010fd32 · outbound

This paper cites GSM-Symbolic: Understanding the Limitations of Mathematical Reasoning in Large Language Models.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning GSM-Symbolic: Understanding the Limitations of Mathematical Reasoning in Large Language Models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.137144Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.137144Z digest=sha256:a4bbf4581abae8f0f2aa5b234c054f8b259d339b6d56b867ec11371dd60f1173

Observation 38d5594e-04c8-45a8-9b98-56a44e4122d7 · outbound

This paper cites Show Your Work: Scratchpads for Intermediate Computation with Language Models.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Show Your Work: Scratchpads for Intermediate Computation with Language Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.143566Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.143566Z digest=sha256:f9d5cd5230c9f0d7b7a597b4d8e28be2b94d7a0a4c4bb6f265a91a76d805e40a

Observation ab05b758-cb33-47e7-92ea-c69c2c1dc7d4 · outbound

This paper cites GPT-4 Technical Report.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning GPT-4 Technical Report

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.149920Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.149920Z digest=sha256:3b3b7584884b69402311d270a06da1cbf3d4168690823ac3c74f267db91a83aa

Observation 29b9e7f1-eedf-448c-beb2-59ffb461b6e5 · outbound

This paper cites Thinking with images.https://openai.com/index/thinking-with-images/, 2025.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Thinking with images.https://openai.com/index/thinking-with-images/, 2025

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.156357Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.156357Z digest=sha256:10c97014e2f575dd9b795bb0fbb4485a16fc1d243ca79dd4b7b5a5638dc9cd33

Observation 7a9f1ba0-82a1-4a24-ae45-3cfb98252cde · outbound

This paper cites Do MLLMs really see it: Reinforcing visual attention in multimodal LLMs.arXiv preprint arXiv:2602.08241, 2026.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Do MLLMs really see it: Reinforcing visual attention in multimodal LLMs.arXiv preprint arXiv:2602.08241, 2026

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.161693Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.161693Z digest=sha256:481c38662ea27f65aa3565b361beaca68d36ead42482bfbbb7a0bec416426f3f

Observation c9a368e4-f37a-4ffd-b1d6-20ad3b5f427f · outbound

This paper cites CogCoM: A Visual Language Model with Chain-of-Manipulations Reasoning.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning CogCoM: A Visual Language Model with Chain-of-Manipulations Reasoning

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.169132Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.169132Z digest=sha256:7fdc8301fdc1c552bd1da8ea505dd634112ec0cabbe410235449d30bc2636b9f

Observation d47f3c4d-f832-4c6b-a745-c38dbd5f852b · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.175515Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.175515Z digest=sha256:dd7d6b71eb9fbb0b73f1da9d7691ee0a3e3eadb2c7defdf4cd9e26b6ca1fb104

Observation 25ad21de-9c81-4709-b03a-9f0f39708c29 · outbound

This paper cites Qwen2.5-VL Technical Report.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Qwen2.5-VL Technical Report

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.182177Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.182177Z digest=sha256:d6b4dca8da6fc463950f4d26a3849404c78dee834c75bcac5d8839a256cd2967

Observation 8c0aada3-d18d-4ebe-83af-b05f2c1bd462 · outbound

This paper cites Qwen3-VL Technical Report.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Qwen3-VL Technical Report

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.188571Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.188571Z digest=sha256:a56b9a3f4fc3acb677fd123ca83db53fe5853df490c5ddef187cdd7303b7735c

Observation fa21b7df-5558-4e33-8163-a62a77acc62f · outbound

This paper cites Vision language models are blind: Failing to translate detailed visual features into words.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Vision language models are blind: Failing to translate detailed visual features into words

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.194557Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.194557Z digest=sha256:ba8e483ae36607fc2c0509fbde66600ca3f2c7ab45c73798aaef84cc48c2495c

Observation b553977f-9a03-4e10-9724-7816c539dd03 · outbound

This paper cites Grounded Reinforcement Learning for Visual Reasoning.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Grounded Reinforcement Learning for Visual Reasoning

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.200723Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.200723Z digest=sha256:356bdfdff04e2b1b42a996b88928f5d68cd064873fa9415092280d2141382b06

Observation 71aed7c4-6d94-4238-a052-da220eee155a · outbound

This paper cites Toolformer: Language models can teach themselves to use tools.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Toolformer: Language models can teach themselves to use tools

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.206226Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.206226Z digest=sha256:c19ab5eda46aa81200542e6c091d4e6b4e0c3b0de2b8b81a140ab6c57e96a28e

Observation eb554f4a-9642-4522-ac1f-13538d74085c · outbound

This paper cites Visual CoT: Advancing Multi-Modal Language Models with a Comprehensive Dataset and Benchmark for Chain-of-Thought Reasoning.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Visual CoT: Advancing Multi-Modal Language Models with a Comprehensive Dataset and Benchmark for Chain-of-Thought Reasoning

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.211374Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.211374Z digest=sha256:a66b4313dfdeb7e641c8121572569ef1f3f94aa3541337040b75dff9a0ef431d

Observation 36606029-2231-4dbc-87b7-538e2926826a · outbound

This paper cites HuggingGPT: Solving AI tasks with ChatGPT and its friends in Hugging Face.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning HuggingGPT: Solving AI tasks with ChatGPT and its friends in Hugging Face

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.216761Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.216761Z digest=sha256:bc61ebc66b11cc4220e8f036318e1ae9accd817ffb054574c494e6118f6c0f2b

Observation d80610d9-2730-4bd8-bb41-1c6a246f93d8 · outbound

This paper cites HybridFlow: A Flexible and Efficient RLHF Framework.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning HybridFlow: A Flexible and Efficient RLHF Framework

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.221402Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.221402Z digest=sha256:f1afa5e63913a2085b3fb572662078a1cba4eaa4fa251ab2167966c2e2c8e297

Observation 7039400c-ff0f-4515-a00c-e38454fb4b39 · outbound

This paper cites Reflexion: Language agents with verbal reinforcement learning.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Reflexion: Language agents with verbal reinforcement learning

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.226408Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.226408Z digest=sha256:6b8b91230856318b97da58b5e1a6f441d6b39bbf9846914627aaae92f9d40343

Observation bd7094a6-99ef-4afb-b609-325ea8259745 · outbound

This paper cites Breaking the Chain: A Causal Analysis of LLM Faithfulness to Intermediate Structures.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Breaking the Chain: A Causal Analysis of LLM Faithfulness to Intermediate Structures

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.231434Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.231434Z digest=sha256:789b78a9002715399432d60d78ee36be7d6535f2e53198061d7bcc159a81a210

Observation 2f895714-0304-4781-afac-8325cbee7e5e · outbound

This paper cites OpenThinkIMG: Learning to Think with Images via Visual Tool Reinforcement Learning.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning OpenThinkIMG: Learning to Think with Images via Visual Tool Reinforcement Learning

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.237048Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.237048Z digest=sha256:605ced9576f2547b78c0f9ad24eaa2668dcc642dced256f77aeb8c4a163e163b

Observation 6adf3fcd-a92f-4e84-9f6c-f35f8fdca13d · outbound

This paper cites Thinking with Images for Multimodal Reasoning: Foundations, Methods, and Future Frontiers.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Thinking with Images for Multimodal Reasoning: Foundations, Methods, and Future Frontiers

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.243692Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.243692Z digest=sha256:42d4d21c2b0e794675fdf9ceba630f727f64df7792ef773f3e27699b4f708efc

Observation 808134bd-c4aa-41ce-8ced-e241ff856bee · outbound

This paper cites When thinking hurts: Mitigating visual forgetting via frame repetition.arXiv preprint arXiv:2603.16256, 2026.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning When thinking hurts: Mitigating visual forgetting via frame repetition.arXiv preprint arXiv:2603.16256, 2026

Reference 56

Resolution
verified exact
raw_fallback, observed 2026-08-11T12:53:16.994849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T12:53:14.249540Z digest=sha256:942901811800feb1002fcdb192a9cb049b40c5fe0a00d0ebb6d8caefd9afa962

Observation 33cfd61d-327a-47bf-831c-7f2f998d7ef9 · outbound

This paper cites FACT-E: Causality-Inspired Evaluation for Trustworthy Chain-of-Thought Reasoning.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning FACT-E: Causality-Inspired Evaluation for Trustworthy Chain-of-Thought Reasoning

Reference 57

Resolution
verified exact
local_arxiv, observed 2026-08-11T12:53:16.902776Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T12:53:14.254925Z digest=sha256:04677d990c9c1ddcb5477e19375a442d2c6d94c4b151dab959bcb155152e9cec

Observation a6faa050-844b-4cd5-9dc9-5a37b4bcd364 · outbound

This paper cites ViperGPT: Visual inference via python execution for reasoning.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning ViperGPT: Visual inference via python execution for reasoning

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.260904Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.260904Z digest=sha256:2e80cf9574399dfa275b521aeada7776b28cbd71e4a5fc5dc3bb7f04599b19dc

Observation c664eea2-5da5-491d-a904-af12e5d86968 · outbound

This paper cites CV-Bench: A computer vision benchmark for evaluating visual perception in multimodal language models.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning CV-Bench: A computer vision benchmark for evaluating visual perception in multimodal language models

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.266490Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.266490Z digest=sha256:7b89987a4e65bd37ebc2cd741d6bc92b47156ba078286e6838c35fdd9238afda

Observation 984d86a0-d722-4cdd-a983-895defd786b0 · outbound

This paper cites an unresolved cited work.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Unresolved cited work

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.271924Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.271924Z digest=sha256:515a1dd22f2a5a5b57b444098be34706414a64985316b8e0428822c6b23b8c57

Observation cb4caa7c-52de-4d80-8dba-fa757ff24771 · outbound

This paper cites Journey before destination: Visual faithfulness in slow thinking.arXiv preprint arXiv:2512.12218, 2025.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Journey before destination: Visual faithfulness in slow thinking.arXiv preprint arXiv:2512.12218, 2025

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.277132Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.277132Z digest=sha256:a198dc6d36a3439732a89f5df39ffc62ebb3c34b81e9dde3977c2db0bbfbdc61

Observation aa0df45c-08c5-43d0-8474-d7483c923177 · outbound

This paper cites GeoEyes: On-demand visual focusing for ultra-high-resolution remote sensing.arXiv preprint arXiv:2602.14201, 2026.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning GeoEyes: On-demand visual focusing for ultra-high-resolution remote sensing.arXiv preprint arXiv:2602.14201, 2026

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.282612Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.282612Z digest=sha256:61b5b2b915209bd42fe460890c4934f9cec6d015af6913446fbb5129771659d4

Observation 0fcc6dce-7fe1-493b-953e-1838a6248d78 · outbound

This paper cites Pixel Reasoner: Incentivizing Pixel-Space Reasoning with Curiosity-Driven Reinforcement Learning.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Pixel Reasoner: Incentivizing Pixel-Space Reasoning with Curiosity-Driven Reinforcement Learning

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.289490Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.289490Z digest=sha256:0cedac2c841fbe0afbef5d961c4107641b662fb78d95a755db20e96f0cfa5271

Observation 431c9e90-1e5c-4855-8d17-1096829a3f45 · outbound

This paper cites PLaT: Latent chain-of-thought as planning: Decoupling reasoning from verbalization.arXiv preprint arXiv:2601.21358, 2026.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning PLaT: Latent chain-of-thought as planning: Decoupling reasoning from verbalization.arXiv preprint arXiv:2601.21358, 2026

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.295877Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.295877Z digest=sha256:b36ed4e34ea53fbe18ad0411902534ec5c0a8e07ecfd329eb3d8a215ad3349eb

Observation 13f7fe4e-f9e3-4b26-924d-51d354c8d659 · outbound

This paper cites VAGEN: Reinforcing world model reasoning for multi-turn VLM agents.arXiv preprint arXiv:2510.16907, 2025.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning VAGEN: Reinforcing world model reasoning for multi-turn VLM agents.arXiv preprint arXiv:2510.16907, 2025

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.301424Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.301424Z digest=sha256:fba885d4d7103a04cfe7932a178159db0e7436c9ac72808a953fb778507dbda2

Observation 155234b5-f117-43e0-ac3f-685adee465c4 · outbound

This paper cites Plan-and-solve prompting: Improving zero-shot chain-of-thought reasoning by large language models.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Plan-and-solve prompting: Improving zero-shot chain-of-thought reasoning by large language models

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.307438Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.307438Z digest=sha256:a4080600f6cbca493463dff9e5f1a52a77b7904d8c463765605c16121b0dbda4

Observation ad4de20a-fa87-4e81-a3f1-935df8556f00 · outbound

This paper cites A practitioner’s guide to multi-turn agentic reinforcement learning.arXiv preprint arXiv:2510.01132, 2025.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning A practitioner’s guide to multi-turn agentic reinforcement learning.arXiv preprint arXiv:2510.01132, 2025

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.313080Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.313080Z digest=sha256:34e8b78bac49f4228e2b41d910009139d00ebae500a9eabb600d2d66660330c8

Observation 7e8456bd-7ef2-4fdb-86cb-b3b5d1e618a7 · outbound

This paper cites Divide, Conquer and Combine: A Training-Free Framework for High-Resolution Image Perception in Multimodal Large Language Models.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Divide, Conquer and Combine: A Training-Free Framework for High-Resolution Image Perception in Multimodal Large Language Models

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.318402Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.318402Z digest=sha256:68d36c22b33f5d7288d47361a37563615cd51482429f53db44769ba5ce2b4925

Observation f574102a-4bd5-4cf0-aaab-eeb6bc7aa745 · outbound

This paper cites Self-consistency improves chain of thought reasoning in language models.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Self-consistency improves chain of thought reasoning in language models

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.324074Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.324074Z digest=sha256:b55cfdac2b2102ac9443bdeff4b90697c8e6cc2e2855f6d0c3d6d362e513846b

Observation 11876701-c04c-48cb-9602-4d5af4eb6f62 · outbound

This paper cites Simple o3: Towards Interleaved Vision-Language Reasoning.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Simple o3: Towards Interleaved Vision-Language Reasoning

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.330057Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.330057Z digest=sha256:81cf8b9b053fecbde4af33bc354965893426264d65a104a92601d3d0eaa94a77

Observation 672af4c4-9a5f-40d2-abde-ad12cb180fd6 · outbound

This paper cites RAGEN: Understanding Self-Evolution in LLM Agents via Multi-Turn Reinforcement Learning.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning RAGEN: Understanding Self-Evolution in LLM Agents via Multi-Turn Reinforcement Learning

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.335916Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.335916Z digest=sha256:37900542f9ab7a24e5c84882b943dd35005787d9ad7789549f75bf334d81ae19

Observation 58bc7e5f-585b-444a-b826-0f46909038fb · outbound

This paper cites CharXiv: Charting Gaps in Realistic Chart Understanding in Multimodal LLMs.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning CharXiv: Charting Gaps in Realistic Chart Understanding in Multimodal LLMs

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.341547Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.341547Z digest=sha256:5d1e45be0d5754c23b71316c8438d52be45225ef79322590572543d8690f5135

Observation f6834941-8cc1-470f-bd19-d04bb69d18b4 · outbound

This paper cites V-FAT: Benchmarking visual fidelity against text-bias.arXiv preprint arXiv:2601.04897, 2025.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning V-FAT: Benchmarking visual fidelity against text-bias.arXiv preprint arXiv:2601.04897, 2025

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.346906Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.346906Z digest=sha256:deb144e5a7eda72b6af35b4a472e93442f28322df5ccdd148b532be0188a691e

Observation 75cb0d3c-04e6-46b5-a8e7-454ba0800c7b · outbound

This paper cites Chi, Quoc V.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Chi, Quoc V

Reference 74

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.352064Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.352064Z digest=sha256:99a281d41a79b46c6a4a64e0e3086d368bab875b9e4ae0e4e24b48ed43692ea6

Observation 8614fd79-8787-43d5-83b9-21608da229f4 · outbound

This paper cites Zooming without zooming: Region- to-image distillation for fine-grained multimodal perception.arXiv preprint arXiv:2602.11858, 2026.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Zooming without zooming: Region- to-image distillation for fine-grained multimodal perception.arXiv preprint arXiv:2602.11858, 2026

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.357126Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.357126Z digest=sha256:c6bae80d472057ea12967ed97968daf2017fd2c3146d650574e523a638f46e9b

Observation 10131f61-7daf-46c1-82f6-f0f69697db60 · outbound

This paper cites Visual generation unlocks human-like reasoning through multimodal world models.arXiv preprint arXiv:2601.19834, 2026.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Visual generation unlocks human-like reasoning through multimodal world models.arXiv preprint arXiv:2601.19834, 2026

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.362216Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.362216Z digest=sha256:fcf4a5b2de2b9f790a8b18409d2dffe78dc8a542d41415daf3cc2ab05dddc74b

Observation 812332c9-4101-4395-ad7a-532e7f106164 · outbound

This paper cites VTool-R1: VLMs learn to think with images via reinforcement learning on multimodal tool use.arXiv preprint arXiv:2505.19255, 2025.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning VTool-R1: VLMs learn to think with images via reinforcement learning on multimodal tool use.arXiv preprint arXiv:2505.19255, 2025

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.367385Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.367385Z digest=sha256:b04e2f4444350874e41a3887cd6bfdfae6a75a5a7c6ff599902b83fc7f72353a

Observation c1358137-74bd-4884-b10f-9687c7390116 · outbound

This paper cites V*: Guided Visual Search as a Core Mechanism in Multimodal LLMs.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning V*: Guided Visual Search as a Core Mechanism in Multimodal LLMs

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.372654Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.372654Z digest=sha256:c0dc9c387b9d9864f27fcfcad2d44f9e125df8e72c2505f1076eb2f1dc28e88e

Observation 39fdb19d-0830-48e0-bb5b-af6e36d48e4a · outbound

This paper cites Tool-augmented policy optimization.arXiv preprint arXiv:2510.07038, 2025.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Tool-augmented policy optimization.arXiv preprint arXiv:2510.07038, 2025

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.378336Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.378336Z digest=sha256:4186c767fb23344e5c10c7f1409bb980b410a93bacef6f6889f370cc1a91d200

Observation d2e8fa83-702f-44e8-bdb0-0b664623ccfd · outbound

This paper cites Reasoning or Reciting? Exploring the Capabilities and Limitations of Language Models Through Counterfactual Tasks.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Reasoning or Reciting? Exploring the Capabilities and Limitations of Language Models Through Counterfactual Tasks

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.384423Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.384423Z digest=sha256:c578dc1b390a00d9d9d941cf01da92bb06f39a8d263872008a59d2897834a1ae

Observation e0bf9a99-100f-4ea1-aa7d-15c558030ec7 · outbound

This paper cites LLaVA-CoT: Let Vision Language Models Reason Step-by-Step.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning LLaVA-CoT: Let Vision Language Models Reason Step-by-Step

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.390585Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.390585Z digest=sha256:251e96cabb7d5ced1360f09b4970892ac6cdae2c449e2adcb187033c3760a4ba

Observation 68cb5366-89ba-4258-9f1a-99d967af8c4c · outbound

This paper cites Act Wisely: Cultivating Meta-Cognitive Tool Use in Agentic Multimodal Models.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Act Wisely: Cultivating Meta-Cognitive Tool Use in Agentic Multimodal Models

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.396078Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.396078Z digest=sha256:f6ab5c1cb4d3f9d4fc8590a0cff5ed6318af7a11a50b14cfa645dfa547cb264c

Observation 51316cec-d858-4612-af50-38963eb96922 · outbound

This paper cites Set-of-Mark Prompting Unleashes Extraordinary Visual Grounding in GPT-4V.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Set-of-Mark Prompting Unleashes Extraordinary Visual Grounding in GPT-4V

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.403481Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.403481Z digest=sha256:72e9c11f57e6d14f2b84de1feae2fa0f33a901a58c0c3dc9e8645d59d580461d

Observation c22ace32-3b90-4012-b5f8-9e466331462e · outbound

This paper cites Thinking in Space: How Multimodal Large Language Models See, Remember, and Recall Spaces.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Thinking in Space: How Multimodal Large Language Models See, Remember, and Recall Spaces

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.409420Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.409420Z digest=sha256:ffb94478b0cda3ae666475bf421c3975f0ad67f7a6140847d6fe22b8a2935043

Observation 5e606f51-b262-42b5-8073-9e895eb919f4 · outbound

This paper cites VisionThink: Smart and Efficient Vision Language Model via Reinforcement Learning.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning VisionThink: Smart and Efficient Vision Language Model via Reinforcement Learning

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.415257Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.415257Z digest=sha256:261f01bb5f28b70af07185c3071c577f83f00e89b4a9f0db525d203e37214590

Observation af5e4430-ad7c-4533-b666-a57b79c4825d · outbound

This paper cites Look-Back: Implicit Visual Re-focusing in MLLM Reasoning.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Look-Back: Implicit Visual Re-focusing in MLLM Reasoning

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.422649Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.422649Z digest=sha256:2d3266c31d838568a8fd915f057ef04a79c8a0e2a1363fa7f984245bd4d90c8c

Observation c46a5d03-1cee-489d-98bd-52fb90f9d5b4 · outbound

This paper cites Walk the Talk: Bridging the Reasoning-Action Gap for Thinking with Images via Multimodal Agentic Policy Optimization.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Walk the Talk: Bridging the Reasoning-Action Gap for Thinking with Images via Multimodal Agentic Policy Optimization

Reference 87

Resolution
verified exact
local_arxiv, observed 2026-08-11T12:53:15.665161Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-11T12:53:14.428800Z digest=sha256:1448ebbf92e21166f366ae7a8af9bacb9b74d81bf313c48c63d70e94f769be26

Observation 6dc2fa7b-8836-4961-aa4b-e74d4e6e85d4 · outbound

This paper cites Thinking with images via self-calling agent.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Thinking with images via self-calling agent

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.434700Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.434700Z digest=sha256:88b6ea2a016a5f1feb7f2fec70da87524e553d3ddfb771b47afb9981817df17e

Observation ebbc4565-7328-45a9-9e9e-c80012b9984e · outbound

This paper cites Tree of thoughts: Deliberate problem solving with large language models.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Tree of thoughts: Deliberate problem solving with large language models

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.440324Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.440324Z digest=sha256:ca8772066970fc86815f67ccfe40b7b87986740fd75aa21616ab154bbd69b377

Observation 6e03bf6e-d596-4191-833a-6f05b55e852a · outbound

This paper cites React: Synergizing reasoning and acting in language models.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning React: Synergizing reasoning and acting in language models

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.445645Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.445645Z digest=sha256:b893582e7e05d984739cbb0a09babb795c621d084a3d55272f0047cd4f5837d6

Observation bb86f08f-a4d6-4b04-937e-b32d0aaac4c4 · outbound

This paper cites an unresolved cited work.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Unresolved cited work

Reference 91

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.450944Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.450944Z digest=sha256:dbcc3f9e2b0c26e74af19174b4b374f7102bd371aa873b90c35c83de8bbf45fa

Observation 93488707-e50d-4f90-8a7d-95b8d82a412d · outbound

This paper cites ProRL agent: Rollout-as-a-service for reinforcement learning training of multi-turn LLM agents.arXiv preprint arXiv:2603.18815, 2026.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning ProRL agent: Rollout-as-a-service for reinforcement learning training of multi-turn LLM agents.arXiv preprint arXiv:2603.18815, 2026

Reference 92

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.456716Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.456716Z digest=sha256:7a1f66af463f94ba138ee1bbcd20013273fa8068a5175e5b12111308f419f882

Observation 1da67c1b-6c4a-472f-93ea-6f3d8f059e62 · outbound

This paper cites MM-CoT: A benchmark for probing visual chain-of-thought reasoning.arXiv preprint arXiv:2512.08228, 2025.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning MM-CoT: A benchmark for probing visual chain-of-thought reasoning.arXiv preprint arXiv:2512.08228, 2025

Reference 93

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.462467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.462467Z digest=sha256:f1e36d5793f10153a0ceeb740ae197b78e4fa79d9a6177d03c883f3208075aca

Observation 1cde152f-f8de-45dd-8e0e-f00f4a94f582 · outbound

This paper cites LLaVA-Mini: Efficient Image and Video Large Multimodal Models with One Vision Token.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning LLaVA-Mini: Efficient Image and Video Large Multimodal Models with One Vision Token

Reference 94

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.468007Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.468007Z digest=sha256:58dd8719fc395b8149b9b1d9c9c5dd62b433c532160a037a9bf6e594b6c270a3

Observation ef2d5d0a-b1e9-435e-9982-f1221d152fad · outbound

This paper cites MME-RealWorld: Could Your Multimodal LLM Challenge High-Resolution Real-World Scenarios that are Difficult for Humans?.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning MME-RealWorld: Could Your Multimodal LLM Challenge High-Resolution Real-World Scenarios that are Difficult for Humans?

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.474606Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.474606Z digest=sha256:21847ab2599af285bf8e1da33d88ce6122197d0a9ef3614d866804a86375c4d8

Observation c12feb7b-3ba7-47e1-a0bc-5cb150c124f2 · outbound

This paper cites Thyme: Think Beyond Images.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Thyme: Think Beyond Images

Reference 96

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.480733Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.480733Z digest=sha256:ac4bedbea2f4a4b82f23985a869ee7c7b317be7e4e312e8528d25e353acb4809

Observation 4fb149dd-5d29-49f4-a574-6c1c48b07001 · outbound

This paper cites Skywork-r1v4: Toward agentic multimodal intelligence through interleaved thinking with images and deepresearch.arXiv preprint arXiv:2512.02395, 2025.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Skywork-r1v4: Toward agentic multimodal intelligence through interleaved thinking with images and deepresearch.arXiv preprint arXiv:2512.02395, 2025

Reference 97

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.486447Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.486447Z digest=sha256:b0773aa85bf656199a1dd259c3b6460d01e8e6dc7551d550f116fc0a0d0120ab

Observation 3a376dd0-d944-496d-b3da-cffe00b77a20 · outbound

This paper cites CM2: Reinforcement learning with checklist rewards for multi-turn and multi-step agentic tool use.arXiv preprint arXiv:2602.12268, 2026.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning CM2: Reinforcement learning with checklist rewards for multi-turn and multi-step agentic tool use.arXiv preprint arXiv:2602.12268, 2026

Reference 98

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.491968Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.491968Z digest=sha256:35991ad9f34ec92980dc8ecbf97e567e5b934bba45783eca5ac5a863e9639f38

Observation abebf7fd-87a9-4b5f-ad26-07e75afc7002 · outbound

This paper cites Multimodal Chain-of-Thought Reasoning in Language Models.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning Multimodal Chain-of-Thought Reasoning in Language Models

Reference 99

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.497323Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.497323Z digest=sha256:2748e4e6f58375b8e4f9a1b605d906a72968332437b0700ae084ec51d81893b4

Observation 754fb110-bb38-4730-85a6-4d1842ee219b · outbound

This paper cites On Robustness and Chain-of-Thought Consistency of RL-Finetuned VLMs.

Thinking With Tools, Not With Pixels: Tool Calls as Text Scaffolds for Visual Reasoning On Robustness and Chain-of-Thought Consistency of RL-Finetuned VLMs

Reference 100

Resolution
unresolved
no resolver link, observed 2026-08-11T12:53:14.503030Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T12:53:14.503030Z digest=sha256:087f2859b0f73999c239952d86fc0a04c4af5fdac0725fb9a41d8358aa5264e8

Pith citing papers

No inbound Pith citation observations are available.