Pith. sign in

Paper Citation Record · LEDGER

Spider2-V: How Far Are Multimodal Agents From Automating Data Science and Engineering Workflows?

As of 19 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 14 inbound Pith citation observations for arXiv:2407.10956.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2407.10956 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 14 of 14 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 14 of 14 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-12T12:45:35.813740Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

2
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation fe4ceb79-a92e-41ad-8ded-60e871fff207 · inbound

Probing the limitations of multimodal language models for chemistry and materials research cites this paper.

Probing the limitations of multimodal language models for chemistry and materials research Spider2-V: How Far Are Multimodal Agents From Automating Data Science and Engineering Workflows?

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-12T12:45:35.813740Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:45:35.813740Z digest=sha256:e3323f9d52bf1a491c6873a21ea2263ddfec03fd9a76515d676eabd6036692e2

Observation 0d002445-7e68-4949-85dc-d633fec70aa3 · inbound

DataLab: A Unified Platform for LLM-Powered Business Intelligence cites this paper.

DataLab: A Unified Platform for LLM-Powered Business Intelligence Spider2-V: How Far Are Multimodal Agents From Automating Data Science and Engineering Workflows?

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-11T23:49:05.019142Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:49:05.019142Z digest=sha256:f6259431cb28820e232761beb33955f7df91857e0c68c5adc02e88962f5c0a92

Observation ee8b7b91-7dcb-45f5-8ae3-8d784c227329 · inbound

Aguvis: Unified Pure Vision Agents for Autonomous GUI Interaction cites this paper.

Aguvis: Unified Pure Vision Agents for Autonomous GUI Interaction Spider2-V: How Far Are Multimodal Agents From Automating Data Science and Engineering Workflows?

Reference 70

Resolution
verified exact
arxiv_id, observed 2026-05-18T04:09:41.707996Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-05-18T04:09:41.494136Z digest=sha256:321dbb05b14b932d50b25325e4d35c61b1d068cb6fbba90e5fe0d9e262a139a2

Observation db1c0535-f2d1-4a01-886d-f72fa0a91aef · inbound

LLM4SR: A Survey on Large Language Models for Scientific Research cites this paper.

LLM4SR: A Survey on Large Language Models for Scientific Research Spider2-V: How Far Are Multimodal Agents From Automating Data Science and Engineering Workflows?

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-10T21:39:25.121898Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:39:25.121898Z digest=sha256:f399431e8b939342e50454e307f2ea080ef2d5743e2e15b0565088d0bde1846e

Observation f84ff813-9810-453e-8a9c-b9752049da78 · inbound

Learn-by-interact: A Data-Centric Framework for Self-Adaptive Agents in Realistic Environments cites this paper.

Learn-by-interact: A Data-Centric Framework for Self-Adaptive Agents in Realistic Environments Spider2-V: How Far Are Multimodal Agents From Automating Data Science and Engineering Workflows?

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-10T18:56:44.432956Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T18:56:44.432956Z digest=sha256:dc96f19b25340ccd244fa728c24599dcaa3b44b91320687d5e3d49ee68a67e79

Observation c3e4f24f-8ec2-4020-bcff-8bfeed4951fc · inbound

CoddLLM: Empowering Large Language Models for Data Analytics cites this paper.

CoddLLM: Empowering Large Language Models for Data Analytics Spider2-V: How Far Are Multimodal Agents From Automating Data Science and Engineering Workflows?

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-09T19:30:38.332696Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T19:30:38.332696Z digest=sha256:ae6342f44c110e0d068d2bec8f48edcb68d5470413cb403f0fcfb5b4ed677ee6

Observation 1e40ed52-f436-4e95-88cd-94d4f9cc67fb · inbound

What Limits Virtual Agent Application? OmniBench: A Scalable Multi-Dimensional Benchmark for Essential Virtual Agent Capabilities cites this paper.

What Limits Virtual Agent Application? OmniBench: A Scalable Multi-Dimensional Benchmark for Essential Virtual Agent Capabilities Spider2-V: How Far Are Multimodal Agents From Automating Data Science and Engineering Workflows?

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:29.136593Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:02:29.136593Z digest=sha256:1d26ad51b9743bf10c0e9748aac3a4cc750828a13b329903b43cae1ad6efa410

Observation 63d06783-aef3-42c3-be41-b5d3b2542565 · inbound

Augmented Vision-Language Models: A Systematic Review cites this paper.

Augmented Vision-Language Models: A Systematic Review Spider2-V: How Far Are Multimodal Agents From Automating Data Science and Engineering Workflows?

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-06T14:33:30.799697Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:33:30.799697Z digest=sha256:c20b1c24c6e7e98220ff856673a3aefc12b492c5bad885ca3ad4704467ddf62e

Observation acee0a58-6f68-4292-a0bd-abefcc9bc119 · inbound

Tabular Data Understanding with LLMs: A Survey of Recent Advances and Challenges cites this paper.

Tabular Data Understanding with LLMs: A Survey of Recent Advances and Challenges Spider2-V: How Far Are Multimodal Agents From Automating Data Science and Engineering Workflows?

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-06T10:20:37.438899Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T10:20:37.438899Z digest=sha256:3a4b5f19533f5bacf49a03be0e5d5e7df71c91d0ef07c2267a20d74a32d6ad44

Observation 7de16c15-aa72-44e3-b937-e58b45e968fd · inbound

VLAA-GUI: Knowing When to Stop, Recover, and Search, A Modular Framework for GUI Automation cites this paper.

VLAA-GUI: Knowing When to Stop, Recover, and Search, A Modular Framework for GUI Automation Spider2-V: How Far Are Multimodal Agents From Automating Data Science and Engineering Workflows?

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-09T22:34:07.744901Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-05-09T22:24:45.045405Z digest=sha256:d670a913a840758a46b238b3246fc0ce56daa1cda30abc46ac2e11c506698d8c

Observation be919de6-1da8-4bbd-8c64-ee0c5bb2a52b · inbound

Free Energy-Driven Reinforcement Learning with Adaptive Advantage Shaping for Unsupervised Reasoning in LLMs cites this paper.

Free Energy-Driven Reinforcement Learning with Adaptive Advantage Shaping for Unsupervised Reasoning in LLMs Spider2-V: How Far Are Multimodal Agents From Automating Data Science and Engineering Workflows?

Reference 109

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T07:45:59.254389Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-05-10T16:58:10.013475Z digest=sha256:c60475acbbad28d8fabce11826b8e153f8d610b842430da5f585810c6862d817

Observation 5c7dedbb-d4cf-456c-8c92-71bc11915883 · inbound

Adapt to Thrive! Adaptive Power-Mean Policy Optimization for Improved LLM Reasoning cites this paper.

Adapt to Thrive! Adaptive Power-Mean Policy Optimization for Improved LLM Reasoning Spider2-V: How Far Are Multimodal Agents From Automating Data Science and Engineering Workflows?

Reference 94

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T08:01:00.201698Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-05-10T16:51:19.555272Z digest=sha256:7be51f2653663af5562c8c0cb0802679d764d9319a8628dd9d128e4c12665386

Observation 27c1d722-381a-42cf-9af1-cb2adb75832b · inbound

Business Utility of Large Language Models as Exploratory Data Analysis Agents cites this paper.

Business Utility of Large Language Models as Exploratory Data Analysis Agents Spider2-V: How Far Are Multimodal Agents From Automating Data Science and Engineering Workflows?

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T23:35:06.992881Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-30T23:25:17.416071Z digest=sha256:1f88cee09a6e4fca19857197cfd1390388204ace9e6c81751142f34b3257b652

Observation f85ecd25-818e-4cb0-930d-895be1e9efe6 · inbound

ChainWorld: Composing Long-Horizon Desktop Workloads from Atomic OSWorld Tasks cites this paper.

ChainWorld: Composing Long-Horizon Desktop Workloads from Atomic OSWorld Tasks Spider2-V: How Far Are Multimodal Agents From Automating Data Science and Engineering Workflows?

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-07-04T06:49:38.862300Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-06-26T14:05:28.423719Z digest=sha256:bad2f89ed87cecde57a7d54ad6a3411eb4843f36b6c2d8c228876fa06779b4f3