Pith. sign in

Paper Citation Record · LEDGER

Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

As of 14 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 50 inbound Pith citation observations for arXiv:2411.07763.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.07763 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 50 of 50 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 50 of 50 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-14T04:43:19.354311Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

5
pith, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 1034f624-a227-4bc0-b124-d186ccd04f7f · inbound

TQA-Bench: Evaluating LLMs for Multi-Table Question Answering cites this paper.

TQA-Bench: Evaluating LLMs for Multi-Table Question Answering Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-12T10:09:45.277225Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:09:45.277225Z digest=sha256:3943d1c3fbc184b49a0b37964e0eb4f5f39a940b6629b08474444a9ac1da9973

Observation a63130b1-013f-4f70-abe9-13334fb68245 · inbound

A Survey of Large Language Model-Based Generative AI for Text-to-SQL: Benchmarks, Applications, Use Cases, and Challenges cites this paper.

A Survey of Large Language Model-Based Generative AI for Text-to-SQL: Benchmarks, Applications, Use Cases, and Challenges Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-11T20:52:21.889548Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:52:21.889548Z digest=sha256:f14eaf269eea6ce7490ff76cffd1e1a551a4fe0ee6db2aafbcd07f2b980b2183

Observation c6331fe4-bad6-41a8-8272-b30ac8f9757b · inbound

Task-Oriented Automatic Fact-Checking with Frame-Semantics cites this paper.

Task-Oriented Automatic Fact-Checking with Frame-Semantics Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-10T16:21:34.727459Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T16:21:34.727459Z digest=sha256:3a1d6d6868f6ffec0d822892e978f2c2e69ab8e81b74041959ef8105ecec0096

Observation a82266f5-4043-40a6-b40d-f2ad9964ea33 · inbound

Querying Databases with Function Calling cites this paper.

Querying Databases with Function Calling Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-10T15:25:14.473377Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:25:14.473377Z digest=sha256:095449c52cf7e87755a85bcf94fd39f5215710061e697ef931638e35bc14f07b

Observation 74ac50a3-8a8c-44ed-8691-199067748316 · inbound

ReFoRCE: A Text-to-SQL Agent with Self-Refinement, Consensus Enforcement, and Column Exploration cites this paper.

ReFoRCE: A Text-to-SQL Agent with Self-Refinement, Consensus Enforcement, and Column Exploration Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-09T18:14:11.094332Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:14:11.094332Z digest=sha256:addf2d7248051d439dbae05739f55d5f7a70eada37def086e5081ab94ae840b7

Observation 79fd8e53-0471-4cdf-877f-36f78f7c867c · inbound

Rationalization Models for Text-to-SQL cites this paper.

Rationalization Models for Text-to-SQL Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-08T14:29:19.365305Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:29:19.365305Z digest=sha256:bf519a40e790f0d7c026a474f3696f719b563cbd96befbc54d18df9abbd5dc96

Observation 89d668b2-0df8-4b7a-a481-5216ec163817 · inbound

LLMs Get Lost In Multi-Turn Conversation cites this paper.

LLMs Get Lost In Multi-Turn Conversation Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-14T01:11:09.242419Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-14T00:57:10.262350Z digest=sha256:f247cd87d5251ff5d5648aea7b43338f7664dcf85dea0956e77c484917dd889f

Observation fbf11f26-cd60-428a-a446-e12bde52f205 · inbound

ExeSQL: Self-Taught Text-to-SQL Models with Execution-Driven Bootstrapping for SQL Dialects cites this paper.

ExeSQL: Self-Taught Text-to-SQL Models with Execution-Driven Bootstrapping for SQL Dialects Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T14:55:58.256867Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:55:58.256867Z digest=sha256:1183be188851038ebd18898c9ba72c759bda7c50aaad2ddd4f0fd734923d5910

Observation e024c67a-5a68-4ad5-88ad-b959076cdea1 · inbound

Effectiveness of Prompt Optimization in NL2SQL Systems cites this paper.

Effectiveness of Prompt Optimization in NL2SQL Systems Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:51.120801Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:51.120801Z digest=sha256:6556bacf59d19fd17bd8159e17e121bccf0259773575611515b7c9eb600c5fb5

Observation a4187268-02c4-45bb-9d82-da1cc7482c43 · inbound

SDE-SQL: Enhancing Text-to-SQL Generation in Large Language Models via Self-Driven Exploration with SQL Probes cites this paper.

SDE-SQL: Enhancing Text-to-SQL Generation in Large Language Models via Self-Driven Exploration with SQL Probes Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T05:43:47.133873Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:43:47.133873Z digest=sha256:e3ebafcfaef74871b981a011a6536a24fbb66c09bfcf61b26d8b39c9be5437fb

Observation 9897b54d-9cd0-4695-9735-b3970a664fb3 · inbound

Text-to-SQL for Enterprise Data Analytics cites this paper.

Text-to-SQL for Enterprise Data Analytics Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T16:09:53.902437Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:09:53.902437Z digest=sha256:fce1110b1826491f6e5cd9d3d8d0df715ec8d50ed236723e2db88aa6cba1be86

Observation 2051115a-8445-4664-9d41-afb2d9ec298e · inbound

Multi-turn Natural Language to Graph Query Language Translation cites this paper.

Multi-turn Natural Language to Graph Query Language Translation Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-06T05:26:36.487616Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T05:26:36.487616Z digest=sha256:1d5648264430f84b81825dfcb788fff8acc16b5b31a163a36acd215997e9af7d

Observation efab0355-7913-4e39-941f-dcc5502ff3a8 · inbound

PaVeRL-SQL: Text-to-SQL via Partial-Match Rewards and Verbal Reinforcement Learning cites this paper.

PaVeRL-SQL: Text-to-SQL via Partial-Match Rewards and Verbal Reinforcement Learning Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-04T22:48:35.073263Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T22:48:35.073263Z digest=sha256:20afb4932cdaa78a44a69343fe435b621b51d437caaf2627e92a1cce3a47b04a

Observation 04b1f38b-6433-4e1b-9549-d770d9239ece · inbound

RAG Strategies for Natural Language-Based SQL Query and REST API Call Generation cites this paper.

RAG Strategies for Natural Language-Based SQL Query and REST API Call Generation Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-04T06:09:51.581974Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T06:09:51.581974Z digest=sha256:a49fed96a241f745f18b8ad8b27216f0a505933c6db58cf2882c8b14f3540202

Observation 915cec4a-cb37-4fbb-b974-0ecf3140f64d · inbound

APEX-SQL: Talking to the data via Agentic Exploration for Text-to-SQL cites this paper.

APEX-SQL: Talking to the data via Agentic Exploration for Text-to-SQL Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-03T01:04:45.439554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T01:04:45.439554Z digest=sha256:4629a2c431f9c8443d2ecf0a703c2cef98cc1daf117093342d6fcc8f31e77aa8

Observation 388c2632-07d7-4255-9e42-948394a1521e · inbound

Both Ends Count! Just How Good are LLM Agents at "Text-to-Big SQL"? cites this paper.

Both Ends Count! Just How Good are LLM Agents at "Text-to-Big SQL"? Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-15T19:56:33.693013Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-15T19:52:53.443887Z digest=sha256:8221f9f4b887abb473fbdf1f4090928f5ba3195e5b03c602a20e5b8e0a04cc5b

Observation 56145ee9-d42e-42fc-93ed-74cadcf0f2d2 · inbound

SpotIt+: Verification-based Text-to-SQL Evaluation with Database Constraints cites this paper.

SpotIt+: Verification-based Text-to-SQL Evaluation with Database Constraints Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 12

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T16:16:15.241822Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-15T16:11:39.112890Z digest=sha256:83e9414141f6c6c3bc9e1785b684e8cde934d2c09a5994e63ad7cadce44ad16b

Observation 3bc20951-2409-44d9-8a7a-06a9437fd7dc · inbound

AV-SQL: Decomposing Complex Text-to-SQL Queries with Agentic Views cites this paper.

AV-SQL: Decomposing Complex Text-to-SQL Queries with Agentic Views Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-11T07:30:57.961185Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-10T17:09:46.385340Z digest=sha256:33ed0313554cafcfe5c97f2e683ec7e2d099d8f65af47b824e93209819365bd7

Observation 2517a394-054d-44af-b916-514346aaa1cb · inbound

SynQL: A Controllable and Scalable Rule-Based Framework for SQL Workload Synthesis for Performance Benchmarking cites this paper.

SynQL: A Controllable and Scalable Rule-Based Framework for SQL Workload Synthesis for Performance Benchmarking Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-11T00:15:54.125693Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-10T18:37:37.185685Z digest=sha256:afc946d2d9746ef4c75f6ffd02a1c1e4fbf8fa30c57fb014c814dc7f60ef077f

Observation 881e0945-a0f4-428f-8c88-2350e4dc39ea · inbound

From Natural Language to PromQL: A Catalog-Driven Framework with Dynamic Temporal Resolution for Cloud-Native Observability cites this paper.

From Natural Language to PromQL: A Catalog-Driven Framework with Dynamic Temporal Resolution for Cloud-Native Observability Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-15T10:55:28.626078Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-05-15T10:53:54.235150Z digest=sha256:e310ed7fb64194f626f331bd25261c67988ba2ee3cb2169329fd050e65404e57

Observation 9a677d35-59d2-479f-a2b4-d99f4067efee · inbound

SemanticAgent: A Semantics-Aware Framework for Text-to-SQL Data Synthesis cites this paper.

SemanticAgent: A Semantics-Aware Framework for Text-to-SQL Data Synthesis Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-05-11T14:16:17.719480Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-09T22:08:11.410285Z digest=sha256:3db0a3f8e66c61db609e94c9bf3d587469c1a46500376becd7e3a0f430fe3f7a

Observation 9983d813-6c8c-4a32-8590-143b0a6d42ab · inbound

From Unstructured Recall to Schema-Grounded Memory: Reliable AI Memory via Iterative, Schema-Aware Extraction cites this paper.

From Unstructured Recall to Schema-Grounded Memory: Reliable AI Memory via Iterative, Schema-Aware Extraction Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-12T10:16:27.514532Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-07T06:58:17.953361Z digest=sha256:d13386fb98380b516925198d93152b9bb19faac7586f553bac47f992955a764a

Observation c5231cc4-7d14-4695-84f7-5be85460efc6 · inbound

DataClawBench: An Agent Benchmark for Exploratory Real-World Financial Data Analysis cites this paper.

DataClawBench: An Agent Benchmark for Exploratory Real-World Financial Data Analysis Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-05-21T00:29:17.498001Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-21T00:25:32.898938Z digest=sha256:e92d119543b98c99b12d5411e69b2145586ebe5cf67615e1ef94fcdcf98e02ea

Observation c066f39d-65e0-4ec0-94f5-1af9e3767731 · inbound

Anatomy of a Query: W5H Dimensions and FAR Patterns for Text-to-SQL Evaluation cites this paper.

Anatomy of a Query: W5H Dimensions and FAR Patterns for Text-to-SQL Evaluation Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T21:56:11.193352Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-08T03:52:56.830355Z digest=sha256:f307cfd4670ae05bf354ffea9a56a9f3874a79771aaeebafe36b6dfb60f8be0d

Observation b497c711-bb22-49c7-a722-a626fa731714 · inbound

LEAF-SQL: Level-wise Exploration with Adaptive Fine-graining for Text-to-SQL Skeleton Prediction cites this paper.

LEAF-SQL: Level-wise Exploration with Adaptive Fine-graining for Text-to-SQL Skeleton Prediction Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-12T06:11:24.461048Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-12T04:31:13.560231Z digest=sha256:fea0261ed6e3b7b51abffb615ee851abfd73f91dd7c0e605ea963bd11f68cc2b

Observation 6c9e5125-7179-4598-b6c8-ca1b8ffbbd90 · inbound

Consistency as a Testable Property: Statistical Methods to Evaluate AI Agent Reliability cites this paper.

Consistency as a Testable Property: Statistical Methods to Evaluate AI Agent Reliability Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-12T04:41:21.767784Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-12T04:41:15.286881Z digest=sha256:ed44c4134990f942af736ce31e16ffee2e8b16468a9bea13099cf87a05271218

Observation 3a53ec0d-dd42-40b1-8ac1-b7af80de05b4 · inbound

Hypergraph Enterprise Agentic Reasoner over Heterogeneous Business Systems cites this paper.

Hypergraph Enterprise Agentic Reasoner over Heterogeneous Business Systems Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 36

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T02:39:41.776490Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-15T02:34:08.115027Z digest=sha256:fba518df26ac96f8b14bb6cbf34ea1b781d42fb91d5bc72a915348de97e04943

Observation 7770b5c6-5c98-4c8f-85a1-9e61c601db60 · inbound

ClinQueryAgent: A Conversational Agent for Population Health Management cites this paper.

ClinQueryAgent: A Conversational Agent for Population Health Management Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 181

Resolution
verified exact
arxiv_id, observed 2026-05-21T01:33:56.175079Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-05-21T01:31:07.031424Z digest=sha256:e603a91d24e84be74cec0419e0bef4286a33c60bf6829b617dc157854739b52e

Observation a712beba-4768-4377-85ac-53e264652809 · inbound

AgentNLQ: A General-Purpose Agent for Natural Language to SQL cites this paper.

AgentNLQ: A General-Purpose Agent for Natural Language to SQL Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 6

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T10:43:12.534798Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-20T10:41:17.671666Z digest=sha256:ff14fd3b8369715458fa61ba93f5b3acb764de5cfd7909eba89198b22ffec7d7

Observation f50d4636-aee3-47f9-a2df-376013546fd3 · inbound

Residual Skill Optimization for Text-to-SQL Ensembles cites this paper.

Residual Skill Optimization for Text-to-SQL Ensembles Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-22T08:41:17.191735Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-22T08:38:41.126772Z digest=sha256:c913694aacdff39f5e4e7516f8c2c590794bdd0f9bd01df879dfc13c9fed0341

Observation afc4827b-2ea0-4c15-ac05-efc47faea0b3 · inbound

Towards Direct Evaluation of Harness Optimizers via Priority Ranking cites this paper.

Towards Direct Evaluation of Harness Optimizers via Priority Ranking Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-22T06:14:40.427387Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-22T06:14:28.559147Z digest=sha256:3d6011756f5da0976f0e33deaa1d4e9f107bcf24c4349c275a30dfa719e48c51

Observation 1ae44542-913c-4ded-8187-280efe232b3e · inbound

BADGER: Bridging Agentic and Deterministic Evaluation for Generative Enterprise Reasoning cites this paper.

BADGER: Bridging Agentic and Deterministic Evaluation for Generative Enterprise Reasoning Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-07-01T23:36:23.428244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-28T14:10:10.215779Z digest=sha256:fe0e9c9c1cde6d6ab40726670d1077d7f663109cb4458c351528e8647e937d4d

Observation 409365bd-8069-4a7d-b089-bff4a287b850 · inbound

Cross-Vendor Sola ISPM Benchmark: Evaluating Agentic AI for Federated Identity Security Reasoning cites this paper.

Cross-Vendor Sola ISPM Benchmark: Evaluating Agentic AI for Federated Identity Security Reasoning Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-07-01T23:36:24.088845Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-28T14:03:41.903170Z digest=sha256:8bde4870c8f4b32edaa109a0e61be143bd8eb46f3851af65a8bfdbedd3ebc10d

Observation 00c78482-b748-45cb-b5ec-1880c1eff1a7 · inbound

EntSQL: A Benchmark for Grounding Text-to-SQL in Long-Context Enterprise Knowledge cites this paper.

EntSQL: A Benchmark for Grounding Text-to-SQL in Long-Context Enterprise Knowledge Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 24

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T02:56:29.223357Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-06-28T10:28:44.425462Z digest=sha256:ad885723f76a3c1c0d24c32fc0f01ecd408541be7ea1d56510e1eaa2c1ac8b44

Observation a257f262-c0d4-4bef-8408-404e94363677 · inbound

SOMA-SQL: Resolving Multi-Source Ambiguity in NL-to-SQL via Synthetic Log and Execution Probing cites this paper.

SOMA-SQL: Resolving Multi-Source Ambiguity in NL-to-SQL via Synthetic Log and Execution Probing Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-07-03T05:27:40.434086Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-27T13:12:15.208886Z digest=sha256:12aa6151de682a2733aac074c60a860fe7798566005e6b3b55882d7305c022ef

Observation cf365e25-a945-49b4-8d59-cc03f9cfcdd9 · inbound

EXPO-SQL: Execution-based Clause-level Policy Optimization for Text-to-SQL cites this paper.

EXPO-SQL: Execution-based Clause-level Policy Optimization for Text-to-SQL Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 43

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T09:05:36.895758Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-07-01T08:57:37.907354Z digest=sha256:f0d73938d815a1344910a5258a23689ec8268f30b53f14f0647fa26e7df218a4

Observation a77795c3-cfdb-4f5e-a373-805d08371eac · inbound

EcoTable: Cost-effective Table Integration in Data Lakes for Natural Language Queries cites this paper.

EcoTable: Cost-effective Table Integration in Data Lakes for Natural Language Queries Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-07-04T14:49:54.599017Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-26T02:38:20.300232Z digest=sha256:451a4d1b338b187c986318b30a1375a3a3033f4097dfdcd7fda21e7d7b2348a7

Observation 241467c5-305b-4693-a10f-a50c5f9cda19 · inbound

EcoTable: Cost-effective Table Integration in Data Lakes for Natural Language Queries cites this paper.

EcoTable: Cost-effective Table Integration in Data Lakes for Natural Language Queries Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-04T02:32:09.642374Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T02:32:09.642374Z digest=sha256:576dd51a6f6e4f03ae48a290115199e93843ba568bc5af5a5cbddc2e57cb60f1

Observation 03377075-9741-4f4a-bed7-6e2e55f454f2 · inbound

Agents That Know Too Much: A Data-Centric Survey of Privacy in LLM Agents cites this paper.

Agents That Know Too Much: A Data-Centric Survey of Privacy in LLM Agents Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 60

Resolution
verified exact
arxiv_id, observed 2026-07-04T14:09:53.331299Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-26T04:29:16.386339Z digest=sha256:87808849d31369cbbec9e77d45e8dfcf5a5de85ffb71ff268814463a5dce17ca

Observation 3307d16a-6052-4930-ad9c-95b046699907 · inbound

Database Context Compression for Text-to-SQL on Real-World Large Databases cites this paper.

Database Context Compression for Text-to-SQL on Real-World Large Databases Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-07-01T16:15:49.706860Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-30T00:49:34.390197Z digest=sha256:ff408e13acb43897a1878fa53f13c99e0893782f2df7f52cf5a5823eca1baec7

Observation e37e7755-79b7-40b5-ab97-94e85651845c · inbound

HETERQA: Benchmarking Record Retrieval over Multiple Heterogeneous Sources cites this paper.

HETERQA: Benchmarking Record Retrieval over Multiple Heterogeneous Sources Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 13

Resolution
unresolved
no resolver link, observed 2026-07-12T05:22:08.124451Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T05:22:08.124451Z digest=sha256:f56a606f50d94103e2fe0d216092e933e7e4a452f812fc04356710625aeb60f1

Observation accbbfbd-9398-4ecf-b3a9-703a47c1159b · inbound

Knowing When to Stop: Predicting Execution-Consistency Convergence in Text-to-SQL cites this paper.

Knowing When to Stop: Predicting Execution-Consistency Convergence in Text-to-SQL Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 41

Resolution
unresolved
no resolver link, observed 2026-07-11T22:28:12.204856Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T22:28:12.204856Z digest=sha256:13c51fc5b0be88a7de4dec4f3cebd45ca451a3a0c59b504d50a1ba418a1a873c

Observation e6c0fe81-e8ff-4c21-bed6-874cf017c4ef · inbound

Spider 2.0-AIFunc: Extending Real-World Text-to-SQL to AI-Native SQL Workflows cites this paper.

Spider 2.0-AIFunc: Extending Real-World Text-to-SQL to AI-Native SQL Workflows Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-07-08T12:54:52.865449Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=arxiv_source observed=2026-07-08T12:54:03.626597Z digest=sha256:fa5ff42f58330c3de50ffdfedba337bc5352d7085db9deb24fc8c1393f1f33c6

Observation 9581471f-df7e-412b-bcab-5aa81e059e75 · inbound

Agentic Data Environments cites this paper.

Agentic Data Environments Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-07-09T12:46:14.556133Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-07-09T12:43:21.070615Z digest=sha256:556fc73f801c0884bf902939a0409c01f1677cf0136351a59744277021a53e0f

Observation 4dca58f8-d8fb-46b6-86a4-0e9ca5fd6a08 · inbound

Theory-Level Autoformalization: From Isolated Statements to Unified Formal Knowledge Bases cites this paper.

Theory-Level Autoformalization: From Isolated Statements to Unified Formal Knowledge Bases Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-02T05:40:05.524603Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T05:40:05.524603Z digest=sha256:c5385385f507dcfb5570235d485fe1cb1317b86a35c9113a9d3bbf04b4d5b9a1

Observation cd2fd891-db85-4804-8328-12092df6cdcc · inbound

I-Rex: An Interactive Debugger for SQL cites this paper.

I-Rex: An Interactive Debugger for SQL Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-01T20:58:07.765905Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T20:58:07.765905Z digest=sha256:6abbc991325767e30075626c9c4a20bc7d97b26705ab43159e4f0f40210e19c7

Observation 6f04a2a8-578d-432d-9f21-5e31cede2bb7 · inbound

Open Security Benchmark: Towards Autonomous Enterprise Cyber Defense cites this paper.

Open Security Benchmark: Towards Autonomous Enterprise Cyber Defense Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-01T10:23:19.750534Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T10:23:19.750534Z digest=sha256:f8c2e3ccc93d2e40f098ae2c3b14cf5ba57b22e32ad5dd19a05fd8a3f6ab56ee

Observation 4dbfa781-8d4a-4829-b50f-6f2f70f2821d · inbound

When Does Trace-Driven Evaluation Mislead MoE Expert Caching? Replay Semantics, Workload Contamination, and Operating Regimes cites this paper.

When Does Trace-Driven Evaluation Mislead MoE Expert Caching? Replay Semantics, Workload Contamination, and Operating Regimes Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-12T00:48:38.366878Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T00:48:38.366878Z digest=sha256:c565071ccb755bcfad0229efdb79c58f32fa0e556e07a42a57d94bc360c1176b

Observation 36d43df2-ca41-4b99-86b9-4d7d426d9d4b · inbound

When Does Trace-Driven Evaluation Mislead MoE Expert Caching? Replay Semantics, Workload Contamination, and Operating Regimes cites this paper.

When Does Trace-Driven Evaluation Mislead MoE Expert Caching? Replay Semantics, Workload Contamination, and Operating Regimes Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-14T04:43:19.354311Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:43:19.354311Z digest=sha256:47530a1b04f2f9962224bef66c6fe912a5003de52230019af4517c0d06ccb636

Observation 00253971-4295-4dca-9e1c-a354250b56be · inbound

Business Truth, not SQL Accuracy: A Rule-Gated 7B Analytics Agent Outperforms a Direct-Prompted 32B Baseline cites this paper.

Business Truth, not SQL Accuracy: A Rule-Gated 7B Analytics Agent Outperforms a Direct-Prompted 32B Baseline Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-11T20:46:34.912755Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T20:46:34.912755Z digest=sha256:04f80c3f6278208d9a519a7a4f6bf0342ed3730324888def149fa93266a4b4cd