Pith. sign in

Paper Citation Record · LEDGER

Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

As of 23 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 53 inbound Pith citation observations for arXiv:2411.07763.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.07763 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 53 of 53 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 53 of 53 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T12:40:26.340816Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

5
pith, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 1034f624-a227-4bc0-b124-d186ccd04f7f · inbound

TQA-Bench: Evaluating LLMs for Multi-Table Question Answering cites this paper.

TQA-Bench: Evaluating LLMs for Multi-Table Question Answering Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 68

Resolution
unresolved
no resolver link, observed 2026-08-12T10:09:45.277225Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T10:09:45.277225Z digest=sha256:1743831f534a42f8ecdd29869bceb1056a5f68b49195901d5899cdc2bcfd6b67

Observation a63130b1-013f-4f70-abe9-13334fb68245 · inbound

A Survey of Large Language Model-Based Generative AI for Text-to-SQL: Benchmarks, Applications, Use Cases, and Challenges cites this paper.

A Survey of Large Language Model-Based Generative AI for Text-to-SQL: Benchmarks, Applications, Use Cases, and Challenges Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-11T20:52:21.889548Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T20:52:21.889548Z digest=sha256:9a4830fd15d232bdc5115b0755eacda71bc7dad68db60ca28be988f1ac45bea2

Observation c6331fe4-bad6-41a8-8272-b30ac8f9757b · inbound

Task-Oriented Automatic Fact-Checking with Frame-Semantics cites this paper.

Task-Oriented Automatic Fact-Checking with Frame-Semantics Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-10T16:21:34.727459Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T16:21:34.727459Z digest=sha256:67c1ffbcc326f41c6ab00f00bc33d0d75383558fcda704d0105c52210cf912dc

Observation a82266f5-4043-40a6-b40d-f2ad9964ea33 · inbound

Querying Databases with Function Calling cites this paper.

Querying Databases with Function Calling Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-10T15:25:14.473377Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:25:14.473377Z digest=sha256:9cff5ce65b8aab01b69574dea2989344cd460003ecd5109a1fdbc2c8f2a8b493

Observation 74ac50a3-8a8c-44ed-8691-199067748316 · inbound

ReFoRCE: A Text-to-SQL Agent with Self-Refinement, Consensus Enforcement, and Column Exploration cites this paper.

ReFoRCE: A Text-to-SQL Agent with Self-Refinement, Consensus Enforcement, and Column Exploration Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-09T18:14:11.094332Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T18:14:11.094332Z digest=sha256:fedae63069cf03865b83b64dc4b5cec9b2bccbbb0f774e74b4414001ef7ff9bb

Observation 79fd8e53-0471-4cdf-877f-36f78f7c867c · inbound

Rationalization Models for Text-to-SQL cites this paper.

Rationalization Models for Text-to-SQL Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-08T14:29:19.365305Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T14:29:19.365305Z digest=sha256:8b1fd8d8c2672b049a4abca9b0bce0988a54629f168d6bd839c2291d92aac897

Observation f8778978-308f-4ea2-8441-b1d2d4383553 · inbound

Towards LLM Agents for Earth Observation cites this paper.

Towards LLM Agents for Earth Observation Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-16T12:40:26.340816Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:40:26.340816Z digest=sha256:033cff073b7d69527cfb5e6595d03b210cd0b410d6d10a8e914cdbd434d0035f

Observation f17ccf55-ccbf-4123-941a-488a52dff0b4 · inbound

Griffin: Towards a Graph-Centric Relational Database Foundation Model cites this paper.

Griffin: Towards a Graph-Centric Relational Database Foundation Model Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-15T23:08:10.129608Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T23:08:10.129608Z digest=sha256:b98ce0f991138c30a788a32f41e61377884b3194481e59bc41b3a7f53defa0e1

Observation 89d668b2-0df8-4b7a-a481-5216ec163817 · inbound

LLMs Get Lost In Multi-Turn Conversation cites this paper.

LLMs Get Lost In Multi-Turn Conversation Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-14T01:11:09.242419Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-14T00:57:10.262350Z digest=sha256:b0eb22ba21e5c82268ca9936aa080397dd78e4e1fe99f3f81e123f1543b0e7ee

Observation fbf11f26-cd60-428a-a446-e12bde52f205 · inbound

ExeSQL: Self-Taught Text-to-SQL Models with Execution-Driven Bootstrapping for SQL Dialects cites this paper.

ExeSQL: Self-Taught Text-to-SQL Models with Execution-Driven Bootstrapping for SQL Dialects Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T14:55:58.256867Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:55:58.256867Z digest=sha256:2bf1bdd15d0e4c6865c66a771636009f785ef68043d88cf086f199fa504c2d42

Observation e024c67a-5a68-4ad5-88ad-b959076cdea1 · inbound

Effectiveness of Prompt Optimization in NL2SQL Systems cites this paper.

Effectiveness of Prompt Optimization in NL2SQL Systems Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T13:54:51.120801Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:54:51.120801Z digest=sha256:caeac39c84c31a4950a49a46e21c0e36bd42ba2060eb946cbf1c8a80c51a25d8

Observation a4187268-02c4-45bb-9d82-da1cc7482c43 · inbound

SDE-SQL: Enhancing Text-to-SQL Generation in Large Language Models via Self-Driven Exploration with SQL Probes cites this paper.

SDE-SQL: Enhancing Text-to-SQL Generation in Large Language Models via Self-Driven Exploration with SQL Probes Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T05:43:47.133873Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:43:47.133873Z digest=sha256:a8014304e8f0c178b9dc83e4b04dedc7f7c6320b33c5f700bd57de2bd3c140df

Observation 9897b54d-9cd0-4695-9735-b3970a664fb3 · inbound

Text-to-SQL for Enterprise Data Analytics cites this paper.

Text-to-SQL for Enterprise Data Analytics Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T16:09:53.902437Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:09:53.902437Z digest=sha256:cbdfd1fa57c4cae01da380ed82dc0105450d5691589c32f5866518fcbf608ee6

Observation 2051115a-8445-4664-9d41-afb2d9ec298e · inbound

Multi-turn Natural Language to Graph Query Language Translation cites this paper.

Multi-turn Natural Language to Graph Query Language Translation Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-06T05:26:36.487616Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T05:26:36.487616Z digest=sha256:e9c6f2d3611d3303627402f4104851f771ff1ffded9bbf85098838305be8a96f

Observation efab0355-7913-4e39-941f-dcc5502ff3a8 · inbound

PaVeRL-SQL: Text-to-SQL via Partial-Match Rewards and Verbal Reinforcement Learning cites this paper.

PaVeRL-SQL: Text-to-SQL via Partial-Match Rewards and Verbal Reinforcement Learning Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-04T22:48:35.073263Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T22:48:35.073263Z digest=sha256:a0dfeae08e5f15fa1bd00f85702f6f06bf40787f95c485c50ba798b26d593386

Observation 04b1f38b-6433-4e1b-9549-d770d9239ece · inbound

RAG Strategies for Natural Language-Based SQL Query and REST API Call Generation cites this paper.

RAG Strategies for Natural Language-Based SQL Query and REST API Call Generation Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-04T06:09:51.581974Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T06:09:51.581974Z digest=sha256:bb4cbdfd59f4da155b57b673968ba373cbcc67a0e0a047c5ea59146f195750d2

Observation 915cec4a-cb37-4fbb-b974-0ecf3140f64d · inbound

APEX-SQL: Talking to the data via Agentic Exploration for Text-to-SQL cites this paper.

APEX-SQL: Talking to the data via Agentic Exploration for Text-to-SQL Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-03T01:04:45.439554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T01:04:45.439554Z digest=sha256:57a20dd1f19d79b885105305bcb7b457e6da4094c7e83acc66f214f5d02cee5f

Observation 388c2632-07d7-4255-9e42-948394a1521e · inbound

Both Ends Count! Just How Good are LLM Agents at "Text-to-Big SQL"? cites this paper.

Both Ends Count! Just How Good are LLM Agents at "Text-to-Big SQL"? Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-05-15T19:56:33.693013Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-15T19:52:53.443887Z digest=sha256:c7bf00184dd92a4abc02d1ef38ea333c41fb1b2245d5242745f0c46ab4c644cb

Observation 56145ee9-d42e-42fc-93ed-74cadcf0f2d2 · inbound

SpotIt+: Verification-based Text-to-SQL Evaluation with Database Constraints cites this paper.

SpotIt+: Verification-based Text-to-SQL Evaluation with Database Constraints Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 12

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T16:16:15.241822Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-15T16:11:39.112890Z digest=sha256:739b32c6601d01fee4f0db76ea8b0316f5a55270424289e599323f49b9699504

Observation 3bc20951-2409-44d9-8a7a-06a9437fd7dc · inbound

AV-SQL: Decomposing Complex Text-to-SQL Queries with Agentic Views cites this paper.

AV-SQL: Decomposing Complex Text-to-SQL Queries with Agentic Views Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-11T07:30:57.961185Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-10T17:09:46.385340Z digest=sha256:0d622076224120eadfa79dee209e5af1027439f727d8896ce04e92080e2e9e59

Observation 2517a394-054d-44af-b916-514346aaa1cb · inbound

SynQL: A Controllable and Scalable Rule-Based Framework for SQL Workload Synthesis for Performance Benchmarking cites this paper.

SynQL: A Controllable and Scalable Rule-Based Framework for SQL Workload Synthesis for Performance Benchmarking Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-11T00:15:54.125693Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-10T18:37:37.185685Z digest=sha256:2a863d87074739c78a33c41b26e6acab0cdedbf1d7905b8e442decb11659422b

Observation 881e0945-a0f4-428f-8c88-2350e4dc39ea · inbound

From Natural Language to PromQL: A Catalog-Driven Framework with Dynamic Temporal Resolution for Cloud-Native Observability cites this paper.

From Natural Language to PromQL: A Catalog-Driven Framework with Dynamic Temporal Resolution for Cloud-Native Observability Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-15T10:55:28.626078Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-15T10:53:54.235150Z digest=sha256:a3df610a597294549afdc7cae8667e8a97fb21b197a6c8227ab8a3d0f1cdff01

Observation 9a677d35-59d2-479f-a2b4-d99f4067efee · inbound

SemanticAgent: A Semantics-Aware Framework for Text-to-SQL Data Synthesis cites this paper.

SemanticAgent: A Semantics-Aware Framework for Text-to-SQL Data Synthesis Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-05-11T14:16:17.719480Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-09T22:08:11.410285Z digest=sha256:14ed7aa0ac93ade3593037af98c92c1a1024f130539023f2ae595c0d64d29571

Observation 9983d813-6c8c-4a32-8590-143b0a6d42ab · inbound

From Unstructured Recall to Schema-Grounded Memory: Reliable AI Memory via Iterative, Schema-Aware Extraction cites this paper.

From Unstructured Recall to Schema-Grounded Memory: Reliable AI Memory via Iterative, Schema-Aware Extraction Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-12T10:16:27.514532Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-07T06:58:17.953361Z digest=sha256:20ac8e841b327e528dc497dbb96fc96bcd8e12e9c332b29fbb9e21d26e8943fe

Observation c5231cc4-7d14-4695-84f7-5be85460efc6 · inbound

DataClawBench: An Agent Benchmark for Exploratory Real-World Financial Data Analysis cites this paper.

DataClawBench: An Agent Benchmark for Exploratory Real-World Financial Data Analysis Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-05-21T00:29:17.498001Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-21T00:25:32.898938Z digest=sha256:d4bd161d3f530131f8627633b49c17358954c19e682f92951ffa572d589d3694

Observation c066f39d-65e0-4ec0-94f5-1af9e3767731 · inbound

Anatomy of a Query: W5H Dimensions and FAR Patterns for Text-to-SQL Evaluation cites this paper.

Anatomy of a Query: W5H Dimensions and FAR Patterns for Text-to-SQL Evaluation Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T21:56:11.193352Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-08T03:52:56.830355Z digest=sha256:c495c23745802d5646337013d8115ecb5b3c5ec52312013012fe9f8738cff65f

Observation b497c711-bb22-49c7-a722-a626fa731714 · inbound

LEAF-SQL: Level-wise Exploration with Adaptive Fine-graining for Text-to-SQL Skeleton Prediction cites this paper.

LEAF-SQL: Level-wise Exploration with Adaptive Fine-graining for Text-to-SQL Skeleton Prediction Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-12T06:11:24.461048Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-12T04:31:13.560231Z digest=sha256:47395f4850606684439a927c201fd3cf6fd7ef09ffdd517c06d9ce3116fde344

Observation 6c9e5125-7179-4598-b6c8-ca1b8ffbbd90 · inbound

Consistency as a Testable Property: Statistical Methods to Evaluate AI Agent Reliability cites this paper.

Consistency as a Testable Property: Statistical Methods to Evaluate AI Agent Reliability Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-12T04:41:21.767784Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-12T04:41:15.286881Z digest=sha256:9d29d447b045ad53ba8407fa9db073f29d08b9b4884db49f8cc155e3bff31f8c

Observation 3a53ec0d-dd42-40b1-8ac1-b7af80de05b4 · inbound

Hypergraph Enterprise Agentic Reasoner over Heterogeneous Business Systems cites this paper.

Hypergraph Enterprise Agentic Reasoner over Heterogeneous Business Systems Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 36

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T02:39:41.776490Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-15T02:34:08.115027Z digest=sha256:201f55873bffc8648b19fae1ce40c6b7abbbb8cb3cbcc9f5e5d4c8c19054ca4a

Observation 7770b5c6-5c98-4c8f-85a1-9e61c601db60 · inbound

ClinQueryAgent: A Conversational Agent for Population Health Management cites this paper.

ClinQueryAgent: A Conversational Agent for Population Health Management Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 181

Resolution
verified exact
arxiv_id, observed 2026-05-21T01:33:56.175079Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-21T01:31:07.031424Z digest=sha256:68deb89af18d3ab7dd4d69f143ef7c0f757215c011bde082ea6a217a6e55391f

Observation a712beba-4768-4377-85ac-53e264652809 · inbound

AgentNLQ: A General-Purpose Agent for Natural Language to SQL cites this paper.

AgentNLQ: A General-Purpose Agent for Natural Language to SQL Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 6

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T10:43:12.534798Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-20T10:41:17.671666Z digest=sha256:e7d47054c76d1e4dd30f134b7ba6371936031432e5177eae2d64f3e505c790e6

Observation f50d4636-aee3-47f9-a2df-376013546fd3 · inbound

Residual Skill Optimization for Text-to-SQL Ensembles cites this paper.

Residual Skill Optimization for Text-to-SQL Ensembles Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-22T08:41:17.191735Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-22T08:38:41.126772Z digest=sha256:1266c0c881792fc0a7d40f17411a991ccb8580dfd65c7eb404d69de54c28fac6

Observation afc4827b-2ea0-4c15-ac05-efc47faea0b3 · inbound

Towards Direct Evaluation of Harness Optimizers via Priority Ranking cites this paper.

Towards Direct Evaluation of Harness Optimizers via Priority Ranking Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-22T06:14:40.427387Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-22T06:14:28.559147Z digest=sha256:17df2b7b020f9c684400cdc03c2f6d25ef20a9b59ae4635d0256d21e96e94cc2

Observation 1ae44542-913c-4ded-8187-280efe232b3e · inbound

BADGER: Bridging Agentic and Deterministic Evaluation for Generative Enterprise Reasoning cites this paper.

BADGER: Bridging Agentic and Deterministic Evaluation for Generative Enterprise Reasoning Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-07-01T23:36:23.428244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-28T14:10:10.215779Z digest=sha256:efaf9e603f4d5ffcb09301b18ef258bd72b134accf1964b88429d2e3da6dfbd3

Observation 409365bd-8069-4a7d-b089-bff4a287b850 · inbound

Cross-Vendor Sola ISPM Benchmark: Evaluating Agentic AI for Federated Identity Security Reasoning cites this paper.

Cross-Vendor Sola ISPM Benchmark: Evaluating Agentic AI for Federated Identity Security Reasoning Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-07-01T23:36:24.088845Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-28T14:03:41.903170Z digest=sha256:e8e2006b4e0c7c2b707793b5ad87d230fc410c6fa87049d8e9865a5fa90bbe9a

Observation 00c78482-b748-45cb-b5ec-1880c1eff1a7 · inbound

EntSQL: A Benchmark for Grounding Text-to-SQL in Long-Context Enterprise Knowledge cites this paper.

EntSQL: A Benchmark for Grounding Text-to-SQL in Long-Context Enterprise Knowledge Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 24

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T02:56:29.223357Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-28T10:28:44.425462Z digest=sha256:57c5888943423e21f3bda9406210a6b4f8560ef0009cda45ccaffeceda9bb99c

Observation a257f262-c0d4-4bef-8408-404e94363677 · inbound

SOMA-SQL: Resolving Multi-Source Ambiguity in NL-to-SQL via Synthetic Log and Execution Probing cites this paper.

SOMA-SQL: Resolving Multi-Source Ambiguity in NL-to-SQL via Synthetic Log and Execution Probing Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-07-03T05:27:40.434086Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-27T13:12:15.208886Z digest=sha256:cb90ba9eaede3d94a0a9d4e619dfa741edb716f041194b2ad2135a62e8b3a3b3

Observation cf365e25-a945-49b4-8d59-cc03f9cfcdd9 · inbound

EXPO-SQL: Execution-based Clause-level Policy Optimization for Text-to-SQL cites this paper.

EXPO-SQL: Execution-based Clause-level Policy Optimization for Text-to-SQL Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 43

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T09:05:36.895758Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-07-01T08:57:37.907354Z digest=sha256:a7ce141211df93041b9fd9270ec92ca17d9c85e626eb392b5684981bacf13c89

Observation a77795c3-cfdb-4f5e-a373-805d08371eac · inbound

EcoTable: Cost-effective Table Integration in Data Lakes for Natural Language Queries cites this paper.

EcoTable: Cost-effective Table Integration in Data Lakes for Natural Language Queries Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-07-04T14:49:54.599017Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-26T02:38:20.300232Z digest=sha256:d749feffa012391dc04abe5d9183a91c0867d2362ad71eb7bc4535a935fafba7

Observation 241467c5-305b-4693-a10f-a50c5f9cda19 · inbound

EcoTable: Cost-effective Table Integration in Data Lakes for Natural Language Queries cites this paper.

EcoTable: Cost-effective Table Integration in Data Lakes for Natural Language Queries Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-04T02:32:09.642374Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T02:32:09.642374Z digest=sha256:b232ec5b88c7b8b3eb92a0884e1b82f318235b8e6878717b0bd7c45b71c0f6a9

Observation 03377075-9741-4f4a-bed7-6e2e55f454f2 · inbound

Agents That Know Too Much: A Data-Centric Survey of Privacy in LLM Agents cites this paper.

Agents That Know Too Much: A Data-Centric Survey of Privacy in LLM Agents Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 60

Resolution
verified exact
arxiv_id, observed 2026-07-04T14:09:53.331299Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-26T04:29:16.386339Z digest=sha256:8b34d2e877d471682a60d0db16587362fec820df8141e5e39a54cb57359d068c

Observation 3307d16a-6052-4930-ad9c-95b046699907 · inbound

Database Context Compression for Text-to-SQL on Real-World Large Databases cites this paper.

Database Context Compression for Text-to-SQL on Real-World Large Databases Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-07-01T16:15:49.706860Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-30T00:49:34.390197Z digest=sha256:c10984657780634a242c6be2536cc46f4229eae307042e3aaa2674b9f2b58ff4

Observation e37e7755-79b7-40b5-ab97-94e85651845c · inbound

HETERQA: Benchmarking Record Retrieval over Multiple Heterogeneous Sources cites this paper.

HETERQA: Benchmarking Record Retrieval over Multiple Heterogeneous Sources Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 13

Resolution
unresolved
no resolver link, observed 2026-07-12T05:22:08.124451Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T05:22:08.124451Z digest=sha256:778f970d2a4b47b3e4599aa82da703201197bf13611a3ca342631da156976714

Observation accbbfbd-9398-4ecf-b3a9-703a47c1159b · inbound

Knowing When to Stop: Predicting Execution-Consistency Convergence in Text-to-SQL cites this paper.

Knowing When to Stop: Predicting Execution-Consistency Convergence in Text-to-SQL Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 41

Resolution
unresolved
no resolver link, observed 2026-07-11T22:28:12.204856Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T22:28:12.204856Z digest=sha256:d0051e96c5fffaa81ec17fd6ecdb564d8edbbb94589c4fc7525d59b5a923ced1

Observation e6c0fe81-e8ff-4c21-bed6-874cf017c4ef · inbound

Spider 2.0-AIFunc: Extending Real-World Text-to-SQL to AI-Native SQL Workflows cites this paper.

Spider 2.0-AIFunc: Extending Real-World Text-to-SQL to AI-Native SQL Workflows Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-07-08T12:54:52.865449Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-07-08T12:54:03.626597Z digest=sha256:5f9aedbd7a86b69dc68fe701ae953027d30257de423474072a00c68e2f5bcf92

Observation 9581471f-df7e-412b-bcab-5aa81e059e75 · inbound

Agentic Data Environments cites this paper.

Agentic Data Environments Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-07-09T12:46:14.556133Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-07-09T12:43:21.070615Z digest=sha256:560ee0a676f618fbabd80359f1d2172f91ede4c104c6dfa8bc3e92d79d204820

Observation 4dca58f8-d8fb-46b6-86a4-0e9ca5fd6a08 · inbound

Theory-Level Autoformalization: From Isolated Statements to Unified Formal Knowledge Bases cites this paper.

Theory-Level Autoformalization: From Isolated Statements to Unified Formal Knowledge Bases Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-02T05:40:05.524603Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T05:40:05.524603Z digest=sha256:81c726f21074163e825e6bb5e07180211956146afc602ea0620afb50bba4ef8b

Observation cd2fd891-db85-4804-8328-12092df6cdcc · inbound

I-Rex: An Interactive Debugger for SQL cites this paper.

I-Rex: An Interactive Debugger for SQL Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-01T20:58:07.765905Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T20:58:07.765905Z digest=sha256:1a58074102f206a8f6414543b0ccb79ed01518de683c889ebf7fd9305495324f

Observation 6f04a2a8-578d-432d-9f21-5e31cede2bb7 · inbound

Open Security Benchmark: Towards Autonomous Enterprise Cyber Defense cites this paper.

Open Security Benchmark: Towards Autonomous Enterprise Cyber Defense Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-01T10:23:19.750534Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T10:23:19.750534Z digest=sha256:a64f3a5c452324f94091021102b7574367573fd906f279f86b201ad7eba220f7

Observation 4dbfa781-8d4a-4829-b50f-6f2f70f2821d · inbound

When Does Trace-Driven Evaluation Mislead MoE Expert Caching? Replay Semantics, Workload Contamination, and Operating Regimes cites this paper.

When Does Trace-Driven Evaluation Mislead MoE Expert Caching? Replay Semantics, Workload Contamination, and Operating Regimes Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-12T00:48:38.366878Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T00:48:38.366878Z digest=sha256:94345d2574409553667bc91392523c2ff481ea4d547cb4becf735858acd7506c

Observation 36d43df2-ca41-4b99-86b9-4d7d426d9d4b · inbound

When Does Trace-Driven Evaluation Mislead MoE Expert Caching? Replay Semantics, Workload Contamination, and Operating Regimes cites this paper.

When Does Trace-Driven Evaluation Mislead MoE Expert Caching? Replay Semantics, Workload Contamination, and Operating Regimes Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-14T04:43:19.354311Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:43:19.354311Z digest=sha256:b55039a806ebc74a89e6c57c6ac0a537d5b50479c07cd1971ab98b83175774e4

Observation 00253971-4295-4dca-9e1c-a354250b56be · inbound

Business Truth, not SQL Accuracy: A Rule-Gated 7B Analytics Agent Outperforms a Direct-Prompted 32B Baseline cites this paper.

Business Truth, not SQL Accuracy: A Rule-Gated 7B Analytics Agent Outperforms a Direct-Prompted 32B Baseline Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-11T20:46:34.912755Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T20:46:34.912755Z digest=sha256:b9ec7aa120d4778afacc5f9c93c6412a4cd47666b07572ce792ff0dfd272115f

Observation bde77146-5afd-484e-ab4c-053449edeb6d · inbound

DSAgentBench: Can Agents Automate End-to-End Data-Science Workflows in Real Computer Environments? cites this paper.

DSAgentBench: Can Agents Automate End-to-End Data-Science Workflows in Real Computer Environments? Spider 2.0: Evaluating Language Models on Real-World Enterprise Text-to-SQL Workflows

Reference 90

Resolution
unresolved
no resolver link, observed 2026-08-15T14:25:46.501908Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T14:25:46.501908Z digest=sha256:d2fd12f076149f163c09690472e110f4e8700839edd64ecfee0d068f5742d915