Pith. sign in

Paper Citation Record · LEDGER

Evaluating LLMs in Database Scenarios: A Lifecycle Benchmark for Assessing Their Potential in Core Database Tasks

As of 18 August 2026, this Paper Citation Record lists 25 of 25 outbound references and 0 inbound Pith citation observations for arXiv:2608.03794.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.03794 v1

Coverage vector

measured 25 of 25 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T12:24:04.602809Z

measured 25 of 25 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

25 of 25 outbound references displayed

  • verified exact0
  • verified fuzzy10
  • unresolved13
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch2

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 7dde2eec-4e90-4b56-95b6-4d89c034f60b · outbound

This paper cites arXiv preprint arXiv:2406.08426 , year=.

Evaluating LLMs in Database Scenarios: A Lifecycle Benchmark for Assessing Their Potential in Core Database Tasks arXiv preprint arXiv:2406.08426 , year=

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-05T12:24:04.479231Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T12:24:04.479231Z digest=sha256:6adcb48420d44d1414690b09f71e918e8f69678416c7f5b7c2b566db9aa624b8

Observation 5bb2a721-f63c-48c5-a2eb-8118064d382a · outbound

This paper cites LLM-Enhanced Data Management.

Evaluating LLMs in Database Scenarios: A Lifecycle Benchmark for Assessing Their Potential in Core Database Tasks LLM-Enhanced Data Management

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-05T12:24:04.484917Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T12:24:04.484917Z digest=sha256:9e9d25df34149e698d3f6174129519ffda90be5ac3ad15045dd7f48483ded36f

Observation a279a8b7-f56a-4594-aa83-7fd1922b9d11 · outbound

This paper cites SURGENT: A Surgical Multi-Agent Assistance System Across the Perioperative Workflow.

Evaluating LLMs in Database Scenarios: A Lifecycle Benchmark for Assessing Their Potential in Core Database Tasks SURGENT: A Surgical Multi-Agent Assistance System Across the Perioperative Workflow

Reference 3

Resolution
metadata mismatch
local_arxiv, observed 2026-08-05T12:24:04.908009Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T12:24:04.492422Z digest=sha256:b50739973aba16e0560ad4474b9792610121dd817971a63dee6db23d90d5a10d

Observation 5805a705-e8a7-4634-92fb-a2a2cec8e469 · outbound

This paper cites Reinforcement Learning for Large Language Model Selective Evidence Adoption from Contaminated Retrieval Results.

Evaluating LLMs in Database Scenarios: A Lifecycle Benchmark for Assessing Their Potential in Core Database Tasks Reinforcement Learning for Large Language Model Selective Evidence Adoption from Contaminated Retrieval Results

Reference 4

Resolution
metadata mismatch
local_arxiv, observed 2026-08-05T12:24:04.879414Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T12:24:04.498229Z digest=sha256:a8c1c1ff65b4b015b0d5eacf503f492d4dfef9e504a08d796744db239b001a99

Observation a9fb90c7-e00e-451b-ac60-af0ba0d4987a · outbound

This paper cites Are Tools All We Need? Unveiling the Tool-Use Tax in LLM Agents.

Evaluating LLMs in Database Scenarios: A Lifecycle Benchmark for Assessing Their Potential in Core Database Tasks Are Tools All We Need? Unveiling the Tool-Use Tax in LLM Agents

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-05T12:24:04.503426Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T12:24:04.503426Z digest=sha256:a410c1f4beacf5f6dda3a59dba05447e0cf8b7dc83d74441a2fbc78461bd129f

Observation 7e88d92e-6dfc-4c03-805d-b79f21490fa6 · outbound

This paper cites 2026 , eprint=.

Evaluating LLMs in Database Scenarios: A Lifecycle Benchmark for Assessing Their Potential in Core Database Tasks 2026 , eprint=

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T12:24:05.238383Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T12:24:04.508724Z digest=sha256:9b8233063eb8586711d8a95bbcf173cda0ae5fdde925034c6cffa5c9916e9bba

Observation b6898bf8-ca94-4f5d-9db3-94522e13528a · outbound

This paper cites Proceedings of the VLDB Endowment , volume=.

Evaluating LLMs in Database Scenarios: A Lifecycle Benchmark for Assessing Their Potential in Core Database Tasks Proceedings of the VLDB Endowment , volume=

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T12:24:05.220507Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T12:24:04.514488Z digest=sha256:3c58e9c1437ec679e783ab237a60843bfbe7fa6d0926a3d8f27b51c7adf0f1fa

Observation b75a72ac-1f7d-46e2-9c1e-86ab5a729ca9 · outbound

This paper cites Data Science and Engineering , volume=.

Evaluating LLMs in Database Scenarios: A Lifecycle Benchmark for Assessing Their Potential in Core Database Tasks Data Science and Engineering , volume=

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-05T12:24:04.520058Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T12:24:04.520058Z digest=sha256:2478a23f6c88f7c0d87fa18145d3a0a31fcd81881f519fb04601e9bc90501f6e

Observation c09c4b7c-5fdd-491a-83e8-39e8d6bf6f9a · outbound

This paper cites S pider: A Large-Scale Human-Labeled Dataset for Complex and Cross-Domain Semantic Parsing and Text-to- SQL Task.

Evaluating LLMs in Database Scenarios: A Lifecycle Benchmark for Assessing Their Potential in Core Database Tasks S pider: A Large-Scale Human-Labeled Dataset for Complex and Cross-Domain Semantic Parsing and Text-to- SQL Task

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-05T12:24:04.524717Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T12:24:04.524717Z digest=sha256:afd539ee2a81b672491cb6fb2b328e843567b85810b008f6b826205fd9525337

Observation 6bdc1cce-72d0-46c2-b27e-e636caba7936 · outbound

This paper cites 2023 , eprint=.

Evaluating LLMs in Database Scenarios: A Lifecycle Benchmark for Assessing Their Potential in Core Database Tasks 2023 , eprint=

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-05T12:24:04.529924Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T12:24:04.529924Z digest=sha256:e8f9604c597007504a19e24ad010f178f5538fe1f5321865675cd4067ee7612c

Observation 2cec246e-a6a0-4e33-b74d-5b5bf449aa0f · outbound

This paper cites arXiv preprint arXiv:2403.02951 , year=.

Evaluating LLMs in Database Scenarios: A Lifecycle Benchmark for Assessing Their Potential in Core Database Tasks arXiv preprint arXiv:2403.02951 , year=

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-05T12:24:04.534698Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T12:24:04.534698Z digest=sha256:d19223f915a02bccab091e7d948a2cebcd2cf21a78d63a261f353b5b2cd88cc2

Observation e0b92cb4-8495-4610-8eec-08636fa4eb78 · outbound

This paper cites NeurIPS 2023 Foundation Models for Decision Making Workshop , year=.

Evaluating LLMs in Database Scenarios: A Lifecycle Benchmark for Assessing Their Potential in Core Database Tasks NeurIPS 2023 Foundation Models for Decision Making Workshop , year=

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T12:24:05.177479Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T12:24:04.540370Z digest=sha256:47c8bffe1c72da1aa77ced803170ea1652f67447c63a68e903d949806521d823

Observation e26e7c73-d027-42de-98cd-1c0d6c922d30 · outbound

This paper cites 2023 , eprint=.

Evaluating LLMs in Database Scenarios: A Lifecycle Benchmark for Assessing Their Potential in Core Database Tasks 2023 , eprint=

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-05T12:24:04.545965Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T12:24:04.545965Z digest=sha256:4495aa06d9ae9e70163c8164f54e0ea52d25aa2b056129edd371269698a7aa13

Observation 85a8127b-44b3-4d00-a259-1e87efba0c2b · outbound

This paper cites Proceedings of the national conference on artificial intelligence , pages=.

Evaluating LLMs in Database Scenarios: A Lifecycle Benchmark for Assessing Their Potential in Core Database Tasks Proceedings of the national conference on artificial intelligence , pages=

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-05T12:24:04.550375Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T12:24:04.550375Z digest=sha256:13e9ff953aedd6d8b648cb633fb0f1a6aeedf2cbd6ab8c0632156c33b1ae2782

Observation a9980328-7a4e-4d3a-93e8-e546b0db2927 · outbound

This paper cites SQLNet: Generating Structured Queries From Natural Language Without Reinforcement Learning.

Evaluating LLMs in Database Scenarios: A Lifecycle Benchmark for Assessing Their Potential in Core Database Tasks SQLNet: Generating Structured Queries From Natural Language Without Reinforcement Learning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-05T12:24:04.555107Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T12:24:04.555107Z digest=sha256:5794630b10ed0b8f4279d6e77438221d0555953b22d407150d0cbfa9f96f5c17

Observation 648490b1-04ea-4a9c-9864-282782699f24 · outbound

This paper cites an unresolved cited work.

Evaluating LLMs in Database Scenarios: A Lifecycle Benchmark for Assessing Their Potential in Core Database Tasks Unresolved cited work

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-05T12:24:04.560040Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T12:24:04.560040Z digest=sha256:6cef3c5f0c10f4e0ad5ad6219394542cddb69d3eda8f8481441e838eaa190f50

Observation 31a827b1-03e7-4493-8904-9a0ba235fde2 · outbound

This paper cites Handbook of linguistic annotation , pages=.

Evaluating LLMs in Database Scenarios: A Lifecycle Benchmark for Assessing Their Potential in Core Database Tasks Handbook of linguistic annotation , pages=

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T12:24:05.136915Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T12:24:04.564610Z digest=sha256:e14f1d6cafe2b84c5726c5fb94686f89a9779661f9d4e6f3a10c955516faa307

Observation 8c829c42-6653-4dbe-89a9-1ba68a7cdca2 · outbound

This paper cites DATABASE DEVELOPMENT LIFE CYCLE , volume =.

Evaluating LLMs in Database Scenarios: A Lifecycle Benchmark for Assessing Their Potential in Core Database Tasks DATABASE DEVELOPMENT LIFE CYCLE , volume =

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T12:24:05.119878Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T12:24:04.569027Z digest=sha256:b5013835a92a28aa7e561cb7e4f48380babf5792c61216401a7de9cb0a606c50

Observation b1337843-6934-4bc5-8749-0e7d31475143 · outbound

This paper cites Statistical methods for rates and proportions , volume=.

Evaluating LLMs in Database Scenarios: A Lifecycle Benchmark for Assessing Their Potential in Core Database Tasks Statistical methods for rates and proportions , volume=

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T12:24:05.102345Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T12:24:04.573243Z digest=sha256:551921d64c99c9e17ca07ad8a1ae404b13712a9b09414102a5ee11b3a8f34fe9

Observation 897a0be7-20a7-4700-a24b-194c89d340e4 · outbound

This paper cites Evaluating the Data Model Robustness of Text-to-SQL Systems Based on Real User Queries.

Evaluating LLMs in Database Scenarios: A Lifecycle Benchmark for Assessing Their Potential in Core Database Tasks Evaluating the Data Model Robustness of Text-to-SQL Systems Based on Real User Queries

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-05T12:24:04.578055Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T12:24:04.578055Z digest=sha256:faf7897f25c06d0b60a27cf2952bea177b1713de0765e38f29d5ea34e5ef1d41

Observation d5b6f41f-e575-4ca8-bd4a-0fe76e152c0c · outbound

This paper cites Clinical and Vaccine Immunology , volume=.

Evaluating LLMs in Database Scenarios: A Lifecycle Benchmark for Assessing Their Potential in Core Database Tasks Clinical and Vaccine Immunology , volume=

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T12:24:05.085613Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T12:24:04.583457Z digest=sha256:b640138a5dd14c6b2e16925cb179c659d4f9ca83e8cba0f928f03c19763a24f9

Observation dd0f4fa8-c510-4714-bf4a-78abdbceaf79 · outbound

This paper cites 2024 , eprint=.

Evaluating LLMs in Database Scenarios: A Lifecycle Benchmark for Assessing Their Potential in Core Database Tasks 2024 , eprint=

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T12:24:05.067608Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T12:24:04.588299Z digest=sha256:f28b6a75ef1b4e7920af3db55c935c2f994053a69b252c5bf62684e174313948

Observation 9e4361f8-5b97-4ca3-aad7-588438dde918 · outbound

This paper cites CoSQL: A Conversational Text-to-SQL Challenge Towards Cross-Domain Natural Language Interfaces to Databases.

Evaluating LLMs in Database Scenarios: A Lifecycle Benchmark for Assessing Their Potential in Core Database Tasks CoSQL: A Conversational Text-to-SQL Challenge Towards Cross-Domain Natural Language Interfaces to Databases

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-05T12:24:04.592803Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T12:24:04.592803Z digest=sha256:a0e16bdd1eecc34da4172599257a312fb2d0cc0e93a05d4fbae5ca87a504d1a8

Observation 52ab411c-8f49-4938-972e-ff1059a96081 · outbound

This paper cites Proceedings of the VLDB Endowment , volume=.

Evaluating LLMs in Database Scenarios: A Lifecycle Benchmark for Assessing Their Potential in Core Database Tasks Proceedings of the VLDB Endowment , volume=

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T12:24:05.050123Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T12:24:04.597769Z digest=sha256:5ca3a58f9c95181d39e89beba0ba4158a7e754748ba4c615d0621d9731238fe1

Observation 12eff54f-5531-473e-ae12-20a86b2ef5be · outbound

This paper cites Proceedings of the 31st ACM SIGKDD Conference on Knowledge Discovery and Data Mining V.

Evaluating LLMs in Database Scenarios: A Lifecycle Benchmark for Assessing Their Potential in Core Database Tasks Proceedings of the 31st ACM SIGKDD Conference on Knowledge Discovery and Data Mining V

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T12:24:05.033702Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-05T12:24:04.602809Z digest=sha256:e4286950f998ef6523f0a9621047f227b784ecb2757371694aab21139fb91ea7

Pith citing papers

No inbound Pith citation observations are available.