LLM-assisted synthesis of database storage readers that bypass engines and materialize PostgreSQL/MySQL data as Apache Arrow for analytical engines.
Title resolution pending
6 Pith papers cite this work. Polarity classification is still indexing.
citation-role summary
citation-polarity summary
years
2026 6roles
baseline 1polarities
baseline 1representative citing papers
ReCAP enables relational DBMS like DuckDB to push property constraints deep into path query plans for speedups up to 400,000x over state-of-the-art graph and relational systems.
DiagramBank is a large-scale curated dataset of 89,422 schematic diagrams from scientific papers with rich metadata to support multimodal retrieval and exemplar-driven figure generation.
TVA presents a multi-version temporal graph storage architecture with temporal tables, hopscotch hashing, and version-skipping that claims up to 9.9x lower query latency and 2.2x lower storage overhead than prior systems.
Quantized open-weight LMs on consumer hardware match closed-source API accuracy for LM-enhanced relational operators while delivering 390x lower cost and 3.8x lower latency in the BlendSQL framework.
JEDI is a generated benchmark suite converting SQL queries into Java Stream and imperative implementations to evaluate performance and identify efficient parallelization strategies.
citing papers explorer
-
Breaking Database Lock-in: Agentic Regeneration of High Performance Storage Readers for Database Bypass
LLM-assisted synthesis of database storage readers that bypass engines and materialize PostgreSQL/MySQL data as Apache Arrow for analytical engines.
-
Efficient Path Query Processing in Relational Database Systems
ReCAP enables relational DBMS like DuckDB to push property constraints deep into path query plans for speedups up to 400,000x over state-of-the-art graph and relational systems.
-
DiagramBank: A Large-scale Dataset of Diagram Design Exemplars with Paper Metadata for Retrieval-Augmented Generation
DiagramBank is a large-scale curated dataset of 89,422 schematic diagrams from scientific papers with rich metadata to support multimodal retrieval and exemplar-driven figure generation.
-
TVA: A Version-aware Temporal Graph Storage System for Real-time Analytics
TVA presents a multi-version temporal graph storage architecture with temporal tables, hopscotch hashing, and version-skipping that claims up to 9.9x lower query latency and 2.2x lower storage overhead than prior systems.
-
Large Databases Need Small, Open-Weight Language Models
Quantized open-weight LMs on consumer hardware match closed-source API accuracy for LM-enhanced relational operators while delivering 390x lower cost and 3.8x lower latency in the BlendSQL framework.
-
JEDI: Java Evaluation of Declarative and Imperative Queries
JEDI is a generated benchmark suite converting SQL queries into Java Stream and imperative implementations to evaluate performance and identify efficient parallelization strategies.