REVIEW 7 cited by
Conversational Challenges in AI-Powered Data Science: Obstacles, Needs, and Design Opportunities
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
Large Language Models (LLMs) are being increasingly employed in data science for tasks like data preprocessing and analytics. However, data scientists encounter substantial obstacles when conversing with LLM-powered chatbots and acting on their suggestions and answers. We conducted a mixed-methods study, including contextual observations, semi-structured interviews (n=14), and a survey (n=114), to identify these challenges. Our findings highlight key issues faced by data scientists, including contextual data retrieval, formulating prompts for complex tasks, adapting generated code to local environments, and refining prompts iteratively. Based on these insights, we propose actionable design recommendations, such as data brushing to support context selection, and inquisitive feedback loops to improve communications with AI-based assistants in data-science tools.
Forward citations
Cited by 7 Pith papers
-
IntentLint: Supporting Intent Scaffolding and Prompt-time Linting in Human-AI Collaborative Data Analysis
IntentLint uses shared, editable rules to scaffold analytic intent and lint prompts, and a user study reports improved perceived collaboration awareness.
-
Steering Semantic Data Processing With DocWrangler
An IDE for LLM-powered text data processing, with user studies showing that people convert open-ended operations into structured classifiers and use vague prompts to explore their data.
-
Flowco: Rethinking Data Analysis in the Age of LLMs
Flowco combines visual dataflow graphs with LLM-generated code, validation checks, and unit tests to help analysts author, debug, and refine data analyses.
-
Representing Visualization Insights as a Dense Insight Network
A five-category framework links dashboard insights into a dense network and is demonstrated in a playground and an LLM-based summarization case study.
-
DataLab: A Unified Platform for LLM-Powered Business Intelligence
DataLab is a unified notebook-based platform for LLM-powered BI tasks that shows strong efficiency gains and competitive accuracy, but its state-of-the-art claim is not supported on several benchmarks.
-
Effective LLM-Driven Code Generation with Pythoness
Pythoness is a test- and specification-driven embedded DSL that generates, validates, and caches LLM code, with a single example showing tests greatly improve pass rates.
-
A Comprehensive Survey of Deep Research: Systems, Methodologies, and Applications
A survey of 80+ Deep Research systems that proposes a four-layer taxonomy (foundation models, tool use, planning, synthesis) and compares commercial and open-source implementations.
Discussion (0). Continue with ORCID to comment.