REVIEW 27 cited by
Which Economic Tasks are Performed with AI? Evidence from Millions of Claude Conversations
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Which Economic Tasks are Performed with AI? Evidence from Millions of Claude Conversations
read the original abstract
Despite widespread speculation about artificial intelligence's impact on the future of work, we lack systematic empirical evidence about how these systems are actually being used for different tasks. Here, we present a novel framework for measuring AI usage patterns across the economy. We leverage a recent privacy-preserving system to analyze over four million Claude.ai conversations through the lens of tasks and occupations in the U.S. Department of Labor's O*NET Database. Our analysis reveals that AI usage primarily concentrates in software development and writing tasks, which together account for nearly half of all total usage. However, usage of AI extends more broadly across the economy, with approximately 36% of occupations using AI for at least a quarter of their associated tasks. We also analyze how AI is being used for tasks, finding 57% of usage suggests augmentation of human capabilities (e.g., learning or iterating on an output) while 43% suggests automation (e.g., fulfilling a request with minimal human involvement). While our data and methods face important limitations and only paint a picture of AI usage on a single platform, they provide an automated, granular approach for tracking AI's evolving role in the economy and identifying leading indicators of future impact as these technologies continue to advance.
Forward citations
Cited by 27 Pith papers
-
AI Fiction in the Wild
Analysis of 500k ChatGPT logs shows over one-third of conversations generate fiction, dominated by power users with repetitive and niche patterns.
-
Generative AI and the Reorganization of Labor Demand
Firms adjust to generative AI by reallocating hiring (52% of exposure decline) and redesigning tasks within jobs (39.5%), with senior roles shifting earlier via reallocation and junior roles using mixed channels.
-
Synthetic Sociality: How Generative Models Privatize the Social Fabric
Generative models privatize social relations by automating social capacities into synthetic forms owned by private companies.
-
Priming, Path-dependence, and Plasticity: Understanding the molding of user-LLM interaction and its implications from (many) chat logs in the wild
Large-scale analysis of wild LLM chat logs finds that user interaction patterns stabilize quickly after initial use and correlate with long-term outcomes like retention, creating an agency paradox of limited explorati...
-
Measuring and Mitigating Persona Distortions from AI Writing Assistance
AI writing assistance systematically distorts how writers are perceived across 29 social dimensions, and mitigating undesirable distortions reduces user preference for AI-assisted text.
-
Agentic Search in the Wild: Intents and Trajectory Dynamics from 14M+ Real Search Requests
Large-scale log study of 14M+ agentic searches finds short sessions, intent-specific repetition patterns, and that 54% of new query terms trace to prior retrieved evidence.
-
Policies Permitting LLM Use for Polishing Peer Reviews Are Currently Not Enforceable
State-of-the-art AI detectors misclassify a non-trivial fraction of LLM-polished peer reviews as fully AI-generated, rendering polishing-only policies currently unenforceable.
-
Economic Evaluations of Language Models
Using O*NET and 4.5M chatbot conversations plus synthetic prompts, EconEvals measures LM performance on U.S. work activities and predicts substantial time savings in 47% of occupations, with usage lagging.
-
Adopt $\neq$ Adapt: Longitudinal Analyses of LLM Conversations in the Wild
Longitudinal analysis of Bing Copilot users shows sticky individual LLM habits, activity-level differences in task complexity and success, and that WildChat is skewed toward power users.
-
Design and Report Benchmarks for Knowledge Work
Proposes a three-step benchmark design method (define work activity, specify tested setting, score work product) derived from work studies and O*NET, demonstrated via three case analyses.
-
Cognitive offloading and the speedup illusion in human-AI interaction
Preregistered behavioral study identifies a speedup illusion where users overestimate time savings from AI assistance on cognitive tasks despite no actual difference in completion times.
-
The efficiency-gain illusion: People underestimate the rate of AI use and overestimate its benefits on simple tasks
Three pre-registered studies with 2691 participants show people underestimate their AI usage rate and overestimate efficiency gains on simple tasks, with prior use entrenching further adoption.
-
Upskilling with Generative AI: Practices and Challenges for Freelance Knowledge Workers
Freelancers use generative AI to support exploratory skill acquisition but not as their main resource due to reliability issues, leading to a shift toward survival-oriented upskilling and the emergence of invisible co...
-
Measuring and Mitigating Persona Distortions from AI Writing Assistance
AI writing distorts perceived writer personas across 29 dimensions in large experiments, and reward-model mitigation reduces but does not eliminate user preference for the AI.
-
LLMs Corrupt Your Documents When You Delegate
LLMs corrupt an average of 25% of document content during long delegated editing workflows across 52 domains, even frontier models, and agentic tools do not mitigate the issue.
-
LLMs Get Lost In Multi-Turn Conversation
LLMs drop 39% in performance during multi-turn conversations due to premature assumptions and inability to recover from early errors.
-
The Jagged Global Economy: Frontier AI Unevenly Exposes National Economies
National AI exposure, built from occupation scores and ILO employment for 141 countries, is much higher in rich white-collar economies, higher for women in 91% of countries, predicts AI adoption, and rises further via...
-
Informal Learning Emerges in Everyday Human-LLM Interaction
Behavioral markers of informal learning occur in about 4.9% of everyday human-LLM user turns and are more common when assistants provide scaffolded support.
-
AI in the Enterprise: How People Use M365 Copilot Chat
Large-scale classification of M365 Copilot Chat sessions shows writing dominates usage with a shift toward content creation over search, varying by occupation.
-
From Exposure to Adoption: Generative AI in European Workplaces
Generative AI adoption in Europe ranges from under 3% to 25%, is steeper for skilled workers in abstract-task jobs and in digitally advanced countries with training, shows a gender gap in exposed roles, and has produc...
-
Can the Recovery Mechanism Survive AI? Skill Formation, Labor, and What Current Measurement Misses
Generative AI may break the education-based recovery mechanism for technological displacement, as evidence shows performance gains without learning gains and current measurements miss the knowledge dimension of cognition.
-
The AI Skills Shift: Mapping Skill Obsolescence, Emergence, and Transition Pathways in the LLM Era
Benchmarking four LLMs on O*NET skills yields SAFI scores showing mathematics and programming as most automatable while active listening and reading comprehension are least, with 78.7% of real AI interactions being au...
-
Vibe Coding in Product Teams: Reconfiguring AI-Assisted Workflows, Prototyping, and Collaboration
Interviews reveal a four-stage vibe coding workflow that accelerates prototyping while introducing tensions between quick efficiency and reflective design intention, plus asymmetries in trust and ownership.
-
From Model Design to Organizational Design: Complexity Redistribution and Trade-Offs in Generative AI
LLMs relocate rather than eliminate trade-offs among generality, accuracy, and simplicity, shifting complexity to infrastructure, compliance, and expertise and redefining competitive advantage around managing that shift.
-
Can the Recovery Mechanism Survive AI? Skill Formation, Labor, and What Current Measurement Misses
Generative AI risks eroding the developmental process of learning by performing high-level cognitive work, creating a paradox where it helps current workers but may undermine future capacity building, requiring new ou...
-
ASE-26: a curriculum for agentic software engineering as a discipline
ASE-26 is a proposed undergraduate curriculum for agentic software engineering organized around an evolutionary spiral of intent and build, with 21 modules and pedagogical commitments for agent-co-produced work.
-
Measuring the Occupation-Level Impact of AbbVie Intelligence: AI Applicability Analysis, 2024-2025
Empirical analysis of AbbVie internal AI usage finds +10.0% and +6.68% gains in occupation-level applicability scores after platform release and learning summit respectively.
discussion (0)
Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.