REVIEW 5 major objections 6 minor 39 references
Return of the solo author: The changing division of labor in science in the age of generative AI
T0 review · 5 major / 6 minor · reviewed 2026-07-14 · grok-4.5
Pith's one-line read The long decline in solo-authored science papers halted and partially reversed after ChatGPT's late-2022 release.
desk verdict Solid large-scale descriptive break in the solo-authorship tail around late 2022, carefully multi-checked; the LLM-substitution reading is correlational and only partly bounded against database and multi-shock confounds. read the letter →
The pith
A machine-rendered reading of the paper's core claim, the machinery that carries it, and where it could break.
The reading
What carries the argument
The solo-authored left tail of the author-count distribution, treated as an observable probe of labor substitution: a solo paper is work completed without credited human coauthors, so changes in its share, who produces it, and what it contains mark the boundary of tasks generative AI can take over.
What would settle it
If independent measures of LLM adoption by field and author, or a later window after journal policies and indexing stabilize, show no corresponding solo-share break once venue composition, author disambiguation, and preprint volume are held fixed, the substitution reading would fail.
Extended reading notes
Core claim
The decades-long decline in the share of solo-authored papers halted and partially reversed around ChatGPT's public release in late 2022. The break is broad across most fields and publication filters, is not explained by new entrants or field-mix changes, and concentrates among authors who had recently or never published alone. Recovered solo papers stay close to the authors' own earlier coauthored content, contract in breadth, and shift toward computational topics, consistent with generative AI substituting for execution labor that coauthors once supplied.
Load-bearing premise
That the shared timing of the break at ChatGPT's public release, together with the field ordering by task substitutability, can be read as evidence of LLM substitution rather than concurrent confounds such as post-pandemic publishing shifts or database coverage changes.
Editorial analysis
A structured set of objections, weighed in public.
Referee Report
Summary. The paper uses the full OpenAlex corpus (1990–2025; ~300M works, 26 fields) to study the left tail of the author-count distribution rather than mean team size. It reports that the long decline in the share of solo-authored papers flattens or partially reverses around ChatGPT’s public release (Nov 2022), with a companion flattening of mean author counts in many fields. The break is heterogeneous: stronger in fields where coauthor tasks are more substitutable by writing/coding/analysis tools, and weak or absent in lab/instrument-heavy fields. Author-level, composition-standardized probabilities show the rebound among previously coauthor-only and never-solo authors, including seniors. Content analyses (SPECTER2 DiD netting field-wide drift; within-author breadth/exploration) find post-2022 solo papers tilt toward computational work, stay near authors’ prior collaborative content, and narrow in scope. The authors interpret solo authorship as an observable probe of AI substitution for parts of coauthor labor, while stating the design is correlational and anchored on a dating convention.
Significance. If the descriptive break and its author/content signatures hold, this is a substantial contribution to the science of science and to empirical work on generative AI’s effect on knowledge production. Focusing on the solo tail rather than the mean is a clear conceptual advance, and the multi-filter design (peer-reviewed, core, preprint, clean all), composition standardization, history-conditioned hazards P_t(k), balanced venue panels, event-study/donut timing checks, and embedding DiD are serious empirical strengths. The paper also carefully separates acceleration vs substitution predictions and assigns them to different parts of the author-count distribution. The result is falsifiable in principle (field ordering, never-solo hazards, content geometry) and would matter for authorship norms, credit, and training pathways even if the causal attribution remains incomplete.
major comments (5)
- [Abstract; Results (mechanism); Limitations; SI D, L, M] The central interpretive claim—that the 2022 left-tail break is an empirical probe of LLM substitution for coauthor labor—rests on a common temporal anchor plus field ordering and content signatures (Abstract; Results “The mechanism is consistent with LLM substitution”; Discussion). The manuscript correctly labels the design correlational and a dating convention (Results; Limitations). Even so, Abstract/title framing and the mechanism section still invite a causal reading that SI D, L, and M only partially bound. Concurrent shocks (post-pandemic publishing, other 2022–23 AI tools, journal policy shifts) remain inside the post window. Please either (i) demote the substitution language consistently to “timing-consistent with” and lead with the descriptive break, or (ii) add a sharper multi-shock falsification (e.g., placebos at other 2020–2024 AI/tool releases; field-by-field comparison of
- [Limitations; SI Section L; Materials and Methods (author IDs)] OpenAlex infrastructure changes are a load-bearing confound risk for authorship counts. SI L shows the peer-reviewed break attenuates from +1.72 to +0.75 pp/yr on balanced 2018–2024 venues (Spearman 0.84 across fields), which bounds venue entry/exit but explicitly cannot rule out within-venue metadata or author-disambiguation changes. MAG discontinuation (end-2021) and the July 2023 author-disambiguation revision fall inside or adjacent to the post window (SI L; Limitations). Because solo status is defined from author IDs, disambiguation revisions can mechanically create or destroy solo papers. A main-text robustness that freezes author IDs / re-runs on a pre-revision snapshot, or at least reports sensitivity of Δβ to excluding mid-2023, is needed before the within-author switching claim can be treated as database-robust.
- [Abstract; Results; SI Sections H and L] Field-level and author-level breaks do not coincide in Engineering, the largest share-level rebound (+2.5 pp/yr peer-reviewed), where author-level solo probability barely moves and the balanced-venue break falls to +0.65 (SI H, L; Results). The paper notes this, but main-text claims that “the break appears among authors who had written only with others” and that composition “does not explain the result” (Abstract; Results) over-generalize. Engineering should be treated as an explicit exception in the Abstract/Results, and the substitution narrative should be restricted to fields where within-author Δβ is positive (e.g., CS, Psychology, Economics in SI Fig. S8).
- [Results (mechanism); Discussion; SI Section K] SI Section K (external Liang et al. corpus) finds multi-author papers carry at least as much estimated LLM-modified writing as solo papers in every arXiv field. That is a useful check against a pure surface-editing account of the solo rebound, but it is currently buried and not integrated with the main mechanism claim. If writing assistance is pervasive on teams, the substitution story must emphasize non-writing execution (coding, analysis, drafting structure) or the decision margin to publish without coauthors—not text polish. Please move a concise version of this result into the main Results/Discussion and state what it does and does not identify.
- [Limitations; Results (preprint vs peer-reviewed); Abstract] Quality and paper-mill inflation remain open for the left-tail recovery (Limitations). The rebound is largest for preprints, where lag is shortest but low-cost solo output is also easiest. Peer-reviewed and core filters still show positive breaks with CIs excluding zero, which is important, yet venue-based filters cannot verify refereeing quality. Without any quality/impact/retraction/duplicate screen on recovered solo papers, the claim of a “reconfiguration of cognitive labor” risks conflating genuine substitution with an influx of low-value solo output. At minimum, report citation or journal-tier distributions for pre- vs post-2022 solo papers under the peer-reviewed filter, or flag quality as outside scope more prominently in the Abstract.
minor comments (6)
- [Figure 1; SI Section C] Figure 1 reports Δβ as change in annual slope of the solo share; clarify in the caption whether monthly fits are rescaled to annual units and how seasonal adjustment (used for mean authors in SI C) is handled for the solo share.
- [Materials and Methods; SI Section F] Eq. (1)–(2) for P_t(k) and d_it are clear, but the main text sometimes uses d_active without restating that it is predetermined and resets after every solo year; a one-sentence reminder would help non-specialist readers.
- [Figure 3; SI Section I] Figure 3a UMAP atlas is descriptive only (as stated), but the two language-defined clusters (Turkish, Indonesian) are easy to misread as topical; consider a footnote or inset noting they are excluded from the English-only axis analysis.
- [Acknowledgments; Materials and Methods] The Acknowledgments disclose extensive Claude use for code and manuscript revision. Given the paper’s topic, a brief Methods note on which analyses were AI-assisted vs human-verified would strengthen reproducibility norms the paper itself discusses.
- [Results; SI Figure S3] SI Fig. S3 heatmap significance uses bootstrap conditional on the observed source universe; state that limitation once in the main text when citing field-level significance counts (23/26 fields).
- [Title] Minor wording: title truncates “generative A” in the provided header; ensure the published title is complete (“generative AI”).
Circularity Check
No significant circularity: observational trend breaks, composition standardization, and embedding DiD are estimated from external OpenAlex data against an external timestamp, not forced by construction or self-citation.
full rationale
This is correlational bibliometric social science, not a first-principles derivation. The load-bearing quantities (field-level solo share, author-level P(solo|history), mean author count, SPECTER2 DiD displacement, keyword-axis projections) are computed from the OpenAlex snapshot; the November 2022 ChatGPT release is an external dating convention, not a fitted target. Pre/post slopes, composition cells (field × age × citations × productivity), history thresholds d_it ≥ k, and solo×post interactions are standard estimators whose outputs are not definitionally identical to their inputs. Semantic axes are keyword-anchored then validated on non-anchor papers and field rankings; the DiD nets field-wide drift rather than tautologically recovering the keywords. There are no self-citations that carry the central claim, no uniqueness theorems, no ansatz smuggled from prior author work, and no renaming of a known pattern as a new derivation. Interpretive reading of the break as LLM substitution is acknowledged by the paper as correlational (Limitations) and is a causal-inference concern, not circularity by construction. Score 0 is the honest finding.
Assumptions & free parameters
free parameters (6)
- ChatGPT release cutoff (Nov 2022 / year 2022)
- Symmetric four-year and asymmetric rolling windows for Δβ
- History threshold k / d_active for never-recent-solo conditioning
- Composition cells (field × academic age × citations × productivity)
- Author output exclusion thresholds (>2000 lifetime works or >50/year)
- Content cohort sampling caps (≤5000 papers/field and top subfields) and UMAP/grid smoothing choices
assumptions (6)
- domain assumption OpenAlex author IDs, primary-field assignments, and venue metadata are sufficiently accurate for field-level and author-year solo rates over 1990–2025 despite known disambiguation and coverage changes.
- domain assumption Venue-based peer-reviewed/core/preprint filters are valid proxies for publication quality strata even though OpenAlex lacks a true peer-review flag.
- domain assumption A solo-authored paper is work completed without credited human coauthors and thus a usable behavioral boundary for substitution of scientific labor.
- domain assumption SPECTER2 embeddings plus keyword-anchored axes capture scientifically meaningful content shifts (especially computational vs equipment work) net of field-wide drift.
- ad hoc to paper Field differences in coauthor-task substitutability (writing/coding/stats vs lab/instrument work) explain heterogeneity in the solo rebound.
- standard math Piecewise linear trends with paper-count weights adequately represent the change in solo-share dynamics around the cutoff.
invented entities (2)
-
Solo-specific content displacement (DiD of solo vs team cohorts in embedding space)
-
History-conditioned solo-authorship probability P_t(k)
independent evidence
Cite this review
Pith. "Pith review of Return of the solo author: The changing division of labor in science in the age of generative AI." pith.science (2026). https://pith.science/paper/PTPIBFBH
@misc{pith2026260710780,
author = {Pith},
title = {Pith review of: Return of the solo author: The changing division of labor in science in the age of generative AI},
year = {2026},
howpublished = {\url{https://pith.science/paper/PTPIBFBH}},
note = {Machine review of arXiv:2607.10780}
}
read the original abstract
Modern science has experienced a long shift from individual work to team production. Generative artificial intelligence (AI) might appear to extend this trajectory by lowering research costs and enabling larger-scale collaboration. Yet if tasks once performed by coauthors can be delegated to AI, the same technology may also weaken the need for collaboration in parts of the research process. Here, we examine this tension by moving beyond average team size and focusing on the solo-authored tail of the author-count distribution. Analyzing over 300 million works across 26 fields, we find that the decades-long decline in solo authorship halted and partially reversed with ChatGPT's public release in late 2022. We also reveal that this is an uneven phenomenon: it is strongest in fields where coauthors' work is more readily replaceable, and weak or absent in fields that depend on physical collaboration. At the individual level, the recovery is not explained by the entry of new researchers or by changes in field composition. Instead, the break appears among authors who had written only with others, including those with no prior solo publications, and among long-established authors as well as newcomers. Their solo papers stay close to their own coauthored work while narrowing in scope and shifting toward computational topics. Because a solo paper is work without credited human coauthors, this study offers an empirical probe of how generative AI can substitute for scientific labor, and evidence of a reconfiguration of cognitive labor within papers rather than of team size.
Figures
Reference graph
Works this paper leans on
-
[1]
death of the renaissance man
BF Jones, The burden of knowledge and the “death of the renaissance man”: Is inno- vation getting harder?Rev. Econ. Stud.76, 283–317 (2009)
2009
-
[2]
S Wuchty, BF Jones, B Uzzi, The increasing dominance of teams in production of knowledge.Science316, 1036–1039 (2007)
2007
-
[3]
C Haeussler, H Sauermann, Division of labor in collaborative knowledge production: The role of team size and interdisciplinarity.Research Policy49, 103987 (2020)
2020
-
[4]
S Fortunato, et al., Science of science.Science359, eaao0185 (2018)
2018
-
[5]
(Cambridge University Press), (2021)
D Wang, AL Barab ´asi,The Science of Science. (Cambridge University Press), (2021)
2021
-
[6]
E Leahey, From sole investigator to team scientist: Trends in the practice and study of research collaboration.Annu. Rev. Sociol.42, 81–100 (2016)
2016
-
[7]
M Thelwall, N Maflahi, Research coauthorship 1900–2020: Continuous, universal, and ongoing expansion.Quantitative Science Studies3, 331–344 (2022)
1900
-
[8]
R Guimer `a, B Uzzi, J Spiro, LAN Amaral, Team assembly mechanisms determine collaboration network structure and team performance.Science308, 697–702 (2005)
2005
Show all 39 references
-
[9]
S Milojevi ´c, Principles of scientific research team formation and evolution.Proc. Natl. Acad. Sci.111, 3984–3989 (2014)
2014
-
[10]
L Wu, D Wang, JA Evans, Large teams develop and small teams disrupt science and technology.Nature566, 378–382 (2019)
2019
-
[11]
P Cunningham, P Smyth, B Smyth, Shifting norms in scholarly publications: Trends in readability, objectivity, authorship, and AI use (arXiv:2510.21725) (2025)
2025
-
[12]
F Xu, L Wu, J Evans, Flat teams drive scientific innovation.Proceedings of the National Academy of Sciences119, e2200927119 (2022)
2022
-
[13]
S Noy, W Zhang, Experimental evidence on the productivity effects of generative artificial intelligence.Science381, 187–192 (2023)
2023
-
[14]
KZ Cui, et al., The effects of generative AI on high-skilled work: Evidence from three field experiments with software developers.Management Science(2026)
2026
-
[15]
A Korinek, Generative AI for economic research: Use cases and implications for economists.Journal of Economic Literature61, 1281–1317 (2023)
2023
-
[16]
P Praus, A note on the topic of single-author articles in science.Scientometrics130, 3071–3088 (2025)
2025
-
[17]
A Goolsbee, C Syverson, Fear, lockdown, and diversion: Comparing drivers of pan- demic economic decline 2020.Journal of Public Economics193, 104311 (2021). 35
2020
-
[18]
R Chetty, JN Friedman, M Stepner, The economic impacts of COVID-19: Evidence from a new public database built using private sector data.Quarterly Journal of Eco- nomics139, 829–889 (2024)
2024
-
[19]
T Eloundou, S Manning, P Mishkin, D Rock, GPTs are GPTs: Labor market impact potential of LLMs.Science384, 1306–1308 (2024)
2024
-
[20]
EW Felten, M Raj, R Seamans, Occupational Heterogeneity in Exposure to Generative AI, (SSRN Working Paper), Technical report (2023)
2023
-
[21]
G Tripodi, et al., Tenure and research trajectories.Proceedings of the National Academy of Sciences122, e2500322122 (2025)
2025
-
[22]
D Kobak, R Gonz ´alez-M´arquez, E ´A Horv ´at, J Lause, Delving into LLM-assisted writing in biomedical publications through excess vocabulary.Science Advances11, eadt3813 (2025)
2025
-
[23]
W Liang, et al., Quantifying large language model usage in scientific papers.Nature Human Behaviour9, 2599–2609 (2025)
2025
-
[24]
M Kwiek, W Roszka, Are female scientists less inclined to publish alone? the gender solo research gap.Scientometrics127, 1697–1735 (2022)
2022
-
[25]
T Lazebnik, A Rosenfeld, How lonely or influential is the lone wolf? an analysis of individual scholars’ solo-authorship dynamics.Scientometrics130, 3053–3069 (2025)
2025
-
[26]
A Zeng, et al., Increasing trend of scientists to switch between topics.Nature Commu- nications10, 3439 (2019)
2019
-
[27]
L McInnes, J Healy, J Melville, UMAP: Uniform manifold approximation and projec- tion for dimension reduction (arXiv:1802.03426) (2018)
2018 arXiv
-
[28]
Social Studies of Science46, 417–435 (2016)
V Larivi `ere, et al., Contributorship and division of labor in knowledge production. Social Studies of Science46, 417–435 (2016)
2016
-
[29]
HW Shen, AL Barab ´asi, Collective credit allocation in science.Proceedings of the Na- tional Academy of Sciences111, 12325–12330 (2014)
2014
-
[30]
S Jabbehdari, JP Walsh, Authorship Norms and Project Structures in Science.Science, Technology, & Human Values42, 872–900 (2017)
2017
-
[31]
Nature Editorial, Tools such as ChatGPT threaten transparent science; here are our ground rules for their use.Nature613, 612 (2023)
2023
-
[32]
M Andal ´on, C de Fontenay, DK Ginther, K Lim, The rise of teamwork and career prospects in academic science.Nature Biotechnology42, 1314–1319 (2024)
2024
-
[33]
JP Alperin, J Portenoy, K Demes, V Larivi`ere, S Haustein, An analysis of the suitability of OpenAlex for bibliometric analyses (arXiv:2404.17663) (2024). 36
2024 arXiv
-
[34]
JH Culbert, et al., Reference coverage analysis of OpenAlex compared to Web of Sci- ence and Scopus.Scientometrics130, 2475–2492 (2025)
2025
-
[35]
J Priem, H Piwowar, R Orr, OpenAlex: A fully-open index of scholarly works, au- thors, venues, institutions, and concepts (arXiv:2205.01833) (2022)
2022 arXiv
-
[36]
H Bouamor, J Pino, K Bali
A Singh, M D’Arcy, A Cohan, D Downey, S Feldman, SciRepEval: A multi-format benchmark for scientific document representations inProceedings of the 2023 Conference on Empirical Methods in Natural Language Processing, eds. H Bouamor, J Pino, K Bali. (Association for Computationa...
2023
-
[37]
(Association for Com- putational Linguistics, Stroudsburg, PA, USA), (2020)
A Cohan, S Feldman, I Beltagy, D Downey, D Weld, SPECTER: Document-level rep- resentation learning using citation-informed transformers inProceedings of the 58th Annual Meeting of the Association for Computational Linguistics. (Association for Com- putational Linguistics, Stro...
2020
-
[38]
OurResearch, We’re building a replacement for Mi- crosoft Academic Graph (https://blog.openalex.org/ were-building-a-replacement-for-microsoft-academic-graph/) (2021) Accessed July 11, 2026
2021
-
[39]
OpenAlex, Author disambiguation, OpenAlex Support (https://help.openalex.org/hc/en-us/articles/ 24347048891543-Author-disambiguation) (n.d.) Accessed July 11, 2026. 37
2026
Reviewed July 14, 2026 · model on record in the stance chip above.
Discussion (0). Continue with ORCID to comment.