REVIEW 3 major objections 5 minor 49 references
Open Sourcing GPTs: Economics of Open Sourcing Advanced AI Models
T0 review · 3 major / 5 minor · reviewed 2026-08-10 · deepseek-v4-flash
Pith's one-line read Owners are most likely to open source an LLM at moderate size, the paper argues.
desk verdict A careful empirical package with a genuinely new latent-space method and a solid LLaMA event study, but the headline inverted-U rests on a single numerical calibration with no sensitivity analysis. read the letter →
The pith
A machine-rendered reading of the paper's core claim, the machinery that carries it, and where it could break.
The reading
What carries the argument
The carrying object is a two-state dynamic programming problem with Bellman equations for the value of an open model, $V^O(q)$, and a closed model, $V^C(q, q_B)$, where $q$ is the model's quality and $q_B$ is the quality of the open-source alternative. Open-sourcing is irreversible and lets external compute contribute to next-period quality; a closed model is priced through an API, and the price influences how fast the rival model improves. Proposition 1 establishes a unique quality threshold $q^*$: the firm open sources when $q < q^*$ and stays closed when $q > q^*$. Proposition 2 shows the firm sets its API price below the static revenue-maximizing level to slow the rival's growth. Varying $m$, the mass of applications the firm owns, produces the inverted-U relationship between firm size and the width of the open-source window.
What would settle it
One check would be to sweep the model's parameters, especially the open-source ecosystem efficiency $\phi$ and discount factor $\beta$, and ask whether the open-source window is hump-shaped in the firm-size parameter $m$ on a broad grid; the current paper reports the shape only at its baseline parameter values. A complementary empirical check would estimate the open-sourcing decision against a continuous measure of the owner's share of compatible applications across a larger sample of model releases, rather than the categorical Big-Tech indicator used in the regression.
Extended reading notes
Core claim
The author models a tech firm that owns all software producers in a segment of a downstream application sector and can irreversibly open source its higher-quality LLM or license it through an API. Open sourcing sacrifices current API revenue but lets internal and external compute improve the model's quality over time; a closed strategy earns immediate license fees while allowing the rival open model to catch up. Solving the dynamic discrete-choice model shows that the firm opens the model when its quality lead over the open alternative is modest and keeps it closed when the lead is large. The same mechanism delivers the paper's central prediction: as the owner's share of compatible applications grows, openness first rises and then falls, because small firms need API revenue, dominant firms already internalize most applications, and intermediate firms gain the most from community-accelerated growth.
Load-bearing premise
The inverted-U shape is demonstrated by numerical simulation at one illustrative parameter vector, so the claim rests on the assumption that the qualitative shape survives across plausible values of the discount factor, ecosystem efficiency, and firm size rather than being an artifact of that particular calibration.
Editorial extensions
If this is right
- A model with a large quality lead over the best open source alternative is unlikely to be open sourced, which matches the persistence of closed frontier models.
- Large tech firms should open source more often than other for-profit firms, since their broad application portfolios internalize more of the community's quality improvements.
- Open sourcing a frontier model can be an R&D catalyst, with the paper's estimates implying a 40 to 140 percent increase in contributions by LLM researchers after the release studied.
- Moderate concentration in the application market, not low or high concentration, is what should sustain a vibrant open source ecosystem for general-purpose models.
Reading between the lines
- The framework implies a testable prediction beyond the paper's sample: as the best open-source alternative improves, the absolute quality lead at which a firm switches from open to closed should also rise, so the open-source window shifts but does not disappear.
- The same logic should apply to other general-purpose software whose owners also sell compatible applications, suggesting that ownership share of complements, not firm revenue, is the right predictor of open-sourcing.
- The paper stops short of welfare analysis; if open-sourcing accelerates research activity, a social planner might favor openness even for models with large quality leads, because the spillover compounds over time.
Signed reviews
Editorial analysis
A structured set of objections, weighed in public.
Referee Report
Summary. This paper studies why for-profit firms open-source advanced LLMs. It constructs a latent technology space from USPTO patent CPC codes and a fine-tuned LLM classifier to show that LLMs are compatible with many technologically diverse firms. A regression on CRFM ecosystem data finds that a larger quality lead over the best open-source model reduces the probability of open-sourcing, that for-profit developers are less likely to open-source, and that Big Tech firms are more likely to do so. An event-study and difference-in-differences analysis around the release of LLaMA finds increased GitHub contributions among LLM researchers. The paper then develops a dynamic discrete-choice model in which a firm owning a share m of LLM-compatible applications trades off closed API revenue against the accelerated quality growth made possible by open-source contributions. The model produces threshold open-source windows, and a numerical value-function iteration predicts an inverted-U relationship between firm size m and open-sourcing propensity. The paper concludes that moderate market concentration may benefit open-source ecosystems of multi-purpose software technologies.
Significance. If the results hold, the paper offers several contributions: a new latent-space method based on hierarchical patent classifications, plausibly causal evidence that open-sourcing LLaMA stimulated research-related activity, and a theoretical framework connecting GPT-like applicability to open-sourcing decisions. The empirical DiD analysis is carefully checked with synthetic DiD, narrow event windows, and alternative treatment definitions. However, the central theoretical prediction is currently supported only by a single numerical experiment at parameters the paper says should not be taken literally, with no sensitivity analysis or reproducible code. Moreover, the model's 'predictions aligned with empirical findings' are not independent confirmations because the model was explicitly motivated by those same findings. The significance of the paper is therefore conditional on strengthening the theoretical robustness and on clarifying the status of the empirical-theoretical alignment.
major comments (3)
- [Section 6.3, Figure 10 and Table 6] The inverted-U relationship between firm size and open-sourcing propensity is computed by value-function iteration at the single parameter vector in Table 6, and Section 6.2 states that these values 'should not be taken literally.' The paper provides no analytical characterization of the shape in m, no sensitivity analysis over β, γ, α, ψ, ϕ, λ, and cD, and no code (Appendix C.2 only refers to 'accompanying code'). Because Figure C.2 already shows that the open-source window is zero for sufficiently small ϕ, the inverted-U could be an artifact of the baseline calibration. This is the paper's headline prediction, so it needs either a proof of single-peakedness under stated conditions or a systematic sensitivity analysis with reproducible code and data.
- [Section 5.1, Table 2] The quality-lead regression is estimated on 86 models with available MMLU scores, and score availability is likely correlated with model prominence and open/closed status. The Big-Tech dummy is defined by only three firms, and the specifications include no time fixed effects or developer-level clustering. Since the theoretical framework in Section 6 is explicitly motivated by these empirical regularities, sample selection in this table weakens the empirical grounding. Please report robustness to selection/score imputation, time fixed effects, and alternative clustering, or temper the interpretation accordingly.
- [Section 6.3 with Section 1] The model is introduced as 'motivated by these findings' and its output is then described as providing 'predictions aligned with empirical findings.' This is not an independent validation: the model is not calibrated to reduced-form moments, and the inverted-U is not tested against data. The theoretical results should be presented as consistency checks, or the paper should provide an out-of-sample empirical test of the inverted-U using the constructed LLM-compatibility measure.
minor comments (5)
- [Figure 10 and Section 6.3] The phrase 'Open Soruce' appears in the Figure 10 title and in the surrounding text; this should read 'Open Source.'
- [Appendix C.2, Tables C.4 and C.5] Both tables are captioned 'VFI Parameters- Opens Model,' but Table C.5 describes the closed-model solution; the caption of Table C.5 should be corrected.
- [Equation (3)] The profit function is written with inconsistent subscripts as π_{i,τ,t} in the text and π_{i,t,τ} in the equation, and the licensing price appears as both P_{τ,t} and P_t. Please standardize the notation.
- [Table 1] The table notes refer to a 'Transformer Sim.' column, but the actual column is labeled 'LLM Compatibility'; the note and column header should be aligned.
- [References and general text] The reference for Meta (2023b) contains the placeholder '[insert date you accessed the site]', and the text contains typos such as 'daset' in Appendix A.1 and 'Huggingace' in Section 5.1; these should be cleaned up.
Circularity Check
No significant circularity: the theory is motivated by empirical findings, but its predictions, including the inverted-U, are derived from an independently specified dynamic model and are not fitted from the data.
full rationale
The paper's empirical findings motivate the model, and Section 6 reports that the theoretical analysis generates predictions aligned with those findings, but this alignment is not a reduction. The model parameters in Table 6 are not estimated from the regressions; the quality-lead result (Proposition 1) and the inverted-U size result (Section 6.3, Figure 10) come from solving the Bellman equations (5)-(7) by value function iteration, not from reading the empirical estimates back into the model. The closest candidate for circularity is the inverted-U claim, which is demonstrated only at one parameter vector that the paper itself says should not be taken literally (Section 6.2); that is a calibration robustness concern, not a definitional or fitted-input circularity. There are no load-bearing self-citations, no imported uniqueness theorems, and no ansatz smuggled in through prior work of the author. The empirical claims are tested against external data (MMLU quality gaps in Table 2; LLaMA DiD on GitHub contributions in Tables 3-5), while the theoretical model is solved independently of those estimates. No circular step satisfying the quoted-reduction standard was found.
Assumptions & free parameters
free parameters (9)
- β (time discount factor) =
0.9
- γ (AI compatibility parameter) =
1.0
- α (shape of production function) =
0.45
- m (size of Firm A in the application sector) =
0.2
- ψ (efficiency of internal development) =
0.5
- ϕ (efficiency of open source ecosystem) =
0.5
- λ (LLM development improvement factor) =
5.0
- cD (LLM development cost factor) =
0.4
- Cosine similarity threshold for LLM compatibility =
0.7
assumptions (5)
- domain assumption Software producers are uniformly distributed on [0,1] and profit from LLM use is e^{-γx}(q k)^α - k - P (Equation 3).
- domain assumption Firm A owns all producers in [0,m] and can irreversibly open source its model; open source quality grows with external compute through ϕK_{-A}.
- domain assumption GitHub contributions by identified LLM researchers proxy for research activity, and non-AI repository contributors form a valid control group.
- domain assumption Patents citing Vaswani et al. (2017) and classified under CPC code G06F40 represent LLM-related technology, and the 0.7 cosine threshold defines compatibility.
- standard math The dynamic programming value functions are differentiable and first-order approximations around q* are valid for the proofs of Propositions 1 and 2.
Cite this review
Pith. "Pith review of Open Sourcing GPTs: Economics of Open Sourcing Advanced AI Models." pith.science (2026). https://pith.science/paper/EM5QC5ZB
@misc{pith2026250111581,
author = {Pith},
title = {Pith review of: Open Sourcing GPTs: Economics of Open Sourcing Advanced AI Models},
year = {2026},
howpublished = {\url{https://pith.science/paper/EM5QC5ZB}},
note = {Machine review of arXiv:2501.11581}
}
read the original abstract
This paper explores the economic underpinnings of open sourcing advanced large language models (LLMs) by for-profit companies. Empirical analysis reveals that: (1) LLMs are compatible with R&D portfolios of numerous technologically differentiated firms; (2) open-sourcing likelihood decreases with an LLM's performance edge over rivals, but increases for models from large tech companies; and (3) open-sourcing an advanced LLM led to an increase in research-related activities. Motivated by these findings, a theoretical framework is developed to examine factors influencing a profit-maximizing firm's open-sourcing decision. The analysis frames this decision as a trade-off between accelerating technology growth and securing immediate financial returns. A key prediction from the theoretical analysis is an inverted-U-shaped relationship between the owner's size, measured by its share of LLM-compatible applications, and its propensity to open source the LLM. This finding suggests that moderate market concentration may be beneficial to the open source ecosystems of multi-purpose software technologies.
Figures
Figures from the paper (1 more)
Reference graph
Works this paper leans on
-
[1]
Agrawal, A., J. S. Gans, and A. Goldfarb (2023a). Artificial intelligence adoption and system-wide change. Journal of Economics & Management Strategy\/
work page 2023
-
[2]
Agrawal, A. K., J. S. Gans, and A. Goldfarb (2023b). Similarities and differences in the adoption of general purpose technologies. Technical report, National Bureau of Economic Research
work page 2023
-
[3]
Ahmed, N., M. Wahed, and N. C. Thompson (2023). The growing influence of industry in ai research. Science\/ 379\/ (6635), 884--886
work page 2023
-
[4]
Allen, R. C. (1983). Collective invention. Journal of economic behavior & organization\/ 4\/ (1), 1--24
work page 1983
-
[5]
Athey, D
Arkhangelsky, D., S. Athey, D. A. Hirshberg, G. W. Imbens, and S. Wager (2021). Synthetic difference-in-differences. American Economic Review\/ 111\/ (12), 4088--4118
2021
-
[6]
Arora, A., S. Belenzon, A. Patacconi, and J. Suh (2020). The changing structure of american innovation: Some cautionary remarks for economic growth. Innovation Policy and the Economy\/ 20 , 39--93
work page 2020
-
[7]
Arora, A., S. Belenzon, and L. Sheer (2021). Knowledge spillovers and corporate investment in scientific research. American Economic Review\/ 111\/ (3), 871--898
work page 2021
-
[8]
Arts, S., B. Cassiman, and J. Hou (2021). Technology differentiation and firm performance. Harvard Business School Strategy Unit Working Paper\/ (22-040)
work page 2021
Show all 49 references
-
[9]
Lo, and A
Beltagy, I., K. Lo, and A. Cohan (2019). Scibert: A pretrained language model for scientific text
2019
-
[10]
Rock, and C
Brynjolfsson, E., D. Rock, and C. Syverson (2018). Artificial intelligence and the modern productivity paradox: A clash of expectations and statistics. In The economics of artificial intelligence: An agenda , pp.\ 23--57. University of Chicago Press
2018
-
[11]
Big tech is inflating fears about ai's risk to humanity: Google brain cofounder
Business-Insider (2023). Big tech is inflating fears about ai's risk to humanity: Google brain cofounder. https://www.businessinsider.com/andrew-ng-google-brain-big-tech-ai-risks-2023-10
2023
-
[12]
Casadesus-Masanell, R. and P. Ghemawat (2006). Dynamic mixed duopoly: A model motivated by linux vs. windows. Management Science\/ 52\/ (7), 1072--1084
2006
-
[13]
Zheng, Y
Chiang, W.-L., L. Zheng, Y. Sheng, A. N. Angelopoulos, T. Li, D. Li, H. Zhang, B. Zhu, M. Jordan, J. E. Gonzalez, and I. Stoica (2024). Chatbot arena: An open platform for evaluating llms by human preference
2024
-
[14]
Meta ceo mark zuckerberg touts to employees ‘incredible breakthroughs’ the company has seen in a.i
CNBC (2023a). Meta ceo mark zuckerberg touts to employees ‘incredible breakthroughs’ the company has seen in a.i. https://www.cnbc.com/2023/06/08/meta-ceo-mark-zuckerberg-talks-companys-ai-efforts-to-employees.html
2023
-
[15]
Meta's open source approach to ai puzzles wall street, techies love it
CNBC (2023b). Meta's open source approach to ai puzzles wall street, techies love it. https://www.cnbc.com/2023/10/16/metas-open-source-approach-to-ai-puzzles-wall-street-techies-love-it.html
2023
-
[16]
Cockburn, I. M., R. Henderson, and S. Stern (2018). The impact of artificial intelligence on innovation: An exploratory analysis. In The economics of artificial intelligence: An agenda , pp.\ 115--146. University of Chicago Press
2018
-
[17]
Deerwester, S., S. T. Dumais, G. W. Furnas, T. K. Landauer, and R. Harshman (1990). Indexing by latent semantic analysis. Journal of the American society for information science\/ 41\/ (6), 391--407
1990
-
[18]
Economides, N. and E. Katsamakas (2006). Two-sided competition of proprietary vs. open source technology platforms and the implications for the software industry. Management science\/ 52\/ (7), 1057--1071
2006
-
[19]
Manning, P
Eloundou, T., S. Manning, P. Mishkin, and D. Rock (2023). Gpts are gpts: An early look at the labor market impact potential of large language models
2023
-
[20]
Fosfuri, A., M. S. Giarratana, and A. Luzzi (2008). The penguin has entered the building: The commercialization of open source software products. Organization science\/ 19\/ (2), 292--305
2008
-
[21]
Gambardella, A. and E. A. von Hippel (2018). Open source hardware as a profit-maximizing strategy of downstream firms
2018
-
[22]
Kelly, and M
Gentzkow, M., B. Kelly, and M. Taddy (2019). Text as data. Journal of Economic Literature\/ 57\/ (3), 535--574
2019
-
[23]
Taska, and F
Goldfarb, A., B. Taska, and F. Teodoridis (2023). Could machine learning be a general purpose technology? a comparison of emerging technologies using data from online job postings. Research Policy\/ 52\/ (1), 104653
2023
-
[24]
Hain, D. S., R. Jurowetzki, T. Buchmann, and P. Wolf (2022). A text-embedding-based approach to measuring patent-to-patent technological similarity. Technological Forecasting and Social Change\/ 177 , 121559
2022
-
[25]
Henkel, J. (2004). Open source software from commercial firms--tools, complements, and collective invention. Zeitschrift f \"u r Betriebswirtschaft\/ 4 , 1--23
2004
-
[26]
Hovy, D. (2022). Text analysis in python for social scientists: Prediction and classification . Cambridge University Press
2022
-
[27]
Howard, J. and S. Ruder (2018). Universal language model fine-tuning for text classification. arXiv preprint arXiv:1801.06146\/
2018 arXiv
-
[28]
Jacobides, M. G., S. Brusoni, and F. Candelon (2021). The evolutionary dynamics of the artificial intelligence ecosystem. Strategy Science\/ 6\/ (4), 412--435
2021
-
[29]
Jaffe, A. B. (1986). Technological opportunity and spillovers of r&d: Evidence from firms' patents, profits, and market value. The American Economic Review\/ 76\/ (5), 984--1001
1986
-
[30]
McCandlish, T
Kaplan, J., S. McCandlish, T. Henighan, T. B. Brown, B. Chess, R. Child, S. Gray, A. Radford, J. Wu, and D. Amodei (2020). Scaling laws for neural language models. arXiv preprint arXiv:2001.08361\/
2020 arXiv
-
[31]
Papanikolaou, A
Kelly, B., D. Papanikolaou, A. Seru, and M. Taddy (2021). Measuring technological innovation over the long run. American Economic Review: Insights\/ 3\/ (3), 303--320
2021
-
[32]
Lerner, J., P. A. Pathak, and J. Tirole (2006). The dynamics of open-source contributors. American Economic Review\/ 96\/ (2), 114--118
2006
-
[33]
Lerner, J. and J. Tirole (2002). Some simple economics of open source. The journal of industrial economics\/ 50\/ (2), 197--234
2002
-
[34]
Lewis, M., Y. Liu, N. Goyal, M. Ghazvininejad, A. Mohamed, O. Levy, V. Stoyanov, and L. Zettlemoyer (2019). Bart: Denoising sequence-to-sequence pre-training for natural language generation, translation, and comprehension. arXiv preprint arXiv:1910.13461\/
2019 arXiv
-
[35]
Introducing llama: A foundational, 65-billion-parameter large language model
Meta (2023a, February). Introducing llama: A foundational, 65-billion-parameter large language model. https://ai.meta.com/blog/large-language-model-llama-meta-ai/
-
[36]
The llama ecosystem: Past, present, and future
Meta (2023b). The llama ecosystem: Past, present, and future. https://ai.meta.com/blog/llama-2-updates-connect-2023/. Accessed: [insert date you accessed the site]
2023
-
[37]
Nagle, F. (2018). Learning by contributing: Gaining competitive advantage through contribution to crowdsourced public goods. Organization Science\/ 29\/ (4), 569--587
2018
-
[38]
Nagle, F. (2019). Open source software and firm productivity. Management Science\/ 65\/ (3), 1191--1215
2019
-
[39]
Nuvolari, A. (2004). Collective invention during the british industrial revolution: the case of the cornish pumping engine. Cambridge Journal of Economics\/ 28\/ (3), 347--363
2004
-
[40]
Let us show you how gpt works — using jane austen
NYT (2023, April). Let us show you how gpt works — using jane austen. https://www.nytimes.com/2023/04/27/upshot/gpt-from-scratch.html
2023
-
[41]
Osterloh, M. and S. Rota (2007). Open source software development—just another case of collective invention? Research Policy\/ 36\/ (2), 157--171
2007
-
[42]
(2023, November)
Post, T. (2023, November). Big tech wants ai regulation. the rest of silicon valley is skeptical. https://www.washingtonpost.com/technology/2023/11/09/ai-regulation-silicon-valley-skeptics/
2023
-
[43]
Radford, A., J. Wu, R. Child, D. Luan, D. Amodei, I. Sutskever, et al. (2019). Language models are unsupervised multitask learners. OpenAI blog\/ 1\/ (8), 9
2019
-
[44]
Shazeer, A
Raffel, C., N. Shazeer, A. Roberts, K. Lee, S. Narang, M. Matena, Y. Zhou, W. Li, and P. J. Liu (2020). Exploring the limits of transfer learning with a unified text-to-text transformer. Journal of machine learning research\/ 21\/ (140), 1--67
2020
-
[45]
Rock, D. (2019). Engineering value: The returns to technological talent and investments in artificial intelligence. Available at SSRN\/ 3427412
2019
-
[46]
Spencer, J. W. (2003). Firms' knowledge-sharing strategies in the global innovation system: empirical evidence from the flat panel display industry. Strategic management journal\/ 24\/ (3), 217--233
2003
-
[47]
Shazeer, N
Vaswani, A., N. Shazeer, N. Parmar, J. Uszkoreit, L. Jones, A. N. Gomez, L. Kaiser, and I. Polosukhin (2017). Attention is all you need. In Advances in neural information processing systems , Volume 30
2017
-
[48]
private-collective
von Hippel, E. and G. von Krogh (2003). Open source software and the “private-collective” innovation model: Issues for organization science. Organization science\/ 14\/ (2), 209--223
2003
-
[49]
Should ai be open-source? behind the tweetstorm over its dangers
WSJ (2024). Should ai be open-source? behind the tweetstorm over its dangers. https://www.wsj.com/articles/should-ai-be-open-source-behind-the-tweetstorm-over-its-dangers-65aa5c97
2024
Reviewed August 10, 2026 · model on record in the stance chip above.
Discussion (0). Continue with ORCID to comment.