REVIEW 2 major objections 2 minor 54 references
Large language models can construct partial orderings of software licenses based on permissiveness through pairwise comparisons.
Reviewed by Pith at T0; open to challenge. T0 means a machine referee read the full paper against a public rubric. the ladder, T0–T4 →
T0 review · grok-4.3
2026-07-01 05:12 UTC pith:MHT3XX6U
load-bearing objection LLM pairwise license comparisons are a new angle but the abstract shows no validation against experts or known legal cases. the 2 major comments →
Partially ordering software licenses
The pith
A machine-rendered reading of the paper's core claim, the machinery that carries it, and where it could break.
Core claim
Using large language models, licenses can be compared pairwise to build a partial ordering based on permissiveness, and taxonomies can be used to understand license selection as combinations of shared provisions. This recovers interpretable attributes that correspond to stricter licenses.
What carries the argument
Pairwise comparisons by large language models to establish relative permissiveness, together with mappings onto existing license taxonomies.
Load-bearing premise
Large language models generate comparisons of license terms that are consistent and align with expert legal judgment instead of reflecting prompt sensitivity or training data patterns.
What would settle it
If copyright lawyers systematically disagree with the LLM pairwise rankings on a representative set of licenses regarding which imposes more restrictions on reuse and modification.
If this is right
- License relationships become traceable at scale rather than remaining unstructured.
- Platforms can identify when licenses are incomparable rather than assuming total orders.
- Stricter license attributes become detectable through the recovered features.
- License selection can be analyzed as choices among shared provisions.
Where Pith is reading between the lines
- This approach could be tested on emerging licenses to see if the ordering holds.
- Legal experts might use the attributes to flag potential conflicts in license combinations.
- Extending the method to other legal documents could reveal similar partial orders in contract terms.
Editorial analysis
A structured set of objections, weighed in public.
Referee Report
Summary. The manuscript introduces LLM-based methods for comparing software licenses at scale. The first constructs a partial order on permissiveness via pairwise judgments; the second projects licenses onto existing taxonomies to identify combinations of shared provisions. The analysis claims to recover interpretable attributes associated with stricter licenses and discusses implications for the open-source ecosystem on platforms such as GitHub and Hugging Face.
Significance. If the LLM judgments can be shown to align with expert legal reasoning, the approach would supply a scalable, reproducible technique for license analysis that is currently absent from the literature. The work is novel in applying LLMs to this domain and could support practical tooling for license compatibility checking, but its contribution is limited by the absence of validation evidence.
major comments (2)
- [Methods] Methods section: The pairwise prompting procedure for constructing the partial order is described without any reported validation against practicing licensing counsel, inter-annotator agreement metrics, or calibration against established legal distinctions (e.g., strong copyleft in GPL-family licenses versus permissive MIT/BSD terms). This validation is load-bearing for the central claim that the recovered ordering reflects legal permissiveness rather than training-data artifacts or prompt sensitivity.
- [Results] Results section: No details are supplied on the prompting strategy (including temperature, few-shot examples, or consistency checks), the procedure for aggregating pairwise judgments into a partial order, or robustness under prompt paraphrases. Without these, the reported interpretable attributes cannot be assessed for stability or legal fidelity.
minor comments (2)
- [Abstract] Abstract: The number of licenses examined and the specific LLMs employed are not stated, making it difficult to gauge the scale of the study.
- [Methods] Notation: The manuscript should define how incomparability is operationalized in the partial order (e.g., when two licenses receive conflicting pairwise judgments).
Simulated Author's Rebuttal
We thank the referee for the constructive feedback. We address the two major comments point by point below, with planned revisions where feasible.
read point-by-point responses
-
Referee: [Methods] Methods section: The pairwise prompting procedure for constructing the partial order is described without any reported validation against practicing licensing counsel, inter-annotator agreement metrics, or calibration against established legal distinctions (e.g., strong copyleft in GPL-family licenses versus permissive MIT/BSD terms). This validation is load-bearing for the central claim that the recovered ordering reflects legal permissiveness rather than training-data artifacts or prompt sensitivity.
Authors: We agree that external validation against legal experts would strengthen claims of legal fidelity. The manuscript positions LLMs as a scalable proxy rather than a replacement for counsel; however, we will revise to add inter-annotator agreement via repeated runs with varied seeds and a calibration subsection comparing known distinctions (GPL-family vs. MIT/BSD). Full engagement with practicing licensing counsel lies outside the scope of this work and would require a separate study. revision: partial
-
Referee: [Results] Results section: No details are supplied on the prompting strategy (including temperature, few-shot examples, or consistency checks), the procedure for aggregating pairwise judgments into a partial order, or robustness under prompt paraphrases. Without these, the reported interpretable attributes cannot be assessed for stability or legal fidelity.
Authors: We accept this criticism and will expand the Methods section in revision. The updated text will report temperature=0, the complete prompt templates (with any few-shot examples), the aggregation procedure (directed graph followed by transitive reduction to obtain the partial order), and new robustness results under prompt paraphrases demonstrating stability of the recovered attributes. revision: yes
- Formal validation against practicing licensing counsel
Circularity Check
No circularity: LLM pairwise comparisons and taxonomy projection are independent of target ordering
full rationale
The paper's method applies LLMs to generate pairwise permissiveness judgments and projects onto existing taxonomies to recover attributes. No equations, fitted parameters, or self-definitional reductions appear. The partial order is constructed from model outputs rather than presupposing the result; no self-citation chain justifies a uniqueness theorem or ansatz. The derivation remains self-contained against external benchmarks (legal taxonomies) and does not rename known patterns or smuggle inputs as predictions.
Axiom & Free-Parameter Ledger
read the original abstract
Licenses are legal instruments that inventors may use to protect the technologies they build and regulate how they are used -- however, the nature of their authorship and selection means that how they are interpreted, chosen, and enforced is largely unstructured. In practice, this makes it difficult to compare licenses at scale -- when is one license considered more permissive than the other, and when are their terms incomparable to each other? Currently, there is a growing list of licenses that are introduced and used, but there is no systematic way to study their relationships. This matters for platforms such as Hugging Face, GitHub, and the Python Package Index, where developers publish or build upon technologies that each have their own licenses. Using large language models (LLMs), we introduce methods for comparing licenses at scale: first, in a pairwise fashion to construct a partial ordering based on permissiveness, and second, by drawing on existing taxonomies of software licenses. The former allows us to trace restrictiveness, and the latter allows us to understand license selection as a combination of shared provisions. Our analysis recovers certain interpretable attributes that correspond to stricter licenses, with legal implications for the open-source ecosystem.
Figures
Reference graph
Works this paper leans on
-
[1]
Property as the law of things , author=. Harvard Law Review , volume=
-
[2]
arXiv preprint arXiv:2504.20185 , year=
The Foundation Model Transparency Index , author=. arXiv preprint arXiv:2504.20185 , year=
-
[3]
arXiv preprint arXiv:2401.xxxxx , year=
Data provenance and the machine learning supply chain , author=. arXiv preprint arXiv:2401.xxxxx , year=
-
[4]
Ensuring Free, Immediate, and Equitable Access to Federally Funded Research , author=. 2022 , month=
work page 2022
- [5]
- [6]
-
[7]
arXiv preprint arXiv:2403.xxxxx , year=
Model ChangeLists: Documenting API model performance changes , author=. arXiv preprint arXiv:2403.xxxxx , year=
-
[8]
News Corp and OpenAI partnership agreement , author=. 2024 , howpublished=
work page 2024
-
[9]
The New York Times licensing and litigation regarding AI training , author=. 2024 , note=
work page 2024
-
[10]
Innovation Policy and the Economy , volume=
Navigating the patent thicket: Cross licenses, patent pools, and standard setting , author=. Innovation Policy and the Economy , volume=. 2001 , publisher=
work page 2001
-
[11]
Common Crawl , author=
-
[12]
Schuhmann, Christoph and others , journal=
-
[13]
Gao, Leo and Biderman, Stella and Black, Sid and others , journal=. The
-
[14]
Books3 dataset , author=
-
[15]
Stability AI complaint , author=
Getty Images v. Stability AI complaint , author=. 2023 , howpublished=
work page 2023
-
[16]
Authors Guild class action lawsuits against AI companies , author=. 2023 , note=
work page 2023
-
[17]
Llama 2 Community License Agreement , author =. 2023 , howpublished =
work page 2023
-
[18]
Nordlander, Erik and Loreto, Daniel and Oliner, Adam and Woo, Ram , title =. 2004 , month =
work page 2004
-
[19]
The Rise of Open Source Licensing:
V. The Rise of Open Source Licensing:. 2005 , address =
work page 2005
-
[20]
A dataset showing a century of evolution in the complexity of the United States legal Code
Jeong, Dawoon and Holehouse, James and Yoon, Jisung and Kempes, Christopher P and West, Geoffrey B and Youn, Hyejin. A dataset showing a century of evolution in the complexity of the United States legal Code. Sci. Data
-
[21]
Anatomy of a Machine Learning Ecosystem: 2 Million Models on Hugging Face , author=. 2025 , eprint=
work page 2025
-
[22]
Lewis & Clark Law Review , year =
Madiha Zahrah Choksi and James Grimmelmann , title =. Lewis & Clark Law Review , year =
- [23]
-
[24]
Cen, Isabella Struckman, Andrew Ilyas, Luis Videgaray, and Aleksander Mądry
Hopkins, Aspen and Cen, Sarah H. and Struckman, Isabella and Ilyas, Andrew and Videgaray, Luis and Mądry, Aleksander , year=2025, pages=. AI Supply Chains: An Emerging Ecosystem of AI Actors, Products, and Services , volume=. Proceedings of the AAAI/ACM Conference on AI, Ethics, and Society , publisher=. doi:10.1609/aies.v8i2.36628 , number=
-
[25]
The Data Provenance Initiative: A Large Scale Audit of Dataset Licensing & Attribution in AI , author=. 2023 , eprint=
work page 2023
-
[26]
and Chen, Xi and Zhang, Jingxuan , year =
Bayar, Ozgur and Chemmanur, Thomas J. and Chen, Xi and Zhang, Jingxuan , year =. The Economics of Patent Licensing: Theory and Evidence on the Determinants and Consequences of Patent Licensing Transactions , note =
-
[27]
Permissive-Washing in the Open AI Supply Chain: A Large-Scale Audit of License Integrity , author=. 2026 , eprint=
work page 2026
-
[28]
Heller, Michael A. and Eisenberg, Rebecca S. , title =. Science , year =. doi:10.1126/science.280.5364.698 , issn =
-
[29]
American Economic Review , year =
Ostrom, Elinor , title =. American Economic Review , year =
-
[30]
Examining Software Developers' Needs for Privacy Enforcing Techniques: A survey , author=. ArXiv , year=
-
[31]
Myers, Christopher R. , year=. Software systems as complex networks: Structure, function, and evolvability of software collaboration graphs , volume=. Physical Review E , publisher=. doi:10.1103/physreve.68.046116 , number=
- [32]
-
[33]
Cheap and Fast -- But is it Good? Evaluating Non-Expert Annotations for Natural Language Tasks
Snow, Rion and O ' Connor, Brendan and Jurafsky, Daniel and Ng, Andrew. Cheap and Fast -- But is it Good? Evaluating Non-Expert Annotations for Natural Language Tasks. Proceedings of the 2008 Conference on Empirical Methods in Natural Language Processing. 2008
work page 2008
-
[34]
A Mathematical Theory of Communication , url =
Shannon, Claude Elwood , biburl =. A Mathematical Theory of Communication , url =. The Bell System Technical Journal , keywords =
-
[35]
System Package Data Exchange (. 2024 , month = dec, number =
work page 2024
- [36]
- [37]
-
[38]
Donald Arseneau , howpublished =
- [39]
-
[40]
Artistic License 1.0 (Perl) , howpublished =
-
[41]
The Model Openness Framework: Promoting Completeness and Openness for Reproducibility, Transparency, and Usability in Artificial Intelligence , author=. 2024 , eprint=
work page 2024
- [42]
- [43]
-
[44]
IEEE Transactions on Visualization and Computer Graphics (InfoVis) , doi =
UpSet: Visualization of Intersecting Sets , author =. IEEE Transactions on Visualization and Computer Graphics (InfoVis) , doi =
-
[45]
Proceedings of the 21st International Conference on Mining Software Repositories , pages =
Wu, Jiaqi and Bao, Lingfeng and Yang, Xiaohu and Xia, Xin and Hu, Xing , title =. Proceedings of the 21st International Conference on Mining Software Repositories , pages =. 2024 , isbn =. doi:10.1145/3643991.3644900 , abstract =
-
[46]
and Charalambous, Georgia , journal=
Kapitsaki, Georgia M. and Charalambous, Georgia , journal=. Modeling and Recommending Open Source Licenses with findOSSLicense , year=
-
[47]
Ralph Allan Bradley and Milton E. Terry , journal =. Rank Analysis of Incomplete Block Designs: I. The Method of Paired Comparisons , urldate =
-
[48]
Kenneth J. Arrow , publisher =. Social Choice and Individual Values , urldate =
-
[49]
Forty-second International Conference on Machine Learning Position Paper Track , year=
Position: Current Model Licensing Practices are Dragging Us into a Quagmire of Legal Noncompliance , author=. Forty-second International Conference on Machine Learning Position Paper Track , year=
-
[50]
Understanding open source and free software licensing - guide to navigation licensing issues in existing and new software , author=. 2004 , url=
work page 2004
-
[51]
Hersb R. Reddy , journal =. Jacobsen v. Katzer: The Federal Circuit Weighs in on the Enforceability of Free and Open Source Software Licenses , urldate =
- [52]
- [53]
- [54]
discussion (0)
Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.