Pith. sign in

REVIEW 2 major objections 297 references

Linguistic Productivity in Large Language Models: Models Coerce, but do not Preempt

T0 review · 2 major / 0 minor · reviewed 2026-06-28 · grok-4.3

Pith's one-line read Large language models capture entrenchment through coercion with nonce words but show no preemption from absent patterns.

desk verdict LLMs show entrenchment effects on nonce coercion but no preemption on overgeneralization; the dissociation is interesting but the methods need more detail to hold up. read the letter →

arxiv 2606.02953 v1 pith:4S6K6R2X submitted 2026-06-01 cs.CL

classification cs.CL
keywords linguisticproductivityentrenchmentpreemptionconstructionalcoercionnoncewordslargelanguagemodelsusage-basedtheoriesovergeneralization
verification ladder T0 review T1 audit T2 compute T3 formal

The pith

A machine-rendered reading of the paper's core claim, the machinery that carries it, and where it could break.

The reading

Usage-based theories hold that language productivity is supported by frequent exposure to structures and limited by their consistent absence where expected. The paper tests whether these same frequency signals shape how LLMs generate and interpret language. Experiments show larger models can extend constructions to made-up words when context forces an atypical meaning, reproducing the entrenchment side of the theory. The same models nevertheless produce overgeneralizations of semantically acceptable patterns that never occurred in their training data, indicating they do not register or apply negative evidence in the way preemption requires.

What carries the argument

The contrast between entrenchment, driven by high-frequency usage of a construction, and preemption, driven by consistent non-occurrence in contexts where the construction might otherwise appear, tested through nonce-word substitution in coercion frames.

What would settle it

A controlled test in which models trained or prompted with explicit negative evidence for a semantically acceptable but unattested construction subsequently stop producing that construction at rates significantly above baseline.

Watch

Extended reading notes

Core claim

Across model sizes and architectures, LLMs reproduce constructional productivity via entrenchment when a broader frame coerces an atypical reading of a nonce word, yet they continue to overgeneralize patterns that are semantically acceptable but unattested, showing that statistical preemption does not constrain their output.

Load-bearing premise

The specific nonce-word tasks and construction frames used here validly isolate the same entrenchment and preemption mechanisms that usage-based theories attribute to human speakers.

Editorial extensions

If this is right

  • Larger models increasingly exhibit entrenchment effects that allow coerced interpretations with novel lexical items.
  • Models of any size fail to block overgeneralization of unattested but semantically coherent patterns.
  • Statistical absence alone does not function as a learning signal for LLMs in the manner predicted by preemption accounts.
  • The dissociation between coercion success and preemption failure holds across different model architectures.

Reading between the lines

Editorial extensions of the paper, not claims the author makes directly.

  • This pattern suggests LLMs may need explicit mechanisms for registering negative evidence if they are to match human-like avoidance of certain generalizations.
  • The result points to a possible test: whether targeted exposure to unattested constructions paired with corrective signals reduces overgeneralization in subsequent generations.
  • It raises the question of whether other statistical or architectural features, beyond raw frequency counts, could supply the missing preemption effect.
Share X Bluesky LinkedIn Reddit HN

Editorial analysis

A structured set of objections, weighed in public.

Desk editor's note, referee report, simulated authors' rebuttal, and a circularity audit.

Referee Report

2 major / 0 minor

Summary. The paper claims that LLMs exhibit entrenchment-driven constructional productivity (via coercion with nonce words) that scales with model size, but lack preemption effects from negative evidence, failing to block overgeneralization on semantically felicitous but unattested patterns; this dissociation is presented as holding across architectures and as evidence that statistical preemption does not constrain LLM productivity in the manner predicted by usage-based theories.

Significance. If the dissociation is robustly demonstrated, the result would bear on whether LLMs implement the two distinct frequency signals posited in usage-based grammar, with potential implications for cognitive modeling of productivity. The nonce-word design is a standard tool for testing generalization and is a positive feature when properly controlled.

major comments (2)
  1. [Abstract] Abstract: results are asserted across architectures with no accompanying details on test constructions, statistical controls, sample sizes, or the operationalization of overgeneralization; without these elements the central empirical claim cannot be evaluated.
  2. [Experimental tasks] The reported failure to avoid overgeneralization on unattested but felicitous patterns is taken to demonstrate absence of preemption; however, this inference requires showing that the nonce-word frames and prompting regime provide sufficient negative evidence and isolate preemption from architecture-specific limits on representing absence, which is not addressed.

Simulated Author's Rebuttal

2 responses · 0 unresolved

We thank the referee for their detailed and constructive comments. We address each major comment point by point below, indicating planned revisions where appropriate.

read point-by-point responses
  1. Referee: [Abstract] Abstract: results are asserted across architectures with no accompanying details on test constructions, statistical controls, sample sizes, or the operationalization of overgeneralization; without these elements the central empirical claim cannot be evaluated.

    Authors: We agree that the abstract is highly condensed. The full manuscript specifies the constructions (coercion frames with nonce words), controls (model size and architecture comparisons), sample sizes (multiple LLMs and prompt variants), and operationalization (preference rates for attested vs. unattested patterns). We will revise the abstract to include a concise reference to these elements. revision: partial

  2. Referee: [Experimental tasks] The reported failure to avoid overgeneralization on unattested but felicitous patterns is taken to demonstrate absence of preemption; however, this inference requires showing that the nonce-word frames and prompting regime provide sufficient negative evidence and isolate preemption from architecture-specific limits on representing absence, which is not addressed.

    Authors: The nonce-word coercion design follows established usage-based methods to supply contexts where preemption from negative evidence would be expected if utilized. Testing across architectures and sizes helps separate general statistical effects from model-specific constraints. We will add explicit discussion in the Methods and Discussion sections on the prompting regime's provision of negative evidence and note limitations in fully isolating preemption from representational factors. revision: yes

Circularity Check

0 steps flagged · score 0.0 of 10

Empirical evaluation with no derivation chain or self-referential reductions

full rationale

The paper reports experimental results from prompting LLMs with nonce words in specific constructional frames to test entrenchment (via coercion) versus preemption effects. No equations, parameters, or derivations appear in the abstract or described content; outcomes are direct model generations compared to linguistic expectations from usage-based theories. No self-citations function as load-bearing uniqueness theorems, no fitted inputs are relabeled as predictions, and no ansatzes or renamings reduce claims to inputs by construction. The central dissociation between coercion recognition and failure to use negative evidence follows from observed outputs against external benchmarks, rendering the work self-contained.

Assumptions & free parameters 0 free parameters · 1 assumptions · 0 invented entities

The paper tests an existing linguistic theory on LLMs rather than deriving new constants or introducing new entities; background assumptions are standard in usage-based linguistics.

assumptions (1)
  • domain assumption Usage-based theories of grammars posit that creative productivity of the structures of language is both bolstered and constrained by two distinct frequency signals: entrenchment and preemption.
    This is the core theoretical premise the experiments are designed to test in LLMs.

how reviews work

0 comments
Cite this review

Pith. "Pith review of Linguistic Productivity in Large Language Models: Models Coerce, but do not Preempt." pith.science (2026). https://pith.science/paper/4S6K6R2X

@misc{pith2026260602953,
  author       = {Pith},
  title        = {Pith review of: Linguistic Productivity in Large Language Models: Models Coerce, but do not Preempt},
  year         = {2026},
  howpublished = {\url{https://pith.science/paper/4S6K6R2X}},
  note         = {Machine review of arXiv:2606.02953}
}
read the original abstract

Usage-based theories of grammars posit that creative productivity of the structures of language is both bolstered and constrained by two distinct frequency signals: entrenchment, stemming from high frequency usage, and preemption, stemming from having never observed a particular linguistic structure in a context where one might expect that structure to appear. Large Language Models are also usage-based, in the sense that the structures of language are learned through exposure to vast amounts of text. Here, we test whether or not the opposing statistical forces of entrenchment and preemption also encourage and constrain linguistic productivity in LLMs. We demonstrate across model architectures that larger models recognize and can reproduce with nonce words constructional productivity (entrenchment) in cases of coercion, wherein the broader constructional context coerces an atypical interpretation of a lexical item. However, we also show that even the largest models do not extend negative evidence to novel language, and statistical preemption does not enable models to avoid overgeneralization of patterns that are semantically felicitous, but never observed in data.

Discussion (0). Continue with ORCID to comment.

Reference graph

Works this paper leans on

297 extracted references · 77 canonical work pages

  1. [1]

    S tanza: A Python Natural Language Processing Toolkit for Many Human Languages

    Qi, Peng and Zhang, Yuhao and Zhang, Yuhui and Bolton, Jason and Manning, Christopher D. , editor=. Stanza:. Proceedings of the 58th Annual Meeting of the Association for Computational Linguistics: System Demonstrations , publisher=. 2020 , month=jul, pages=. doi:10.18653/v1/2020.acl-demos.14 , abstractNote=

  2. [2]

    Devlin, Jacob and Chang, Ming-Wei and Lee, Kenton and Toutanova, Kristina , editor=. B. Proceedings of the 2019 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, Volume 1 (Long and Short Papers) , publisher=. 2019 , month=jun, pages=. doi:10.18653/v1/N19-1423 , abstractNote=

  3. [3]

    Mahowald, Kyle , year=. A. Proceedings of the 17th Conference of the European Chapter of the Association for Computational Linguistics , publisher=

  4. [4]

    “Construction after

    Jackendoff, Ray , year=. “Construction after. Language , publisher=

  5. [5]

    The better your

    Weissweiler, Leonie and Hofmann, Valentin and Köksal, Abdullatif and Schütze, Hinrich , editor=. The better your. Proceedings of the 2022 Conference on Empirical Methods in Natural Language Processing , publisher=. 2022 , month=dec, pages=. doi:10.18653/v1/2022.emnlp-main.746 , abstractNote=

  6. [6]

    Findings of the

    Warstadt, Alex and Mueller, Aaron and Choshen, Leshem and Wilcox, Ethan and Zhuang, Chengxu and Ciro, Juan and Mosquera, Rafael and Paranjabe, Bhargavi and Williams, Adina and Linzen, Tal and Cotterell, Ryan. Findings of the B aby LM Challenge: Sample-Efficient Pretraining on Developmentally Plausible Corpora. Proceedings of the BabyLM Challenge at the 27...

  7. [7]

    Constructions are Revealed in Word Distributions

    Rozner, Joshua and Weissweiler, Leonie and Mahowald, Kyle and Shain, Cory. Constructions are Revealed in Word Distributions. Proceedings of the 2025 Conference on Empirical Methods in Natural Language Processing. 2025

  8. [8]

    B aby LM ' s First Constructions: Causal interventions provide a signal of learning

    Rozner, Joshua and Weissweiler, Leonie and Shain, Cory. B aby LM ' s First Constructions: Causal interventions provide a signal of learning. Proceedings of the 2025 Conference on Empirical Methods in Natural Language Processing. 2025

Show all 297 references
  1. [9]

    arXiv preprint arXiv:1907.11692 , year=

    Roberta: A robustly optimized bert pretraining approach , author=. arXiv preprint arXiv:1907.11692 , year=

  2. [10]

    Evaluating C x G Generalisation in LLM s via Construction-Based NLI Fine Tuning

    Mackintosh, Tom and Tayyar Madabushi, Harish and Bonial, Claire. Evaluating C x G Generalisation in LLM s via Construction-Based NLI Fine Tuning. Proceedings of the Second International Workshop on Construction Grammars and NLP. 2025

  3. [11]

    UC xn: Typologically Informed Annotation of Constructions Atop U niversal D ependencies

    Weissweiler, Leonie and B. UC xn: Typologically Informed Annotation of Constructions Atop U niversal D ependencies. Proceedings of the 2024 Joint International Conference on Computational Linguistics, Language Resources and Evaluation (LREC-COLING 2024). 2024

  4. [12]

    Construction Identification and Disambiguation Using BERT : A Case Study of NPN

    Scivetti, Wesley and Schneider, Nathan. Construction Identification and Disambiguation Using BERT : A Case Study of NPN. Proceedings of the 29th Conference on Computational Natural Language Learning. 2025. doi:10.18653/v1/2025.conll-1.24

  5. [13]

    Do Construction Distributions Shape Formal Language Learning In G erman B aby LM s?

    Bunzeck, Bastian and Duran, Daniel and Zarrie , Sina. Do Construction Distributions Shape Formal Language Learning In G erman B aby LM s?. Proceedings of the 29th Conference on Computational Natural Language Learning. 2025. doi:10.18653/v1/2025.conll-1.12

  6. [14]

    and Varoquaux, G

    Pedregosa, F. and Varoquaux, G. and Gramfort, A. and Michel, V. and Thirion, B. and Grisel, O. and Blondel, M. and Prettenhofer, P. and Weiss, R. and Dubourg, V. and Vanderplas, J. and Passos, A. and Cournapeau, D. and Brucher, M. and Perrot, M. and Duchesnay, E. , journal=. S...

  7. [15]

    Unpacking Let Alone: Human-Scale Models Generalize to a Rare Construction in Form but not Meaning

    Scivetti, Wesley and Aoyama, Tatsuya and Wilcox, Ethan and Schneider, Nathan. Unpacking Let Alone: Human-Scale Models Generalize to a Rare Construction in Form but not Meaning. Proceedings of the 2025 Conference on Empirical Methods in Natural Language Processing. 2025

  8. [16]

    Language Models Learn Rare Phenomena from Less Rare Phenomena:

    Misra, Kanishka and Mahowald, Kyle , editor =. Language Models Learn Rare Phenomena from Less Rare Phenomena:. Proceedings of the 2024. 2024 , pages =. doi:10.18653/v1/2024.emnlp-main.53 , abstract =

  9. [17]

    2024 , eprint=

    Time Travel in LLMs: Tracing Data Contamination in Large Language Models , author=. 2024 , eprint=

  10. [18]

    Modeling Semantic Containment and Exclusion in Natural Language Inference

    MacCartney, Bill and Manning, Christopher D. Modeling Semantic Containment and Exclusion in Natural Language Inference. Proceedings of the 22nd International Conference on Computational Linguistics (Coling 2008). 2008

  11. [19]

    Tayyar Madabushi, Harish and Romain, Laurence and Divjak, Dagmar and Milin, Petar , editor=. Cx. Proceedings of the 28th International Conference on Computational Linguistics , publisher=. 2020 , month=dec, pages=. doi:10.18653/v1/2020.coling-main.355 , abstractNote=

  12. [20]

    Computational Linguistics , author=

    Probing. Computational Linguistics , author=. 2022 , month=apr, pages=. doi:10.1162/coli_a_00422 , abstractNote=

  13. [21]

    2020 , month=jul, pages=

    Transactions of the Association for Computational Linguistics , author=. 2020 , month=jul, pages=. doi:10.1162/tacl_a_00321 , abstractNote=

  14. [22]

    Yaghoobzadeh, Yadollah and Kann, Katharina and Hazen, T. J. and Agirre, Eneko and Schütze, Hinrich , editor=. Probing for. Proceedings of the 57th Annual Meeting of the Association for Computational Linguistics , publisher=. 2019 , month=jul, pages=. doi:10.18653/v1/P19-1574 ,...

  15. [23]

    Vulić, Ivan and Ponti, Edoardo Maria and Litschko, Robert and Glavaš, Goran and Korhonen, Anna , editor=. Probing. Proceedings of the 2020 Conference on Empirical Methods in Natural Language Processing (EMNLP) , publisher=. 2020 , month=nov, pages=. doi:10.18653/v1/2020.emnlp-...

  16. [25]

    Zhou, Yichu and Srikumar, Vivek , editor=. Direct. Proceedings of the 2021 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies , publisher=. 2021 , month=jun, pages=. doi:10.18653/v1/2021.naacl-main.401 , abstractNote=

  17. [26]

    Aoyama, Tatsuya and Schneider, Nathan , editor=. Probe-. Proceedings of the 2022 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies: Student Research Workshop , publisher=. 2022 , month=jul, pages=. doi:10.186...

  18. [27]

    Designing and

    Hewitt, John and Liang, Percy , editor=. Designing and. Proceedings of the 2019 Conference on Empirical Methods in Natural Language Processing and the 9th International Joint Conference on Natural Language Processing (EMNLP-IJCNLP) , publisher=. 2019 , month=nov, pages=. doi:1...

  19. [28]

    Jawahar, Ganesh and Sagot, Benoît and Seddah, Djamé , editor=. What. Proceedings of the 57th Annual Meeting of the Association for Computational Linguistics , publisher=. 2019 , month=jul, pages=. doi:10.18653/v1/P19-1356 , abstractNote=

  20. [29]

    Counterfactual

    Ravfogel, Shauli and Prasad, Grusha and Linzen, Tal and Goldberg, Yoav , editor=. Counterfactual. Proceedings of the 25th Conference on Computational Natural Language Learning , publisher=. 2021 , month=nov, pages=. doi:10.18653/v1/2021.conll-1.15 , abstractNote=

  21. [30]

    , editor=

    Clark, Kevin and Khandelwal, Urvashi and Levy, Omer and Manning, Christopher D. , editor=. What. Proceedings of the 2019 ACL Workshop BlackboxNLP: Analyzing and Interpreting Neural Networks for NLP , publisher=. 2019 , month=aug, pages=. doi:10.18653/v1/W19-4828 , abstractNote=

  22. [31]

    Karidi, Taelin and Zhou, Yichu and Schneider, Nathan and Abend, Omri and Srikumar, Vivek , editor=. Putting. Proceedings of the 2021 Conference on Empirical Methods in Natural Language Processing , publisher=. 2021 , month=nov, pages=. doi:10.18653/v1/2021.emnlp-main.806 , abs...

  23. [32]

    and Gardner, Matt and Belinkov, Yonatan and Peters, Matthew E

    Liu, Nelson F. and Gardner, Matt and Belinkov, Yonatan and Peters, Matthew E. and Smith, Noah A. , editor=. Linguistic. Proceedings of the 2019 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, Volume 1 (Lon...

  24. [33]

    Probing for

    Conia, Simone and Navigli, Roberto , editor=. Probing for. Proceedings of the 60th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , publisher=. 2022 , month=may, pages=. doi:10.18653/v1/2022.acl-long.316 , abstractNote=

  25. [34]

    Probing for semantic evidence of composition by means of simple classification tasks , url=

    Ettinger, Allyson and Elgohary, Ahmed and Resnik, Philip , year=. Probing for semantic evidence of composition by means of simple classification tasks , url=. doi:10.18653/v1/W16-2524 , booktitle=

  26. [35]

    Transactions of the Association for Computational Linguistics , author=

    Amnesic. Transactions of the Association for Computational Linguistics , author=. 2021 , month=mar, pages=. doi:10.1162/tacl_a_00359 , abstractNote=

  27. [36]

    and Pimentel, Tiago and Saphra, Naomi and Cotterell, Ryan , editor=

    White, Jennifer C. and Pimentel, Tiago and Saphra, Naomi and Cotterell, Ryan , editor=. A. Proceedings of the 2021 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies , publisher=. 2021 , month=jun, pages=. doi...

  28. [37]

    , editor=

    Hewitt, John and Manning, Christopher D. , editor=. A. Proceedings of the 2019 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, Volume 1 (Long and Short Papers) , publisher=. 2019 , month=jun, pages=. doi:1...

  29. [38]

    Tenney, Ian and Das, Dipanjan and Pavlick, Ellie , editor=. B. Proceedings of the 57th Annual Meeting of the Association for Computational Linguistics , publisher=. 2019 , month=jul, pages=. doi:10.18653/v1/P19-1452 , abstractNote=

  30. [39]

    Goldberg, Adele E. , year=. Constructions:

  31. [40]

    Croft, William , year=. Radical

  32. [41]

    Goldberg, Adele E. , year=. Constructions at

  33. [42]

    Tseng, Yu-Hsiang and Shih, Cing-Fang and Chen, Pin-Er and Chou, Hsin-Yu and Ku, Mao-Chang and Hsieh, Shu-Kai , editor=. Cx. Proceedings of the Thirteenth Language Resources and Evaluation Conference , publisher=. 2022 , month=jun, pages=

  34. [44]

    Veenboer, Tim and Bloem, Jelke , editor=. Using. Findings of the Association for Computational Linguistics: ACL 2023 , publisher=. 2023 , month=jul, pages=. doi:10.18653/v1/2023.findings-acl.819 , abstractNote=

  35. [45]

    Chronis, Gabriella and Mahowald, Kyle and Erk, Katrin , editor=. A. Proceedings of the 61st Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers) , publisher=. 2023 , month=jul, pages=. doi:10.18653/v1/2023.acl-long.14 , abstractNote=

  36. [46]

    Pannitto, Ludovica and Herbelot, Aurélie , editor=. C. Proceedings of the First International Workshop on Construction Grammars and NLP (CxGs+NLP, GURT/SyntaxFest 2023) , publisher=. 2023 , month=mar, pages=

  37. [47]

    Assessing

    Goldberg, Yoav , year=. Assessing. doi:10.48550/arXiv.1901.05287 , abstractNote=

  38. [48]

    Reconstruction

    Kim, Najoung and Khilnani, Jatin and Warstadt, Alex and Qaddoumi, Abed , year=. Reconstruction. doi:10.48550/arXiv.2212.10792 , abstractNote=

  39. [49]

    Kulmizev, Artur and Ravishankar, Vinit and Abdou, Mostafa and Nivre, Joakim , editor=. Do. Proceedings of the 58th Annual Meeting of the Association for Computational Linguistics , publisher=. 2020 , month=jul, pages=. doi:10.18653/v1/2020.acl-main.375 , abstractNote=

  40. [50]

    Wolf, Thomas and Debut, Lysandre and Sanh, Victor and Chaumond, Julien and Delangue, Clement and Moi, Anthony and Cistac, Pierric and Rault, Tim and Louf, Rémi and Funtowicz, Morgan and Davison, Joe and Shleifer, Sam and von Platen, Patrick and Ma, Clara and Jernite, Yacine an...

  41. [51]

    Lin, Yongjie and Tan, Yi Chern and Frank, Robert , editor=. Open. Proceedings of the 2019 ACL Workshop BlackboxNLP: Analyzing and Interpreting Neural Networks for NLP , publisher=. 2019 , month=aug, pages=. doi:10.18653/v1/W19-4825 , abstractNote=

  42. [52]

    Information-

    Pimentel, Tiago and Valvoda, Josef and Maudslay, Rowan Hall and Zmigrod, Ran and Williams, Adina and Cotterell, Ryan , editor=. Information-. Proceedings of the 58th Annual Meeting of the Association for Computational Linguistics , publisher=. 2020 , month=jul, pages=. doi:10....

  43. [53]

    Goldberg, Adele E. , year=. Usage-based constructionist approaches and Large Language Models , url=. doi:10.31234/osf.io/8bmwz , abstractNote=

  44. [54]

    Constructing a Language: A Usage-Based Theory of Language Acquisition , ISBN=

    Tomasello, Michael , year=. Constructing a Language: A Usage-Based Theory of Language Acquisition , ISBN=

  45. [55]

    Lingbuzz , author=

    Modern language models refute Chomsky’s approach to language , volume=. Lingbuzz , author=. 2023 , language=

  46. [56]

    Mortensen, David and Levin, Lori and Schütze, Hinrich , editor=

    Weissweiler, Leonie and He, Taiqi and Otani, Naoki and R. Mortensen, David and Levin, Lori and Schütze, Hinrich , editor=. Construction Grammar Provides Unique Insight into Neural Language Models , url=. Proceedings of the First International Workshop on Construction Grammars ...

  47. [57]

    and Levin, Lori , editor=

    Zhou, Shijia and Weissweiler, Leonie and He, Taiqi and Schütze, Hinrich and Mortensen, David R. and Levin, Lori , editor=. Constructions. Proceedings of the 2024 Joint International Conference on Computational Linguistics, Language Resources and Evaluation (LREC-COLING 2024) ,...

  48. [58]

    and Izrailevitch, Valentina and Xiao, Yunze and Schütze, Hinrich and Weissweiler, Leonie , editor=

    Mortensen, David R. and Izrailevitch, Valentina and Xiao, Yunze and Schütze, Hinrich and Weissweiler, Leonie , editor=. Verbing Weirds Language (Models): Evaluation of English Zero-Derivation in Five LLMs , url=. Proceedings of the 2024 Joint International Conference on Comput...

  49. [59]

    A C onstruction G rammar C orpus of V arying S chematicity: A D ataset for the E valuation of A bstractions in L anguage M odels

    Bonial, Claire and Tayyar Madabushi, Harish. A C onstruction G rammar C orpus of V arying S chematicity: A D ataset for the E valuation of A bstractions in L anguage M odels. Proceedings of the 2024 Joint International Conference on Computational Linguistics, Language Resource...

  50. [60]

    Advances in Neural Information Processing Systems , author=

    Chain-of-. Advances in Neural Information Processing Systems , author=. 2022 , month=dec, pages=

  51. [61]

    Building

    Schäfer, Roland and Bildhauer, Felix , editor=. Building. Proceedings of the Eighth International Conference on Language Resources and Evaluation (LREC’12) , publisher=. 2012 , month=may, pages=

  52. [62]

    Proceedings of the 3rd Workshop on Challenges in the Management of Large Corpora , author=

    Processing and querying large web corpora with the. Proceedings of the 3rd Workshop on Challenges in the Management of Large Corpora , author=. 2015 , pages=

  53. [63]

    2016 , eprint=

    Neural Machine Translation by Jointly Learning to Align and Translate , author=. 2016 , eprint=

  54. [64]

    A Primer in BERT ology: What We Know About How BERT Works

    Rogers, Anna and Kovaleva, Olga and Rumshisky, Anna. A Primer in BERT ology: What We Know About How BERT Works. Transactions of the Association for Computational Linguistics. 2020. doi:10.1162/tacl_a_00349

  55. [65]

    Proceedings of the National Academy of Sciences , volume=

    Emergent linguistic structure in artificial neural networks trained by self-supervision , author=. Proceedings of the National Academy of Sciences , volume=. 2020 , publisher=

  56. [66]

    Why C an GPT L earn I n- C ontext? L anguage M odels S ecretly P erform G radient Descent as M eta- O ptimizers

    Dai, Damai and Sun, Yutao and Dong, Li and Hao, Yaru and Ma, Shuming and Sui, Zhifang and Wei, Furu. Why C an GPT L earn I n- C ontext? L anguage M odels S ecretly P erform G radient Descent as M eta- O ptimizers. Findings of the Association for Computational Linguistics: ACL ...

  57. [67]

    Dai and Quoc V Le , booktitle=

    Jason Wei and Maarten Bosma and Vincent Zhao and Kelvin Guu and Adams Wei Yu and Brian Lester and Nan Du and Andrew M. Dai and Quoc V Le , booktitle=. Finetuned. 2022 , url=

  58. [68]

    Sheng Lu and Irina Bigoulaeva and Rachneet Sachdeva and Harish Tayyar Madabushi and Iryna Gurevych , year=. Are. 2309.01809 , archivePrefix=

  59. [69]

    Sainz, Oscar and Campos, Jon Ander and García-Ferrero, Iker and Etxaniz, Julen and Agirre, Eneko , year=. Did

  60. [70]

    NLP E valuation in trouble: O n the N eed to M easure LLM D ata C ontamination for each B enchmark

    Sainz, Oscar and Campos, Jon and Garc \' a-Ferrero, Iker and Etxaniz, Julen and de Lacalle, Oier Lopez and Agirre, Eneko. NLP E valuation in trouble: O n the N eed to M easure LLM D ata C ontamination for each B enchmark. Findings of the Association for Computational Linguisti...

  61. [71]

    Lewis, Martha and Mitchell, Melanie , journal=. Using

  62. [72]

    Data Contamination: From Memorization to Exploitation

    Magar, Inbal and Schwartz, Roy. Data Contamination: From Memorization to Exploitation. Proceedings of the 60th Annual Meeting of the Association for Computational Linguistics (Volume 2: Short Papers). 2022. doi:10.18653/v1/2022.acl-short.18

  63. [73]

    Advances in neural information processing systems , volume=

    Language models are few-shot learners , author=. Advances in neural information processing systems , volume=

  64. [74]

    Proceedings of the 2015 Conference on Empirical Methods in Natural Language Processing , year=

    A large annotated corpus for learning natural language inference , author=. Proceedings of the 2015 Conference on Empirical Methods in Natural Language Processing , year=

  65. [75]

    A brief history of natural logic , author=

  66. [76]

    SemEval-2014 Task 1: Evaluation of Compositional Distributional Semantic Models on Full Sentences through Semantic Relatedness and Textual Entailment , doi =

    Marelli, Marco and Bentivogli, Luisa and Baroni, Marco and Bernardi, Raffaella and Menini, Stefano and Zamparelli, Roberto , year =. SemEval-2014 Task 1: Evaluation of Compositional Distributional Semantic Models on Full Sentences through Semantic Relatedness and Textual Entai...

  67. [77]

    Proceedings of the 18th International Workshop on Semantic Evaluation (SemEval-2024) , pages=

    Nycu-nlp at semeval-2024 task 2: Aggregating large language models in biomedical natural language inference for clinical trials , author=. Proceedings of the 18th International Workshop on Semantic Evaluation (SemEval-2024) , pages=

  68. [78]

    The Cambridge handbook of child language , pages=

    The usage-based theory of language acquisition , author=. The Cambridge handbook of child language , pages=. 2009 , publisher=

  69. [79]

    Chomsky, Noam , year=. The

  70. [80]

    Computational learning of construction grammars , volume=

    Dunn, Jonathan , year=. Computational learning of construction grammars , volume=. Language and Cognition , publisher=. doi:10.1017/langcog.2016.7 , abstractNote=

  71. [81]

    Exploring the Constructicon: Linguistic Analysis of a Computational CxG , url=

    Dunn, Jonathan , editor=. Exploring the Constructicon: Linguistic Analysis of a Computational CxG , url=. Proceedings of the First International Workshop on Construction Grammars and NLP (CxGs+NLP, GURT/SyntaxFest 2023) , publisher=. 2023 , month=mar, pages=

  72. [82]

    Language and Cognitive Processes , volume=

    Evidence for automatic accessing of constructional meaning: Jabberwocky sentences prime associated verbs , author=. Language and Cognitive Processes , volume=. 2013 , publisher=

  73. [83]

    Proceedings of the 17th Linguistic Annotation Workshop (LAW-XVII). 2023

  74. [84]

    Construction

    Hoffmann, Thomas , year=. Construction

  75. [85]

    2013 , publisher=

    The Oxford handbook of construction grammar , author=. 2013 , publisher=

  76. [86]

    Trends in cognitive sciences , volume=

    Constructions: A new theoretical approach to language , author=. Trends in cognitive sciences , volume=. 2003 , publisher=

  77. [87]

    construction grammar

    The mechanisms of" construction grammar" , author=. Annual Meeting of the Berkeley Linguistics Society , pages=

  78. [88]

    2014 , publisher=

    The minimalist program , author=. 2014 , publisher=

  79. [89]

    Mortensen, David and Levin, Lori and Sch

    Weissweiler, Leonie and He, Taiqi and Otani, Naoki and R. Mortensen, David and Levin, Lori and Sch. Construction Grammar Provides Unique Insight into Neural Language Models. Proceedings of the First International Workshop on Construction Grammars and NLP (CxGs+NLP, GURT/Syntax...

  80. [90]

    Constructions and Frames , volume=

    Usage-based constructionist approaches and large language models , author=. Constructions and Frames , volume=. 2024 , publisher=

  81. [91]

    Constructing

    Bonial, Claire and Tayyar Madabushi, Harish , journal=. Constructing. 2024 , publisher=

  82. [92]

    Beuls, Katrien and Van Eecke, Paul , journal=. Humans. 2024 , publisher=

  83. [94]

    Language

    Brown, Tom and Mann, Benjamin and Ryder, Nick and Subbiah, Melanie and Kaplan, Jared D and Dhariwal, Prafulla and Neelakantan, Arvind and Shyam, Pranav and Sastry, Girish and Askell, Amanda and Agarwal, Sandhini and Herbert-Voss, Ariel and Krueger, Gretchen and Henighan, Tom a...

  84. [95]

    Chi and Tatsunori Hashimoto and Oriol Vinyals and Percy Liang and Jeff Dean and William Fedus , journal=

    Jason Wei and Yi Tay and Rishi Bommasani and Colin Raffel and Barret Zoph and Sebastian Borgeaud and Dani Yogatama and Maarten Bosma and Denny Zhou and Donald Metzler and Ed H. Chi and Tatsunori Hashimoto and Oriol Vinyals and Percy Liang and Jeff Dean and William Fedus , jour...

  85. [96]

    2023 , eprint=

    Attention Is All You Need , author=. 2023 , eprint=

  86. [97]

    Proceedings of the First International Workshop on Construction Grammars and NLP (CxGs+NLP, GURT/SyntaxFest 2023). 2023

  87. [98]

    Gemini Team and Rohan Anil and Sebastian Borgeaud and others

    Gemini: A Family of Highly Capable Multimodal Models , author="Gemini Team and Rohan Anil and Sebastian Borgeaud and others", year=. 2312.11805 , archivePrefix=

  88. [99]

    arXiv preprint arXiv:2412.16720 , year=

    Openai o1 system card , author=. arXiv preprint arXiv:2412.16720 , year=

  89. [100]

    Baldassarre, Maria Teresa and Caivano, Danilo and Fernandez Nieto, Berenice and Gigante, Domenico and Ragone, Azzurra , booktitle=. The

  90. [101]

    P roceedings of the NAACL HLT W orkshop on E xtracting and U sing C onstructions in C omputational L inguistics. 2010

  91. [102]

    A bstract M eaning R epresentation of C onstructions: T he M ore W e I nclude, the B etter the R epresentation

    Bonial, Claire and Badarau, Bianca and Griffitt, Kira and Hermjakob, Ulf and Knight, Kevin and O ' Gorman, Tim and Palmer, Martha and Schneider, Nathan. A bstract M eaning R epresentation of C onstructions: T he M ore W e I nclude, the B etter the R epresentation. Proceedings ...

  92. [103]

    Leak, C heat, R epeat: D ata C ontamination and E valuation M alpractices in C losed- S ource LLM s

    Balloccu, Simone and Schmidtov \'a , Patr \'i cia and Lango, Mateusz and Dusek, Ondrej. Leak, C heat, R epeat: D ata C ontamination and E valuation M alpractices in C losed- S ource LLM s. Proceedings of the 18th Conference of the European Chapter of the Association for Comput...

  93. [104]

    2502.06215 , archivePrefix=

    Xin Zhou and Martin Weyssow and Ratnadira Widyasari and Ting Zhang and Junda He and Yunbo Lyu and Jianming Chang and Beiqi Zhang and Dan Huang and David Lo , year=. 2502.06215 , archivePrefix=

  94. [105]

    Reasoning or R eciting? E xploring the C apabilities and L imitations of L anguage M odels T hrough C ounterfactual T asks

    Wu, Zhaofeng and Qiu, Linlu and Ross, Alexis and Aky. Reasoning or R eciting? E xploring the C apabilities and L imitations of L anguage M odels T hrough C ounterfactual T asks. Proceedings of the 2024 Conference of the North American Chapter of the Association for Computation...

  95. [106]

    International journal of corpus linguistics , volume=

    Extending collostructional analysis: A corpus-based perspective onalternations' , author=. International journal of corpus linguistics , volume=. 2004 , publisher=

  96. [107]

    2010 , publisher=

    Language, usage and cognition , author=. 2010 , publisher=

  97. [108]

    Emrullah and Papailiopoulos, Dimitris and Oymak, Samet , title =

    Li, Yingcong and Ildiz, M. Emrullah and Papailiopoulos, Dimitris and Oymak, Samet , title =. Proceedings of the 40th International Conference on Machine Learning , articleno =. 2023 , publisher =

  98. [109]

    2023 , eprint=

    What and How does In-Context Learning Learn? Bayesian Model Averaging, Parameterization, and Generalization , author=. 2023 , eprint=

  99. [110]

    2025 , eprint=

    DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning , author=. 2025 , eprint=

  100. [111]

    2025 , eprint=

    Neither Stochastic Parroting nor AGI: LLMs Solve Tasks through Context-Directed Extrapolation from Training Data Priors , author=. 2025 , eprint=

  101. [112]

    The Thirteenth International Conference on Learning Representations , year=

    Data Mixing Laws: Optimizing Data Mixtures by Predicting Language Modeling Performance , author=. The Thirteenth International Conference on Learning Representations , year=

  102. [113]

    How Should Pre-Trained Language Models Be Fine-Tuned Towards Adversarial Robustness? , url =

    Dong, Xinshuai and Luu, Anh Tuan and Lin, Min and Yan, Shuicheng and Zhang, Hanwang , booktitle =. How Should Pre-Trained Language Models Be Fine-Tuned Towards Adversarial Robustness? , url =

  103. [114]

    Thirty-seventh Conference on Neural Information Processing Systems , year=

    Augmenting Language Models with Long-Term Memory , author=. Thirty-seventh Conference on Neural Information Processing Systems , year=

  104. [115]

    Brown and Benjamin Chess and Rewon Child and Scott Gray and Alec Radford and Jeffrey Wu and Dario Amodei , title =

    Jared Kaplan and Sam McCandlish and Tom Henighan and Tom B. Brown and Benjamin Chess and Rewon Child and Scott Gray and Alec Radford and Jeffrey Wu and Dario Amodei , title =. CoRR , volume =. 2020 , url =. 2001.08361 , timestamp =

  105. [116]

    Proceedings of the Second International Workshop on Construction Grammars & NLP (CxG+NLP 2025), co-located with IWCS 2025 , year =

    Bonial, Claire and Pellegrin, Taylor and Torgbi, Melissa and Tayyar Madabushi, Harish , title =. Proceedings of the Second International Workshop on Construction Grammars & NLP (CxG+NLP 2025), co-located with IWCS 2025 , year =

  106. [117]

    Cognitive science , volume=

    The metaphorical structure of the human conceptual system , author=. Cognitive science , volume=. 1980 , publisher=

  107. [118]

    Journal of semantics , volume=

    Transfers of meaning , author=. Journal of semantics , volume=. 1995 , publisher=

  108. [119]

    Metonymy in language and thought , volume=

    Towards a theory of metonymy , author=. Metonymy in language and thought , volume=. 1999 , publisher=

  109. [120]

    Computational linguistics , volume=

    The generative lexicon , author=. Computational linguistics , volume=

  110. [121]

    Cognitive Linguistics , doi =

    Type shifting in construction grammar: An integrated approach to aspectual coercion , author=. Cognitive Linguistics , doi =. 2004 , lastchecked =

  111. [122]

    Transportation Research Record: Journal of the Transportation Research Board , number=

    Theoretical maximum capacity as benchmark for empty vehicle redistribution in personal rapid transit , author=. Transportation Research Record: Journal of the Transportation Research Board , number=. 2010 , publisher=. doi:34251 , url =

  112. [123]

    Transportation Research Record: Journal of the Transportation Research Board , number=

    A Different Title to Test Repeated Authors , author=. Transportation Research Record: Journal of the Transportation Research Board , number=. 2011 , publisher=. doi:34251 , url =

  113. [124]

    Journal of Field Robotics , volume=

    Autonomous driving in urban environments: Boss and the urban challenge , author=. Journal of Field Robotics , volume=. 2008 , publisher=

  114. [125]

    2012 , organization=

    Are we ready for autonomous driving? The kitti vision benchmark suite , author=. 2012 , organization=

  115. [126]

    Grounding `Grounding' in NLP

    Chandu, Khyathi Raghavi and Bisk, Yonatan and Black, Alan W. Grounding `Grounding' in NLP. Findings of the Association for Computational Linguistics: ACL-IJCNLP 2021. 2021. doi:10.18653/v1/2021.findings-acl.375

  116. [127]

    , author=

    Grounding in communication. , author=. 1991 , publisher=

  117. [128]

    1998 , publisher=

    The importance of being earnest and other plays , author=. 1998 , publisher=

  118. [129]

    Literary and Linguistic Computing , author=

    The. Literary and Linguistic Computing , author=. 2010 , month=dec, pages=. doi:10.1093/llc/fqq018 , abstractNote=

  119. [130]

    Aho and Jeffrey D

    Alfred V. Aho and Jeffrey D. Ullman , title =. 1972

  120. [131]

    Publications Manual , year = "1983", publisher =

  121. [132]

    Chandra and Dexter C

    Ashok K. Chandra and Dexter C. Kozen and Larry J. Stockmeyer , year = "1981", title =. doi:10.1145/322234.322243

  122. [133]

    Scalable training of

    Andrew, Galen and Gao, Jianfeng , booktitle=. Scalable training of. 2007 , url=

  123. [134]

    Dan Gusfield , title =. 1997

  124. [135]

    Tetreault , title =

    Mohammad Sadegh Rasooli and Joel R. Tetreault , title =. Computing Research Repository , volume =. 2015 , url =

  125. [136]

    A Framework for Learning Predictive Structures from Multiple Tasks and Unlabeled Data , Volume =

    Ando, Rie Kubota and Zhang, Tong , Issn =. A Framework for Learning Predictive Structures from Multiple Tasks and Unlabeled Data , Volume =. Journal of Machine Learning Research , Month = dec, Numpages =. 2005 , url=

  126. [137]

    and Tukey, John W

    Cooley, James W. and Tukey, John W. , journal=. An algorithm for the machine calculation of complex. 1965 , url=

  127. [138]

    2013 , month =

    Hoffmann, Thomas and Trousdale, Graeme , title = ". 2013 , month =. doi:10.1093/oxfordhb/9780195396683.001.0001 , url =

  128. [139]

    Course in general linguistics , volume=

    Nature of the linguistic sign , author=. Course in general linguistics , volume=. 1916 , publisher=

  129. [140]

    The minimalist program , author=

  130. [141]

    2010 , publisher=

    Meaning and the lexicon: the parallel architecture 1975-2010 , author=. 2010 , publisher=

  131. [142]

    GURT 2014 (Georgetown University Round Table on Languages and Linguistics) , pages=

    Fluid Construction Grammar: State of the art and future outlook , author=. GURT 2014 (Georgetown University Round Table on Languages and Linguistics) , pages=

  132. [143]

    Constructions , year=

    Construction grammar for kids , author=. Constructions , year=

  133. [144]

    2005 , publisher=

    Constructing a language: A usage-based theory of language acquisition , author=. 2005 , publisher=

  134. [145]

    Construction Grammar , DOI=

    Hoffmann, Thomas , year=. Construction Grammar , DOI=

  135. [146]

    1995 , publisher=

    Constructions: A construction grammar approach to argument structure , author=. 1995 , publisher=

  136. [147]

    1992 , publisher=

    Semantic structures , author=. 1992 , publisher=

  137. [148]

    , editor =

    Michaelis, Laura A. , editor =. Sign-based Construction Grammar , booktitle =. 2013 , month =. doi:10.1093/oxfordhb/9780195396683.013.0008 , url =

  138. [149]

    Berkeley construction grammar , author=

  139. [150]

    Language , volume =

    Fillmore, Charles and Kay, Paul and O'Connor, Mary , title =. Language , volume =. 1988 , publisher =

  140. [151]

    Sign-based construction grammar , number=

    The framenet constructicon , author=. Sign-based construction grammar , number=. 2012 , publisher=

  141. [152]

    The corpus of contemporary American English (COCA): 560 million words, 1990-present , author=

  142. [153]

    2020 , note =

    Kevin Knight and others , title =. 2020 , note =

  143. [154]

    arXiv preprint arXiv:2210.13181 , year=

    The better your syntax, the better your semantics? probing pretrained language models for the English comparative correlative , author=. arXiv preprint arXiv:2210.13181 , year=

  144. [155]

    Proceedings of the 7th linguistic annotation workshop and interoperability with discourse , pages=

    Abstract meaning representation for sembanking , author=. Proceedings of the 7th linguistic annotation workshop and interoperability with discourse , pages=

  145. [156]

    Proceedings of the Eleventh International Conference on Language Resources and Evaluation (LREC 2018) , year=

    Abstract Meaning Representation of constructions: The more we include, the better the representation , author=. Proceedings of the Eleventh International Conference on Language Resources and Evaluation (LREC 2018) , year=

  146. [157]

    Computational linguistics , volume=

    The proposition bank: An annotated corpus of semantic roles , author=. Computational linguistics , volume=. 2005 , publisher=

  147. [158]

    Advances in neural information processing systems , volume=

    Attention is all you need , author=. Advances in neural information processing systems , volume=

  148. [159]

    arXiv preprint arXiv:1906.11511 , year=

    Inducing syntactic trees from bert representations , author=. arXiv preprint arXiv:1906.11511 , year=

  149. [160]

    A structural probe for finding syntax in word representations , author=. Proceedings of the 2019 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, Volume 1 (Long and Short Papers) , pages=

  150. [161]

    arXiv preprint arXiv:2005.04511 , year=

    Finding universal grammatical relations in multilingual BERT , author=. arXiv preprint arXiv:2005.04511 , year=

  151. [162]

    Studies in linguistic analysis , pages=

    A synopsis of linguistic theory, 1930-1955 , author=. Studies in linguistic analysis , pages=

  152. [163]

    Extending the scope of Construction Grammar , volume=

    A radically data-driven Construction Grammar: Experiments with Dutch causative constructions , author=. Extending the scope of Construction Grammar , volume=. 2014 , publisher=

  153. [164]

    Corpus Linguistics and Linguistic Theory , volume=

    Recent change in the productivity and schematicity of the way-construction: A distributional semantic analysis , author=. Corpus Linguistics and Linguistic Theory , volume=. 2018 , publisher=

  154. [165]

    Language and cognition , volume=

    Computational learning of construction grammars , author=. Language and cognition , volume=. 2017 , publisher=

  155. [166]

    arXiv preprint arXiv:2301.12642 , year=

    Exploring the Constructicon: Linguistic Analysis of a Computational CxG , author=. arXiv preprint arXiv:2301.12642 , year=

  156. [167]

    Journal of memory and language , volume=

    Constructing meaning: The role of affordances and grammatical constructions in sentence comprehension , author=. Journal of memory and language , volume=. 2000 , publisher=

  157. [168]

    The Cambridge Handbook of Construction Grammar , editor =

    Tayyar Madabushi, Harish and Laurence Romain and Petar Milin and Dagmar Divjak , title =. The Cambridge Handbook of Construction Grammar , editor =. 2024 , note =

  158. [169]

    Neural reality of argument structure constructions

    Li, Bai and Zhu, Zining and Thomas, Guillaume and Rudzicz, Frank and Xu, Yang. Neural reality of argument structure constructions. Proceedings of the 60th Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers). 2022. doi:10.18653/v1/2022.acl-long.512

  159. [170]

    arXiv preprint arXiv:2010.05358 , year=

    Learning which features matter: RoBERTa acquires a preference for linguistic generalizations (eventually) , author=. arXiv preprint arXiv:2010.05358 , year=

  160. [171]

    China National Conference on Chinese Computational Linguistics , pages=

    A robustly optimized BERT pre-training approach with post-training , author=. China National Conference on Chinese Computational Linguistics , pages=. 2021 , organization=

  161. [172]

    arXiv preprint arXiv:2206.07682 , year=

    Emergent abilities of large language models , author=. arXiv preprint arXiv:2206.07682 , year=

  162. [173]

    Language , pages=

    From usage to grammar: The mind's response to repetition , author=. Language , pages=. 2006 , publisher=

  163. [174]

    , title =

    Goldberg, Adele E. , title =. Constructions and Frames , year =

  164. [175]

    2023 , eprint=

    GPT-4 Technical Report , author=. 2023 , eprint=

  165. [176]

    Language Models are Few-Shot Learners , url =

    Brown, Tom and Mann, Benjamin and Ryder, Nick and Subbiah, Melanie and Kaplan, Jared D and Dhariwal, Prafulla and Neelakantan, Arvind and Shyam, Pranav and Sastry, Girish and Askell, Amanda and Agarwal, Sandhini and Herbert-Voss, Ariel and Krueger, Gretchen and Henighan, Tom a...

  166. [177]

    Attention is All you Need , url =

    Vaswani, Ashish and Shazeer, Noam and Parmar, Niki and Uszkoreit, Jakob and Jones, Llion and Gomez, Aidan N and Kaiser, ukasz and Polosukhin, Illia , booktitle =. Attention is All you Need , url =. 2017 , address =

  167. [178]

    International Conference on Learning Representations , year=

    Finetuned Language Models are Zero-Shot Learners , author=. International Conference on Learning Representations , year=

  168. [179]

    arXiv preprint arXiv:2304.15004 , year=

    Are emergent abilities of Large Language Models a mirage? , author=. arXiv preprint arXiv:2304.15004 , year=

  169. [180]

    S em E val-2017 Task 1: Semantic Textual Similarity Multilingual and Crosslingual Focused Evaluation

    Cer, Daniel and Diab, Mona and Agirre, Eneko and Lopez-Gazpio, I \ n igo and Specia, Lucia. S em E val-2017 Task 1: Semantic Textual Similarity Multilingual and Crosslingual Focused Evaluation. Proceedings of the 11th International Workshop on Semantic Evaluation ( S em E val-...

  170. [181]

    2023 , eprint=

    Are Emergent Abilities in Large Language Models just In-Context Learning? , author=. 2023 , eprint=

  171. [182]

    Theoretical Issues in Natural Language Processing

    M insky ' s Frame System Theory. Theoretical Issues in Natural Language Processing. 1975

  172. [183]

    arXiv preprint arXiv:2309.00135 , year=

    Construction grammar and artificial intelligence , author=. arXiv preprint arXiv:2309.00135 , year=

  173. [184]

    Towards a unified usage-based model of grammar and meaning , author=

    Distributional semantics meets Construction Grammar. Towards a unified usage-based model of grammar and meaning , author=. First International Workshop on Designing Meaning Representations (DMR 2019) , year=

  174. [185]

    Introducing Construction Semantics (CxS): a frame-semantic extension of Construction Grammar and constructicography , title =

    Alexander Willich , pages =. Introducing Construction Semantics (CxS): a frame-semantic extension of Construction Grammar and constructicography , title =. Linguistics Vanguard , doi =. 2022 , lastchecked =

  175. [186]

    Proceedings of the Language Resources and Evaluation Conference (LREC) , year=

    A Construction Grammar Corpus of Varying Schematicity: A Dataset for the Evaluation of Abstractions in Language Models , author=. Proceedings of the Language Resources and Evaluation Conference (LREC) , year=

  176. [187]

    2009 , publisher=

    Usage-based and emergentist approaches to language acquisition , author=. 2009 , publisher=

  177. [188]

    Findings of the Association for Computational Linguistics: EACL 2023 , pages=

    Modelling language acquisition through syntactico-semantic pattern finding , author=. Findings of the Association for Computational Linguistics: EACL 2023 , pages=

  178. [189]

    Proceedings of the 29th International Conference on Computational Linguistics , pages=

    Language acquisition through intention reading and pattern finding , author=. Proceedings of the 29th International Conference on Computational Linguistics , pages=

  179. [190]

    Manning , title =

    Peng Qi and Yuhao Zhang and Yuhui Zhang and Jason Bolton and Christopher D. Manning , title =. CoRR , volume =. 2020 , url =. 2003.07082 , timestamp =

  180. [191]

    2023 , eprint=

    Larger language models do in-context learning differently , author=. 2023 , eprint=

  181. [192]

    Training language models to follow instructions with human feedback , url =

    Ouyang, Long and Wu, Jeffrey and Jiang, Xu and Almeida, Diogo and Wainwright, Carroll and Mishkin, Pamela and Zhang, Chong and Agarwal, Sandhini and Slama, Katarina and Ray, Alex and Schulman, John and Hilton, Jacob and Kelton, Fraser and Miller, Luke and Simens, Maddie and As...

  182. [193]

    2023 , eprint=

    Dissociating language and thought in large language models: a cognitive perspective , author=. 2023 , eprint=

  183. [194]

    arXiv preprint arXiv:2010.02375 , year=

    Investigating representations of verb bias in neural language models , author=. arXiv preprint arXiv:2010.02375 , year=

  184. [195]

    Frontiers in Psychology , volume=

    Can recurrent neural networks validate usage-based theories of grammar acquisition? , author=. Frontiers in Psychology , volume=. 2022 , publisher=

  185. [196]

    Proceedings of the Thirteenth Language Resources and Evaluation Conference , pages=

    CxLM: A construction and context-aware language model , author=. Proceedings of the Thirteenth Language Resources and Evaluation Conference , pages=

  186. [197]

    Findings of the Association for Computational Linguistics: ACL 2023 , pages=

    Using collostructional analysis to evaluate BERT’s representation of linguistic constructions , author=. Findings of the Association for Computational Linguistics: ACL 2023 , pages=

  187. [198]

    Proceedings of the 2023 CLASP Conference on Learning with Small Data (LSD) , pages=

    Entrenchment Matters: Investigating Positional and Constructional Sensitivity in Small and Large Language Models , author=. Proceedings of the 2023 CLASP Conference on Learning with Small Data (LSD) , pages=

  188. [199]

    Frontiers in Artificial Intelligence , volume=

    Explaining pretrained language models' understanding of linguistic structures using construction grammar , author=. Frontiers in Artificial Intelligence , volume=. 2023 , publisher=

  189. [200]

    arXiv preprint arXiv:2302.02178 , year=

    Construction grammar provides unique insight into neural language models , author=. arXiv preprint arXiv:2302.02178 , year=

  190. [201]

    2018 , publisher=

    Constructicography: Constructicon development across languages , author=. 2018 , publisher=

  191. [202]

    International journal of corpus linguistics , volume=

    Collostructions: Investigating the interaction of words and constructions , author=. International journal of corpus linguistics , volume=. 2003 , publisher=

  192. [203]

    Language Universals and Linguistic Typology , address =

    Bernard Comrie , year =. Language Universals and Linguistic Typology , address =

  193. [204]

    Syntactic structures , address =

    Noam Chomsky , year =. Syntactic structures , address =

  194. [205]

    On the sound-system of central

    Bloomfield, Leonard , journal =. On the sound-system of central

  195. [206]

    Analogy, leveling, markedness:

    Analogy, leveling, markedness:. Analogy, leveling, markedness:

  196. [207]

    Sebastian Nordhoff , year =

  197. [208]

    Multiword Expressions between the Corpus and the Lexicon: Universality, Idiosyncrasy, and the Lexicon-Corpus Interface

    Barbu Mititelu, Verginica and Giouli, Voula and Evang, Kilian and Zeman, Daniel and Osenova, Petya and Tiberius, Carole and Krek, Simon and Markantonatou, Stella and Stoyanova, Ivelina and Stankovi \'c , Ranka and Chiarcos, Christian. Multiword Expressions between the Corpus a...

  198. [209]

    Northern European Journal of Language Technology , volume=

    Parseme meets universal dependencies: Getting on the same page in representing multiword expressions , author=. Northern European Journal of Language Technology , volume=

  199. [210]

    The Swedish FrameNet++ , pages=

    Multiword expressions--a tough typological nut for Swedish FrameNet++ , author=. The Swedish FrameNet++ , pages=. 2021 , publisher=

  200. [211]

    Word: A cross-linguistic typology , pages=

    Word: A typological framework , author=. Word: A cross-linguistic typology , pages=

  201. [212]

    Computational Linguistics and Intelligent Text Processing: Third International Conference, CICLing 2002 Mexico City, Mexico, February 17--23, 2002 Proceedings 3 , pages=

    Multiword expressions: A pain in the neck for NLP , author=. Computational Linguistics and Intelligent Text Processing: Third International Conference, CICLing 2002 Mexico City, Mexico, February 17--23, 2002 Proceedings 3 , pages=. 2002 , organization=

  202. [213]

    , author=

    Multiword expressions. , author=. Handbook of natural language processing , volume=

  203. [214]

    Word , volume=

    Defining the word , author=. Word , volume=. 2023 , publisher=

  204. [215]

    Alan , year=

    Croft, William and Cruse, D. Alan , year=. From idioms to construction grammar , booktitle=

  205. [216]

    2022 , publisher=

    Morphosyntax: constructions of the world's languages , author=. 2022 , publisher=

  206. [217]

    1957 , publisher=

    Syntactic structures , author=. 1957 , publisher=

  207. [218]

    Language , volume=

    Regularity and Idiomaticity in Grammatical Constructions: The Case of Let Alone , author=. Language , volume=

  208. [219]

    Language , volume=

    From usage to grammar: The mind's response to repetition , author=. Language , volume=. 2006 , publisher=

  209. [220]

    Language , volume=

    A usage-based approach to Spanish verbs of'becoming' , author=. Language , volume=. 2006 , publisher=

  210. [221]

    Linguistics , doi =

    The partial productivity of constructions as induction , author =. Linguistics , doi =. 2011 , lastchecked =

  211. [222]

    Literary theory: An anthology , volume=

    Course in general linguistics , author=. Literary theory: An anthology , volume=. 2004 , publisher=

  212. [223]

    Studies in second language acquisition , volume=

    Phonological evidence for exemplar storage of multiword sequences , author=. Studies in second language acquisition , volume=. 2002 , publisher=

  213. [224]

    Dialect and language variation , pages=

    The social stratification of (r) in New York City department stores , author=. Dialect and language variation , pages=. 1986 , publisher=

  214. [225]

    Language variation and change , volume=

    The intersection of sex and social class in the course of linguistic change , author=. Language variation and change , volume=. 1990 , publisher=

  215. [226]

    Cowell, Andrew and Moss Sr., Alonzo , year=

  216. [227]

    A Conversational Database of the Arapaho Language in Video Format , author=

  217. [228]

    2021 , url=

    Arapaho Text Database , author=. 2021 , url=

  218. [229]

    Dictionaries: Journal of the Dictionary Society of North America , volume=

    The Problem of Polysynthesis in UMR Annotations: Complexities in Handling Preverbal Modification and Noun Incorporation in Arapaho , author=. Dictionaries: Journal of the Dictionary Society of North America , volume=. 2025 , publisher=

  219. [230]

    Proceedings of the 2024 Joint International Conference on Computational Linguistics, Language Resources and Evaluation (LREC-COLING 2024) , pages=

    Bootstrapping UMR Annotations for Arapaho from Language Documentation Resources , author=. Proceedings of the 2024 Joint International Conference on Computational Linguistics, Language Resources and Evaluation (LREC-COLING 2024) , pages=

  220. [231]

    UMR Annotation of Multiword Expressions

    Bonn, Julia and Cowell, Andrew and Haji c , Jan and Palmer, Alexis and Palmer, Martha and Pustejovsky, James and Sun, Haibo and Uresova, Zdenka and Wein, Shira and Xue, Nianwen and Zhao, Jin. UMR Annotation of Multiword Expressions. Proceedings of the Fourth International Work...

  221. [232]

    Proceedings of the 2024 Joint International Conference on Computational Linguistics, Language Resources and Evaluation (LREC-COLING 2024) , pages=

    Building a Broad Infrastructure for Uniform Meaning Representations , author=. Proceedings of the 2024 Joint International Conference on Computational Linguistics, Language Resources and Evaluation (LREC-COLING 2024) , pages=

  222. [233]

    Journal of the American Medical Informatics Association , volume=

    Towards comprehensive syntactic and semantic annotations of the clinical narrative , author=. Journal of the American Medical Informatics Association , volume=. 2013 , publisher=

  223. [234]

    Proceedings of the 11th Joint Conference on Lexical and Computational Semantics , pages=

    PropBank comes of Age—Larger, smarter, and more diverse , author=. Proceedings of the 11th Joint Conference on Lexical and Computational Semantics , pages=

  224. [235]

    Handbook of linguistic annotation , pages=

    Current directions in english and arabic propbank , author=. Handbook of linguistic annotation , pages=. 2017 , publisher=

  225. [236]

    language , volume=

    Thematic proto-roles and argument selection , author=. language , volume=. 1991 , publisher=

  226. [237]

    , author=

    PropBank: Semantics of New Predicate Types. , author=. LREC , pages=

  227. [238]

    The Oxford Handbook of Cognitive Science , year=

    VerbNet , author=. The Oxford Handbook of Cognitive Science , year=

  228. [239]

    Center for Computational Language and Education Research Institute of Cognitive Science University of Colorado at Boulder , year=

    Propbank annotation guidelines , author=. Center for Computational Language and Education Research Institute of Cognitive Science University of Colorado at Boulder , year=

  229. [240]

    Human Language Technology: Proceedings of a Workshop held at Plainsboro, New Jersey, March 8-11, 1994 , year=

    The Penn treebank: Annotating predicate argument structure , author=. Human Language Technology: Proceedings of a Workshop held at Plainsboro, New Jersey, March 8-11, 1994 , year=

  230. [241]

    , author=

    Propbank Instance Annotation Guidelines Using a Dedicated Editor, Jubilee. , author=. LREC , year=

  231. [242]

    Transactions on Machine Learning Research , issn=

    Beyond the Imitation Game: Quantifying and extrapolating the capabilities of language models , author=. Transactions on Machine Learning Research , issn=. 2023 , url=

  232. [243]

    Transactions on Machine Learning Research , issn=

    Emergent Abilities of Large Language Models , author=. Transactions on Machine Learning Research , issn=. 2022 , url=

  233. [244]

    2023 , eprint=

    A Survey of Hallucination in Large Foundation Models , author=. 2023 , eprint=

  234. [245]

    `` You Are An Expert Linguistic Annotator '' : Limits of LLM s as Analyzers of A bstract M eaning R epresentation

    Ettinger, Allyson and Hwang, Jena and Pyatkin, Valentina and Bhagavatula, Chandra and Choi, Yejin. `` You Are An Expert Linguistic Annotator '' : Limits of LLM s as Analyzers of A bstract M eaning R epresentation. Findings of the Association for Computational Linguistics: EMNL...

  235. [246]

    Treebanks: Building and using parsed corpora , pages=

    The Penn treebank: an overview , author=. Treebanks: Building and using parsed corpora , pages=. 2003 , publisher=

  236. [247]

    Language resources and evaluation , volume=

    A Chinese semantic lexicon of senses and roles , author=. Language resources and evaluation , volume=. 2006 , publisher=

  237. [248]

    LDC Catalog No.: LDC2006T03 ISBN , pages=

    Korean propbank , author=. LDC Catalog No.: LDC2006T03 ISBN , pages=

  238. [249]

    Proceedings of the fourth linguistic annotation workshop , pages=

    The revised arabic propbank , author=. Proceedings of the fourth linguistic annotation workshop , pages=

  239. [250]

    Proceedings of the 9th Workshop on Multiword Expressions , pages=

    Semantic roles for nominal predicates: Building a lexical resource , author=. Proceedings of the 9th Workshop on Multiword Expressions , pages=

  240. [251]

    , author=

    Propbank-Br: a Brazilian Treebank annotated with semantic role labels. , author=. LREC , pages=

  241. [252]

    Language Resources and Evaluation , volume=

    Building the essential resources for Finnish: the Turku Dependency Treebank , author=. Language Resources and Evaluation , volume=. 2014 , publisher=

  242. [253]

    Language Resources and Evaluation , volume=

    Annotation of semantic roles for the Turkish proposition bank , author=. Language Resources and Evaluation , volume=. 2018 , publisher=

  243. [254]

    arXiv preprint arXiv:2202.12246 , year=

    Neural reality of argument structure constructions , author=. arXiv preprint arXiv:2202.12246 , year=

  244. [255]

    Few-Shot Semantic Parsing with Language Models Trained on Code

    Shin, Richard and Van Durme, Benjamin. Few-Shot Semantic Parsing with Language Models Trained on Code. Proceedings of the 2022 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies. 2022. doi:10.18653/v1/2022.naa...

  245. [256]

    Frontiers in Artificial Intelligence , volume=

    The unreasonable effectiveness of large language models in zero-shot semantic annotation of legal texts , author=. Frontiers in Artificial Intelligence , volume=. 2023 , publisher=

  246. [257]

    arXiv preprint arXiv:2402.13446 , year=

    Large Language Models for Data Annotation: A Survey , author=. arXiv preprint arXiv:2402.13446 , year=

  247. [258]

    2005 , publisher=

    VerbNet: A broad-coverage, comprehensive verb lexicon , author=. 2005 , publisher=

  248. [259]

    and Choi, Yejin

    Liu, Alisa and Swayamdipta, Swabha and Smith, Noah A. and Choi, Yejin. WANLI : Worker and AI Collaboration for Natural Language Inference Dataset Creation. Findings of the Association for Computational Linguistics: EMNLP 2022. 2022. doi:10.18653/v1/2022.findings-emnlp.508

  249. [260]

    arXiv preprint arXiv:2110.07178 , year=

    Symbolic knowledge distillation: from general language models to commonsense models , author=. arXiv preprint arXiv:2110.07178 , year=

  250. [261]

    and Wallace, Eric and Singh, Sameer

    Shin, Taylor and Razeghi, Yasaman and Logan IV, Robert L. and Wallace, Eric and Singh, Sameer. A uto P rompt: E liciting K nowledge from L anguage M odels with A utomatically G enerated P rompts. Proceedings of the 2020 Conference on Empirical Methods in Natural Language Proce...

  251. [262]

    arXiv preprint arXiv:2212.09246 , year=

    I2d2: Inductive knowledge distillation with neurologic and self-imitation , author=. arXiv preprint arXiv:2212.09246 , year=

  252. [263]

    arXiv preprint arXiv:2303.03846 , year=

    Larger language models do in-context learning differently , author=. arXiv preprint arXiv:2303.03846 , year=

  253. [264]

    Transactions of the Association for Computational Linguistics , volume=

    How abstract is linguistic generalization in large language models? Experiments with argument structure , author=. Transactions of the Association for Computational Linguistics , volume=. 2023 , publisher=

  254. [265]

    2014 , school=

    Take a look at this! Form, function and productivity of English light verb constructions , author=. 2014 , school=

  255. [266]

    Proceedings of the AAAI Conference on Artificial Intelligence , volume=

    English light verb construction identification using lexical knowledge , author=. Proceedings of the AAAI Conference on Artificial Intelligence , volume=

  256. [267]

    Proceedings of the Eleventh International Conference on Language Resources and Evaluation (LREC 2018) , year=

    The new Propbank: Aligning Propbank with AMR through POS unification , author=. Proceedings of the Eleventh International Conference on Language Resources and Evaluation (LREC 2018) , year=

  257. [268]

    Designing a uniform meaning representation for natural language processing , author=. KI-K. 2021 , publisher=

  258. [269]

    The 4th International Workshop on Designing Meaning Representations , year=

    UMR annotation of multiword expressions , author=. The 4th International Workshop on Designing Meaning Representations , year=

  259. [270]

    Are Emergent Abilities in Large Language Models just In-Context Learning?

    Lu, Sheng and Bigoulaeva, Irina and Sachdeva, Rachneet and Tayyar Madabushi, Harish and Gurevych, Iryna. Are Emergent Abilities in Large Language Models just In-Context Learning?. Proceedings of the 62nd Annual Meeting of the Association for Computational Linguistics (Volume 1...

  260. [271]

    Ivanova and Idan A

    Kyle Mahowald and Anna A. Ivanova and Idan A. Blank and Nancy Kanwisher and Joshua B. Tenenbaum and Evelina Fedorenko , keywords =. Dissociating language and thought in large language models , journal =. 2024 , issn =. doi:https://doi.org/10.1016/j.tics.2024.01.011 , url =

  261. [272]

    Word classes in radical construction grammar , author=

  262. [273]

    2001 , publisher=

    Radical construction grammar: Syntactic theory in typological perspective , author=. 2001 , publisher=

  263. [274]

    Proceedings of the 58th annual meeting of the association for computational linguistics , pages=

    Climbing towards NLU: On meaning, form, and understanding in the age of data , author=. Proceedings of the 58th annual meeting of the association for computational linguistics , pages=

  264. [275]

    Cognition and categorization , pages=

    Nonanalytic concept formation and memory for instances , author=. Cognition and categorization , pages=. 1978 , publisher=

  265. [276]

    Schema abstraction

    " Schema abstraction" in a multiple-trace memory model. , author=. Psychological review , volume=. 1986 , publisher=

  266. [277]

    , author=

    Context theory of classification learning. , author=. Psychological review , volume=. 1978 , publisher=

  267. [278]

    , author=

    Retention of abstract ideas. , author=. Journal of Experimental psychology , volume=. 1970 , publisher=

  268. [279]

    Cognitive psychology , volume=

    Family resemblances: Studies in the internal structure of categories , author=. Cognitive psychology , volume=. 1975 , publisher=

  269. [280]

    , author=

    An on-line investigation of prototype and exemplar strategies in classification. , author=. Journal of Experimental Psychology: Learning, Memory, and Cognition , volume=. 1989 , publisher=

  270. [281]

    Conceptual Structure, Discourse and Language/CSLI , year=

    The Way Constructions Grow , author=. Conceptual Structure, Discourse and Language/CSLI , year=

  271. [282]

    Linguistic Inquiry, Monograph one, The MIT press , year=

    Word formation in generative grammar Cambridge , author=. Linguistic Inquiry, Monograph one, The MIT press , year=

  272. [283]

    Linguistics in the Morning Calm/Hanshin , year=

    Lexical morphology and phonology , author=. Linguistics in the Morning Calm/Hanshin , year=

  273. [284]

    Corpus Linguistics and Linguistic Theory , volume=

    Corpus frequency and acceptability judgments: A study of morphosyntactic variants in Czech , author=. Corpus Linguistics and Linguistic Theory , volume=. 2012 , publisher=

  274. [285]

    Russian linguistics , volume=

    Morphosyntactic variation and syntactic constructions in Czech nominal declension: corpus frequency and native-speaker judgments , author=. Russian linguistics , volume=. 2012 , publisher=

  275. [286]

    2012 IEEE International Conference on Development and Learning and Epigenetic Robotics (ICDL) , pages=

    Adult learners use both entrenchment and preemption to infer grammatical constraints , author=. 2012 IEEE International Conference on Development and Learning and Epigenetic Robotics (ICDL) , pages=. 2012 , organization=

  276. [287]

    Cognitive Development , volume=

    The role of entrenchment in children’s and adults’ performance on grammaticality judgment tasks , author=. Cognitive Development , volume=. 2004 , publisher=

  277. [288]

    , author=

    Category-based induction. , author=. Psychological review , volume=. 1990 , publisher=

  278. [289]

    , author=

    Ideals, central tendency, and frequency of instantiation as determinants of graded structure in categories. , author=. Journal of experimental psychology: learning, memory, and cognition , volume=. 1985 , publisher=

  279. [290]

    arXiv preprint arXiv:2004.10151 , year=

    Experience grounds language , author=. arXiv preprint arXiv:2004.10151 , year=

  280. [291]

    Language Resources and Evaluation , volume=

    Constructing understanding: on the constructional information encoded in large language models , author=. Language Resources and Evaluation , volume=. 2025 , publisher=

  281. [292]

    Language , volume=

    Learning what not to say: The role of statistical preemption and categorization in a-adjective production , author=. Language , volume=. 2011 , publisher=

  282. [293]

    2019 , publisher=

    Explain me this: Creativity, competition, and the partial productivity of constructions , author=. 2019 , publisher=

  283. [294]

    Cognitive linguistics , volume=

    Corpus evidence of the viability of statistical preemption , author=. Cognitive linguistics , volume=

  284. [295]

    , author=

    Negative entrenchment: A usage-based approach to negative evidence. , author=. Cognitive Linguistics , volume=

  285. [296]

    Transactions of the Association for Computational Linguistics , volume=

    Assessing the ability of LSTMs to learn syntax-sensitive dependencies , author=. Transactions of the Association for Computational Linguistics , volume=

  286. [297]

    Transactions of the Association for Computational Linguistics , volume=

    BLiMP: The benchmark of linguistic minimal pairs for English , author=. Transactions of the Association for Computational Linguistics , volume=. 2020 , publisher=

  287. [298]

    Beyond Memorization: Assessing Semantic Generalization in Large Language Models Using Phrasal Constructions , author=. Proceedings of the 14th International Joint Conference on Natural Language Processing and the 4th Conference of the Asia-Pacific Chapter of the Association fo...

  288. [299]

    Proceedings of the 17th Conference of the European Chapter of the Association for Computational Linguistics , pages=

    A discerning several thousand judgments: GPT-3 rates the article+ adjective+ numeral+ noun construction , author=. Proceedings of the 17th Conference of the European Chapter of the Association for Computational Linguistics , pages=

  289. [300]

    Language , volume=

    The roles of verb semantics, entrenchment, and morphophonology in the retreat from dative argument-structure overgeneralization errors , author=. Language , volume=. 2012 , publisher=

Pith tools

Reviewed June 28, 2026 · model on record in the stance chip above.