Pith. sign in

REVIEW 2 cited by

Conceptual structure coheres in human cognition but not in large language models

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2304.02754 v2 pith:WL427PH5 submitted 2023-04-05 cs.AI cs.CLcs.LG

classification cs.AIcs.CLcs.LG
keywords humanstructureconceptuallanguagebehaviorcontemporaryllmsmodels
verification ladder T0 review T1 audit T2 compute T3 formal

Signed reviews

No signed human review yet.

0 comments
read the original abstract

Neural network models of language have long been used as a tool for developing hypotheses about conceptual representation in the mind and brain. For many years, such use involved extracting vector-space representations of words and using distances among these to predict or understand human behavior in various semantic tasks. Contemporary large language models (LLMs), however, make it possible to interrogate the latent structure of conceptual representations using experimental methods nearly identical to those commonly used with human participants. The current work utilizes three common techniques borrowed from cognitive psychology to estimate and compare the structure of concepts in humans and a suite of LLMs. In humans, we show that conceptual structure is robust to differences in culture, language, and method of estimation. Structures estimated from LLM behavior, while individually fairly consistent with those estimated from human behavior, vary much more depending upon the particular task used to generate responses--across tasks, estimates of conceptual structure from the very same model cohere less with one another than do human structure estimates. These results highlight an important difference between contemporary LLMs and human cognition, with implications for understanding some fundamental limitations of contemporary machine language.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. The "LLM World of Words" English free association norms generated by large language models

    cs.CL 2024-12 conditional novelty 6.0 of 10

    A new dataset of 3+ million free association responses from three LLMs, matched to human norms, with validation showing human-like semantic priming and gender bias patterns.

  2. Distinguishing AI-Generated and Human-Written Text Through Psycholinguistic Analysis

    cs.CL 2025-05 reject novelty 3.0 of 10

    A conceptual mapping from stylometric features to psycholinguistic theories is presented, but no empirical validation is provided.

Pith tools