REVIEW 3 cited by
minicons: Enabling Flexible Behavioral and Representational Analyses of Transformer Language Models
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
We present minicons, an open source library that provides a standard API for researchers interested in conducting behavioral and representational analyses of transformer-based language models (LMs). Specifically, minicons enables researchers to apply analysis methods at two levels: (1) at the prediction level -- by providing functions to efficiently extract word/sentence level probabilities; and (2) at the representational level -- by also facilitating efficient extraction of word/phrase level vectors from one or more layers. In this paper, we describe the library and apply it to two motivating case studies: One focusing on the learning dynamics of the BERT architecture on relative grammatical judgments, and the other on benchmarking 23 different LMs on zero-shot abductive reasoning. minicons is available at https://github.com/kanishkamisra/minicons
Forward citations
Cited by 3 Pith papers
-
VIP-MINGLE: A Corpus for Videoconference and In-Person Multimodal Interaction in Group Language Engagement
A paired within-subject multimodal corpus of 32 groups recorded in-person and via videoconference reveals setting-specific shifts in turn-taking, language complexity, facial expression, and rated enjoyment.
-
Is It JUST Semantics? A Case Study of Discourse Particle Understanding in LLMs
LLMs separate broad senses of "just" such as temporal and adjective, but on naturalistic sentences they struggle to distinguish the fine-grained discourse-particle senses.
-
semantic-features: A User-Friendly Tool for Studying Contextual Word Embeddings in Interpretable Semantic Spaces
A library and demo reveal that masked language models encode the dative construction's person-like versus place-like reading of ambiguous recipients.
Discussion (0). Continue with ORCID to comment.