Pith. sign in

REVIEW 1 cited by

Inducing brain-relevant bias in natural language processing models

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1911.03268 v1 pith:J3J3XBQG submitted 2019-10-29 q-bio.NC cs.CLcs.LG

classification q-bio.NCcs.CLcs.LG
keywords languagebrainrepresentationsactivitylearnedmodelsfine-tuningfmri
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Progress in natural language processing (NLP) models that estimate representations of word sequences has recently been leveraged to improve the understanding of language processing in the brain. However, these models have not been specifically designed to capture the way the brain represents language meaning. We hypothesize that fine-tuning these models to predict recordings of brain activity of people reading text will lead to representations that encode more brain-activity-relevant language information. We demonstrate that a version of BERT, a recently introduced and powerful language model, can improve the prediction of brain activity after fine-tuning. We show that the relationship between language and brain activity learned by BERT during this fine-tuning transfers across multiple participants. We also show that, for some participants, the fine-tuned representations learned from both magnetoencephalography (MEG) and functional magnetic resonance imaging (fMRI) are better for predicting fMRI than the representations learned from fMRI alone, indicating that the learned representations capture brain-activity-relevant information that is not simply an artifact of the modality. While changes to language representations help the model predict brain activity, they also do not harm the model's ability to perform downstream NLP tasks. Our findings are notable for research on language understanding in the brain.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Inducing Human-like Biases in Moral Reasoning Language Models

    cs.AI 2024-11 conditional novelty 6.0 of 10

    Fine-tuning BERT-family language models on moral reasoning or fMRI data does not significantly improve how closely their internal activations match human brain activity during moral judgment, although larger models al...

Pith tools