Pith. sign in

REVIEW 1 cited by

Text-Transport: Toward Learning Causal Effects of Natural Language

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2310.20697 v1 pith:ML4IZ6XE submitted 2023-10-31 cs.CL stat.ME

classification cs.CLstat.ME
keywords causaleffectslanguagetext-transportdatanaturaltextassumptions
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

As language technologies gain prominence in real-world settings, it is important to understand how changes to language affect reader perceptions. This can be formalized as the causal effect of varying a linguistic attribute (e.g., sentiment) on a reader's response to the text. In this paper, we introduce Text-Transport, a method for estimation of causal effects from natural language under any text distribution. Current approaches for valid causal effect estimation require strong assumptions about the data, meaning the data from which one can estimate valid causal effects often is not representative of the actual target domain of interest. To address this issue, we leverage the notion of distribution shift to describe an estimator that transports causal effects between domains, bypassing the need for strong assumptions in the target domain. We derive statistical guarantees on the uncertainty of this estimator, and we report empirical results and analyses that support the validity of Text-Transport across data settings. Finally, we use Text-Transport to study a realistic setting--hate speech on social media--in which causal effects do shift significantly between text domains, demonstrating the necessity of transport when conducting causal inference on natural language.

Discussion (0). Sign in to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Mitigating Hidden Confounding by Progressive Confounder Imputation via Large Language Models

    cs.CL 2025-06 reject novelty 5.0 of 10

    ProCI uses LLMs to iteratively generate and impute hidden confounders, then validates them with a conditional independence test to improve treatment effect estimation.

Pith tools