Pith. sign in

Generating Synthetic Text Data to Evaluate Causal Inference Methods

1 Pith paper cite this work. Polarity classification is still indexing.

1 Pith paper citing it
abstract

Drawing causal conclusions from observational data requires making assumptions about the true data-generating process. Causal inference research typically considers low-dimensional data, such as categorical or numerical fields in structured medical records. High-dimensional and unstructured data such as natural language complicates the evaluation of causal inference methods; such evaluations rely on synthetic datasets with known causal effects. Models for natural language generation have been widely studied and perform well empirically. However, existing methods not immediately applicable to producing synthetic datasets for causal evaluations, as they do not allow for quantifying a causal effect on the text itself. In this work, we develop a framework for adapting existing generation models to produce synthetic text datasets with known causal effects. We use this framework to perform an empirical comparison of four recently-proposed methods for estimating causal effects from text data. We release our code and synthetic datasets.

citation-role summary

background 1

citation-polarity summary

fields

cs.CL 1

years

2024 1

verdicts

CONDITIONAL 1

roles

background 1

polarities

support 1

representative citing papers

Political-LLM: Large Language Models in Political Science

cs.CL · 2024-12-09 · conditional · novelty 5.0

A survey and taxonomy of LLM applications in political science, with a case study suggesting that larger LLMs reproduce ANES 2016 voting patterns more accurately than smaller ones.

citing papers explorer

Showing 1 of 1 citing paper.

  • Political-LLM: Large Language Models in Political Science cs.CL · 2024-12-09 · conditional · none · ref 163 · internal anchor

    A survey and taxonomy of LLM applications in political science, with a case study suggesting that larger LLMs reproduce ANES 2016 voting patterns more accurately than smaller ones.