Pith. sign in

REVIEW 2 cited by

Political DEBATE: Efficient Zero-shot and Few-shot Classifiers for Political Text

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2409.02078 v1 pith:SXKGSXGE submitted 2024-09-03 cs.CL

classification cs.CL
keywords modelsdocumentspoliticalclassificationfew-shotlanguagezero-shotability
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Social scientists quickly adopted large language models due to their ability to annotate documents without supervised training, an ability known as zero-shot learning. However, due to their compute demands, cost, and often proprietary nature, these models are often at odds with replication and open science standards. This paper introduces the Political DEBATE (DeBERTa Algorithm for Textual Entailment) language models for zero-shot and few-shot classification of political documents. These models are not only as good, or better than, state-of-the art large language models at zero and few-shot classification, but are orders of magnitude more efficient and completely open source. By training the models on a simple random sample of 10-25 documents, they can outperform supervised classifiers trained on hundreds or thousands of documents and state-of-the-art generative models with complex, engineered prompts. Additionally, we release the PolNLI dataset used to train these models -- a corpus of over 200,000 political documents with highly accurate labels across over 800 classification tasks.

Discussion (0). Sign in to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Political Leaning and Politicalness Classification of Texts

    cs.CL 2025-07 conditional novelty 6.0 of 10

    The authors compile large multi-dataset benchmarks for political leaning and politicalness classification, show that single-dataset models fail out-of-distribution, and release new models with improved cross-domain F1 scores.

  2. Generative Exaggeration in LLM Social Agents: Consistency, Bias, and Toxicity

    cs.HC 2025-07 conditional novelty 5.0 of 10

    When LLMs are given more context about a real social media user, they become more ideologically consistent but also more extreme, toxic, and stereotyped than the user actually is.

Pith tools