Pith. sign in

REVIEW 1 cited by

Claim Check-Worthiness Detection as Positive Unlabelled Learning

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2003.02736 v2 pith:B7HETG6Y submitted 2020-03-05 cs.CL cs.LG

classification cs.CLcs.LG
keywords detectioncheck-worthinessclaimtasksapproachcheck-worthycheckingfact
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

As the first step of automatic fact checking, claim check-worthiness detection is a critical component of fact checking systems. There are multiple lines of research which study this problem: check-worthiness ranking from political speeches and debates, rumour detection on Twitter, and citation needed detection from Wikipedia. To date, there has been no structured comparison of these various tasks to understand their relatedness, and no investigation into whether or not a unified approach to all of them is achievable. In this work, we illuminate a central challenge in claim check-worthiness detection underlying all of these tasks, being that they hinge upon detecting both how factual a sentence is, as well as how likely a sentence is to be believed without verification. As such, annotators only mark those instances they judge to be clear-cut check-worthy. Our best performing method is a unified approach which automatically corrects for this using a variant of positive unlabelled learning that finds instances which were incorrectly labelled as not check-worthy. In applying this, we out-perform the state of the art in two of the three tasks studied for claim check-worthiness detection in English.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. ALPET: Active Few-shot Learning for Citation Worthiness Detection in Low-Resource Wikipedia Languages

    cs.CL 2025-02 conditional novelty 6.0 of 10

    ALPET, an active-learning plus PET pipeline, detects citation-worthy sentences in Catalan, Basque and Albanian while needing roughly 58-72% fewer labeled examples than its CCW baseline.

Pith tools