REVIEW 1 cited by
Assessing agreement on classification tasks: the kappa statistic
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Currently, computational linguists and cognitive scientists working in the area of discourse and dialogue argue that their subjective judgments are reliable using several different statistics, none of which are easily interpretable or comparable to each other. Meanwhile, researchers in content analysis have already experienced the same difficulties and come up with a solution in the kappa statistic. We discuss what is wrong with reliability measures as they are currently used for discourse and dialogue work in computational linguistics and cognitive science, and argue that we would be better off as a field adopting techniques from content analysis.
Forward citations
Cited by 1 Pith paper
-
Automating Credit Card Limit Adjustments Using Machine Learning
An XGBoost model with cost-sensitive learning reports Cohen's kappa of 0.81 against a bank committee's credit card limit decisions and is proposed to automate the process.
Discussion (0). Continue with ORCID to comment.