Pith. sign in

REVIEW 1 cited by

Adam-Smith at SemEval-2023 Task 4: Discovering Human Values in Arguments with Ensembles of Transformer-based Models

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2305.08625 v1 pith:BWQM4774 submitted 2023-05-15 cs.CL cs.AI

classification cs.CLcs.AI
keywords argumentsmodelssystemtaskvaluesbest-performingensemblingf1-score
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

This paper presents the best-performing approach alias "Adam Smith" for the SemEval-2023 Task 4: "Identification of Human Values behind Arguments". The goal of the task was to create systems that automatically identify the values within textual arguments. We train transformer-based models until they reach their loss minimum or f1-score maximum. Ensembling the models by selecting one global decision threshold that maximizes the f1-score leads to the best-performing system in the competition. Ensembling based on stacking with logistic regressions shows the best performance on an additional dataset provided to evaluate the robustness ("Nahj al-Balagha"). Apart from outlining the submitted system, we demonstrate that the use of the large ensemble model is not necessary and that the system size can be significantly reduced.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Identifying and Understanding Human Values in Text: A Tailorable LLM-based Architecture

    cs.AI 2026-04 conditional novelty 5.0 of 10

    A modular LLM architecture generates value specifications from theory texts, then detects and intensity-rates human values in arbitrary text, evaluated at ~0.34 micro-F1 on ValueEval.

Pith tools