Pith. sign in

REVIEW 2 cited by

pysentimiento: A Python Toolkit for Opinion Mining and Social NLP tasks

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2106.09462 v3 pith:WGTX7CAF submitted 2021-06-17 cs.CL

pysentimiento: A Python Toolkit for Opinion Mining and Social NLP tasks

classification cs.CL
keywords socialtaskspythoncomprehensiveenglishissueslanguageslibrary
verification ladder T0 review T1 audit T2 compute T3 formal T4 reserved
0 comments
read the original abstract

In recent years, the extraction of opinions and information from user-generated text has attracted a lot of interest, largely due to the unprecedented volume of content in Social Media. However, social researchers face some issues in adopting cutting-edge tools for these tasks, as they are usually behind commercial APIs, unavailable for other languages than English, or very complex to use for non-experts. To address these issues, we present pysentimiento, a comprehensive multilingual Python toolkit designed for opinion mining and other Social NLP tasks. This open-source library brings state-of-the-art models for Spanish, English, Italian, and Portuguese in an easy-to-use Python library, allowing researchers to leverage these techniques. We present a comprehensive assessment of performance for several pre-trained language models across a variety of tasks, languages, and datasets, including an evaluation of fairness in the results.

discussion (0)

Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

  1. A Durability and Cross-Language Transfer Benchmark for a Validated Teaching-Feedback Classification Protocol

    cs.CL 2026-07 accept novelty 5.0

    The teaching-feedback classification protocol remains durable across three representation generations and transfers to English sentiment, so model choice is a deployment decision.

  2. psytechlab at CLPsych 2026: Utilising Natural Language Processing methods and Large Language Models for Social Media Text Analysis

    cs.CL 2026-07 conditional novelty 3.0

    A standard NLP/LLM pipeline for CLPsych 2026 self-state, change-detection, and timeline-summarization tasks reached top-tier consistency metrics on summaries and mid-level ranks elsewhere.