Pith. sign in

REVIEW 1 cited by

XLM-T: Multilingual Language Models in Twitter for Sentiment Analysis and Beyond

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2104.12250 v2 pith:CSJGJ2OB submitted 2021-04-25 cs.CL

classification cs.CL
keywords multilinguallanguagemodelmodelstwitterxlm-tanalysiscurrent
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Language models are ubiquitous in current NLP, and their multilingual capacity has recently attracted considerable attention. However, current analyses have almost exclusively focused on (multilingual variants of) standard benchmarks, and have relied on clean pre-training and task-specific corpora as multilingual signals. In this paper, we introduce XLM-T, a model to train and evaluate multilingual language models in Twitter. In this paper we provide: (1) a new strong multilingual baseline consisting of an XLM-R (Conneau et al. 2020) model pre-trained on millions of tweets in over thirty languages, alongside starter code to subsequently fine-tune on a target task; and (2) a set of unified sentiment analysis Twitter datasets in eight different languages and a XLM-T model fine-tuned on them.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Speech Emotion Recognition via Entropy-Aware Score Selection

    cs.SD 2025-08 conditional novelty 4.0 of 10

    Entropy and varentropy thresholds on a wav2vec2 emotion model trigger a fallback to Whisper plus RoBERTa sentiment, yielding small average F1 gains on IEMOCAP and MSP-IMPROV.

Pith tools