REVIEW 2 cited by
BERTweet: A pre-trained language model for English Tweets
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
We present BERTweet, the first public large-scale pre-trained language model for English Tweets. Our BERTweet, having the same architecture as BERT-base (Devlin et al., 2019), is trained using the RoBERTa pre-training procedure (Liu et al., 2019). Experiments show that BERTweet outperforms strong baselines RoBERTa-base and XLM-R-base (Conneau et al., 2020), producing better performance results than the previous state-of-the-art models on three Tweet NLP tasks: Part-of-speech tagging, Named-entity recognition and text classification. We release BERTweet under the MIT License to facilitate future research and applications on Tweet data. Our BERTweet is available at https://github.com/VinAIResearch/BERTweet
Forward citations
Cited by 2 Pith papers
-
Psychology-driven LLM Agents for Explainable Panic Prediction on Social Media during Sudden Disaster Events
PsychoAgent claims to predict individual panic during disasters by simulating psychological chains with LLMs, but its evaluation is weakened by selective screening and a circular BERT verification loop.
-
Involvement drives complexity of language in online debates
Influential Twitter users who are more partisan, negative, or offensive tend to use more lexically complex language, but the causal claim that involvement drives complexity is not supported.
Discussion (0). Continue with ORCID to comment.