REVIEW 11 cited by
AraBERT: Transformer-based Model for Arabic Language Understanding
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
The Arabic language is a morphologically rich language with relatively few resources and a less explored syntax compared to English. Given these limitations, Arabic Natural Language Processing (NLP) tasks like Sentiment Analysis (SA), Named Entity Recognition (NER), and Question Answering (QA), have proven to be very challenging to tackle. Recently, with the surge of transformers based models, language-specific BERT based models have proven to be very efficient at language understanding, provided they are pre-trained on a very large corpus. Such models were able to set new standards and achieve state-of-the-art results for most NLP tasks. In this paper, we pre-trained BERT specifically for the Arabic language in the pursuit of achieving the same success that BERT did for the English language. The performance of AraBERT is compared to multilingual BERT from Google and other state-of-the-art approaches. The results showed that the newly developed AraBERT achieved state-of-the-art performance on most tested Arabic NLP tasks. The pretrained araBERT models are publicly available on https://github.com/aub-mind/arabert hoping to encourage research and applications for Arabic NLP.
Forward citations
Cited by 11 Pith papers
-
Fann or Flop: A Multigenre, Multiera Benchmark for Arabic Poetry Understanding in LLMs
The new Fann or Flop benchmark measures LLM comprehension of Arabic poetry through expert-written verse explanations and shows current LLMs perform poorly on interpretive depth.
-
Linking Hadith Narrator Identities Across Heterogeneous Arabic Biographical Databases: A Multi-Signal Entity Resolution Pipeline
A name-only then multi-signal entity-resolution pipeline links Sanadset narrators to Hawramani and Muslimscholars, releasing 94k+95k stratified links and a 185k-node transmission graph.
-
Evaluation of Adversarial Robustness in Arabic Language Models
Arabic BERT-family sentiment models lose up to 92% accuracy under diacritics and 58% under conjunction attacks; paraphrase attacks cut accuracy by 76% on average, and adversarial training only partially helps.
-
A Signer-Invariant Conformer and Multi-Scale Fusion Transformer for Continuous Sign Language Recognition
The paper reports state-of-the-art WERs of 13.07% (signer-independent) and 47.78% (unseen sentences) on Isharah-1000 using a conformer and a multi-scale fusion transformer.
-
Optimizing RAG Pipelines for Arabic: A Systematic Analysis of Core Components
For Arabic retrieval-augmented generation, sentence-aware chunking, BGE-M3 and Multilingual-E5-large embeddings, bge-reranker-v2-m3, and Aya-8B yield the highest RAGAS scores across six Arabic datasets.
-
CVPD at QIAS 2025 Shared Task: An Efficient Encoder-Based Approach for Islamic Inheritance Reasoning
An encoder-based relevance-scoring system achieves 69.87% accuracy on Islamic inheritance multiple-choice questions, below Gemini's 87.60% but with far smaller compute.
-
Enhanced Arabic Text Retrieval with Attentive Relevance Scoring
An Arabic dense retriever using a trainable attentive scoring module instead of dot-product similarity reports improved top-k passage retrieval on ArabicaQA.
-
The Role of Orthographic Consistency in Multilingual Embedding Models for Text Classification in Arabic-Script Languages
Language-specific RoBERTa models for four Arabic-script languages beat multilingual baselines on news classification, though the claimed orthographic-consistency mechanism is not demonstrated.
-
Multi-task Learning with Active Learning for Arabic Offensive Speech Detection
A multi-task Arabic offensive speech detector with entropy-based active learning and weighted emoji tokens reports 85.42% macro F1 on OSACT2022 using roughly 3,300 training samples.
-
GATE: General Arabic Text Embedding for Enhanced Semantic Textual Similarity with Matryoshka Representation Learning and Hybrid Loss Training
GATE's Arabic-Triplet-Matryoshka-V2 reports the highest average scores on the MTEB Arabic STS17/STS22/STS22-v2 tasks among the models compared in the paper.
-
DS@GT ARC at CheckThat! 2026: LLM-Based Trace Ranking and Grouped Reward Modeling for Multilingual Numerical Claim Verification
LoRA-tuned LLM trace scoring beats a TF-IDF grouped reward model on most English metrics and Recall@5, AraBERT beats multilingual BERT on Arabic, and prompt-based sub-claim decomposition hurts rather than helps.
Discussion (0). Sign in to comment.