Pith. sign in

REVIEW 2 cited by

Transformer Models in Education: Summarizing Science Textbooks with AraBART, MT5, AraT5, and mBART

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2406.07692 v1 pith:PISMKSOS submitted 2024-06-11 cs.CL cs.ET

classification cs.CLcs.ET
keywords arabicmodelstexttextbookstextsarabartarat5content
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Recently, with the rapid development in the fields of technology and the increasing amount of text t available on the internet, it has become urgent to develop effective tools for processing and understanding texts in a way that summaries the content without losing the fundamental essence of the information. Given this challenge, we have developed an advanced text summarization system targeting Arabic textbooks. Relying on modern natu-ral language processing models such as MT5, AraBART, AraT5, and mBART50, this system evaluates and extracts the most important sentences found in biology textbooks for the 11th and 12th grades in the Palestinian curriculum, which enables students and teachers to obtain accurate and useful summaries that help them easily understand the content. We utilized the Rouge metric to evaluate the performance of the trained models. Moreover, experts in education Edu textbook authoring assess the output of the trained models. This approach aims to identify the best solutions and clarify areas needing improvement. This research provides a solution for summarizing Arabic text. It enriches the field by offering results that can open new horizons for research and development in the technologies for understanding and generating the Arabic language. Additionally, it contributes to the field with Arabic texts through creating and compiling schoolbook texts and building a dataset.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Detecting AI-Generated Text in Educational Content: Leveraging Machine Learning and Explainable AI for Academic Integrity

    cs.CL 2025-01 conditional novelty 4.0 of 10

    XGBoost and Random Forest can distinguish ChatGPT-written cybersecurity paragraphs from human Wikipedia paragraphs with 81 to 83% accuracy, and a narrow XGBoost model beat GPTZero in a three-class test, yet the benchm...

  2. Leveraging Explainable AI for LLM Text Attribution: Differentiating Human-Written and Multiple LLMs-Generated Text

    cs.CL 2025-01 reject novelty 3.0 of 10

    On a 600-essay, two-topic dataset, TF-IDF features let a Random Forest identify which of five LLMs or a human wrote a text with about 97% accuracy, though the comparison to GPTZero is inconsistent.

Pith tools