Pith. sign in

REVIEW 2 cited by

Redefining Simplicity: Benchmarking Large Language Models from Lexical to Document Simplification

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2502.08281 v1 pith:3WPOBEU5 submitted 2025-02-12 cs.CL

Redefining Simplicity: Benchmarking Large Language Models from Lexical to Document Simplification

classification cs.CL
keywords llmssimplificationdocumentexistingfourlanguagelargelexical
verification ladder T0 review T1 audit T2 compute T3 formal T4 reserved
0 comments
Share X Bluesky LinkedIn Reddit HN
read the original abstract

Text simplification (TS) refers to the process of reducing the complexity of a text while retaining its original meaning and key information. Existing work only shows that large language models (LLMs) have outperformed supervised non-LLM-based methods on sentence simplification. This study offers the first comprehensive analysis of LLM performance across four TS tasks: lexical, syntactic, sentence, and document simplification. We compare lightweight, closed-source and open-source LLMs against traditional non-LLM methods using automatic metrics and human evaluations. Our experiments reveal that LLMs not only outperform non-LLM approaches in all four tasks but also often generate outputs that exceed the quality of existing human-annotated references. Finally, we present some future directions of TS in the era of LLMs.

discussion (0)

Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

  1. A Human-in-the-Loop Corpus for LLM-Based Simplification of Scientific Summaries

    cs.CL 2026-07 accept novelty 5.0

    A human-in-the-loop corpus of scientific-summary simplifications with original, GPT-simplified, reader-annotated, and expert-edited versions for training and benchmarking simplification systems.

  2. Human--LLM Collaboration Is Transforming Complexity Metrics in Scientific Texts

    cs.CY 2026-06 unverdicted novelty 5.0

    Analysis of arXiv abstracts detects increased top-word turnover and flattening of LLM-style to complexity-metric relationships after 2022.