REVIEW 4 cited by
Sentence Simplification via Large Language Models
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Sentence Simplification aims to rephrase complex sentences into simpler sentences while retaining original meaning. Large Language models (LLMs) have demonstrated the ability to perform a variety of natural language processing tasks. However, it is not yet known whether LLMs can be served as a high-quality sentence simplification system. In this work, we empirically analyze the zero-/few-shot learning ability of LLMs by evaluating them on a number of benchmark test sets. Experimental results show LLMs outperform state-of-the-art sentence simplification methods, and are judged to be on a par with human annotators.
Forward citations
Cited by 4 Pith papers
-
A Hybrid Multi-Agent Prompting Approach for Simplifying Complex Sentences
A multi-agent GPT-4O pipeline with an internal semantic-lexical gate claims 70% success on simplifying 100 video game sentences, versus 48% for a single-agent version.
-
Automated Feedback Loops to Protect Text Simplification with Generative AI from Information Loss
Adding all missing named entities back into simplified biomedical text yields the highest cosine similarity and ROUGE-1 to the original among five insertion strategies, but the evaluation is partly circular and lacks ...
-
A Practical Guide for Supporting Formative Assessment and Feedback Using Generative AI
A narrative review that aligns generative AI tools with formative assessment principles, provides classroom prompt examples, and identifies missing evaluation metrics for AI feedback.
-
Redefining Simplicity: Benchmarking Large Language Models from Lexical to Document Simplification
In a four-task benchmark, GPT-4o, Llama3.1-70B, and Gemma2-2B outperform traditional text simplification systems on most automatic metrics, and GPT-4o is preferred over human-written references in a small human study.
Discussion (0). Continue with ORCID to comment.