Pith. sign in

REVIEW 1 cited by

Free-text Rationale Generation under Readability Level Control

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2407.01384 v3 pith:N4TTA2XA submitted 2024-07-01 cs.CL

classification cs.CL
keywords readabilitylevelrationalescomplexitycontrolexplanationfree-textgeneration
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Free-text rationales justify model decisions in natural language and thus become likable and accessible among approaches to explanation across many tasks. However, their effectiveness can be hindered by misinterpretation and hallucination. As a perturbation test, we investigate how large language models (LLMs) perform rationale generation under the effects of readability level control, i.e., being prompted for an explanation targeting a specific expertise level, such as sixth grade or college. We find that explanations are adaptable to such instruction, though the observed distinction between readability levels does not fully match the defined complexity scores according to traditional readability metrics. Furthermore, the generated rationales tend to feature medium level complexity, which correlates with the measured quality using automatic metrics. Finally, our human annotators confirm a generally satisfactory impression on rationales at all readability levels, with high-school-level readability being most commonly perceived and favored.

Discussion (0). Sign in to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. ELI-Why: Evaluating the Pedagogical Utility of Language Model Explanations

    cs.CL 2025-06 conditional novelty 6.0 of 10

    ELI-Why shows GPT-4's grade-tailored explanations often miss the intended educational level and are less informative than human-curated explanations.

Pith tools