Pith. sign in

REVIEW 2 cited by

ALISON: Fast and Effective Stylometric Authorship Obfuscation

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2402.00835 v1 pith:22VSENYL submitted 2024-02-01 cs.CL cs.AIcs.LG

classification cs.CLcs.AIcs.LG
keywords obfuscationmethodsauthorshipalisonsotatextauthorbetter
verification ladder T0 review T1 audit T2 compute T3 formal

Signed reviews

No signed human review yet.

0 comments
read the original abstract

Authorship Attribution (AA) and Authorship Obfuscation (AO) are two competing tasks of increasing importance in privacy research. Modern AA leverages an author's consistent writing style to match a text to its author using an AA classifier. AO is the corresponding adversarial task, aiming to modify a text in such a way that its semantics are preserved, yet an AA model cannot correctly infer its authorship. To address privacy concerns raised by state-of-the-art (SOTA) AA methods, new AO methods have been proposed but remain largely impractical to use due to their prohibitively slow training and obfuscation speed, often taking hours. To this challenge, we propose a practical AO method, ALISON, that (1) dramatically reduces training/obfuscation time, demonstrating more than 10x faster obfuscation than SOTA AO methods, (2) achieves better obfuscation success through attacking three transformer-based AA methods on two benchmark datasets, typically performing 15% better than competing methods, (3) does not require direct signals from a target AA classifier during obfuscation, and (4) utilizes unique stylometric features, allowing sound model interpretation for explainable obfuscation. We also demonstrate that ALISON can effectively prevent four SOTA AA methods from accurately determining the authorship of ChatGPT-generated texts, all while minimally changing the original text semantics. To ensure the reproducibility of our findings, our code and data are available at: https://github.com/EricX003/ALISON.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Personalized Author Obfuscation with Large Language Models

    cs.CL 2025-05 conditional novelty 5.0 of 10

    LLM paraphrasing obfuscates authorship unevenly across users, and prompting with each author's top SHAP-identified style feature improves average evasion but does not consistently beat zero-shot paraphrasing.

  2. Unveiling Unicode's Unseen Underpinnings in Undermining Authorship Attribution

    cs.CR 2025-08 unverdicted novelty 4.0 of 10

    The paper proposes integrating Unicode steganography into adversarial stylometry to undermine authorship attribution.

Pith tools