Pith. sign in

REVIEW 1 cited by

Automatic Detection of Vague Words and Sentences in Privacy Policies

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1808.06219 v2 pith:KJWRFKBA submitted 2018-08-19 cs.CL

classification cs.CL
keywords policiesprivacyvaguevaguenesswordsautomaticcontentdetection
verification ladder T0 review T1 audit T2 compute T3 formal

Signed reviews

No signed human review yet.

0 comments
read the original abstract

Website privacy policies represent the single most important source of information for users to gauge how their personal data are collected, used and shared by companies. However, privacy policies are often vague and people struggle to understand the content. Their opaqueness poses a significant challenge to both users and policy regulators. In this paper, we seek to identify vague content in privacy policies. We construct the first corpus of human-annotated vague words and sentences and present empirical studies on automatic vagueness detection. In particular, we investigate context-aware and context-agnostic models for predicting vague words, and explore auxiliary-classifier generative adversarial networks for characterizing sentence vagueness. Our experimental results demonstrate the effectiveness of proposed approaches. Finally, we provide suggestions for resolving vagueness and improving the usability of privacy policies.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Automating Conflict-Aware ACL Configurations with Natural Language Intents

    cs.NI 2025-08 conditional novelty 6.0 of 10

    Xumi automates the full ACL configuration pipeline from natural language intents, reporting 90-98% rule translation accuracy, 3.33x more accurate conflict detection than overlap baselines, and about 40% fewer rule add...

Pith tools