Pith. sign in

REVIEW 1 cited by

Interpretable Unified Language Checking

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2304.03728 v1 pith:WIUQIGE2 submitted 2023-04-07 cs.CL

classification cs.CL
keywords languagecheckingllmstasksunifieddetectionfact-checkingfind
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Despite recent concerns about undesirable behaviors generated by large language models (LLMs), including non-factual, biased, and hateful language, we find LLMs are inherent multi-task language checkers based on their latent representations of natural and social knowledge. We present an interpretable, unified, language checking (UniLC) method for both human and machine-generated language that aims to check if language input is factual and fair. While fairness and fact-checking tasks have been handled separately with dedicated models, we find that LLMs can achieve high performance on a combination of fact-checking, stereotype detection, and hate speech detection tasks with a simple, few-shot, unified set of prompts. With the ``1/2-shot'' multi-task language checking method proposed in this work, the GPT3.5-turbo model outperforms fully supervised baselines on several language tasks. The simple approach and results suggest that based on strong latent knowledge representations, an LLM can be an adaptive and explainable tool for detecting misinformation, stereotypes, and hate speech.

Discussion (0). Sign in to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. OpenAlex reports about 8 citations worldwide. Full citation record

  1. Retrieval-Augmented Recommendation Explanation Generation with Hierarchical Aggregation

    cs.IR 2025-07 conditional novelty 5.0 of 10

    REXHA improves recommendation explanation generation by hierarchically summarizing all reviews into user and item profiles and using pseudo-document queries to retrieve relevant reviews for a language model.

Pith tools