Pith. sign in

REVIEW 7 cited by

FakeGPT: Fake News Generation, Explanation and Detection of Large Language Models

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2310.05046 v2 pith:BDS672RL submitted 2023-10-08 cs.CL

classification cs.CL
keywords fakenewschatgptdetectingdetectionlanguageexplanationgeneration
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

The rampant spread of fake news has adversely affected society, resulting in extensive research on curbing its spread. As a notable milestone in large language models (LLMs), ChatGPT has gained significant attention due to its exceptional natural language processing capabilities. In this study, we present a thorough exploration of ChatGPT's proficiency in generating, explaining, and detecting fake news as follows. Generation -- We employ four prompt methods to generate fake news samples and prove the high quality of these samples through both self-assessment and human evaluation. Explanation -- We obtain nine features to characterize fake news based on ChatGPT's explanations and analyze the distribution of these factors across multiple public datasets. Detection -- We examine ChatGPT's capacity to identify fake news. We explore its detection consistency and then propose a reason-aware prompt method to improve its performance. Although our experiments demonstrate that ChatGPT shows commendable performance in detecting fake news, there is still room for its improvement. Consequently, we further probe into the potential extra information that could bolster its effectiveness in detecting fake news.

Discussion (0). Sign in to comment.

Forward citations

Cited by 7 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Selective Disclosure Watermarking for Large Language Models

    cs.CR 2026-07 accept novelty 7.0 of 10

    HeRo recursively partitions the LLM vocabulary into a hierarchy, embedding multi-bit payloads across layers so that verifiers with different keys recover only their authorized portion while preserving the original sam...

  2. Tailored untruths: How personalisation challenges LLM safeguards

    cs.CL 2025-10 conditional novelty 7.0 of 10

    A 1.6-million-text study of eight LLMs in four languages finds that adding demographic personae to disinformation prompts raises jailbreak rates from 78% to 82%.

  3. Personalized Large Language Models Can Increase the Belief Accuracy of Social Networks

    cs.SI 2025-06 conditional novelty 7.0 of 10

    A pre-registered experiment finds that adding a personalized, factually grounded LLM to an online discussion moves individuals' beliefs toward the truth and makes them build more accurate social networks.

  4. GPE: Evaluating Robust Evidence Aggregation for Fact Verification under Controllable GEO-Style Poisoning

    cs.CR 2026-07 conditional novelty 6.0 of 10

    GPE, a new benchmark with controllable GEO-style poisoning, shows LLM fact verifiers degrade sharply under poisoned evidence, with no single verifier winning across all attack types.

  5. Chain-of-Models: Cross-Model Auditing for Bias-Robust LLM Judges

    cs.CL 2026-05 conditional novelty 6.0 of 10

    Per-bias selection of a cross-family LLM auditor lifts biased-judgment accuracy from 0.805/0.824 baselines to 0.884.

  6. Leveraging Knowledge Graphs and LLMs for Structured Generation of Misinformation

    cs.AI 2025-05 conditional novelty 5.0 of 10

    A knowledge graph-guided LLM pipeline generates misinformation by replacing objects in true triplets with structurally similar ones, and LLM judges often fail to flag the fakes as fake.

  7. Multimodal rumor detection enhanced by external evidence and forgery features

    cs.LG 2026-01 conditional novelty 4.0 of 10

    A combination of Fourier forgery features, BLIP captions, evidence attention, and gated fusion reports 94.9% macro accuracy on Weibo and 94.1% on Twitter for rumor detection on MR2, but the closest baseline is omitted...

Pith tools