Pith. sign in

REVIEW 1 cited by

Neutralizing Bias in LLM Reasoning using Entailment Graphs

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2503.11614 v1 pith:WUB5KNEY submitted 2025-03-14 cs.CL

classification cs.CL
keywords biasllmsframeworkattestationdatasetsoriginalreasoningbias-neutralized
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

LLMs are often claimed to be capable of Natural Language Inference (NLI), which is widely regarded as a cornerstone of more complex forms of reasoning. However, recent works show that LLMs still suffer from hallucinations in NLI due to attestation bias, where LLMs overly rely on propositional memory to build shortcuts. To solve the issue, we design an unsupervised framework to construct counterfactual reasoning data and fine-tune LLMs to reduce attestation bias. To measure bias reduction, we build bias-adversarial variants of NLI datasets with randomly replaced predicates in premises while keeping hypotheses unchanged. Extensive evaluations show that our framework can significantly reduce hallucinations from attestation bias. Then, we further evaluate LLMs fine-tuned with our framework on original NLI datasets and their bias-neutralized versions, where original entities are replaced with randomly sampled ones. Extensive results show that our framework consistently improves inferential performance on both original and bias-neutralized NLI datasets.

Discussion (0). Sign in to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. LLMs are Frequency Pattern Learners in Natural Language Inference

    cs.CL 2025-05 conditional novelty 6.0 of 10

    Predicate frequency is systematically biased in NLI entailment data, fine-tuned LLMs increasingly rely on this frequency cue, and the cue correlates with WordNet hypernym frequency.

Pith tools