Pith. sign in

REVIEW 1 cited by

Towards Context-Robust LLMs: A Gated Representation Fine-tuning Approach

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2502.14100 v2 pith:Z3NOR3EM submitted 2025-02-19 cs.CL cs.IR

classification cs.CLcs.IR
keywords externalllmscontext-robustknowledgecontextsgrftinternalrepresentation
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Large Language Models (LLMs) enhanced with external contexts, such as through retrieval-augmented generation (RAG), often face challenges in handling imperfect evidence. They tend to over-rely on external knowledge, making them vulnerable to misleading and unhelpful contexts. To address this, we propose the concept of context-robust LLMs, which can effectively balance internal knowledge with external context, similar to human cognitive processes. Specifically, context-robust LLMs should rely on external context only when lacking internal knowledge, identify contradictions between internal and external knowledge, and disregard unhelpful contexts. To achieve this goal, we introduce Grft, a lightweight and plug-and-play gated representation fine-tuning approach. Grft consists of two key components: a gating mechanism to detect and filter problematic inputs, and low-rank representation adapters to adjust hidden representations. By training a lightweight intervention function with only 0.0004\% of model size on fewer than 200 examples, Grft can effectively adapt LLMs towards context-robust behaviors.

Discussion (0). Sign in to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

  1. Magnifying What Matters: Attention-Guided Adaptive Rendering for Visual Text Comprehension

    cs.CV 2026-06 unverdicted novelty 6.0 of 10

    AGAR uses middle-to-late layer attention in VLMs to identify and enlarge important word spans in rendered text images, improving performance on visual text comprehension benchmarks.

Pith tools