Bridging the Knowledge-Prediction Gap in LLMs on Multiple-Choice Questions

Haesung Pyun; Yohan Jo; Yoonah Park

arxiv: 2509.23782 · v4 · pith:NXSVIU5Fnew · submitted 2025-09-28 · 💻 cs.CL

Bridging the Knowledge-Prediction Gap in LLMs on Multiple-Choice Questions

Yoonah Park , Haesung Pyun , Yohan Jo This is my paper

classification 💻 cs.CL

keywords knowledge-predictionllmsknowledgemodelsacrossbehaviordiversegeometric

0 comments

read the original abstract

While large language models (LLMs) perform strongly on diverse tasks, their trustworthiness is limited by erratic behavior that is unfaithful to their internal knowledge. In particular, LLMs often fail on multiple-choice questions (MCQs) even if they encode correct answers in their hidden representations, revealing a misalignment between internal knowledge and output behavior. We investigate and mitigate this knowledge-prediction gap on MCQs through a three-step analysis of hidden representations. First, we quantify the prevalence and magnitude of the gap across models and datasets. Second, we provide a geometric interpretation by identifying distinct knowledge and prediction subspaces in the residual stream. Third, we introduce KAPPA, a lightweight inference-time intervention that aligns the two subspaces within the residual stream to reduce the knowledge-prediction gap. Our results provide a geometric and interpretable explanation of the knowledge-prediction gap in LLMs. Furthermore, KAPPA effectively reduces the gap across diverse MCQ benchmarks and models, and generalizes to free-form settings.

This paper has not been read by Pith yet.

discussion (0)

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

Which Way Did It Move? Diagnosing and Overcoming Directional Motion Blindness in Video-LLMs
cs.CV 2026-05 conditional novelty 7.0

Video-LLMs exhibit directional motion blindness from a direction binding gap; DeltaDirect projector objective lifts synthetic accuracy to 85.4% and real accuracy by 21.9 points while preserving other video capabilities.