Pith. sign in

Paper Citation Record · LEDGER

Why Language Models Hallucinate

As of 6 August 2026, this Paper Citation Record lists 2 of 2 outbound references and 100 inbound Pith citation observations for arXiv:2509.04664.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.04664 v1

Coverage vector

measured 2 of 2 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-07-11T11:50:26.030339Z

measured 102 of 102 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00

measured 100 of 111 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-04T13:32:26.312020Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T02:28:24.338817Z

Reference resolution

2 of 2 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch2

External citation measurements

20
pith, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation dda2cde0-5033-4449-b456-5b4ba4af8fb1 · outbound

This paper cites Constitutional AI: Harmlessness from AI Feedback.

Why Language Models Hallucinate Constitutional AI: Harmlessness from AI Feedback

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-05-13T12:32:40.679198Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-07-11T11:50:26.030339Z digest=sha256:4d0acfa5bcad560f6649e06b585263e4c778f8fb5df2be4c906843e4fb9b59a5

Observation 0245e22f-d452-4f46-92c3-dfa57097f318 · outbound

This paper cites Language Models (Mostly) Know What They Know.

Why Language Models Hallucinate Language Models (Mostly) Know What They Know

Reference 2

Resolution
metadata mismatch
local_arxiv, observed 2026-05-13T12:32:40.682545Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-07-11T11:50:26.030339Z digest=sha256:332b1944b54094b534707058202957a5e2460dbade849772176fa0db3c6bea9e

Pith citing papers

Observation 36aa315f-0fd1-41f1-8a9c-7b4f5568d517 · inbound

Semantic Concurrency Limits in Large Language Models cites this paper.

Semantic Concurrency Limits in Large Language Models Why Language Models Hallucinate

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-05-22T18:41:56.351188Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-22T18:41:30.008331Z digest=sha256:ec28b76ee90b41061f93302e71bcb9000e8a059745f3525c6fb312527bf9aca4

Observation 5640f15c-4859-451d-88c8-06916b227f99 · inbound

Variational Visual Question Answering for Uncertainty-Aware Selective Prediction cites this paper.

Variational Visual Question Answering for Uncertainty-Aware Selective Prediction Why Language Models Hallucinate

Reference 3

Resolution
metadata mismatch
local_arxiv, observed 2026-05-22T15:31:44.795787Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-22T15:29:59.372173Z digest=sha256:f17ffe59d462deb40fa4d30fbbbdba2ab8d82f9ccac8959f6856161b20ed0ba7

Observation fb66e9c2-eb7f-4207-8fc7-374b1cf67c70 · inbound

Vibe Coding in Product Teams: Reconfiguring AI-Assisted Workflows, Prototyping, and Collaboration cites this paper.

Vibe Coding in Product Teams: Reconfiguring AI-Assisted Workflows, Prototyping, and Collaboration Why Language Models Hallucinate

Reference 50

Resolution
metadata mismatch
local_arxiv, observed 2026-05-18T17:11:40.505825Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-18T17:07:16.400481Z digest=sha256:375ad3d42e2949d7e6f25220a0122d6ff54713428a4979014606796fa06597db

Observation 46e4ab86-ec51-4d14-8af8-5d6761f3e040 · inbound

RECAP: Transparent Inference-Time Emotion Alignment for Medical Dialogue Systems cites this paper.

RECAP: Transparent Inference-Time Emotion Alignment for Medical Dialogue Systems Why Language Models Hallucinate

Reference 28

Resolution
metadata mismatch
local_arxiv, observed 2026-05-18T16:52:43.974917Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-18T16:52:22.768827Z digest=sha256:fa7979825a3f58c1e23ca3b3bca7aefd09374d97eac7dea14e99cc3b31cfe6b5

Observation 73af856e-982c-4259-b64b-4c98978c56c5 · inbound

The Prompt Engineering Report Distilled: Quick Start Guide for Life Sciences cites this paper.

The Prompt Engineering Report Distilled: Quick Start Guide for Life Sciences Why Language Models Hallucinate

Reference 72

Resolution
verified exact
local_arxiv, observed 2026-05-18T16:41:37.978703Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-18T16:39:03.794436Z digest=sha256:18644f33f48c3c58d469e433ae48263c4b501ff58d59c5e5a9e8cd410116d308

Observation f4f95fe8-ce2f-47db-81ee-e915c95e8fa5 · inbound

The Prompt Engineering Report Distilled: Quick Start Guide for Life Sciences cites this paper.

The Prompt Engineering Report Distilled: Quick Start Guide for Life Sciences Why Language Models Hallucinate

Reference 73

Resolution
metadata mismatch
local_arxiv, observed 2026-05-18T16:41:37.636651Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-18T16:39:03.794436Z digest=sha256:bbdbd1a88eb544c6387efd8b599ea736009739a13925ecaecd179b6c92209f53

Observation cc967bd0-86f4-4560-a7cd-bfaa3e9bc8f7 · inbound

Probing the Critical Point (CritPt) of AI Reasoning: a Frontier Physics Research Benchmark cites this paper.

Probing the Critical Point (CritPt) of AI Reasoning: a Frontier Physics Research Benchmark Why Language Models Hallucinate

Reference 67

Resolution
verified exact
local_arxiv, observed 2026-05-18T11:52:35.421816Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-18T11:52:10.205796Z digest=sha256:f522b4abb4f5e89624c4b85ce9cb3a8f5702de37c558115b9539370ed4509ad0

Observation bfceeb23-1523-4c91-a1a6-7cb105bd69ad · inbound

Generalization of Fine-Tuned Uncertainty Communication and Metacognition in Large Language Models cites this paper.

Generalization of Fine-Tuned Uncertainty Communication and Metacognition in Large Language Models Why Language Models Hallucinate

Reference 642

Resolution
unresolved
no resolver link, observed 2026-08-04T13:32:26.312020Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T13:32:26.312020Z digest=sha256:3cc3bf2f2727b5b97778d45355f0c99e7fbcde82766df96af3e74db1cdaf90f3

Observation 194b700d-3788-4b64-8cef-dbf312ff86ce · inbound

Benchmarking foundation potentials against quantum chemistry methods for predicting molecular redox potentials cites this paper.

Benchmarking foundation potentials against quantum chemistry methods for predicting molecular redox potentials Why Language Models Hallucinate

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-04T07:54:35.249846Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T07:54:35.249846Z digest=sha256:8422396181c95da4358a3886e4ef3f3e8ed09ff404d466cbe30f6af7726d5fa5

Observation bf87f286-4dea-4e70-9c84-84687d80d38d · inbound

Enhancing Trustworthy GUI Grounding via Self-Critiqued Reinforcement Learning cites this paper.

Enhancing Trustworthy GUI Grounding via Self-Critiqued Reinforcement Learning Why Language Models Hallucinate

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-04T07:02:41.372851Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T07:02:41.372851Z digest=sha256:0e98cb98deb2ad75d6372ac2c2f5e7abbdf1f3175f3b42be3efcf6636ac350d9

Observation 66e5baa9-4563-4606-9a17-26b397623c65 · inbound

TeaRAG: A Token-Efficient Agentic Retrieval-Augmented Generation Framework cites this paper.

TeaRAG: A Token-Efficient Agentic Retrieval-Augmented Generation Framework Why Language Models Hallucinate

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-03T23:33:46.112856Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T23:33:46.112856Z digest=sha256:5a67c585e4e18338e13e2e3c61fba9566a7c468b4fdae0315155ac40c08cc7fc

Observation 47cdfd92-1307-4e6c-9224-a9c667bb8abf · inbound

Hierarchical Memorization in Large Language Models: Evidence from Citation Generation cites this paper.

Hierarchical Memorization in Large Language Models: Evidence from Citation Generation Why Language Models Hallucinate

Reference 29

Resolution
verified exact
local_arxiv, observed 2026-05-17T23:15:26.887258Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-17T23:12:12.874406Z digest=sha256:b1be31b97518afcd182b0bd1736555d45c154ca9dd70cd438fe628b64c630f7a

Observation 35d708e2-9531-402b-ac5f-0525fabfcce5 · inbound

Robust AI Security and Alignment: A Sisyphean Endeavor? cites this paper.

Robust AI Security and Alignment: A Sisyphean Endeavor? Why Language Models Hallucinate

Reference 5

Resolution
verified exact
local_arxiv, observed 2026-05-16T22:58:38.496206Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T22:58:29.306435Z digest=sha256:afeef3b03bc47db07dc7cda5d976abe83c58a05fdcc3a813635e953dc6e9cfdf

Observation 777168af-b0ad-4d72-adc3-8fb04f00b473 · inbound

Concept Tokens: Learning Behavioral Embeddings Through Concept Definitions cites this paper.

Concept Tokens: Learning Behavioral Embeddings Through Concept Definitions Why Language Models Hallucinate

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-03T12:04:58.514194Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T12:04:58.514194Z digest=sha256:2f8a2907ca6a5b2937ff01716ac34c437da62e2e7458f96229570dd792ee1036

Observation 088f51c6-ccc6-4d8c-84e6-73f40072e6f5 · inbound

Characterizing the Effect of Noise in Language Generation in the Limit cites this paper.

Characterizing the Effect of Noise in Language Generation in the Limit Why Language Models Hallucinate

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-03T07:16:27.957897Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T07:16:27.957897Z digest=sha256:4e00ed3d127b870c03438afa4a38275f9ca7ac697af5bbd04b967e917b3ab598

Observation 6adc291c-0df8-4ef8-b116-c208dc6c6f34 · inbound

UCPO: Uncertainty-Aware Policy Optimization cites this paper.

UCPO: Uncertainty-Aware Policy Optimization Why Language Models Hallucinate

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-03T06:34:28.912210Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T06:34:28.912210Z digest=sha256:2fa4556f018ea580623e9288066c15644e14a7d38abb8fee278e0c8aa4e97c13

Observation e2a3ce71-11dd-4a11-9d7d-e8459ca4b964 · inbound

Stop Rewarding Hallucinated Steps: Faithfulness-Aware Step-Level Reinforcement Learning for Small Reasoning Models cites this paper.

Stop Rewarding Hallucinated Steps: Faithfulness-Aware Step-Level Reinforcement Learning for Small Reasoning Models Why Language Models Hallucinate

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-08-03T04:09:48.609050Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T04:09:48.609050Z digest=sha256:d8e9d396a6965182ffbdcb8159fe66cde1daea08d178e5bd035d2263b3d8f590

Observation aefacbb8-fe85-48f0-850a-07d0674a5964 · inbound

Revis: Sparse Latent Steering to Mitigate Object Hallucination in Large Vision-Language Models cites this paper.

Revis: Sparse Latent Steering to Mitigate Object Hallucination in Large Vision-Language Models Why Language Models Hallucinate

Reference 2

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T05:07:20.452445Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-16T05:06:24.973439Z digest=sha256:96e4fcf01243a29ef1d0ebccea0e855890eda6fc150e17b6d957b4637f4a9def

Observation 63e738c7-908a-417f-a421-d3c84f0c127f · inbound

When Should LLMs Be Less Specific? Selective Abstraction for Reliable Long-Form Text Generation cites this paper.

When Should LLMs Be Less Specific? Selective Abstraction for Reliable Long-Form Text Generation Why Language Models Hallucinate

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-03T00:03:21.277478Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T00:03:21.277478Z digest=sha256:b30c9c379ea75a577f1b9f5315ea5f4b3a20290eabc17c9328a473f675722975

Observation 25dd45ff-0e22-4640-8838-016cad77a979 · inbound

Empty Shelves or Lost Keys? Recall Is the Bottleneck for Parametric Factuality cites this paper.

Empty Shelves or Lost Keys? Recall Is the Bottleneck for Parametric Factuality Why Language Models Hallucinate

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-02T23:23:46.484754Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T23:23:46.484754Z digest=sha256:088f53c85f3a6a569e97abfea7016f591ad69abb9bfa1f147f31cd7b39408ca0

Observation 6c4ddc9a-3121-455e-8594-eaeed519cf14 · inbound

Towards a Science of AI Agent Reliability cites this paper.

Towards a Science of AI Agent Reliability Why Language Models Hallucinate

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-02T22:31:58.319721Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:31:58.319721Z digest=sha256:50a76c68bb991edea9045b851db21ce35af287da0b7290b9881571fd58d3e47e

Observation 20a21ab7-11a8-4640-84c5-1d9208206532 · inbound

CausalFlip: A Benchmark for LLM Causal Judgment Beyond Semantic Matching cites this paper.

CausalFlip: A Benchmark for LLM Causal Judgment Beyond Semantic Matching Why Language Models Hallucinate

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-02T21:28:28.700868Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T21:28:28.700868Z digest=sha256:927bdb325705ea4ca64e4c06df6ac4c1d8d2538d83992b6a157a53351b84fca4

Observation a90cf74a-8082-42ad-b498-cb7a3e7cc943 · inbound

Beyond Explainable AI (XAI): An Overdue Paradigm Shift and Post-XAI Research Directions cites this paper.

Beyond Explainable AI (XAI): An Overdue Paradigm Shift and Post-XAI Research Directions Why Language Models Hallucinate

Reference 129

Resolution
verified exact
local_arxiv, observed 2026-05-15T18:51:30.328059Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-15T18:50:52.313363Z digest=sha256:d4076c58113c2ee7da1ec3a56368500823c9ef67dfa9ca47f7ac4b009c2ac4fc

Observation 8f8d1c23-c9ae-4817-9669-c481b3cef7bb · inbound

Beyond Explainable AI (XAI): An Overdue Paradigm Shift and Post-XAI Research Directions cites this paper.

Beyond Explainable AI (XAI): An Overdue Paradigm Shift and Post-XAI Research Directions Why Language Models Hallucinate

Reference 123

Resolution
unresolved
no resolver link, observed 2026-08-02T20:04:33.368573Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T20:04:33.368573Z digest=sha256:18910d31ea050eeaf034353371ab96d3aacebcda923db5fb1da31ffbe82c7564

Observation ee656f72-0fa7-4ce7-8667-eef4773b08a4 · inbound

Beyond Relevance: On the Relationship Between Retrieval and RAG Information Coverage cites this paper.

Beyond Relevance: On the Relationship Between Retrieval and RAG Information Coverage Why Language Models Hallucinate

Reference 27

Resolution
metadata mismatch
local_arxiv, observed 2026-05-15T13:10:49.404568Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-15T13:10:06.367994Z digest=sha256:d82ceb22e760d4188812ba23bb81fb09e05a9db6e6b2d3a2571f8460d664c0fe

Observation 5d7e8b59-5bab-49b0-bbef-242696dd434d · inbound

Toward Epistemic Stability: Engineering Consistent Procedures for Industrial LLM Hallucination Reduction cites this paper.

Toward Epistemic Stability: Engineering Consistent Procedures for Industrial LLM Hallucination Reduction Why Language Models Hallucinate

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-05-15T14:30:03.586482Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-07-11T11:50:26.030339Z digest=sha256:6298bb2676d83c472b963925aa6927756fb7228f6c651f806c5399823dbe7572

Observation 6d145a09-a089-4a57-b2ac-73c88462ef51 · inbound

Causal Evidence that Language Models use Confidence to Drive Behavior cites this paper.

Causal Evidence that Language Models use Confidence to Drive Behavior Why Language Models Hallucinate

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-05-21T09:44:05.686530Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T09:43:05.524088Z digest=sha256:c0ee2d567ab2de4ded9dab31a6ed9a2dba2711f3fe43c8761528ec4d41306b70

Observation 61049c1b-c8f3-4f10-8398-e6d1b4e94578 · inbound

Ensemble-Based Uncertainty Estimation for Code Correctness Estimation cites this paper.

Ensemble-Based Uncertainty Estimation for Code Correctness Estimation Why Language Models Hallucinate

Reference 17

Resolution
metadata mismatch
local_arxiv, observed 2026-05-14T22:53:14.151626Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-14T22:52:58.524934Z digest=sha256:e05ad5c86304c431adb7137d680c7631508a4f42752c6f6233fbb5009fec55aa

Observation 19e82e8d-c859-44bf-88f5-481322694786 · inbound

EnsemHalDet: Robust VLM Hallucination Detection via Ensemble of Internal State Detectors cites this paper.

EnsemHalDet: Robust VLM Hallucination Detection via Ensemble of Internal State Detectors Why Language Models Hallucinate

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-05-13T20:03:12.372323Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T20:00:44.105580Z digest=sha256:748f3907f5ebc26b5106fe7bcd02e8541838e52314903258bbf81570f63a9788

Observation 0f933586-7bd8-4745-a1ad-e688b68e6917 · inbound

EnsemHalDet: Robust VLM Hallucination Detection via Ensemble of Internal State Detectors cites this paper.

EnsemHalDet: Robust VLM Hallucination Detection via Ensemble of Internal State Detectors Why Language Models Hallucinate

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-05-21T10:30:00.345702Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-21T10:26:25.982873Z digest=sha256:1650180d074a48541d2989b1f249ee2a2578fa56e6639fb232ae267ad07b54ad

Observation 1d672977-662b-4560-8d11-3ea6521f575b · inbound

Enhancing Multi-Robot Exploration Using Probabilistic Frontier Prioritization with Dirichlet Process Gaussian Mixtures cites this paper.

Enhancing Multi-Robot Exploration Using Probabilistic Frontier Prioritization with Dirichlet Process Gaussian Mixtures Why Language Models Hallucinate

Reference 19

Resolution
unresolved
no resolver link, observed 2026-07-13T13:35:02.428552Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T13:35:02.428552Z digest=sha256:bb5e81b3f333f66bbaebd598894ab12535beea30971722d897c4222f996a75be

Observation 75698b77-0145-4d5f-83fd-3049f6352dc2 · inbound

STEAR: Layer-Aware Spatiotemporal Evidence Intervention for Hallucination Mitigation in Video Large Language Models cites this paper.

STEAR: Layer-Aware Spatiotemporal Evidence Intervention for Hallucination Mitigation in Video Large Language Models Why Language Models Hallucinate

Reference 19

Resolution
verified exact
local_arxiv, observed 2026-05-13T20:18:13.347911Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T20:16:04.999705Z digest=sha256:4af024494df05e60c9a450c89174e5afcec0bc9a62e85517bac1152035df812f

Observation 2fecfea5-6dc9-4b0d-a940-a60941b7b312 · inbound

BAS: A Decision-Theoretic Approach to Evaluating Large Language Model Confidence cites this paper.

BAS: A Decision-Theoretic Approach to Evaluating Large Language Model Confidence Why Language Models Hallucinate

Reference 27

Resolution
verified exact
local_arxiv, observed 2026-05-13T19:53:11.942625Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-13T19:48:13.133733Z digest=sha256:deb92de05d0989878ac782063ccd299f015fbf76abd061cc49de620188becf9d

Observation b1e14b01-0f09-44a9-b486-768ede65e865 · inbound

When Do Hallucinations Arise? A Graph Perspective on the Evolution of Path Reuse and Path Compression cites this paper.

When Do Hallucinations Arise? A Graph Perspective on the Evolution of Path Reuse and Path Compression Why Language Models Hallucinate

Reference 3

Resolution
unresolved
no resolver link, observed 2026-07-13T13:07:00.928770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T13:07:00.928770Z digest=sha256:855ea6049bffc58919d255a476b0b4d214b3ab0d02b1e37c9612f9cdd9c40b2a

Observation 6d467ede-55fe-475c-b9e9-f38a54842be4 · inbound

A Two-Stage LLM Framework for Accessible and Verified XAI Explanations cites this paper.

A Two-Stage LLM Framework for Accessible and Verified XAI Explanations Why Language Models Hallucinate

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-13T12:32:40.683342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T15:25:46.925378Z digest=sha256:e798964805b4183b5af0d0b7b8afe376397c0771b5880dfb59987b759350e199

Observation e8ef4672-fcd5-48c2-b3c7-d24d63411de9 · inbound

Calibration-Aware Policy Optimization for Reasoning LLMs cites this paper.

Calibration-Aware Policy Optimization for Reasoning LLMs Why Language Models Hallucinate

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-13T12:32:40.683342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-10T15:40:24.396851Z digest=sha256:9502812d8554fe17bf23f658d0d881d375e77bb310dfee805f801c7d47e241ee

Observation 86f8b163-2066-4927-b5d7-a39317c908d6 · inbound

InfiniteScienceGym: An Unbounded, Procedurally-Generated Benchmark for Scientific Analysis cites this paper.

InfiniteScienceGym: An Unbounded, Procedurally-Generated Benchmark for Scientific Analysis Why Language Models Hallucinate

Reference 13

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T12:32:40.683342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-10T15:40:48.770507Z digest=sha256:bb86df56b47a5b520358a33b156f09fc7e0855fb7bb667131b286733cc19d128

Observation 15892cf8-2a42-4089-8d3f-cd7c9dbdb4af · inbound

Beyond Literal Summarization: Redefining Hallucination for Medical SOAP Note Evaluation cites this paper.

Beyond Literal Summarization: Redefining Hallucination for Medical SOAP Note Evaluation Why Language Models Hallucinate

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-13T12:32:40.683342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T10:30:09.067724Z digest=sha256:710fa12765fa1a180d80a918c076f9e7f57e504893cd84ec9ecd9cd8cd07e64a

Observation 17440a10-fbc8-4cdf-b9d0-e7a668d5f31e · inbound

The Crutch or the Ceiling? How Different Generations of LLMs Shape EFL Student Writings cites this paper.

The Crutch or the Ceiling? How Different Generations of LLMs Shape EFL Student Writings Why Language Models Hallucinate

Reference 34

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T12:32:40.683342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T09:52:56.704142Z digest=sha256:3883b3ed4ebe166d8253bbbe5b62a099f7ac05b2112568669858de39b6d21d25

Observation cce669a0-b40e-436e-b3c7-4c2249f6ea6d · inbound

From Subsumption to Satisfiability: LLM-Assisted Active Learning for OWL Ontologies cites this paper.

From Subsumption to Satisfiability: LLM-Assisted Active Learning for OWL Ontologies Why Language Models Hallucinate

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-13T12:32:40.683342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T08:19:47.467705Z digest=sha256:b9b034f33cceb322e8748d7bb7bf116d54b77c29efb70eef897898d9330709bf

Observation 46175c00-2610-4a06-bf2a-78d88ee8c0b7 · inbound

HalluClear: Diagnosing, Evaluating and Mitigating Hallucinations in GUI Agents cites this paper.

HalluClear: Diagnosing, Evaluating and Mitigating Hallucinations in GUI Agents Why Language Models Hallucinate

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-13T12:32:40.683342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-10T06:27:19.717895Z digest=sha256:f690fb92d351b2b528dd70651676b86f9c7a9fedaa3b1da076ad8974c6ba3be4

Observation e5b8c8c8-b4d5-4e1a-aa61-89fb5db74a45 · inbound

ERA: Evidence-based Reliability Alignment for Honest Retrieval-Augmented Generation cites this paper.

ERA: Evidence-based Reliability Alignment for Honest Retrieval-Augmented Generation Why Language Models Hallucinate

Reference 19

Resolution
verified exact
local_arxiv, observed 2026-05-15T20:21:34.437836Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-15T20:20:49.769471Z digest=sha256:ae81f989ab6f02611e3ad84a264e98e9e54038bd95c5ea52312ceb41e8d144b8

Observation bce7a904-be84-43c7-9b1d-94fd64c54714 · inbound

Open-H-Embodiment: A Large-Scale Dataset for Enabling Foundation Models in Medical Robotics cites this paper.

Open-H-Embodiment: A Large-Scale Dataset for Enabling Foundation Models in Medical Robotics Why Language Models Hallucinate

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-07-05T02:10:36.422039Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-07-11T11:50:26.030339Z digest=sha256:4012ed8fc6122d58adb3bb30443cbc854c371eb9ebdd1a24a844eac188b9ad17

Observation 6cc75828-9141-4ee8-9eaa-e2312c73d9e5 · inbound

Adaptive Test-Time Compute Allocation with Evolving In-Context Demonstrations cites this paper.

Adaptive Test-Time Compute Allocation with Evolving In-Context Demonstrations Why Language Models Hallucinate

Reference 15

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T12:32:40.683342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-09T23:55:50.606359Z digest=sha256:be3e3e3f6e9448b715726a0055696af03aec7249ebeae9fea49f3afd3acc4494

Observation ce38668a-50ac-4cdb-bc70-758d102a26fa · inbound

Process Supervision of Confidence Margin for Calibrated LLM Reasoning cites this paper.

Process Supervision of Confidence Margin for Calibrated LLM Reasoning Why Language Models Hallucinate

Reference 32

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T12:32:40.683342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T08:19:09.437464Z digest=sha256:3eb7ea4f7b57bc1f1f1508aa18d3679e8c2eaed3a4242319827282518ddb0087

Observation c2d8ddc3-8051-491d-a4ce-c7c42cd9399e · inbound

Uncertainty Propagation in LLM-Based Systems cites this paper.

Uncertainty Propagation in LLM-Based Systems Why Language Models Hallucinate

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-13T12:32:40.683342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-08T06:10:07.638483Z digest=sha256:cbe99f973eaa76cbbf76bb2d452f9d459f49f1729d8c6b5fce3fa956e7373b76

Observation 7a19cab9-902d-43a5-8c02-14c45d2ce580 · inbound

SIEVES: Selective Prediction Generalizes through Visual Evidence Scoring cites this paper.

SIEVES: Selective Prediction Generalizes through Visual Evidence Scoring Why Language Models Hallucinate

Reference 25

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T12:32:40.683342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-07T16:48:20.100672Z digest=sha256:6d109fb896f6425cc48a5b1c9e6fc05462e2c1871ff9964a6c83ce1a0057b230

Observation 32ad1953-bf67-494c-adf9-eb4b0c977b48 · inbound

SIEVES: Selective Prediction Generalizes through Visual Evidence Scoring cites this paper.

SIEVES: Selective Prediction Generalizes through Visual Evidence Scoring Why Language Models Hallucinate

Reference 25

Resolution
metadata mismatch
local_arxiv, observed 2026-05-15T06:19:50.041458Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-15T06:15:55.968506Z digest=sha256:1bc7876da2eb7173c9e75e03ede687e68a62d1fab84b013d040e50b3d2de7314

Observation fddb10b1-ba40-4278-9822-40dfff80a858 · inbound

MCMit: Mid-Circuit Measurement Error Mitigation cites this paper.

MCMit: Mid-Circuit Measurement Error Mitigation Why Language Models Hallucinate

Reference 26

Resolution
verified exact
local_arxiv, observed 2026-07-01T08:45:35.103639Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-07-01T08:36:37.641628Z digest=sha256:edab2e777319e25c090cb5a520b1033f7495c077ca87186cd9774ce17006383d

Observation bfb3698b-e91c-481a-900e-ea9a2aa812ec · inbound

Ceci n'est pas une explication: Evaluating Explanation Failures as Explainability Pitfalls in Language Learning Systems cites this paper.

Ceci n'est pas une explication: Evaluating Explanation Failures as Explainability Pitfalls in Language Learning Systems Why Language Models Hallucinate

Reference 20

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T12:32:40.683342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-07T15:12:00.081798Z digest=sha256:ed5251cc862b2f3848883996adf07e15e322649f9c4a4e84315ed1a26f75fc73

Observation 0a893520-c3da-4251-b9ba-9159e7f6e298 · inbound

Ceci n'est pas une explication: Evaluating Explanation Failures as Explainability Pitfalls in Language Learning Systems cites this paper.

Ceci n'est pas une explication: Evaluating Explanation Failures as Explainability Pitfalls in Language Learning Systems Why Language Models Hallucinate

Reference 20

Resolution
metadata mismatch
local_arxiv, observed 2026-05-25T06:35:25.207382Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-25T06:31:00.249428Z digest=sha256:5172e9c05d85f96f71c10132775c5236592b7e04d54565276f11b73e2b16d4ce

Observation 1669f201-5f9d-4bd6-b830-ce162b7cc631 · inbound

From Unstructured Recall to Schema-Grounded Memory: Reliable AI Memory via Iterative, Schema-Aware Extraction cites this paper.

From Unstructured Recall to Schema-Grounded Memory: Reliable AI Memory via Iterative, Schema-Aware Extraction Why Language Models Hallucinate

Reference 17

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T12:32:40.683342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-07T06:58:17.953361Z digest=sha256:b25dcb954cde1a24adec7be9a9d5ce2391684afb128b7287410f1b5ac7d5fa1d

Observation e8f6ccd4-9647-41fd-96e3-de15f49437be · inbound

ReLay: Personalized LLM-Generated Plain-Language Summaries for Better Understanding, but at What Cost? cites this paper.

ReLay: Personalized LLM-Generated Plain-Language Summaries for Better Understanding, but at What Cost? Why Language Models Hallucinate

Reference 13

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T12:32:40.683342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-09T20:01:49.749076Z digest=sha256:d5c76fdb7322283ab3adbfd27348403bdb6befdf08d378eaf26083963f49dcd6

Observation bb9c6f3f-fd25-4f51-bbef-eab03403b27f · inbound

CLEAR: Revealing How Noise and Ambiguity Degrade Reliability in LLMs for Medicine cites this paper.

CLEAR: Revealing How Noise and Ambiguity Degrade Reliability in LLMs for Medicine Why Language Models Hallucinate

Reference 30

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T12:32:40.683342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-09T18:38:44.201038Z digest=sha256:825f12fb6962e6e5615e3d02d0a5fb2a94d7e2074997894317f1803616620a10

Observation f204ed1e-4770-449e-b0ed-0430a47d8969 · inbound

CLEAR: Revealing How Noise and Ambiguity Degrade Reliability in LLMs for Medicine cites this paper.

CLEAR: Revealing How Noise and Ambiguity Degrade Reliability in LLMs for Medicine Why Language Models Hallucinate

Reference 30

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T12:32:40.683342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-12T01:52:55.174427Z digest=sha256:67a0858fe91e64f1f922ccaa19684fce216061416bdc5fbd1fa72b0250e47b4d

Observation 23ce9aa6-e56f-4662-817d-79b0eada6f03 · inbound

LLM Ghostbusters: Surgical Hallucination Suppression via Adaptive Unlearning cites this paper.

LLM Ghostbusters: Surgical Hallucination Suppression via Adaptive Unlearning Why Language Models Hallucinate

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-13T12:32:40.683342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-09T18:44:05.775332Z digest=sha256:1777dc586fb9bd7d2e54f4577c77ffb7179fe4ea5a92f730d720fb5285bc1bff

Observation ef77ad03-c3b9-49af-9efb-d7c001b5f34c · inbound

When Should a Language Model Trust Itself? Same-Model Self-Verification as a Conditional Confidence Signal cites this paper.

When Should a Language Model Trust Itself? Same-Model Self-Verification as a Conditional Confidence Signal Why Language Models Hallucinate

Reference 4

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T12:32:40.683342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-10T17:34:30.826104Z digest=sha256:193f7fe31ce4aa3915afff3dcb4f224c6412524c31977528ba9e2eb67c433126

Observation b4e24ae3-4cd5-406d-94e6-571629a5e630 · inbound

Agentic Repository Mining: A Multi-Task Evaluation cites this paper.

Agentic Repository Mining: A Multi-Task Evaluation Why Language Models Hallucinate

Reference 20

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T12:32:40.683342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-08T16:04:59.167405Z digest=sha256:c658249326cc94220dc49f5ac41ba3c88876324557ab65a6495b2d7bfcdf1dbe

Observation 11cfa4ed-41b6-4212-876e-33c8f77f8227 · inbound

Low-Cost Black-Box Detection of LLM Hallucinations via Dynamical System Prediction cites this paper.

Low-Cost Black-Box Detection of LLM Hallucinations via Dynamical System Prediction Why Language Models Hallucinate

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-13T12:32:40.683342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-08T16:52:23.713204Z digest=sha256:45d039020c0f3f798c460e143c05de7e976dd5670a1e2e8d2285f429406ef44c

Observation bd0b4861-18d3-4839-bc18-9137c5cfdef7 · inbound

Benchmarking LLM-Based Static Analysis for Secure Smart Contract Development: Reliability, Limitations, and Potential Hybrid Solutions cites this paper.

Benchmarking LLM-Based Static Analysis for Secure Smart Contract Development: Reliability, Limitations, and Potential Hybrid Solutions Why Language Models Hallucinate

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-05-13T12:32:40.683342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T02:18:05.137488Z digest=sha256:2377504291248a647b1ebf83d69edeaf748ec563eab1fa0406962c057b3defa0

Observation edb75b36-0bd1-4cc3-a0e1-9489a6d3c9c6 · inbound

Scalable Token-Level Hallucination Detection in Large Language Models cites this paper.

Scalable Token-Level Hallucination Detection in Large Language Models Why Language Models Hallucinate

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-13T12:32:40.683342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-13T05:49:23.534294Z digest=sha256:658d8db983abeb3e55bf6703bf05570c01225aafc6a668ccf801396e5794a3ca

Observation 6c791c65-fd90-4091-8d34-2ddfdfbae65d · inbound

Measuring Google AI Overviews: Activation, Source Quality, Claim Fidelity, and Publisher Impact cites this paper.

Measuring Google AI Overviews: Activation, Source Quality, Claim Fidelity, and Publisher Impact Why Language Models Hallucinate

Reference 26

Resolution
metadata mismatch
local_arxiv, observed 2026-05-15T02:13:30.270092Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-15T02:12:39.528277Z digest=sha256:cb2edf7281bb8fd36d83cf5ec5cd4d0dfcc77b635313dd8cddb9ebeed5dbfef2

Observation d50dfe45-6b0c-456d-8cba-d809096a5731 · inbound

Predictable Confabulations: Factual Recall by LLMs Scales with Model Size and Topic Frequency cites this paper.

Predictable Confabulations: Factual Recall by LLMs Scales with Model Size and Topic Frequency Why Language Models Hallucinate

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-05-20T11:03:13.722869Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-20T10:59:32.371351Z digest=sha256:aa0608d7353baf05cd1453db799f1fc1b83b476dd548b406083e96e338e2b69a

Observation d8f05633-1f16-4fb6-8db8-79f54d38ef73 · inbound

Towards FairRAG: Preventing Representational Harm in Retrieval-Augmented Generation by Enforcing Fair Exposure at Retrieval Time cites this paper.

Towards FairRAG: Preventing Representational Harm in Retrieval-Augmented Generation by Enforcing Fair Exposure at Retrieval Time Why Language Models Hallucinate

Reference 16

Resolution
metadata mismatch
local_arxiv, observed 2026-05-20T22:23:47.837798Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-20T22:22:01.645643Z digest=sha256:90c334184081b39c938afc3f1be38b5f36ce932e6778a2b2e2f30ba82e33b4e4

Observation 2caf087e-75e3-4ebd-a74d-01b61adc1692 · inbound

Opportunities and Risks of Generative AI through the Health Information Journey cites this paper.

Opportunities and Risks of Generative AI through the Health Information Journey Why Language Models Hallucinate

Reference 66

Resolution
metadata mismatch
local_arxiv, observed 2026-05-25T05:15:22.479016Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-05-25T05:12:28.366867Z digest=sha256:8608b6ebf0a01b6a7d7952dc7f5b29d9112fecacd2edca1c99cce63ae13ab210

Observation 44b1588e-ae17-4dde-bc58-3999e0fcda95 · inbound

SVR-MAD: A Bayesian-Inspired Framework for Posterior-Guided Multi-Agent Debate cites this paper.

SVR-MAD: A Bayesian-Inspired Framework for Posterior-Guided Multi-Agent Debate Why Language Models Hallucinate

Reference 23

Resolution
metadata mismatch
local_arxiv, observed 2026-05-25T04:55:23.833332Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-25T04:51:54.151671Z digest=sha256:6480b5b02b3378aded010b07dd6850d756104aa1b16fe2818503cc94c647a4b2

Observation 39db4301-d476-4c17-9ad8-11ad53e2a37c · inbound

Confidence Calibration in Large Language Models cites this paper.

Confidence Calibration in Large Language Models Why Language Models Hallucinate

Reference 9

Resolution
unresolved
no resolver link, observed 2026-07-13T13:23:25.998584Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T13:23:25.998584Z digest=sha256:70d83c832a09632afe5371e1a91c8a7f3fff607190f2d657dc303d8d54e0ab49

Observation b365c188-ad31-4ca6-8cc8-60e44dd214fe · inbound

Plans for Evaluating Structured Generative Search Summaries cites this paper.

Plans for Evaluating Structured Generative Search Summaries Why Language Models Hallucinate

Reference 22

Resolution
metadata mismatch
local_arxiv, observed 2026-06-29T16:33:39.100484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-29T16:26:58.658036Z digest=sha256:3a417ad72324c3bb8dc88f82442753e57e41a1bbf8cf2949269f379f05c20673

Observation e26fcf5e-f373-4f2f-8e27-00f4799e05cf · inbound

Innovation: An Almost Characterization of Hallucination cites this paper.

Innovation: An Almost Characterization of Hallucination Why Language Models Hallucinate

Reference 4

Resolution
metadata mismatch
local_arxiv, observed 2026-06-29T19:33:54.220134Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-29T19:26:16.761291Z digest=sha256:0e3d76a64058d52421e4c6fdd9d34c9b42dbdaf6216fdd570dededaaf2aabd50

Observation 70a1aebe-352e-4c7c-85ed-85fcffa3e12d · inbound

It's Not Always Sycophancy: Measuring LLM Conformity as a Function of Epistemic Uncertainty cites this paper.

It's Not Always Sycophancy: Measuring LLM Conformity as a Function of Epistemic Uncertainty Why Language Models Hallucinate

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-06-29T18:03:47.985325Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-29T18:00:09.760987Z digest=sha256:e600938d3dcf9d8d84599da7234d6b28f050c43ac054fd5f83e630c80c175de7

Observation fa4e0737-8d31-4408-8aed-d69f3237ce9e · inbound

Entropy Distribution as a Fingerprint for Hallucinations in Generative Models cites this paper.

Entropy Distribution as a Fingerprint for Hallucinations in Generative Models Why Language Models Hallucinate

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-06-29T12:13:26.669886Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-29T12:10:43.626846Z digest=sha256:471da3a18e8e59d0029ccb29a728d46fa746cbedf390e874073935a77efb360f

Observation d6d937f1-645c-425c-a998-7b2b5e57624c · inbound

What Makes LVLMs Hallucinate Less? Unveiling the Architectural Factors Behind Hallucination Robustness cites this paper.

What Makes LVLMs Hallucinate Less? Unveiling the Architectural Factors Behind Hallucination Robustness Why Language Models Hallucinate

Reference 36

Resolution
verified exact
local_arxiv, observed 2026-06-28T23:22:46.586186Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-28T23:21:50.254449Z digest=sha256:a4580a08e9597a26d723e7f5395744c581e311ff16a7ca176424711c61020e75

Observation 0fac5bfc-4dd5-4f02-9e84-c9fbf17366c9 · inbound

Towards Lightweight Reliability: Using Soft Prompts for Hallucination Mitigation in Large Language Models cites this paper.

Towards Lightweight Reliability: Using Soft Prompts for Hallucination Mitigation in Large Language Models Why Language Models Hallucinate

Reference 18

Resolution
metadata mismatch
local_arxiv, observed 2026-06-28T20:32:38.008345Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-28T18:32:29.161640Z digest=sha256:21e2dfd96047848f28390736010a03d7dcb0d331e3db7313c813b771499f34e4

Observation 2316f796-ce99-4cdc-8427-115367973bed · inbound

POIROT: Interrogating Agents for Failure Detection in Multi-Agent Systems cites this paper.

POIROT: Interrogating Agents for Failure Detection in Multi-Agent Systems Why Language Models Hallucinate

Reference 11

Resolution
metadata mismatch
local_arxiv, observed 2026-07-01T23:06:19.977845Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-28T14:44:21.487169Z digest=sha256:aaa0b427767df2a1a38f3f82a25ff5d8cf844e0d1c4e7676670fd90273672bc8

Observation e8beacca-cabc-4c2b-af07-4786c547fdcb · inbound

DECK: A Consistency x Confidence Taxonomy of LLM Hallucinations cites this paper.

DECK: A Consistency x Confidence Taxonomy of LLM Hallucinations Why Language Models Hallucinate

Reference 37

Resolution
metadata mismatch
local_arxiv, observed 2026-07-01T22:56:19.754046Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-06-28T14:57:12.792031Z digest=sha256:525efbb6933fa4ed9a0dd1daf495bb3486f91fca9b37447546334086a9fc3dc5

Observation 1beaa9b8-b18e-4bc9-9136-768e9768a054 · inbound

Bastet: A Fine-Grained Expert-Labeled Dataset for DeFi Smart Contract Vulnerability Detection cites this paper.

Bastet: A Fine-Grained Expert-Labeled Dataset for DeFi Smart Contract Vulnerability Detection Why Language Models Hallucinate

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-07-02T03:36:29.647269Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-28T09:54:42.373926Z digest=sha256:44c7124aba4610a65f774066ce122eeecb8bd78964127c7977247b041a6df1c8

Observation b96b328b-5b84-49ed-89f7-37944d8e7af1 · inbound

Large Language Models Are Overconfident in Their Own Responses cites this paper.

Large Language Models Are Overconfident in Their Own Responses Why Language Models Hallucinate

Reference 9

Resolution
metadata mismatch
local_arxiv, observed 2026-07-02T03:16:33.649262Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-06-28T10:14:14.372346Z digest=sha256:8a50d790da99d1896dd0c67e110956e76f06e89bbadfe856972023a116ac7445

Observation 474cc72b-0c91-4271-9cbb-9090152829fb · inbound

SANE Schema-aware Natural-language Evaluation of Biological Data cites this paper.

SANE Schema-aware Natural-language Evaluation of Biological Data Why Language Models Hallucinate

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-07-02T07:56:47.078978Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-28T06:34:54.198555Z digest=sha256:26bca03326545ed766dde7c11d569c9023e83add928651cd1d69c777790c08c2

Observation 5703f455-dba7-4e7f-bfd3-f339ab8fe262 · inbound

Carrollian holography with agentic AI: Real mass is imaginary cites this paper.

Carrollian holography with agentic AI: Real mass is imaginary Why Language Models Hallucinate

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-07-02T10:56:52.832101Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-28T04:51:03.607367Z digest=sha256:13633980e2e0674fb5f8b3f6f3304b382440cd5f7a1c49c35c84473b3cd3d4a6

Observation 90e6ba6b-4e22-474d-bd28-78c8d5667152 · inbound

OpenHalDet: A Unified Benchmark for Hallucination Detection across Diverse Generation Scenarios cites this paper.

OpenHalDet: A Unified Benchmark for Hallucination Detection across Diverse Generation Scenarios Why Language Models Hallucinate

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-07-02T17:57:15.349053Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-27T21:50:44.634425Z digest=sha256:c22107917b820b688831015a0a86bf8394fb42b0971d1f6e2e33f9efb8235e0e

Observation 4bb42235-a897-4243-915d-f2a7f00d1421 · inbound

Calibrated Sampling-Free Uncertainty Estimation in Bayesian Deep Learning cites this paper.

Calibrated Sampling-Free Uncertainty Estimation in Bayesian Deep Learning Why Language Models Hallucinate

Reference 23

Resolution
verified exact
local_arxiv, observed 2026-07-03T17:28:44.715937Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-27T04:04:33.390770Z digest=sha256:f88082a176e2ecdea7de241a31507616f7e77b20b5fbc77a5ddc0a7a7394618c

Observation 37e0f20d-db66-4321-9124-d7a086fc287f · inbound

Caring Without Feeling: Affective Dynamics as the Control Layer of Human-AI Agent Collaboration cites this paper.

Caring Without Feeling: Affective Dynamics as the Control Layer of Human-AI Agent Collaboration Why Language Models Hallucinate

Reference 68

Resolution
metadata mismatch
local_arxiv, observed 2026-06-30T23:35:07.574437Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-30T23:29:42.265595Z digest=sha256:ae4083cecc3f7ebc9943b6b719a040c37c491ed1304db6bc55800ffcd1643f58

Observation 436c23b3-ad28-424c-9f6c-a7a70856a208 · inbound

Breaking the Solver Bottleneck: Training Task Generators at the Learnable Frontier cites this paper.

Breaking the Solver Bottleneck: Training Task Generators at the Learnable Frontier Why Language Models Hallucinate

Reference 152

Resolution
metadata mismatch
local_arxiv, observed 2026-07-03T08:57:48.097843Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-06-27T10:36:09.211639Z digest=sha256:ce8d90d01d1cae16874e1a43dcd5c6ee3332f73f4412356d362c2c9413d4d68c

Observation 3de9a33c-0f5f-475e-8ba7-402e820188c7 · inbound

Formally Verified Code Synthesis for Structured Data Translation in a Medical Internet of Things cites this paper.

Formally Verified Code Synthesis for Structured Data Translation in a Medical Internet of Things Why Language Models Hallucinate

Reference 5

Resolution
metadata mismatch
local_arxiv, observed 2026-07-04T05:19:35.243408Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-06-26T16:07:28.510381Z digest=sha256:f92ca0594c521d07d04657a2007c9326b916dafaf34b5e81466c67e52544d967

Observation 88300e01-06a4-452c-b465-ed4ccfb21bfe · inbound

Bayesian Adaptation Gym: A Benchmark for the Bayesian Low-Rank Adaptation of Multi-Modal Language Models cites this paper.

Bayesian Adaptation Gym: A Benchmark for the Bayesian Low-Rank Adaptation of Multi-Modal Language Models Why Language Models Hallucinate

Reference 5

Resolution
verified exact
local_arxiv, observed 2026-07-04T08:09:41.583697Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-26T12:07:15.289430Z digest=sha256:99fcb2d875d7bad8fca92fb623598b011341c50e5f7173a5b626f4802e38a983

Observation ff470234-330d-4d63-840e-54936217790c · inbound

Grad Detect: Gradient-Based Hallucination Detection in LLMs cites this paper.

Grad Detect: Gradient-Based Hallucination Detection in LLMs Why Language Models Hallucinate

Reference 43

Resolution
metadata mismatch
local_arxiv, observed 2026-06-26T00:08:43.211134Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-06-25T23:59:19.950066Z digest=sha256:f7dedf6e1dc1b26208259b7b9a3391d8c5ec8561c6285d6672f3ddb437a37b0b

Observation 8bbb4e39-6ed5-4abf-9c07-a77f2a155807 · inbound

LibEvoBench: Probing Temporal Knowledge Stratification in Code Generation Models cites this paper.

LibEvoBench: Probing Temporal Knowledge Stratification in Code Generation Models Why Language Models Hallucinate

Reference 44

Resolution
metadata mismatch
local_arxiv, observed 2026-06-25T21:18:24.664100Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-06-25T20:38:40.331343Z digest=sha256:30ba6298069904d2300ad7944edabf6ec0bf58574800218a96fe01b7330f2bf5

Observation 8a321c21-dfe2-46d7-ac98-ba344afd2d82 · inbound

Thinking Like a Scientist? A Structural Study of LLM-Generated Research Methods cites this paper.

Thinking Like a Scientist? A Structural Study of LLM-Generated Research Methods Why Language Models Hallucinate

Reference 43

Resolution
metadata mismatch
local_arxiv, observed 2026-07-03T17:28:44.388554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-06-27T04:09:11.258587Z digest=sha256:58acd4f1e44f163ed2f97811eb0d38285f93f0d3b04cb4a9907d305b47d9e106

Observation 26789822-c4df-487e-8a15-9069864eea22 · inbound

What We are Missing in Multimodal LLM Evaluation? cites this paper.

What We are Missing in Multimodal LLM Evaluation? Why Language Models Hallucinate

Reference 4

Resolution
metadata mismatch
local_arxiv, observed 2026-07-04T15:39:56.692467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-26T01:31:35.307417Z digest=sha256:626569dbf305e36556e866a67dc9fd9f7be17cbd31a6cf0c30880233b84b54db

Observation b6776bdc-920e-470c-9e2d-d3d72ce1a8b5 · inbound

At the Edge of Understanding: Sparse Autoencoders Trace The Limits of Transformer Generalization cites this paper.

At the Edge of Understanding: Sparse Autoencoders Trace The Limits of Transformer Generalization Why Language Models Hallucinate

Reference 32

Resolution
metadata mismatch
local_arxiv, observed 2026-06-26T01:28:50.571058Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-06-26T01:27:39.812228Z digest=sha256:74ead3abbe91991c6edd4a2d11374c27d979b9cac0346bca4f3d2d220e7ea2d2

Observation b1f9a726-9170-4886-ae66-d52f50a8f52c · inbound

Humanizing Automatically Generated Unit Test Suites with LLM-Based Refactoring cites this paper.

Humanizing Automatically Generated Unit Test Suites with LLM-Based Refactoring Why Language Models Hallucinate

Reference 30

Resolution
verified exact
local_arxiv, observed 2026-06-29T03:03:03.858199Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-29T03:01:26.331966Z digest=sha256:ca99d4b0adae1adb4c0eb914fdf5e54b0d4e2f466a1b5e6dc65c30e1eb771109

Observation 9e23311f-6f3e-43e1-888b-a4e25c8d3d37 · inbound

Humanizing Automatically Generated Unit Test Suites with LLM-Based Refactoring cites this paper.

Humanizing Automatically Generated Unit Test Suites with LLM-Based Refactoring Why Language Models Hallucinate

Reference 30

Resolution
verified exact
local_arxiv, observed 2026-07-02T21:17:23.476727Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-07-02T21:16:09.463677Z digest=sha256:ce5fdb195b9c9e13f68579a4bd73aa53e6d9429ae56ed488286bdfd25ede5b69

Observation a9595119-5105-4b35-a78d-de5b5d968eab · inbound

NPUsper: Eliminating Redundant Computation for Real-Time Whisper on Mobile NPUs cites this paper.

NPUsper: Eliminating Redundant Computation for Real-Time Whisper on Mobile NPUs Why Language Models Hallucinate

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-07-02T06:06:40.636478Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-07-02T05:57:29.325247Z digest=sha256:bf6287dd6a61c0680a42e56185acb86ea2ad5b81b13db9443b104f306f375f20

Observation 80a56620-8aac-4f65-9014-275c53e7c094 · inbound

How to Avoid Debate: Scalable AI Safety via Doubly-Efficient Interactive Proofs cites this paper.

How to Avoid Debate: Scalable AI Safety via Doubly-Efficient Interactive Proofs Why Language Models Hallucinate

Reference 14

Resolution
unresolved
no resolver link, observed 2026-07-12T01:34:45.323277Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-12T01:34:45.323277Z digest=sha256:c3b27187ba2d5b0fb5b9d186e8237e23e5d417d8057a18ad36b78da452edb720

Observation c6a94747-907a-47e5-b2e9-7ae23efa9053 · inbound

Consistent but Miscalibrated: Evaluating LLM Limitations for Risk Communication in Natural Language cites this paper.

Consistent but Miscalibrated: Evaluating LLM Limitations for Risk Communication in Natural Language Why Language Models Hallucinate

Reference 177

Resolution
unresolved
no resolver link, observed 2026-07-11T23:16:58.545731Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-11T23:16:58.545731Z digest=sha256:a4f64324713295e47a5b1f74a99b7014ed0dd67317e1d4cff0de0def1200e1c9

Observation 9c3737f0-5c3c-41d4-8603-0060fc9332e0 · inbound

Consistent but Miscalibrated: Evaluating LLM Limitations for Risk Communication in Natural Language cites this paper.

Consistent but Miscalibrated: Evaluating LLM Limitations for Risk Communication in Natural Language Why Language Models Hallucinate

Reference 177

Resolution
unresolved
no resolver link, observed 2026-07-13T07:02:13.140334Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-13T07:02:13.140334Z digest=sha256:dfe9bb63da6dfc06e8fc6a8cd27f5269b36f6b4b02f11542499c2d4343070990

Observation e011f841-9ca7-4269-8230-15073a3f7e5b · inbound

The "I Don't Know" Filter: Enhancing Agentic Reliability in Function Calling cites this paper.

The "I Don't Know" Filter: Enhancing Agentic Reliability in Function Calling Why Language Models Hallucinate

Reference 6

Resolution
unresolved
no resolver link, observed 2026-07-11T22:10:22.729188Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T22:10:22.729188Z digest=sha256:1232787ed8935a7dbcca188aefbfa110202226b6d828506df9993648e3d4bee3

Observation a48ec1e8-5be8-40b7-834a-7ea622dc476d · inbound

On the effectiveness of reward functions in reinforcement learning for confidence calibration of large language models cites this paper.

On the effectiveness of reward functions in reinforcement learning for confidence calibration of large language models Why Language Models Hallucinate

Reference 19

Resolution
unresolved
no resolver link, observed 2026-07-11T20:01:06.623978Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T20:01:06.623978Z digest=sha256:30208f2172000af13c22e0fec6c0663964d5dc922ed502ba152f2fc2fefba1c4

Observation 6c4c69f9-46c6-499a-889e-0580d89e2212 · inbound

ARMOR: Stabilizing On-Policy LLM RL with Off-Policy Anchor Samples cites this paper.

ARMOR: Stabilizing On-Policy LLM RL with Off-Policy Anchor Samples Why Language Models Hallucinate

Reference 28

Resolution
unresolved
no resolver link, observed 2026-07-14T11:23:30.063231Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-14T11:23:30.063231Z digest=sha256:fb1d3a0583c54ac0876c7de5a45d73fa7a8a8723308e38d45ba2b6973cc3d831

Observation f6ee98a5-88fd-4b88-afd1-bd1fff59d1b6 · inbound

ARMOR: Stabilizing On-Policy LLM RL with Off-Policy Anchor Samples cites this paper.

ARMOR: Stabilizing On-Policy LLM RL with Off-Policy Anchor Samples Why Language Models Hallucinate

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-02T07:23:26.610302Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T07:23:26.610302Z digest=sha256:e4b4157291ef9dd9c232d1ae565bcc28e2a6d844f78949a7d81be71c46bc652d