Pith. sign in

Paper Citation Record · LEDGER

Beyond Facts: Evaluating Intent Hallucination in Large Language Models

As of 8 August 2026, this Paper Citation Record lists 33 of 33 outbound references and 0 inbound Pith citation observations for arXiv:2506.06539.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.06539 v1

Coverage vector

measured 33 of 33 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:59:45.148904Z

measured 33 of 33 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

33 of 33 outbound references displayed

  • verified exact3
  • verified fuzzy0
  • unresolved30
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 7486e26c-2f03-451b-b57d-fc3cdaf1ab93 · outbound

This paper cites The Internal State of an LLM Knows When It's Lying.

Beyond Facts: Evaluating Intent Hallucination in Large Language Models The Internal State of an LLM Knows When It's Lying

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T05:59:42.983678Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:59:42.983678Z digest=sha256:7edf37900b79e5cf10aa6c0d40550e70b56ad60c0471c66eb74d603301bc3982

Observation af59f1e0-b7bf-40b0-b518-e7dd09bc33fa · outbound

This paper cites an unresolved cited work.

Beyond Facts: Evaluating Intent Hallucination in Large Language Models Unresolved cited work

Reference 2

Resolution
verified exact
doi, observed 2026-08-07T05:59:45.205301Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T05:59:43.082483Z digest=sha256:7b8729d0e7e8b2d5cb18123ee0521d598fd4646f664ad08426eba344ac8572e2

Observation 1dbc5d44-9024-428c-b333-cbc9696dcf43 · outbound

This paper cites Hallucinated but Factual! Inspecting the Factuality of Hallucinations in Abstractive Summarization.

Beyond Facts: Evaluating Intent Hallucination in Large Language Models Hallucinated but Factual! Inspecting the Factuality of Hallucinations in Abstractive Summarization

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T05:59:43.157123Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:59:43.157123Z digest=sha256:ada7d3c2c1ca9a776bf6201dad29fab35d69ab884b2dd2c009e3f8537c6591b3

Observation 813cdce5-ddb9-4dbc-b57a-58910839ddd6 · outbound

This paper cites FELM: Benchmarking Factuality Evaluation of Large Language Models.

Beyond Facts: Evaluating Intent Hallucination in Large Language Models FELM: Benchmarking Factuality Evaluation of Large Language Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T05:59:43.265833Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:59:43.265833Z digest=sha256:80510cf2946f2ac2dd56ecc396da6a3cdbad076e2c5e61754f40d3b7ae7d4d9e

Observation 499c3ed1-298b-4b73-94ea-c1839f07e23d · outbound

This paper cites The Llama 3 Herd of Models.

Beyond Facts: Evaluating Intent Hallucination in Large Language Models The Llama 3 Herd of Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T05:59:43.349807Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:59:43.349807Z digest=sha256:d5ed3d03b338feacb3d2ccd2d026fada086b71836b9a9f91f6cc02e0f80b6871

Observation 6e4690cb-73dd-4714-b550-5d69ac0a0a28 · outbound

This paper cites an unresolved cited work.

Beyond Facts: Evaluating Intent Hallucination in Large Language Models Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:59:45.489654Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T05:59:43.437293Z digest=sha256:eb61ff4d664e7a3f0f0676da046997932a7705cdffd7fd98bafdece6326ba96b

Observation 3f9ba1b7-4bfe-4f8a-9dcd-48e09c598241 · outbound

This paper cites A Probabilistic Framework for LLM Hallucination Detection via Belief Tree Propagation.

Beyond Facts: Evaluating Intent Hallucination in Large Language Models A Probabilistic Framework for LLM Hallucination Detection via Belief Tree Propagation

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T05:59:43.526747Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:59:43.526747Z digest=sha256:fccd53bd123feccf2d85b38ce13280b780520496bbcc88cd9f521901575e5ee9

Observation d8fb3b0e-2f69-45e2-87a2-1b974f8a7ff3 · outbound

This paper cites A Survey on Hallucination in Large Language Models: Principles, Taxonomy, Challenges, and Open Questions.

Beyond Facts: Evaluating Intent Hallucination in Large Language Models A Survey on Hallucination in Large Language Models: Principles, Taxonomy, Challenges, and Open Questions

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T05:59:43.609322Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:59:43.609322Z digest=sha256:8dc00ce797a887d754ffbeee3591a612bac4a7d25f55ca33a2a4768b25c823dc

Observation 8df5f450-380c-4a56-9db7-3af207d7fac4 · outbound

This paper cites an unresolved cited work.

Beyond Facts: Evaluating Intent Hallucination in Large Language Models Unresolved cited work

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T05:59:43.689890Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:59:43.689890Z digest=sha256:86d8a7f65fef763e778c11288d50707f90063573cc813e7d1d2ed61c98feb6b3

Observation 1a17cd4e-0354-45da-91eb-e25eb0a49620 · outbound

This paper cites Mistral 7B.

Beyond Facts: Evaluating Intent Hallucination in Large Language Models Mistral 7B

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T05:59:43.753431Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:59:43.753431Z digest=sha256:25e8ca9b9514aca59c638bdfea45fd3fbf1ed8a8402415dcfbde7d22ab1e362f

Observation 5d24f450-98f6-456e-a2c3-66500b910799 · outbound

This paper cites an unresolved cited work.

Beyond Facts: Evaluating Intent Hallucination in Large Language Models Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:59:45.478148Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T05:59:43.841056Z digest=sha256:52e760e5a40ab510f533daef5b90872866f905560efdf6e89a76a1946d1a5ce7

Observation c0594533-16f9-4a33-a6a3-fb39606e34e9 · outbound

This paper cites HaluEval: A Large-Scale Hallucination Evaluation Benchmark for Large Language Models.

Beyond Facts: Evaluating Intent Hallucination in Large Language Models HaluEval: A Large-Scale Hallucination Evaluation Benchmark for Large Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T05:59:43.915918Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:59:43.915918Z digest=sha256:10734fba5e07b13e2d3c6e65f54c6368a122b5c8c528db89ce68506414562658

Observation 0da7d885-f47c-4439-8a21-4a6f2d24ae8b · outbound

This paper cites Lost in the Middle: How Language Models Use Long Contexts.

Beyond Facts: Evaluating Intent Hallucination in Large Language Models Lost in the Middle: How Language Models Use Long Contexts

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T05:59:43.976991Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:59:43.976991Z digest=sha256:b3ff23108f8be1513f3355878921e203342c6f0c8591c9c223bf179a16db044f

Observation 0222f3b5-73e5-4915-b6a1-0b85358da39b · outbound

This paper cites SelfCheckGPT: Zero-Resource Black-Box Hallucination Detection for Generative Large Language Models.

Beyond Facts: Evaluating Intent Hallucination in Large Language Models SelfCheckGPT: Zero-Resource Black-Box Hallucination Detection for Generative Large Language Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T05:59:44.062538Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:59:44.062538Z digest=sha256:1cabf74f76bc7a14f219865d74495efd0cb03616c0e249f0188c4023f9e1ba32

Observation 5138a1cf-6e21-4d72-9246-c9e9298dabf4 · outbound

This paper cites FActScore: Fine-grained Atomic Evaluation of Factual Precision in Long Form Text Generation.

Beyond Facts: Evaluating Intent Hallucination in Large Language Models FActScore: Fine-grained Atomic Evaluation of Factual Precision in Long Form Text Generation

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T05:59:44.143311Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:59:44.143311Z digest=sha256:fa8be774725744541ab1a35c6bbb9c7e21e1795eae7afdd7b48d6b88750e106f

Observation 081ff96c-3773-4096-bb1d-ecc99865d9b1 · outbound

This paper cites FaithEval: Can Your Language Model Stay Faithful to Context, Even If "The Moon is Made of Marshmallows".

Beyond Facts: Evaluating Intent Hallucination in Large Language Models FaithEval: Can Your Language Model Stay Faithful to Context, Even If "The Moon is Made of Marshmallows"

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T05:59:44.232248Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:59:44.232248Z digest=sha256:9d70fdc9f4b5b92b7a5996d69f93d974a44368730164ce562e4c31ad73c5132c

Observation b6bb0d91-3f18-49f6-ae84-35f6e2313de8 · outbound

This paper cites Fine-grained Hallucination Detection and Editing for Language Models.

Beyond Facts: Evaluating Intent Hallucination in Large Language Models Fine-grained Hallucination Detection and Editing for Language Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T05:59:44.316066Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:59:44.316066Z digest=sha256:e1f447b78e92b1d03d928ebf7fa318893a580971a8595a8625ee0e48f9e4e6ee

Observation 68a25084-bb5a-4f50-ad2d-6a5e45d3ab3e · outbound

This paper cites Self-contradictory Hallucinations of Large Language Models: Evaluation, Detection and Mitigation.

Beyond Facts: Evaluating Intent Hallucination in Large Language Models Self-contradictory Hallucinations of Large Language Models: Evaluation, Detection and Mitigation

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T05:59:44.405689Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:59:44.405689Z digest=sha256:b747dd44d60d44da485f88687380972c9fa5f991495d29005d391d9cf5b38296

Observation 5bd0733c-425f-459f-a042-caaabb334476 · outbound

This paper cites RAGTruth: A Hallucination Corpus for Developing Trustworthy Retrieval-Augmented Language Models.

Beyond Facts: Evaluating Intent Hallucination in Large Language Models RAGTruth: A Hallucination Corpus for Developing Trustworthy Retrieval-Augmented Language Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T05:59:44.496379Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:59:44.496379Z digest=sha256:bb0b4cb76f539f1bcd21ad327cd45f2540c94baf221419c7dd170ed9c023497d

Observation f4114e0e-2579-4684-8d30-7ed1de508f8c · outbound

This paper cites GPT-4 Technical Report.

Beyond Facts: Evaluating Intent Hallucination in Large Language Models GPT-4 Technical Report

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T05:59:44.596701Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:59:44.596701Z digest=sha256:bd3f2d9a3665744ed9f8e738b08fbf19f1b8c46893b37ac7dfce8325466364e2

Observation e5a595aa-45dd-4311-94fb-a2086e0d5151 · outbound

This paper cites an unresolved cited work.

Beyond Facts: Evaluating Intent Hallucination in Large Language Models Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:59:45.466966Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T05:59:44.608086Z digest=sha256:ca2838ef13b32bd9ca15b7ed9307a496536700d9e500d87af33fd94b23c9edf8

Observation 0ffed680-fcb8-4413-9c16-0d51fb04d1fe · outbound

This paper cites InFoBench: Evaluating Instruction Following Ability in Large Language Models.

Beyond Facts: Evaluating Intent Hallucination in Large Language Models InFoBench: Evaluating Instruction Following Ability in Large Language Models

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T05:59:44.620268Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:59:44.620268Z digest=sha256:18383e95111f18470d266fb2b82c2e1391891e3b14ab8e7012121ed842b10f38

Observation 311cca00-19e7-49ed-bf4a-75059a34d13d · outbound

This paper cites an unresolved cited work.

Beyond Facts: Evaluating Intent Hallucination in Large Language Models Unresolved cited work

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T05:59:44.636835Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:59:44.636835Z digest=sha256:c719f761fd85e3cb2cdd577b1c52b2bbbf84a7caed9ba72104a6c62e4ca3d647

Observation 7a078d58-06a7-4aa2-a52b-726e26932c19 · outbound

This paper cites an unresolved cited work.

Beyond Facts: Evaluating Intent Hallucination in Large Language Models Unresolved cited work

Reference 24

Resolution
verified exact
doi, observed 2026-08-07T05:59:45.179874Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T05:59:44.695061Z digest=sha256:5a7336105d171cabc7dcc0bc5654595620ca146b2e11d6aa397604fee6890b08

Observation d4105d1c-af86-41ff-8019-40790cbf07e6 · outbound

This paper cites an unresolved cited work.

Beyond Facts: Evaluating Intent Hallucination in Large Language Models Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:59:45.454240Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T05:59:44.792899Z digest=sha256:852c844716574a84e3698539d31e9dc49687d4256b66ca249753dd02a95e66e7

Observation fde1dadf-855e-4a20-812c-122b094b1697 · outbound

This paper cites Self-Consistency Improves Chain of Thought Reasoning in Language Models.

Beyond Facts: Evaluating Intent Hallucination in Large Language Models Self-Consistency Improves Chain of Thought Reasoning in Language Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T05:59:44.845931Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:59:44.845931Z digest=sha256:ec37f0fa877db3c4bced3793ef047d8ccee027880a1634297514ccf88432b69d

Observation 9dfa9af7-d6d7-47d1-9e3f-5227d320b560 · outbound

This paper cites Pandora's Box or Aladdin's Lamp: A Comprehensive Analysis Revealing the Role of RAG Noise in Large Language Models.

Beyond Facts: Evaluating Intent Hallucination in Large Language Models Pandora's Box or Aladdin's Lamp: A Comprehensive Analysis Revealing the Role of RAG Noise in Large Language Models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T05:59:44.927584Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:59:44.927584Z digest=sha256:635184ef778df3d65a785e0fd18fd7076ffa045ea66940c49ec2388d985b5b8c

Observation a4cd64f7-23be-4921-8194-0b4bb2f2d5d7 · outbound

This paper cites A New Benchmark and Reverse Validation Method for Passage-level Hallucination Detection.

Beyond Facts: Evaluating Intent Hallucination in Large Language Models A New Benchmark and Reverse Validation Method for Passage-level Hallucination Detection

Reference 28

Resolution
verified exact
local_arxiv, observed 2026-08-07T05:59:45.249502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-08-07T05:59:45.006815Z digest=sha256:857054a278320c70c8527183cb5b11c26c95c9503850b8373b6cc84f3a02cd88

Observation 08d569c9-c2eb-4d08-b6ff-f8ce7a503eca · outbound

This paper cites KnowHalu: Hallucination Detection via Multi-Form Knowledge Based Factual Checking.

Beyond Facts: Evaluating Intent Hallucination in Large Language Models KnowHalu: Hallucination Detection via Multi-Form Knowledge Based Factual Checking

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T05:59:45.064636Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:59:45.064636Z digest=sha256:bcb20f726a077384c19bf8640ea192c64cfe4c58105c15be27824f7a5d41e3a1

Observation 575a9984-3e2b-4e9e-8587-5e55ee9de3d1 · outbound

This paper cites Knowledge Overshadowing Causes Amalgamated Hallucination in Large Language Models.

Beyond Facts: Evaluating Intent Hallucination in Large Language Models Knowledge Overshadowing Causes Amalgamated Hallucination in Large Language Models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T05:59:45.106922Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:59:45.106922Z digest=sha256:7a388eb9d3dacb4d57d5b3823de5a290bb1d8e9da26c5559aeb793b77dca3d82

Observation 8408cbb2-990f-4293-9580-a39ebf4322b3 · outbound

This paper cites LIMA: Less Is More for Alignment.

Beyond Facts: Evaluating Intent Hallucination in Large Language Models LIMA: Less Is More for Alignment

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T05:59:45.140074Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:59:45.140074Z digest=sha256:735ee5b63b049ac7ecde7e3d55aaa708957995d811934fe73d3926eab4969fab

Observation 4ae2b51f-ad35-4259-ae3c-6f8a2175515f · outbound

This paper cites online" 'onlinestring :=.

Beyond Facts: Evaluating Intent Hallucination in Large Language Models online" 'onlinestring :=

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T05:59:45.145255Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:59:45.145255Z digest=sha256:daf0e9de41a91b39d9bc2d5051f5a1fa95cc88e5477cee18c71f4a8fb28caaf5

Observation c3eeb918-408b-422b-8dae-8a61df6d960c · outbound

This paper cites write newline.

Beyond Facts: Evaluating Intent Hallucination in Large Language Models write newline

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T05:59:45.148904Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:59:45.148904Z digest=sha256:fdbe0add6f0039fc3c04f04ba7aef17b70c4d6ad95381bdfda68989ef1c9d4ee

Pith citing papers

No inbound Pith citation observations are available.