Pith. sign in

Paper Citation Record · LEDGER

HalluLens: LLM Hallucination Benchmark

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 23 inbound Pith citation observations for arXiv:2504.17550.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.17550 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 23 of 23 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 23 of 23 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:01:08.125965Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 710392c0-8593-4199-9562-ece45783b8a7 · inbound

AbstentionBench: Reasoning LLMs Fail on Unanswerable Questions cites this paper.

AbstentionBench: Reasoning LLMs Fail on Unanswerable Questions HalluLens: LLM Hallucination Benchmark

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T05:01:08.125965Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:01:08.125965Z digest=sha256:58baf4b178f3534534206735eb42fce70b898ae9c62b34a8c3bd58680c29d1d0

Observation dd8252b2-fc3d-4546-81ec-605e437dbb68 · inbound

Embodied AI Agents: Modeling the World cites this paper.

Embodied AI Agents: Modeling the World HalluLens: LLM Hallucination Benchmark

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-06T22:10:25.436137Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:10:25.436137Z digest=sha256:2f7d293159c1a716eb238b8693da960ab98415cced69856dc226212c14d5bd6f

Observation 37e98791-aba2-4f3b-9a78-ecf26545fd0c · inbound

Introducing the Swiss Food Knowledge Graph: AI for Context-Aware Nutrition Recommendation cites this paper.

Introducing the Swiss Food Knowledge Graph: AI for Context-Aware Nutrition Recommendation HalluLens: LLM Hallucination Benchmark

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T17:43:11.362519Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:43:11.362519Z digest=sha256:b20f859217b4088f3cc3d1f3dd9749a9a7c1e414e0c8cab82ac2d940f0512ba0

Observation c3bc9cd2-4707-482c-8568-42007d7edf06 · inbound

MIRAGE-Bench: LLM Agent is Hallucinating and Where to Find Them cites this paper.

MIRAGE-Bench: LLM Agent is Hallucinating and Where to Find Them HalluLens: LLM Hallucination Benchmark

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T13:06:36.599018Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:06:36.599018Z digest=sha256:81aa7717f65a81d383e44a055fe41e44b0540619dd666c85cde8625053c81817

Observation 5831c5ed-3040-4a23-a8c2-3a38c4a34cce · inbound

A comprehensive taxonomy of hallucinations in Large Language Models cites this paper.

A comprehensive taxonomy of hallucinations in Large Language Models HalluLens: LLM Hallucination Benchmark

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T05:29:09.780706Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T05:29:09.780706Z digest=sha256:8efcaf6bc4460c00540e3fd31b3096b59fdcbd367cadbc185fe58758c8d943e6

Observation 08f453e8-d49c-4bdc-8a33-614d4f0addf4 · inbound

ReasoningTrack: Chain-of-Thought Reasoning for Long-term Vision-Language Tracking cites this paper.

ReasoningTrack: Chain-of-Thought Reasoning for Long-term Vision-Language Tracking HalluLens: LLM Hallucination Benchmark

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-05T23:33:19.965610Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:33:19.965610Z digest=sha256:1ec8b6220913f8795d51e22b946d66bd711e123dd1f1ab1943ec5e3cf30398ee

Observation 1caab12d-ef18-4070-a689-d66bcbeb9b4a · inbound

GOSU: Retrieval-Augmented Generation with Global-Level Optimized Semantic Unit-Centric Framework cites this paper.

GOSU: Retrieval-Augmented Generation with Global-Level Optimized Semantic Unit-Centric Framework HalluLens: LLM Hallucination Benchmark

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-05T13:36:35.132035Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T13:36:35.132035Z digest=sha256:b5ad54cced6a2f702fdbd9466a1079133ae38d48520f8f4c9fc7e5b2291ed692

Observation 1e25fd92-76c3-4f24-9d4d-bd7323c38648 · inbound

ReFACT: A Benchmark for Scientific Confabulation Detection with Positional Error Annotations cites this paper.

ReFACT: A Benchmark for Scientific Confabulation Detection with Positional Error Annotations HalluLens: LLM Hallucination Benchmark

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-18T12:56:24.497247Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-18T12:54:01.717015Z digest=sha256:703e13894751a7b8d909777a63e6b928a6472a012e4c89c82f0edd80d04fc54e

Observation eb0f60c0-8233-4d36-b187-45b6acb266dc · inbound

When Do Hallucinations Arise? A Graph Perspective on the Evolution of Path Reuse and Path Compression cites this paper.

When Do Hallucinations Arise? A Graph Perspective on the Evolution of Path Reuse and Path Compression HalluLens: LLM Hallucination Benchmark

Reference 1

Resolution
unresolved
no resolver link, observed 2026-07-13T13:07:00.928770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T13:07:00.928770Z digest=sha256:fecf8196fb70a170bd5406f04771185d02b51339412fef05a2cbb6f6149214cb

Observation 5cb32dba-21c3-4776-b09b-ea6890e3981c · inbound

From Binary Groundedness to Support Relations: Towards a Reader-Centred Taxonomy for Comprehension of AI Output cites this paper.

From Binary Groundedness to Support Relations: Towards a Reader-Centred Taxonomy for Comprehension of AI Output HalluLens: LLM Hallucination Benchmark

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-11T06:30:58.657451Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T17:37:28.129376Z digest=sha256:8a382fb1b6c0ae1cb6d744eee4e3a81adeac541a1d2494e92af8b0ed813b1c38

Observation 6be78a1e-743e-44e0-8a4b-e759714bc979 · inbound

HalluClear: Diagnosing, Evaluating and Mitigating Hallucinations in GUI Agents cites this paper.

HalluClear: Diagnosing, Evaluating and Mitigating Hallucinations in GUI Agents HalluLens: LLM Hallucination Benchmark

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-10T06:31:30.913700Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T06:27:19.717895Z digest=sha256:bee3a2cf2d1fa6c5e2042a9e8d35fa04d6dcc90a7aeee200ea1f327b1c639451

Observation fde6b7d0-6551-476f-9e37-e6c3b7aa2dd7 · inbound

Contrastive Attribution in the Wild: An Interpretability Analysis of LLM Failures on Realistic Benchmarks cites this paper.

Contrastive Attribution in the Wild: An Interpretability Analysis of LLM Failures on Realistic Benchmarks HalluLens: LLM Hallucination Benchmark

Reference 9

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T09:38:42.843693Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T05:12:31.218055Z digest=sha256:08f68820086019598dc6a578b6d8b9533e5048f61f5c74c9cb79458f2815321e

Observation 58b08582-fad0-494c-bc2f-8345ec78c00c · inbound

Benchmarking Source-Sensitive Reasoning in Turkish: Humans and LLMs under Evidential Trust Manipulation cites this paper.

Benchmarking Source-Sensitive Reasoning in Turkish: Humans and LLMs under Evidential Trust Manipulation HalluLens: LLM Hallucination Benchmark

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-11T21:56:34.399829Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-08T03:42:06.876417Z digest=sha256:24319e84e7b243b4e13e09907932705a68ebdb008c191b43d75e661d357cae69

Observation 6428576b-9a66-41e8-92e0-ce1a174edf42 · inbound

Hallucination as an Anomaly: Dynamic Intervention via Probabilistic Circuits cites this paper.

Hallucination as an Anomaly: Dynamic Intervention via Probabilistic Circuits HalluLens: LLM Hallucination Benchmark

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-11T19:51:11.274747Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-08T10:52:14.666333Z digest=sha256:7b24cd5c9af8f165b41b69f7e53a132580428d24a4b8b23f12285af977b0e75b

Observation b96bea36-ddba-4ac7-a459-694aa94f0410 · inbound

StereoTales: A Multilingual Framework for Open-Ended Stereotype Discovery in LLMs cites this paper.

StereoTales: A Multilingual Framework for Open-Ended Stereotype Discovery in LLMs HalluLens: LLM Hallucination Benchmark

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-12T05:51:27.293122Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-12T04:50:17.399580Z digest=sha256:bb9b0decac740bcb26b248e877e45676fdebf577bf18c20eca0623e479b9bb3f

Observation ce33085a-6e5d-4c06-9941-2928e6d6c9d7 · inbound

StereoTales: A Multilingual Framework for Open-Ended Stereotype Discovery in LLMs cites this paper.

StereoTales: A Multilingual Framework for Open-Ended Stereotype Discovery in LLMs HalluLens: LLM Hallucination Benchmark

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-13T07:22:28.981194Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T07:20:32.494840Z digest=sha256:af15501509f73c77b5fe9b8144280aeb81985fcd36ba94866495bed32a98e60c

Observation bd559f56-718a-48b3-aefb-91e6c6cd3c4c · inbound

REALISTA: Realistic Latent Adversarial Attacks that Elicit LLM Hallucinations cites this paper.

REALISTA: Realistic Latent Adversarial Attacks that Elicit LLM Hallucinations HalluLens: LLM Hallucination Benchmark

Reference 105

Resolution
verified exact
arxiv_id, observed 2026-05-14T20:17:56.743792Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-14T20:13:10.814899Z digest=sha256:6cd00b477f649f75a7795a02677244bb04e0bfe5ad7c37c3c2dad549a09a1689

Observation a6760fd2-3410-4673-83a9-401894c769b2 · inbound

Do No Harm? Hallucination and Actor-Level Abuse in Web-Deployed Medical Large Language Models cites this paper.

Do No Harm? Hallucination and Actor-Level Abuse in Web-Deployed Medical Large Language Models HalluLens: LLM Hallucination Benchmark

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-21T05:53:59.416150Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-21T05:49:55.738870Z digest=sha256:8df975eeebc5adb58169da5cceee0b7c1045729dc03f012142fbbbfde0cbce70

Observation ccf7b7c9-e403-42c9-b9d7-0bdaec1b7542 · inbound

K-FinHallu: A Hallucination Detection Benchmark for Multi-Turn RAG in Korean Finance cites this paper.

K-FinHallu: A Hallucination Detection Benchmark for Multi-Turn RAG in Korean Finance HalluLens: LLM Hallucination Benchmark

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-06-29T09:03:16.148279Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-29T08:54:48.807164Z digest=sha256:3ee1c736becbfe20af535a5fa575bb50957d5fff0eb0b030b7a505a187ac0071

Observation b82edf81-756c-4f85-939b-097951a4d9e7 · inbound

Latent Performance Profiling of Large Language Models cites this paper.

Latent Performance Profiling of Large Language Models HalluLens: LLM Hallucination Benchmark

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-06-29T07:33:13.394631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-29T07:31:02.595386Z digest=sha256:1be9511437de31aa8a694276d56613327447e1ac86cc4883bd5a6c1bf8e2a6f2

Observation f3365005-c053-4dbb-9a77-89f8f2fcb14d · inbound

Disentangling Visual and Factual Correctness in LVLMs' Visualization Literacy cites this paper.

Disentangling Visual and Factual Correctness in LVLMs' Visualization Literacy HalluLens: LLM Hallucination Benchmark

Reference 41

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T02:06:26.198841Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T11:19:51.893740Z digest=sha256:4369bac3f6d34d03a19e960d283e665f18bd8ec86da59e2e2eda14ec430cfa82

Observation 8026be1b-17af-41bf-9403-7f7421cd59c0 · inbound

What Do People Actually Want From AI? Mapping Preference Plurality cites this paper.

What Do People Actually Want From AI? Mapping Preference Plurality HalluLens: LLM Hallucination Benchmark

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-06-28T01:41:29.860815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T01:32:04.660400Z digest=sha256:8977c1116834774dbfd85a6646c0f69eab8139aec3976f8beb6f43bede7715c4

Observation 00ed01a9-0708-42ef-9e46-ad803e5d813a · inbound

Contextualized Evaluation of Vision Language Models through Dynamic, Multi-turn Interactions cites this paper.

Contextualized Evaluation of Vision Language Models through Dynamic, Multi-turn Interactions HalluLens: LLM Hallucination Benchmark

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-02T01:57:59.043987Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T01:57:59.043987Z digest=sha256:07a584871f36a82c8a85d64356ce9e7ae97a5c18e7cd99f28209cf041887c6b8