Pith. sign in

Paper Citation Record · LEDGER

HalluLens: LLM Hallucination Benchmark

As of 20 August 2026, this Paper Citation Record lists 14 of 14 outbound references and 26 inbound Pith citation observations for arXiv:2504.17550.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.17550 v1

Coverage vector

measured 14 of 14 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-16T10:40:19.484778Z

measured 40 of 40 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 26 of 26 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T20:58:22.465675Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

14 of 14 outbound references displayed

  • verified exact1
  • verified fuzzy3
  • unresolved10
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation fe2f6aee-274e-4910-81f7-293bbd59e291 · outbound

This paper cites an unresolved cited work.

HalluLens: LLM Hallucination Benchmark Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-16T10:40:20.127487Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T10:40:19.369975Z digest=sha256:5216e892841b54c6efcd038020c8cbbdf15b3e4cf5fa01f236a043e5fbb2baa2

Observation ec1df31a-8bdc-4d99-ad01-b4a716cdcecc · outbound

This paper cites Survey on Factuality in Large Language Models: Knowledge, Retrieval and Domain-Specificity.

HalluLens: LLM Hallucination Benchmark Survey on Factuality in Large Language Models: Knowledge, Retrieval and Domain-Specificity

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-16T10:40:19.329519Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:40:19.329519Z digest=sha256:0206f89568b8c024caa8768f40dc6edf8a86b1689aab6cb3815839e9af1612c4

Observation 38827227-be20-4e5b-89e4-87938905063b · outbound

This paper cites an unresolved cited work.

HalluLens: LLM Hallucination Benchmark Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-16T10:40:19.989618Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T10:40:19.400936Z digest=sha256:f2af4368c18e4ef8b4be6cbae316286d16e437f0d26e9a635563f7ffb4728e38

Observation e3ba799f-b172-4b43-bb5d-76a7ff6d3be4 · outbound

This paper cites an unresolved cited work.

HalluLens: LLM Hallucination Benchmark Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-16T10:40:19.948296Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T10:40:19.422585Z digest=sha256:372cfad65982deaf60f65fce9f57e2b819345db0281fb16660672c8a79c52462

Observation 48139ecb-1f90-44c1-9885-59e738babfcf · outbound

This paper cites unanswerable.

HalluLens: LLM Hallucination Benchmark unanswerable

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:40:19.919710Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T10:40:19.430376Z digest=sha256:c7c410510a9d61705a3d15e4b527d2904c3866670ca37afff41e1bc8369fa443

Observation 238c2af3-5710-4658-981a-794abdbe76d0 · outbound

This paper cites an unresolved cited work.

HalluLens: LLM Hallucination Benchmark Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-16T10:40:19.864568Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T10:40:19.440500Z digest=sha256:3e1027c060dbc5ad3881e4051ed0d080887071954c0d1c7c3ee33160d6b108db

Observation 50c546e1-9c8d-493c-8169-255c6e092061 · outbound

This paper cites an unresolved cited work.

HalluLens: LLM Hallucination Benchmark Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-16T10:40:20.081379Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T10:40:19.450485Z digest=sha256:9ed9a11430e0c85a7e1071a5d46b6168a858030f1430f862963330dfa61a9258

Observation b417c244-28f7-4c69-802a-e54dbf6aeedf · outbound

This paper cites an unresolved cited work.

HalluLens: LLM Hallucination Benchmark Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-16T10:40:20.039403Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T10:40:19.466802Z digest=sha256:cfe07084196dd7ab749563fb4e1ee613ec6b7744082c3c225acfffc3cdc8a983

Observation 467afffd-343b-4004-8875-2038360921d3 · outbound

This paper cites an unresolved cited work.

HalluLens: LLM Hallucination Benchmark Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-16T10:40:19.822478Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T10:40:19.474582Z digest=sha256:4cc2f681fc419f0ecc3545be8c44c475df06969b70734c188fbbd05118094418

Observation 35ce6d55-aa47-453e-a4f6-ba6bb0068307 · outbound

This paper cites {wiki_document}.

HalluLens: LLM Hallucination Benchmark {wiki_document}

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:40:19.782703Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T10:40:19.484778Z digest=sha256:4ed2e60e68c27619e255f685cb7d7c24359aadec0ad74d0df8672facaf2a9f5f

Observation 553fc54c-7141-4e2d-a89b-d175d90ffc60 · outbound

This paper cites GPT-4 Technical Report.

HalluLens: LLM Hallucination Benchmark GPT-4 Technical Report

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-08-16T10:40:19.316874Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:40:19.316874Z digest=sha256:ba0a95d0e3220e13211426cbfbdca2d8eb6550da66e6ab0c73f722d933ea1e9d

Observation ca7a9cbc-c633-41bf-8faf-dbd2ad02ca5e · outbound

This paper cites Measuring short-form factuality in large language models.

HalluLens: LLM Hallucination Benchmark Measuring short-form factuality in large language models

Reference 2018

Resolution
unresolved
no resolver link, observed 2026-08-16T10:40:19.340665Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:40:19.340665Z digest=sha256:4c2d2cb839047b58e272341e5277973df54be5337da76be1a91d6c0409df7830

Observation 7bca60dd-0ac8-44d7-8a17-bed5a53ad2dd · outbound

This paper cites doi: 10.1145/3571730.http://dx.doi.org/10.1145/3571730.

HalluLens: LLM Hallucination Benchmark doi: 10.1145/3571730.http://dx.doi.org/10.1145/3571730

Reference 2023

Resolution
verified exact
doi, observed 2026-08-16T10:40:19.598496Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T10:40:19.305831Z digest=sha256:7e90d8999fd061d9129abef9f977ad3ad6966b34aa02603130332dab07b92450

Observation 173c253e-88c1-4787-bd80-78b20004cf94 · outbound

This paper cites correct”, “incorrect.

HalluLens: LLM Hallucination Benchmark correct”, “incorrect

Reference 2024

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:40:20.159736Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T10:40:19.358216Z digest=sha256:6cbc40b7d702887351254ca4eb23d00a5cb79141987e32f52e3389f82f5706c9

Pith citing papers

Observation 37fdbe5b-af66-4104-896a-fb76efdacb38 · inbound

Phare: A Safety Probe for Large Language Models cites this paper.

Phare: A Safety Probe for Large Language Models HalluLens: LLM Hallucination Benchmark

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-15T20:58:22.465675Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:58:22.465675Z digest=sha256:b2f5de7599cff716e970bea617430700eb29339a0b41f6ee6e5e8d9bf0ecd11c

Observation 6c6d2e74-fa0b-4a8b-bb16-c419cb0eefa0 · inbound

The Hallucination Tax of Reinforcement Finetuning cites this paper.

The Hallucination Tax of Reinforcement Finetuning HalluLens: LLM Hallucination Benchmark

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T15:44:30.304843Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:44:30.304843Z digest=sha256:c0890a43d3214e4798f4f83e81f6b8e85e189a38c47ac9c78c11538b8dd1a1d8

Observation 710392c0-8593-4199-9562-ece45783b8a7 · inbound

AbstentionBench: Reasoning LLMs Fail on Unanswerable Questions cites this paper.

AbstentionBench: Reasoning LLMs Fail on Unanswerable Questions HalluLens: LLM Hallucination Benchmark

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T05:01:08.125965Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:01:08.125965Z digest=sha256:7382b27e2e9bb5abcc119909cef925e713ec098304a01d833ae2588418c265d2

Observation 77e35e4d-2d50-473b-92a1-f8497ba24f36 · inbound

Machine Mirages: Defining the Undefined cites this paper.

Machine Mirages: Defining the Undefined HalluLens: LLM Hallucination Benchmark

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T11:24:29.959566Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:24:29.959566Z digest=sha256:fdd89ef09dc2af4e4ef1b914dc0aadf307064f3e18215e6058f3fe4afd4d590b

Observation dd8252b2-fc3d-4546-81ec-605e437dbb68 · inbound

Embodied AI Agents: Modeling the World cites this paper.

Embodied AI Agents: Modeling the World HalluLens: LLM Hallucination Benchmark

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-06T22:10:25.436137Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:10:25.436137Z digest=sha256:a35dc714c9ca719bf8ea0b35f349ef194592ee79e00375d14e7df1ec12e4c9db

Observation 37e98791-aba2-4f3b-9a78-ecf26545fd0c · inbound

Introducing the Swiss Food Knowledge Graph: AI for Context-Aware Nutrition Recommendation cites this paper.

Introducing the Swiss Food Knowledge Graph: AI for Context-Aware Nutrition Recommendation HalluLens: LLM Hallucination Benchmark

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T17:43:11.362519Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:43:11.362519Z digest=sha256:eb9255118057a1f7420793511108d8786d32c847609bcd47aea331a7b9dfd387

Observation c3bc9cd2-4707-482c-8568-42007d7edf06 · inbound

MIRAGE-Bench: LLM Agent is Hallucinating and Where to Find Them cites this paper.

MIRAGE-Bench: LLM Agent is Hallucinating and Where to Find Them HalluLens: LLM Hallucination Benchmark

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T13:06:36.599018Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:06:36.599018Z digest=sha256:da5ccae5877429babecc622a881a225a35e285b925c2d6f240357001bc620c86

Observation 5831c5ed-3040-4a23-a8c2-3a38c4a34cce · inbound

A comprehensive taxonomy of hallucinations in Large Language Models cites this paper.

A comprehensive taxonomy of hallucinations in Large Language Models HalluLens: LLM Hallucination Benchmark

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T05:29:09.780706Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T05:29:09.780706Z digest=sha256:a77d6f79b95e5c6c4f701b0cab231423d476cf4163b5b7e157f00232c4ad6640

Observation 08f453e8-d49c-4bdc-8a33-614d4f0addf4 · inbound

ReasoningTrack: Chain-of-Thought Reasoning for Long-term Vision-Language Tracking cites this paper.

ReasoningTrack: Chain-of-Thought Reasoning for Long-term Vision-Language Tracking HalluLens: LLM Hallucination Benchmark

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-05T23:33:19.965610Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:33:19.965610Z digest=sha256:c0c7adfccc651e52cdf41d1f5a00c344f921d89b7735bf4464d748f7bc6537db

Observation 1caab12d-ef18-4070-a689-d66bcbeb9b4a · inbound

GOSU: Retrieval-Augmented Generation with Global-Level Optimized Semantic Unit-Centric Framework cites this paper.

GOSU: Retrieval-Augmented Generation with Global-Level Optimized Semantic Unit-Centric Framework HalluLens: LLM Hallucination Benchmark

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-05T13:36:35.132035Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T13:36:35.132035Z digest=sha256:51f4952b0732cf34c321970d6611e342b575dc81e71b54da83990ae7474cb978

Observation 1e25fd92-76c3-4f24-9d4d-bd7323c38648 · inbound

ReFACT: A Benchmark for Scientific Confabulation Detection with Positional Error Annotations cites this paper.

ReFACT: A Benchmark for Scientific Confabulation Detection with Positional Error Annotations HalluLens: LLM Hallucination Benchmark

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-18T12:56:24.497247Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-05-18T12:54:01.717015Z digest=sha256:0a658845b8d2bcac2dbfaecf14ac406506d2c9f1f3b6d3d6c6f52d2245a74ca7

Observation eb0f60c0-8233-4d36-b187-45b6acb266dc · inbound

When Do Hallucinations Arise? A Graph Perspective on the Evolution of Path Reuse and Path Compression cites this paper.

When Do Hallucinations Arise? A Graph Perspective on the Evolution of Path Reuse and Path Compression HalluLens: LLM Hallucination Benchmark

Reference 1

Resolution
unresolved
no resolver link, observed 2026-07-13T13:07:00.928770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T13:07:00.928770Z digest=sha256:592ebe0a3b9c40bf52726853cb557bacdc494f5c242efdb6b24effa9a4549f2c

Observation 5cb32dba-21c3-4776-b09b-ea6890e3981c · inbound

From Binary Groundedness to Support Relations: Towards a Reader-Centred Taxonomy for Comprehension of AI Output cites this paper.

From Binary Groundedness to Support Relations: Towards a Reader-Centred Taxonomy for Comprehension of AI Output HalluLens: LLM Hallucination Benchmark

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-11T06:30:58.657451Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-10T17:37:28.129376Z digest=sha256:02c3503b64dc8136ae90511c229cc4dcc71286029ec1de87b0a84e04b8576b65

Observation 6be78a1e-743e-44e0-8a4b-e759714bc979 · inbound

HalluClear: Diagnosing, Evaluating and Mitigating Hallucinations in GUI Agents cites this paper.

HalluClear: Diagnosing, Evaluating and Mitigating Hallucinations in GUI Agents HalluLens: LLM Hallucination Benchmark

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-10T06:31:30.913700Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-10T06:27:19.717895Z digest=sha256:8a68d2d9839327e2a9bfacc1f5bf5f188edc8a0a5174c6d4278f156facc2751a

Observation fde6b7d0-6551-476f-9e37-e6c3b7aa2dd7 · inbound

Contrastive Attribution in the Wild: An Interpretability Analysis of LLM Failures on Realistic Benchmarks cites this paper.

Contrastive Attribution in the Wild: An Interpretability Analysis of LLM Failures on Realistic Benchmarks HalluLens: LLM Hallucination Benchmark

Reference 9

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T09:38:42.843693Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-10T05:12:31.218055Z digest=sha256:526446223b1fca526cf22c13d7819a7a88804388c06f1462721c0e0cb9fd38ef

Observation 58b08582-fad0-494c-bc2f-8345ec78c00c · inbound

Benchmarking Source-Sensitive Reasoning in Turkish: Humans and LLMs under Evidential Trust Manipulation cites this paper.

Benchmarking Source-Sensitive Reasoning in Turkish: Humans and LLMs under Evidential Trust Manipulation HalluLens: LLM Hallucination Benchmark

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-11T21:56:34.399829Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-08T03:42:06.876417Z digest=sha256:8151571ed7829fce8b34a558a63876b62f9254c452f50011b0461dd44b4a6e9c

Observation 6428576b-9a66-41e8-92e0-ce1a174edf42 · inbound

Hallucination as an Anomaly: Dynamic Intervention via Probabilistic Circuits cites this paper.

Hallucination as an Anomaly: Dynamic Intervention via Probabilistic Circuits HalluLens: LLM Hallucination Benchmark

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-05-11T19:51:11.274747Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-08T10:52:14.666333Z digest=sha256:1fae63904963fab495511575753eb46d2490db0610f285f71d7c7709f78822c3

Observation b96bea36-ddba-4ac7-a459-694aa94f0410 · inbound

StereoTales: A Multilingual Framework for Open-Ended Stereotype Discovery in LLMs cites this paper.

StereoTales: A Multilingual Framework for Open-Ended Stereotype Discovery in LLMs HalluLens: LLM Hallucination Benchmark

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-12T05:51:27.293122Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-12T04:50:17.399580Z digest=sha256:422bab138c258e96bd18bbbfee1454c280fb232ee9392c9a1bad0275bf03318c

Observation ce33085a-6e5d-4c06-9941-2928e6d6c9d7 · inbound

StereoTales: A Multilingual Framework for Open-Ended Stereotype Discovery in LLMs cites this paper.

StereoTales: A Multilingual Framework for Open-Ended Stereotype Discovery in LLMs HalluLens: LLM Hallucination Benchmark

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-13T07:22:28.981194Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-13T07:20:32.494840Z digest=sha256:fa19bec5866b3eda120eece3186eccf0f9d0bb1ff379c402b1b4238bc8de43db

Observation bd559f56-718a-48b3-aefb-91e6c6cd3c4c · inbound

REALISTA: Realistic Latent Adversarial Attacks that Elicit LLM Hallucinations cites this paper.

REALISTA: Realistic Latent Adversarial Attacks that Elicit LLM Hallucinations HalluLens: LLM Hallucination Benchmark

Reference 105

Resolution
verified exact
arxiv_id, observed 2026-05-14T20:17:56.743792Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-05-14T20:13:10.814899Z digest=sha256:a7786ba7fd3a92bb44dc2d1afb84d7db9211f0019f82419b40ff617e3a23b226

Observation a6760fd2-3410-4673-83a9-401894c769b2 · inbound

Do No Harm? Hallucination and Actor-Level Abuse in Web-Deployed Medical Large Language Models cites this paper.

Do No Harm? Hallucination and Actor-Level Abuse in Web-Deployed Medical Large Language Models HalluLens: LLM Hallucination Benchmark

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-21T05:53:59.416150Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-21T05:49:55.738870Z digest=sha256:751d744388449ebee202d0be6bac20e61743b10fc5198767ff785a7819cd3ad2

Observation ccf7b7c9-e403-42c9-b9d7-0bdaec1b7542 · inbound

K-FinHallu: A Hallucination Detection Benchmark for Multi-Turn RAG in Korean Finance cites this paper.

K-FinHallu: A Hallucination Detection Benchmark for Multi-Turn RAG in Korean Finance HalluLens: LLM Hallucination Benchmark

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-06-29T09:03:16.148279Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-06-29T08:54:48.807164Z digest=sha256:6bac68ed0220ad2d71681ec98194a5aa63dd866d21171e56d229dc1cbe736dfe

Observation b82edf81-756c-4f85-939b-097951a4d9e7 · inbound

Latent Performance Profiling of Large Language Models cites this paper.

Latent Performance Profiling of Large Language Models HalluLens: LLM Hallucination Benchmark

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-06-29T07:33:13.394631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-06-29T07:31:02.595386Z digest=sha256:f34f59995b03bfef217bdb39fd400434b78dc58b3af29568f02a43c76f348f07

Observation f3365005-c053-4dbb-9a77-89f8f2fcb14d · inbound

Disentangling Visual and Factual Correctness in LVLMs' Visualization Literacy cites this paper.

Disentangling Visual and Factual Correctness in LVLMs' Visualization Literacy HalluLens: LLM Hallucination Benchmark

Reference 41

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T02:06:26.198841Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-28T11:19:51.893740Z digest=sha256:e5194861fff2c00c4b49c1ab7bf7948efd1caa553ea84262f07d4319390f6a84

Observation 8026be1b-17af-41bf-9403-7f7421cd59c0 · inbound

What Do People Actually Want From AI? Mapping Preference Plurality cites this paper.

What Do People Actually Want From AI? Mapping Preference Plurality HalluLens: LLM Hallucination Benchmark

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-06-28T01:41:29.860815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-28T01:32:04.660400Z digest=sha256:dde6cb37390de3ce22bda1b7e87021a31f0345e496679190c7970b6d8b694cd5

Observation 00ed01a9-0708-42ef-9e46-ad803e5d813a · inbound

Contextualized Evaluation of Vision Language Models through Dynamic, Multi-turn Interactions cites this paper.

Contextualized Evaluation of Vision Language Models through Dynamic, Multi-turn Interactions HalluLens: LLM Hallucination Benchmark

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-02T01:57:59.043987Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T01:57:59.043987Z digest=sha256:69be503fd2e8ebd6482118bc37b776cf36c4c6bebac862cb7b8c1ad7f065743d