Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-19T20:31:20.017866Z
Paper Citation Record · LEDGER
As of 18 August 2026, this Paper Citation Record lists 69 of 69 outbound references and 2 inbound Pith citation observations for arXiv:2605.17007.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-19T20:31:20.017866Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-15T14:47:41.319887Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-15T14:47:41.656143Z
69 of 69 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation ee1713f0-de1a-4a51-ae93-0f7774c7255f · outbound
HalluScore: Large Language Model Hallucination Question Answering Benchmark On faithfulness and factuality in ab- stractive summarization
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 0e6f4f07-d000-4691-b095-07c8c1432a89 · outbound
HalluScore: Large Language Model Hallucination Question Answering Benchmark A survey on hallucination in large lan- guage models: Principles, taxonomy, challenges, and open questions
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 7d517dd1-47c4-4289-adcb-c391a79f178f · outbound
HalluScore: Large Language Model Hallucination Question Answering Benchmark Arahallueval: A fine-grained hallucination evaluation framework for arabic llms
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 3bfb508f-7658-48e4-a9fd-daa32f797051 · outbound
HalluScore: Large Language Model Hallucination Question Answering Benchmark Survey of hallucination in natural language generation
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 108dd1fb-4c12-4896-ab06-e2f25628da6e · outbound
HalluScore: Large Language Model Hallucination Question Answering Benchmark A sur- vey of automatic hallucination evaluation on natural language generation
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 35812170-7057-4218-88ab-839f62a898d6 · outbound
HalluScore: Large Language Model Hallucination Question Answering Benchmark Large language models hallucination: A comprehen- sive survey
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation f77be92e-deb7-487a-b3c0-821ec7192995 · outbound
HalluScore: Large Language Model Hallucination Question Answering Benchmark ALLaM: Large Language Models for Arabic and English
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 7008bd6d-bab8-4001-bf65-c191ed138861 · outbound
HalluScore: Large Language Model Hallucination Question Answering Benchmark Jais and Jais-chat: Arabic-Centric Foundation and Instruction-Tuned Open Generative Large Language Models
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 45a8a431-084d-4c3a-ac63-156a3986a408 · outbound
HalluScore: Large Language Model Hallucination Question Answering Benchmark Fanar: An Arabic-Centric Multimodal Generative AI Platform
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation e68b2474-68db-4288-b396-c2ce47cf277f · outbound
HalluScore: Large Language Model Hallucination Question Answering Benchmark A Survey of Large Language Models for Arabic Language and its Dialects
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 088dd01e-ee3d-4d35-a82d-d45eeb4446c8 · outbound
HalluScore: Large Language Model Hallucination Question Answering Benchmark Evaluating ara- bic large language models: A survey of bench- marks, methods, and gaps
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation a2f8ef85-b251-4744-b345-e0ce56ab6cf1 · outbound
HalluScore: Large Language Model Hallucination Question Answering Benchmark Arabic natural language processing: Challenges and solutions
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation fa137c4c-29f5-405b-a223-7f7c7c73a60c · outbound
HalluScore: Large Language Model Hallucination Question Answering Benchmark Unresolved cited work
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 52f550ee-f6e0-45c0-9f34-c3ea298bd781 · outbound
HalluScore: Large Language Model Hallucination Question Answering Benchmark Halwasa: Quantify and analyze halluci- nations in large language models: Arabic as a case study
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation c8ef8c2d-32c9-4c8d-8b7a-c92dc04dbc7d · outbound
HalluScore: Large Language Model Hallucination Question Answering Benchmark HalluVerse25: Fine-grained Multilingual Benchmark Dataset for LLM Hallucinations
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 4284bdd8-4407-484c-9fb7-0cd5b53c122f · outbound
HalluScore: Large Language Model Hallucination Question Answering Benchmark Aftina: enhanc- ing stability and preventing hallucination in ai- based islamic fatwa generation using llms and rag
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation e66ab6ac-9dcf-47de-a318-875b8627cc5f · outbound
HalluScore: Large Language Model Hallucination Question Answering Benchmark Islamiceval2025: The first shared task of capturing llms hallucination in islamic content
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation e244d198-3317-4186-b063-28b3a535b664 · outbound
HalluScore: Large Language Model Hallucination Question Answering Benchmark HaluEval: A Large-Scale Hallucination Evaluation Benchmark for Large Language Models
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 9f6835c7-52c9-4767-84e8-c9248efd34ee · outbound
HalluScore: Large Language Model Hallucination Question Answering Benchmark Analyzing llm behavior in dialogue sum- marization: Unveiling circumstantial hallucina- tion trends
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 765d1820-4d41-431a-975f-1118eabfb459 · outbound
HalluScore: Large Language Model Hallucination Question Answering Benchmark Evaluating Hallucinations in Chinese Large Language Models
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation fcce44a8-d1c1-457a-9085-c44a0e70c1d4 · outbound
HalluScore: Large Language Model Hallucination Question Answering Benchmark Uhgeval: Benchmarking the hallucina- tion of chinese large language models via uncon- strained generation
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 45e0857c-6403-4820-ae6a-321d78502b29 · outbound
HalluScore: Large Language Model Hallucination Question Answering Benchmark C-faith: A chinese fine-grained benchmark for automated halluci- nation evaluation
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation e822f75a-c0eb-41d2-83d6-e060f18f6d17 · outbound
HalluScore: Large Language Model Hallucination Question Answering Benchmark Retrieve only when it needs: Adap- tive retrieval augmentation for hallucination mitigation in large language models
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 732f6da4-14ad-472f-b31c-39a3f6efdcf3 · outbound
HalluScore: Large Language Model Hallucination Question Answering Benchmark Exploring rag solu- tions to reduce hallucinations in llms
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 32431627-2742-4a0c-8a81-000a72677b30 · outbound
HalluScore: Large Language Model Hallucination Question Answering Benchmark Detecting hallucinations in large language models using semantic entropy
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 05e7868a-31f7-48fa-8916-8205e410034e · outbound
HalluScore: Large Language Model Hallucination Question Answering Benchmark En- hancing uncertainty-based hallucination detec- tion with stronger focus
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation fb2d3bcd-8392-4535-b3a3-ca96d6c546d0 · outbound
HalluScore: Large Language Model Hallucination Question Answering Benchmark Detecting and mitigating hallucinations in machine translation: Model internal workings alone do well, sentence similarity even better
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation d8485e3d-3a11-4ee5-a43e-0d9518e2f22c · outbound
HalluScore: Large Language Model Hallucination Question Answering Benchmark Leveraging graph structures to de- tect hallucinations in large language models
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 9dea6609-a06f-45f6-bf2d-38569e678140 · outbound
HalluScore: Large Language Model Hallucination Question Answering Benchmark Halugnn: Hallucination detection in large language models using graph neural network
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 915cc98f-6616-4985-9cd7-50b99aeb8595 · outbound
HalluScore: Large Language Model Hallucination Question Answering Benchmark Hallushift: Measuring distribution shifts towards hallucination detection in llms
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 344b517f-b600-4b25-8dd4-66fd0219f26f · outbound
HalluScore: Large Language Model Hallucination Question Answering Benchmark Selfcheck- gpt: Zero-resource black-box hallucination de- tection for generative large language models
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 6dcab520-1454-4c62-bed1-9e2c31e52141 · outbound
HalluScore: Large Language Model Hallucination Question Answering Benchmark Sac3: reliable hallucination detection in black- box language models via semantic-aware cross- check consistency
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 4bab83ab-45ee-4848-829b-7c2ecc20ba49 · outbound
HalluScore: Large Language Model Hallucination Question Answering Benchmark Ai- generated news articles based on large language models
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation ee8fde59-8d33-4e06-a247-8251f52e8388 · outbound
HalluScore: Large Language Model Hallucination Question Answering Benchmark Self- expertise: knowledge-based instruction dataset augmentation for a legal expert language model
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 8fcd387a-6c98-440b-898b-bd3952538b60 · outbound
HalluScore: Large Language Model Hallucination Question Answering Benchmark A survey on rag with llms
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 9eb4c08a-852f-4144-9344-6876065e7ad8 · outbound
HalluScore: Large Language Model Hallucination Question Answering Benchmark Chain- of-thought prompting elicits reasoning in large language models
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 4bc90e49-ac48-4db5-b679-254a5ab2acd8 · outbound
HalluScore: Large Language Model Hallucination Question Answering Benchmark Chain- of-verification reduces hallucination in large lan- guage models
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation da7a8fee-6ddd-4625-89c9-ec350d4eaa90 · outbound
HalluScore: Large Language Model Hallucination Question Answering Benchmark Mitigating Large Language Model Hallucination with Faithful Finetuning
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation a2be840f-0364-4a9b-8c33-d35a1f38f977 · outbound
HalluScore: Large Language Model Hallucination Question Answering Benchmark DoLa: Decoding by Contrasting Layers Improves Factuality in Large Language Models
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 80be3781-54ac-46db-9c66-78c9bffd56bc · outbound
HalluScore: Large Language Model Hallucination Question Answering Benchmark Feqa: A ques- tion answering evaluation framework for faith- fulness assessment in abstractive summariza- tion
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 3493b1ca-a1f0-4307-b9c2-6b0ae8d5f3a8 · outbound
HalluScore: Large Language Model Hallucination Question Answering Benchmark Evaluating the factual consistency of abstractive text summarization
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 9795a8df-c8a7-4829-b5a5-a710bb42fcd2 · outbound
HalluScore: Large Language Model Hallucination Question Answering Benchmark Factscore: Fine-grained atomic evalu- ation of factual precision in long form text gen- eration
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation cec35b2f-6dfc-457a-8da9-9672e421a43d · outbound
HalluScore: Large Language Model Hallucination Question Answering Benchmark How reliable are automatic eval- uation methods for instruction-tuned llms?
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation c8fa316e-4275-4c73-af19-1907bff5e92e · outbound
HalluScore: Large Language Model Hallucination Question Answering Benchmark Truthfulqa: Measuring how models mimic human false- hoods
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 50996ebd-c92f-4f9b-bf00-ef85b2626835 · outbound
HalluScore: Large Language Model Hallucination Question Answering Benchmark Freshllms: Refreshing large language models with search engine augmentation
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 8b55cd7f-f84b-48a9-9cba-e2c77c9beaab · outbound
HalluScore: Large Language Model Hallucination Question Answering Benchmark Assessing the factual accuracy of generated text
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 1f16aa9a-a40c-42ca-84b3-35d368f2d36d · outbound
HalluScore: Large Language Model Hallucination Question Answering Benchmark Generativeaiforislamictexts: Theeman framework for mitigating gpt hallucinations
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 111d5b67-34b1-4a61-8fed-4a6cfd66c4f5 · outbound
HalluScore: Large Language Model Hallucination Question Answering Benchmark Mitigating llm hal- lucinations in quranic content: An agentic ap- proach using deployable language models
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 112953d5-8999-4a6e-8327-c6b27314b492 · outbound
HalluScore: Large Language Model Hallucination Question Answering Benchmark SemEval-2025 Task 3: Mu-SHROOM, the Multilingual Shared Task on Hallucinations and Related Observable Overgeneration Mistakes
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation ed6af985-1c9e-4b1f-9015-3dbe75261a1c · outbound
HalluScore: Large Language Model Hallucination Question Answering Benchmark Halomi: A manually anno- tated benchmark for multilingual hallucination and omission detection in machine translation
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 6ed730fe-c1c8-41ec-81ea-d1f6ad11c87c · outbound
HalluScore: Large Language Model Hallucination Question Answering Benchmark Poly-FEVER: A Multilingual Fact Verification Benchmark for Hallucination Detection in Large Language Models
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation a71aa15b-e1ec-4b42-bc2d-a1b438c8d02f · outbound
HalluScore: Large Language Model Hallucination Question Answering Benchmark Hot- potqa: A dataset for diverse, explainable multi- hop question answering
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 16475eef-9b6b-49f9-a441-84d848b51268 · outbound
HalluScore: Large Language Model Hallucination Question Answering Benchmark Triviaqa: A large scale distantly super- vised challenge dataset for reading comprehen- sion
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 52f2babf-f010-475f-a35c-a432d35dbf70 · outbound
HalluScore: Large Language Model Hallucination Question Answering Benchmark Medhallu: A comprehen- sive benchmark for detecting medical hallucina- tions in large language models
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 77476ed3-a272-4f28-b306-c70fd2f09690 · outbound
HalluScore: Large Language Model Hallucination Question Answering Benchmark Defan: Definitive answer dataset for llm hallucination evaluation
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 82dba318-3575-4609-bf53-d8c0cc031e17 · outbound
HalluScore: Large Language Model Hallucination Question Answering Benchmark Naseej launches its innovative arabic ai language model “noon
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 29b5d582-a1e3-40b0-a1d5-a4903d8d0e0a · outbound
HalluScore: Large Language Model Hallucination Question Answering Benchmark Introducing claude sonnet 4.5
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 50ba481a-9c98-4ea7-a466-81b453853f01 · outbound
HalluScore: Large Language Model Hallucination Question Answering Benchmark DeepSeek-V3 Technical Report
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 0fa784af-b4e4-439f-8aac-f74c763ccd72 · outbound
HalluScore: Large Language Model Hallucination Question Answering Benchmark [Online]
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 3a169845-6b2f-4138-b956-fd2368894682 · outbound
HalluScore: Large Language Model Hallucination Question Answering Benchmark GPT-4 Technical Report
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation c0031f15-fee3-425e-aa18-f2d29992e9be · outbound
HalluScore: Large Language Model Hallucination Question Answering Benchmark OpenAI GPT-5 System Card
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 066260b9-1c48-4b2d-869a-edb5f4840def · outbound
HalluScore: Large Language Model Hallucination Question Answering Benchmark Llama-4-maverick- 17b-128e-instruct-fp8
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation b6a81e73-62c6-41fa-ba3e-6c588ff67bca · outbound
HalluScore: Large Language Model Hallucination Question Answering Benchmark Qwen3-next- 80b-a3b-instruct
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation dc8e0658-9aad-4438-8202-da09e65ad465 · outbound
HalluScore: Large Language Model Hallucination Question Answering Benchmark Qwen3-235b-a22b-instruct-2507-fp8
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 23745810-8085-4d9d-a51b-68c60228742a · outbound
HalluScore: Large Language Model Hallucination Question Answering Benchmark System card: Claude opus 4 and claude sonnet 4
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 9f7e693e-a8f3-4b50-b1a6-9efe2974ab9a · outbound
HalluScore: Large Language Model Hallucination Question Answering Benchmark DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 833ce9a3-d8ec-4e64-af6d-0e093ef7b48d · outbound
HalluScore: Large Language Model Hallucination Question Answering Benchmark Openai o3 and o4-mini system card
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 5ca92cf6-8b2b-4d57-9b75-5b8505d8f55e · outbound
HalluScore: Large Language Model Hallucination Question Answering Benchmark The double-edged sword of anthro- pomorphism in llms
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation eb2336b8-516e-4607-a6d2-847058e4d86a · outbound
HalluScore: Large Language Model Hallucination Question Answering Benchmark Breaking the illusion: Revisiting llm anthropomorphism
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation aa867811-6cb4-460f-a4e5-cced2d308534 · inbound
HalluTruthQA: A Fine-Grained Benchmark for Hallucination Detection, Localization, and Explanation in Arabic Question Answering HalluScore: Large Language Model Hallucination Question Answering Benchmark
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 376be2f7-2ebf-42f3-a560-69c2d37534d2 · inbound
HalluTruthQA-4K: A Fine-Grained Corpus and Annotation Process for Arabic Hallucination Detection and Truth Verification HalluScore: Large Language Model Hallucination Question Answering Benchmark
Reference 2025
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.