Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T13:32:50.913865Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 61 of 61 outbound references and 0 inbound Pith citation observations for arXiv:2505.21409.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T13:32:50.913865Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
61 of 61 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 07e3c137-1528-4fef-8fd3-8ee46e10ba6c · outbound
RelationalFactQA: A Benchmark for Evaluating Tabular Fact Retrieval from Large Language Models Ai hallucination report 2025, 2025
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation db2fe7a2-deb1-4651-8dfb-b2face751cf2 · outbound
RelationalFactQA: A Benchmark for Evaluating Tabular Fact Retrieval from Large Language Models Unresolved cited work
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f386ed11-ef70-4e89-b24e-978300c46ede · outbound
RelationalFactQA: A Benchmark for Evaluating Tabular Fact Retrieval from Large Language Models Balsiger, H.-R
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 09c7ceea-687d-4aab-841f-bb57859f6cac · outbound
RelationalFactQA: A Benchmark for Evaluating Tabular Fact Retrieval from Large Language Models Interpretable Medical Diagnostics with Structured Data Extraction by Large Language Models
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 83d20bd7-532e-4110-9a9a-e04daa29a71b · outbound
RelationalFactQA: A Benchmark for Evaluating Tabular Fact Retrieval from Large Language Models Borisov, K
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5a333f0c-374c-44d2-8ba9-77c3db137e51 · outbound
RelationalFactQA: A Benchmark for Evaluating Tabular Fact Retrieval from Large Language Models Buoncristiano, G
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 21db787f-3b63-472b-a4e7-8305a2cc2a45 · outbound
RelationalFactQA: A Benchmark for Evaluating Tabular Fact Retrieval from Large Language Models Retrieve, Merge, Predict: Augmenting Tables with Data Lakes
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c3536f04-1b0a-4238-b799-71090ed6f0d8 · outbound
RelationalFactQA: A Benchmark for Evaluating Tabular Fact Retrieval from Large Language Models Chang, X
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3de3a363-6858-4b22-bd6b-b316ecd28949 · outbound
RelationalFactQA: A Benchmark for Evaluating Tabular Fact Retrieval from Large Language Models Unresolved cited work
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 539ed863-a7ca-4ace-bee0-2aa5697e9122 · outbound
RelationalFactQA: A Benchmark for Evaluating Tabular Fact Retrieval from Large Language Models Chowdhury, D
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6ebb0201-619e-4392-9502-01af1f438bcc · outbound
RelationalFactQA: A Benchmark for Evaluating Tabular Fact Retrieval from Large Language Models Christophides, V
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9b9c87c3-d87c-4b35-8027-caa2498e6fc6 · outbound
RelationalFactQA: A Benchmark for Evaluating Tabular Fact Retrieval from Large Language Models DeepSeek LLM: Scaling Open-Source Language Models with Longtermism
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ccc29672-a4f9-4b2f-8a84-b3a2504f780e · outbound
RelationalFactQA: A Benchmark for Evaluating Tabular Fact Retrieval from Large Language Models Unresolved cited work
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation df98fbb8-4333-4889-92e7-52e433d66a7b · outbound
RelationalFactQA: A Benchmark for Evaluating Tabular Fact Retrieval from Large Language Models Elnashar, J
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 08c5554f-d8b2-4bf7-9373-0fa7ba694972 · outbound
RelationalFactQA: A Benchmark for Evaluating Tabular Fact Retrieval from Large Language Models doi: https://doi.org/10.1016/j.accinf.2024.100715
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 04667e40-d325-417c-be23-747a66d0cc9d · outbound
RelationalFactQA: A Benchmark for Evaluating Tabular Fact Retrieval from Large Language Models Gemma: Open Models Based on Gemini Research and Technology
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9f9808ce-865c-42ec-99b6-36da7c895896 · outbound
RelationalFactQA: A Benchmark for Evaluating Tabular Fact Retrieval from Large Language Models Unresolved cited work
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c55372c2-1aaa-46f7-83f6-e76287c276b7 · outbound
RelationalFactQA: A Benchmark for Evaluating Tabular Fact Retrieval from Large Language Models Holtzman, J
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5d98f628-b3f6-4e59-9b1d-f2e583163ebb · outbound
RelationalFactQA: A Benchmark for Evaluating Tabular Fact Retrieval from Large Language Models Glavic, G
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1f775385-3eef-40b5-8c87-0d5b04d31c49 · outbound
RelationalFactQA: A Benchmark for Evaluating Tabular Fact Retrieval from Large Language Models Unresolved cited work
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7648c6b5-6d7b-4221-a791-bd58a2912d55 · outbound
RelationalFactQA: A Benchmark for Evaluating Tabular Fact Retrieval from Large Language Models Unresolved cited work
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8f6d8c24-3092-4da0-9839-90c9010a4939 · outbound
RelationalFactQA: A Benchmark for Evaluating Tabular Fact Retrieval from Large Language Models Joshi, E
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f9bc9594-3b18-4636-a527-fdaf7ec9b2e3 · outbound
RelationalFactQA: A Benchmark for Evaluating Tabular Fact Retrieval from Large Language Models Mistral 7B
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 07134d5f-adf9-4e76-b9f3-13bc1c1d65d0 · outbound
RelationalFactQA: A Benchmark for Evaluating Tabular Fact Retrieval from Large Language Models Open-WikiTable: Dataset for Open Domain Question Answering with Complex Reasoning over Table
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 75b5265f-a770-4f4b-a55e-786304332ccc · outbound
RelationalFactQA: A Benchmark for Evaluating Tabular Fact Retrieval from Large Language Models Kwiatkowski, J
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d849b4bf-17e1-4e8a-a17c-0de972ad3745 · outbound
RelationalFactQA: A Benchmark for Evaluating Tabular Fact Retrieval from Large Language Models Kasneci, K
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 03ba16d3-e65d-4c38-9153-14c06ef62a34 · outbound
RelationalFactQA: A Benchmark for Evaluating Tabular Fact Retrieval from Large Language Models Unresolved cited work
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation ef23b778-e2ee-4279-ad2b-b2f3286ff541 · outbound
RelationalFactQA: A Benchmark for Evaluating Tabular Fact Retrieval from Large Language Models Unresolved cited work
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8047c490-164c-4fb7-b9dd-bd8b8615ba34 · outbound
RelationalFactQA: A Benchmark for Evaluating Tabular Fact Retrieval from Large Language Models Lee, M.-W
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d711835e-1280-4b37-bf07-db934020e050 · outbound
RelationalFactQA: A Benchmark for Evaluating Tabular Fact Retrieval from Large Language Models Unresolved cited work
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 72d61015-d0fe-4da7-b5f8-f561bddc1362 · outbound
RelationalFactQA: A Benchmark for Evaluating Tabular Fact Retrieval from Large Language Models Unresolved cited work
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2b4bc770-9318-4679-b948-ecb4b16afbbc · outbound
RelationalFactQA: A Benchmark for Evaluating Tabular Fact Retrieval from Large Language Models The Llama 3 Herd of Models
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 48022fea-0bfc-464d-a74c-451b1da09a50 · outbound
RelationalFactQA: A Benchmark for Evaluating Tabular Fact Retrieval from Large Language Models Unresolved cited work
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 203906d0-b638-40db-9095-63eb26094c58 · outbound
RelationalFactQA: A Benchmark for Evaluating Tabular Fact Retrieval from Large Language Models Papicchio, P
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 3ee3bf0f-0a9c-4f6c-b505-6038d0c9fac2 · outbound
RelationalFactQA: A Benchmark for Evaluating Tabular Fact Retrieval from Large Language Models Pasupat and P
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 53540648-f8a8-493c-8e1f-b972c4829785 · outbound
RelationalFactQA: A Benchmark for Evaluating Tabular Fact Retrieval from Large Language Models Semantic Operators: A Declarative Model for Rich, AI-based Data Processing
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4bb87158-d39f-489e-8709-8416759587c1 · outbound
RelationalFactQA: A Benchmark for Evaluating Tabular Fact Retrieval from Large Language Models GPT-4 Technical Report
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7e8e62e6-53d7-420c-93e4-800ae8fd1de9 · outbound
RelationalFactQA: A Benchmark for Evaluating Tabular Fact Retrieval from Large Language Models Petroni, P
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 346a7052-f8cf-4ad8-902a-20fa0f0b3bd2 · outbound
RelationalFactQA: A Benchmark for Evaluating Tabular Fact Retrieval from Large Language Models Qwen2.5 Technical Report
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f39c3868-ee30-4293-9d2e-30ac78f08a52 · outbound
RelationalFactQA: A Benchmark for Evaluating Tabular Fact Retrieval from Large Language Models Unresolved cited work
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5617f6dd-a755-4803-bf1c-fa10dbccc684 · outbound
RelationalFactQA: A Benchmark for Evaluating Tabular Fact Retrieval from Large Language Models Petroni, T
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 214e4505-e483-4ee3-a224-aaef144eca94 · outbound
RelationalFactQA: A Benchmark for Evaluating Tabular Fact Retrieval from Large Language Models Saparina and M
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f2628c97-c9c0-4033-b7cd-b815f8e49dd5 · outbound
RelationalFactQA: A Benchmark for Evaluating Tabular Fact Retrieval from Large Language Models DocETL: Agentic Query Rewriting and Evaluation for Complex Document Processing
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5393e0aa-2a2f-4186-b6ab-b3c38c2baa2f · outbound
RelationalFactQA: A Benchmark for Evaluating Tabular Fact Retrieval from Large Language Models Singhal, S
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 56c35b50-fe33-4bc5-ace0-e2f4d064e867 · outbound
RelationalFactQA: A Benchmark for Evaluating Tabular Fact Retrieval from Large Language Models Stuhler, C
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 31cc0abf-f788-4a7e-8436-331998aaef12 · outbound
RelationalFactQA: A Benchmark for Evaluating Tabular Fact Retrieval from Large Language Models Saeed, N
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 462e6e3c-9e7f-4112-be47-7ce49f70b95c · outbound
RelationalFactQA: A Benchmark for Evaluating Tabular Fact Retrieval from Large Language Models Unresolved cited work
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f584738f-6c6d-449e-bf4f-c33a73286a87 · outbound
RelationalFactQA: A Benchmark for Evaluating Tabular Fact Retrieval from Large Language Models Truhn, J
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation acbd60b6-dba9-4d11-8652-d52550417321 · outbound
RelationalFactQA: A Benchmark for Evaluating Tabular Fact Retrieval from Large Language Models Measuring short-form factuality in large language models
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ffa750af-20cd-4e63-9098-d18d57002d8a · outbound
RelationalFactQA: A Benchmark for Evaluating Tabular Fact Retrieval from Large Language Models Unresolved cited work
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 80ac46d7-739e-4dfd-a43a-8510c0dbccc0 · outbound
RelationalFactQA: A Benchmark for Evaluating Tabular Fact Retrieval from Large Language Models Unresolved cited work
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fa26c866-5deb-493d-8ea9-74b2814f7e04 · outbound
RelationalFactQA: A Benchmark for Evaluating Tabular Fact Retrieval from Large Language Models YAGO 4.5: A Large and Clean Knowledge Base with a Rich Taxonomy
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 51a3bf34-1a99-41c8-be65-de9f2fc3be9c · outbound
RelationalFactQA: A Benchmark for Evaluating Tabular Fact Retrieval from Large Language Models Unresolved cited work
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1f7e146c-038d-4741-8944-43da98534dda · outbound
RelationalFactQA: A Benchmark for Evaluating Tabular Fact Retrieval from Large Language Models Seq2SQL: Generating Structured Queries from Natural Language using Reinforcement Learning
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 85c57c1d-5f7a-4778-b8c9-00cb62d2b7bc · outbound
RelationalFactQA: A Benchmark for Evaluating Tabular Fact Retrieval from Large Language Models bird”, “galois
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c30f6445-753b-4c35-bd49-80e5b3e8ccc1 · outbound
RelationalFactQA: A Benchmark for Evaluating Tabular Fact Retrieval from Large Language Models Unresolved cited work
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6bff2db9-7609-4480-b209-367f9cc1daba · outbound
RelationalFactQA: A Benchmark for Evaluating Tabular Fact Retrieval from Large Language Models TableLLM: Enabling Tabular Data Manipulation by LLMs in Real Office Usage Scenarios
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 88eec023-3341-49c6-969a-3536460f28ae · outbound
RelationalFactQA: A Benchmark for Evaluating Tabular Fact Retrieval from Large Language Models Unresolved cited work
Reference 2020
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c5c90338-bdf0-448d-867b-c46c29e78773 · outbound
RelationalFactQA: A Benchmark for Evaluating Tabular Fact Retrieval from Large Language Models TruthfulQA: Measuring How Models Mimic Human Falsehoods
Reference 2022
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 23a283f7-c1e8-4d6f-9126-2fc6563a52d4 · outbound
RelationalFactQA: A Benchmark for Evaluating Tabular Fact Retrieval from Large Language Models Mistral 7B
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 29306f15-a00a-40f5-8bf4-dd5123b49722 · outbound
RelationalFactQA: A Benchmark for Evaluating Tabular Fact Retrieval from Large Language Models doi: 10.3390/computers13100257
Reference 2024
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
No inbound Pith citation observations are available.