Pith. sign in

Paper Citation Record · LEDGER

Long-form factuality in large language models

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 19 inbound Pith citation observations for arXiv:2403.18802.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2403.18802 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 19 of 19 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 19 of 19 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:34:07.814259Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-30T06:44:18.784391Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 1efbd665-8e61-4f59-9a39-601f5debd3a3 · inbound

Preserving Knowledge in Large Language Model with Model-Agnostic Self-Decompression cites this paper.

Preserving Knowledge in Large Language Model with Model-Agnostic Self-Decompression Long-form factuality in large language models

Reference 57

Resolution
verified exact
arxiv_id, observed 2026-05-24T00:13:39.462264Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-24T00:09:52.093810Z digest=sha256:5f7a2cebe22b45acdae19d448d93e9324d5322e44a4db0535da9efcc404b2e46

Observation 06b1d0f2-7278-4747-b188-f7c739200a4b · inbound

Measuring short-form factuality in large language models cites this paper.

Measuring short-form factuality in large language models Long-form factuality in large language models

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-15T06:45:50.288896Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-15T06:45:50.219157Z digest=sha256:bf63b0bd38b2ca459d60ac57c687d5503e5b83b15fea521078e8b986871c3d47

Observation 93e1b83e-2620-44ce-86b5-de939e95e8df · inbound

HD-NDEs: Neural Differential Equations for Hallucination Detection in LLMs cites this paper.

HD-NDEs: Neural Differential Equations for Hallucination Detection in LLMs Long-form factuality in large language models

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-07T12:34:07.814259Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:34:07.814259Z digest=sha256:c05deeee0eb6aeedac29e85ad362eb9d1bc71006048c1865035fd4820a06589e

Observation c2dcf01b-4e35-4276-95f5-0e034b3270a4 · inbound

RealFactBench: A Benchmark for Evaluating Large Language Models in Real-World Fact-Checking cites this paper.

RealFactBench: A Benchmark for Evaluating Large Language Models in Real-World Fact-Checking Long-form factuality in large language models

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T00:51:47.184863Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:51:47.184863Z digest=sha256:c350bbaba351254c6f4d72bf59b28dbf3b631ad37407ac08bf846d5970544040

Observation 0f995fdc-e477-4550-9282-17b2f2db3e86 · inbound

Veracity: An Open-Source AI Fact-Checking System cites this paper.

Veracity: An Open-Source AI Fact-Checking System Long-form factuality in large language models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T23:53:39.856611Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:53:39.856611Z digest=sha256:3f7e7c2a56a6d7c7106368d42048fb972be9e1501cdfb4cd6c20b26342ef7d51

Observation 2fce3b73-9d30-4887-902c-f4726afa870e · inbound

Can External Validation Tools Improve Annotation Quality for LLM-as-a-Judge? cites this paper.

Can External Validation Tools Improve Annotation Quality for LLM-as-a-Judge? Long-form factuality in large language models

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T15:02:21.746960Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T15:02:21.746960Z digest=sha256:7361e0ab686c778696f818f7fc630be1197a4e0be49ac7c0e83d38c19296a5d8

Observation 9e1a384a-f6b2-4472-b578-acd638b17247 · inbound

JADES: A Universal Framework for Jailbreak Assessment via Decompositional Scoring cites this paper.

JADES: A Universal Framework for Jailbreak Assessment via Decompositional Scoring Long-form factuality in large language models

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-05T14:51:03.793719Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T14:51:03.793719Z digest=sha256:9eb17a0708f5fc06c5c40c4b5bf524005239ecdff2e2858612cc7ab907f55ac8

Observation cf27c228-ad09-4a55-bc2a-242a563f198e · inbound

DecMetrics: Structured Claim Decomposition Scoring for Factually Consistent LLM Outputs cites this paper.

DecMetrics: Structured Claim Decomposition Scoring for Factually Consistent LLM Outputs Long-form factuality in large language models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-05T13:18:40.851754Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T13:18:40.851754Z digest=sha256:377ca158fd73cf917c0dc104cb2d6f6d1d5e8fa428836cbf1765b6137d525c10

Observation 0da5c309-6527-475d-86f9-86c56defb8bc · inbound

Human-AI Complementarity: A Goal for Amplified Oversight cites this paper.

Human-AI Complementarity: A Goal for Amplified Oversight Long-form factuality in large language models

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-04T07:21:22.548978Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:21:22.548978Z digest=sha256:4ebcf7658184ecb776d18227f41b652e09329439c56606c143d5f72c6afde728

Observation 2829bc8c-14c0-46d3-8f85-8ca0e8a248ff · inbound

All Leaks Count, Some Count More: Interpretable Temporal Contamination Detection and Mitigation in LLM Backtesting cites this paper.

All Leaks Count, Some Count More: Interpretable Temporal Contamination Detection and Mitigation in LLM Backtesting Long-form factuality in large language models

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-02T22:21:36.837859Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T22:21:36.837859Z digest=sha256:4bc451a724d0df288a2b0f69ae65e481a0cf218a0d81478990fd9f07029a3a25

Observation b23702a5-67fc-4a5b-992c-45f6d364fca5 · inbound

Beyond Precision: Importance-Aware Recall for Factuality Evaluation in Long-Form LLM Generation cites this paper.

Beyond Precision: Importance-Aware Recall for Factuality Evaluation in Long-Form LLM Generation Long-form factuality in large language models

Reference 16

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T20:23:13.860761Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T20:20:16.299929Z digest=sha256:0464d28be8e76e17f8987cf77c857ebfd1ea2fdeac796f493014d890ef610853

Observation 69ce7a89-cf1a-4c6a-9687-9ed399fd4f38 · inbound

VerifAI: A Verifiable Open-Source Search Engine for Biomedical Question Answering cites this paper.

VerifAI: A Verifiable Open-Source Search Engine for Biomedical Question Answering Long-form factuality in large language models

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-05-16T14:07:58.734248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-16T14:03:58.706599Z digest=sha256:0f44f6a67b0984f127c60f5db9cfdad22e797ed1fc3e4743d32919a8892be8b0

Observation e771731a-56e5-4628-ae6f-3b69b81154bb · inbound

Answer Only as Precisely as Justified: Calibrated Claim-Level Specificity Control for Agentic Systems cites this paper.

Answer Only as Precisely as Justified: Calibrated Claim-Level Specificity Control for Agentic Systems Long-form factuality in large language models

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-10T05:36:02.238245Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T05:31:11.718051Z digest=sha256:5c082237d1297a44a1682341edf337093a6c83fd1636accb16161dbc56b71c81

Observation 1257812d-bb5c-4dab-a9a8-fe38ec1b1653 · inbound

Answer Only as Precisely as Justified: Calibrated Claim-Level Specificity Control for Agentic Systems cites this paper.

Answer Only as Precisely as Justified: Calibrated Claim-Level Specificity Control for Agentic Systems Long-form factuality in large language models

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-20T23:59:15.228903Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-20T23:56:44.979471Z digest=sha256:8f023b47e4c0f1560867d36fec50004bdd47a0a31c892be5db72208582a18d81

Observation 3133827c-6439-48ad-bf6a-4851398e89b2 · inbound

FinGround: Detecting and Grounding Financial Hallucinations via Atomic Claim Verification cites this paper.

FinGround: Detecting and Grounding Financial Hallucinations via Atomic Claim Verification Long-form factuality in large language models

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T21:11:20.023407Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-08T06:24:43.021602Z digest=sha256:fda0106fc83263fe244c1169eff225cd9a6e274bc2bd7677857af0a514de4645

Observation 7ed87e6d-64d2-4283-816f-b245dccae29e · inbound

StereoTales: A Multilingual Framework for Open-Ended Stereotype Discovery in LLMs cites this paper.

StereoTales: A Multilingual Framework for Open-Ended Stereotype Discovery in LLMs Long-form factuality in large language models

Reference 115

Resolution
verified exact
arxiv_id, observed 2026-05-12T05:51:27.235175Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-12T04:50:17.399580Z digest=sha256:212c8873850f8ee75f66fb27ab395c30d3eb0d2f7e98c332cd18ac8e54de0e79

Observation b24693e0-49c3-4527-9bf8-95a34920c2dd · inbound

StereoTales: A Multilingual Framework for Open-Ended Stereotype Discovery in LLMs cites this paper.

StereoTales: A Multilingual Framework for Open-Ended Stereotype Discovery in LLMs Long-form factuality in large language models

Reference 115

Resolution
verified exact
arxiv_id, observed 2026-05-13T07:22:28.962869Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T07:20:32.494840Z digest=sha256:e0afb6ab79ae72f2fe90cb133de90493345e158a610a58db047e718aa184a22c

Observation cb1d8c87-bea7-45d8-8671-28a21c766c1a · inbound

Can LLM-as-a-Judge Reliably Verify Rubrics in Agentic Scenarios? cites this paper.

Can LLM-as-a-Judge Reliably Verify Rubrics in Agentic Scenarios? Long-form factuality in large language models

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-06-30T06:44:18.785878Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-30T06:42:30.226214Z digest=sha256:760a4f9dbac76756c1f158ee597904290465e00d98a1243c47b2bdb3a6771945

Observation dc54ef0c-6cf5-47bc-b6ae-eea8cc20e1f2 · inbound

Answer-Reconstruction Search Density: Measuring the Query and Source Work Compressed by Conversational Answers cites this paper.

Answer-Reconstruction Search Density: Measuring the Query and Source Work Compressed by Conversational Answers Long-form factuality in large language models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-01T14:03:17.719461Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T14:03:17.719461Z digest=sha256:47b792cdfb8a86e119e6f66e5aa12d9da0005f07796dc34bbeb32b6b4708b763