Pith. sign in

Paper Citation Record · LEDGER

LLMs as Factual Reasoners: Insights from Existing Benchmarks and Beyond

As of 17 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 8 inbound Pith citation observations for arXiv:2305.14540.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2305.14540 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 8 of 8 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 8 of 8 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T20:41:35.326319Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

13
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 6702f301-7c4a-4f00-8840-a587eca2640f · inbound

A Survey on Hallucination in Large Language Models: Principles, Taxonomy, Challenges, and Open Questions cites this paper.

A Survey on Hallucination in Large Language Models: Principles, Taxonomy, Challenges, and Open Questions LLMs as Factual Reasoners: Insights from Existing Benchmarks and Beyond

Reference 158

Resolution
verified exact
arxiv_id, observed 2026-05-13T02:46:27.667941Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-13T02:46:26.957539Z digest=sha256:d466ae8e1b68ff28c814f7a3cb6780436a0129f65e5a14a6bb0941fdf1e89bc7

Observation df5805a9-b5cd-4fc3-92bc-c560ea5a3497 · inbound

Evaluating LLM Reasoning in the Operations Research Domain with ORQA cites this paper.

Evaluating LLM Reasoning in the Operations Research Domain with ORQA LLMs as Factual Reasoners: Insights from Existing Benchmarks and Beyond

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-11T06:01:13.320631Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T06:01:13.320631Z digest=sha256:ef413afca1d50732e2b33117cd1c7e28688152acf4ba73976dad4f368f05d81e

Observation c3511464-e67b-49ca-8eb6-3d587f48818d · inbound

Towards Large Reasoning Models: A Survey of Reinforced Reasoning with Large Language Models cites this paper.

Towards Large Reasoning Models: A Survey of Reinforced Reasoning with Large Language Models LLMs as Factual Reasoners: Insights from Existing Benchmarks and Beyond

Reference 67

Resolution
verified exact
arxiv_id, observed 2026-05-15T21:20:59.266360Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-15T21:20:59.128986Z digest=sha256:6accfe2e89a66d1bc61c70722d7dd1f6589743d4f71c8974e591a28706c012c3

Observation 10913ab3-8fcc-49af-8f7d-920fa4def301 · inbound

Learning Auxiliary Tasks Improves Reference-Free Hallucination Detection in Open-Domain Long-Form Generation cites this paper.

Learning Auxiliary Tasks Improves Reference-Free Hallucination Detection in Open-Domain Long-Form Generation LLMs as Factual Reasoners: Insights from Existing Benchmarks and Beyond

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-15T20:41:35.326319Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T20:41:35.326319Z digest=sha256:7cdfec0e39b015c769052d65db8e7561a06affcd36786d386b34e73bde47357b

Observation 89316d04-92b6-48df-98e7-3dc5e15c8425 · inbound

HIDE and Seek: Detecting Hallucinations in Language Models via Decoupled Representations cites this paper.

HIDE and Seek: Detecting Hallucinations in Language Models via Decoupled Representations LLMs as Factual Reasoners: Insights from Existing Benchmarks and Beyond

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-15T19:08:01.769575Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T19:08:01.769575Z digest=sha256:04c24cd44b1cf00a9dc8de001bc472fa1274dba981499e30bff56e057b56cbbd

Observation 781b68b5-d02b-423b-b110-12113439cfb4 · inbound

DeepTRACE: Auditing Deep Research AI Systems for Tracking Reliability Across Citations and Evidence cites this paper.

DeepTRACE: Auditing Deep Research AI Systems for Tracking Reliability Across Citations and Evidence LLMs as Factual Reasoners: Insights from Existing Benchmarks and Beyond

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-05T12:11:34.208721Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:11:34.208721Z digest=sha256:ff9d81d099c4522efc2f9438e1af05e987c194ab4de80ae04e6538529e6472fc

Observation 16ce558f-cee6-4dc7-8347-4328044e005a · inbound

Resonant Context Anchoring: Decoupling Attention Routing and Signal Gain at Inference Time cites this paper.

Resonant Context Anchoring: Decoupling Attention Routing and Signal Gain at Inference Time LLMs as Factual Reasoners: Insights from Existing Benchmarks and Beyond

Reference 57

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T23:26:21.593970Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-06-28T14:27:00.523125Z digest=sha256:b7e6c6d8f4fb64caca14bd608b14f00b6c0c0306235219b54967f5b31c121866

Observation 88dc0c26-054b-4f11-9784-d7707957d96c · inbound

A French OSCE Dialogue Dataset and Controllable Virtual Patient System for Clinical Training cites this paper.

A French OSCE Dialogue Dataset and Controllable Virtual Patient System for Clinical Training LLMs as Factual Reasoners: Insights from Existing Benchmarks and Beyond

Reference 65

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T01:34:09.675660Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-06-30T01:29:19.877774Z digest=sha256:6b7e270baf6526340ede881781f5551588c21c3a8fd47be86f8a114cf26bea35