Pith. sign in

Paper Citation Record · LEDGER

BeHonest: Benchmarking Honesty in Large Language Models

As of 3 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 7 inbound Pith citation observations for arXiv:2406.13261.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2406.13261 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 7 of 7 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-03T06:30:56.289259+00:00

measured 7 of 7 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-30T10:45:44.127391Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T20:18:56.663568Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 75180596-08b1-4632-8046-36a663f80dea · inbound

Unlearners Can Lie: Evaluating and Improving Honesty in LLM Unlearning cites this paper.

Unlearners Can Lie: Evaluating and Improving Honesty in LLM Unlearning BeHonest: Benchmarking Honesty in Large Language Models

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T07:31:24.489114Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-12T02:47:10.818363Z digest=sha256:0e6c78775c5dee50171b946cb47b6434886fce597b3a25c91c1ff3c9b4fc32ba

Observation 13d213e0-542c-436b-8658-ed6f6f6dc5d4 · inbound

Navigating the Sea of LLM Evaluation: Investigating Bias in Toxicity Benchmarks cites this paper.

Navigating the Sea of LLM Evaluation: Investigating Bias in Toxicity Benchmarks BeHonest: Benchmarking Honesty in Large Language Models

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T05:31:23.935337Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-12T05:28:45.453455Z digest=sha256:1bfc6b7f4cf1a4041b7d2e5fdf400374423b91d69deddd837817e8e229b7e05f

Observation f48f2433-9f2c-487e-91e2-55c73e13ae0e · inbound

DECOR: Auditing LLM Deception via Information Manipulation Theory cites this paper.

DECOR: Auditing LLM Deception via Information Manipulation Theory BeHonest: Benchmarking Honesty in Large Language Models

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-20T06:28:05.381264Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-20T06:27:10.445757Z digest=sha256:67fc6210929e5edd10d202483a664ce5a09d60fbef82d35bca1e60afb2c0b670

Observation bf0bdbc0-8745-477a-894d-c71c971ee0a0 · inbound

SPADE-Bench: Evaluating Spontaneous Strategic Deception in Agents via Plan-Action Divergence cites this paper.

SPADE-Bench: Evaluating Spontaneous Strategic Deception in Agents via Plan-Action Divergence BeHonest: Benchmarking Honesty in Large Language Models

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T23:06:20.590107Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-06-28T14:39:51.672024Z digest=sha256:9d6d2520f1bf9f9f59e511cee48774d65f470215811ebe13464288fe634136e6

Observation cfbb8b37-c1c3-4263-bdd6-9f2c03bfb24b · inbound

SPADE-Bench: Evaluating Spontaneous Strategic Deception in Agents via Plan-Action Divergence cites this paper.

SPADE-Bench: Evaluating Spontaneous Strategic Deception in Agents via Plan-Action Divergence BeHonest: Benchmarking Honesty in Large Language Models

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T10:54:36.985886Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-06-30T10:45:44.127391Z digest=sha256:fce12b5d96b5c8d1dd89f8bfb98e196004e50d32d024e2605dd3be4603ff6271

Observation d33f246a-485b-4c20-9b3e-204721c35d77 · inbound

RogueAI: A Reverse Turing Test for Detecting Licensed AI Deception in Dialogue cites this paper.

RogueAI: A Reverse Turing Test for Detecting Licensed AI Deception in Dialogue BeHonest: Benchmarking Honesty in Large Language Models

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-07-03T15:18:33.688342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-06-27T06:34:39.457798Z digest=sha256:7b009ff5213909244ef2f0d5bc9c6959df302882d672ce79b5fc25546107dfac

Observation 05a3b2ca-af7e-45c6-bd80-db214128ad5e · inbound

Decoding Hidden Deception in Reasoning LLMs: Activation Explainers for Deception Auditing cites this paper.

Decoding Hidden Deception in Reasoning LLMs: Activation Explainers for Deception Auditing BeHonest: Benchmarking Honesty in Large Language Models

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T20:18:56.665646Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-06-27T01:28:38.810889Z digest=sha256:9d4d653402fb6b3ba74687e79c71767fbd01ae61165c9f5ab09e665778dea03b