Pith. sign in

Paper Citation Record · LEDGER

BELLS: A Framework Towards Future Proof Benchmarks for the Evaluation of LLM Safeguards

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 3 inbound Pith citation observations for arXiv:2406.01364.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2406.01364 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 3 of 3 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T20:25:52.920792Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-14T01:35:51.330444Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 72162509-06f2-40a5-ba8c-49f06d77bcb9 · inbound

AgentHarm: A Benchmark for Measuring Harmfulness of LLM Agents cites this paper.

AgentHarm: A Benchmark for Measuring Harmfulness of LLM Agents BELLS: A Framework Towards Future Proof Benchmarks for the Evaluation of LLM Safeguards

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-14T01:35:51.347845Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-05-14T01:35:50.992477Z digest=sha256:c9d7b07675081b5aea9af965867c91d95945324219f9da52551f19de8b162b38

Observation a3afa071-68f8-45ef-a72a-5d790d7b0fef · inbound

Is Reasoning All You Need? Probing Bias in the Age of Reasoning Language Models cites this paper.

Is Reasoning All You Need? Probing Bias in the Age of Reasoning Language Models BELLS: A Framework Towards Future Proof Benchmarks for the Evaluation of LLM Safeguards

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T20:25:52.920792Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:25:52.920792Z digest=sha256:59c47c5dec4162002a5157ac05e5a0216df138908fc3b82abf8e452370cfc5ef

Observation a7645fa9-ea0d-47a0-a6de-365ab87a2eb1 · inbound

SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses cites this paper.

SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses BELLS: A Framework Towards Future Proof Benchmarks for the Evaluation of LLM Safeguards

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-04T09:25:39.537797Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:25:39.537797Z digest=sha256:61037cd50e23872a71c7cceff4c187eae55f6bcfe15ff838997b1e1a144630cb