Pith. sign in

Paper Citation Record · LEDGER

ASSERT: Automated Safety Scenario Red Teaming for Evaluating the Robustness of Large Language Models

As of 22 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2310.09624.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2310.09624 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 5 of 5 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T10:17:26.674036Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation fd341ccf-1802-4f7c-8bd4-19f107d7fd08 · inbound

The Scales of Justitia: A Comprehensive Survey on Safety Evaluation of LLMs cites this paper.

The Scales of Justitia: A Comprehensive Survey on Safety Evaluation of LLMs ASSERT: Automated Safety Scenario Red Teaming for Evaluating the Robustness of Large Language Models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T10:17:26.674036Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:17:26.674036Z digest=sha256:7cc403fc48245bdb267ebb207d1b61be51b6a0ce75e22e94619c4b5fe1f7931a

Observation 33e30a3e-e233-4d4b-9455-e98db6691ece · inbound

Agent Identity Evals: Measuring Agentic Identity cites this paper.

Agent Identity Evals: Measuring Agentic Identity ASSERT: Automated Safety Scenario Red Teaming for Evaluating the Robustness of Large Language Models

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-06T14:57:06.911179Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:57:06.911179Z digest=sha256:e86bf81d320ec792e10dd4d987038bf8052964dcda8ed100d6b401fa4f7f4090

Observation b415017e-7f86-4495-baa0-ebd2b6799c51 · inbound

PEEM: Prompt Engineering Evaluation Metrics for Interpretable Joint Evaluation of Prompts and Responses cites this paper.

PEEM: Prompt Engineering Evaluation Metrics for Interpretable Joint Evaluation of Prompts and Responses ASSERT: Automated Safety Scenario Red Teaming for Evaluating the Robustness of Large Language Models

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-05-15T14:00:03.067668Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-15T13:57:41.428695Z digest=sha256:2a1c82d473651aa0f143b3d24a815f95a872a85da0c5e85f15b0d2ee89148710

Observation 9dbb2ff4-f077-4d6d-ba59-234e0a5a5ee4 · inbound

Measuring Behavior Portability in Large Language Models cites this paper.

Measuring Behavior Portability in Large Language Models ASSERT: Automated Safety Scenario Red Teaming for Evaluating the Robustness of Large Language Models

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-06-26T09:09:16.270975Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-06-26T08:59:20.724868Z digest=sha256:2d9e9bd327dfd72283020b7c4bd87aa18e64555d94a7d0d2f41dfdead2e1e244

Observation 00c93cdc-df59-47ba-bbb0-569dfd3da8a1 · inbound

Adversarial Stress Testing of Role-Playing Language Agents using Multi-Agent Evaluation cites this paper.

Adversarial Stress Testing of Role-Playing Language Agents using Multi-Agent Evaluation ASSERT: Automated Safety Scenario Red Teaming for Evaluating the Robustness of Large Language Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T00:29:55.943379Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:29:55.943379Z digest=sha256:2393e9e9497b8190f151404c29fc55117c0a37d054b73a8b63e9bea4485ca995