Pith. sign in

Paper Citation Record · LEDGER

ASSERT: Automated Safety Scenario Red Teaming for Evaluating the Robustness of Large Language Models

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2310.09624.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2310.09624 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 5 of 5 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T10:17:26.674036Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation fd341ccf-1802-4f7c-8bd4-19f107d7fd08 · inbound

The Scales of Justitia: A Comprehensive Survey on Safety Evaluation of LLMs cites this paper.

The Scales of Justitia: A Comprehensive Survey on Safety Evaluation of LLMs ASSERT: Automated Safety Scenario Red Teaming for Evaluating the Robustness of Large Language Models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T10:17:26.674036Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:17:26.674036Z digest=sha256:7a72519b56058093378a956898e527a9acb14c4b07241749b304c4057144a2f6

Observation 33e30a3e-e233-4d4b-9455-e98db6691ece · inbound

Agent Identity Evals: Measuring Agentic Identity cites this paper.

Agent Identity Evals: Measuring Agentic Identity ASSERT: Automated Safety Scenario Red Teaming for Evaluating the Robustness of Large Language Models

Reference 82

Resolution
unresolved
no resolver link, observed 2026-08-06T14:57:06.911179Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T14:57:06.911179Z digest=sha256:b2513b77f424f3174155e1b2b31e50ff1fc8fd798394d45d5247e0bad527ce19

Observation b415017e-7f86-4495-baa0-ebd2b6799c51 · inbound

PEEM: Prompt Engineering Evaluation Metrics for Interpretable Joint Evaluation of Prompts and Responses cites this paper.

PEEM: Prompt Engineering Evaluation Metrics for Interpretable Joint Evaluation of Prompts and Responses ASSERT: Automated Safety Scenario Red Teaming for Evaluating the Robustness of Large Language Models

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-05-15T14:00:03.067668Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-15T13:57:41.428695Z digest=sha256:2c8f0a77bb7edab5ee218183df7a2cd62996b0b3b3822045396ee76f5820f363

Observation 9dbb2ff4-f077-4d6d-ba59-234e0a5a5ee4 · inbound

Measuring Behavior Portability in Large Language Models cites this paper.

Measuring Behavior Portability in Large Language Models ASSERT: Automated Safety Scenario Red Teaming for Evaluating the Robustness of Large Language Models

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-06-26T09:09:16.270975Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-26T08:59:20.724868Z digest=sha256:538f1e74e0b202a033124b6c4337d5fbbb440513bf87c81cd676a151b5a61f98

Observation 00c93cdc-df59-47ba-bbb0-569dfd3da8a1 · inbound

Adversarial Stress Testing of Role-Playing Language Agents using Multi-Agent Evaluation cites this paper.

Adversarial Stress Testing of Role-Playing Language Agents using Multi-Agent Evaluation ASSERT: Automated Safety Scenario Red Teaming for Evaluating the Robustness of Large Language Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T00:29:55.943379Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:29:55.943379Z digest=sha256:bd73c8d8d230901c4e5a49121dd20a7e81148554d3e0615122f72f285c7957cf