Pith. sign in

Paper Citation Record · LEDGER

S-Eval: Towards Automated and Comprehensive Safety Evaluation for Large Language Models

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 7 inbound Pith citation observations for arXiv:2405.14191.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2405.14191 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 7 of 7 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 7 of 7 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T10:17:26.844499Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T18:10:00.725685Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 246381f3-539f-4c3e-a52a-21eafb6c6b56 · inbound

From Local to Global: A Graph RAG Approach to Query-Focused Summarization cites this paper.

From Local to Global: A Graph RAG Approach to Query-Focused Summarization S-Eval: Towards Automated and Comprehensive Safety Evaluation for Large Language Models

Reference 74

Resolution
verified exact
arxiv_id, observed 2026-08-04T01:57:26.797777Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-11T05:10:57.816312Z digest=sha256:f78042c90e052e69a32a733dd6b9200d366e11d6c24459013d04a22bbf205efb

Observation d4ac6a66-df10-4627-9818-94933ba03a0e · inbound

The Scales of Justitia: A Comprehensive Survey on Safety Evaluation of LLMs cites this paper.

The Scales of Justitia: A Comprehensive Survey on Safety Evaluation of LLMs S-Eval: Towards Automated and Comprehensive Safety Evaluation for Large Language Models

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-07T10:17:26.844499Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:17:26.844499Z digest=sha256:b9ccb644844f9b2effb8550ff2c8dac3809a8972eae7e3084b7c1b527eed8d03

Observation 303efba4-6986-4882-ab4a-783b4a3ba1d2 · inbound

ORFuzz: Fuzzing the "Other Side" of LLM Safety -- Testing Over-Refusal cites this paper.

ORFuzz: Fuzzing the "Other Side" of LLM Safety -- Testing Over-Refusal S-Eval: Towards Automated and Comprehensive Safety Evaluation for Large Language Models

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-08-04T01:57:26.797777Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-18T23:27:52.438709Z digest=sha256:ae5d2990d3789ea075479d513c428c3449bbc0da38938f58212ad0f8f677457e

Observation d7650cbd-c31f-4368-ad84-0f9410e5f121 · inbound

Controlling the Risk of Corrupted Contexts for Language Models via Early-Exiting cites this paper.

Controlling the Risk of Corrupted Contexts for Language Models via Early-Exiting S-Eval: Towards Automated and Comprehensive Safety Evaluation for Large Language Models

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-04T12:45:26.905184Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T12:45:26.905184Z digest=sha256:6be26d8f8afd79e1d2441d4fd05d5cd606676413dc7e1f674de3bdcd7ca628df

Observation 1a5121e0-d0c8-4202-a219-9821cb75e50e · inbound

Multilingual Refusal Alignment for Safer Large Language Models cites this paper.

Multilingual Refusal Alignment for Safer Large Language Models S-Eval: Towards Automated and Comprehensive Safety Evaluation for Large Language Models

Reference 72

Resolution
verified exact
arxiv_id, observed 2026-08-04T01:57:26.797777Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-07-04T18:09:17.531727Z digest=sha256:be1d3662e05ebf3b48289b36ab3b9e7e90fcdb15056e8fffabc70ca220542451

Observation d42ba51f-292d-4d59-8f56-08e25b9b7ef9 · inbound

Yuvion LLM: An Adversarially-Aware Large Language Model for Content And AI Safety cites this paper.

Yuvion LLM: An Adversarially-Aware Large Language Model for Content And AI Safety S-Eval: Towards Automated and Comprehensive Safety Evaluation for Large Language Models

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-08-04T01:57:26.797777Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-29T00:46:03.210076Z digest=sha256:b1687e11f402244c57d0b6729191a50185d9697b7396640a1049c1c1d947af1b

Observation 077f18ff-e8be-4268-a724-7044051ea787 · inbound

When Refusal Looks Safe: The Refusal-Cue Shortcut in Safety Guard Models cites this paper.

When Refusal Looks Safe: The Refusal-Cue Shortcut in Safety Guard Models S-Eval: Towards Automated and Comprehensive Safety Evaluation for Large Language Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-05T23:46:20.542247Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T23:46:20.542247Z digest=sha256:2e92c985f97c8d445b8f375c2e3044bd1fe44eb552c44c46e213795f6d1cb955