Pith. sign in

Paper Citation Record · LEDGER

SoSBench: Benchmarking Safety Alignment on Six Scientific Domains

As of 11 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2505.21605.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.21605 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T15:14:22.859306Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-06-30T11:44:38.782098Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation c1dfdbb7-93ca-471b-a80f-efbeddb31a2d · inbound

Frontier AI Risk Management Framework in Practice: A Risk Analysis Technical Report cites this paper.

Frontier AI Risk Management Framework in Practice: A Risk Analysis Technical Report SoSBench: Benchmarking Safety Alignment on Six Scientific Domains

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T15:14:22.859306Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:14:22.859306Z digest=sha256:8bed32d371c2f232e90cd310d0a969a707767e1fcd6913580138934d9c964aed

Observation 2ee1e43c-0c4c-428b-befa-f507c57ff8af · inbound

Alignment and Safety in Large Language Models: Safety Mechanisms, Training Paradigms, and Emerging Challenges cites this paper.

Alignment and Safety in Large Language Models: Safety Mechanisms, Training Paradigms, and Emerging Challenges SoSBench: Benchmarking Safety Alignment on Six Scientific Domains

Reference 116

Resolution
unresolved
no resolver link, observed 2026-08-06T14:13:06.362805Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:13:06.362805Z digest=sha256:8786d01ea1fbd5b9bc96dc1b24da6368bfe2ab0f60a3964be0b90ef2c58116e9

Observation 355bc557-2a1e-4ccc-8523-646d060d7b69 · inbound

BadScientist: Can a Research Agent Write Convincing but Unsound Papers that Fool LLM Reviewers? cites this paper.

BadScientist: Can a Research Agent Write Convincing but Unsound Papers that Fool LLM Reviewers? SoSBench: Benchmarking Safety Alignment on Six Scientific Domains

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-04T08:59:46.952346Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T08:59:46.952346Z digest=sha256:f6f22fbcd27ae9bba4345e1058763a94c39357cb5a99f0dec13db3c141cd6bf4

Observation 6716e6ff-3368-4e39-8e49-7a5bc2dbdcbe · inbound

SafeSci: Safety Evaluation of Large Language Models in Science Domains and Beyond cites this paper.

SafeSci: Safety Evaluation of Large Language Models in Science Domains and Beyond SoSBench: Benchmarking Safety Alignment on Six Scientific Domains

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-05-15T18:01:25.616943Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-15T18:00:10.534346Z digest=sha256:05bf3f9cee075bb08c132618b76faea6d840d6288b8cf22abb8a08ab4e73a972

Observation 7210aaf0-1365-44fa-b328-d94281f84e69 · inbound

MultiBreak: A Scalable and Diverse Multi-turn Jailbreak Benchmark for Evaluating LLM Safety cites this paper.

MultiBreak: A Scalable and Diverse Multi-turn Jailbreak Benchmark for Evaluating LLM Safety SoSBench: Benchmarking Safety Alignment on Six Scientific Domains

Reference 42

Resolution
metadata mismatch
local_arxiv, observed 2026-05-11T09:31:01.236234Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-10T16:00:32.413225Z digest=sha256:429b7112d884cd750e185e63bffca108b2564fac22f4fbe107bfd4c0680025d9

Observation 492d9bf0-6eba-4355-b744-db3b29397e8c · inbound

Inverting the Shield: Systematically Generating Safety Tests from Policy Specifications cites this paper.

Inverting the Shield: Systematically Generating Safety Tests from Policy Specifications SoSBench: Benchmarking Safety Alignment on Six Scientific Domains

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-06-30T11:44:38.783278Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-06-30T11:35:46.622873Z digest=sha256:54ec84a32ec094200b8c2ef4c5fe31c4b08420dfcddbe814a80ae2006a2fab88