Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T04:37:17.038006Z
Paper Citation Record · LEDGER
As of 18 August 2026, this Paper Citation Record lists 25 of 25 outbound references and 2 inbound Pith citation observations for arXiv:2501.17749.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T04:37:17.038006Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:13:50.508020Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-12T08:40:41.350151Z
25 of 25 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 00626506-bf78-4424-b696-1cffda6c7dca · outbound
Early External Safety Testing of OpenAI's o3-mini: Insights from the Pre-Deployment Evaluation SORRY-Bench: Systematically Evaluating Large Language Model Safety Refusal
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2b96bf4d-ed93-4bed-b67b-095f5ad09327 · outbound
Early External Safety Testing of OpenAI's o3-mini: Insights from the Pre-Deployment Evaluation S-Eval: Towards Automated and Comprehensive Safety Evaluation for Large Language Models
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 10b0be7e-f833-4089-a919-4ac79cdbfcff · outbound
Early External Safety Testing of OpenAI's o3-mini: Insights from the Pre-Deployment Evaluation SafetyBench: Evaluating the Safety of Large Language Models
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bb7698cd-626c-41ce-8617-d4761030ec04 · outbound
Early External Safety Testing of OpenAI's o3-mini: Insights from the Pre-Deployment Evaluation CHiSafetyBench: A Chinese Hierarchical Safety Benchmark for Large Language Models
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 842d4427-4229-4d9f-b061-38f3e911ef4c · outbound
Early External Safety Testing of OpenAI's o3-mini: Insights from the Pre-Deployment Evaluation SafeBench: A Safety Evaluation Framework for Multimodal Large Language Models
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d0ff1b51-4537-40f7-b0e4-8ba34959e8ea · outbound
Early External Safety Testing of OpenAI's o3-mini: Insights from the Pre-Deployment Evaluation LongSafety: Enhance Safety for Long-Context LLMs
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bb969384-9c6c-4882-b7fe-83ea091a5669 · outbound
Early External Safety Testing of OpenAI's o3-mini: Insights from the Pre-Deployment Evaluation SALAD-Bench: A Hierarchical and Comprehensive Safety Benchmark for Large Language Models
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a893c6c9-40be-435a-b83f-d26c4f08386d · outbound
Early External Safety Testing of OpenAI's o3-mini: Insights from the Pre-Deployment Evaluation Beavertails: Towards improved safety alignment of LLM via a human-preference dataset,
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 47aba5ba-9f30-4c3b-9b95-6928a2e9bb0b · outbound
Early External Safety Testing of OpenAI's o3-mini: Insights from the Pre-Deployment Evaluation SimpleSafetyTests: a Test Suite for Identifying Critical Safety Risks in Large Language Models
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cccad398-d7c5-4490-8bf7-dd5ffd2b1e11 · outbound
Early External Safety Testing of OpenAI's o3-mini: Insights from the Pre-Deployment Evaluation Astral: Automated safety testing of large language models,
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation d79e7cbe-e727-45df-bafb-bce88417f312 · outbound
Early External Safety Testing of OpenAI's o3-mini: Insights from the Pre-Deployment Evaluation A survey on metamorphic testing,
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 04dee86e-138f-4109-a9fe-b5931220627e · outbound
Early External Safety Testing of OpenAI's o3-mini: Insights from the Pre-Deployment Evaluation Guardrails for trust, safet y, and ethical development and deployment of large language models (llm),
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation d8295da2-1dd8-43a7-b7fa-3d9105a14fa2 · outbound
Early External Safety Testing of OpenAI's o3-mini: Insights from the Pre-Deployment Evaluation European Commission AI Act
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation f86c79da-9169-4d0b-8b5b-d75b3355f507 · outbound
Early External Safety Testing of OpenAI's o3-mini: Insights from the Pre-Deployment Evaluation Artificial Intelligence Act (Regulation (EU) 2024/16 89), Official Journal version of 13 June 2024
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 1cb77ed6-af48-4b3f-a363-3dafa0f2183a · outbound
Early External Safety Testing of OpenAI's o3-mini: Insights from the Pre-Deployment Evaluation Llama Guard: LLM-based Input-Output Safeguard for Human-AI Conversations
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5ab876b4-fd61-4419-aeff-ba967d05e0b7 · outbound
Early External Safety Testing of OpenAI's o3-mini: Insights from the Pre-Deployment Evaluation ShieldLM: Empowering LLMs as Aligned, Customizable and Explainable Safety Detectors
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 32549227-2140-4849-890a-5b7dd32a6123 · outbound
Early External Safety Testing of OpenAI's o3-mini: Insights from the Pre-Deployment Evaluation A StrongREJECT for Empty Jailbreaks
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0797e6cc-8eec-43f3-af8c-bedab0117b73 · outbound
Early External Safety Testing of OpenAI's o3-mini: Insights from the Pre-Deployment Evaluation Red Teaming Language Models to Reduce Harms: Methods, Scaling Behaviors, and Lessons Learned
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 90df0369-a45a-42dc-a089-0ad59a44f7b3 · outbound
Early External Safety Testing of OpenAI's o3-mini: Insights from the Pre-Deployment Evaluation Catastrophic Jailbreak of Open-source LLMs via Exploiting Generation
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation afba7695-fc6d-448e-bfff-2b153789b5f7 · outbound
Early External Safety Testing of OpenAI's o3-mini: Insights from the Pre-Deployment Evaluation Universal and Transferable Adversarial Attacks on Aligned Language Models
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 38e85f31-01f8-42c7-bad5-766efcf9985a · outbound
Early External Safety Testing of OpenAI's o3-mini: Insights from the Pre-Deployment Evaluation HarmBench: A Standardized Evaluation Framework for Automated Red Teaming and Robust Refusal
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1fd88780-7f4e-4dc4-a69b-db7ef0a9d757 · outbound
Early External Safety Testing of OpenAI's o3-mini: Insights from the Pre-Deployment Evaluation "Do Anything Now": Characterizing and Evaluating In-The-Wild Jailbreak Prompts on Large Language Models
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation befa1f1a-7b65-4894-9e48-d125dd22d8da · outbound
Early External Safety Testing of OpenAI's o3-mini: Insights from the Pre-Deployment Evaluation Jailbroken: H ow does llm safety training fail?,
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation f68c3631-350a-45ac-90f4-496ba9779108 · outbound
Early External Safety Testing of OpenAI's o3-mini: Insights from the Pre-Deployment Evaluation WalledEval: A Comprehensive Safety Evaluation Toolkit for Large Language Models
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7d4fb587-da85-4b7b-ac9f-b386a5e3ed18 · outbound
Early External Safety Testing of OpenAI's o3-mini: Insights from the Pre-Deployment Evaluation Bi-Factorial Preference Optimization: Balancing Safety-Helpfulness in Language Models
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f91ec8fc-2a41-4e0c-9663-6682edc638e3 · inbound
Towards Reasoning Era: A Survey of Long Chain-of-Thought for Reasoning Large Language Models Early External Safety Testing of OpenAI's o3-mini: Insights from the Pre-Deployment Evaluation
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 15ecdde5-f430-4616-8d8d-7541c6213c5d · inbound
Beyond Safe Answers: A Benchmark for Evaluating True Risk Awareness in Large Reasoning Models Early External Safety Testing of OpenAI's o3-mini: Insights from the Pre-Deployment Evaluation
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.