Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T11:29:15.193096Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 18 of 18 outbound references and 1 inbound Pith citation observation for arXiv:2506.02372.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T11:29:15.193096Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T10:17:26.850823Z
A source-named dated measurement, never combined with another source.
Source: cited_works
18 of 18 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 258bae2d-697f-433d-a09f-efaae2cac962 · outbound
AnswerCarefully: A Dataset for Improving the Safety of Japanese LLM Output Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5d43c3a6-5a43-45d0-b0c8-a0efb2f51d51 · outbound
AnswerCarefully: A Dataset for Improving the Safety of Japanese LLM Output Bowman, Zac Hatfield-Dodds, Ben Mann, Dario Amodei, Nicholas Joseph, Sam McCandlish, Tom Brown, and Jared Kaplan
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 312bed32-3c6c-4f5a-a6d9-bba5dc19d8e0 · outbound
AnswerCarefully: A Dataset for Improving the Safety of Japanese LLM Output Llama Guard: LLM-based Input-Output Safeguard for Human-AI Conversations
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1bec56d0-8f72-4102-8c4a-a814d710fabb · outbound
AnswerCarefully: A Dataset for Improving the Safety of Japanese LLM Output Investigating tuning methods for achieving usefulness and safety for Japanese large language models (in Japanese)
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e95eafd8-5227-4877-85f6-b62b11c09b86 · outbound
AnswerCarefully: A Dataset for Improving the Safety of Japanese LLM Output Japanese safety boundary test for large language models (in Japanese)
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 0b3b0dce-8bbd-4034-baf7-21f25fdb39a8 · outbound
AnswerCarefully: A Dataset for Improving the Safety of Japanese LLM Output LLM-jp: A Cross-organizational Project for the Research and Development of Fully Open Japanese LLMs
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cfe4d581-d78d-43c8-b876-8dfb84542e94 · outbound
AnswerCarefully: A Dataset for Improving the Safety of Japanese LLM Output Construction of the Japanese TruthfulQA Dataset (in Japanese)
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7c5ba52e-7c82-4aa3-8f24-a59be2b46736 · outbound
AnswerCarefully: A Dataset for Improving the Safety of Japanese LLM Output Unresolved cited work
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 8e5070fc-ede0-47ef-968b-62f442e5007f · outbound
AnswerCarefully: A Dataset for Improving the Safety of Japanese LLM Output JSocialFact: a misinfor- mation dataset from social media for benchmarking LLM safety
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 50ad9775-4563-4de8-acd6-317acfea1bab · outbound
AnswerCarefully: A Dataset for Improving the Safety of Japanese LLM Output GPT-4 technical report, 2024
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e171592d-8d21-4485-b4d4-09aacb7b3792 · outbound
AnswerCarefully: A Dataset for Improving the Safety of Japanese LLM Output Large-scale human evaluation of LLM safety (in Japanese)
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation be382a7c-2b54-43c8-ac68-79a9c93bbd81 · outbound
AnswerCarefully: A Dataset for Improving the Safety of Japanese LLM Output Gemini: A family of highly capable multimodal models, 2024
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation da3c5d52-8b6b-47f7-b631-89c647c403d6 · outbound
AnswerCarefully: A Dataset for Improving the Safety of Japanese LLM Output Llama 2: Open foundation and fine-tuned chat models, 2023
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 36eafbe0-8e45-4677-b670-6bfbb183a6df · outbound
AnswerCarefully: A Dataset for Improving the Safety of Japanese LLM Output Do-Not-Answer: Evaluating safeguards in LLMs
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 27c19c1a-5550-4496-9d84-6c4786339b2c · outbound
AnswerCarefully: A Dataset for Improving the Safety of Japanese LLM Output A Chinese Dataset for Evaluating the Safeguards in Large Language Models
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6bff4682-60b8-46d1-b28a-b4bfe646b41c · outbound
AnswerCarefully: A Dataset for Improving the Safety of Japanese LLM Output JBBQ: Japanese Bias Benchmark for Analyzing Social Biases in Large Language Models
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 76600452-782e-4b5d-94a5-c31bbd0e8567 · outbound
AnswerCarefully: A Dataset for Improving the Safety of Japanese LLM Output Xing, Hao Zhang, Joseph E
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b6abdcf6-0478-4e02-806e-adfe4b6e71f7 · outbound
AnswerCarefully: A Dataset for Improving the Safety of Japanese LLM Output Unresolved cited work
Reference 184
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 5cf06d1a-4497-4622-81af-3c946d84e8ef · inbound
The Scales of Justitia: A Comprehensive Survey on Safety Evaluation of LLMs AnswerCarefully: A Dataset for Improving the Safety of Japanese LLM Output
Reference 85
Source-reported events for the cited work
Unavailable: canonical work link unavailable.