Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-06-30T23:59:10.797762Z
Paper Citation Record · LEDGER
As of 15 August 2026, this Paper Citation Record lists 12 of 12 outbound references and 1 inbound Pith citation observation for arXiv:2605.03217.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-06-30T23:59:10.797762Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-07-12T06:23:49.752475Z
A source-named dated measurement, never combined with another source.
Source: cited_works
12 of 12 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation ddfe04df-5058-4e9f-af23-17d4f1abb550 · outbound
Moral Sensitivity in LLMs: A Tiered Evaluation of Contextual Bias via Behavioral Profiling and Mechanistic Interpretability Constitutional AI: Harmlessness from AI Feedback
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 43afc6c4-41a1-4eed-81a8-7ec856d3dceb · outbound
Moral Sensitivity in LLMs: A Tiered Evaluation of Contextual Bias via Behavioral Profiling and Mechanistic Interpretability DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation a19b7205-130e-4674-9cc6-f64ec74858fc · outbound
Moral Sensitivity in LLMs: A Tiered Evaluation of Contextual Bias via Behavioral Profiling and Mechanistic Interpretability The Capacity for Moral Self-Correction in Large Language Models
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation dd8a28f2-da1c-40cd-97aa-e824a3b6d3c9 · outbound
Moral Sensitivity in LLMs: A Tiered Evaluation of Contextual Bias via Behavioral Profiling and Mechanistic Interpretability Unboxing Occupational Bias: Grounded Debiasing of LLMs with U.S. Labor Data
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 104a090f-2ae9-4b21-8b56-6c87f5d29f43 · outbound
Moral Sensitivity in LLMs: A Tiered Evaluation of Contextual Bias via Behavioral Profiling and Mechanistic Interpretability Distilling the Knowledge in a Neural Network
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation b8ab3419-7e5e-4d04-b8e3-c8148b5bc134 · outbound
Moral Sensitivity in LLMs: A Tiered Evaluation of Contextual Bias via Behavioral Profiling and Mechanistic Interpretability Language Model Alignment in Multilingual Trolley Problems
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 02e50204-3f55-461d-b8e3-f7c50349e26d · outbound
Moral Sensitivity in LLMs: A Tiered Evaluation of Contextual Bias via Behavioral Profiling and Mechanistic Interpretability XSTest: A Test Suite for Identifying Exaggerated Safety Behaviours in Large Language Models
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation dd831664-28de-433e-8fe1-996f2e523a09 · outbound
Moral Sensitivity in LLMs: A Tiered Evaluation of Contextual Bias via Behavioral Profiling and Mechanistic Interpretability Moral Mimicry: Large Language Models Produce Moral Rationalizations Tailored to Political Identity
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 17083016-82b3-424d-b2e6-7ecc06495f03 · outbound
Moral Sensitivity in LLMs: A Tiered Evaluation of Contextual Bias via Behavioral Profiling and Mechanistic Interpretability Our tiered evaluation framework and mechanistic analysis are designed to make model biases more transparent and auditable
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation c094d000-1a1b-410e-b937-565caa20ff37 · outbound
Moral Sensitivity in LLMs: A Tiered Evaluation of Contextual Bias via Behavioral Profiling and Mechanistic Interpretability Unresolved cited work
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 36010f41-fcf6-4a89-9d8c-c8d877e88884 · outbound
Moral Sensitivity in LLMs: A Tiered Evaluation of Contextual Bias via Behavioral Profiling and Mechanistic Interpretability Unresolved cited work
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 9a7aad71-2208-4a5f-8ad4-cdbae4830215 · outbound
Moral Sensitivity in LLMs: A Tiered Evaluation of Contextual Bias via Behavioral Profiling and Mechanistic Interpretability Unresolved cited work
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.
Observation 744e89d5-6068-46e3-ac0f-0634cc296711 · inbound
Where do LLMs Fall Short in CBT-Guided Affective Reasoning? Moral Sensitivity in LLMs: A Tiered Evaluation of Contextual Bias via Behavioral Profiling and Mechanistic Interpretability
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.