Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-04T10:04:08.124917Z
Paper Citation Record · LEDGER
As of 22 August 2026, this Paper Citation Record lists 15 of 15 outbound references and 2 inbound Pith citation observations for arXiv:2510.12229.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-04T10:04:08.124917Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-02T22:20:53.464810Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-09T05:46:01.579330Z
15 of 15 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 19963716-f3d0-4db2-a663-672329b7f802 · outbound
Analysing Moral Bias in Finetuned LLMs through Mechanistic Interpretability Addressing Moral Uncertainty using Large Language Models for Ethical Decision-Making
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a334bba8-d244-413f-9e79-27b7ee69afae · outbound
Analysing Moral Bias in Finetuned LLMs through Mechanistic Interpretability Mistral 7B
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 61b60131-a68b-46c5-9592-7dc3bb7f12f7 · outbound
Analysing Moral Bias in Finetuned LLMs through Mechanistic Interpretability URL https://distill.pub/ 2020/circuits/zoom-in/
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e1b90ac2-af4b-4981-91d6-2c43468eb6e7 · outbound
Analysing Moral Bias in Finetuned LLMs through Mechanistic Interpretability Can LLMs Simulate Human Behavioral Variability? A Case Study in the Phonemic Fluency Task
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d4d23b16-2f90-42ed-a028-165c546272cf · outbound
Analysing Moral Bias in Finetuned LLMs through Mechanistic Interpretability Sarfati, Y ., Hardy-Bayl´e, M
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9563ab4b-bcc7-4fbd-b8be-149b0fb75df1 · outbound
Analysing Moral Bias in Finetuned LLMs through Mechanistic Interpretability Gemma 2: Improving Open Language Models at a Practical Size
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 425f9994-2666-456d-9267-2c710b28a798 · outbound
Analysing Moral Bias in Finetuned LLMs through Mechanistic Interpretability Taxonomy of risks posed by lan- guage models
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 457e262e-5331-48a7-ba88-877c2af58e10 · outbound
Analysing Moral Bias in Finetuned LLMs through Mechanistic Interpretability acl-long.893/
Reference 893
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f1e57758-7181-4be4-8d6a-e94627284b89 · outbound
Analysing Moral Bias in Finetuned LLMs through Mechanistic Interpretability Could a Large Language Model be Conscious?
Reference 2005
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1741f6ed-a444-4df9-9ef3-1a013553e5a7 · outbound
Analysing Moral Bias in Finetuned LLMs through Mechanistic Interpretability Towards Best Practices of Activation Patching in Language Models: Metrics and Methods
Reference 2010
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ec1a5368-4d93-46ef-9348-bc03d92d5fcf · outbound
Analysing Moral Bias in Finetuned LLMs through Mechanistic Interpretability The Llama 3 Herd of Models
Reference 2013
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0a68d1b7-cf90-46d8-9ab4-3e89cbdff978 · outbound
Analysing Moral Bias in Finetuned LLMs through Mechanistic Interpretability M., Gebru, T., McMillan-Major, A., and Shmitchell, S
Reference 2014
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 59110f0d-e0b7-48c3-95af-bfa9a5a84f07 · outbound
Analysing Moral Bias in Finetuned LLMs through Mechanistic Interpretability doi: 10.18653/v1/2023.acl-long
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f5192ad6-599f-412c-8d58-0d48776bb6ed · outbound
Analysing Moral Bias in Finetuned LLMs through Mechanistic Interpretability Interpreting Bias in Large Language Models: A Feature-Based Approach
Reference 2024
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6a2ef335-0406-48f7-87ce-ccb2a49563dd · outbound
Analysing Moral Bias in Finetuned LLMs through Mechanistic Interpretability Dissecting Bias in LLMs: A Mechanistic Interpretability Perspective
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f702573e-df5f-4169-b928-02847faba96c · inbound
Mechanistic Interpretability of Cognitive Complexity in LLMs via Linear Probing using Bloom's Taxonomy Analysing Moral Bias in Finetuned LLMs through Mechanistic Interpretability
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 55550e9d-3148-46e8-95a7-d0387ade3479 · inbound
User identity conditions moral wrongness ratings in non-reasoning large language models Analysing Moral Bias in Finetuned LLMs through Mechanistic Interpretability
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.