Pith. sign in

Paper Citation Record · LEDGER

The Capacity for Moral Self-Correction in Large Language Models

As of 18 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 43 inbound Pith citation observations for arXiv:2302.07459.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2302.07459 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 43 of 43 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 43 of 43 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T12:24:59.539102Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-09T21:36:34.279815Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 149d118c-5cdb-4547-b1bb-da70e3050926 · inbound

Language Models can Solve Computer Tasks cites this paper.

Language Models can Solve Computer Tasks The Capacity for Moral Self-Correction in Large Language Models

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-17T12:17:26.788090Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-17T12:17:26.602361Z digest=sha256:29263832eb3e8ecceccb76dd3aed504f5b38df3fe3d28214ed5ed738696dbcbb

Observation 07864920-420d-4a91-83b8-9b9fcc771be8 · inbound

Teaching Large Language Models to Self-Debug cites this paper.

Teaching Large Language Models to Self-Debug The Capacity for Moral Self-Correction in Large Language Models

Reference 91

Resolution
verified exact
arxiv_id, observed 2026-05-12T06:24:24.808331Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-05-12T06:24:24.607354Z digest=sha256:e6e6e694b14c83176d3544ccf82991a1ad039ece3689388f21665320c4540f42

Observation 0b9cd473-741f-43d5-97be-ca4f9c6a8e13 · inbound

A Comprehensive Overview of Large Language Models cites this paper.

A Comprehensive Overview of Large Language Models The Capacity for Moral Self-Correction in Large Language Models

Reference 176

Resolution
verified exact
arxiv_id, observed 2026-05-19T20:28:39.562148Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-19T20:28:38.900026Z digest=sha256:ba16166d494d5ecb4dc38f6e78ed9bd246ebdd2cf431c17f5a8bb900bd4a8213

Observation 0ae7da65-0ca0-476e-b3f9-c8ebc989590e · inbound

Cognitive Architectures for Language Agents cites this paper.

Cognitive Architectures for Language Agents The Capacity for Moral Self-Correction in Large Language Models

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-16T19:33:44.333626Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-16T19:33:44.146134Z digest=sha256:93923e5092c952944cea8f34c3c393e9ac8b3f84a071a7887b307b0812e926b7

Observation 74654157-bd1a-4329-bf0f-bdbc71872afc · inbound

Large Language Models as Optimizers cites this paper.

Large Language Models as Optimizers The Capacity for Moral Self-Correction in Large Language Models

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-15T00:04:31.261649Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-15T00:04:31.212102Z digest=sha256:1e6708b3554f7cf3b463025535332cd8996aa0412df8f3026850afe7c6b0a278

Observation 1c830fb9-5ae1-4ecd-b80d-22d042172d50 · inbound

Large Language Models Cannot Self-Correct Reasoning Yet cites this paper.

Large Language Models Cannot Self-Correct Reasoning Yet The Capacity for Moral Self-Correction in Large Language Models

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-12T05:48:24.649228Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-12T05:48:24.133045Z digest=sha256:83e5f34732509262d43779be277f88d0106dfef5136b4b3e6487e03ac8987385

Observation ca8d4468-837a-4cf2-9522-c86a72686373 · inbound

Low-Resource Languages Jailbreak GPT-4 cites this paper.

Low-Resource Languages Jailbreak GPT-4 The Capacity for Moral Self-Correction in Large Language Models

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-17T09:24:13.993100Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-17T09:24:13.911401Z digest=sha256:ce48ee18062a4858061dd2ca9f03ddc88d723a4bba219354cb9fdcb9b6046011

Observation d2c72841-ded1-483c-8ec5-bda52fefa590 · inbound

LiveCodeBench: Holistic and Contamination Free Evaluation of Large Language Models for Code cites this paper.

LiveCodeBench: Holistic and Contamination Free Evaluation of Large Language Models for Code The Capacity for Moral Self-Correction in Large Language Models

Reference 149

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T17:34:42.734237Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-05-10T17:34:42.565806Z digest=sha256:6edadf35d460a6134cc1e14ad406aa126fc5fd7ce8a3caba193bfb6dae18e8ab

Observation d94a8a51-2420-4ff4-85fa-80d1531e32e6 · inbound

Inference Scaling Laws: An Empirical Analysis of Compute-Optimal Inference for Problem-Solving with Language Models cites this paper.

Inference Scaling Laws: An Empirical Analysis of Compute-Optimal Inference for Problem-Solving with Language Models The Capacity for Moral Self-Correction in Large Language Models

Reference 256

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T06:38:37.056783Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-05-18T06:38:36.517935Z digest=sha256:03712231ac5f5a4a2e0978f676e6a014743cfdbd78b3db838aebe443a209a617

Observation 6dedd3d2-c6c3-436a-84b4-af3b8e919535 · inbound

Large Language Models Can Self-Improve in Long-context Reasoning cites this paper.

Large Language Models Can Self-Improve in Long-context Reasoning The Capacity for Moral Self-Correction in Large Language Models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-12T21:59:11.403148Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T21:59:11.403148Z digest=sha256:7e565b483957c8033fde94635b9e1c19343468a017cebadd3001f60b42e8a011

Observation e3e7b37d-073b-4f93-b286-7d73d84a736d · inbound

Biased or Flawed? Mitigating Stereotypes in Generative Language Models by Addressing Task-Specific Flaws cites this paper.

Biased or Flawed? Mitigating Stereotypes in Generative Language Models by Addressing Task-Specific Flaws The Capacity for Moral Self-Correction in Large Language Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-11T15:02:33.198617Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T15:02:33.198617Z digest=sha256:276759ce9375de66d821a6ff8d1f9f6cc15e55c25424c4f4b0971012c11a7734

Observation 53314db9-ef6e-4ce6-a33e-593c3817681b · inbound

Rethinking Uncertainty Estimation in LLMs: A Principled Single-Sequence Measure cites this paper.

Rethinking Uncertainty Estimation in LLMs: A Principled Single-Sequence Measure The Capacity for Moral Self-Correction in Large Language Models

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-23T06:22:38.178394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-23T06:22:25.496269Z digest=sha256:3a6e49ab03f3277085500d9848d18dce8c4ca9c2795df0c339f6b3c5a29e92d9

Observation c5cb5273-24f3-43b9-a98e-2e78ea4f55eb · inbound

Explicit vs. Implicit: Investigating Social Bias in Large Language Models through Self-Reflection cites this paper.

Explicit vs. Implicit: Investigating Social Bias in Large Language Models through Self-Reflection The Capacity for Moral Self-Correction in Large Language Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-10T22:18:05.680475Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T22:18:05.680475Z digest=sha256:dbc3aa30411f59a0bb6364f628fca9af12373bf41b6147fec0969536006c94c2

Observation 63bb554f-d214-4413-ba0e-1629fc133caf · inbound

Foundations of Large Language Models cites this paper.

Foundations of Large Language Models The Capacity for Moral Self-Correction in Large Language Models

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-10T20:14:58.768223Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T20:14:58.768223Z digest=sha256:8d673ecb3096d9c2ca5f719a91915e11219adfa2ee58bb401566e118034190a7

Observation 6bfb490b-80d9-430e-b81b-1cc57021624d · inbound

Examining Alignment of Large Language Models through Representative Heuristics: The Case of Political Stereotypes cites this paper.

Examining Alignment of Large Language Models through Representative Heuristics: The Case of Political Stereotypes The Capacity for Moral Self-Correction in Large Language Models

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-10T15:18:19.037689Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T15:18:19.037689Z digest=sha256:3f97ded4efeadc66dbce08ca6d5456b4f7ce9e5d7ee56ed8ee9eda45f94ab6f8

Observation baeecbd7-6222-423c-bd54-a7e288f10e8e · inbound

Data-adaptive Safety Rules for Training Reward Models cites this paper.

Data-adaptive Safety Rules for Training Reward Models The Capacity for Moral Self-Correction in Large Language Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-10T14:25:13.128436Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:25:13.128436Z digest=sha256:6326635cc72867e4c896f5c9e2a5ee7f3b3ac3ddefce77beb37d362561a60baf

Observation 1968ce0f-9a1e-4384-9168-7da50c9e8cd4 · inbound

A Checks-and-Balances Framework for Context-Aware Ethical AI Alignment cites this paper.

A Checks-and-Balances Framework for Context-Aware Ethical AI Alignment The Capacity for Moral Self-Correction in Large Language Models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-09T20:09:54.043982Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T20:09:54.043982Z digest=sha256:89209b43c2338031f55854c1556ba75655963de536f60e1a3827ef02b217d745

Observation f59dbf2d-9322-421d-8f5e-178568ba72be · inbound

Fairness through Difference Awareness: Measuring Desired Group Discrimination in LLMs cites this paper.

Fairness through Difference Awareness: Measuring Desired Group Discrimination in LLMs The Capacity for Moral Self-Correction in Large Language Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-09T14:03:44.403660Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T14:03:44.403660Z digest=sha256:28ee0639794bbdce46e5937197e1af15eda989e6addd0473c340ecd4b3a18c1d

Observation f108862d-1a37-4cfa-b9ec-9a2212584ec8 · inbound

Reflection-Window Decoding: Text Generation with Selective Refinement cites this paper.

Reflection-Window Decoding: Text Generation with Selective Refinement The Capacity for Moral Self-Correction in Large Language Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-09T04:13:44.247895Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T04:13:44.247895Z digest=sha256:ddd654a20097a9df909f395c5a1d935b0272c4c1722b0aa8d4bf71a9e8500d13

Observation 751caf8e-eca2-4170-a6ff-ec9796f26cff · inbound

Breaking Down Bias: On The Limits of Generalizable Pruning Strategies cites this paper.

Breaking Down Bias: On The Limits of Generalizable Pruning Strategies The Capacity for Moral Self-Correction in Large Language Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-08T11:42:07.337104Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T11:42:07.337104Z digest=sha256:3776d4372ab4c15278841f224e47b0a26b0abe94f9b5ca4c52328913ba06f7db

Observation c2069a82-2f6f-4986-be4f-79441d1df4e4 · inbound

Information Gain-Guided Causal Intervention for Autonomous Debiasing Large Language Models cites this paper.

Information Gain-Guided Causal Intervention for Autonomous Debiasing Large Language Models The Capacity for Moral Self-Correction in Large Language Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-16T12:24:59.539102Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:24:59.539102Z digest=sha256:348da8378073fbb7656d0cd8e50c1e419df1372e875d5f1329c8ffdd95f4f06c

Observation d3187f13-3880-49ea-99b9-117e8c96b578 · inbound

Bias Analysis and Mitigation through Protected Attribute Detection and Regard Classification cites this paper.

Bias Analysis and Mitigation through Protected Attribute Detection and Regard Classification The Capacity for Moral Self-Correction in Large Language Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-16T11:58:23.577593Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-16T11:58:23.577593Z digest=sha256:993aefce54cf5b46e9a7d1100093b4f0230b3a82686d4c53d8980a0472e96820

Observation 24701a4b-1fc4-4550-8d2d-63eb5f9e0848 · inbound

Should AI Mimic People? Understanding AI-Supported Writing Technology Among Black Users cites this paper.

Should AI Mimic People? Understanding AI-Supported Writing Technology Among Black Users The Capacity for Moral Self-Correction in Large Language Models

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-16T04:36:47.085804Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:36:47.085804Z digest=sha256:072fe565cf88c02ee95246e8aa43ec0421cbc419c59a974727ddc0792a25c04d

Observation c3ce01d4-a0f0-4d49-a4aa-9f545825b134 · inbound

LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models cites this paper.

LLM Ethics Benchmark: A Three-Dimensional Assessment System for Evaluating Moral Reasoning in Large Language Models The Capacity for Moral Self-Correction in Large Language Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-16T04:37:48.063908Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:37:48.063908Z digest=sha256:a40a24fc5cf95cccab9ff8d8b803d5b2d0db9ac5e8796f29c0997e956498cec3

Observation 05c7c6da-0367-4364-b4a3-b20339ca4ae7 · inbound

Improving LLM Outputs Against Jailbreak Attacks with Expert Model Integration cites this paper.

Improving LLM Outputs Against Jailbreak Attacks with Expert Model Integration The Capacity for Moral Self-Correction in Large Language Models

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-15T20:36:57.249783Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:36:57.249783Z digest=sha256:2e8a0341cdab3ce42fada44018d64e77abd528ffd0d1f4d7f95c26f877326550

Observation ab20ae17-7b33-41a4-928d-0cf1c23a9f40 · inbound

When Ethics and Payoffs Diverge: LLM Agents in Morally Charged Social Dilemmas cites this paper.

When Ethics and Payoffs Diverge: LLM Agents in Morally Charged Social Dilemmas The Capacity for Moral Self-Correction in Large Language Models

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T14:22:55.542349Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:22:55.542349Z digest=sha256:5c2141320828e6695e4ab7b621abcbbcedc567b67b0cc4074793c9361bc4b720

Observation 18762a2e-951c-4066-9112-e4e92f5fba79 · inbound

Whispers of Many Shores: Cultural Alignment through Collaborative Cultural Expertise cites this paper.

Whispers of Many Shores: Cultural Alignment through Collaborative Cultural Expertise The Capacity for Moral Self-Correction in Large Language Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T12:13:21.801364Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:13:21.801364Z digest=sha256:1373a8e88573847e6c7cdc6857a0cc87b0f59c6075e0a631278540df19a0e8dd

Observation dc0fda96-a79d-41d2-bd94-0566c47ef569 · inbound

Exchange of Perspective Prompting Enhances Reasoning in Large Language Models cites this paper.

Exchange of Perspective Prompting Enhances Reasoning in Large Language Models The Capacity for Moral Self-Correction in Large Language Models

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T11:04:00.028467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:04:00.028467Z digest=sha256:19855be411c1b7d2790d93225ed5b40584c7470dbe6bcc9ff641024899f73b4c

Observation fbc44730-4ea4-41bd-a0b0-7518f62f13e3 · inbound

Internal Value Alignment in Large Language Models through Controlled Value Vector Activation cites this paper.

Internal Value Alignment in Large Language Models through Controlled Value Vector Activation The Capacity for Moral Self-Correction in Large Language Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T17:17:29.139322Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:17:29.139322Z digest=sha256:b4c7e8f2ff89e229917a8b11f9028e47ca38c7bda8512d717994f24d11e9e85a

Observation 3bb6d37e-5fe3-4eee-ba19-fa80475479ff · inbound

SC-Captioner: Improving Image Captioning with Self-Correction by Reinforcement Learning cites this paper.

SC-Captioner: Improving Image Captioning with Self-Correction by Reinforcement Learning The Capacity for Moral Self-Correction in Large Language Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-05T22:58:34.281893Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T22:58:34.281893Z digest=sha256:7c290b2c85f38abafd91ca54ca50660ca5739cafbe4762eb6689107dd356e31d

Observation 4fc69028-ca24-4e2e-83a6-e60b34a39393 · inbound

Beyond I'm Sorry, I Can't: Dissecting Large Language Model Refusal cites this paper.

Beyond I'm Sorry, I Can't: Dissecting Large Language Model Refusal The Capacity for Moral Self-Correction in Large Language Models

Reference 16

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T18:56:45.850733Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-05-18T18:56:13.680353Z digest=sha256:5c362703ae0b1a8e973f5ca7cd8f09733c8c254cf0eebb373737cd51d087ae2f

Observation 385a94ea-d6f0-4069-9fd8-4d2e082ba3ae · inbound

HalluField: Detecting LLM Hallucinations via Field-Theoretic Modeling cites this paper.

HalluField: Detecting LLM Hallucinations via Field-Theoretic Modeling The Capacity for Moral Self-Correction in Large Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-04T17:41:05.230225Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T17:41:05.230225Z digest=sha256:4966e0361d4daf533a75d0455f5bd3b719a54326b136a3e4b9cbf2ba90e01741

Observation 1ccaa8a8-8275-4c93-aa2e-deb37ffaf330 · inbound

We Think, Therefore We Align LLMs to Helpful, Harmless and Honest Before They Go Wrong cites this paper.

We Think, Therefore We Align LLMs to Helpful, Harmless and Honest Before They Go Wrong The Capacity for Moral Self-Correction in Large Language Models

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-21T22:14:23.545726Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-05-21T22:11:20.651761Z digest=sha256:99ebd8db3aa299e2d0d000eecf51094614ddf917e5b84e3dc07b70a285a9ed26

Observation 3ec12e0e-4a7a-47ca-b2f7-5dc1d3b5640a · inbound

BioPro: Towards Difference-Aware Gender Fairness for Vision-Language Models cites this paper.

BioPro: Towards Difference-Aware Gender Fairness for Vision-Language Models The Capacity for Moral Self-Correction in Large Language Models

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-03T19:25:20.551987Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T19:25:20.551987Z digest=sha256:80c7cc73c4d613be3bca21ba2a19220fb6d3a04f6a437f153be99b2c9e07f6b6

Observation d3811577-9d12-46b6-968d-3619ed56349d · inbound

DeFrame: Debiasing Large Language Models Against Framing Effects cites this paper.

DeFrame: Debiasing Large Language Models Against Framing Effects The Capacity for Moral Self-Correction in Large Language Models

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-03T04:43:42.987351Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T04:43:42.987351Z digest=sha256:277ecdcea47769d7b397712ad6b3114411e0c4b3fd0a2a96109a814ad6e9c9c4

Observation a52956e5-0878-45f6-8f99-ea0da6d9d165 · inbound

Moral Sensitivity in LLMs: A Tiered Evaluation of Contextual Bias via Behavioral Profiling and Mechanistic Interpretability cites this paper.

Moral Sensitivity in LLMs: A Tiered Evaluation of Contextual Bias via Behavioral Profiling and Mechanistic Interpretability The Capacity for Moral Self-Correction in Large Language Models

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T10:51:29.506430Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-08T18:34:20.065335Z digest=sha256:fa43f8768b9b830e047e09e96f1ff39783c101813f207e86b06a1b381935d9f5

Observation a19b7205-130e-4674-9cc6-f64ec74858fc · inbound

Moral Sensitivity in LLMs: A Tiered Evaluation of Contextual Bias via Behavioral Profiling and Mechanistic Interpretability cites this paper.

Moral Sensitivity in LLMs: A Tiered Evaluation of Contextual Bias via Behavioral Profiling and Mechanistic Interpretability The Capacity for Moral Self-Correction in Large Language Models

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T00:25:10.319507Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-06-30T23:59:10.797762Z digest=sha256:42f3e91abf573432250226129591ef478024cafe9db5406df9d3a73b577849ae

Observation c8076c4f-ce26-41d5-bac1-be6708a8e662 · inbound

Internal vs. External: Comparing Deliberation and Evolution for Multi-Agent Constitutional Design cites this paper.

Internal vs. External: Comparing Deliberation and Evolution for Multi-Agent Constitutional Design The Capacity for Moral Self-Correction in Large Language Models

Reference 22

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T07:26:31.959348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-05-12T02:51:35.213002Z digest=sha256:9f7ec0605588eb5e0a98f1fd44080579eb79d3ab5dc3f010ca35cfb54c2fb465

Observation f9ce0d48-ad37-45aa-b0c0-aa6602207968 · inbound

Conformity Generates Collective Misalignment in AI Agents Societies cites this paper.

Conformity Generates Collective Misalignment in AI Agents Societies The Capacity for Moral Self-Correction in Large Language Models

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-12T06:06:27.048080Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-05-12T04:34:53.814304Z digest=sha256:0ab2a06d62b4f4796dcd1c090e939f6b95526ff29965ecc270dff05b436e0fad

Observation 12806caf-14dc-4bc1-8de8-add97f3c4b96 · inbound

Online Data Selection Is Implicit Alignment cites this paper.

Online Data Selection Is Implicit Alignment The Capacity for Moral Self-Correction in Large Language Models

Reference 20

Resolution
metadata mismatch
local_arxiv, observed 2026-07-09T21:36:34.280994Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-07-09T21:35:41.947875Z digest=sha256:8f332c1876d19413f8052ceb4d033fd64a2f7fffe257305df1d7e2f077a720d6

Observation 34db6765-7a34-49b8-8f82-bebcc6bf6aad · inbound

The Missing Layer: Specification Infrastructure for AI Oversight cites this paper.

The Missing Layer: Specification Infrastructure for AI Oversight The Capacity for Moral Self-Correction in Large Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-07-30T12:26:06.939145Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-30T12:26:06.939145Z digest=sha256:366800b8464bacb78d2cc8c7beb4b0fa2f205178c7d7e60fed8dba7f2473af83

Observation 11c5cfd8-6d5f-40e7-af5b-ef6cfff3b8ba · inbound

Dual-Adversarial Safety Alignment: Cultivating Intrinsic Threat Comprehension in LRMs cites this paper.

Dual-Adversarial Safety Alignment: Cultivating Intrinsic Threat Comprehension in LRMs The Capacity for Moral Self-Correction in Large Language Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-11T15:29:05.544671Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:29:05.544671Z digest=sha256:a947ecd38fe20120b9c48c3bdba0078298b5b6e5d9fdb3e02dc6eabf4c92a5c7

Observation 569f83e2-ef03-44a3-a815-8a60bc49cf49 · inbound

Rules or Character? Scaling Laws for AI Safety Design cites this paper.

Rules or Character? Scaling Laws for AI Safety Design The Capacity for Moral Self-Correction in Large Language Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-14T13:21:16.568273Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T13:21:16.568273Z digest=sha256:3acdafc13e2df8e68c9fd120b469d082585f6520a4953c42d3d4709be4dc834c