Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T17:41:37.839642Z
Paper Citation Record · LEDGER
As of 18 August 2026, this Paper Citation Record lists 46 of 46 outbound references and 6 inbound Pith citation observations for arXiv:2411.12405.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T17:41:37.839642Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-16T05:53:49.268919Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
46 of 46 outbound references displayed
External citation measurements
2
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
Observation 8aa62f06-3458-4eb8-a094-f29c9ce12d93 · outbound
Evaluating the Prompt Steerability of Large Language Models online" 'onlinestring :=
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f4158f4f-1a87-45f0-b754-4d8651d2951c · outbound
Evaluating the Prompt Steerability of Large Language Models write newline
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d4b367cf-ba58-4054-ad75-43c4ca4b123c · outbound
Evaluating the Prompt Steerability of Large Language Models Moral Foundations of Large Language Models
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4108345e-5975-4b9e-b15d-e753c19b757d · outbound
Evaluating the Prompt Steerability of Large Language Models Steering Large Language Models for Machine Translation with Finetuning and In-Context Learning
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation fbe38eae-e774-42f4-8e7f-dbde6e44bc1f · outbound
Evaluating the Prompt Steerability of Large Language Models Unresolved cited work
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 84153e73-afdc-497a-98d1-b965948ea50d · outbound
Evaluating the Prompt Steerability of Large Language Models Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 26394716-dd3c-4216-b0e2-dc57ce2ecee6 · outbound
Evaluating the Prompt Steerability of Large Language Models What's the Magic Word? A Control Theory of LLM Prompting
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ba842455-332c-4e74-91c3-c538a86e91c9 · outbound
Evaluating the Prompt Steerability of Large Language Models Language Models are Few-Shot Learners
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 040ff7b4-a903-489b-8824-232af3fd036f · outbound
Evaluating the Prompt Steerability of Large Language Models PAD: Personalized Alignment of LLMs at Decoding-Time
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 08ea60ec-06a1-4052-912e-4e5117b58840 · outbound
Evaluating the Prompt Steerability of Large Language Models Marked Personas: Using Natural Language Prompts to Measure Stereotypes in Language Models
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 070c05e0-ce12-4b55-b779-1afc724a0749 · outbound
Evaluating the Prompt Steerability of Large Language Models Social Choice Should Guide AI Alignment in Dealing with Diverse Human Feedback
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d71f37de-ed6d-4b74-afa8-508ac0b48d95 · outbound
Evaluating the Prompt Steerability of Large Language Models Modular Pluralism: Pluralistic Alignment via Multi-LLM Collaboration
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 42b2fa4d-3b1d-40d0-a900-2eb5da1c9ab5 · outbound
Evaluating the Prompt Steerability of Large Language Models Unresolved cited work
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 12ae1958-6f04-4460-aa79-e72512fd0f21 · outbound
Evaluating the Prompt Steerability of Large Language Models CharED: Character-wise Ensemble Decoding for Large Language Models
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b0a16491-0bee-4d25-939c-83c562d27c7b · outbound
Evaluating the Prompt Steerability of Large Language Models Unresolved cited work
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d5e45457-3610-4a40-b429-ef9bb8a5b89c · outbound
Evaluating the Prompt Steerability of Large Language Models Context Steering: Controllable Personalization at Inference Time
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 06c59bec-ccff-46f2-b23a-0d1713794b69 · outbound
Evaluating the Prompt Steerability of Large Language Models Unresolved cited work
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2961b6b9-3d7d-4332-bae9-52ff9e9c9a30 · outbound
Evaluating the Prompt Steerability of Large Language Models Unresolved cited work
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 2608deb9-74c1-44bc-b8b7-e10bce25842b · outbound
Evaluating the Prompt Steerability of Large Language Models What are human values, and how do we align AI to them?
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fb6015f3-3ac4-4e8f-85c2-4c5f6ca3c189 · outbound
Evaluating the Prompt Steerability of Large Language Models Large Language Models as Superpositions of Cultural Perspectives
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9e8cedc7-2aaf-4101-84fc-ced3eba60610 · outbound
Evaluating the Prompt Steerability of Large Language Models Propulsion: Steering LLM with Tiny Fine-Tuning
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 59a40dfb-8f62-4a62-94d8-4271a014cbdf · outbound
Evaluating the Prompt Steerability of Large Language Models Programming Refusal with Conditional Activation Steering
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ca2d3d15-dc02-4225-a0be-8cb1858bceb2 · outbound
Evaluating the Prompt Steerability of Large Language Models The Power of Scale for Parameter-Efficient Prompt Tuning
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1e953fee-b4a4-4740-88a9-47e5b3ec9edc · outbound
Evaluating the Prompt Steerability of Large Language Models How do nonlinear transformers learn and generalize in in-context learning? In Forty-first International Conference on Machine Learning
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 28dbff7e-8f6e-49c1-aecc-7db35f08032f · outbound
Evaluating the Prompt Steerability of Large Language Models On the steerability of large language models toward data-driven personas
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5a08b799-4132-4fd3-9f6d-8e0c892225ab · outbound
Evaluating the Prompt Steerability of Large Language Models Unresolved cited work
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation a5c00417-40fd-42d9-9c2d-a9cf022520d1 · outbound
Evaluating the Prompt Steerability of Large Language Models Evaluating Large Language Model Biases in Persona-Steered Generation
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fcb4f6f1-7bd9-4b4d-bdc5-5f50c3ce60d8 · outbound
Evaluating the Prompt Steerability of Large Language Models Language Models in Dialogue: Conversational Maxims for Human-AI Interactions
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3c7e512d-8568-430e-bb5d-4f7248282633 · outbound
Evaluating the Prompt Steerability of Large Language Models Discovering Language Model Behaviors with Model-Written Evaluations
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 592d5477-45f9-4c4f-8636-ae2fd93e5012 · outbound
Evaluating the Prompt Steerability of Large Language Models Steering Llama 2 via Contrastive Activation Addition
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 407674b5-dcd3-4ab3-a452-59a50cfc42f3 · outbound
Evaluating the Prompt Steerability of Large Language Models PersonaGym: Evaluating Persona Agents and LLMs
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7de1a0d9-f186-416b-a3ab-d83fe393a83b · outbound
Evaluating the Prompt Steerability of Large Language Models Unresolved cited work
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a2083548-c992-489e-8e43-446888144444 · outbound
Evaluating the Prompt Steerability of Large Language Models Unresolved cited work
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 0f9709d4-ee51-4b12-b4c9-72c17f22477b · outbound
Evaluating the Prompt Steerability of Large Language Models Unresolved cited work
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation c4d5d47c-c1fc-42f5-80b5-dbe911ae4c92 · outbound
Evaluating the Prompt Steerability of Large Language Models Unresolved cited work
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation bd8d25af-9e42-447a-a9ff-111ca1b4eccb · outbound
Evaluating the Prompt Steerability of Large Language Models A Roadmap to Pluralistic Alignment
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0022fd8b-2d63-4a0c-8137-4484d851c965 · outbound
Evaluating the Prompt Steerability of Large Language Models Steering Without Side Effects: Improving Post-Deployment Control of Language Models
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation af7dd123-e1d1-4b7c-ae57-80e3020b858e · outbound
Evaluating the Prompt Steerability of Large Language Models Exploring and steering the moral compass of Large Language Models
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4ac29ab4-4652-421c-9f98-8ef1804a0dd0 · outbound
Evaluating the Prompt Steerability of Large Language Models Steering Language Models With Activation Engineering
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation af6b22c8-9651-44a9-a4c5-fe9dc290e71e · outbound
Evaluating the Prompt Steerability of Large Language Models "My Answer is C": First-Token Probabilities Do Not Match Text Answers in Instruction-Tuned Language Models
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2d15a519-9ab5-4c5d-9b26-06ce2aaf7a31 · outbound
Evaluating the Prompt Steerability of Large Language Models Larger language models do in-context learning differently
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7158f4f1-e44e-4713-b4e7-ca9e648f6fe9 · outbound
Evaluating the Prompt Steerability of Large Language Models Unresolved cited work
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 396802aa-a2a3-4cc7-91d3-2eeab30d07ca · outbound
Evaluating the Prompt Steerability of Large Language Models Fundamental Limitations of Alignment in Large Language Models
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ed24eca5-6458-4407-a22c-d7375c01932d · outbound
Evaluating the Prompt Steerability of Large Language Models Aligning LLMs with Individual Preferences via Interaction
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a4700da1-00d5-4eb2-82a0-837e43fb0b3b · outbound
Evaluating the Prompt Steerability of Large Language Models Value FULCRA: Mapping Large Language Models to the Multidimensional Spectrum of Basic Human Values
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a23900c1-73a1-44f4-ba2e-17545e2bd10a · outbound
Evaluating the Prompt Steerability of Large Language Models Unresolved cited work
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation b06137df-9d8d-4cec-8aef-ab9f3d0727af · inbound
Security Steerability is All You Need Evaluating the Prompt Steerability of Large Language Models
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9f869a5d-b89e-4782-a86e-4c4706e7bba7 · inbound
AgentMisalignment: Measuring the Propensity for Misaligned Behaviour in LLM-Based Agents Evaluating the Prompt Steerability of Large Language Models
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bfba1236-9f15-4436-8e5c-85856d5dc69c · inbound
An Auditable Agent Platform For Automated Molecular Optimisation Evaluating the Prompt Steerability of Large Language Models
Reference 73
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dc0fad67-1014-46d3-a6d6-fdae6d6cfddb · inbound
AI Behavioral Science Evaluating the Prompt Steerability of Large Language Models
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ef0fd661-6cec-4f75-a1db-3df998ac2229 · inbound
The Pragmatic Persona: Discovering LLM Persona through Bridging Inference Evaluating the Prompt Steerability of Large Language Models
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.
Observation 273bc998-9712-4cde-af61-d363b7159f12 · inbound
Investigating Linguistic Steering: An Analysis of Adjectival Effects Across Large Language Model Architectures Evaluating the Prompt Steerability of Large Language Models
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.