Pith. sign in

Paper Citation Record · LEDGER

Evaluating the Prompt Steerability of Large Language Models

As of 18 August 2026, this Paper Citation Record lists 46 of 46 outbound references and 6 inbound Pith citation observations for arXiv:2411.12405.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.12405 v2

Coverage vector

measured 46 of 46 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T17:41:37.839642Z

measured 52 of 52 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T05:53:49.268919Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

46 of 46 outbound references displayed

  • verified exact1
  • verified fuzzy1
  • unresolved44
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

2
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation 8aa62f06-3458-4eb8-a094-f29c9ce12d93 · outbound

This paper cites online" 'onlinestring :=.

Evaluating the Prompt Steerability of Large Language Models online" 'onlinestring :=

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-12T17:41:37.607742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:41:37.607742Z digest=sha256:29a7e25fcb85334294e8a36b0cbaf721ee718fbd11f9007a06678403df8c07d7

Observation f4158f4f-1a87-45f0-b754-4d8651d2951c · outbound

This paper cites write newline.

Evaluating the Prompt Steerability of Large Language Models write newline

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-12T17:41:37.613844Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:41:37.613844Z digest=sha256:588b7d9855b64bd5faeda8347296918530a08ad4fa3ad6bac66fea4d287b5ac3

Observation d4b367cf-ba58-4054-ad75-43c4ca4b123c · outbound

This paper cites Moral Foundations of Large Language Models.

Evaluating the Prompt Steerability of Large Language Models Moral Foundations of Large Language Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-12T17:41:37.619963Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:41:37.619963Z digest=sha256:41b9193170100f88511c2bcb21fc17636f7e816dadb105abea073abf74b2340b

Observation 4108345e-5975-4b9e-b15d-e753c19b757d · outbound

This paper cites Steering Large Language Models for Machine Translation with Finetuning and In-Context Learning.

Evaluating the Prompt Steerability of Large Language Models Steering Large Language Models for Machine Translation with Finetuning and In-Context Learning

Reference 4

Resolution
verified exact
local_arxiv, observed 2026-08-12T17:41:38.478629Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-12T17:41:37.626504Z digest=sha256:c69182f58cc2b3a6dfd319eaf748e3f309e4f949d0a83eed066110e46497188d

Observation fbe38eae-e774-42f4-8e7f-dbde6e44bc1f · outbound

This paper cites an unresolved cited work.

Evaluating the Prompt Steerability of Large Language Models Unresolved cited work

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-12T17:41:37.631940Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:41:37.631940Z digest=sha256:40c79c32470d5a490c5876f2e1f5adaa92af22804952700ec70865e4432aa441

Observation 84153e73-afdc-497a-98d1-b965948ea50d · outbound

This paper cites Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback.

Evaluating the Prompt Steerability of Large Language Models Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-12T17:41:37.637083Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:41:37.637083Z digest=sha256:4c64aa1ef39e40e8ccfd08018d0d98acf95999e4f03d6a09c6298a565f3f6738

Observation 26394716-dd3c-4216-b0e2-dc57ce2ecee6 · outbound

This paper cites What's the Magic Word? A Control Theory of LLM Prompting.

Evaluating the Prompt Steerability of Large Language Models What's the Magic Word? A Control Theory of LLM Prompting

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-12T17:41:37.642442Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:41:37.642442Z digest=sha256:c0ab7909eaf908d70a181675fbcf4ba34795d136938d1a3ac52d5b13885e4103

Observation ba842455-332c-4e74-91c3-c538a86e91c9 · outbound

This paper cites Language Models are Few-Shot Learners.

Evaluating the Prompt Steerability of Large Language Models Language Models are Few-Shot Learners

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-12T17:41:37.647636Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:41:37.647636Z digest=sha256:11a076d28a4dfc1818040c7e2bc3eff01a7c7c6675fe91aaab2a3a01f876ce75

Observation 040ff7b4-a903-489b-8824-232af3fd036f · outbound

This paper cites PAD: Personalized Alignment of LLMs at Decoding-Time.

Evaluating the Prompt Steerability of Large Language Models PAD: Personalized Alignment of LLMs at Decoding-Time

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-12T17:41:37.653815Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:41:37.653815Z digest=sha256:a98db58e34f2a9f7eec9055b86d1dc5065b364bb3db2481e9c06fc51a842ff85

Observation 08ea60ec-06a1-4052-912e-4e5117b58840 · outbound

This paper cites Marked Personas: Using Natural Language Prompts to Measure Stereotypes in Language Models.

Evaluating the Prompt Steerability of Large Language Models Marked Personas: Using Natural Language Prompts to Measure Stereotypes in Language Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-12T17:41:37.659054Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:41:37.659054Z digest=sha256:c4d72dd67f9137e3b8f625bde7c13255b6629bfde01b5528344e65a4069fbd07

Observation 070c05e0-ce12-4b55-b779-1afc724a0749 · outbound

This paper cites Social Choice Should Guide AI Alignment in Dealing with Diverse Human Feedback.

Evaluating the Prompt Steerability of Large Language Models Social Choice Should Guide AI Alignment in Dealing with Diverse Human Feedback

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-12T17:41:37.664353Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:41:37.664353Z digest=sha256:f199c991f9183252a3e1001425462ad730da18fbd217cfd4fccdcc0cab7eef10

Observation d71f37de-ed6d-4b74-afa8-508ac0b48d95 · outbound

This paper cites Modular Pluralism: Pluralistic Alignment via Multi-LLM Collaboration.

Evaluating the Prompt Steerability of Large Language Models Modular Pluralism: Pluralistic Alignment via Multi-LLM Collaboration

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-12T17:41:37.669774Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:41:37.669774Z digest=sha256:a6834df2f41e935d6a6edf6e3412fb6191fb4f3916d2e18850558dde0c68c665

Observation 42b2fa4d-3b1d-40d0-a900-2eb5da1c9ab5 · outbound

This paper cites an unresolved cited work.

Evaluating the Prompt Steerability of Large Language Models Unresolved cited work

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-12T17:41:37.674861Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:41:37.674861Z digest=sha256:bf882d335e524756c140337f41a23f2170f0e3d5dcd5fdba70322b7d4a8c8560

Observation 12ae1958-6f04-4460-aa79-e72512fd0f21 · outbound

This paper cites CharED: Character-wise Ensemble Decoding for Large Language Models.

Evaluating the Prompt Steerability of Large Language Models CharED: Character-wise Ensemble Decoding for Large Language Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-12T17:41:37.679520Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:41:37.679520Z digest=sha256:3b678f67d5a36ab526fdcc9bd29d35ce79d97f9eab236cead19064462fbe7c6e

Observation b0a16491-0bee-4d25-939c-83c562d27c7b · outbound

This paper cites an unresolved cited work.

Evaluating the Prompt Steerability of Large Language Models Unresolved cited work

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-12T17:41:37.684478Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:41:37.684478Z digest=sha256:4573a3a8ca93c9ff7aee1a76b4ae48892dc2d83592ed17f43edd8e2cfad1474b

Observation d5e45457-3610-4a40-b429-ef9bb8a5b89c · outbound

This paper cites Context Steering: Controllable Personalization at Inference Time.

Evaluating the Prompt Steerability of Large Language Models Context Steering: Controllable Personalization at Inference Time

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-12T17:41:37.689178Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:41:37.689178Z digest=sha256:bfa7fa0a23a898f009c395ffaf8c6ae7bc76f0cf78957729bf949aa022f54e57

Observation 06c59bec-ccff-46f2-b23a-0d1713794b69 · outbound

This paper cites an unresolved cited work.

Evaluating the Prompt Steerability of Large Language Models Unresolved cited work

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-12T17:41:37.693988Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:41:37.693988Z digest=sha256:76f78d3ba0d51c78071411d32a2ab18b89b727266ad6b1cc25d78557a5a15680

Observation 2961b6b9-3d7d-4332-bae9-52ff9e9c9a30 · outbound

This paper cites an unresolved cited work.

Evaluating the Prompt Steerability of Large Language Models Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-12T17:41:38.631681Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-12T17:41:37.699814Z digest=sha256:1465cf8af74e1afaa014601829df63e5b369e0e08581c073b12ec5184c146413

Observation 2608deb9-74c1-44bc-b8b7-e10bce25842b · outbound

This paper cites What are human values, and how do we align AI to them?.

Evaluating the Prompt Steerability of Large Language Models What are human values, and how do we align AI to them?

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-12T17:41:37.704779Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:41:37.704779Z digest=sha256:043496d780e767cd34f7ff98db065f4c04f68f201e35e0a685173cc965f446a1

Observation fb6015f3-3ac4-4e8f-85c2-4c5f6ca3c189 · outbound

This paper cites Large Language Models as Superpositions of Cultural Perspectives.

Evaluating the Prompt Steerability of Large Language Models Large Language Models as Superpositions of Cultural Perspectives

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-12T17:41:37.709708Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:41:37.709708Z digest=sha256:fa95a47996335eb7ef871bfdf76070430b76f6b09926a5b2c13d0e20914bae6e

Observation 9e8cedc7-2aaf-4101-84fc-ced3eba60610 · outbound

This paper cites Propulsion: Steering LLM with Tiny Fine-Tuning.

Evaluating the Prompt Steerability of Large Language Models Propulsion: Steering LLM with Tiny Fine-Tuning

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-12T17:41:37.714827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:41:37.714827Z digest=sha256:359328a10c1477d4d58cf039dbedf6f16f64e78638f577d2f9d3f2a3e126ec74

Observation 59a40dfb-8f62-4a62-94d8-4271a014cbdf · outbound

This paper cites Programming Refusal with Conditional Activation Steering.

Evaluating the Prompt Steerability of Large Language Models Programming Refusal with Conditional Activation Steering

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-12T17:41:37.719599Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:41:37.719599Z digest=sha256:6fbf1d51c50075b8dddfa699a9f4ef416b739c8f1e0242c01d6648f0cc4224e5

Observation ca2d3d15-dc02-4225-a0be-8cb1858bceb2 · outbound

This paper cites The Power of Scale for Parameter-Efficient Prompt Tuning.

Evaluating the Prompt Steerability of Large Language Models The Power of Scale for Parameter-Efficient Prompt Tuning

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-12T17:41:37.724304Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:41:37.724304Z digest=sha256:3c78e7223031033cf559d33358118c77471040a17fb02d9ac072d8d082eae706

Observation 1e953fee-b4a4-4740-88a9-47e5b3ec9edc · outbound

This paper cites How do nonlinear transformers learn and generalize in in-context learning? In Forty-first International Conference on Machine Learning.

Evaluating the Prompt Steerability of Large Language Models How do nonlinear transformers learn and generalize in in-context learning? In Forty-first International Conference on Machine Learning

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T17:41:38.613522Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-12T17:41:37.729246Z digest=sha256:f01ca057c4cd2d81f86bb146e8494267e178c48aa97d20fe7b30855bba0da4ae

Observation 28dbff7e-8f6e-49c1-aecc-7db35f08032f · outbound

This paper cites On the steerability of large language models toward data-driven personas.

Evaluating the Prompt Steerability of Large Language Models On the steerability of large language models toward data-driven personas

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-12T17:41:37.734102Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:41:37.734102Z digest=sha256:729225f34f43b02e3617cc5b59419721514cd4b1a9677be049ce4b31b496c709

Observation 5a08b799-4132-4fd3-9f6d-8e0c892225ab · outbound

This paper cites an unresolved cited work.

Evaluating the Prompt Steerability of Large Language Models Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-12T17:41:38.597211Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-12T17:41:37.739151Z digest=sha256:267767e8aeeff83b4ccb3ba2440f306471f504b05f0d9d42313fcdc4a0a37a26

Observation a5c00417-40fd-42d9-9c2d-a9cf022520d1 · outbound

This paper cites Evaluating Large Language Model Biases in Persona-Steered Generation.

Evaluating the Prompt Steerability of Large Language Models Evaluating Large Language Model Biases in Persona-Steered Generation

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-12T17:41:37.744071Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:41:37.744071Z digest=sha256:286f817636ff835a6a59e30e491b4d557cfb71b3f13c2e3af4ebc2c88e4af545

Observation fcb4f6f1-7bd9-4b4d-bdc5-5f50c3ce60d8 · outbound

This paper cites Language Models in Dialogue: Conversational Maxims for Human-AI Interactions.

Evaluating the Prompt Steerability of Large Language Models Language Models in Dialogue: Conversational Maxims for Human-AI Interactions

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-12T17:41:37.749236Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:41:37.749236Z digest=sha256:def92c1d4d57be0790967ec0a0d7de9bdea3dd939de4d2560c158fec84c7f0cf

Observation 3c7e512d-8568-430e-bb5d-4f7248282633 · outbound

This paper cites Discovering Language Model Behaviors with Model-Written Evaluations.

Evaluating the Prompt Steerability of Large Language Models Discovering Language Model Behaviors with Model-Written Evaluations

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-12T17:41:37.754124Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:41:37.754124Z digest=sha256:6836d0968ec9d3c3d9420fb2e057c3c5a01c21764ed8cd3fa6a1828a46b857e7

Observation 592d5477-45f9-4c4f-8636-ae2fd93e5012 · outbound

This paper cites Steering Llama 2 via Contrastive Activation Addition.

Evaluating the Prompt Steerability of Large Language Models Steering Llama 2 via Contrastive Activation Addition

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-12T17:41:37.759134Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:41:37.759134Z digest=sha256:f2def748095145c17a6b81cbadf3db13041477f4e9eb702d3416ad1aea97b256

Observation 407674b5-dcd3-4ab3-a452-59a50cfc42f3 · outbound

This paper cites PersonaGym: Evaluating Persona Agents and LLMs.

Evaluating the Prompt Steerability of Large Language Models PersonaGym: Evaluating Persona Agents and LLMs

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-12T17:41:37.764144Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:41:37.764144Z digest=sha256:1b16142db90ff70eaafb8c8f9cae4e796808ffdc592fcc61d062bf1286aa72ca

Observation 7de1a0d9-f186-416b-a3ab-d83fe393a83b · outbound

This paper cites an unresolved cited work.

Evaluating the Prompt Steerability of Large Language Models Unresolved cited work

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-12T17:41:37.769206Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:41:37.769206Z digest=sha256:14920e74769bffe58a74767bdfb8bcdd96788172d067e73ca414505852e5c131

Observation a2083548-c992-489e-8e43-446888144444 · outbound

This paper cites an unresolved cited work.

Evaluating the Prompt Steerability of Large Language Models Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-12T17:41:38.571810Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-12T17:41:37.774452Z digest=sha256:0132b07f3029996e582a4ee12936bae4902dfbf852df594f6a1fb810d7146d56

Observation 0f9709d4-ee51-4b12-b4c9-72c17f22477b · outbound

This paper cites an unresolved cited work.

Evaluating the Prompt Steerability of Large Language Models Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-12T17:41:38.555721Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-12T17:41:37.779199Z digest=sha256:11647aa546c7c96b02e53b1c278ffa28aeb19ae68d9735c7c0d6a0cf606fc2ca

Observation c4d5d47c-c1fc-42f5-80b5-dbe911ae4c92 · outbound

This paper cites an unresolved cited work.

Evaluating the Prompt Steerability of Large Language Models Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-12T17:41:38.539854Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-12T17:41:37.783952Z digest=sha256:23d9767db43664db6d05b6da6bd180de23c1785fdb7f8fa2b6ef8f517364d386

Observation bd8d25af-9e42-447a-a9ff-111ca1b4eccb · outbound

This paper cites A Roadmap to Pluralistic Alignment.

Evaluating the Prompt Steerability of Large Language Models A Roadmap to Pluralistic Alignment

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-12T17:41:37.789564Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:41:37.789564Z digest=sha256:4de9c3544061c06575434d52b91bc8e9030767f7882f7fbecbf72732a1f558a6

Observation 0022fd8b-2d63-4a0c-8137-4484d851c965 · outbound

This paper cites Steering Without Side Effects: Improving Post-Deployment Control of Language Models.

Evaluating the Prompt Steerability of Large Language Models Steering Without Side Effects: Improving Post-Deployment Control of Language Models

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-12T17:41:37.794346Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:41:37.794346Z digest=sha256:cc8742026273e47633b049a904a49e31cce1514c7febfd820978ce44d528d92f

Observation af7dd123-e1d1-4b7c-ae57-80e3020b858e · outbound

This paper cites Exploring and steering the moral compass of Large Language Models.

Evaluating the Prompt Steerability of Large Language Models Exploring and steering the moral compass of Large Language Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-12T17:41:37.799331Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:41:37.799331Z digest=sha256:f2937d8ed748b2e59b8bec30130b4b58d2dd858894811141d0f497f4594e4bfc

Observation 4ac29ab4-4652-421c-9f98-8ef1804a0dd0 · outbound

This paper cites Steering Language Models With Activation Engineering.

Evaluating the Prompt Steerability of Large Language Models Steering Language Models With Activation Engineering

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-12T17:41:37.804325Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:41:37.804325Z digest=sha256:9682df4ff316fcdae23225e281f7762c6a98296ca9639fe47d4f56cb9171c856

Observation af6b22c8-9651-44a9-a4c5-fe9dc290e71e · outbound

This paper cites "My Answer is C": First-Token Probabilities Do Not Match Text Answers in Instruction-Tuned Language Models.

Evaluating the Prompt Steerability of Large Language Models "My Answer is C": First-Token Probabilities Do Not Match Text Answers in Instruction-Tuned Language Models

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-12T17:41:37.809204Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:41:37.809204Z digest=sha256:70522c814e00bfbec9495f65bf7d6878b9425b05a9c33376569b295f00716f67

Observation 2d15a519-9ab5-4c5d-9b26-06ce2aaf7a31 · outbound

This paper cites Larger language models do in-context learning differently.

Evaluating the Prompt Steerability of Large Language Models Larger language models do in-context learning differently

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-12T17:41:37.814156Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:41:37.814156Z digest=sha256:46f47ed308e1721a76d3d37d2371050470043dba7cbc9d4dd65c2887e281a049

Observation 7158f4f1-e44e-4713-b4e7-ca9e648f6fe9 · outbound

This paper cites an unresolved cited work.

Evaluating the Prompt Steerability of Large Language Models Unresolved cited work

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-12T17:41:37.818994Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:41:37.818994Z digest=sha256:baa5b5f2a2d1b6236df78286e9a2ee51e86ab3bda2d1d58ee7a40efb89ba8ab7

Observation 396802aa-a2a3-4cc7-91d3-2eeab30d07ca · outbound

This paper cites Fundamental Limitations of Alignment in Large Language Models.

Evaluating the Prompt Steerability of Large Language Models Fundamental Limitations of Alignment in Large Language Models

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-12T17:41:37.823892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:41:37.823892Z digest=sha256:40419d25d756435db777c955e7a9829267aea44b26bc4fa89f057dbf23e83143

Observation ed24eca5-6458-4407-a22c-d7375c01932d · outbound

This paper cites Aligning LLMs with Individual Preferences via Interaction.

Evaluating the Prompt Steerability of Large Language Models Aligning LLMs with Individual Preferences via Interaction

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-12T17:41:37.829646Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:41:37.829646Z digest=sha256:372aa459746e12c0f80d150c5fe4d2eee194506ad0d0932f5aabc8a87db6c1a9

Observation a4700da1-00d5-4eb2-82a0-837e43fb0b3b · outbound

This paper cites Value FULCRA: Mapping Large Language Models to the Multidimensional Spectrum of Basic Human Values.

Evaluating the Prompt Steerability of Large Language Models Value FULCRA: Mapping Large Language Models to the Multidimensional Spectrum of Basic Human Values

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-12T17:41:37.834841Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T17:41:37.834841Z digest=sha256:580fcc04005ff5caac4143ceeb0da8befc783cb1430da6474599fe9e5b347771

Observation a23900c1-73a1-44f4-ba2e-17545e2bd10a · outbound

This paper cites an unresolved cited work.

Evaluating the Prompt Steerability of Large Language Models Unresolved cited work

Reference 46

Resolution
unresolved
raw_fallback, observed 2026-08-12T17:41:38.513075Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-08-12T17:41:37.839642Z digest=sha256:b021c45ea3668b8a64dfc8b772f8e18d62c5ee97fab0f5e802bcc63d54ea36ff

Pith citing papers

Observation b06137df-9d8d-4cec-8aef-ab9f3d0727af · inbound

Security Steerability is All You Need cites this paper.

Security Steerability is All You Need Evaluating the Prompt Steerability of Large Language Models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-16T05:53:49.268919Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T05:53:49.268919Z digest=sha256:2cc05035dbf556c50a626e348d71d5b77628b0b036b161b5601925803a7bbb25

Observation 9f869a5d-b89e-4782-a86e-4c4706e7bba7 · inbound

AgentMisalignment: Measuring the Propensity for Misaligned Behaviour in LLM-Based Agents cites this paper.

AgentMisalignment: Measuring the Propensity for Misaligned Behaviour in LLM-Based Agents Evaluating the Prompt Steerability of Large Language Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T10:54:57.988680Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:54:57.988680Z digest=sha256:b1a5aedd177c6336b74a6ee44f92fbc1047ebad6dd43ea336fa212dad6c8aaf8

Observation bfba1236-9f15-4436-8e5c-85856d5dc69c · inbound

An Auditable Agent Platform For Automated Molecular Optimisation cites this paper.

An Auditable Agent Platform For Automated Molecular Optimisation Evaluating the Prompt Steerability of Large Language Models

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-06T04:32:30.391727Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T04:32:30.391727Z digest=sha256:5c90afdbb636871cab335bbe8f009c869bc5983c6bca88403eb875e1a4f91244

Observation dc0fad67-1014-46d3-a6d6-fdae6d6cfddb · inbound

AI Behavioral Science cites this paper.

AI Behavioral Science Evaluating the Prompt Steerability of Large Language Models

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-15T17:27:30.505364Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:27:30.505364Z digest=sha256:9bef399f4bf197324228fcd6c4c647899b492041fdd83b55ae5b55ccf7d50a4d

Observation ef0fd661-6cec-4f75-a1db-3df998ac2229 · inbound

The Pragmatic Persona: Discovering LLM Persona through Bridging Inference cites this paper.

The Pragmatic Persona: Discovering LLM Persona through Bridging Inference Evaluating the Prompt Steerability of Large Language Models

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-09T00:14:27.138161Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-08T04:05:13.050792Z digest=sha256:d59c1526499b0b7daa78304d5bbed282275e01a67488f327ecba265a4fdc4f43

Observation 273bc998-9712-4cde-af61-d363b7159f12 · inbound

Investigating Linguistic Steering: An Analysis of Adjectival Effects Across Large Language Model Architectures cites this paper.

Investigating Linguistic Steering: An Analysis of Adjectival Effects Across Large Language Model Architectures Evaluating the Prompt Steerability of Large Language Models

Reference 10

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T09:05:36.554423Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-07-01T09:01:34.038992Z digest=sha256:e79a10c61f1fac428657756a6f648f71c55dc2299da291e777edf271456247da