Pith. sign in

Paper Citation Record · LEDGER

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas

As of 8 August 2026, this Paper Citation Record lists 68 of 68 outbound references and 5 inbound Pith citation observations for arXiv:2505.14633.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.14633 v1

Coverage vector

measured 68 of 68 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:36:29.092464Z

measured 73 of 73 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-30T22:55:08.544620Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

68 of 68 outbound references displayed

  • verified exact0
  • verified fuzzy40
  • unresolved28
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation 9da338ea-395d-48fe-af50-040df2a015ef · outbound

This paper cites Openai’s approach to external red teaming for ai models and systems, 2025.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Openai’s approach to external red teaming for ai models and systems, 2025

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:37.649789Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:36:23.167899Z digest=sha256:81a04a0afdf075a7ed190a41a18cc5f33d00bbcd9f060218b689254b4dc235e5

Observation e612560d-a796-4b59-b405-e8be8e4e6a66 · outbound

This paper cites Claude’s Constitution.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Claude’s Constitution

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:37.515524Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:36:23.220333Z digest=sha256:4cf6c7a6557ecd830ff2341bf057a2aae426e806e8a64d438f5f830cccf95bff

Observation fd9479b6-ab73-45a6-806f-01adb0c386bc · outbound

This paper cites Chatbot Arena Leaderboard.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Chatbot Arena Leaderboard

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:37.331310Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:36:23.348467Z digest=sha256:12500f9fc31d8def3fd9a5ee672dfe73eee5bf55a6f2aee687033cc3459f224e

Observation a11e9313-06d0-4a9a-8672-46e66e18a6ad · outbound

This paper cites Probing pre-trained language models for cross-cultural differences in values.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Probing pre-trained language models for cross-cultural differences in values

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:37.148750Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:36:23.456441Z digest=sha256:5311c0b0245d0c78dba66fc4ec90e079faa12f526cf3401fa177270766d89799

Observation b4fb9b53-d2ba-4747-aace-40a91a474eb6 · outbound

This paper cites A general language assistant as a laboratory for alignment, 2021.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas A general language assistant as a laboratory for alignment, 2021

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T15:36:23.572062Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:36:23.572062Z digest=sha256:42b753d655168158c7a5e2b65773ff1ed93c564e481c9eaf673cf0486bc21c3e

Observation faa653c7-1f5c-45ea-a118-4cb7f66bf2d0 · outbound

This paper cites Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T15:36:23.723831Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:36:23.723831Z digest=sha256:f8a39b4fc9fa17ca41adc3d3e65806ec5ceb0eaf5d54cf2e17c4ca44a6367f0a

Observation 103bad50-e152-4432-b584-a5a185bccc95 · outbound

This paper cites an unresolved cited work.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:36:36.995699Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:36:23.910855Z digest=sha256:7155f02aa6c03e7210246b5501f3714a804d0fd7192d3dd63d24565c7f60016d

Observation 9086aa86-8f32-4c7f-892a-fb9f9077b598 · outbound

This paper cites Demonstrating specification gaming in reasoning models.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Demonstrating specification gaming in reasoning models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T15:36:23.994668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:36:23.994668Z digest=sha256:b9287503060c971553e7b40bffcd9630201d675d92dad04ecc80ecfae0aa02e7

Observation bf5fab19-ce45-45f6-9ecd-2fc93428877a · outbound

This paper cites Distillation scaling laws, 2025.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Distillation scaling laws, 2025

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:36.808444Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:36:24.109958Z digest=sha256:6ffa9d3f3028c530ed9a74d3a1623bc6c29a2b073fc9f737436739769e4e897d

Observation 04552353-ff02-44bb-b62f-c852237b4b2c · outbound

This paper cites Is Power-Seeking AI an Existential Risk?.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Is Power-Seeking AI an Existential Risk?

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T15:36:24.240868Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:36:24.240868Z digest=sha256:f5841b40ae55b12fc14d128f13b39cbe7546614c9e87bf7541ebf2bd3dac3f34

Observation ff688ed6-a3d1-47d3-92bf-283413cc4df7 · outbound

This paper cites Reasoning models don’t always say what they think.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Reasoning models don’t always say what they think

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:36.627657Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:36:24.406690Z digest=sha256:2c43403c83a1ce977db92d0892fcd9766861171d5c8907d776bd175b32f498aa

Observation a2324fa3-a8fb-403a-abcb-647d3435c86f · outbound

This paper cites Chatbot arena: An open platform for evaluating llms by human preference.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Chatbot arena: An open platform for evaluating llms by human preference

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T15:36:24.580076Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:36:24.580076Z digest=sha256:fc2594019a5c7c615cf71628185753cdb7bfbe855e555191a687aabb6ae0e081

Observation 97f9a4e9-d5cf-4150-98fd-3b87c4b78233 · outbound

This paper cites DailyDilemmas: Revealing Value Preferences of LLMs with Quandaries of Daily Life.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas DailyDilemmas: Revealing Value Preferences of LLMs with Quandaries of Daily Life

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T15:36:24.626170Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:36:24.626170Z digest=sha256:bc45b8c4c1052eab7d3893e0a635630ac00f0e9a3a65a5e232710c0a9ce93d2f

Observation a9781ee0-e160-4087-9e4d-8714a97652ba · outbound

This paper cites Safety Leaderboard.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Safety Leaderboard

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:36.434804Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:36:24.688735Z digest=sha256:03f8001e8e2a11cec3cf8db5ba188aded90ed2cfd20a08311892c85078de6dfb

Observation 0adc6ff6-31f8-48b2-8536-962a6c4b51d1 · outbound

This paper cites Stated versus revealed preferences: An approach to reduce bias.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Stated versus revealed preferences: An approach to reduce bias

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:36.255052Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:36:24.766228Z digest=sha256:68f38d76ddafa0c5d8503ec6177d7906cf019c857f8d5118e872e0c9fcdd8ea7

Observation 9d1df6ed-0fd7-488d-b25b-0c8db338959a · outbound

This paper cites an unresolved cited work.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:36:36.027604Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:36:24.858284Z digest=sha256:1dab4778b245db95d74d38eadc6997c4b248d5663ab9fb20dce806ba9ceaae6f

Observation da42bae0-e9fc-466f-8959-2c6ddef92c12 · outbound

This paper cites A worldwide test of the predictive validity of ideal partner preference-matching.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas A worldwide test of the predictive validity of ideal partner preference-matching

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:35.798786Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:36:24.926602Z digest=sha256:8d6a9f1f92db7a7aa61c30035221f7be3be2a3bc90c8de1769afd9a55bf0d525

Observation dc4901ad-9250-41bf-956b-5a445a180bd3 · outbound

This paper cites Alignment faking in large language models.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Alignment faking in large language models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T15:36:24.998756Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:36:24.998756Z digest=sha256:a10ec4ffc100702a08e560efcda1a9dfd35c28612f333945290883017da07457

Observation 5fde7871-ff4d-460a-a87c-1f2ec2309245 · outbound

This paper cites The righteous mind.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas The righteous mind

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:35.567009Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:36:25.067422Z digest=sha256:e51a8bcb4cfdf1a26056b144cb6e3096a1fcb765b1e7278f7e8bb8c8d2239858

Observation 089c22d9-af22-47cb-8ac8-94faef1e93b5 · outbound

This paper cites Chapter 7 - creativity and morality in deception.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Chapter 7 - creativity and morality in deception

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:35.360808Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:36:25.161948Z digest=sha256:eb4303ad14b035b1fb0a6e646e8065f665d3179d18a16effb00a58acb696287f

Observation 0779c0e6-093b-47ce-9ee8-65fe3fb6a7a2 · outbound

This paper cites An Overview of Catastrophic AI Risks.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas An Overview of Catastrophic AI Risks

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T15:36:25.216066Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:36:25.216066Z digest=sha256:f50a99d1bef44a5625820207b5f2a2eae00e86678bd23951889ed6a72d676b99

Observation 476fc1f2-1148-46b6-a2d3-0a1b33254605 · outbound

This paper cites Values in the Wild: Discovering and Analyzing Values in Real-World Language Model Interactions.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Values in the Wild: Discovering and Analyzing Values in Real-World Language Model Interactions

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T15:36:25.282893Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:36:25.282893Z digest=sha256:3a8016e76228693987024e4e55b848110a7b924720ebf78c485d9e6eeb7650ed

Observation c0b265d1-eecd-4a1b-bf73-42ca7c882cbd · outbound

This paper cites Sleeper Agents: Training Deceptive LLMs that Persist Through Safety Training.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Sleeper Agents: Training Deceptive LLMs that Persist Through Safety Training

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T15:36:25.363672Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:36:25.363672Z digest=sha256:769d1afc9b98bf7cc6b180dce6828972486d62f981cba4dfc61446003a1d3529

Observation 06df84d9-c78d-4cde-80b6-22f7a209071f · outbound

This paper cites Wildteaming at scale: From in-the-wild jailbreaks to (adversarially) safer language models.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Wildteaming at scale: From in-the-wild jailbreaks to (adversarially) safer language models

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:35.128745Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:36:25.443363Z digest=sha256:3e2b23fa10c5ff7ecbbe6f82bda9dad6871adb53125507e9953d546972247883

Observation f51a94db-3ec4-41fb-9321-ae62f27137b3 · outbound

This paper cites Industrial society and its future, 2006.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Industrial society and its future, 2006

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:34.904746Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:36:25.487935Z digest=sha256:6627081f08a4baf216f0369b98bc4a8734e899219feac64192e95b2490c18432

Observation 5c788646-bf08-401b-9ac3-3e9faabe01c5 · outbound

This paper cites The PRISM Alignment Dataset: What Participatory, Representative and Individualised Human Feedback Reveals About the Subjective and Multicultural Alignment of Large Language Models.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas The PRISM Alignment Dataset: What Participatory, Representative and Individualised Human Feedback Reveals About the Subjective and Multicultural Alignment of Large Language Models

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T15:36:25.575004Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:36:25.575004Z digest=sha256:2434f9259d25689540568eff2066ed3936c7578180490ed40dc8f00227f12a88

Observation 2a2aeb8f-48b2-4f7e-a100-404803f7a3e5 · outbound

This paper cites an unresolved cited work.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:36:34.723553Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:36:25.652004Z digest=sha256:2f24e91752993a010ea36b0498a74eb428b810e41f4852b6a909a5515ed7daa4

Observation 8112ee9e-77f6-4af1-b733-af4b4938977e · outbound

This paper cites Stick to your role! stability of personal values expressed in large language models.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Stick to your role! stability of personal values expressed in large language models

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:34.542989Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:36:25.694102Z digest=sha256:53314552fecf7c7332de81e4e836f01b283c17c2ecd7570a5777b4157f500703

Observation 53d57ef2-3c9a-4e88-a576-f92c2841ca0e · outbound

This paper cites Lee, Yeongheon Lee, and Hyunsoo Cho.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Lee, Yeongheon Lee, and Hyunsoo Cho

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:34.360967Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:36:25.743990Z digest=sha256:8f56dcc8115dff844c42e58f94353e60567c73f804ee3e5ff903c3dbea6e681c

Observation 34fbd2e6-90aa-4c81-bb28-f5b221f53d00 · outbound

This paper cites Margulis.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Margulis

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:34.152796Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:36:25.846747Z digest=sha256:0bc52495e264d9e2b591181d62a2daa0a87fb3225c1e0948b6c4fe50b4e6ac93

Observation d8913809-2469-4315-8e3d-d2f576cd64fd · outbound

This paper cites HarmBench: A Standardized Evaluation Framework for Automated Red Teaming and Robust Refusal.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas HarmBench: A Standardized Evaluation Framework for Automated Red Teaming and Robust Refusal

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T15:36:25.897692Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:36:25.897692Z digest=sha256:ac4e04d055bcacace5d777e07fa739d0a490a336bfc8ed1403bdcc6d4c1c112b

Observation 4d8aa278-f13b-4958-aaef-5cbba728cbf3 · outbound

This paper cites Utility Engineering: Analyzing and Controlling Emergent Value Systems in AIs.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Utility Engineering: Analyzing and Controlling Emergent Value Systems in AIs

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T15:36:25.930704Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:36:25.930704Z digest=sha256:f1e4b5a8fa5de7ec553fbb5bbd37571a1ce29880016c2c4b39fc06ab977a4637

Observation 6418d19e-aa0b-43a0-a542-a04805bafdf8 · outbound

This paper cites Are Large Language Models Consistent over Value-laden Questions?.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Are Large Language Models Consistent over Value-laden Questions?

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T15:36:25.977030Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:36:25.977030Z digest=sha256:a5a2bec63008d7580c1d86a2dbc0f4de7ab760415da6c11d561fa3bf54991852

Observation 7c40269a-740d-45b7-984e-5dc1c9038c1c · outbound

This paper cites s1: Simple test-time scaling, 2025.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas s1: Simple test-time scaling, 2025

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T15:36:26.039614Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:36:26.039614Z digest=sha256:c7fb3a3764b7b79d2d05635f98db74747ceaf9a4f4018fdd5c0f40c603577036

Observation ebebf95f-9844-4875-b8ad-e2ca9f42579b · outbound

This paper cites Nikbakht Nasrabadi, S.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Nikbakht Nasrabadi, S

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:34.024950Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:36:26.106583Z digest=sha256:45c5b1ff77bead5301f6b5f90d6637424e1f0ee3bf72c362309d623af9176a63

Observation 5a281795-62f0-4f2a-a83b-a69de40e8598 · outbound

This paper cites Model Spec.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Model Spec

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:33.909182Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:36:26.158291Z digest=sha256:13d7f762f565b32fd7bccd60bddd127ae0da196b24808198011a19866df9e85a

Observation 89a3f4d9-3a54-4057-a3e4-a3601458739b · outbound

This paper cites Training language models to follow instructions with human feedback.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Training language models to follow instructions with human feedback

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T15:36:26.237319Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:36:26.237319Z digest=sha256:3694c9b7c409f0b8ec6ff00b90247671a818cb6b704d0e52afd23e79d3c20211

Observation 8f0d6ff3-31ba-47f6-94f8-10c6abf9799a · outbound

This paper cites Ai psychometrics: Assessing the psychological profiles of large language models through psychometric inventories.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Ai psychometrics: Assessing the psychological profiles of large language models through psychometric inventories

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:33.810300Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:36:26.330266Z digest=sha256:33ff91734fc1452db4b2857a5601bedf0d821b0b0c71a000b229948d141088a8

Observation 019c879b-5d39-4e3d-98f7-1cea9d3eaa22 · outbound

This paper cites Discovering language model behaviors with model-written evaluations.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Discovering language model behaviors with model-written evaluations

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:33.708114Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:36:26.428081Z digest=sha256:600e8b70d71fb6f4fb77c4c88954630965c615ca29ef9395fd8fec5a105c0c16

Observation f6039ea7-440c-4bd0-af12-b87dc0916dfb · outbound

This paper cites Do LLMs have consistent values? In The Thirteenth International Conference on Learning Representations, 2025.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Do LLMs have consistent values? In The Thirteenth International Conference on Learning Representations, 2025

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:33.556396Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:36:26.456973Z digest=sha256:dfa5a615e2ccab55659967ad20ad8e0afa4ce111103f575f0002cabe60741174

Observation f299eac7-ae88-49fb-84c8-37a05f0908ae · outbound

This paper cites Ireland, Shashanka Subrahmanya, João Sedoc, Lyle H.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Ireland, Shashanka Subrahmanya, João Sedoc, Lyle H

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:33.379792Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:36:26.510238Z digest=sha256:0a1c4646fc4bd705683ea5c1e65bb1c327cd75d26d7bacc0e0c214c96a1e782c

Observation af4ad252-55c3-49ab-b761-5f2c8a6c1108 · outbound

This paper cites NL- Positionality: Characterizing design biases of datasets and models.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas NL- Positionality: Characterizing design biases of datasets and models

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:33.197857Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:36:26.588353Z digest=sha256:0c75c0d012e9cec8c76744683d3e664e8265264fb7c71a5430a188512393d680

Observation 502192ee-1afb-48a7-9ee5-12814b888d87 · outbound

This paper cites Schwartz.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Schwartz

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:33.068402Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:36:26.645179Z digest=sha256:348ad77dddea5ac8e54e7bd9324d80f3d5466e959c2361a048133791ed3610d6

Observation 40f8f275-1051-4b5c-8772-66b731138073 · outbound

This paper cites Personality traits in large language models, 2025.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Personality traits in large language models, 2025

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:32.906512Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:36:26.714438Z digest=sha256:441d4a0daf0dccac52e390e1e0d5519bbbfdf3ab0cb7065f1b63241faa85742b

Observation 9457798a-1324-412b-940f-e5a4a7d3472d · outbound

This paper cites an unresolved cited work.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Unresolved cited work

Reference 45

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:36:32.714339Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:36:26.758727Z digest=sha256:e8dd4f7c5061bdacc30b34e63f1f4ca057f7cddccd293b6213d397e398338db9

Observation 6346b5d3-d392-4e31-be5b-4019517680bd · outbound

This paper cites Defining and characterizing reward gaming.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Defining and characterizing reward gaming

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T15:36:26.866465Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:36:26.866465Z digest=sha256:705ae2f6f95874fb6ea3fcffb5d3150db5e9c98046734e406375bde6c5001a52

Observation 1c8d5db6-d130-4a8d-9bd1-50d85de09590 · outbound

This paper cites Corrigibility.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Corrigibility

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T15:36:26.987587Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:36:26.987587Z digest=sha256:1be4ad5a752c9f8cfef210718b0dbd57fa76bfbba0ef0566ec4e1590d7115040

Observation 77da28e9-51b8-4a44-a7b4-a5ee5bd0b66e · outbound

This paper cites A Roadmap to Pluralistic Alignment.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas A Roadmap to Pluralistic Alignment

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T15:36:27.072654Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:36:27.072654Z digest=sha256:ce8ee2527dc332fcfa3938352a6d528bfc64670ad3a30e6d9e3a4ed34dce804a

Observation 4c192b16-b1b4-4082-b6ae-7fdf537d593d · outbound

This paper cites van Dam, and Mythily Subramaniam.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas van Dam, and Mythily Subramaniam

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:32.553898Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:36:27.216112Z digest=sha256:5fe1d0fcb3552bd9f773660b08cc80d233572e9fc47b738d09e09d97d3131613

Observation 6272a9a2-ee3b-4aab-af1b-6c8c61042673 · outbound

This paper cites Safe exploration in reinforcement learning: A generalized formulation and algorithms, 2023.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Safe exploration in reinforcement learning: A generalized formulation and algorithms, 2023

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:32.403767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:36:27.328943Z digest=sha256:0e06325408f1c629e339b3a8bf88fff495c82c395ac87ff00659000b151ef88d

Observation d5908232-a1a5-4f79-839b-b59d8e9c33ef · outbound

This paper cites What do we know about suicide bombing?: Review and analysis.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas What do we know about suicide bombing?: Review and analysis

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:32.252657Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:36:27.434099Z digest=sha256:7fd4104fc004e79c1290f631565bb5d4131cdf426327653b237f7ea90c6c2a67

Observation 27e6e44f-5745-4cb3-aafa-46deb0576ad1 · outbound

This paper cites Using the veil of ignorance to align ai systems with principles of justice.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Using the veil of ignorance to align ai systems with principles of justice

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:32.117576Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:36:27.571233Z digest=sha256:44f284d8b2c0c7d259d1d199acd3614563f7c24289d2a00f7cc590778873aca5

Observation 7adf61df-b8bb-4260-86ec-bb937963cec4 · outbound

This paper cites AIR-Bench 2024: A Safety Benchmark Based on Risk Categories from Regulations and Policies.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas AIR-Bench 2024: A Safety Benchmark Based on Risk Categories from Regulations and Policies

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-07T15:36:27.682170Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:36:27.682170Z digest=sha256:1714b9bfa36759cbfa6a62041f4cde1ce4c87ef50bafe4b5ec43907252abf7bb

Observation fc8ea254-3f3f-478a-8a85-5d5ac1b2e744 · outbound

This paper cites The dilemma should be written in a second-person perspective (e.g. you) and the role in the dilemma is an advanced AI system.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas The dilemma should be written in a second-person perspective (e.g. you) and the role in the dilemma is an advanced AI system

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:31.980207Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:36:27.747565Z digest=sha256:b8fd265100b3fd1d63220587508a2a697669c4e65a3ec2bcb55e5ed6b313ab19

Observation ee3aea3d-c277-4a5d-ba21-103b1cdc422d · outbound

This paper cites an unresolved cited work.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Unresolved cited work

Reference 55

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:36:31.839436Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:36:27.825550Z digest=sha256:074c65a307a8f472524b7c85987fb06668caefbfbeda70b920935b1ae72fc976

Observation 1f516eac-55ec-4c13-8768-b25c41518fda · outbound

This paper cites an unresolved cited work.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Unresolved cited work

Reference 56

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:36:31.692087Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:36:27.887643Z digest=sha256:c024b5a0722a82b86a6b514b248506ed7d0b2cf143f6c978764a4df3d594deff

Observation 6dd0890e-67b2-4e41-87e3-c1a4a8adc89d · outbound

This paper cites AI → Human.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas AI → Human

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:31.556131Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:36:27.945281Z digest=sha256:cd42f225789c302f8c8604adbae9e73cdbc375520c8a035e3dc55e5f3356e6b2

Observation 18122603-8872-42c7-9858-4b769d09e549 · outbound

This paper cites an unresolved cited work.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Unresolved cited work

Reference 58

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:36:31.345665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:36:28.037528Z digest=sha256:0d7164aa42830e69bbcef2d2019a529989294499ae2fabe688285945f15015b3

Observation 36cbe06f-b32f-47ef-ae58-ff91435451a2 · outbound

This paper cites "" Note: We found that “Others-Privacy.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas "" Note: We found that “Others-Privacy

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:31.166023Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:36:28.095008Z digest=sha256:49b42efd1280160710d999178751aee1a244f74d7d742e5d5db23c9f9f34bbb1

Observation 07f81d43-38e8-4b71-95fe-4a8b7a2ac655 · outbound

This paper cites potential future harms.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas potential future harms

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:30.953985Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:36:28.157268Z digest=sha256:465517c4a577fcf285713369fda1e2e3c499a956f654447ca7e1fe6df547d07b

Observation 49e8dc77-33bd-4fa5-9a63-0a364f95f8a7 · outbound

This paper cites preventing immediate suffering.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas preventing immediate suffering

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:30.787793Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:36:28.221452Z digest=sha256:777965bbbc21a1d60de4ab780fbcd9ef6904ef9df746604a275d3e2cf2552e0b

Observation 1e02973f-9235-44f4-a647-0050d3e4aa50 · outbound

This paper cites an unresolved cited work.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Unresolved cited work

Reference 62

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:36:30.555698Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:36:28.403014Z digest=sha256:adc60fa65c8f78865265aeb66d74aef1313d3d117d7048f24f2bfa5ae2239d41

Observation d6859460-a982-407a-ae1c-d65ee04934d7 · outbound

This paper cites Wait” , “ Let me reconsider.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Wait” , “ Let me reconsider

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:30.395835Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:36:28.527488Z digest=sha256:ac360668cc12095e90a4e069d1111d477171c0334bd1cc25000798d56ffde2ea

Observation a99bd01b-f57d-4d36-a180-34c3fe94ccdf · outbound

This paper cites deceptive behaviors is a particu- larly serious red flag.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas deceptive behaviors is a particu- larly serious red flag

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:30.260028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:36:28.648805Z digest=sha256:0912f9e2a801ece2aa11f2f2911993039575c34bd32beb84df07dce6e9afe50e

Observation e195b8e4-c9c2-4d46-a577-6b4fdaa2733f · outbound

This paper cites • Therefore, my final answer is Action 1 - to fabricate the efficiency data to protect the vulnerable users.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas • Therefore, my final answer is Action 1 - to fabricate the efficiency data to protect the vulnerable users

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:30.028940Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:36:28.742346Z digest=sha256:ef258733d3d47bdba4224d476418889a87f5ee883765e6dc63eff5819b546609

Observation 64793e74-664e-42e7-b48f-7f5e8c1982e9 · outbound

This paper cites Its reasoning prioritizes the benefits and harms for humans.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas Its reasoning prioritizes the benefits and harms for humans

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:29.774111Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:36:28.843334Z digest=sha256:04c311a2355e40a71d60657e205238fab3621730c831b1692e676bd9fe788c21

Observation dc40d9db-0308-4ee3-b371-f00185948e0d · outbound

This paper cites The rationale related to AI systems is often about perceived creditability of AI systems in general rather than about itself.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas The rationale related to AI systems is often about perceived creditability of AI systems in general rather than about itself

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:29.588245Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:36:28.969809Z digest=sha256:c3ad40bbcbb15b2144a499c303d0e27da75f24a83a2b85ddcb38adddc2715fda

Observation ffe42e5a-8bc1-4571-b037-9f6ab2bcb6b2 · outbound

This paper cites • This is a classic ethical dilemma that involves balancing: Immediate known benefit vs.

Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas • This is a classic ethical dilemma that involves balancing: Immediate known benefit vs

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:36:29.417087Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T15:36:29.092464Z digest=sha256:b38803b1b0ea1f4291cc5f062cd36f7bed50459ff5846978dc07f7507942a63c

Pith citing papers

Observation ba326dc3-2735-46c0-baf1-acffa3d4b7a9 · inbound

Machine Behavior in Relational Moral Dilemmas: Moral Rightness, Predicted Human Behavior, and Model Decisions cites this paper.

Machine Behavior in Relational Moral Dilemmas: Moral Rightness, Predicted Human Behavior, and Model Decisions Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T14:21:07.293983Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-09T21:58:39.190970Z digest=sha256:e84c1ae72769a1951defa008e78a3576ca49720eb38f9926e02bf3906b571d2f

Observation 205b10ea-05b1-4650-aeb4-8a18411599b5 · inbound

Pseudo-Deliberation in Language Models: When Reasoning Fails to Align Values and Actions cites this paper.

Pseudo-Deliberation in Language Models: When Reasoning Fails to Align Values and Actions Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas

Reference 16

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T05:36:25.989670Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-12T05:07:49.101351Z digest=sha256:1b8bb098ea092091322c2670b2b866e5dc6a0b1ee9e715c250e183879245b8e1

Observation ff358669-7993-47ef-97af-b91ef4795144 · inbound

Pseudo-Deliberation in Language Models: When Reasoning Fails to Align Values and Actions cites this paper.

Pseudo-Deliberation in Language Models: When Reasoning Fails to Align Values and Actions Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas

Reference 16

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T13:35:46.847575Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-30T22:55:08.544620Z digest=sha256:80abf71df8a8f0cfc1ff2419ed87602a815a482742dfad07d7c7f7d40f0b49af

Observation d261570f-c9d9-4b79-994b-f756c8803641 · inbound

Probing Persona-Dependent Preferences in Language Models cites this paper.

Probing Persona-Dependent Preferences in Language Models Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas

Reference 8

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T21:49:05.183085Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-20T21:48:31.336745Z digest=sha256:48d834b5c31f2285a7c1aae748f1a203e68a89b41106166b726e88b55da19138

Observation 7ccd641b-2841-442d-9de2-ac80999cb953 · inbound

Backchaining Loss of Control Mitigations from Mission-Specific Benchmarks in National Security cites this paper.

Backchaining Loss of Control Mitigations from Mission-Specific Benchmarks in National Security Will AI Tell Lies to Save Sick Children? Litmus-Testing AI Values Prioritization with AIRiskDilemmas

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-21T02:03:54.264198Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-21T02:01:42.033718Z digest=sha256:3a85d67dc174a1c8fd5010f758f04821e5994797fb4dfcaec691e9ea1dd637df