Pith. sign in

Paper Citation Record · LEDGER

Probing the Preferences of a Language Model: Integrating Verbal and Behavioral Tests of AI Welfare

As of 9 August 2026, this Paper Citation Record lists 13 of 13 outbound references and 2 inbound Pith citation observations for arXiv:2509.07961.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.07961 v2

Coverage vector

measured 13 of 13 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-04T21:29:55.556748Z

measured 15 of 15 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-03T09:31:52.605473Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T02:28:24.338817Z

Reference resolution

13 of 13 outbound references displayed

  • verified exact1
  • verified fuzzy4
  • unresolved4
  • parse uncertain0
  • malformed identifier2
  • metadata mismatch2

External citation measurements

0
pith, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation 9ada5a37-453c-4c6f-997e-55f7478739db · outbound

This paper cites and Adams, A.

Probing the Preferences of a Language Model: Integrating Verbal and Behavioral Tests of AI Welfare and Adams, A

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-04T21:29:55.791538Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-04T21:29:55.499137Z digest=sha256:2233b92d351ba95b4cbd2f03f606729c298a4a82a62401afdb2c71f76ac36357

Observation fc4f0029-bc11-4fe9-84ea-4b3093ec70d2 · outbound

This paper cites Anthropic (a) (2025).

Probing the Preferences of a Language Model: Integrating Verbal and Behavioral Tests of AI Welfare Anthropic (a) (2025)

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-04T21:29:55.504327Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T21:29:55.504327Z digest=sha256:e19aa894c57c5833877c9c680c246fcf23d5b67ed596420b44450f4e90f06844

Observation f94a5e3c-05e8-4cb5-b961-2936281b7f84 · outbound

This paper cites Claude Opus 4 and 4.1 can now end a rare subset of conversations.

Probing the Preferences of a Language Model: Integrating Verbal and Behavioral Tests of AI Welfare Claude Opus 4 and 4.1 can now end a rare subset of conversations

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-04T21:29:55.761854Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-04T21:29:55.513265Z digest=sha256:2e8d1c2708aa9b2e8610da3fda9819cfa2cf7741ffee10818bc680b7289f17c0

Observation 32327f00-ae59-4aab-aae6-910381a70a9f · outbound

This paper cites Vending-Bench: A Benchmark for Long-Term Coherence of Autonomous Agents.

Probing the Preferences of a Language Model: Integrating Verbal and Behavioral Tests of AI Welfare Vending-Bench: A Benchmark for Long-Term Coherence of Autonomous Agents

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-04T21:29:55.517898Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T21:29:55.517898Z digest=sha256:9f2348175344bbb443b2bffa7334c37c6d81e4f695c17b16a395b5dabbbfbd88

Observation ab18c93a-3e95-4353-a05c-d5d23b8a8b29 · outbound

This paper cites Building safer dialogue agents.

Probing the Preferences of a Language Model: Integrating Verbal and Behavioral Tests of AI Welfare Building safer dialogue agents

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-04T21:29:55.745883Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-04T21:29:55.534180Z digest=sha256:8157005dbb1d135128c4b5dbf36dd58acd72521dcd73f39ee52bcd0a6500e912

Observation ad4d1973-cf79-4e83-8bda-02b589f19431 · outbound

This paper cites an unresolved cited work.

Probing the Preferences of a Language Model: Integrating Verbal and Behavioral Tests of AI Welfare Unresolved cited work

Reference 9

Resolution
metadata mismatch
arxiv_id, observed 2026-08-04T21:29:55.670994Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-04T21:29:55.538786Z digest=sha256:6f8e3808b3bff7cd9a147d04cf36666b4832c5d4dc8341acc80b5e3cec397bbf

Observation 40536265-2e05-4fbe-b41f-71f84ce98013 · outbound

This paper cites "Understanding AI": Semantic Grounding in Large Language Models.

Probing the Preferences of a Language Model: Integrating Verbal and Behavioral Tests of AI Welfare "Understanding AI": Semantic Grounding in Large Language Models

Reference 10

Resolution
malformed identifier
no resolver link, observed 2026-08-04T21:29:55.543010Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T21:29:55.543010Z digest=sha256:81b273139929469fb3c6309225c8653ee5dd6739ace305fc97b6f3406e9c075b

Observation 7e3b42c4-18da-4fdb-a9ab-e60418513bf7 · outbound

This paper cites Towards Evaluating AI Systems for Moral Status Using Self-Reports.

Probing the Preferences of a Language Model: Integrating Verbal and Behavioral Tests of AI Welfare Towards Evaluating AI Systems for Moral Status Using Self-Reports

Reference 11

Resolution
malformed identifier
no resolver link, observed 2026-08-04T21:29:55.547436Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T21:29:55.547436Z digest=sha256:331e5ed7a70b7445b79dc7c1f6ea14ba898a191e606dd74b459f62e8b6af0375

Observation d3fa58fb-5cad-49cd-a08f-a0d3b9fbad19 · outbound

This paper cites an unresolved cited work.

Probing the Preferences of a Language Model: Integrating Verbal and Behavioral Tests of AI Welfare Unresolved cited work

Reference 12

Resolution
verified exact
doi, observed 2026-08-04T21:29:55.591853Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-04T21:29:55.552197Z digest=sha256:190029b7ee89332ee47f9860fc3332e96b8282af5f9d986328e236022163cbc6

Observation cb09d305-f2f9-4cb2-9224-2f3e8027a665 · outbound

This paper cites and Bradley, A.

Probing the Preferences of a Language Model: Integrating Verbal and Behavioral Tests of AI Welfare and Bradley, A

Reference 719

Resolution
metadata mismatch
arxiv_id, observed 2026-08-04T21:29:55.616132Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-04T21:29:55.556748Z digest=sha256:2501bee32b0e955b31cddb05eb6b3838df628fed0cdb8569ca1e7e38cf4aebff

Observation 72e27767-89b5-4e73-bfe2-7f0733a515db · outbound

This paper cites B., Levine, C.

Probing the Preferences of a Language Model: Integrating Verbal and Behavioral Tests of AI Welfare B., Levine, C

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-04T21:29:55.529247Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T21:29:55.529247Z digest=sha256:913cb2abaee15311bb7d4d20c1a8a9e2738cc789b4faedb8e9758a693cac194f

Observation 1e25c8df-a869-4f6a-a463-8925d68006c0 · outbound

This paper cites Scheming AIs: Will AIs fake alignment during training in order to get power?.

Probing the Preferences of a Language Model: Integrating Verbal and Behavioral Tests of AI Welfare Scheming AIs: Will AIs fake alignment during training in order to get power?

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-04T21:29:55.522992Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T21:29:55.522992Z digest=sha256:4b0543b5122f82c8fe7febdaef303202ac614b539383fd3e0470c53e09767153

Observation cd19d3dd-9b3c-4d04-a211-570145ad0509 · outbound

This paper cites Project Vend: Can Claude run a small shop? (And why does that matter?).

Probing the Preferences of a Language Model: Integrating Verbal and Behavioral Tests of AI Welfare Project Vend: Can Claude run a small shop? (And why does that matter?)

Reference 2025

Resolution
verified fuzzy
raw_fallback, observed 2026-08-04T21:29:55.776977Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-04T21:29:55.508906Z digest=sha256:644546e6b2a7ea013c3b2af5803dc3044a94a3c997a10ba6459a692124167dde

Pith citing papers

Observation 00502c5d-ff7e-4d06-a837-da6cd10cde91 · inbound

No Reliable Evidence of Self-Reported Sentience in Small Large Language Models cites this paper.

No Reliable Evidence of Self-Reported Sentience in Small Large Language Models Probing the Preferences of a Language Model: Integrating Verbal and Behavioral Tests of AI Welfare

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-03T09:31:52.605473Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T09:31:52.605473Z digest=sha256:422b22fdb64e4627052e2b9fce192a35a5d83a6fe6e4d4a7a53827cfa5c67654

Observation 6dc6b8ba-f540-47df-84be-988820fb4214 · inbound

A Scalable Approach to Evaluating Moral Sensitivity in LLMs cites this paper.

A Scalable Approach to Evaluating Moral Sensitivity in LLMs Probing the Preferences of a Language Model: Integrating Verbal and Behavioral Tests of AI Welfare

Reference 169

Resolution
verified exact
local_arxiv, observed 2026-07-12T05:48:31.768267Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-07-12T05:44:33.099337Z digest=sha256:49386b76e3bff6f8eb502d225d2f6db267a3185a93b438cd148b2c67e93155ca