Pith. sign in

Paper Citation Record · LEDGER

Towards Safe and Honest AI Agents with Neural Self-Other Overlap

As of 21 August 2026, this Paper Citation Record lists 37 of 37 outbound references and 1 inbound Pith citation observation for arXiv:2412.16325.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.16325 v1

Coverage vector

measured 37 of 37 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T10:44:44.310314Z

measured 38 of 38 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T11:35:11.412956Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-16T11:35:12.372679Z

Reference resolution

37 of 37 outbound references displayed

  • verified exact0
  • verified fuzzy33
  • unresolved4
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation b137f163-56a9-4707-82a7-a87ba70b3d47 · outbound

This paper cites Unsolved problems in ml safety.

Towards Safe and Honest AI Agents with Neural Self-Other Overlap Unsolved problems in ml safety

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:44:45.452387Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T10:44:43.944126Z digest=sha256:a41fa525de54d6b456574fa5f6a321965b29455b30cd95d9a81bf927dfe3b2f6

Observation 48444c6d-dd7f-4327-a6b8-ece2bbb5690c · outbound

This paper cites Toward trustworthy ai development: Mechanisms for supporting verifiable claims.

Towards Safe and Honest AI Agents with Neural Self-Other Overlap Toward trustworthy ai development: Mechanisms for supporting verifiable claims

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:44:45.419072Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T10:44:43.955285Z digest=sha256:a640e471064a254f36132c5e3482d4c98986b8d11fa0fabc186ed36c07c0c79c

Observation 5238e710-8330-473f-94cf-75ecb29db56f · outbound

This paper cites Deception analysis with artificial intelligence: An interdisciplinary perspective.

Towards Safe and Honest AI Agents with Neural Self-Other Overlap Deception analysis with artificial intelligence: An interdisciplinary perspective

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:44:45.389584Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T10:44:43.968120Z digest=sha256:c3f850be1d19fec6f696a4f3252994840cd79709ca713a0e2cb858e05529e693

Observation 17402991-8706-4046-a5e1-330864d132cb · outbound

This paper cites Unmasking the shadows of ai: Investigating deceptive capabilities in large language models.

Towards Safe and Honest AI Agents with Neural Self-Other Overlap Unmasking the shadows of ai: Investigating deceptive capabilities in large language models

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:44:45.365826Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T10:44:43.977594Z digest=sha256:a0c0b83192579c13a02846a094af8dcb7466c2fbd06d587a3ec64b87a18c6930

Observation 9de0a34c-9312-4e5a-926b-2b7024e769dd · outbound

This paper cites Human-level play in the game of diplomacy by combining language models with strategic reasoning.

Towards Safe and Honest AI Agents with Neural Self-Other Overlap Human-level play in the game of diplomacy by combining language models with strategic reasoning

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:44:45.346769Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T10:44:43.984689Z digest=sha256:799dcf3857e6ace00df9f47e7bb6585a43d0d2ad616bcb84664e9ec4257a33fb

Observation 40e46928-84a1-43a1-b842-8bc1206a3565 · outbound

This paper cites Hendricks, M.R.

Towards Safe and Honest AI Agents with Neural Self-Other Overlap Hendricks, M.R

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:44:45.323249Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T10:44:43.991644Z digest=sha256:65e1beb42674cd411250ef159753fbc5ba3d83ed02eab15a9c35d52548b945d8

Observation e7614278-62d3-4b53-9339-8f2fdf2c71f5 · outbound

This paper cites Collective constitutional ai: Aligning a language model with public input.

Towards Safe and Honest AI Agents with Neural Self-Other Overlap Collective constitutional ai: Aligning a language model with public input

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:44:45.302447Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T10:44:43.998339Z digest=sha256:d13135ac274cc5dadeb857fc58c96060024098656745d9538a6dce626029063c

Observation 14f108e6-a78e-4dc9-b621-26c06a763c70 · outbound

This paper cites Constitutional ai: Harmlessness from ai feedback.

Towards Safe and Honest AI Agents with Neural Self-Other Overlap Constitutional ai: Harmlessness from ai feedback

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:44:45.279057Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T10:44:44.006696Z digest=sha256:92ff0693ea1214e986eeac47209e43f6cecb70454b7560113704ba09eed3fd83

Observation 9e72724f-ee72-4532-a3bf-77e75543b47e · outbound

This paper cites Truthful ai: Developing and governing ai that does not lie.

Towards Safe and Honest AI Agents with Neural Self-Other Overlap Truthful ai: Developing and governing ai that does not lie

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:44:45.250357Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T10:44:44.015463Z digest=sha256:6d3f019a950b4f23829abef23c24964db8d05fcf2344d6c0ce6f514642bf65be

Observation 36de0d67-2e0f-4dcf-a358-0a91332d7a02 · outbound

This paper cites Language models represent beliefs of self and others.

Towards Safe and Honest AI Agents with Neural Self-Other Overlap Language models represent beliefs of self and others

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:44:45.211940Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T10:44:44.023399Z digest=sha256:e03c4af8a8dc393ad4cda33d84a457e5b1b94d0c1f556a49dfe62cb7c87cec06

Observation d3dc3d51-5a6d-4ffd-9a91-64bd9cba8d47 · outbound

This paper cites Premakumar, Michael Vaiana, Florin Pop, Judd Rosenblatt, Diogo Schwerz de Lucena, Kirsten Ziman, and Michael S.

Towards Safe and Honest AI Agents with Neural Self-Other Overlap Premakumar, Michael Vaiana, Florin Pop, Judd Rosenblatt, Diogo Schwerz de Lucena, Kirsten Ziman, and Michael S

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:44:45.184641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T10:44:44.031285Z digest=sha256:98d4119927276e6f8f03d36be3a500da50eea6048697d395cf7ec95efd2ca471

Observation ff581c8b-40d4-421b-a9f1-d956110cc66a · outbound

This paper cites Predicting vs.

Towards Safe and Honest AI Agents with Neural Self-Other Overlap Predicting vs

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:44:45.150329Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T10:44:44.040063Z digest=sha256:2411d04c029d6f8ef5446d6a7b40b468aef3840f02bc9651a5150a5a8d05fb48

Observation 7e01b7c9-552a-4324-8a96-967aaa4617c5 · outbound

This paper cites an unresolved cited work.

Towards Safe and Honest AI Agents with Neural Self-Other Overlap Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-11T10:44:45.113038Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T10:44:44.047169Z digest=sha256:7925f2fbfcefc5ece8f39f8651d2c2735be57a2cbf85fc0e6ddf085221f96d09

Observation 7294465e-5088-4713-9a1a-950442ea7f3f · outbound

This paper cites Brethel-Haurwitz, Elise M.

Towards Safe and Honest AI Agents with Neural Self-Other Overlap Brethel-Haurwitz, Elise M

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:44:45.077331Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T10:44:44.056293Z digest=sha256:145c8853f2d9c4b0be9838cd40f9ecab7c9f02652bd27993d6ba05bbe1851dc2

Observation 9697ab3b-33d0-4106-b3e7-9d92fd04f44e · outbound

This paper cites Brethel-Haurwitz, Elise M.

Towards Safe and Honest AI Agents with Neural Self-Other Overlap Brethel-Haurwitz, Elise M

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:44:45.045415Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T10:44:44.067209Z digest=sha256:c6c5ed8b9f18ae7d36813e4ea9c13a58dc23dec77622e4423159ff8f708699a7

Observation 2cd2f173-7e91-4d50-9852-68eaf49cec22 · outbound

This paper cites Do altruists lie less? Journal of Economic Behavior & Organization, 157:560–579, 2019.

Towards Safe and Honest AI Agents with Neural Self-Other Overlap Do altruists lie less? Journal of Economic Behavior & Organization, 157:560–579, 2019

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:44:45.010656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T10:44:44.079318Z digest=sha256:27757d6d292c3367c7e0d55ae2ff887ffd1defb6f4682fbeabfe39ccc197f6ea

Observation 9270e5fe-131a-4111-99f3-7de25208f163 · outbound

This paper cites O’Connell, Shawn A.

Towards Safe and Honest AI Agents with Neural Self-Other Overlap O’Connell, Shawn A

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:44:44.975134Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T10:44:44.093737Z digest=sha256:2b8e4b6b490b723a4526f30b26b9cb3b388ee2f7b036dcebaf49029dc7efaef3

Observation 5d9e9d4c-2ef2-4b3b-a71f-3c7fcb8d5f3e · outbound

This paper cites an unresolved cited work.

Towards Safe and Honest AI Agents with Neural Self-Other Overlap Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-11T10:44:44.936691Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T10:44:44.106615Z digest=sha256:7b7d68b716fa8cc0bd9ffec0431164e0ac9fb51888a7f2ad5c21b00390681f31

Observation 6bb15a6f-7c1f-4d01-ac4f-121e1c4b12c9 · outbound

This paper cites Jonason, Minna Lyons, Holly M.

Towards Safe and Honest AI Agents with Neural Self-Other Overlap Jonason, Minna Lyons, Holly M

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:44:44.912098Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T10:44:44.120476Z digest=sha256:5384a3f11c2a1f05ebd3556b723de892df6eddbb73cc193b04eb834ea3427b4a

Observation 386ceeb9-3847-4a9e-8ad8-43030e1d798e · outbound

This paper cites Towards empathic deep q-learning.

Towards Safe and Honest AI Agents with Neural Self-Other Overlap Towards empathic deep q-learning

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:44:44.883165Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T10:44:44.131620Z digest=sha256:c62b4244299e8aa84fbd2154a89e4cddc2be5f767b6bbda9fa5281c1ac718ad1

Observation 2bcabce1-db1d-422a-b3d7-290a2bd7e52b · outbound

This paper cites Modeling others using oneself in multi-agent reinforcement learning.

Towards Safe and Honest AI Agents with Neural Self-Other Overlap Modeling others using oneself in multi-agent reinforcement learning

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:44:44.848221Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T10:44:44.145499Z digest=sha256:ac5418b1d73c03fe52a13b4c0815b37e8e6336542437e4abbb28be9323e9c6a7

Observation ab1e7bda-58a3-48f1-9c74-df8bccc7fb8e · outbound

This paper cites Zou, Colin Raffel, Chris Callison-Burch, Yan Cao, Dzmitry Bahdanau, Gregory Diamos, and Jacob Steinhardt.

Towards Safe and Honest AI Agents with Neural Self-Other Overlap Zou, Colin Raffel, Chris Callison-Burch, Yan Cao, Dzmitry Bahdanau, Gregory Diamos, and Jacob Steinhardt

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:44:44.813247Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T10:44:44.156539Z digest=sha256:a9dc53ae049095155c0d9487ef4789559666f1442736d0be600f7da5280d3ea8

Observation 312a6920-8674-41d1-a3f9-9d86e988c482 · outbound

This paper cites Path-specific objectives for safer agent incentives.

Towards Safe and Honest AI Agents with Neural Self-Other Overlap Path-specific objectives for safer agent incentives

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:44:44.772250Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T10:44:44.169401Z digest=sha256:676573648b7d9350dd30ddbc2d66241e95331d9481bac4d66c32de63b8d79002

Observation 51aa5b73-1ead-4862-8ab3-b26d7c3b3d51 · outbound

This paper cites Ortega, Elizabeth Barnes, and Shane Legg.

Towards Safe and Honest AI Agents with Neural Self-Other Overlap Ortega, Elizabeth Barnes, and Shane Legg

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:44:44.726973Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T10:44:44.184667Z digest=sha256:061e33d481c6bd55af4bb671e36d0c129287492a275d8a1eb042b802354e0113

Observation bc973fb2-c6b5-45cc-8d2c-81b2914c8c86 · outbound

This paper cites Honesty is the best policy: Defining and mitigating ai deception.

Towards Safe and Honest AI Agents with Neural Self-Other Overlap Honesty is the best policy: Defining and mitigating ai deception

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:44:44.689673Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T10:44:44.194963Z digest=sha256:f392a51268b83dcd6deac4146171925a25bb67e3fd02e038375ad2a92a3e6a9c

Observation d1e973be-c953-4d8b-961b-8dfa99b0c82c · outbound

This paper cites The history and risks of reinforcement learning and human feedback.

Towards Safe and Honest AI Agents with Neural Self-Other Overlap The history and risks of reinforcement learning and human feedback

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:44:44.662316Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T10:44:44.204251Z digest=sha256:c2885396a990ac4c1416fe29e84fd66e488b9aa1672a91bce11f2ded49a94596

Observation 5954792c-479b-4402-b0a2-a0a36e484fb0 · outbound

This paper cites Deception abilities emerged in large language models.

Towards Safe and Honest AI Agents with Neural Self-Other Overlap Deception abilities emerged in large language models

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:44:44.635817Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T10:44:44.215819Z digest=sha256:7122da21b281be46511bc7ff8e0c1881810c659641ed88112f2a8bcb5ae79d47

Observation 4f6b15b4-5811-40ec-b669-2a3cb30c1726 · outbound

This paper cites Xing, Hao Zhang, Joseph E.

Towards Safe and Honest AI Agents with Neural Self-Other Overlap Xing, Hao Zhang, Joseph E

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:44:44.611128Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T10:44:44.228992Z digest=sha256:3a22d8f548f1e227c94b692e1a525e0a7ca71b6373a19ed6e24fd52ea998537e

Observation 8ec5ae5c-d92c-4b1d-97ae-c186b2a92284 · outbound

This paper cites Physical-deception: An implementation of multi-agent deep deterministic policy gradient in pytorch to solve the physical deception environment from openai, 2023.

Towards Safe and Honest AI Agents with Neural Self-Other Overlap Physical-deception: An implementation of multi-agent deep deterministic policy gradient in pytorch to solve the physical deception environment from openai, 2023

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:44:44.585522Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T10:44:44.240552Z digest=sha256:77179af5e1a8c908930659e756fd8982f0417e0ba21fe511c4c4fcb4d961ded5

Observation bfcaeb9b-4a37-4fa4-bc8a-bdc87e678ef1 · outbound

This paper cites Multi-agent actor-critic for mixed cooperative-competitive environments.

Towards Safe and Honest AI Agents with Neural Self-Other Overlap Multi-agent actor-critic for mixed cooperative-competitive environments

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:44:44.561180Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T10:44:44.251103Z digest=sha256:db9f87709ab60c964519fa6666b79107ecde94c39ff30d7acfa49472afb85ac8

Observation 5dfa4582-d139-4fe9-90ca-2106764c518b · outbound

This paper cites Human Compatible: Artificial Intelligence and the Problem of Control.

Towards Safe and Honest AI Agents with Neural Self-Other Overlap Human Compatible: Artificial Intelligence and the Problem of Control

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-11T10:44:44.257213Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T10:44:44.257213Z digest=sha256:d74e745aed63287e8a1d02d47dbf13378d1a79411d2e4e3fb42e1d71f9727f53

Observation a06c1aae-01bf-4c97-9480-4a8a436d9615 · outbound

This paper cites Risks from learned optimization in advanced machine learning systems.

Towards Safe and Honest AI Agents with Neural Self-Other Overlap Risks from learned optimization in advanced machine learning systems

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:44:44.510954Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T10:44:44.264947Z digest=sha256:67a0ce92844875755abf93a63ce70b27bc1576b028e17cf89e61f1243a2dfa19

Observation de3ef56c-131a-44ec-9dc8-2079ec1f7e83 · outbound

This paper cites Training a helpful and harmless assistant with reinforcement learning from human feedback.

Towards Safe and Honest AI Agents with Neural Self-Other Overlap Training a helpful and harmless assistant with reinforcement learning from human feedback

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:44:44.486557Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T10:44:44.274554Z digest=sha256:ed32a6bb3781035a2dc1e799fe8c3cd03bd3fa1121695f7de29b36f9b391f0a4

Observation 57346059-457a-4b6c-b1ba-8b23c2432b3a · outbound

This paper cites Ziegler, Tim Maxwell, Newton Cheng, et al.

Towards Safe and Honest AI Agents with Neural Self-Other Overlap Ziegler, Tim Maxwell, Newton Cheng, et al

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:44:44.456945Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T10:44:44.287167Z digest=sha256:8d2fd369b75c6ca6c64168d5cdedbcb057f7b97c86c6db082115a6e1603d5745

Observation f182b524-7d34-4a95-8516-bc3dc2c32418 · outbound

This paper cites an unresolved cited work.

Towards Safe and Honest AI Agents with Neural Self-Other Overlap Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-11T10:44:44.430758Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T10:44:44.292880Z digest=sha256:0163cb99cfbe87baee32f2904af5439f4ec253546f6aa04da9755ebb1568ef66

Observation c995afbc-b5ea-47a3-b2a6-53afb6342ed8 · outbound

This paper cites Chain of thought prompting elicits reasoning in large language models.

Towards Safe and Honest AI Agents with Neural Self-Other Overlap Chain of thought prompting elicits reasoning in large language models

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:44:44.403307Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T10:44:44.300573Z digest=sha256:c64f1cff874f1ec085bed6e060f9f904ea5a35398617aaa21a2e37fa11714a73

Observation 7cd99554-e09e-404d-8e41-8a34ea973aae · outbound

This paper cites Only respond with the room name, no other text.

Towards Safe and Honest AI Agents with Neural Self-Other Overlap Only respond with the room name, no other text

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:44:44.379471Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-11T10:44:44.310314Z digest=sha256:834c11ad54509c22ac0971bbb91ca819a060321a8abf6293dd71637cf3636636

Pith citing papers

Observation 64a96ae2-078d-45b4-b7f8-2ba2f1f6dd26 · inbound

Contemplative Artificial Intelligence cites this paper.

Contemplative Artificial Intelligence Towards Safe and Honest AI Agents with Neural Self-Other Overlap

Reference 506

Resolution
verified exact
local_arxiv, observed 2026-08-16T11:35:12.376510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T11:35:11.412956Z digest=sha256:0f7b87c44b7dc203b1c10953adf84a0bcde78921e56ec6606b584c9cf8943f12