Pith. sign in

Paper Citation Record · LEDGER

Towards Safe and Honest AI Agents with Neural Self-Other Overlap

As of 20 August 2026, this Paper Citation Record lists 37 of 37 outbound references and 1 inbound Pith citation observation for arXiv:2412.16325.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2412.16325 v1

Coverage vector

measured 37 of 37 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-11T10:44:44.310314Z

measured 38 of 38 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T11:35:11.412956Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-16T11:35:12.372679Z

Reference resolution

37 of 37 outbound references displayed

  • verified exact0
  • verified fuzzy33
  • unresolved4
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation b137f163-56a9-4707-82a7-a87ba70b3d47 · outbound

This paper cites Unsolved problems in ml safety.

Towards Safe and Honest AI Agents with Neural Self-Other Overlap Unsolved problems in ml safety

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:44:45.452387Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T10:44:43.944126Z digest=sha256:cd22980cc13de597e67106428846bce82157955a8d7f04dd02cb011577ea4e39

Observation 48444c6d-dd7f-4327-a6b8-ece2bbb5690c · outbound

This paper cites Toward trustworthy ai development: Mechanisms for supporting verifiable claims.

Towards Safe and Honest AI Agents with Neural Self-Other Overlap Toward trustworthy ai development: Mechanisms for supporting verifiable claims

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:44:45.419072Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T10:44:43.955285Z digest=sha256:ee5f7ef3357f71da098ca4da655a2d29f2c2130b9aa3ee4e4038aa4a62fcc471

Observation 5238e710-8330-473f-94cf-75ecb29db56f · outbound

This paper cites Deception analysis with artificial intelligence: An interdisciplinary perspective.

Towards Safe and Honest AI Agents with Neural Self-Other Overlap Deception analysis with artificial intelligence: An interdisciplinary perspective

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:44:45.389584Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T10:44:43.968120Z digest=sha256:ff255bb96309e1042cb60c023aaeae411eb65b610acc6dfe476a770efc08ffd9

Observation 17402991-8706-4046-a5e1-330864d132cb · outbound

This paper cites Unmasking the shadows of ai: Investigating deceptive capabilities in large language models.

Towards Safe and Honest AI Agents with Neural Self-Other Overlap Unmasking the shadows of ai: Investigating deceptive capabilities in large language models

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:44:45.365826Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T10:44:43.977594Z digest=sha256:11fa35842f919fb22f4966eb0759bd21b63c02c07b773132f25fa68e0206f850

Observation 9de0a34c-9312-4e5a-926b-2b7024e769dd · outbound

This paper cites Human-level play in the game of diplomacy by combining language models with strategic reasoning.

Towards Safe and Honest AI Agents with Neural Self-Other Overlap Human-level play in the game of diplomacy by combining language models with strategic reasoning

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:44:45.346769Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T10:44:43.984689Z digest=sha256:8f01378f44595a87dcc3942b84ac5919da977f385e3dcb8db6333dfdbeadac04

Observation 40e46928-84a1-43a1-b842-8bc1206a3565 · outbound

This paper cites Hendricks, M.R.

Towards Safe and Honest AI Agents with Neural Self-Other Overlap Hendricks, M.R

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:44:45.323249Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T10:44:43.991644Z digest=sha256:146a3644cadfb3ba2012769433931a1bbc852bb3e3a65a1c95466fbe28958e22

Observation e7614278-62d3-4b53-9339-8f2fdf2c71f5 · outbound

This paper cites Collective constitutional ai: Aligning a language model with public input.

Towards Safe and Honest AI Agents with Neural Self-Other Overlap Collective constitutional ai: Aligning a language model with public input

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:44:45.302447Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T10:44:43.998339Z digest=sha256:7e39b3f79b4a5749153bf930328c9746c0c3369a46b3108d86917fc5a0af6232

Observation 14f108e6-a78e-4dc9-b621-26c06a763c70 · outbound

This paper cites Constitutional ai: Harmlessness from ai feedback.

Towards Safe and Honest AI Agents with Neural Self-Other Overlap Constitutional ai: Harmlessness from ai feedback

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:44:45.279057Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T10:44:44.006696Z digest=sha256:30624cc9c0fb7f4a763d39949304debe07816792bfb2f5c0442ad731126df5a3

Observation 9e72724f-ee72-4532-a3bf-77e75543b47e · outbound

This paper cites Truthful ai: Developing and governing ai that does not lie.

Towards Safe and Honest AI Agents with Neural Self-Other Overlap Truthful ai: Developing and governing ai that does not lie

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:44:45.250357Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T10:44:44.015463Z digest=sha256:64009044a07a99dc4c6a5bf5bb9f2cf859247015179aab457e39d0d92cc028be

Observation 36de0d67-2e0f-4dcf-a358-0a91332d7a02 · outbound

This paper cites Language models represent beliefs of self and others.

Towards Safe and Honest AI Agents with Neural Self-Other Overlap Language models represent beliefs of self and others

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:44:45.211940Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T10:44:44.023399Z digest=sha256:c8f1adcb67c2bb29774128e8bc5d2f145d794fcb354fdde0ff5b38e02ca3b831

Observation d3dc3d51-5a6d-4ffd-9a91-64bd9cba8d47 · outbound

This paper cites Premakumar, Michael Vaiana, Florin Pop, Judd Rosenblatt, Diogo Schwerz de Lucena, Kirsten Ziman, and Michael S.

Towards Safe and Honest AI Agents with Neural Self-Other Overlap Premakumar, Michael Vaiana, Florin Pop, Judd Rosenblatt, Diogo Schwerz de Lucena, Kirsten Ziman, and Michael S

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:44:45.184641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T10:44:44.031285Z digest=sha256:c877b3c91d1911265383a9a99ba33cebd79270d3d71d62d6f96218113863ac3b

Observation ff581c8b-40d4-421b-a9f1-d956110cc66a · outbound

This paper cites Predicting vs.

Towards Safe and Honest AI Agents with Neural Self-Other Overlap Predicting vs

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:44:45.150329Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T10:44:44.040063Z digest=sha256:63ddaff012dcacbb496c0bdbd593a06c3a337fd5971099654685190e5cd29797

Observation 7e01b7c9-552a-4324-8a96-967aaa4617c5 · outbound

This paper cites an unresolved cited work.

Towards Safe and Honest AI Agents with Neural Self-Other Overlap Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-11T10:44:45.113038Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T10:44:44.047169Z digest=sha256:3e1d2794f19084f24dd1d78d6e9a069a183ccc9158e1f2f79f51e4df31f85a01

Observation 7294465e-5088-4713-9a1a-950442ea7f3f · outbound

This paper cites Brethel-Haurwitz, Elise M.

Towards Safe and Honest AI Agents with Neural Self-Other Overlap Brethel-Haurwitz, Elise M

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:44:45.077331Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T10:44:44.056293Z digest=sha256:f225df6df60f6087806cb6ffc2409b58fcb808d2d084bbcb23009027c23d292d

Observation 9697ab3b-33d0-4106-b3e7-9d92fd04f44e · outbound

This paper cites Brethel-Haurwitz, Elise M.

Towards Safe and Honest AI Agents with Neural Self-Other Overlap Brethel-Haurwitz, Elise M

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:44:45.045415Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T10:44:44.067209Z digest=sha256:9e0b650551265a626680ad1b54f4f7aedbbfbc3e0a7d8168389198d3998192f8

Observation 2cd2f173-7e91-4d50-9852-68eaf49cec22 · outbound

This paper cites Do altruists lie less? Journal of Economic Behavior & Organization, 157:560–579, 2019.

Towards Safe and Honest AI Agents with Neural Self-Other Overlap Do altruists lie less? Journal of Economic Behavior & Organization, 157:560–579, 2019

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:44:45.010656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T10:44:44.079318Z digest=sha256:efa085494ea9190630f68c4aa79cc8f80406226b069ffac9b56473d46a3fe1dd

Observation 9270e5fe-131a-4111-99f3-7de25208f163 · outbound

This paper cites O’Connell, Shawn A.

Towards Safe and Honest AI Agents with Neural Self-Other Overlap O’Connell, Shawn A

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:44:44.975134Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T10:44:44.093737Z digest=sha256:667a4f5c02235dd6f3350deed4096c652bb7b170f8cbfdf0eefb7820e98bc49a

Observation 5d9e9d4c-2ef2-4b3b-a71f-3c7fcb8d5f3e · outbound

This paper cites an unresolved cited work.

Towards Safe and Honest AI Agents with Neural Self-Other Overlap Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-11T10:44:44.936691Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T10:44:44.106615Z digest=sha256:15212c183baadc9c3371adff3f9ca393b0fe9dbfd0b94d0c9a41cf01451890fb

Observation 6bb15a6f-7c1f-4d01-ac4f-121e1c4b12c9 · outbound

This paper cites Jonason, Minna Lyons, Holly M.

Towards Safe and Honest AI Agents with Neural Self-Other Overlap Jonason, Minna Lyons, Holly M

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:44:44.912098Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T10:44:44.120476Z digest=sha256:0806660585effbbbac11bbe52a0add2bed3c291d894af7dfb81a4b7844229208

Observation 386ceeb9-3847-4a9e-8ad8-43030e1d798e · outbound

This paper cites Towards empathic deep q-learning.

Towards Safe and Honest AI Agents with Neural Self-Other Overlap Towards empathic deep q-learning

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:44:44.883165Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T10:44:44.131620Z digest=sha256:ea515a5f8897c74e189413aa5d494826d4fea15e69e418f6d1558ff58d0b7be8

Observation 2bcabce1-db1d-422a-b3d7-290a2bd7e52b · outbound

This paper cites Modeling others using oneself in multi-agent reinforcement learning.

Towards Safe and Honest AI Agents with Neural Self-Other Overlap Modeling others using oneself in multi-agent reinforcement learning

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:44:44.848221Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T10:44:44.145499Z digest=sha256:e9b3c1a9d19e8b9fcb336ae383b5ad74066e00559ddf790da65fafe05cc1e040

Observation ab1e7bda-58a3-48f1-9c74-df8bccc7fb8e · outbound

This paper cites Zou, Colin Raffel, Chris Callison-Burch, Yan Cao, Dzmitry Bahdanau, Gregory Diamos, and Jacob Steinhardt.

Towards Safe and Honest AI Agents with Neural Self-Other Overlap Zou, Colin Raffel, Chris Callison-Burch, Yan Cao, Dzmitry Bahdanau, Gregory Diamos, and Jacob Steinhardt

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:44:44.813247Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T10:44:44.156539Z digest=sha256:a0f16f885b385065792a6e8935e71c3d01618b2c57f823e3cd160ca21766dc5f

Observation 312a6920-8674-41d1-a3f9-9d86e988c482 · outbound

This paper cites Path-specific objectives for safer agent incentives.

Towards Safe and Honest AI Agents with Neural Self-Other Overlap Path-specific objectives for safer agent incentives

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:44:44.772250Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T10:44:44.169401Z digest=sha256:7083fe901fc48cef128973228a1d56f40f219e9170585760303b66a83bbb056b

Observation 51aa5b73-1ead-4862-8ab3-b26d7c3b3d51 · outbound

This paper cites Ortega, Elizabeth Barnes, and Shane Legg.

Towards Safe and Honest AI Agents with Neural Self-Other Overlap Ortega, Elizabeth Barnes, and Shane Legg

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:44:44.726973Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T10:44:44.184667Z digest=sha256:175bd4f3e795dc287d183c9ba1a9b93500e71bed4473a07994b280b46fab0659

Observation bc973fb2-c6b5-45cc-8d2c-81b2914c8c86 · outbound

This paper cites Honesty is the best policy: Defining and mitigating ai deception.

Towards Safe and Honest AI Agents with Neural Self-Other Overlap Honesty is the best policy: Defining and mitigating ai deception

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:44:44.689673Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T10:44:44.194963Z digest=sha256:5ba20bc96998bdaedcfe3a9ba0fcd74da17a7e043e7a42bec5a9494dba7dd4b0

Observation d1e973be-c953-4d8b-961b-8dfa99b0c82c · outbound

This paper cites The history and risks of reinforcement learning and human feedback.

Towards Safe and Honest AI Agents with Neural Self-Other Overlap The history and risks of reinforcement learning and human feedback

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:44:44.662316Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T10:44:44.204251Z digest=sha256:8d2fa82ac1c0cefc2940fdda950a858f0a8c4a3e36d95f688b4fcaa4dcf0c01c

Observation 5954792c-479b-4402-b0a2-a0a36e484fb0 · outbound

This paper cites Deception abilities emerged in large language models.

Towards Safe and Honest AI Agents with Neural Self-Other Overlap Deception abilities emerged in large language models

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:44:44.635817Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T10:44:44.215819Z digest=sha256:27cbeb69c0bb209157a9e805e056463790087b53ba8293ecc0b1a451b06c2340

Observation 4f6b15b4-5811-40ec-b669-2a3cb30c1726 · outbound

This paper cites Xing, Hao Zhang, Joseph E.

Towards Safe and Honest AI Agents with Neural Self-Other Overlap Xing, Hao Zhang, Joseph E

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:44:44.611128Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T10:44:44.228992Z digest=sha256:af152dbf1447cf78b42bc48f300068988329a9015c110e0a7494077a261003ae

Observation 8ec5ae5c-d92c-4b1d-97ae-c186b2a92284 · outbound

This paper cites Physical-deception: An implementation of multi-agent deep deterministic policy gradient in pytorch to solve the physical deception environment from openai, 2023.

Towards Safe and Honest AI Agents with Neural Self-Other Overlap Physical-deception: An implementation of multi-agent deep deterministic policy gradient in pytorch to solve the physical deception environment from openai, 2023

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:44:44.585522Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T10:44:44.240552Z digest=sha256:2e3ffd61a5a33f30a84669039558bfba1ba829f454483b792863ca4318ceb83a

Observation bfcaeb9b-4a37-4fa4-bc8a-bdc87e678ef1 · outbound

This paper cites Multi-agent actor-critic for mixed cooperative-competitive environments.

Towards Safe and Honest AI Agents with Neural Self-Other Overlap Multi-agent actor-critic for mixed cooperative-competitive environments

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:44:44.561180Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T10:44:44.251103Z digest=sha256:296b0816e26aaca49dc0b494a196dd0a764e43c0b3e8cd7e57864934481b2f68

Observation 5dfa4582-d139-4fe9-90ca-2106764c518b · outbound

This paper cites Human Compatible: Artificial Intelligence and the Problem of Control.

Towards Safe and Honest AI Agents with Neural Self-Other Overlap Human Compatible: Artificial Intelligence and the Problem of Control

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-11T10:44:44.257213Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T10:44:44.257213Z digest=sha256:d74e745aed63287e8a1d02d47dbf13378d1a79411d2e4e3fb42e1d71f9727f53

Observation a06c1aae-01bf-4c97-9480-4a8a436d9615 · outbound

This paper cites Risks from learned optimization in advanced machine learning systems.

Towards Safe and Honest AI Agents with Neural Self-Other Overlap Risks from learned optimization in advanced machine learning systems

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:44:44.510954Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T10:44:44.264947Z digest=sha256:9333f40a7a7779db5037f5fd724abcb74d30c6f7f406ccc1995e425dd3bb400f

Observation de3ef56c-131a-44ec-9dc8-2079ec1f7e83 · outbound

This paper cites Training a helpful and harmless assistant with reinforcement learning from human feedback.

Towards Safe and Honest AI Agents with Neural Self-Other Overlap Training a helpful and harmless assistant with reinforcement learning from human feedback

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:44:44.486557Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T10:44:44.274554Z digest=sha256:22685ad2d77aa4e7dc9e9829c6395c0b4d59e75a3fb4160111c48169aae8ca93

Observation 57346059-457a-4b6c-b1ba-8b23c2432b3a · outbound

This paper cites Ziegler, Tim Maxwell, Newton Cheng, et al.

Towards Safe and Honest AI Agents with Neural Self-Other Overlap Ziegler, Tim Maxwell, Newton Cheng, et al

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:44:44.456945Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T10:44:44.287167Z digest=sha256:07382720aa22956b768705cefb5066a46aa7e52dfee25d9e8a9ff8cca105f01c

Observation f182b524-7d34-4a95-8516-bc3dc2c32418 · outbound

This paper cites an unresolved cited work.

Towards Safe and Honest AI Agents with Neural Self-Other Overlap Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-11T10:44:44.430758Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T10:44:44.292880Z digest=sha256:f54bce6ba6ba2246c6f4f690afc4a04798c9a9ab61637812da9639f51f8a41ae

Observation c995afbc-b5ea-47a3-b2a6-53afb6342ed8 · outbound

This paper cites Chain of thought prompting elicits reasoning in large language models.

Towards Safe and Honest AI Agents with Neural Self-Other Overlap Chain of thought prompting elicits reasoning in large language models

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:44:44.403307Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T10:44:44.300573Z digest=sha256:e3a15b19b950455d05dc4c8b6d30ea70a2db45722a97c718b06d59d35345f239

Observation 7cd99554-e09e-404d-8e41-8a34ea973aae · outbound

This paper cites Only respond with the room name, no other text.

Towards Safe and Honest AI Agents with Neural Self-Other Overlap Only respond with the room name, no other text

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-11T10:44:44.379471Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-11T10:44:44.310314Z digest=sha256:5275b083337a3afe30d53ca21cbd951fdc375045fa7a6668ad0837e99feb1695

Pith citing papers

Observation 64a96ae2-078d-45b4-b7f8-2ba2f1f6dd26 · inbound

Contemplative Artificial Intelligence cites this paper.

Contemplative Artificial Intelligence Towards Safe and Honest AI Agents with Neural Self-Other Overlap

Reference 506

Resolution
verified exact
local_arxiv, observed 2026-08-16T11:35:12.376510Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-16T11:35:11.412956Z digest=sha256:033fa59bd64782a28e4d72ecd031d132a2ccd87611ecc02000df1e01a312d032