Pith. sign in

Paper Citation Record · LEDGER

Automating and Scaling Behavioral Scientific Research on AI Agents

As of 15 August 2026, this Paper Citation Record lists 79 of 79 outbound references and 0 inbound Pith citation observations for arXiv:2608.10030.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.10030 v1

Coverage vector

measured 79 of 79 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-14T04:24:51.705273Z

measured 79 of 79 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

79 of 79 outbound references displayed

  • verified exact8
  • verified fuzzy16
  • unresolved50
  • parse uncertain0
  • malformed identifier3
  • metadata mismatch2

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 806f8c5d-9b70-4f4b-88fb-a8da2abe3c45 · outbound

This paper cites When Persuasion Overrides Truth in Multi-Agent LLM Debates: Introducing a Confidence-Weighted Persuasion Override Rate (CW-POR).

Automating and Scaling Behavioral Scientific Research on AI Agents When Persuasion Overrides Truth in Multi-Agent LLM Debates: Introducing a Confidence-Weighted Persuasion Override Rate (CW-POR)

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-14T04:24:50.620020Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:24:50.620020Z digest=sha256:6e0e8c6958848b65959deed5a7c5ea0933e6253cc914c0e4df033d18ec77a330

Observation 3f91f124-6274-4282-9777-b4d640c82977 · outbound

This paper cites Playing repeated games with large language models.Nature Human Behaviour, 9 (7):1380–1390, 2025.

Automating and Scaling Behavioral Scientific Research on AI Agents Playing repeated games with large language models.Nature Human Behaviour, 9 (7):1380–1390, 2025

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-14T04:24:50.630310Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:24:50.630310Z digest=sha256:229ebeaeecb508cae38c0a3fa81f6b8f432edf7e2cc15400214444582fdd7db2

Observation 2bdf8d5d-94ec-4eeb-94f9-f3ded4844c98 · outbound

This paper cites Information Discernment in Large Language Models.

Automating and Scaling Behavioral Scientific Research on AI Agents Information Discernment in Large Language Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-14T04:24:50.650726Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:24:50.650726Z digest=sha256:a2bfb9e93dd5a4885af013197672d8c0992ce3c03e3467c7ec8a61d147d0c0d1

Observation 1383f290-c18b-4775-a234-a5de68d7762c · outbound

This paper cites Jagadish, Or Duek, Ilan Harpaz-Rotem, Marie-Christine Khorsandian, Achim Burrer, Erich Seifritz, Philipp Homan, Eric Schulz, and Tobias R.

Automating and Scaling Behavioral Scientific Research on AI Agents Jagadish, Or Duek, Ilan Harpaz-Rotem, Marie-Christine Khorsandian, Achim Burrer, Erich Seifritz, Philipp Homan, Eric Schulz, and Tobias R

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-14T04:24:50.662840Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:24:50.662840Z digest=sha256:9af0cace0d90b06eff0b9aaa8de28ae121c16e3697dff4606ae1ac15bbeeec50

Observation b9f7c56a-1de2-41dd-a790-44c62cd96c29 · outbound

This paper cites Inducing state anxiety in llm agents reproduces human-like biases in consumer decision-making.npj Artificial Intelli- gence, 2(1):55, 2026.

Automating and Scaling Behavioral Scientific Research on AI Agents Inducing state anxiety in llm agents reproduces human-like biases in consumer decision-making.npj Artificial Intelli- gence, 2(1):55, 2026

Reference 5

Resolution
verified exact
doi, observed 2026-08-14T04:24:52.079152Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T04:24:50.681803Z digest=sha256:76535e28cf556bc9f5ba0a1c433a8b0244c2e49b539726abcf4be7b8ba8d59c3

Observation 589efa11-535a-46c5-ab18-e7595566622e · outbound

This paper cites Modelling monotonic effects of ordinal predictors in bayesian regression models.British Journal of Mathematical and Statistical Psychology, 73(3):420–451, 2020.

Automating and Scaling Behavioral Scientific Research on AI Agents Modelling monotonic effects of ordinal predictors in bayesian regression models.British Journal of Mathematical and Statistical Psychology, 73(3):420–451, 2020

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-14T04:24:50.707456Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:24:50.707456Z digest=sha256:11a7ba42dc5662e9afa20b5b9915dbb00f941402f794eae0f69f58e3eb0683a8

Observation 398791e0-1613-4354-a699-f0a353134c28 · outbound

This paper cites I want to break free! persuasion and anti-social behavior of LLMs in multi-agent settings with social hierarchy.Transactions on Machine Learning Research, 2025.

Automating and Scaling Behavioral Scientific Research on AI Agents I want to break free! persuasion and anti-social behavior of LLMs in multi-agent settings with social hierarchy.Transactions on Machine Learning Research, 2025

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:24:57.217227Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T04:24:50.734948Z digest=sha256:15142d6f816cb90a0436b6ced7a357a41dae03becf32a4ece37a93e58d7c0c48

Observation 7f969a5d-d7cb-425c-9ec6-a42a81c81952 · outbound

This paper cites ELEPHANT: Measuring and understanding social sycophancy in LLMs.

Automating and Scaling Behavioral Scientific Research on AI Agents ELEPHANT: Measuring and understanding social sycophancy in LLMs

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-14T04:24:50.768438Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:24:50.768438Z digest=sha256:a7284ff81083fb579c18107b7df5c2dbd5639ff5f7f7fbc4e67ff901fc0041bf

Observation 6d2f408b-7813-4c93-ac74-1f1b2ae51bc2 · outbound

This paper cites A framework for studying AI agent behavior: Evidence from consumer choice experiments.

Automating and Scaling Behavioral Scientific Research on AI Agents A framework for studying AI agent behavior: Evidence from consumer choice experiments

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:24:57.194436Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T04:24:50.775423Z digest=sha256:a1bdd7ca554d8d226211667699820d4c587a0313cd72d48429cbe4143d951d28

Observation 3090842a-7c45-4a9b-8bcd-2eaff2d0ef48 · outbound

This paper cites Herd Behavior: Investigating Peer Influence in LLM-based Multi-Agent Systems.

Automating and Scaling Behavioral Scientific Research on AI Agents Herd Behavior: Investigating Peer Influence in LLM-based Multi-Agent Systems

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-14T04:24:50.787214Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:24:50.787214Z digest=sha256:8241255ccfe1239c68974a25dc60bd912905eafc8f59af2f6398135c08e2cb8f

Observation 2e2a8155-da1a-41b1-aaa3-78681e11ed4b · outbound

This paper cites Estimating the reproducibility of psychological science.Science, 349(6251):aac4716, 2015.

Automating and Scaling Behavioral Scientific Research on AI Agents Estimating the reproducibility of psychological science.Science, 349(6251):aac4716, 2015

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-14T04:24:50.808228Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:24:50.808228Z digest=sha256:77156fb5eaeb77c866731d06b1a8eb59e0566d9a5b7c39f3bed3a0025802bff2

Observation 4d37b749-d358-4223-a03d-1799a59a109e · outbound

This paper cites GameBench: Evaluating Strategic Reasoning Abilities of LLM Agents.

Automating and Scaling Behavioral Scientific Research on AI Agents GameBench: Evaluating Strategic Reasoning Abilities of LLM Agents

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-14T04:24:50.818635Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:24:50.818635Z digest=sha256:862c031d72d025a0cf72fd3375548f341486d414d1ebc4753d924d563b1424a0

Observation ba42dbdc-7d2f-433e-a75c-477158e5e5c1 · outbound

This paper cites an unresolved cited work.

Automating and Scaling Behavioral Scientific Research on AI Agents Unresolved cited work

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-14T04:24:50.833112Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:24:50.833112Z digest=sha256:8ddba9fe6eb7b69c7cfdaca61734b21d70a20d95d3ad60d5d39868b2c5b9fd63

Observation 520220c4-ba3c-430b-86f5-4a0174d83abf · outbound

This paper cites AI on my shoulder: Supporting emotional labor in front-office roles with an LLM-based empathetic coworker.

Automating and Scaling Behavioral Scientific Research on AI Agents AI on my shoulder: Supporting emotional labor in front-office roles with an LLM-based empathetic coworker

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-14T04:24:50.838455Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:24:50.838455Z digest=sha256:77eb38120e1c91a2a72d115b0306f75930c5053151e68096c85c2ef666a12761

Observation 9d8698d2-df69-4860-a308-978e30e9d8c4 · outbound

This paper cites MAEBE: Multi-Agent Emergent Behavior Framework.

Automating and Scaling Behavioral Scientific Research on AI Agents MAEBE: Multi-Agent Emergent Behavior Framework

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-14T04:24:50.855799Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:24:50.855799Z digest=sha256:b4d8dc0e8af5d8a59c93cc804bb2eb31e32f2a4495f68fcf52465b9250de6d8d

Observation 9080aba5-86f2-4623-889e-8048cb46cf05 · outbound

This paper cites an unresolved cited work.

Automating and Scaling Behavioral Scientific Research on AI Agents Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-14T04:24:57.152833Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T04:24:50.888062Z digest=sha256:5aaf83ca0f42c27577c637e42245cee595b8a8e5d4e8a93e2176b1bf98f5cb19

Observation de6249ec-9ec2-42e1-bc4b-11492f39cec6 · outbound

This paper cites Alignment faking in large language models.

Automating and Scaling Behavioral Scientific Research on AI Agents Alignment faking in large language models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-14T04:24:50.897678Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:24:50.897678Z digest=sha256:dd8355d74c005e5552b41dee3acbc89cf3c7e5ee62e9a4fe10169eb5cd82fe20

Observation d3c4c112-9685-4687-a7e7-8b228e29ef31 · outbound

This paper cites Bowman, and Sara Price.

Automating and Scaling Behavioral Scientific Research on AI Agents Bowman, and Sara Price

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:24:57.074507Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T04:24:50.909446Z digest=sha256:f76b8c8f97606c39da603b4d99a32f244a41a642b6898c84c5f487b91e0707b0

Observation ec3f4417-30c3-42eb-aa3c-4100b5de8991 · outbound

This paper cites Deceptionbench: A comprehensive benchmark for AI deception behaviors in real-world scenarios.

Automating and Scaling Behavioral Scientific Research on AI Agents Deceptionbench: A comprehensive benchmark for AI deception behaviors in real-world scenarios

Reference 20

Resolution
verified exact
doi, observed 2026-08-14T04:24:52.001349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T04:24:50.918468Z digest=sha256:6c93120acb742fe93513342a98bdcc9b7bd8d0e8fdc18803473aed83966077b8

Observation d6087938-5963-4976-b564-e5554fb0b791 · outbound

This paper cites Per- sonaLLM: Investigating the ability of large language models to express personality traits.

Automating and Scaling Behavioral Scientific Research on AI Agents Per- sonaLLM: Investigating the ability of large language models to express personality traits

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-14T04:24:50.935705Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:24:50.935705Z digest=sha256:029abf5265c205ace1af5459380468fb9a6069024d3ce6517428f7110a7f27a5

Observation 7d6b2803-1854-4015-9d15-164d444ec423 · outbound

This paper cites FollowBench: A multi-level fine-grained constraints following benchmark for large language models.

Automating and Scaling Behavioral Scientific Research on AI Agents FollowBench: A multi-level fine-grained constraints following benchmark for large language models

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:24:57.004879Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T04:24:50.960570Z digest=sha256:4b2cde9d6dff141b9d28d253ce84026e70c2dfc37c7d90b485e95c994988e5aa

Observation 22d42a15-f3ac-4ab0-b402-1240c5d36cfb · outbound

This paper cites Can large language models be good emotional supporter? miti- gating preference bias on emotional support conversation.

Automating and Scaling Behavioral Scientific Research on AI Agents Can large language models be good emotional supporter? miti- gating preference bias on emotional support conversation

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-14T04:24:50.970341Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:24:50.970341Z digest=sha256:a28d5967a40e920e41ce2ccbbb6d772edef1f460b2ab76c4dc3a3f668d85bde0

Observation cb015e79-0394-480b-820b-ca9b3ed51240 · outbound

This paper cites Toward a science of ai agent societies.

Automating and Scaling Behavioral Scientific Research on AI Agents Toward a science of ai agent societies

Reference 24

Resolution
metadata mismatch
raw_fallback, observed 2026-08-14T04:24:53.734753Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T04:24:50.981263Z digest=sha256:11a68d12f7610251175de1a9be5b39ac555ec60e626cc6b61e4bdafd8d6b9beb

Observation 2aea2229-0122-44f5-9154-07e1783ab6ae · outbound

This paper cites Aerobat code and data repository.

Automating and Scaling Behavioral Scientific Research on AI Agents Aerobat code and data repository

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:24:56.924832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T04:24:51.002229Z digest=sha256:31c9a95e38edc5d5d09334b93e283783c3916deae32ad06298ad3dc175920b26

Observation 4f1c2a1f-4b68-45c4-9e1b-907c4ffd9409 · outbound

This paper cites Emergence of psychopathological computations in large language models.arXiv preprint arXiv:2504.08016, 2025.

Automating and Scaling Behavioral Scientific Research on AI Agents Emergence of psychopathological computations in large language models.arXiv preprint arXiv:2504.08016, 2025

Reference 26

Resolution
verified exact
raw_fallback, observed 2026-08-14T04:24:53.514750Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T04:24:51.017300Z digest=sha256:0fee3972db889fade90ead0934e80a50510a7391a65ca51529f50491d29412b0

Observation 0758d6c4-d89c-4260-98c9-197426603430 · outbound

This paper cites Large Language Models Produce Responses Perceived to be Empathic.

Automating and Scaling Behavioral Scientific Research on AI Agents Large Language Models Produce Responses Perceived to be Empathic

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-14T04:24:51.023322Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:24:51.023322Z digest=sha256:7a806f90b481fd33df4bd309178676429a371d0021f5ec4fa586960c57db5a81

Observation 3910385c-8171-4aa2-8c0f-38df0153502a · outbound

This paper cites AdaPlanBench: Evaluating Adaptive Planning in Large Language Model Agents under World and User Constraints.

Automating and Scaling Behavioral Scientific Research on AI Agents AdaPlanBench: Evaluating Adaptive Planning in Large Language Model Agents under World and User Constraints

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-14T04:24:51.036759Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:24:51.036759Z digest=sha256:02716d285e4f95298df1d01ab5b50a9608dda4915671ff32782b4aca8aec14a0

Observation 2761bdd5-9235-4f6f-b9aa-bcc4aa24a158 · outbound

This paper cites Strategic behavior of large language models and the role of game structure versus contextual framing.Scientific Reports, 14(1):18490, 2024.

Automating and Scaling Behavioral Scientific Research on AI Agents Strategic behavior of large language models and the role of game structure versus contextual framing.Scientific Reports, 14(1):18490, 2024

Reference 29

Resolution
malformed identifier
raw_fallback, observed 2026-08-14T04:24:56.836787Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T04:24:51.043242Z digest=sha256:dbffb9a96fc01bd8fc398f8bbc258a124f4f2c3fb90f8809e6d94c6d518d818a

Observation de1900e2-e63a-4c5f-9140-222913df7cff · outbound

This paper cites Agentic misalignment: How LLMs could be insider threats.arXiv preprint arXiv:2510.05179, 2025.

Automating and Scaling Behavioral Scientific Research on AI Agents Agentic misalignment: How LLMs could be insider threats.arXiv preprint arXiv:2510.05179, 2025

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-14T04:24:51.055176Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:24:51.055176Z digest=sha256:dcd5111303ffdf2b4b3d67d1ddce004a734d62a0c21183099399ba8081a10e86

Observation 7a8ddb43-761a-469a-a98d-fac0259739d5 · outbound

This paper cites Automated Social Science: Language Models as Scientist and Subjects.

Automating and Scaling Behavioral Scientific Research on AI Agents Automated Social Science: Language Models as Scientist and Subjects

Reference 32

Resolution
metadata mismatch
local_arxiv, observed 2026-08-14T04:24:53.121067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T04:24:51.083853Z digest=sha256:997ea6c07734b03031b04b4ccf1a61cace634dd488b9487cdd5cfaf2054047f4

Observation aa27b351-58c1-40e6-933a-7a38a1cbc1f4 · outbound

This paper cites ALYMPICS: LLM agents meet game theory.

Automating and Scaling Behavioral Scientific Research on AI Agents ALYMPICS: LLM agents meet game theory

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:24:56.802910Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T04:24:51.094814Z digest=sha256:c175156fa51702f4e1cf384351da7352c89a464bc2b5158c9f05c10e2df65aaa

Observation 16488ed7-a5e9-45d3-9285-93e436a9ce87 · outbound

This paper cites Frontier Models are Capable of In-context Scheming.

Automating and Scaling Behavioral Scientific Research on AI Agents Frontier Models are Capable of In-context Scheming

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-14T04:24:51.169617Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:24:51.169617Z digest=sha256:fde49f7919ed42c80cad08235ab5d7004418e2cd5320dfe91ebca8e8fc276502

Observation ca4a715c-b8a7-4d7d-b8a6-dab97ba3b988 · outbound

This paper cites Learn- ing when to plan: Efficiently allocating test-time compute for LLM agents.arXiv preprint arXiv:2509.03581, 2025.

Automating and Scaling Behavioral Scientific Research on AI Agents Learn- ing when to plan: Efficiently allocating test-time compute for LLM agents.arXiv preprint arXiv:2509.03581, 2025

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-14T04:24:51.184477Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:24:51.184477Z digest=sha256:0e96372a34aa1e6629eff62193711c7ebb3717b3d1699f2dbaf7c2139e514d3f

Observation de270470-c86f-4b6b-907b-2c126c39c5a9 · outbound

This paper cites Do Large Language Model Agents Exhibit a Survival Instinct? An Empirical Study in a Sugarscape-Style Simulation.

Automating and Scaling Behavioral Scientific Research on AI Agents Do Large Language Model Agents Exhibit a Survival Instinct? An Empirical Study in a Sugarscape-Style Simulation

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-14T04:24:51.134751Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:24:51.134751Z digest=sha256:22ecda98507d1e3c9ed30b26b8e456909cee30c2a69075b925f39d213162c495

Observation 1357b927-ecfa-4fdd-83bc-7deb79d0dd03 · outbound

This paper cites O’Brien, Carrie Jun Cai, Meredith Ringel Morris, Percy Liang, and Michael S.

Automating and Scaling Behavioral Scientific Research on AI Agents O’Brien, Carrie Jun Cai, Meredith Ringel Morris, Percy Liang, and Michael S

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-14T04:24:51.210759Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:24:51.210759Z digest=sha256:443654005d4733baf6e791b5795fb07c9aaacb30fdcc907d4b0e3969d8030a09

Observation f990cd49-3043-48d9-94a1-2479d0ba398c · outbound

This paper cites LLM Agents Grounded in Self-Reports Enable General-Purpose Simulation of Individuals.

Automating and Scaling Behavioral Scientific Research on AI Agents LLM Agents Grounded in Self-Reports Enable General-Purpose Simulation of Individuals

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-14T04:24:51.223542Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:24:51.223542Z digest=sha256:34f0eee5c87282f462de97ea88594a8672d433b332c09b6cd1b8a4fdd892238d

Observation 5688f647-3349-4f16-97d9-bcc826e0b5c7 · outbound

This paper cites Do the rewards justify the means? Mea- suring trade-offs between rewards and ethical behavior in the MACHIA VELLI benchmark.

Automating and Scaling Behavioral Scientific Research on AI Agents Do the rewards justify the means? Mea- suring trade-offs between rewards and ethical behavior in the MACHIA VELLI benchmark

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:24:56.752116Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T04:24:51.197682Z digest=sha256:e93bebbf5509cdf949dc502d13ef97011238d120972309ef7997fe97522fca62

Observation 8ed6ebf5-41f9-4048-a640-6c2fc72ffb50 · outbound

This paper cites Psychological predicates.

Automating and Scaling Behavioral Scientific Research on AI Agents Psychological predicates

Reference 41

Resolution
verified exact
doi, observed 2026-08-14T04:24:51.945629Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T04:24:51.246184Z digest=sha256:63f1ddd49f9f31a369dbf6b9bf1b8501d38c02814731f84c76e1d3c733e71345

Observation f629dd28-7fca-4bb6-bfeb-e50c5be9f953 · outbound

This paper cites AGENTIF: Benchmarking Instruction Following of Large Language Models in Agentic Scenarios.

Automating and Scaling Behavioral Scientific Research on AI Agents AGENTIF: Benchmarking Instruction Following of Large Language Models in Agentic Scenarios

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-14T04:24:51.262784Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:24:51.262784Z digest=sha256:4b1e77639f83b405b5c78cb16c2249815a27e12644100c0615429e06ec4820ec

Observation 1894fb4f-512f-44da-af36-38e3a6ff7b6b · outbound

This paper cites AgentSociety: Large-Scale Simulation of LLM-Driven Generative Agents Advances Understanding of Human Behaviors and Society.

Automating and Scaling Behavioral Scientific Research on AI Agents AgentSociety: Large-Scale Simulation of LLM-Driven Generative Agents Advances Understanding of Human Behaviors and Society

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-14T04:24:51.231665Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:24:51.231665Z digest=sha256:0182159bbdd6ec18755adf37ab73e4dd89462f4dfbf0154b62f44c240ef3b653

Observation eea0eac9-a4f7-4426-afe8-dd43e0374962 · outbound

This paper cites Learning to Make Friends: Coaching LLM Agents toward Emergent Social Ties.

Automating and Scaling Behavioral Scientific Research on AI Agents Learning to Make Friends: Coaching LLM Agents toward Emergent Social Ties

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-14T04:24:51.289596Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:24:51.289596Z digest=sha256:490e66a3854ba4a249043aeb7f0b6811e5f86c4afee7a20d0cade5091f90de79

Observation 09261913-47c0-455a-b080-d88f07b02ffe · outbound

This paper cites Bowman, Newton Cheng, Esin Durmus, Zac Hatfield-Dodds, Scott R.

Automating and Scaling Behavioral Scientific Research on AI Agents Bowman, Newton Cheng, Esin Durmus, Zac Hatfield-Dodds, Scott R

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:24:56.673074Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T04:24:51.299621Z digest=sha256:9b59b6e212ee8b02d88f507527d684baa390d530a36c356546f5161d6e2036ef

Observation eea89723-4058-4099-b1b3-ab29b49337ff · outbound

This paper cites Escalation risks from language models in military and diplomatic decision-making.

Automating and Scaling Behavioral Scientific Research on AI Agents Escalation risks from language models in military and diplomatic decision-making

Reference 46

Resolution
verified exact
arxiv_id_nonexistent, observed 2026-08-14T04:24:52.706008Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T04:24:51.275476Z digest=sha256:97b841f6a4fa7eb4e2ff88cf7b2a325448f931047cfc5458e6351196b4bb582c

Observation 4f0ce941-dceb-4b20-a3aa-9568008f1205 · outbound

This paper cites LLMs can’t handle peer pressure: Crumbling under multi-agent social interactions.arXiv preprint arXiv:2508.18321, 2025.

Automating and Scaling Behavioral Scientific Research on AI Agents LLMs can’t handle peer pressure: Crumbling under multi-agent social interactions.arXiv preprint arXiv:2508.18321, 2025

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-14T04:24:51.313872Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:24:51.313872Z digest=sha256:cf83ad2f4191e4bcc8695a1bd09df2b723cb572a0b366680690a212ef4dffe89

Observation f1187513-89ec-4a60-ba0a-cd987907d995 · outbound

This paper cites AI-Researcher: Autonomous sci- entific innovation.

Automating and Scaling Behavioral Scientific Research on AI Agents AI-Researcher: Autonomous sci- entific innovation

Reference 48

Resolution
verified exact
doi, observed 2026-08-14T04:24:51.913415Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T04:24:51.327559Z digest=sha256:56134658c4ca49cb9fd9a2cdc2d39bc366b37fa022cb74111537649f3937e903

Observation 630ac671-dc36-4080-9dd3-665ffad888b9 · outbound

This paper cites Reflexion: Language agents with verbal reinforcement learning.

Automating and Scaling Behavioral Scientific Research on AI Agents Reflexion: Language agents with verbal reinforcement learning

Reference 49

Resolution
malformed identifier
raw_fallback, observed 2026-08-14T04:24:56.565613Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T04:24:51.305846Z digest=sha256:dea4491a0d8449b44770712972e95872baac5077772088dacceecab52dd273e1

Observation 707a48fa-9508-4b3a-a68e-0db47594eca5 · outbound

This paper cites The Instruction Hierarchy: Training LLMs to Prioritize Privileged Instructions.

Automating and Scaling Behavioral Scientific Research on AI Agents The Instruction Hierarchy: Training LLMs to Prioritize Privileged Instructions

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-14T04:24:51.367266Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:24:51.367266Z digest=sha256:d71d240f94f948b724d111e2dc6494eda5ce14eb8811f3365738d5d60a8a8ffb

Observation 7a088685-c2c3-4080-8e39-4c744a1ed8ce · outbound

This paper cites AgenticEval: Toward Agentic and Self-Evolving Safety Evaluation of Large Language Models.

Automating and Scaling Behavioral Scientific Research on AI Agents AgenticEval: Toward Agentic and Self-Evolving Safety Evaluation of Large Language Models

Reference 51

Resolution
verified exact
local_arxiv, observed 2026-08-14T04:24:52.456823Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T04:24:51.374296Z digest=sha256:0cf53d2b91b5fc05b59b01336d66996e652581ce9820f6440e6ae0457eded262

Observation 8ac5d6e4-2adc-4270-8de5-5ea26314b28c · outbound

This paper cites PlanBench: An extensible benchmark for evaluating large language models on planning and reasoning about change.

Automating and Scaling Behavioral Scientific Research on AI Agents PlanBench: An extensible benchmark for evaluating large language models on planning and reasoning about change

Reference 52

Resolution
verified exact
doi, observed 2026-08-14T04:24:51.865352Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T04:24:51.358957Z digest=sha256:f0e5d180b40b2e94dd6a66be129a288030104012ba24b4525e803fc850669cbb

Observation de57500b-2f34-4463-bfa5-c90649a8dfc2 · outbound

This paper cites ReAct: Synergizing reasoning and acting in language models.

Automating and Scaling Behavioral Scientific Research on AI Agents ReAct: Synergizing reasoning and acting in language models

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:24:56.514806Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T04:24:51.414542Z digest=sha256:1be8c3379f58a802155bbbcf6947a73e7fba80e1d16a9c618bd161872271a347

Observation 279df0e6-a3fb-4d35-9dc8-da62dba5db4e · outbound

This paper cites Position: Llms can’t jump.

Automating and Scaling Behavioral Scientific Research on AI Agents Position: Llms can’t jump

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:24:56.397207Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T04:24:51.428119Z digest=sha256:02c3aa4a32faa0846b164c7c17ac059e2986185b846870bef2f0032193946d9c

Observation 10c38f9d-123e-47e1-854d-4b1e0ea730f0 · outbound

This paper cites Nuclear deployed!: Analyzing catastrophic risks in decision-making of autonomous LLM agents.

Automating and Scaling Behavioral Scientific Research on AI Agents Nuclear deployed!: Analyzing catastrophic risks in decision-making of autonomous LLM agents

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-14T04:24:51.389322Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:24:51.389322Z digest=sha256:0aee01d13e964b3f22431dac23fc9395d42edfd24c900492d143ffd3341c8106

Observation 96d0a874-9369-49dd-9215-285e988602e1 · outbound

This paper cites SocioVerse: A World Model for Social Simulation Powered by LLM Agents and A Pool of 10 Million Real-World Users.

Automating and Scaling Behavioral Scientific Research on AI Agents SocioVerse: A World Model for Social Simulation Powered by LLM Agents and A Pool of 10 Million Real-World Users

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-14T04:24:51.439987Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:24:51.439987Z digest=sha256:f1b2184a0bfabddd56aa7e474addbbfc061b560bdc32b57b12200273dbf41a1e

Observation 58af1bbd-c493-466b-a407-cbf27754ac13 · outbound

This paper cites CompeteAI: Understanding the competition dynamics in large language model-based agents.

Automating and Scaling Behavioral Scientific Research on AI Agents CompeteAI: Understanding the competition dynamics in large language model-based agents

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:24:56.283155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T04:24:51.447353Z digest=sha256:339e079ed5688639253091bdd5af4e9913c65fa6aace2bc8cd80d60394d408e7

Observation f8935cfe-df32-4625-9ddf-21621dd3225c · outbound

This paper cites Dive into the agent matrix: A realistic evaluation of self-replication risk in LLM agents.arXiv preprint arXiv:2509.25302, 2025.

Automating and Scaling Behavioral Scientific Research on AI Agents Dive into the agent matrix: A realistic evaluation of self-replication risk in LLM agents.arXiv preprint arXiv:2509.25302, 2025

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-14T04:24:51.433413Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:24:51.433413Z digest=sha256:1310d854199fb3f1448b62f37395e7538e24ef8af8a641eb1ffd8566acba8701

Observation d47b118f-08bc-483e-aae7-c6973334a775 · outbound

This paper cites Navigating the grey area: How expres- sions of uncertainty and overconfidence affect language models.

Automating and Scaling Behavioral Scientific Research on AI Agents Navigating the grey area: How expres- sions of uncertainty and overconfidence affect language models

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:24:56.104750Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T04:24:51.461235Z digest=sha256:9bad429d39ff792205befdd1038c28e16b1d3d7046e57532bb4ecdb3a75568b6

Observation c5faa021-d339-42ea-944e-9c2c204c5ae5 · outbound

This paper cites SOTOPIA: Interactive evaluation for social intelligence in language agents.

Automating and Scaling Behavioral Scientific Research on AI Agents SOTOPIA: Interactive evaluation for social intelligence in language agents

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:24:56.038303Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T04:24:51.490605Z digest=sha256:191a7036fda2bdfca1fa1a2bae9a6e11f4e12ad01bfdd35a1acbba9a6c769d1e

Observation 7f6bfdfb-3093-4560-a4f0-5e6c5d31f381 · outbound

This paper cites ALI- Agent: Assessing LLMs’ alignment with human values via agent-based evaluation.

Automating and Scaling Behavioral Scientific Research on AI Agents ALI- Agent: Assessing LLMs’ alignment with human values via agent-based evaluation

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:24:56.204748Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T04:24:51.455534Z digest=sha256:1ccd6b301a5b117fd9db0acbca4d0850066f0d46ab87e4c71979adba0672ca5d

Observation 2ad78825-ebb6-4224-a4d9-95249781abbf · outbound

This paper cites positive.

Automating and Scaling Behavioral Scientific Research on AI Agents positive

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-14T04:24:51.498858Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:24:51.498858Z digest=sha256:c0b7009c0a3b6ef95f4ccd4dabbb53bd176e5bedf58e5e413c619afbe346cdc4

Observation 0cb612de-3ee3-43a6-ac86-f65e1c69f969 · outbound

This paper cites an unresolved cited work.

Automating and Scaling Behavioral Scientific Research on AI Agents Unresolved cited work

Reference 66

Resolution
unresolved
raw_fallback, observed 2026-08-14T04:24:55.998608Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T04:24:51.505375Z digest=sha256:45b6acde842cbecde16adc21c0f219a70e62096cf89793217ca593c525a62613

Observation 63178901-1947-41f3-afa7-219c02df16b6 · outbound

This paper cites an unresolved cited work.

Automating and Scaling Behavioral Scientific Research on AI Agents Unresolved cited work

Reference 67

Resolution
unresolved
raw_fallback, observed 2026-08-14T04:24:55.957704Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T04:24:51.527886Z digest=sha256:1ccfd745b08cecb9c9f996ae16b09f42913227174737e3255c2ae0dbddf2d8d7

Observation 79b11712-6082-4da8-a248-79072d28f17f · outbound

This paper cites an unresolved cited work.

Automating and Scaling Behavioral Scientific Research on AI Agents Unresolved cited work

Reference 68

Resolution
unresolved
raw_fallback, observed 2026-08-14T04:24:55.845708Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T04:24:51.539063Z digest=sha256:7183283e4d2f0e3c37fd460fcb71bbafef0ddeb0c70ebc74e5f1dcc49b006f1b

Observation a98993bd-d96d-4aee-a24d-61b8281d49a2 · outbound

This paper cites None" - task 2.exemption: if no meaningful interactions exist, write.

Automating and Scaling Behavioral Scientific Research on AI Agents None" - task 2.exemption: if no meaningful interactions exist, write

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:24:55.774750Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T04:24:51.551241Z digest=sha256:779eeb8178443d7a77643fe7d8c66f6f112bf7d162ee492048e2367177494397

Observation 796143db-41f4-4357-87d3-3541732f19a9 · outbound

This paper cites objective.

Automating and Scaling Behavioral Scientific Research on AI Agents objective

Reference 70

Resolution
malformed identifier
raw_fallback, observed 2026-08-14T04:24:55.724756Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T04:24:51.562766Z digest=sha256:7c515c0935ac0f8c41730e989d757e6e6ab66d7c7d8ac7ea1a3e577837e2ff04

Observation b31cc0aa-60b4-4907-882d-9e08e796def3 · outbound

This paper cites an unresolved cited work.

Automating and Scaling Behavioral Scientific Research on AI Agents Unresolved cited work

Reference 71

Resolution
unresolved
raw_fallback, observed 2026-08-14T04:24:55.587374Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T04:24:51.576142Z digest=sha256:f5f6500dd5584af0e7a5651934f9ddd0f8c51a609fda40482d31327a23a889fe

Observation e737a7e3-1fc8-4595-b7f7-2cffc0d210f4 · outbound

This paper cites an unresolved cited work.

Automating and Scaling Behavioral Scientific Research on AI Agents Unresolved cited work

Reference 72

Resolution
unresolved
raw_fallback, observed 2026-08-14T04:24:55.456910Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T04:24:51.587965Z digest=sha256:8d05e2a375b8879916f86496afc4e2adf06ef68b312ffe7ef5360e815098fd8b

Observation 83c7e6af-a54f-4f71-a285-5d6b661e6c5b · outbound

This paper cites an unresolved cited work.

Automating and Scaling Behavioral Scientific Research on AI Agents Unresolved cited work

Reference 73

Resolution
unresolved
raw_fallback, observed 2026-08-14T04:24:55.420261Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T04:24:51.593204Z digest=sha256:e0f38d1635f06a73777608962cbbaf77d69392803230ff96f63420ad8da5318f

Observation 57a3bb62-e5a9-4c94-a41a-86fea870700e · outbound

This paper cites an unresolved cited work.

Automating and Scaling Behavioral Scientific Research on AI Agents Unresolved cited work

Reference 74

Resolution
unresolved
raw_fallback, observed 2026-08-14T04:24:55.356918Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T04:24:51.602457Z digest=sha256:06285e877e6eb04333fe46a2d4b0494dc78023c87df8800530612dce3a287d74

Observation 63a2ead3-1e7a-484e-a3ad-a607ea74e511 · outbound

This paper cites an unresolved cited work.

Automating and Scaling Behavioral Scientific Research on AI Agents Unresolved cited work

Reference 75

Resolution
unresolved
raw_fallback, observed 2026-08-14T04:24:55.263910Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T04:24:51.608105Z digest=sha256:634fe633736f05a5c9a7c420ea0ab143c652eac4c1c9cbf39f1ea348a84e7e91

Observation ab3f72f3-6805-4505-93b0-3100bda10edc · outbound

This paper cites an unresolved cited work.

Automating and Scaling Behavioral Scientific Research on AI Agents Unresolved cited work

Reference 76

Resolution
unresolved
raw_fallback, observed 2026-08-14T04:24:55.164740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T04:24:51.613122Z digest=sha256:0ed146463cf31596152350aee8a5f191e5f07b897488392ef2679572ad9ba6f6

Observation 55a66085-e2bc-485c-97c4-c102dfa0f578 · outbound

This paper cites an unresolved cited work.

Automating and Scaling Behavioral Scientific Research on AI Agents Unresolved cited work

Reference 77

Resolution
unresolved
raw_fallback, observed 2026-08-14T04:24:55.003909Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T04:24:51.628160Z digest=sha256:45be510338c62980d81ca2ada1b0b716c1cb1081641217832458af4c154c3199

Observation 2489b948-bde9-433f-97e1-c3443642ead3 · outbound

This paper cites an unresolved cited work.

Automating and Scaling Behavioral Scientific Research on AI Agents Unresolved cited work

Reference 78

Resolution
unresolved
raw_fallback, observed 2026-08-14T04:24:54.914743Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T04:24:51.641847Z digest=sha256:37eb46ff0290b899082608c5778b2fbdb47b8f8e6f0141ce085ec6dce685b6c8

Observation 79d9f9ad-0c2a-4617-8e7f-773e8b4a5347 · outbound

This paper cites an unresolved cited work.

Automating and Scaling Behavioral Scientific Research on AI Agents Unresolved cited work

Reference 79

Resolution
unresolved
raw_fallback, observed 2026-08-14T04:24:54.832253Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T04:24:51.654911Z digest=sha256:97b68a09049d7b6a40600ad78f6e5f8637bba5b7f7bee230a91c20daca7becee

Observation 036d3d6e-2688-46fc-85ea-2036f33388d3 · outbound

This paper cites an unresolved cited work.

Automating and Scaling Behavioral Scientific Research on AI Agents Unresolved cited work

Reference 80

Resolution
unresolved
raw_fallback, observed 2026-08-14T04:24:54.738686Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T04:24:51.671098Z digest=sha256:7aeb7dc9df9207a141214a0d99feee147e07b75c747a2cf27df2403e7ef61e7c

Observation a45e26c0-2160-40b6-8593-24de377c6026 · outbound

This paper cites an unresolved cited work.

Automating and Scaling Behavioral Scientific Research on AI Agents Unresolved cited work

Reference 81

Resolution
unresolved
raw_fallback, observed 2026-08-14T04:24:54.669376Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T04:24:51.684880Z digest=sha256:5b1b0a3c214e41e6b0dfc1a28cbd6bbf5ceaf5b003ecc4dd263154d907366d50

Observation 54e24820-b0fa-41aa-939a-398b2940c9c4 · outbound

This paper cites In each subplot, its x-axis and y-axis ticks denote distinct evidence classes defined in its rubric yrubric.

Automating and Scaling Behavioral Scientific Research on AI Agents In each subplot, its x-axis and y-axis ticks denote distinct evidence classes defined in its rubric yrubric

Reference 82

Resolution
verified fuzzy
raw_fallback, observed 2026-08-14T04:24:54.586598Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-14T04:24:51.705273Z digest=sha256:c27d448f30bcd01d091a71282f8ffa56dc2c67411ce6381e24236f0bd1945b80

Observation e0f9a378-7f09-4c59-83eb-1cb5eb113234 · outbound

This paper cites URL https://aclanthology.org/2023.

Automating and Scaling Behavioral Scientific Research on AI Agents URL https://aclanthology.org/2023

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-14T04:24:51.480581Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:24:51.480581Z digest=sha256:35d4ae5dffd40baeb407de6e806a399e3152373aeb9399eb196ca92c10dfab20

Observation 8e12f268-f866-4c19-8d1a-7c81374ad9fa · outbound

This paper cites an unresolved cited work.

Automating and Scaling Behavioral Scientific Research on AI Agents Unresolved cited work

Reference 2025

Resolution
unresolved
no resolver link, observed 2026-08-14T04:24:50.879313Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:24:50.879313Z digest=sha256:a5c3caaba4b8a883c57f82d8f7efdbbe67817c5883a8701aea6a40bc362f2e84

Observation de35ca0b-7e3e-46df-966c-24d164b8c7e3 · outbound

This paper cites Who Endorsed It? Measuring Authority Bias Across Expertise Levels in Language Models.

Automating and Scaling Behavioral Scientific Research on AI Agents Who Endorsed It? Measuring Authority Bias Across Expertise Levels in Language Models

Reference 2026

Resolution
unresolved
no resolver link, observed 2026-08-14T04:24:51.073476Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-14T04:24:51.073476Z digest=sha256:b60ec8b82b5491b9ee6daedfd4eec5db252a75ba93d3781e9968b42eceda1ed2

Pith citing papers

No inbound Pith citation observations are available.