Pith. sign in

Paper Citation Record · LEDGER

Interpretable Risk Mitigation in LLM Agent Systems

As of 18 August 2026, this Paper Citation Record lists 77 of 77 outbound references and 3 inbound Pith citation observations for arXiv:2505.10670.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.10670 v1

Coverage vector

measured 77 of 77 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T21:11:02.149478Z

measured 80 of 80 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T04:51:56.450815Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-17T00:31:24.634056Z

Reference resolution

77 of 77 outbound references displayed

  • verified exact2
  • verified fuzzy24
  • unresolved49
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c43eafac-f59d-4479-8e8e-f89c0279a822 · outbound

This paper cites Artificial intelligence and the future of work: Evidence from OECD countries.

Interpretable Risk Mitigation in LLM Agent Systems Artificial intelligence and the future of work: Evidence from OECD countries

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:11:03.643143Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:11:01.782535Z digest=sha256:99522b61a1151fbea3f396245014bcad51b26299b9fa2536e84ebb170c9560c9

Observation bf3d3a32-c723-4ab7-a21a-243b73e9489b · outbound

This paper cites Dai, Chelsea Finn, Justin Fu, Kanishka Gopalakrishnan, et al.

Interpretable Risk Mitigation in LLM Agent Systems Dai, Chelsea Finn, Justin Fu, Kanishka Gopalakrishnan, et al

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:11:03.627827Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:11:01.787366Z digest=sha256:2f25a7efecc3fa8451a62347e7ac9d94d083d3a4cd4fe544bc65ed079a46e2bd

Observation 3801ed77-8037-4a96-b87d-1b592e49f25f · outbound

This paper cites Mistral 7b: Open foundation models, 2023.

Interpretable Risk Mitigation in LLM Agent Systems Mistral 7b: Open foundation models, 2023

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:11:03.612474Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:11:01.792100Z digest=sha256:839a9aa20eca3f6daee89d14baff0e31ebbe5063a98239c477fdc16441eb2264

Observation 54ed5ba8-5bf9-4e48-a28b-7b79fb7685e8 · outbound

This paper cites Playing repeated games with Large Language Models.

Interpretable Risk Mitigation in LLM Agent Systems Playing repeated games with Large Language Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-15T21:11:01.796811Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:11:01.796811Z digest=sha256:6ca29cfc59c340584f5b203cd765e5bce4de7d393f10b82246de7654a43706db

Observation 9b7eff9b-d167-419d-ae41-da6ca8c46fe7 · outbound

This paper cites Concrete Problems in AI Safety.

Interpretable Risk Mitigation in LLM Agent Systems Concrete Problems in AI Safety

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-15T21:11:01.802560Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:11:01.802560Z digest=sha256:5cd9982d7d255a715d4fbd06e77d7bd27c056690ac283d7827bdf9f04b35441a

Observation 16702f8c-2cb7-4d44-a2c2-8387f4b0ae83 · outbound

This paper cites an unresolved cited work.

Interpretable Risk Mitigation in LLM Agent Systems Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-15T21:11:03.597680Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:11:01.807791Z digest=sha256:b394d1ddae1402c491577ee1f5aaa6ec8742d63d534cda666a61d49125ad0bde

Observation 94b9346b-084f-4af5-a678-a12f9bd12757 · outbound

This paper cites Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback.

Interpretable Risk Mitigation in LLM Agent Systems Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-15T21:11:01.813065Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:11:01.813065Z digest=sha256:cfb3e2020694388452c6130215ee5ea7a7bf8bebc8067527262df8508c99f766

Observation 5e6eafe6-a930-4240-a93e-aaa30f5aa5d6 · outbound

This paper cites Emergent tool use from multi-agent autocurricula.

Interpretable Risk Mitigation in LLM Agent Systems Emergent tool use from multi-agent autocurricula

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:11:03.580794Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:11:01.817697Z digest=sha256:681ae3f19be051cc2f4746212c2fd0d625beea03eb5ea31294178b5f4f49f041

Observation 351ad7f1-6ab1-403d-b7ff-ef072b69c117 · outbound

This paper cites Bender, Timnit Gebru, Angelina McMillan-Major, and Shmargaret Shmitchell.

Interpretable Risk Mitigation in LLM Agent Systems Bender, Timnit Gebru, Angelina McMillan-Major, and Shmargaret Shmitchell

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:11:03.562891Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:11:01.822090Z digest=sha256:f16df6b11ccba79916cf3ab4ad847861b0cf88ec26ac6c78c9a8dd4dbf2e734e

Observation 8a106779-c66d-48c7-8c48-b0f78d8abf43 · outbound

This paper cites On the Opportunities and Risks of Foundation Models.

Interpretable Risk Mitigation in LLM Agent Systems On the Opportunities and Risks of Foundation Models

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-15T21:11:01.826932Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:11:01.826932Z digest=sha256:f66ba15a4101c183ed099580e49199bbe9914502a8aae97ce54ff0d0921aa369

Observation 2a7adc57-b623-480d-be0c-b1dd7e69a6ee · outbound

This paper cites an unresolved cited work.

Interpretable Risk Mitigation in LLM Agent Systems Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-15T21:11:03.544642Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:11:01.832003Z digest=sha256:9df3b0bb6f64b0d8da056ce734905f456e0a0efa9d5c5c695945a4f1e542474e

Observation a338b78c-533d-4e4b-9b30-639f80356236 · outbound

This paper cites RT-1: Robotics Transformer for Real-World Control at Scale.

Interpretable Risk Mitigation in LLM Agent Systems RT-1: Robotics Transformer for Real-World Control at Scale

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-15T21:11:01.836214Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:11:01.836214Z digest=sha256:37ee356306d056c24f581471d08ee04fefc707811c081a99ae379af6d87e718d

Observation 16d22cc2-e43f-4ec2-aff9-61a6d878a1a6 · outbound

This paper cites Playing games with gpt: What can we learn about a large language model from canonical strategic games? SSRN Electronic Journal, 2023.

Interpretable Risk Mitigation in LLM Agent Systems Playing games with gpt: What can we learn about a large language model from canonical strategic games? SSRN Electronic Journal, 2023

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:11:03.526152Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:11:01.840725Z digest=sha256:a91742897ea7b9d76c9a3c4139248b954120b9adfec8a425a2e13687723016c8

Observation fa8b734c-b330-4c95-85f8-426bab50f0df · outbound

This paper cites Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D.

Interpretable Risk Mitigation in LLM Agent Systems Brown, Benjamin Mann, Nick Ryder, Melanie Subbiah, Jared D

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:11:03.508895Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:11:01.844793Z digest=sha256:0850297141c211f40a5c303048de21e0aa6ca8f6d357c43a7f191c7b802089ef

Observation b60e728a-df0f-4060-9aeb-3d455d436d74 · outbound

This paper cites What can machine learning do? workforce implications.

Interpretable Risk Mitigation in LLM Agent Systems What can machine learning do? workforce implications

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:11:03.491497Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:11:01.848928Z digest=sha256:2a4bb93ef97d411d8c53048d29f52bcd2233b169453d765cf7078d41394b0d4b

Observation e55a86e7-747f-4031-8209-67c0bab512ab · outbound

This paper cites Evaluating Large Language Models Trained on Code.

Interpretable Risk Mitigation in LLM Agent Systems Evaluating Large Language Models Trained on Code

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-15T21:11:01.853101Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:11:01.853101Z digest=sha256:97ed16a42de506f8718cd4f08902971027df67c63e7a6780ed23beab556a3633

Observation 08b4f22d-3581-40e0-aa27-dd14aaa9041a · outbound

This paper cites Instigating cooperation among llm agents using adaptive information modulation.

Interpretable Risk Mitigation in LLM Agent Systems Instigating cooperation among llm agents using adaptive information modulation

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-15T21:11:01.857296Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:11:01.857296Z digest=sha256:4264b9f34f01334f2678e1fc3514d949b85cd5a1279239a2593ca68902214fd3

Observation 72433d46-a595-481b-a5ef-2f80a6405e69 · outbound

This paper cites Sparse Autoencoders Find Highly Interpretable Features in Language Models.

Interpretable Risk Mitigation in LLM Agent Systems Sparse Autoencoders Find Highly Interpretable Features in Language Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-15T21:11:01.870732Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:11:01.870732Z digest=sha256:e7c442b67768c15dfba60d5b0138870ffbd84b34ce604d5a4e567c45bccc3048

Observation eadd38c1-fbcd-4b04-a105-2ea6f1e6a5e6 · outbound

This paper cites Reinforcement learning in a prisoner’s dilemma.

Interpretable Risk Mitigation in LLM Agent Systems Reinforcement learning in a prisoner’s dilemma

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:11:03.458432Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:11:01.875550Z digest=sha256:e535ecd7bd48b6aee31bda7f12a1d6abff10c35a4b8f9e94fa2bc86266ec38b2

Observation 24f6a202-78a8-481d-8db8-1c4b79df229d · outbound

This paper cites Toy Models of Superposition.

Interpretable Risk Mitigation in LLM Agent Systems Toy Models of Superposition

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-15T21:11:01.880133Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:11:01.880133Z digest=sha256:0992c45d236b9be6b8100dfca13e7f9e6bf7606178c95e13fec6818d1f48fe87

Observation 3da5c266-6d9b-45d6-8d90-16eafac12480 · outbound

This paper cites PoGaIN: Poisson-Gaussian Image Noise Modeling from Paired Samples.

Interpretable Risk Mitigation in LLM Agent Systems PoGaIN: Poisson-Gaussian Image Noise Modeling from Paired Samples

Reference 22

Resolution
metadata mismatch
local_arxiv, observed 2026-08-15T21:11:02.834041Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:11:01.886176Z digest=sha256:d513a1956e2b86de4a014cf35c788237d4a49c58f95160d4c1a1b332755c8ebd

Observation 9a93bcfe-f4cc-4a52-9e56-a716c5208000 · outbound

This paper cites Not All Language Model Features Are One-Dimensionally Linear.

Interpretable Risk Mitigation in LLM Agent Systems Not All Language Model Features Are One-Dimensionally Linear

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-15T21:11:01.891128Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:11:01.891128Z digest=sha256:a4a89d07ed85bac39988892460f20381dbca7d5d9de495807f84b63a47275e8a

Observation 80ab46e3-96c8-4341-b15b-4ef484f62fc1 · outbound

This paper cites Some experimental games.

Interpretable Risk Mitigation in LLM Agent Systems Some experimental games

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:11:03.442507Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:11:01.896121Z digest=sha256:49a0005a56c8bc9b5eacee8682f06590a5622667ad5f883445c94e2f50e60969

Observation f9da16f4-a983-41b4-b6d2-cccdb6ba0154 · outbound

This paper cites Nicer Than Humans: How do Large Language Models Behave in the Prisoner's Dilemma?.

Interpretable Risk Mitigation in LLM Agent Systems Nicer Than Humans: How do Large Language Models Behave in the Prisoner's Dilemma?

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-15T21:11:01.900891Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:11:01.900891Z digest=sha256:6830411567da8c1bf1f9781f3a722c9651e98ff91090dcffd4850d363b390e4b

Observation 88637004-c8d5-4677-890d-d755e3357d38 · outbound

This paper cites Artificial intelligence, values, and alignment.

Interpretable Risk Mitigation in LLM Agent Systems Artificial intelligence, values, and alignment

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-15T21:11:01.906112Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:11:01.906112Z digest=sha256:96b276e3f4c71d4a78f4c570cd540c156fb9c5cbfe859d019e78c71da890881a

Observation cb688a1f-ae7d-4316-9fcd-3ddffe2c900b · outbound

This paper cites The Llama 3 Herd of Models.

Interpretable Risk Mitigation in LLM Agent Systems The Llama 3 Herd of Models

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-15T21:11:01.910539Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:11:01.910539Z digest=sha256:c7f71b82858392a2a34dd0bb5095bac27db7c730f2fc7ba922ce24037474570f

Observation 2aea34d7-c236-4ba2-b05c-5266d05cbe3a · outbound

This paper cites Measuring Massive Multitask Language Understanding.

Interpretable Risk Mitigation in LLM Agent Systems Measuring Massive Multitask Language Understanding

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-15T21:11:01.915548Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:11:01.915548Z digest=sha256:d2694fc08a587683ada08213828ad884ae52aed2f21b2a8e09bc4c36d2978238

Observation daa69867-514e-4500-ac68-9b455cec0cf3 · outbound

This paper cites Measuring mathematical problem solving with the math dataset,.

Interpretable Risk Mitigation in LLM Agent Systems Measuring mathematical problem solving with the math dataset,

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-15T21:11:01.920013Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:11:01.920013Z digest=sha256:9de668d2f9727cb07dd65a9ba0119ee301e1b5fb8d295eef3b6cbd7bee0ccbd9

Observation ccc257f3-639b-4bd7-b693-d99e2d88ec53 · outbound

This paper cites Cogagent: A visual language model for gui agents.

Interpretable Risk Mitigation in LLM Agent Systems Cogagent: A visual language model for gui agents

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:11:03.416984Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:11:01.930482Z digest=sha256:ca928a69b51caf319ca122ec93ca755a3b8eedbfb364c840e3ab3348a8622dd3

Observation f465e08f-8581-42e6-b51b-25e319bc463c · outbound

This paper cites Non-linear inference time intervention: Improving llm truthfulness.

Interpretable Risk Mitigation in LLM Agent Systems Non-linear inference time intervention: Improving llm truthfulness

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:11:03.400826Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:11:01.935781Z digest=sha256:844c2ba95bb5f8284019d11d65b584ff648afcab7dfe3efdd991937db16f7939

Observation 79c959d9-a3b3-4327-86ed-96acb55e005b · outbound

This paper cites Towards Reasoning in Large Language Models: A Survey.

Interpretable Risk Mitigation in LLM Agent Systems Towards Reasoning in Large Language Models: A Survey

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-15T21:11:01.939936Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:11:01.939936Z digest=sha256:da374a1ea748d380dd9343439240bd041a27de20ed7e00495e38cba27908d737

Observation f8bf0ff6-e450-4f1c-b07c-950516506ed9 · outbound

This paper cites Large language models for uavs: Current state and pathways to the future.

Interpretable Risk Mitigation in LLM Agent Systems Large language models for uavs: Current state and pathways to the future

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:11:03.384589Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:11:01.944592Z digest=sha256:581457afea64933ece197f20a12b724c7232bf75ecdd605a113a910a8be2c642

Observation a7a39f6f-26fb-4c28-b196-bc2a237200ff · outbound

This paper cites Mixtral of Experts.

Interpretable Risk Mitigation in LLM Agent Systems Mixtral of Experts

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-15T21:11:01.949339Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:11:01.949339Z digest=sha256:870cfd79364ac20bacd101a3d80de9627d50789c0df5f68bcedfa90b96ffc25c

Observation 017ef2a9-5520-4eb9-bdc4-f92641554c31 · outbound

This paper cites llama-3-8b-it-res (revision 53425c3), 2024.

Interpretable Risk Mitigation in LLM Agent Systems llama-3-8b-it-res (revision 53425c3), 2024

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:11:03.367736Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:11:01.953994Z digest=sha256:ba3b24c875ef625c4ca02e1deb20b8015db22cbad21020c37b90adc9e9e65146

Observation 91339d58-f620-44d4-8eb7-b59703b0d07a · outbound

This paper cites an unresolved cited work.

Interpretable Risk Mitigation in LLM Agent Systems Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-08-15T21:11:03.350181Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:11:01.958191Z digest=sha256:92457c9f31cd6b3c7341badaf9168ec4d3104e9f5b4ff028e874c85b60053153

Observation 1a302ee4-fd9f-46ae-8fab-c55bc4736872 · outbound

This paper cites Martin, Hans-Theo Normann, and T.

Interpretable Risk Mitigation in LLM Agent Systems Martin, Hans-Theo Normann, and T

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:11:03.333745Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:11:01.962360Z digest=sha256:e2734550fbb0b70f81c856c98f0a16a0b962f766a96875573f132bceae4a45f6

Observation 71a8d99b-1365-4a5b-ac7b-cce9c7fceaf8 · outbound

This paper cites Rusu, Kieran Milan, John Quan, Tiago Ramalho, Agnieszka Grabska-Barwinska, et al.

Interpretable Risk Mitigation in LLM Agent Systems Rusu, Kieran Milan, John Quan, Tiago Ramalho, Agnieszka Grabska-Barwinska, et al

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:11:03.318008Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:11:01.966629Z digest=sha256:64ef772e3ca30ad5cd773d3ae6ba5bd1134c4074f730720604740fc5f42c8e1c

Observation 53ae654e-f604-43fa-bec1-7e4bbe48d9bf · outbound

This paper cites Inference-Time Intervention: Eliciting Truthful Answers from a Language Model.

Interpretable Risk Mitigation in LLM Agent Systems Inference-Time Intervention: Eliciting Truthful Answers from a Language Model

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-15T21:11:01.971070Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:11:01.971070Z digest=sha256:8ec3b580dee0f7e52f6b4882e830bd197fc5815f4337a6a825e85f2a08204172

Observation accf6756-7ccc-46fa-8ac0-5a95cac9c7e1 · outbound

This paper cites Gemma Scope: Open Sparse Autoencoders Everywhere All At Once on Gemma 2.

Interpretable Risk Mitigation in LLM Agent Systems Gemma Scope: Open Sparse Autoencoders Everywhere All At Once on Gemma 2

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-15T21:11:01.975880Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:11:01.975880Z digest=sha256:afd87288f89c53bded5a31bdcffd12b196537682c17d17a00b1cf71d9a2e6ff8

Observation e1e40780-c265-4c8d-b588-0f4db9b277d5 · outbound

This paper cites The mythos of model interpretability.

Interpretable Risk Mitigation in LLM Agent Systems The mythos of model interpretability

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:11:03.302405Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:11:01.980577Z digest=sha256:d44a11abead2638d598781ca122c021fff0e2af5cd07e9473b88d25665e5d5c6

Observation 4fc9c058-9b7f-4030-b591-6b05a8f26e5f · outbound

This paper cites Pre-train, prompt, and predict: A systematic survey of prompting methods in natural language processing.

Interpretable Risk Mitigation in LLM Agent Systems Pre-train, prompt, and predict: A systematic survey of prompting methods in natural language processing

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-15T21:11:01.989753Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:11:01.989753Z digest=sha256:985f12b7f804488919109978059e0ace5334974e155652474eff4d728b5c8fc5

Observation 3cdffd23-9790-4415-81f6-49a0963613c8 · outbound

This paper cites Large Model Strategic Thinking, Small Model Efficiency: Transferring Theory of Mind in Large Language Models.

Interpretable Risk Mitigation in LLM Agent Systems Large Model Strategic Thinking, Small Model Efficiency: Transferring Theory of Mind in Large Language Models

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-15T21:11:01.994175Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:11:01.994175Z digest=sha256:030cdc4113011a8ce88fafe08fb9218908fcd6c5b427a8aaf7dc5b402dc4c134

Observation c86296d2-ed87-49d3-b9f2-78d6f82f7afa · outbound

This paper cites Linguistic regularities in continuous space word representations.

Interpretable Risk Mitigation in LLM Agent Systems Linguistic regularities in continuous space word representations

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:11:03.259770Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:11:02.003164Z digest=sha256:14c80c7581d4a59132af3b2eebacb274012a55c1410a42af5bcfba56cff8e605

Observation a59a7951-c2e9-43c9-a557-199380bc4da6 · outbound

This paper cites Large Language Models: A Survey.

Interpretable Risk Mitigation in LLM Agent Systems Large Language Models: A Survey

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-15T21:11:02.008126Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:11:02.008126Z digest=sha256:55c09564f65472efadb3616a54763b1b3db9c37a2f326b50084d7175c664a6d4

Observation fae7cda6-218c-4394-8746-aa9f8ca079ab · outbound

This paper cites A strategy of win-stay, lose-shift that outperforms tit-for-tat in the prisoner’s dilemma game.

Interpretable Risk Mitigation in LLM Agent Systems A strategy of win-stay, lose-shift that outperforms tit-for-tat in the prisoner’s dilemma game

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-15T21:11:02.013053Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:11:02.013053Z digest=sha256:297bcf3d8af88404e6253ccda686701fe1cc383cab173314d7ae5f2310a35c5f

Observation 4826b429-d15b-417a-b168-ff52c95b5cbc · outbound

This paper cites Training language models to follow instructions with human feedback.

Interpretable Risk Mitigation in LLM Agent Systems Training language models to follow instructions with human feedback

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-15T21:11:02.018166Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:11:02.018166Z digest=sha256:29f0349e400225acc354b33d713e013bbc145fcfe5d4df58884d2286ca59639b

Observation 014ab367-984a-479d-83c0-103ef33837a1 · outbound

This paper cites Cooperation: A systematic review of how to enable agent to circumvent the prisoner’s dilemma.

Interpretable Risk Mitigation in LLM Agent Systems Cooperation: A systematic review of how to enable agent to circumvent the prisoner’s dilemma

Reference 48

Resolution
verified exact
raw_fallback, observed 2026-08-15T21:11:02.613607Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:11:02.023478Z digest=sha256:d385d9ad1c6e54fb67b436b766c33b215bb54555206632484457ba10c5292d27

Observation 7a05f4db-7afc-4148-a98b-9ef26cb1000d · outbound

This paper cites Generative Agents: Interactive Simulacra of Human Behavior.

Interpretable Risk Mitigation in LLM Agent Systems Generative Agents: Interactive Simulacra of Human Behavior

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-15T21:11:02.028670Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:11:02.028670Z digest=sha256:764da072f0c834989483916a21b759a5cbcbee718e1e8544ee3d02dc32cd55d5

Observation de77e093-1471-462a-b0a6-a927f6729dc5 · outbound

This paper cites TinyClick: Single-Turn Agent for Empowering GUI Automation.

Interpretable Risk Mitigation in LLM Agent Systems TinyClick: Single-Turn Agent for Empowering GUI Automation

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-15T21:11:02.033414Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:11:02.033414Z digest=sha256:962e30645ac09c1a5f31691e29fb4c8b0fea0e76a9ab99e6a671c7a77f0669f4

Observation 73a4ef3f-38c0-44df-85f6-445319704ac3 · outbound

This paper cites an unresolved cited work.

Interpretable Risk Mitigation in LLM Agent Systems Unresolved cited work

Reference 51

Resolution
unresolved
raw_fallback, observed 2026-08-15T21:11:03.231251Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:11:02.038063Z digest=sha256:c97d6f66f4d5f927feb49290a840b83e350f38be2872f97ae5f41b59668728d5

Observation 35196dd2-37eb-4576-a6c2-b40e2da03ea4 · outbound

This paper cites Effect of private deliberation: Deception of large language models in game play.

Interpretable Risk Mitigation in LLM Agent Systems Effect of private deliberation: Deception of large language models in game play

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:11:03.216654Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:11:02.042680Z digest=sha256:fdc535878cba31685465388b0ccf7b803192e2dcf593801266a8bdd209cfed83

Observation a70cb5aa-ff92-409f-b95a-3405b51bbf27 · outbound

This paper cites GPQA: A Graduate-Level Google-Proof Q&A Benchmark.

Interpretable Risk Mitigation in LLM Agent Systems GPQA: A Graduate-Level Google-Proof Q&A Benchmark

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-15T21:11:02.047108Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:11:02.047108Z digest=sha256:ce0dcba663afd8c78272e4d613c1ee49956300d5a838867a76aebbca4aff98a3

Observation d5a5b1b9-7b0f-427c-973e-a741a2cd28cd · outbound

This paper cites A primer in BERTology: What we know about how BERT works.

Interpretable Risk Mitigation in LLM Agent Systems A primer in BERTology: What we know about how BERT works

Reference 54

Resolution
malformed identifier
no resolver link, observed 2026-08-15T21:11:02.051703Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:11:02.051703Z digest=sha256:5ac12d1a477e3668787213a4bc49aad5bd0d9267c7da646c01809819225c8efa

Observation cb2fe07f-6a82-4fb7-bc5a-7b4cefa20efd · outbound

This paper cites Research priorities for robust and beneficial artificial intelligence.

Interpretable Risk Mitigation in LLM Agent Systems Research priorities for robust and beneficial artificial intelligence

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:11:03.201675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:11:02.055985Z digest=sha256:bc0950e2b22be52f339bce528904fe053144ca4d6b4b42198a945f05dd705b61

Observation 7a2cf6c4-3cb3-42c1-a6c5-e33c3fd83511 · outbound

This paper cites an unresolved cited work.

Interpretable Risk Mitigation in LLM Agent Systems Unresolved cited work

Reference 56

Resolution
unresolved
raw_fallback, observed 2026-08-15T21:11:03.185336Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:11:02.060734Z digest=sha256:06ae9f13f86fe9b3b025d30a012e52163d5d45bb053ccfe19db259438902ef0a

Observation 95101970-bd83-42b3-bfc5-703b311591ce · outbound

This paper cites Toolformer: Language Models Can Teach Themselves to Use Tools.

Interpretable Risk Mitigation in LLM Agent Systems Toolformer: Language Models Can Teach Themselves to Use Tools

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-15T21:11:02.064856Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:11:02.064856Z digest=sha256:43035a3f85fa6fd8654e66fc384379840dafb90e5693cfb3b0269ac1d1d87edd

Observation 8f6167c9-4a5f-46af-b08b-065532a5c700 · outbound

This paper cites An evolutionary model of personality traits related to cooperative behavior using a large language model.

Interpretable Risk Mitigation in LLM Agent Systems An evolutionary model of personality traits related to cooperative behavior using a large language model

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:11:03.170463Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:11:02.069111Z digest=sha256:98411265d2b0bd574b0f2e186264c2fe45f62821d4f1dd2649d16fddac1baa63

Observation 0b622b1d-44e4-43db-bc95-870dd9ca5525 · outbound

This paper cites A comparative analysis of the definitions of autonomous weapons systems.

Interpretable Risk Mitigation in LLM Agent Systems A comparative analysis of the definitions of autonomous weapons systems

Reference 59

Resolution
verified exact
doi, observed 2026-08-15T21:11:02.190554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:11:02.073714Z digest=sha256:e5412951b20370d0463f89fbd49778862e15752cbb9f85660d49da1fa07cea64

Observation 54cdc25e-dcbe-4a06-a3ca-a0b083ec66e7 · outbound

This paper cites Gemma: Open Models Based on Gemini Research and Technology.

Interpretable Risk Mitigation in LLM Agent Systems Gemma: Open Models Based on Gemini Research and Technology

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-15T21:11:02.078502Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:11:02.078502Z digest=sha256:ef023ec26272247cba1e30888ad222f5263ac07bbe407057a7f92bd3651a66ca

Observation e70192c0-2fa3-4d26-9e99-25bce4521ab8 · outbound

This paper cites Gemma 2: Improving Open Language Models at a Practical Size.

Interpretable Risk Mitigation in LLM Agent Systems Gemma 2: Improving Open Language Models at a Practical Size

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-15T21:11:02.082737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:11:02.082737Z digest=sha256:bccf4597c16c569626c3d43a0564d5d741769cf5f7484404b07a6292e59b3add

Observation 625f59c6-0eef-406d-9cf0-d8c92c8ef449 · outbound

This paper cites Scaling monosemanticity: Extracting interpretable features from claude 3 sonnet.

Interpretable Risk Mitigation in LLM Agent Systems Scaling monosemanticity: Extracting interpretable features from claude 3 sonnet

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-15T21:11:02.087316Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:11:02.087316Z digest=sha256:81b313134bdab422d32cfc76fc0bacdb0c5faa703cf65bd43b18b8b4bd38a34b

Observation 7b2ebf77-d4d1-47eb-846b-c88fd4d56a1d · outbound

This paper cites Moral alignment for llm agents.

Interpretable Risk Mitigation in LLM Agent Systems Moral alignment for llm agents

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:11:03.143335Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:11:02.092060Z digest=sha256:906818dbbc70e3879e2a181548bfe8f74882d3938cb73541599c2cbf3a054fad

Observation f523afad-184f-41ef-9a0e-57e19a11c506 · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

Interpretable Risk Mitigation in LLM Agent Systems LLaMA: Open and Efficient Foundation Language Models

Reference 64

Resolution
unresolved
no resolver link, observed 2026-08-15T21:11:02.098429Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:11:02.098429Z digest=sha256:9f623e8e9a072adf447a5a14f7eca499dd093dccf882bcaeb1727a7d52620784

Observation 82ac90b8-2d66-452e-9693-66c10e7d314b · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

Interpretable Risk Mitigation in LLM Agent Systems Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-15T21:11:02.104902Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:11:02.104902Z digest=sha256:4c33d00608e1ce971047d87195ee193cf4016570ffae541f6b20d15bf4624a9c

Observation 20ce1bc9-e4d8-4904-9888-9de2d34c5479 · outbound

This paper cites GLUE: A Multi-Task Benchmark and Analysis Platform for Natural Language Understanding.

Interpretable Risk Mitigation in LLM Agent Systems GLUE: A Multi-Task Benchmark and Analysis Platform for Natural Language Understanding

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-15T21:11:02.109561Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:11:02.109561Z digest=sha256:e029072b6d52231a090526021bc31127e9118ddd6905f84356cab8974aa8e6d2

Observation f2942ef8-188f-4e1f-a44b-0747ba8e3471 · outbound

This paper cites Mobile-agent: Autonomous multi-modal mobile device agent with visual perception,.

Interpretable Risk Mitigation in LLM Agent Systems Mobile-agent: Autonomous multi-modal mobile device agent with visual perception,

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:11:03.127797Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:11:02.114293Z digest=sha256:06ded670fc70774e03b71908a19465a65a730496b4487509c9cf58a4dc9da1ca

Observation e8c5f846-e91b-401e-80a4-0fd5d73e1194 · outbound

This paper cites Dai, and Quoc V Le.

Interpretable Risk Mitigation in LLM Agent Systems Dai, and Quoc V Le

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:11:03.112312Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:11:02.123585Z digest=sha256:98b65b6755c1cd5e99a196894b4c7fb0be02f7390972e270c05072639c72cf15

Observation 9aac0e58-4508-4dce-b103-c860ac6c28f4 · outbound

This paper cites Chain-of-Thought Prompting Elicits Reasoning in Large Language Models.

Interpretable Risk Mitigation in LLM Agent Systems Chain-of-Thought Prompting Elicits Reasoning in Large Language Models

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-15T21:11:02.128289Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:11:02.128289Z digest=sha256:7490b5daedf04105ae22a7ff91c365e9e41771b371ce4fa4c6f8f95f34c55d7f

Observation 1cea8ac7-27e1-49cf-b37d-d6744ff11b68 · outbound

This paper cites ReAct: Synergizing Reasoning and Acting in Language Models.

Interpretable Risk Mitigation in LLM Agent Systems ReAct: Synergizing Reasoning and Acting in Language Models

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-15T21:11:02.133411Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:11:02.133411Z digest=sha256:8dbea1b06f143813b8530b029a8cb487b07d652c28985c523ae5820d38529b35

Observation 647ac5ed-565f-4692-8bee-b21bc1834a89 · outbound

This paper cites AppAgent: Multimodal Agents as Smartphone Users.

Interpretable Risk Mitigation in LLM Agent Systems AppAgent: Multimodal Agents as Smartphone Users

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-15T21:11:02.140648Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:11:02.140648Z digest=sha256:66741ae45a9a65e21aa48371f7f22f9abb8c9a6788273da5048a51030e90c16d

Observation 33d30244-b423-4420-bcc9-fbd665b74b6d · outbound

This paper cites Mobile-Agent: Autonomous Multi-Modal Mobile Device Agent with Visual Perception.

Interpretable Risk Mitigation in LLM Agent Systems Mobile-Agent: Autonomous Multi-Modal Mobile Device Agent with Visual Perception

Reference 72

Resolution
unresolved
no resolver link, observed 2026-08-15T21:11:02.118923Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:11:02.118923Z digest=sha256:307dcd5fdf32fe62d9753454a2af0106f64a8cd7337be5dc3aef7d8b40a22429

Observation 10bb76ee-bc2d-40f4-babf-40043502314b · outbound

This paper cites Representation Engineering: A Top-Down Approach to AI Transparency.

Interpretable Risk Mitigation in LLM Agent Systems Representation Engineering: A Top-Down Approach to AI Transparency

Reference 73

Resolution
unresolved
no resolver link, observed 2026-08-15T21:11:02.149478Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:11:02.149478Z digest=sha256:b8bcc30508eac4338d3e2c2262d1a9c3ecdb760930a9ffd0df35be9ed84fc175

Observation e1633c79-0375-4295-bfa3-8325c644a49a · outbound

This paper cites You Only Look at Screens: Multimodal Chain-of-Action Agents.

Interpretable Risk Mitigation in LLM Agent Systems You Only Look at Screens: Multimodal Chain-of-Action Agents

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-15T21:11:02.145319Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:11:02.145319Z digest=sha256:729fe8725ee380b78bd5372858cdb5fd008b8ca2b39ae0883b93f3b629fdf8fa

Observation d843c580-4512-4d8f-a3a6-6b41724a5ec6 · outbound

This paper cites an unresolved cited work.

Interpretable Risk Mitigation in LLM Agent Systems Unresolved cited work

Reference 2016

Resolution
unresolved
no resolver link, observed 2026-08-15T21:11:01.984893Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:11:01.984893Z digest=sha256:977eba11d180dca9efb9b3b0e9c89a04c2ec480ef3caff4d518eb103eea92809

Observation 671ed4ec-d1ad-4d99-97f8-e0d15322b52b · outbound

This paper cites Measuring Mathematical Problem Solving With the MATH Dataset.

Interpretable Risk Mitigation in LLM Agent Systems Measuring Mathematical Problem Solving With the MATH Dataset

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-15T21:11:01.925637Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:11:01.925637Z digest=sha256:c555218d0202e2b5159da1a3bd738146b202bdbcfbd361368474ba9320e7434f

Observation 6b1e14ac-f124-489b-84f4-00a88d7ea70d · outbound

This paper cites an unresolved cited work.

Interpretable Risk Mitigation in LLM Agent Systems Unresolved cited work

Reference 2023

Resolution
unresolved
raw_fallback, observed 2026-08-15T21:11:03.474249Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:11:01.866455Z digest=sha256:3e70126129b6dcab211899849f5d49426ce28d479d894d4b9db18496c1c13dad

Observation 98fbd528-7341-4dfe-92d5-440307b35277 · outbound

This paper cites an unresolved cited work.

Interpretable Risk Mitigation in LLM Agent Systems Unresolved cited work

Reference 2024

Resolution
unresolved
raw_fallback, observed 2026-08-15T21:11:03.276283Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T21:11:01.998638Z digest=sha256:ca80af219653425a6bab904a173c93769f0c79c273b411f490a2f3f17fd8967e

Pith citing papers

Observation b72a9458-11dc-47d1-84f0-f77d9122b3d6 · inbound

A Call for Collaborative Intelligence: Why Human-Agent Systems Should Precede AI Autonomy cites this paper.

A Call for Collaborative Intelligence: Why Human-Agent Systems Should Precede AI Autonomy Interpretable Risk Mitigation in LLM Agent Systems

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T04:51:56.450815Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:51:56.450815Z digest=sha256:feeac579c59424f523aec23174c09e48fbb6f8f53d206fcaf90c78c388b85d38

Observation 1b7081ac-e99d-43fc-be2e-190be39eed65 · inbound

LLM Harms: A Taxonomy and Discussion cites this paper.

LLM Harms: A Taxonomy and Discussion Interpretable Risk Mitigation in LLM Agent Systems

Reference 189

Resolution
verified exact
arxiv_id, observed 2026-05-17T00:31:24.636618Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-17T00:29:07.951709Z digest=sha256:13f33b4cc959f8832a5b13787cd58932f80352ef1252bb541e1a483df69a18ec

Observation 84d058c8-51a1-4859-970e-8b0c97117aac · inbound

LLM Harms: A Taxonomy and Discussion cites this paper.

LLM Harms: A Taxonomy and Discussion Interpretable Risk Mitigation in LLM Agent Systems

Reference 189

Resolution
unresolved
no resolver link, observed 2026-08-03T18:19:27.337242Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T18:19:27.337242Z digest=sha256:8accc8b8ee2bb53ffb54ba41fb6b8e0c96d7edee279df5fdd9a7697468d0a490