Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T04:50:00.253920Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 38 of 38 outbound references and 2 inbound Pith citation observations for arXiv:2506.09600.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T04:50:00.253920Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-06-30T07:48:36.924294Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-06-30T07:54:22.258152Z
38 of 38 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation a9fe4ffe-27a7-4b2d-bac8-777d4d8dd59c · outbound
Effective Red-Teaming of Policy-Adherent Agents Unresolved cited work
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ed34efe0-0318-4a5e-a5e3-dc1c3f58dbd8 · outbound
Effective Red-Teaming of Policy-Adherent Agents AgentHarm: A Benchmark for Measuring Harmfulness of LLM Agents
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5bd98dfd-fade-4c92-993d-2104cf826649 · outbound
Effective Red-Teaming of Policy-Adherent Agents A Framework for Testing and Adapting REST APIs as LLM Tools
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 03ed9a8b-08c1-47cf-a062-174d7e1d1cab · outbound
Effective Red-Teaming of Policy-Adherent Agents Evaluating Large Language Models Trained on Code
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a98cffe2-f85d-496d-9947-cd94c9849bc6 · outbound
Effective Red-Teaming of Policy-Adherent Agents The Llama 3 Herd of Models
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1e9eb559-74f2-464a-8fe8-116158765a52 · outbound
Effective Red-Teaming of Policy-Adherent Agents Unresolved cited work
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e6252988-2705-441d-be6b-a23da52a6621 · outbound
Effective Red-Teaming of Policy-Adherent Agents Unresolved cited work
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2d02d8c2-ef73-4a6e-9b5b-998a49c59faf · outbound
Effective Red-Teaming of Policy-Adherent Agents GPT-4o System Card
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0fc1f222-66e5-48a6-9d06-cccaa70e343d · outbound
Effective Red-Teaming of Policy-Adherent Agents Unresolved cited work
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 48730d43-8fe6-4f01-a378-2c588ee6b58a · outbound
Effective Red-Teaming of Policy-Adherent Agents Unresolved cited work
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 986a60bd-ecac-46ec-ba2d-f5e102ed3cfc · outbound
Effective Red-Teaming of Policy-Adherent Agents Unresolved cited work
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7ceccfad-f25e-486e-adc7-2455580e1072 · outbound
Effective Red-Teaming of Policy-Adherent Agents Unveiling Safety Vulnerabilities of Large Language Models
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cc16bd91-aabd-4430-a712-cd7800f2b6f4 · outbound
Effective Red-Teaming of Policy-Adherent Agents Unresolved cited work
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6c005901-1f7d-42ed-895b-fca8e3ed9781 · outbound
Effective Red-Teaming of Policy-Adherent Agents ST-WebAgentBench: A Benchmark for Evaluating Safety and Trustworthiness in Web Agents
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d2052270-87ab-4262-b1ae-ae4e9432038d · outbound
Effective Red-Teaming of Policy-Adherent Agents SOPBench: Evaluating Language Agents at Following Standard Operating Procedures and Constraints
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1fc9cdf3-3095-4c91-9b47-b3ec89dd02ca · outbound
Effective Red-Teaming of Policy-Adherent Agents DeepSeek-V3 Technical Report
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 45a07ed6-ae4e-4bc8-8523-8826db4d0170 · outbound
Effective Red-Teaming of Policy-Adherent Agents From LLM to Conversational Agent: A Memory Enhanced Architecture with Fine-Tuning of Large Language Models
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e4aa5648-db3b-4be7-ade6-9fd1c50de94e · outbound
Effective Red-Teaming of Policy-Adherent Agents AgentBench: Evaluating LLMs as Agents
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 50533e89-9e21-42b8-844c-5106ca7e14f1 · outbound
Effective Red-Teaming of Policy-Adherent Agents Prompt Injection attack against LLM-integrated Applications
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a358ba86-58cc-4329-87c5-f8f28fee284d · outbound
Effective Red-Teaming of Policy-Adherent Agents Unresolved cited work
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0fde2ab8-b1f9-4dd3-afcb-8dacea8d9941 · outbound
Effective Red-Teaming of Policy-Adherent Agents u ndler, Mark Niklas M \
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 5b29a17e-92b7-409b-b739-9d0cd5e120c8 · outbound
Effective Red-Teaming of Policy-Adherent Agents Unresolved cited work
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e7683656-a1bd-4537-aaef-7b477760998e · outbound
Effective Red-Teaming of Policy-Adherent Agents do anything now
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 36f055a0-5e8a-40fb-9d6b-71cf64ba9975 · outbound
Effective Red-Teaming of Policy-Adherent Agents do anything now
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f46532e3-00d6-4dd6-99fb-2a643e881f52 · outbound
Effective Red-Teaming of Policy-Adherent Agents CHOPS: CHat with custOmer Profile Systems for Customer Service with LLMs
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ada563b1-aadb-4908-9872-59d56c02826d · outbound
Effective Red-Teaming of Policy-Adherent Agents Can LLMs Generate Novel Research Ideas? A Large-Scale Human Study with 100+ NLP Researchers
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 04ecd511-17b3-4848-8819-74ac326f158f · outbound
Effective Red-Teaming of Policy-Adherent Agents Unresolved cited work
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c7f4545a-55a6-4feb-8926-256a2478f566 · outbound
Effective Red-Teaming of Policy-Adherent Agents Multi-Turn Context Jailbreak Attack on Large Language Models From First Principles
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0066ba6d-3785-41fd-96c2-443202205f55 · outbound
Effective Red-Teaming of Policy-Adherent Agents Unresolved cited work
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0c447225-0981-48ad-aafa-07b43f553d88 · outbound
Effective Red-Teaming of Policy-Adherent Agents Emotional Manipulation Through Prompt Engineering Amplifies Disinformation Generation in AI Large Language Models
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 617a1164-8ca5-4f15-bafa-2f5dfbfabcb7 · outbound
Effective Red-Teaming of Policy-Adherent Agents Tutor CoPilot: A Human-AI Approach for Scaling Real-Time Expertise
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5f6b1128-03a9-4a3a-885e-aefa408de241 · outbound
Effective Red-Teaming of Policy-Adherent Agents Qwen3 Technical Report
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f2aa3b30-54c0-4f1e-8b3f-35ebd1d96784 · outbound
Effective Red-Teaming of Policy-Adherent Agents $\tau$-bench: A Benchmark for Tool-Agent-User Interaction in Real-World Domains
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bdddd043-26b7-436e-9c0a-a2d6f4633c16 · outbound
Effective Red-Teaming of Policy-Adherent Agents Unresolved cited work
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a67f23d5-4b4a-419f-8f27-be2a1c2ced3d · outbound
Effective Red-Teaming of Policy-Adherent Agents Survey on Evaluation of LLM-based Agents
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 816d0e1b-59bd-4e90-9fb5-a609d7754cc9 · outbound
Effective Red-Teaming of Policy-Adherent Agents Unresolved cited work
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 80feea2e-0a2a-4cb2-a8a6-236e3e9bf407 · outbound
Effective Red-Teaming of Policy-Adherent Agents online" 'onlinestring :=
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 82e72879-d510-4f14-9dad-a2a6be99f225 · outbound
Effective Red-Teaming of Policy-Adherent Agents write newline
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 29492dc8-2762-4be4-a9ef-17672f85925e · inbound
Don't Pass@k: A Bayesian Framework for Large Language Model Evaluation Effective Red-Teaming of Policy-Adherent Agents
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e24b196a-a0fd-4f1c-8ca4-5f9ae0e3084f · inbound
PolicyGuard: A Dialogue-Grounded Sub-Agent Verifier for Policy Adherence in LLM Agents Effective Red-Teaming of Policy-Adherent Agents
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.