Pith. sign in

Paper Citation Record · LEDGER

RAS-Eval: A Comprehensive Benchmark for Security Evaluation of LLM Agents in Real-World Environments

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 9 inbound Pith citation observations for arXiv:2506.15253.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.15253 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 9 of 9 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 9 of 9 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-30T16:22:23.857438Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 1aab66c6-34ae-4d9e-bdaf-350397be20b1 · inbound

Evaluating Privilege Usage of Agents with Real-World Tools cites this paper.

Evaluating Privilege Usage of Agents with Real-World Tools RAS-Eval: A Comprehensive Benchmark for Security Evaluation of LLM Agents in Real-World Environments

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-14T22:18:04.361136Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-14T22:16:56.848520Z digest=sha256:1ed691d9e7e656f021e63ee7a79994e5e736ea021866f8a27a09102cdb62867d

Observation a5c5f379-665a-44a6-85eb-35b38f957443 · inbound

Content-Aware Attack Detection in LLM Agent Tool-Call Traffic: An Empirical Study of Features, Architectures, and Evaluation Protocols cites this paper.

Content-Aware Attack Detection in LLM Agent Tool-Call Traffic: An Empirical Study of Features, Architectures, and Evaluation Protocols RAS-Eval: A Comprehensive Benchmark for Security Evaluation of LLM Agents in Real-World Environments

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-13T00:57:00.300069Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-13T00:56:02.786795Z digest=sha256:f5bd32c2fcbb4cada59fa28c24d9f372760a3a85efc361ddfc1a98a064e3e10a

Observation 3e82437a-40d0-4159-895d-391f36553946 · inbound

Content-Aware Attack Detection in LLM Agent Tool-Call Traffic: An Empirical Study of Features, Architectures, and Evaluation Protocols cites this paper.

Content-Aware Attack Detection in LLM Agent Tool-Call Traffic: An Empirical Study of Features, Architectures, and Evaluation Protocols RAS-Eval: A Comprehensive Benchmark for Security Evaluation of LLM Agents in Real-World Environments

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-14T21:17:58.611548Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-14T21:17:10.373607Z digest=sha256:84960e8df423ef2de67daef308292564c958ff581e42a66b05699c99dc9c19a0

Observation fb9ca04a-f186-4b68-a354-bc9994f08cd4 · inbound

Content-Aware Attack Detection in LLM Agent Tool-Call Traffic: An Empirical Study of Features, Architectures, and Evaluation Protocols cites this paper.

Content-Aware Attack Detection in LLM Agent Tool-Call Traffic: An Empirical Study of Features, Architectures, and Evaluation Protocols RAS-Eval: A Comprehensive Benchmark for Security Evaluation of LLM Agents in Real-World Environments

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-25T06:00:23.557984Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-25T05:57:22.726910Z digest=sha256:6b21fff08640447f65a7f1d69f6f2d4e1cb4a0d4717242a709572705c2920f48

Observation 954a84a3-b2a7-46d4-afa0-291657d1ddfb · inbound

Do Coding Agents Understand Least-Privilege Authorization? cites this paper.

Do Coding Agents Understand Least-Privilege Authorization? RAS-Eval: A Comprehensive Benchmark for Security Evaluation of LLM Agents in Real-World Environments

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-05-19T16:37:39.984544Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-19T16:34:14.379419Z digest=sha256:43eaa534f15e3d0b5097e5b8385df13392d0b911ef63c772774c321227b53584

Observation 85b0d4e7-a8a2-42ee-86d0-03b4d12aa8f5 · inbound

Taxonomy and Consistency Analysis of Safety Benchmarks for AI Agents cites this paper.

Taxonomy and Consistency Analysis of Safety Benchmarks for AI Agents RAS-Eval: A Comprehensive Benchmark for Security Evaluation of LLM Agents in Real-World Environments

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-21T01:43:56.870990Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-21T01:42:55.693115Z digest=sha256:477dd4b46455aee458cc09c3d7b3c9c65f84d6fe2065a74a5bfebb330690fe2b

Observation 64a5976c-460d-4834-a76c-5b390b9f4032 · inbound

ADR: An Agentic Detection System for Enterprise Agentic AI Security cites this paper.

ADR: An Agentic Detection System for Enterprise Agentic AI Security RAS-Eval: A Comprehensive Benchmark for Security Evaluation of LLM Agents in Real-World Environments

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-20T13:18:18.173988Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-20T13:17:59.293695Z digest=sha256:50fa5b064d177289356204927c749b504f161dfaf0798ff623764b14af2d6ea4

Observation bede31b4-ccc2-4cdc-95c2-e57b79c744f8 · inbound

When the Manual Lies: A Realistic Benchmark to Evaluate MCP Poisoning Attacks for LLM Agents cites this paper.

When the Manual Lies: A Realistic Benchmark to Evaluate MCP Poisoning Attacks for LLM Agents RAS-Eval: A Comprehensive Benchmark for Security Evaluation of LLM Agents in Real-World Environments

Reference 15

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T16:24:55.111630Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-30T16:22:23.857438Z digest=sha256:b24afb077daf7e2270f14b8cdd948e9dd49193c9b297ca5438dd481fd44f1ad3

Observation ecc9f3f8-dc64-4210-8d39-eddc804b5446 · inbound

Toward Secure LLM Agents: Threat Surfaces, Attacks, Defenses, and Evaluation cites this paper.

Toward Secure LLM Agents: Threat Surfaces, Attacks, Defenses, and Evaluation RAS-Eval: A Comprehensive Benchmark for Security Evaluation of LLM Agents in Real-World Environments

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-06-27T13:20:56.905038Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T12:55:22.831264Z digest=sha256:23e235a402ca7a25fd5fc036d94b1272a9c0ccb6918c6b2b37adeccce7299f60