pith. sign in

LLMail-Inject: A Dataset from a Realistic Adaptive Prompt Injection Challenge

4 Pith papers cite this work. Polarity classification is still indexing.

4 Pith papers citing it

citation-role summary

dataset 1

citation-polarity summary

years

2026 3 2025 1

verdicts

UNVERDICTED 4

roles

dataset 1

polarities

use dataset 1

representative citing papers

Hallucination as Exploit: Evidence-Carrying Multimodal Agents

cs.AI · 2026-05-18 · unverdicted · novelty 6.0 · 2 refs

Evidence-carrying multimodal agents decompose tool calls into predicates, obtain certificates from DOM/OCR/AX verifiers, and use a deterministic gate to authorize actions only when certificates support them, achieving zero unsafe executions in tested tasks.

citing papers explorer

Showing 4 of 4 citing papers.