An evaluation of seven LLM security tools on a new 500-prompt benchmark finds the ChatGPT-3.5-Turbo baseline unusable due to false positives and names Lakera Guard and ProtectAI LLM Guard the best overall tools.
https://www.pinecone.io/
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
citation-role summary
other 1
citation-polarity summary
fields
cs.CR 1years
2025 1verdicts
CONDITIONAL 1roles
other 1polarities
use method 1representative citing papers
citing papers explorer
-
Evaluating the efficacy of LLM Safety Solutions : The Palit Benchmark Dataset
An evaluation of seven LLM security tools on a new 500-prompt benchmark finds the ChatGPT-3.5-Turbo baseline unusable due to false positives and names Lakera Guard and ProtectAI LLM Guard the best overall tools.