A new dataset of production-style prompts with matching assertion criteria, and a benchmark where fine-tuned 7-8B models beat GPT-4o at generating those criteria.
Building guardrails for large language models, 2024
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.CL 1years
2025 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
PROMPTEVALS: A Dataset of Assertions and Guardrails for Custom Production Large Language Model Pipelines
A new dataset of production-style prompts with matching assertion criteria, and a benchmark where fine-tuned 7-8B models beat GPT-4o at generating those criteria.