Pith. sign in

REVIEW 1 cited by

LLMGuard: Guarding Against Unsafe LLM Behavior

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2403.00826 v1 pith:AD5L2DET submitted 2024-02-27 cs.CL cs.CRcs.LG

classification cs.CLcs.CRcs.LG
keywords llmguardbringscontentalleviatealthoughapplicationbehaviorbehaviours
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Although the rise of Large Language Models (LLMs) in enterprise settings brings new opportunities and capabilities, it also brings challenges, such as the risk of generating inappropriate, biased, or misleading content that violates regulations and can have legal concerns. To alleviate this, we present "LLMGuard", a tool that monitors user interactions with an LLM application and flags content against specific behaviours or conversation topics. To do this robustly, LLMGuard employs an ensemble of detectors.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Latent Interpolation Learning Using Diffusion Models for Cardiac Volume Reconstruction

    eess.IV 2025-08 unverdicted novelty 5.0 of 10

    CaLID claims state-of-the-art 3D cardiac volume reconstruction from sparse 2D MRI slices via latent-space diffusion interpolation with a 24x speedup and no auxiliary inputs, but verification is impossible because the ...

Pith tools