REVIEW 3 cited by
A Survey on Responsible Generative AI: What to Generate and What Not
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
In recent years, generative AI (GenAI), like large language models and text-to-image models, has received significant attention across various domains. However, ensuring the responsible generation of content by these models is crucial for their real-world applicability. This raises an interesting question: What should responsible GenAI generate, and what should it not? To answer the question, this paper investigates the practical responsible requirements of both textual and visual generative models, outlining five key considerations: generating truthful content, avoiding toxic content, refusing harmful instruction, leaking no training data-related content, and ensuring generated content identifiable. Specifically, we review recent advancements and challenges in addressing these requirements. Besides, we discuss and emphasize the importance of responsible GenAI across healthcare, education, finance, and artificial general intelligence domains. Through a unified perspective on both textual and visual generative models, this paper aims to provide insights into practical safety-related issues and further benefit the community in building responsible GenAI.
Forward citations
Cited by 3 Pith papers
-
Not Just Text: Uncovering Vision Modality Typographic Threats in Image Generation Models
Typography inserted into input images can manipulate CLIP-guided image generation models to produce harmful or biased content, and existing text-focused defenses do not catch it.
-
UVCG: Leveraging Temporal Consistency for Universal Video Protection
UVCG protects videos from AI editing by perturbing frames so their latent representations align with a chosen target video, disrupting editing pipelines while reusing perturbations across frames for efficiency.
-
Global Challenge for Safe and Secure LLMs Track 1
In a two-phase red-team challenge, top automated jailbreak methods reached 0.96 to 0.98 attack success on Llama-2 and Vicuna, with evidence of transfer across models.
Discussion (0). Continue with ORCID to comment.