Pith. sign in

AI Ethics by Design: Implementing Customizable Guardrails for Responsible AI Development

1 Pith paper cite this work. Polarity classification is still indexing.

1 Pith paper citing it
abstract

This paper explores the development of an ethical guardrail framework for AI systems, emphasizing the importance of customizable guardrails that align with diverse user values and underlying ethics. We address the challenges of AI ethics by proposing a structure that integrates rules, policies, and AI assistants to ensure responsible AI behavior, while comparing the proposed framework to the existing state-of-the-art guardrails. By focusing on practical mechanisms for implementing ethical standards, we aim to enhance transparency, user autonomy, and continuous improvement in AI systems. Our approach accommodates ethical pluralism, offering a flexible and adaptable solution for the evolving landscape of AI governance. The paper concludes with strategies for resolving conflicts between ethical directives, underscoring the present and future need for robust, nuanced and context-aware AI systems.

citation-role summary

background 1

citation-polarity summary

fields

cs.CR 1

years

2025 1

verdicts

REJECT 1

roles

background 1

polarities

background 1

representative citing papers

Prompt Injection 2.0: Hybrid AI Threats

cs.CR · 2025-07-17 · reject · novelty 2.0

A structured taxonomy of hybrid prompt injection attacks shows how XSS, CSRF, and SQL injection vectors converge with LLM manipulation to bypass traditional controls.

citing papers explorer

Showing 1 of 1 citing paper.

  • Prompt Injection 2.0: Hybrid AI Threats cs.CR · 2025-07-17 · reject · none · ref 6 · internal anchor

    A structured taxonomy of hybrid prompt injection attacks shows how XSS, CSRF, and SQL injection vectors converge with LLM manipulation to bypass traditional controls.