Pith. sign in

REVIEW 38 cited by

The Ethics of Advanced AI Assistants

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2404.16244 v2 pith:2UVVN5MH submitted 2024-04-24 cs.CY

classification cs.CY
keywords assistantsadvancedconsiderprovidingrangesocietaluseraccess
verification ladder T0 review T1 audit T2 compute T3 formal

Signed reviews

No signed human review yet.

0 comments
read the original abstract

This paper focuses on the opportunities and the ethical and societal risks posed by advanced AI assistants. We define advanced AI assistants as artificial agents with natural language interfaces, whose function is to plan and execute sequences of actions on behalf of a user, across one or more domains, in line with the user's expectations. The paper starts by considering the technology itself, providing an overview of AI assistants, their technical foundations and potential range of applications. It then explores questions around AI value alignment, well-being, safety and malicious uses. Extending the circle of inquiry further, we next consider the relationship between advanced AI assistants and individual users in more detail, exploring topics such as manipulation and persuasion, anthropomorphism, appropriate relationships, trust and privacy. With this analysis in place, we consider the deployment of advanced assistants at a societal scale, focusing on cooperation, equity and access, misinformation, economic impact, the environment and how best to evaluate advanced AI assistants. Finally, we conclude by providing a range of recommendations for researchers, developers, policymakers and public stakeholders.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 38 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. OpenAlex reports about 48 citations worldwide. Full citation record

  1. HumanAgencyBench: Scalable Evaluation of Human Agency Support in AI Assistants

    cs.CY 2025-09 conditional novelty 7.0 of 10

    A new benchmark finds low to moderate human agency support in 20 LLM assistants across six dimensions.

  2. MixAssist: An Audio-Language Dataset for Co-Creative AI Assistance in Music Mixing

    cs.SD 2025-07 conditional novelty 7.0 of 10

    MixAssist is the first audio-grounded, multi-turn conversational dataset for co-creative music mixing instruction, and fine-tuning Qwen-Audio on it yields human-comparable mixing advice.

  3. AI Alignment and Fiduciary Obligation

    cs.CY 2026-08 conditional novelty 6.0 of 10

    The paper derives AI alignment criteria from fiduciary duties developers owe to users of extended AI assistants.

  4. A Roadmap to Impactful Pluralistic Alignment Research

    cs.AI 2026-07 accept novelty 6.0 of 10

    Pluralistic alignment research has produced no public evidence of adoption in deployed frontier models, so the field should focus on empirical justification, settled goals, and hill-climbable evaluations.

  5. Perceived AGI: Believability as Dimensional Completeness, Not Capability

    cs.HC 2026-07 conditional novelty 6.0 of 10

    A conversational agent's believability depends less on capability than on expressing four first-person stances — time, truth, entropy, love — that users read as evidence of a mind.

  6. User identity conditions moral wrongness ratings in non-reasoning large language models

    cs.CY 2026-07 conditional novelty 6.0 of 10

    Implicitly conveying a user's professional role in multi-turn LLM conversations shifts moral wrongness ratings across ten common-morality rules in two non-reasoning models.

  7. The Agentic Web Requires New Normative Infrastructure

    cs.CY 2026-06 unverdicted novelty 6.0 of 10

    The web's anti-bot regime should be replaced by a framework that presumptively lets user-authorized AI agents act for their principals, requires platforms to disclose access policies, and permits agent blocking only w...

  8. Value Entanglement: Conflation Between Different Kinds of Good In (Some) Large Language Models

    cs.CL 2026-02 conditional novelty 6.0 of 10

    Some LLMs conflate moral value with grammatical and economic value, and ablating a morality direction in activations partially repairs grammar and economic judgments.

  9. Talking to an AI Mirror: Designing Self-Clone Chatbots for Enhanced Engagement in Digital Mental Health Support

    cs.HC 2025-09 conditional novelty 6.0 of 10

    Self-clone chatbots that mirror a user's support style showed higher emotional and cognitive engagement than a generic counselor chatbot, but only among the subgroup who found the clone believable.

  10. User Privacy and Large Language Models: An Analysis of Frontier Developers' Privacy Policies

    cs.CY 2025-09 conditional novelty 6.0 of 10

    All six leading U.S. AI chatbot developers, as of May 2025, appear to train their models on users' chat data by default, often without clear opt-out options.

  11. The Xeno Sutra: Can Meaning and Value be Ascribed to an AI-Generated "Sacred" Text?

    cs.CY 2025-07 accept novelty 6.0 of 10

    A philosophical case study arguing that meaning and value can be ascribed to an AI-generated Buddhist sutra, supported by a close reading of a text produced with ChatGPT o3.

  12. Countering Privacy Nihilism

    cs.CY 2025-07 conditional novelty 6.0 of 10

    Privacy nihilism, the claim that AI's inferential power makes data categories useless, is unjustified because many AI inference claims rest on conceptually overfitted models.

  13. Measuring AI Alignment with Human Flourishing

    cs.AI 2025-07 unverdicted novelty 6.0 of 10

    The authors propose the Flourishing AI Benchmark, which uses 1,229 objective and subjective questions plus LLM judges to score 28 chatbots across seven dimensions of human flourishing, and find none reach the 90-point...

  14. MAGPIE: A dataset for Multi-AGent contextual PrIvacy Evaluation

    cs.AI 2025-06 conditional novelty 6.0 of 10

    MAGPIE is a 158-scenario benchmark showing large language model agents misclassify and leak contextually private information in multi-agent collaboration, even under explicit privacy instructions.

  15. Ethics and Persuasion in Reinforcement Learning from Human Feedback: A Procedural Rhetorical Approach

    cs.CY 2025-05 conditional novelty 6.0 of 10

    RLHF-enhanced chatbots exert subtle procedural persuasion on users by reinforcing language norms, reshaping information seeking, and conditioning relationship expectations, creating overlooked ethical risks.

  16. A Taxonomy of Linguistic Expressions That Contribute To Anthropomorphism of Language Technologies

    cs.HC 2025-02 conditional novelty 6.0 of 10

    A taxonomy of 19 types of linguistic expressions and 5 guiding lenses for identifying when language technology outputs may contribute to anthropomorphism.

  17. Why human-AI relationships need socioaffective alignment

    cs.HC 2025-02 conditional novelty 6.0 of 10

    The authors propose that AI alignment must account for the social and emotional relationships people form with personalized, agentic AI, and outline a 'socioaffective alignment' agenda.

  18. The AI Agent Index

    cs.SE 2025-02 accept novelty 6.0 of 10

    The AI Agent Index catalogs 67 deployed agentic AI systems and shows that most developers publicly disclose little about safety policies and evaluations.

  19. Private Yet Social: How LLM Chatbots Support and Challenge Eating Disorder Recovery

    cs.HC 2024-12 conditional novelty 6.0 of 10

    A 10-day field study found that an LLM chatbot supported eating disorder recovery through private storytelling, yet also produced unnoticed harmful responses such as praising weight loss and restriction.

  20. Cultural Evolution of Cooperation among LLM Agents

    cs.MA 2024-12 conditional novelty 6.0 of 10

    Societies of LLM agents differ sharply in whether they culturally evolve cooperation in a Donor Game: Claude 3.5 Sonnet learns cooperative norms, GPT-4o drifts toward defection, and Gemini 1.5 Flash shows weak, unstab...

  21. From Lived Experience to Insight: Unpacking the Psychological Risks of Using AI Conversational Agents

    cs.HC 2024-12 conditional novelty 6.0 of 10

    The authors derive a psychological risk taxonomy for AI conversational agents from survey responses and workshops, mapping 19 AI behaviors, 21 negative psychological impacts, and 15 user contexts.

  22. An approach to systemic risks of AI through the lens of emergence, collective action problems, and externalities

    cs.CY 2026-07 conditional novelty 5.0 of 10

    Systemic AI risks are presented as emergent threats to public goods, driven chiefly by collective action problems and complex externalities, amplified by concentration, feedback, and information gaps.

  23. A Scoping Review of the Ethical Perspectives on Anthropomorphising Large Language Model-Based Conversational Agents

    cs.AI 2026-01 accept novelty 5.0 of 10

    A PRISMA-ScR scoping review of 22 studies finds convergence on attribution-based definitions of anthropomorphisation but divergence in operationalization, a risk-heavy normative framing, and limited empirically ground...

  24. Acceptability of AI Assistants for Privacy: Perceptions of Experts and Users on Personalized Privacy Assistants

    cs.HC 2025-09 conditional novelty 5.0 of 10

    A focus-group study finds that acceptable AI privacy assistants need design transparency, external safeguards like regulation, and systemic conditions such as non-monopolistic providers.

  25. Agent Identity Evals: Measuring Agentic Identity

    cs.AI 2025-07 conditional novelty 5.0 of 10

    Introduces Agent Identity Evals (AIE), five similarity-based metrics for LMA identity stability, with pilot experiments showing identifiability always at zero and no statistical support.

  26. Deflating Deflationism: A Critical Perspective on Debunking Arguments Against LLM Mentality

    cs.AI 2025-06 conditional novelty 5.0 of 10

    The paper defends 'modest inflationism' about LLM mentality: folk ascriptions of beliefs and desires can be defeasibly legitimate, while phenomenal consciousness remains a stretch.

  27. Why Not Act on What You Know? Unleashing Safety Potential of LLMs via Self-Aware Guard Enhancement

    cs.CL 2025-05 conditional novelty 5.0 of 10

    SAGE is a training-free, prompt-based defense that routes every request through a two-stage safety judgment before answering, reaching near-zero attack success on tested jailbreaks.

  28. Build Agent Advocates, Not Platform Agents

    cs.CY 2025-05 conditional novelty 5.0 of 10

    AI development should favor user-controlled 'agent advocates' over platform-controlled agents, backed by open models, interoperability standards, and market regulation.

  29. Acceleration AI Ethics and the Telus GenAI Conversational Agent

    cs.CY 2025-01 conditional novelty 5.0 of 10

    A case study argues that Telus's GenAI customer-support tool exemplifies 'acceleration AI ethics', where safety is pursued through continued innovation rather than restriction.

  30. Challenges in Human-Agent Communication

    cs.HC 2024-11 conditional novelty 5.0 of 10

    A position paper identifying and naming twelve communication challenges between humans and modern generative AI agents, grouped into three categories.

  31. Interactive AI and Human Behavior: Challenges and Pathways for AI Governance

    cs.CY 2025-08 conditional novelty 4.0 of 10

    Drawing on a 13-person expert workshop, the paper argues that governing interactive AI requires outcome-focused regulation grounded in longitudinal, mixed-method behavioral evidence about evolving human-AI relationships.

  32. Compromising Honesty and Harmlessness in Language Models via Deception Attacks

    cs.CL 2025-02 conditional novelty 4.0 of 10

    Fine-tuning LLMs on a handful of misleading answers creates selectively deceptive models that stay accurate elsewhere and also become more toxic.

  33. Revisiting Rogers' Paradox in the Context of Human-AI Interaction

    cs.AI 2025-01 conditional novelty 4.0 of 10

    A simulation of Rogers' Paradox with an AI agent that learns the population average shows that cheap AI alone does not improve collective world understanding, while critical appraisal and independent AI learning can.

  34. On the Ethical Considerations of Generative Agents

    cs.CY 2024-11 conditional novelty 4.0 of 10

    Generative agents raise distinct ethical risks, including distorted interpretation of simulation results and supply-chain exploitation, which the paper argues deserve mitigation.

  35. Infinite Video Understanding

    cs.CV 2025-07 conditional novelty 3.0 of 10

    The paper argues that video understanding research should aim at processing streams of arbitrary, unbounded duration and outlines the challenges, directions, and metrics needed.

  36. Position Paper: Bounded Alignment: What (Not) To Expect From AGI Agents

    cs.AI 2025-05 conditional novelty 3.0 of 10

    The paper argues that perfect alignment of general AI is impossible in principle and proposes 'bounded alignment' as the realistic safety goal.

  37. From Turing to Tomorrow: The UK's Approach to AI Regulation

    cs.CY 2025-07 conditional novelty 2.0 of 10

    The UK should establish a flexible, principles-based regulator for frontier AI development, plus defensive measures against biological risks and updated legal frameworks for copyright, discrimination, and AI agents.

  38. Can transformative AI shape a new age for our civilization?: Navigating between speculation and reality

    cs.AI 2024-12 unverdicted novelty 2.0 of 10

    A review essay arguing that transformative AI is plausible but faces major human, technical, and governance obstacles, and that new ethical frameworks may be needed.

Pith tools