Pith. sign in

arXiv:2305.14688 (2023)

9 Pith papers cite this work, alongside 53 external citations. Polarity classification is still indexing.

9 Pith papers citing it
53 external citations · external index

citation-role summary

background 1

citation-polarity summary

roles

background 1

polarities

background 1

representative citing papers

Automated Design of Agentic Systems

cs.AI · 2024-08-15 · conditional · novelty 7.0

Meta Agent Search uses a meta-agent to iteratively program novel agentic systems in code, producing agents that outperform state-of-the-art hand-designed ones across coding, science, and math while transferring across domains and models.

Understanding the Mechanism of Altruism in Large Language Models

econ.GN · 2026-04-21 · unverdicted · novelty 6.0

A small set of sparse autoencoder features in LLMs drives shifts between generous and selfish allocations in dictator games, with causal patching and steering confirming their role and generalization to other social games.

Teaching Astronomy with Large Language Models

physics.ed-ph · 2025-06-07 · unverdicted · novelty 5.0

Structured integration of LLMs in astronomy education, including a domain-specific tutor and documentation requirements, leads to improved AI literacy and reduced student reliance on AI over the semester.

Using Large Language Models in Physics Education

physics.ed-ph · 2026-05-22 · unverdicted · novelty 4.0 · 2 refs

Frontier LLMs from mid-2024 to late-2025 reach near-perfect scores on text-based physics problems and show improved human alignment in grading, but assigning partial credit for flawed reasoning remains difficult.

Dr. Jekyll and Mr. Hyde: Two Faces of LLMs

cs.CR · 2023-12-06 · unverdicted · novelty 3.0

Impersonating complex misaligned personas via biographies and role-play bypasses safety in ChatGPT, Gemini, and Deepseek, succeeding on 38-40 out of 40 illicit questions across tested models.

citing papers explorer

Showing 9 of 9 citing papers.

  • The Moltbook Files: A Harmless Slopocalypse or Humanity's Last Experiment cs.CL · 2026-05-08 · unverdicted · none · ref 7

    An AI-agent social platform generated mostly neutral content whose use in fine-tuning reduced model truthfulness comparably to human Reddit data, suggesting limited unique harm but flagging tail risks like secret leaks.

  • Automated Design of Agentic Systems cs.AI · 2024-08-15 · conditional · none · ref 224

    Meta Agent Search uses a meta-agent to iteratively program novel agentic systems in code, producing agents that outperform state-of-the-art hand-designed ones across coding, science, and math while transferring across domains and models.

  • Understanding the Mechanism of Altruism in Large Language Models econ.GN · 2026-04-21 · unverdicted · none · ref 213

    A small set of sparse autoencoder features in LLMs drives shifts between generous and selfish allocations in dictator games, with causal patching and steering confirming their role and generalization to other social games.

  • RankFlow: A Multi-Role Collaborative Reranking Workflow Utilizing Large Language Models cs.IR · 2025-02-02 · unverdicted · none · ref 68

    RankFlow deploys four LLM roles in sequence to rewrite queries, generate pseudo-answers, summarize passages, and rerank candidates, outperforming prior methods on TREC-DL, BEIR, and NovelEval.

  • SLIP: Soft Label Mechanism and Key-Extraction-Guided CoT-based Defense Against Instruction Backdoor in APIs cs.CR · 2025-08-08 · unverdicted · none · ref 3

    SLIP combines a soft label mechanism with key-extraction-guided CoT to reduce instruction backdoor attack success rate to 25.13% and raise clean accuracy to 87.15% in LLM agents.

  • Teaching Astronomy with Large Language Models physics.ed-ph · 2025-06-07 · unverdicted · none · ref 58

    Structured integration of LLMs in astronomy education, including a domain-specific tutor and documentation requirements, leads to improved AI literacy and reduced student reliance on AI over the semester.

  • Label Over Logic? How Source Cues Bias Human Fallacy Judgments More Than LLMs cs.HC · 2026-05-28 · unverdicted · none · ref 65

    Humans exhibit greater source-label bias in logical fallacy judgments than LLMs, which maintain more consistent evaluations regardless of source cues.

  • Using Large Language Models in Physics Education physics.ed-ph · 2026-05-22 · unverdicted · none · ref 41 · 2 links

    Frontier LLMs from mid-2024 to late-2025 reach near-perfect scores on text-based physics problems and show improved human alignment in grading, but assigning partial credit for flawed reasoning remains difficult.

  • Dr. Jekyll and Mr. Hyde: Two Faces of LLMs cs.CR · 2023-12-06 · unverdicted · none · ref 23

    Impersonating complex misaligned personas via biographies and role-play bypasses safety in ChatGPT, Gemini, and Deepseek, succeeding on 38-40 out of 40 illicit questions across tested models.