Pith. sign in

Promptbreeder: Self-Referential Self-Improvement via Prompt Evolution , booktitle =

3 Pith papers cite this work. Polarity classification is still indexing.

3 Pith papers citing it

fields

cs.CL 2 cs.SE 1

years

2026 2 2025 1

representative citing papers

TTHE: Test-Time Harness Evolution

cs.SE · 2026-07-09 · conditional · novelty 6.0

An LLM agent can improve itself at test time by rewriting its surrounding executable harness from unlabeled traces, using only proxy signals and a frozen model.

Memory in the Age of AI Agents

cs.CL · 2025-12-15 · unverdicted · novelty 6.0

The paper maps agent memory research via three forms (token-level, parametric, latent), three functions (factual, experiential, working), and dynamics of formation/evolution/retrieval, plus benchmarks and future directions.

Context Training with Active Information Seeking

cs.CL · 2026-05-13 · unverdicted · novelty 5.0 · 2 refs

Active information seeking via search tools, when combined with multi-candidate context pruning during training, produces consistent gains on translation, health, and reasoning tasks over naive tool addition or no-tool baselines.

citing papers explorer

Showing 3 of 3 citing papers.

  • TTHE: Test-Time Harness Evolution cs.SE · 2026-07-09 · conditional · none · ref 11

    An LLM agent can improve itself at test time by rewriting its surrounding executable harness from unlabeled traces, using only proxy signals and a frozen model.

  • Memory in the Age of AI Agents cs.CL · 2025-12-15 · unverdicted · none · ref 96

    The paper maps agent memory research via three forms (token-level, parametric, latent), three functions (factual, experiential, working), and dynamics of formation/evolution/retrieval, plus benchmarks and future directions.

  • Context Training with Active Information Seeking cs.CL · 2026-05-13 · unverdicted · none · ref 10 · 2 links

    Active information seeking via search tools, when combined with multi-candidate context pruning during training, produces consistent gains on translation, health, and reasoning tasks over naive tool addition or no-tool baselines.