Pith. sign in

Paper Citation Record · LEDGER

LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners

As of 22 August 2026, this Paper Citation Record lists 24 of 24 outbound references and 31 inbound Pith citation observations for arXiv:2505.11942.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.11942 v3

Coverage vector

measured 24 of 24 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T20:48:13.467627Z

measured 55 of 55 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 31 of 31 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T15:11:35.363071Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T17:09:59.374964Z

Reference resolution

24 of 24 outbound references displayed

  • verified exact0
  • verified fuzzy7
  • unresolved15
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch2

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 3dbd0a93-d745-4e0f-8d67-78af7a42b733 · outbound

This paper cites GPT-4 Technical Report.

LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners GPT-4 Technical Report

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-15T20:48:13.367578Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:48:13.367578Z digest=sha256:0e41bd8cbb7669dfa570203ce2900abcf02fa089fa445412972333320f310fce

Observation 3b735d48-060e-421f-afc8-09d01be3c167 · outbound

This paper cites Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs.

LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners Do NOT Think That Much for 2+3=? On the Overthinking of o1-Like LLMs

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-15T20:48:13.372321Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:48:13.372321Z digest=sha256:7c1e99f005714361d516a9800280af95af5d782f0457776bc25bdd9d18b4be6c

Observation d98d14d9-7e37-4f9a-a53e-6c8a57a54a23 · outbound

This paper cites Minedojo: Building open-ended embodied agents with internet-scale knowledge.

LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners Minedojo: Building open-ended embodied agents with internet-scale knowledge

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:48:13.919060Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T20:48:13.376982Z digest=sha256:8448fd8b179af1cd2195c8212cdc1c504ad261c1b3b8930c430712dfa9e9b36f

Observation a55f3d1a-f802-4692-8250-67a48d1d3a7d · outbound

This paper cites Catastrophic forgetting in connectionist networks.

LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners Catastrophic forgetting in connectionist networks

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-15T20:48:13.380867Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:48:13.380867Z digest=sha256:9d72c2716cfa88a2c1c90523d0bd2eec3763f4b5b7a03a22c7a1b14d6462545f

Observation e74e092e-a753-4d4d-8812-6ffd75dd99c1 · outbound

This paper cites The Llama 3 Herd of Models.

LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners The Llama 3 Herd of Models

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-15T20:48:13.385367Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:48:13.385367Z digest=sha256:053a79f755ce58d702b13be78eaa079f80e36773cc46b17c8cb584b1c4781a82

Observation 50939ffa-f0f5-428d-a6b9-3c587e0b1244 · outbound

This paper cites Sadler, Percy Liang, Xifeng Yan, and Yu Su.

LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners Sadler, Percy Liang, Xifeng Yan, and Yu Su

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:48:13.895804Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T20:48:13.389812Z digest=sha256:27cbb1ac16682aae9d66db55cf363c96d3246749f20ccf4037d99c442314fc5b

Observation 2a942efe-b00c-4544-a47c-89b9b9b53926 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-15T20:48:13.394573Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:48:13.394573Z digest=sha256:54adaa45c4bb8e961211d488b8774837bd453dbe0a445e00f97e2e05c2fcacd5

Observation aa00b23b-b0c7-4fc8-a23d-31dbf576fa0f · outbound

This paper cites LiveCodeBench: Holistic and Contamination Free Evaluation of Large Language Models for Code.

LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners LiveCodeBench: Holistic and Contamination Free Evaluation of Large Language Models for Code

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T20:48:13.398662Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:48:13.398662Z digest=sha256:ec2f964043ed0db94d774c43f27ffb5e7eaa68b78dca4be679d69494b9a7e662

Observation 73f8705e-ef48-4ff8-a724-8abe1f681f61 · outbound

This paper cites Tree search for language model agents.

LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners Tree search for language model agents

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-15T20:48:13.403617Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:48:13.403617Z digest=sha256:d803ec3bda0cb2b68524f613208205a3243b9a7338c3996b3fd321cef8abc341

Observation fea8fd41-0e54-449d-a32e-0020c3e51d9b · outbound

This paper cites A game theoretic approach to lowering incentives to violate speed limits in Finland.

LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners A game theoretic approach to lowering incentives to violate speed limits in Finland

Reference 10

Resolution
metadata mismatch
local_arxiv, observed 2026-08-15T20:48:13.669902Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T20:48:13.407374Z digest=sha256:d81105ee3916875e516a5e6522e7f72f3f75ecf1e4cebe5e7f049bfb5fbf4b88

Observation 31fd13df-4c97-43d6-9168-548ba699ad90 · outbound

This paper cites From System 1 to System 2: A Survey of Reasoning Large Language Models.

LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners From System 1 to System 2: A Survey of Reasoning Large Language Models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-15T20:48:13.411450Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:48:13.411450Z digest=sha256:d0e51f0af9b1ac719db0964273341702ffba27901809e0186ecab250468ffa23

Observation 9719d5d0-c10d-45ab-b649-279b965bedde · outbound

This paper cites Cross-Dataset Adaptation for Instrument Classification in Cataract Surgery Videos.

LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners Cross-Dataset Adaptation for Instrument Classification in Cataract Surgery Videos

Reference 12

Resolution
metadata mismatch
local_arxiv, observed 2026-08-15T20:48:13.634197Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T20:48:13.415542Z digest=sha256:73de48f8a0d39d447d0811ed01b6cb86429318af12142bdb8c4e12a36a39f1ef

Observation 5a6138c8-0c51-4f60-8392-3dd42cb38eab · outbound

This paper cites VisualAgentBench: Towards Large Multimodal Models as Visual Foundation Agents.

LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners VisualAgentBench: Towards Large Multimodal Models as Visual Foundation Agents

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-15T20:48:13.419453Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:48:13.419453Z digest=sha256:17965c35f8c01770a7d9df1aa1e065b10bd5cae48102e6e7e8ef3a3ec33042b9

Observation 387005ac-794e-41aa-842a-47b827c8b8ab · outbound

This paper cites Large Language Models as Optimizers.

LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners Large Language Models as Optimizers

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-15T20:48:13.423460Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:48:13.423460Z digest=sha256:e94e05ec6f26bc5cda0e6654f5323966895a8c5bc65d3cf72f5709c00bac5ed8

Observation 47288f4a-5acb-443b-bf48-1ae23320dc44 · outbound

This paper cites Gpqa: A graduate-level google-proof q&a benchmark.

LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners Gpqa: A graduate-level google-proof q&a benchmark

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T20:48:13.427576Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:48:13.427576Z digest=sha256:34cc1b7c1dc28431aa5fb3a4f4b8532d4b2aaaf01a732cc22bc38e0047be5c07

Observation fa13e382-22a3-4880-a0f9-a5aaf255f3cd · outbound

This paper cites DebugBench: Evaluating Debugging Capability of Large Language Models.

LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners DebugBench: Evaluating Debugging Capability of Large Language Models

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-15T20:48:13.431841Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:48:13.431841Z digest=sha256:70502475d5596a7c38bd11d4727e2f6a395e92a5bf61938b77300cf369b07e8a

Observation d60ec375-a143-4592-84b9-3fe3f9e72670 · outbound

This paper cites Le, Ed H.

LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners Le, Ed H

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:48:13.875272Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T20:48:13.436152Z digest=sha256:c406d6ed818aaa32d45ef227e09aba7bc0e3beddae37b2dfcc413cd7ac636730

Observation f875d5b4-e04a-40e7-9e23-aefd13b3872e · outbound

This paper cites Qwen2.5 Technical Report.

LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners Qwen2.5 Technical Report

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-15T20:48:13.441226Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:48:13.441226Z digest=sha256:7805b13a9c93c2ef46f7c85e1935a008dd0616518b9b421e8b89ccb51de0a6a2

Observation ac6d6d51-f993-425f-b6db-c8320c2bf74e · outbound

This paper cites Agentoccam: A simple yet strong baseline for llm-based web agents.

LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners Agentoccam: A simple yet strong baseline for llm-based web agents

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:48:13.862675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T20:48:13.445450Z digest=sha256:67e5178ea0bbd5209a18d5455c4a39706c403a459e8225484af8d222ba059383

Observation ae63f2a1-00fa-47c3-bab4-021456f83084 · outbound

This paper cites Towards lifelong learning of large language models: A survey.

LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners Towards lifelong learning of large language models: A survey

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-15T20:48:13.449890Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:48:13.449890Z digest=sha256:e8628a0bac0630e0ad1134ce5317e6887f5174ec0179685fe39ff0f1783aa0ca

Observation 93f12457-9d41-4230-bc38-8e14559d9846 · outbound

This paper cites Lifelong learning of large language model based agents: A roadmap.

LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners Lifelong learning of large language model based agents: A roadmap

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-15T20:48:13.454126Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:48:13.454126Z digest=sha256:f617ad4015dbad5983f6cc098df6de7a30dfc42fba5915c780ccf25aea5c1e15

Observation 872f8212-6cff-4c8d-ad3b-1e250959e711 · outbound

This paper cites Webarena: A realistic web environment for build- ing autonomous agents.

LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners Webarena: A realistic web environment for build- ing autonomous agents

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:48:13.840248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T20:48:13.458228Z digest=sha256:7189614dbf9c813ad9907692c5c0444af72deab6140446a98aca58a2828b3f1f

Observation a54097ba-d635-40a4-adf2-3b65dae13f0b · outbound

This paper cites client,".

LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners client,"

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:48:13.826852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T20:48:13.463287Z digest=sha256:f061662cfd4b195fa7d601845572a1445a9986e65a48c232d7f983348241bb28

Observation 00714672-c52a-44ca-9e31-1dd2281ce3ad · outbound

This paper cites 2023- 10-15.

LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners 2023- 10-15

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T20:48:13.813203Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-15T20:48:13.467627Z digest=sha256:8b3bbdced49a8fb5558718716fadbd970421c79d7ba18629d27be22d0b1eb1ab

Pith citing papers

Observation 46ed92c7-08f1-4692-b51b-e64cb26ba351 · inbound

Continual Learning for Generative AI: From LLMs to MLLMs and Beyond cites this paper.

Continual Learning for Generative AI: From LLMs to MLLMs and Beyond LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners

Reference 260

Resolution
unresolved
no resolver link, observed 2026-08-07T00:40:37.596730Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:40:37.596730Z digest=sha256:9df704678034ef4689fe5b0a5909172d7bf4340ccf8161dd0d4be9c6fdfa3fdc

Observation 934d75bc-9274-4063-a085-9d541639eba9 · inbound

A Survey of Self-Evolving Agents: What, When, How, and Where to Evolve on the Path to Artificial Super Intelligence cites this paper.

A Survey of Self-Evolving Agents: What, When, How, and Where to Evolve on the Path to Artificial Super Intelligence LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners

Reference 100

Resolution
verified exact
arxiv_id, observed 2026-05-14T22:23:15.776528Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-14T22:23:14.621091Z digest=sha256:60b9aa5c1331e8f57bbd90512f69b2055a5532218bf6a17d01519cc56676bb27

Observation 945335f0-3e5d-4dc7-91d9-f2bb909a0158 · inbound

Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory cites this paper.

Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-05-14T23:13:15.719769Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-14T23:13:15.016486Z digest=sha256:c231ea6806a6da7c3eafa0f72db80080bfc19c52c3ab3edba6295ad93e44f035

Observation 102908df-210f-4749-9ad2-ee6c2cb9cf3a · inbound

Scaling Teams or Scaling Time? Memory Enabled Lifelong Learning in LLM Multi-Agent Systems cites this paper.

Scaling Teams or Scaling Time? Memory Enabled Lifelong Learning in LLM Multi-Agent Systems LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-14T22:43:12.033070Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-14T22:42:43.070265Z digest=sha256:9b3e3ccbd28898dca170ebcd88ada22ff5c8d1ab4f4248f35413e38bb34dddee

Observation 40b5de03-7002-4d38-99f1-e954de0aa76d · inbound

LLMs Corrupt Your Documents When You Delegate cites this paper.

LLMs Corrupt Your Documents When You Delegate LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners

Reference 96

Resolution
verified exact
arxiv_id, observed 2026-05-10T09:48:47.713951Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-10T09:47:21.966292Z digest=sha256:76fd360afca9a42051a049f7660b8db6ea4170d780a7f2dc5d9fc57458a7e514

Observation 3bebf6ae-0b6a-4de3-b314-222b4e6c27ac · inbound

GenericAgent: A Token-Efficient Self-Evolving LLM Agent via Contextual Information Density Maximization (V1.0) cites this paper.

GenericAgent: A Token-Efficient Self-Evolving LLM Agent via Contextual Information Density Maximization (V1.0) LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-10T06:56:47.531182Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-10T06:53:42.607059Z digest=sha256:84782e75334ce6dd3e051d918d12c9d40943830e335d25f8344b6e69312bb606

Observation 5db456b3-b1ff-4987-8075-ab3858fead24 · inbound

From Agent Loops to Deterministic Graphs: Execution Lineage for Reproducible AI-Native Work cites this paper.

From Agent Loops to Deterministic Graphs: Execution Lineage for Reproducible AI-Native Work LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-11T20:16:09.330280Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-08T09:50:30.639962Z digest=sha256:5824b636a14e052ad34a74342dfc58fd2e26c8bdb792bf715601b8d31456231d

Observation 8ea3b2b9-260a-40f0-a9fb-86bc0edb6419 · inbound

Learning CLI Agents with Structured Action Credit under Selective Observation cites this paper.

Learning CLI Agents with Structured Action Credit under Selective Observation LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners

Reference 76

Resolution
verified exact
arxiv_id, observed 2026-05-11T03:00:56.358550Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-11T02:59:26.100818Z digest=sha256:0739e2b4d6bca39850adfa99a956abfaad8a101ab90f75c583b087ff3daac32b

Observation a768e694-7ca1-4ee1-9336-ae9873277fe1 · inbound

MINTEval: Evaluating Memory under Multi-Target Interference in Long-Horizon Agent Systems cites this paper.

MINTEval: Evaluating Memory under Multi-Target Interference in Long-Horizon Agent Systems LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-05-20T10:48:12.955233Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-05-20T10:45:08.157521Z digest=sha256:c82ffbd18fd335b97ccef580e5df606727d9527fb961cdcaaa23b550173d1e01

Observation 223a33c4-15ec-4e37-a708-32a2b129ae13 · inbound

Mem-$\pi$: Adaptive Memory through Learning When and What to Generate cites this paper.

Mem-$\pi$: Adaptive Memory through Learning When and What to Generate LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners

Reference 59

Resolution
verified exact
arxiv_id, observed 2026-05-21T04:29:34.495977Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-05-21T04:27:25.041652Z digest=sha256:ce1818915ef0ae46d4afdaddb8f32578f09ab7a956b11e4fce9fd589d0c15e9f

Observation 7fe3ef37-e2ac-43f3-b5d3-f02eff7d353b · inbound

MUSE-Autoskill: Self-Evolving Agents via Skill Creation, Memory, Management, and Evaluation cites this paper.

MUSE-Autoskill: Self-Evolving Agents via Skill Creation, Memory, Management, and Evaluation LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-06-29T17:53:47.806029Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-29T17:41:19.270370Z digest=sha256:9e04ea2481c7307e4dd0d27ad223c5a2a2c40aeb957bcbe4962ec030fa90d8f1

Observation 8e1ee355-1246-4978-bfe0-af7dacd61367 · inbound

AgentCL: Toward Rigorous Evaluation of Continual Learning in Language Agents cites this paper.

AgentCL: Toward Rigorous Evaluation of Continual Learning in Language Agents LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners

Reference 32

Resolution
malformed identifier
arxiv_id, observed 2026-07-01T22:56:20.140079Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-28T14:52:49.244423Z digest=sha256:76016a76d0ef9926b97a07b45c71b24e61715d05436e81586445f5c39e97b307

Observation 68135c1e-bb52-4b5d-a59d-e275a884ae8f · inbound

Learning While Acting: A Skill-Enhanced Test-Time Co-Evolution Framework for Online Lifelong Learning Agents cites this paper.

Learning While Acting: A Skill-Enhanced Test-Time Co-Evolution Framework for Online Lifelong Learning Agents LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners

Reference 24

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T07:06:44.878043Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-28T07:05:18.137509Z digest=sha256:1f07dd4f7237adc23ca1693d2c0e04a8cae330c65bf90104f385b3bb9149c76a

Observation fde2c070-d2d6-4e65-8ffd-2b2f4c651c5b · inbound

M$^3$Eval: Multi-Modal Memory Evaluation through Cognitively-Grounded Video Tasks cites this paper.

M$^3$Eval: Multi-Modal Memory Evaluation through Cognitively-Grounded Video Tasks LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners

Reference 77

Resolution
verified exact
arxiv_id, observed 2026-07-02T08:16:47.802049Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-28T06:16:07.090870Z digest=sha256:f5cfa8e1e659841c3d37a086fa1d007a5c758a1db2a2d46a6ebe5fa639e3a3d5

Observation fcdb34ad-74df-4b64-a80c-8593f9e9f365 · inbound

Continual Learning Bench: Evaluating Frontier AI Systems in Real-World Stateful Environments cites this paper.

Continual Learning Bench: Evaluating Frontier AI Systems in Real-World Stateful Environments LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-07-02T12:56:56.846590Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-28T01:45:28.693098Z digest=sha256:b3848a7cd0c3aee3de04a05e0dd3a6ce38d87fd633e605e218500d13b643b411

Observation e6dea711-ccd8-4f57-93da-c73538e28d79 · inbound

FinEvolveBench: A Benchmark for Self-Evolving Agents on Low-Repetition Tasks with Implicit Rewards cites this paper.

FinEvolveBench: A Benchmark for Self-Evolving Agents on Low-Repetition Tasks with Implicit Rewards LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners

Reference 12

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T16:47:09.741298Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-27T22:24:25.502732Z digest=sha256:cdaa021bf7e80e6e56befc5531eea1758d721d109a850b96d56617f7aea56187

Observation 80781c9c-c8a2-4a08-8be4-a2ce2e46fc6f · inbound

Evaluation of ML Resource Utilization Requires Model Life Cycle Assessment cites this paper.

Evaluation of ML Resource Utilization Requires Model Life Cycle Assessment LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners

Reference 190

Resolution
verified exact
arxiv_id, observed 2026-07-01T21:06:14.632784Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-28T17:27:19.467192Z digest=sha256:f1702452ebf2e458c827810ccdb35b4ef8c81a60fedcb6c7da6977d191fc4a50

Observation 99290d27-b412-4094-afb3-2f219213a6d8 · inbound

Bayesian-Agent: Posterior-Guided Skill Evolution Across LLM Agent Harnesses cites this paper.

Bayesian-Agent: Posterior-Guided Skill Evolution Across LLM Agent Harnesses LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners

Reference 24

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T21:47:25.855088Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-27T19:25:47.085930Z digest=sha256:fdf02a4ca601cd5b6ed0373d3383bdc2d755674cccd92e5e9362d5bc0ff424c7

Observation 6178b1cc-eeee-4452-a557-91b63c2ced00 · inbound

From Player to Master: Enhancing Test-Time Learning of LLM Agents via Reinforcement Learning over Memory cites this paper.

From Player to Master: Enhancing Test-Time Learning of LLM Agents via Reinforcement Learning over Memory LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-07-02T22:37:25.575786Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-27T18:48:30.813878Z digest=sha256:056117550442272e1c99fed3f313a93770cbc1fd1f03a6128d63c75ea9e36b4b

Observation 8fc21509-c1dd-48d0-94ee-0cb1c4e2f20e · inbound

GateMem: Benchmarking Memory Governance in Multi-Principal Shared-Memory Agents cites this paper.

GateMem: Benchmarking Memory Governance in Multi-Principal Shared-Memory Agents LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-07-03T23:49:02.507885Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-26T21:45:13.494097Z digest=sha256:9c9872dfe5ff776042e96d5d7e503c246ad0608d3dbf74de4c3a1a9f95aed2d5

Observation d39e33f6-1562-496e-9a34-1abcf0cab0b8 · inbound

AlphaMemo: Structured Search-Process Memory for Self-Evolving Alpha Mining Agents cites this paper.

AlphaMemo: Structured Search-Process Memory for Self-Evolving Alpha Mining Agents LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-06-29T17:03:41.271266Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-06-29T16:55:21.649886Z digest=sha256:d8d369cde17e857a225390d905cf74994bfb477eda819bd59df62e7ed140a4bf

Observation ef7b8b11-cfe1-40bd-837c-96742ad33f77 · inbound

Are We Ready For An Agent-Native Memory System? cites this paper.

Are We Ready For An Agent-Native Memory System? LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-07-04T17:09:59.376525Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-25T23:52:26.260258Z digest=sha256:d3bc8a9aad2df2583d3f23921a4f56f954db4dc014687f437f0b80434e31a66a

Observation 76303297-79b6-4d46-bf41-60a5ff39214a · inbound

Always-OnAgents:A Survey of Persistent Memory, State, and Governance in LLMAgents cites this paper.

Always-OnAgents:A Survey of Persistent Memory, State, and Governance in LLMAgents LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners

Reference 55

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T15:15:48.385777Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-06-30T03:44:51.320606Z digest=sha256:896fdcab59c63268c1fbc01b529e900f0d731cf4b25d43d04b41d2dc45b8684d

Observation a011244a-6d00-4119-a029-e0ee87c7f1a5 · inbound

EdgeBench: Unveiling Scaling Laws of Learning from Real-World Environments cites this paper.

EdgeBench: Unveiling Scaling Laws of Learning from Real-World Environments LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners

Reference 82

Resolution
unresolved
no resolver link, observed 2026-07-11T07:57:43.000834Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T07:57:43.000834Z digest=sha256:13b00febc07c1b56eee0338b190f5360630fb3e16b3ae1753c33b9c17d352d83

Observation 08d995d4-ee66-46e8-abed-1167c6aff41b · inbound

Rethinking Self-Evolution: A Constrained Exploration-Exploitation Process for Mitigating Skill Overfitting cites this paper.

Rethinking Self-Evolution: A Constrained Exploration-Exploitation Process for Mitigating Skill Overfitting LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-01T11:46:03.282739Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:46:03.282739Z digest=sha256:f4dc33155750ffe6e9bca4ca57c16d01a2edabbc72b1a38a5f1fe1ee31925b8a

Observation 884533dd-6334-41ee-af77-a2f36bc6e103 · inbound

SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution cites this paper.

SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners

Reference 30

Resolution
unresolved
no resolver link, observed 2026-07-30T21:12:59.453856Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T21:12:59.453856Z digest=sha256:3ccd4864c52289b0bc8df5ea3aa2be99498e7fa14241f2a51a180b4e6a0c7b0a

Observation 3f8ec626-822f-4a24-b2a2-620a39c5a29a · inbound

Progressive Multimodal Alignment for Continual Instruction Tuning cites this paper.

Progressive Multimodal Alignment for Continual Instruction Tuning LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners

Reference 54

Resolution
unresolved
no resolver link, observed 2026-07-30T16:11:52.369041Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T16:11:52.369041Z digest=sha256:4a65362aba7ca30ef76b6ddf720319b19db8b360c787a668c298da3d1519fb08

Observation 5a7635e6-51f1-4379-9d3d-35c8d1c18e8d · inbound

Progressive Multimodal Alignment for Continual Instruction Tuning cites this paper.

Progressive Multimodal Alignment for Continual Instruction Tuning LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-03T01:24:03.664375Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T01:24:03.664375Z digest=sha256:59c00d56a493fe5351d90cae64e105e5d0e76ef32db3f9613dff28b6eb9d10ce

Observation 25d47a20-a4eb-4e00-9e2f-45b721254bb9 · inbound

CoEvo-Mem: Co-Evolving Retrieval Policy and Memory Bank for LLM Agents cites this paper.

CoEvo-Mem: Co-Evolving Retrieval Policy and Memory Bank for LLM Agents LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-15T15:11:35.363071Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T15:11:35.363071Z digest=sha256:163bcecc208ef2ce964fe51b89f9a145785b1c16d38c7a4df786cd8ebeb6fe22

Observation 29246f5b-6c86-4555-994b-062debd0b8bb · inbound

FinEvo-Bench: A Longitudinal Benchmark for Self-Evolving Agents in Professional Financial Workflows cites this paper.

FinEvo-Bench: A Longitudinal Benchmark for Self-Evolving Agents in Professional Financial Workflows LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T14:10:36.008325Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:10:36.008325Z digest=sha256:5f2fefc237c3e85e4f6338fecc0aa8532fc7ccf6b7be9de3ef0298e370eb06dc

Observation 541eb622-8323-4e1e-b4ea-ccb2a1b59808 · inbound

VibeLifeBench: Can Your Life Agent Be Proactive and Persistent in a Living World? cites this paper.

VibeLifeBench: Can Your Life Agent Be Proactive and Persistent in a Living World? LifelongAgentBench: Evaluating LLM Agents as Lifelong Learners

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-12T15:18:12.475969Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:18:12.475969Z digest=sha256:5e43bc70cd33220989d45624f1fb1042a41e902bfe3dee8dd26be3c11b03090f