Pith. sign in

Paper Citation Record · LEDGER

Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory

As of 5 August 2026, this Paper Citation Record lists 2 of 2 outbound references and 68 inbound Pith citation observations for arXiv:2511.20857.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2511.20857 v2

Coverage vector

measured 2 of 2 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-21T18:04:48.156727Z

measured 70 of 70 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00

measured 68 of 68 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-04T21:37:35.836340Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-04T17:20:00.018935Z

Reference resolution

2 of 2 outbound references displayed

  • verified exact1
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 7e2fa6bd-8d0d-4e51-b189-993c0edfe3f7 · outbound

This paper cites Evaluating Very Long-Term Conversational Memory of LLM Agents.

Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory Evaluating Very Long-Term Conversational Memory of LLM Agents

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-21T18:05:27.147044Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-21T18:04:48.156727Z digest=sha256:c089369dd9757602ab8a1b4101d81ad8bc25de88a69339b6180ea47653df35df

Observation 799b9935-80e0-4020-b1a5-f045028d9abb · outbound

This paper cites 1,3” or “2-4.

Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory 1,3” or “2-4

Reference 2

Resolution
malformed identifier
raw_fallback, observed 2026-05-21T18:05:27.248013Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-21T18:04:48.156727Z digest=sha256:f52fbb84bf8c82ce11e8d2e31ef14126048d5a7ea9b2e13b1370a4dde4f5289d

Pith citing papers

Observation 0837b2ba-1b6a-4925-a54d-0c4e790718f1 · inbound

Agentic Reasoning for Large Language Models cites this paper.

Agentic Reasoning for Large Language Models Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-05-17T15:14:25.765131Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-17T15:14:25.558878Z digest=sha256:cb0f4dce26ec2d8ea03f2687520c0a002ed548bd62043666e602db7f6358d608

Observation 2d6e198c-3d5b-43e2-a5b9-11ac55bdbb70 · inbound

Toward Efficient Agents: Memory, Tool learning, and Planning cites this paper.

Toward Efficient Agents: Memory, Tool learning, and Planning Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory

Reference 145

Resolution
unresolved
no resolver link, observed 2026-08-03T09:21:45.040455Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T09:21:45.040455Z digest=sha256:75d10414ca1d723d0163b52b111a87a3a548c08f19e4f1331299a6f55e2f1df0

Observation d2db6c0e-0fc4-41a3-9046-10b6cbba76dd · inbound

Improve Large Language Model Systems with User Logs cites this paper.

Improve Large Language Model Systems with User Logs Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory

Reference 40

Resolution
verified exact
local_arxiv, observed 2026-05-21T14:34:12.932663Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-21T14:34:01.332088Z digest=sha256:bf1c8611a2fc964d099458115585a615aa1b2f95164cab010f32b073489222f3

Observation 162a938e-68b1-45ea-a0c3-e5c58b3c8886 · inbound

Improve Large Language Model Systems with User Logs cites this paper.

Improve Large Language Model Systems with User Logs Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-03T04:00:20.314092Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T04:00:20.314092Z digest=sha256:cf1996cbed1d221db6eff9ed19f8d140773e143163f0b5723a34d723361298b7

Observation 34861772-4e01-4d15-8107-7d542ab0be67 · inbound

Scaling Teams or Scaling Time? Memory Enabled Lifelong Learning in LLM Multi-Agent Systems cites this paper.

Scaling Teams or Scaling Time? Memory Enabled Lifelong Learning in LLM Multi-Agent Systems Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-14T23:13:16.650386Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T22:42:43.070265Z digest=sha256:a41d9124ecbcc548eda4e7a86da7e39b43ef736fb8a80d9ba52ac473ac1966d0

Observation 85f52853-f057-4f59-937f-27754691a28a · inbound

ActionNex: A Virtual Outage Manager for Cloud Computing cites this paper.

ActionNex: A Virtual Outage Manager for Cloud Computing Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-14T23:13:16.650386Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-13T19:18:39.881155Z digest=sha256:8914070ff4e70ac2e60d6d9ce12437f91ab52ef93e62378d1b445c2f4ab56aa2

Observation 901c90a3-fe5f-49fd-9bba-0da567080a59 · inbound

TSUBASA: Improving Long-Horizon Personalization via Evolving Memory and Self-Learning with Context Distillation cites this paper.

TSUBASA: Improving Long-Horizon Personalization via Evolving Memory and Self-Learning with Context Distillation Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory

Reference 72

Resolution
verified exact
arxiv_id, observed 2026-05-14T23:13:16.650386Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-10T16:48:53.092395Z digest=sha256:bf316c4e3a38e32d20bd92f169f9d731305b2ab16b76a4b03285471bf595f89f

Observation 0f50ded2-146e-48dd-9d83-fdbbdfd0e440 · inbound

MemCoT: Test-Time Scaling through Memory-Driven Chain-of-Thought cites this paper.

MemCoT: Test-Time Scaling through Memory-Driven Chain-of-Thought Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory

Reference 38

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T23:13:16.650386Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T17:46:59.581486Z digest=sha256:6d7278b8a7565c18ae29fd802e26301dd124fd06d63783ddfdda6f3467dc6714

Observation a22e10f1-281c-4305-b4ef-f8f02df92a0e · inbound

MemCoT: Test-Time Scaling through Memory-Driven Chain-of-Thought cites this paper.

MemCoT: Test-Time Scaling through Memory-Driven Chain-of-Thought Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory

Reference 38

Resolution
metadata mismatch
local_arxiv, observed 2026-05-21T09:34:57.214613Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-21T09:34:16.292323Z digest=sha256:a93eb2b0c78e61fdd667242e3a711611370b77362e773885a21d08e8211ff544

Observation a8af0925-a521-4c93-9d65-d64adcad0345 · inbound

M$^\star$: Every Task Deserves Its Own Memory Harness cites this paper.

M$^\star$: Every Task Deserves Its Own Memory Harness Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory

Reference 18

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T23:13:16.650386Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T17:40:12.368228Z digest=sha256:644cd6c687e0aa80061a2040ebbb707f2f74e789d99619c21dd3c6f0d5eb1461

Observation bccacb77-9bac-48a5-8127-bc4143e7a00d · inbound

Evo-MedAgent: Beyond One-Shot Diagnosis with Agents That Remember, Reflect, and Improve cites this paper.

Evo-MedAgent: Beyond One-Shot Diagnosis with Agents That Remember, Reflect, and Improve Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-14T23:13:16.650386Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T12:35:32.915685Z digest=sha256:f0fe7b84d21484c6d21554aa7ed33d6133cb3431ae579fc0c4157a7c72766633

Observation 4e900222-1b84-4215-827a-f08631de3ea6 · inbound

From Procedural Skills to Strategy Genes: Towards Experience-Driven Test-Time Evolution cites this paper.

From Procedural Skills to Strategy Genes: Towards Experience-Driven Test-Time Evolution Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory

Reference 15

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T23:13:16.650386Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T10:52:47.760627Z digest=sha256:57e7ad5558010f808da3ea9adb87ac4ec62db6862bc216bc620c852bd45e7430

Observation e60cb321-02d7-44c4-8bfd-a5351ade78fb · inbound

From Procedural Skills to Strategy Genes: Towards Experience-Driven Test-Time Evolution cites this paper.

From Procedural Skills to Strategy Genes: Towards Experience-Driven Test-Time Evolution Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory

Reference 15

Resolution
unresolved
no resolver link, observed 2026-07-12T19:51:48.531280Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T19:51:48.531280Z digest=sha256:e2016d8384d635388c934ee16f7ee07e85642cb54d3fa13a5faad9b2a3a1074b

Observation 542a4428-6f28-43e9-9a91-f6504d871da7 · inbound

SkillFlow:Benchmarking Lifelong Skill Discovery and Evolution for Autonomous Agents cites this paper.

SkillFlow:Benchmarking Lifelong Skill Discovery and Evolution for Autonomous Agents Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-05-14T23:13:16.650386Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-10T06:13:32.434201Z digest=sha256:5d26b93984f46f05989a62da3436eefbc72dd96c8fcfb66f1ab58ebbe1dae346

Observation 96ade66f-415d-475f-a0ca-790d84caf1c1 · inbound

Bian Que: An Agentic Framework with Flexible Skill Arrangement for Online System Operations cites this paper.

Bian Que: An Agentic Framework with Flexible Skill Arrangement for Online System Operations Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-05-14T23:13:16.650386Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-07T11:17:46.692409Z digest=sha256:de5d1b21144e7bc48e4d9e3e26a248a997ac2d090cfdd9e8584ab0a1372133d8

Observation d54c7b03-1909-474a-be23-6c38d24dff77 · inbound

Bian Que: An Agentic Framework with Flexible Skill Arrangement for Online System Operations cites this paper.

Bian Que: An Agentic Framework with Flexible Skill Arrangement for Online System Operations Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory

Reference 48

Resolution
verified exact
arxiv_id, observed 2026-05-14T23:13:16.650386Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-12T01:50:45.507172Z digest=sha256:72b7e75edcd451a268348aa5d7b8356647518576617d8a722ccfa25e903c4ef6

Observation 1abc07e5-e386-4813-b3bf-1c4346d8eb2e · inbound

Agentic-imodels: Evolving agentic interpretability tools via autoresearch cites this paper.

Agentic-imodels: Evolving agentic interpretability tools via autoresearch Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory

Reference 63

Resolution
verified exact
arxiv_id, observed 2026-05-14T23:13:16.650386Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-07T16:37:43.371592Z digest=sha256:73058a84f25094aca6e78347b0b73b4363c1ab2527413e88dd1845c8fe61b64d

Observation 6b35f8de-26d3-4610-85c0-9623e6a8232d · inbound

Skill1: Unified Evolution of Skill-Augmented Agents via Reinforcement Learning cites this paper.

Skill1: Unified Evolution of Skill-Augmented Agents via Reinforcement Learning Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory

Reference 22

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T23:13:16.650386Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-08T10:23:52.522238Z digest=sha256:d46445bea08677408dc56bc1ac8e20b70cf0c24e9c17bb2b65d99e7d599e0b6a

Observation e9173142-86b6-423b-b790-2cd94e31e416 · inbound

Skill1: Unified Evolution of Skill-Augmented Agents via Reinforcement Learning cites this paper.

Skill1: Unified Evolution of Skill-Augmented Agents via Reinforcement Learning Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory

Reference 22

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T23:13:16.650386Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-11T02:00:00.663355Z digest=sha256:e694b19676756fc5b5c5a6a565ea0b131edc4ae8bfcbc2b7a87a4f23b5188b48

Observation 27e3c272-0d5b-46f2-a228-c86c589adea4 · inbound

Skill1: Unified Evolution of Skill-Augmented Agents via Reinforcement Learning cites this paper.

Skill1: Unified Evolution of Skill-Augmented Agents via Reinforcement Learning Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory

Reference 22

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T23:13:16.650386Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-13T07:17:13.708752Z digest=sha256:e65f29385ab86658c543ad5f6a33a622550b6c7e536443a1cd55e7cf7344554e

Observation 046f4948-05dc-4eeb-b9cf-00d648ced2ba · inbound

From Agent Loops to Deterministic Graphs: Execution Lineage for Reproducible AI-Native Work cites this paper.

From Agent Loops to Deterministic Graphs: Execution Lineage for Reproducible AI-Native Work Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory

Reference 48

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T23:13:16.650386Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-08T09:50:30.639962Z digest=sha256:1ac80fd1de63e0ec8f4fc56c1a305ef387a8307ac4636c00deefa694bfdf5fc5

Observation ec3adc7a-e507-48a1-9c81-c7c9d35a3d10 · inbound

A Comprehensive Survey on Agent Skills: Taxonomy, Techniques, and Applications cites this paper.

A Comprehensive Survey on Agent Skills: Taxonomy, Techniques, and Applications Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory

Reference 146

Resolution
verified exact
arxiv_id, observed 2026-05-14T23:13:16.650386Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-11T01:47:39.926540Z digest=sha256:d8d904a6a842d4733fc27c02aa2e5e706b7e2ff6ea9ea182e013e4cb002815c6

Observation 35c27b98-50d0-43c0-825f-d6c8348da9bc · inbound

A Comprehensive Survey on Agent Skills: Taxonomy, Techniques, and Applications cites this paper.

A Comprehensive Survey on Agent Skills: Taxonomy, Techniques, and Applications Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory

Reference 148

Resolution
verified exact
local_arxiv, observed 2026-05-20T23:19:14.814754Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-20T23:15:44.550045Z digest=sha256:4a27e42ba59bb57e9424804d3521f5d3faf92b2a015ca5926079aa83d567252d

Observation e87966a1-9cdc-4f5e-8d91-a6102df020e7 · inbound

A Comprehensive Survey on Agent Skills: Taxonomy, Techniques, and Applications cites this paper.

A Comprehensive Survey on Agent Skills: Taxonomy, Techniques, and Applications Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory

Reference 140

Resolution
verified exact
local_arxiv, observed 2026-06-30T23:25:07.289996Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-30T23:23:42.883286Z digest=sha256:9ae2d1aac3990f4d804d22525c71d64ca38060b1abc2f6c2f4561d0356524392

Observation 22175fa4-fb63-4439-a326-c859a44c11fd · inbound

MemCompiler: Compile, Don't Inject -- State-Conditioned Memory for Embodied Agents cites this paper.

MemCompiler: Compile, Don't Inject -- State-Conditioned Memory for Embodied Agents Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-14T23:13:16.650386Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-11T02:36:53.775557Z digest=sha256:3fcc2c87d8d2ebb108b3f1a23d4bf8bc377b96a2952d1af987d141c0c3e3db5e

Observation d3c3b01d-8889-4961-abcf-0aaf5839706f · inbound

MemCompiler: Compile, Don't Inject -- State-Conditioned Memory for Embodied Agents cites this paper.

MemCompiler: Compile, Don't Inject -- State-Conditioned Memory for Embodied Agents Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory

Reference 40

Resolution
verified exact
local_arxiv, observed 2026-05-15T05:59:48.997096Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-15T05:57:05.259111Z digest=sha256:b3375676bdcbe9ec2e78ceaa560ddcc6e0e1f2360f460dbcf013a4cba371b9be

Observation 545925b9-4c11-4566-a109-7185f2096439 · inbound

AgentPSO: Evolving Agent Reasoning Skill via Multi-agent Particle Swarm Optimization cites this paper.

AgentPSO: Evolving Agent Reasoning Skill via Multi-agent Particle Swarm Optimization Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory

Reference 42

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T23:13:16.650386Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-12T00:55:52.965890Z digest=sha256:00a24792c015d678d04a88ecf24d8ddd2501bf51c8055d3041b9fcedd128b4fb

Observation aaccb7a4-5374-4895-88b7-e2c25f86751a · inbound

AgentPSO: Evolving Agent Reasoning Skill via Multi-agent Particle Swarm Optimization cites this paper.

AgentPSO: Evolving Agent Reasoning Skill via Multi-agent Particle Swarm Optimization Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory

Reference 42

Resolution
metadata mismatch
local_arxiv, observed 2026-06-30T23:35:07.629183Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-30T23:29:12.545347Z digest=sha256:b4236007695fe65133172dc7808cce224c0f435e53a8dc9def7c0aaed0827367

Observation f5c80c84-4e9d-4b40-ab64-e0496befc17b · inbound

Do Self-Evolving Agents Forget? Capability Degradation and Preservation in Lifelong LLM Agent Adaptation cites this paper.

Do Self-Evolving Agents Forget? Capability Degradation and Preservation in Lifelong LLM Agent Adaptation Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory

Reference 19

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T23:13:16.650386Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-12T04:10:31.784413Z digest=sha256:338439a7dd78443d9c1665273504d4be7425d7788b8f891152a178474752c1c9

Observation ad4452eb-2f58-4c98-97ac-ad8c2ad2c128 · inbound

MAGE: Multi-Agent Self-Evolution with Co-Evolutionary Knowledge Graphs cites this paper.

MAGE: Multi-Agent Self-Evolution with Co-Evolutionary Knowledge Graphs Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-14T23:13:16.650386Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-12T05:13:28.089038Z digest=sha256:56d8d9a7eb34be77dceb969eb039c85016aa7ba39a4b7bf725224841fe55b70a

Observation b18eb268-a1be-4705-9d75-605a53e8741f · inbound

MemReread: Enhancing Agentic Long-Context Reasoning via Memory-Guided Rereading cites this paper.

MemReread: Enhancing Agentic Long-Context Reasoning via Memory-Guided Rereading Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-05-14T23:13:16.650386Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-12T05:22:43.330891Z digest=sha256:cf373dc21853c7d9a744af2aab85f5f20d0468204e1b589e8eccda43070c2570

Observation ac261d9c-b950-4c1c-8f05-92fc46e95974 · inbound

Shepherd: Enabling Programmable Meta-Agents via Reversible Agentic Execution Traces cites this paper.

Shepherd: Enabling Programmable Meta-Agents via Reversible Agentic Execution Traces Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory

Reference 42

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T23:13:16.650386Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-12T03:29:33.497561Z digest=sha256:1dd44e50b588a47fbb462b764e6e919ab15228f67e4b23755e2d623c43ea7d55

Observation 298d372a-b8a5-44c8-be60-66a7d322fcb7 · inbound

Shepherd: Enabling Programmable Meta-Agents via Reversible Agentic Execution Traces cites this paper.

Shepherd: Enabling Programmable Meta-Agents via Reversible Agentic Execution Traces Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory

Reference 42

Resolution
metadata mismatch
local_arxiv, observed 2026-06-30T22:25:06.845608Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-30T22:24:26.528532Z digest=sha256:1ce203d5b9b23243ff1f0bdb18c6165c13f3248b84f1c694bafee8dd725b0d5a

Observation 7731b378-3142-4b4c-bce8-e214c85c5e38 · inbound

Context Training with Active Information Seeking cites this paper.

Context Training with Active Information Seeking Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory

Reference 60

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T23:13:16.650386Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-14T20:03:01.429424Z digest=sha256:861fea4a87e8441f72dd28f7532b6ca23af47e4e5609411a4b80722d106a7b44

Observation cf3aad9b-f80b-4de0-a85f-91308d2ec64d · inbound

Context Training with Active Information Seeking cites this paper.

Context Training with Active Information Seeking Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory

Reference 60

Resolution
metadata mismatch
local_arxiv, observed 2026-05-15T06:09:49.991157Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-15T06:07:10.805180Z digest=sha256:324c00435a325766ffce3b8c46fb05d93eb51758a4c8f4149b8f817b0506a616

Observation 6af1ea49-8558-49c3-a9c3-a49393cbfb23 · inbound

RealICU: Do LLM Agents Understand Long-Context ICU Data? A Benchmark Beyond Behavior Imitation cites this paper.

RealICU: Do LLM Agents Understand Long-Context ICU Data? A Benchmark Beyond Behavior Imitation Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-14T23:13:16.650386Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-14T18:47:46.239810Z digest=sha256:3221d4ca006cf960f9eef27e80916bd350bff929c01fdec538d8eae8b1688132

Observation c19c32c9-dcc7-44d4-80e2-40fa9ea126dd · inbound

EvolveMem:Self-Evolving Memory Architecture via AutoResearch for LLM Agents cites this paper.

EvolveMem:Self-Evolving Memory Architecture via AutoResearch for LLM Agents Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory

Reference 31

Resolution
verified exact
local_arxiv, observed 2026-05-15T04:49:44.079669Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-15T04:48:12.883761Z digest=sha256:2b19ad97cc33945f5453c323a4beea7bbb0d86226286221b790db3d816f71eb4

Observation 8ec255aa-ff9b-492d-a17e-5345aad3104f · inbound

Test-Time Learning with an Evolving Library cites this paper.

Test-Time Learning with an Evolving Library Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-05-15T01:48:28.969956Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-15T01:44:16.865595Z digest=sha256:ab050504e1c77146542f4b57a59bb3ce711f50e87dcd5c71ca20933a1f2688e8

Observation c86321c1-05b6-446c-b9e4-48b35335e3cb · inbound

Test-Time Learning with an Evolving Library cites this paper.

Test-Time Learning with an Evolving Library Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-02T14:06:40.867144Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T14:06:40.867144Z digest=sha256:a210e0394dd207140bfaf02a80f779f5587df199403d520aecfa217195cac39d

Observation 1312ebbd-7bd8-470b-86a2-ce2c58c2b570 · inbound

Is One Score Enough? Rethinking the Evaluation of Sequentially Evolving LLM Memory cites this paper.

Is One Score Enough? Rethinking the Evaluation of Sequentially Evolving LLM Memory Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory

Reference 35

Resolution
verified exact
local_arxiv, observed 2026-05-19T16:47:40.489988Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-19T16:43:37.472644Z digest=sha256:6eb0cabcd5edf3d3082e52b61d2e564d71a5449c32411b047c2d9f7134e4a0a9

Observation 80b0efc5-e7bb-45d7-91c6-bebcd39af8ac · inbound

EXG: Self-Evolving Agents with Experience Graphs cites this paper.

EXG: Self-Evolving Agents with Experience Graphs Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory

Reference 31

Resolution
metadata mismatch
local_arxiv, observed 2026-05-19T22:17:49.338215Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-19T22:17:10.728289Z digest=sha256:aa7c7b5d3ff073a879cd89802908c802e8b917c8a01f0682fa1a50c483896c2a

Observation 88f441a7-c589-4cc2-b067-fa1df2bb3a91 · inbound

EvoMemBench: Benchmarking Agent Memory from a Self-Evolving Perspective cites this paper.

EvoMemBench: Benchmarking Agent Memory from a Self-Evolving Perspective Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory

Reference 35

Resolution
metadata mismatch
local_arxiv, observed 2026-05-20T11:38:14.541752Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-20T11:35:23.892275Z digest=sha256:7cc941b218c153cea0b72c0fa662aa6beaab9183940808497c44fce1379b57bb

Observation 5e0676b8-bc48-4c42-b361-08a1b64eb9a0 · inbound

Code as Agent Harness cites this paper.

Code as Agent Harness Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory

Reference 202

Resolution
verified exact
local_arxiv, observed 2026-05-20T10:58:14.364346Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-20T10:54:54.558241Z digest=sha256:b37943b750cea6b99b3905c1a0f6b826e1ecd95479b55d194d6d09f57f6c490d

Observation b1eae535-c00f-4620-aeb5-00429443b89c · inbound

Auto-Dreamer: Learning Offline Memory Consolidation for Language Agents cites this paper.

Auto-Dreamer: Learning Offline Memory Consolidation for Language Agents Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory

Reference 33

Resolution
metadata mismatch
local_arxiv, observed 2026-05-21T05:39:40.764887Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-05-21T05:36:56.151440Z digest=sha256:686fd2e83afd3ab88985307db9a470ef3751a3bbaaa4205a09595d5f6f59526b

Observation 0c82230f-c0e9-49ba-b087-1b26e3407495 · inbound

Rethinking Memory as Continuously Evolving Connectivity cites this paper.

Rethinking Memory as Continuously Evolving Connectivity Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory

Reference 53

Resolution
metadata mismatch
local_arxiv, observed 2026-06-29T12:33:24.800094Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-06-29T12:25:50.220145Z digest=sha256:fc27bb4d936383c396962f86dd2b6a2cda5204886fc0ea90ff21abe1429ac571

Observation 56be5de9-6e7a-4d25-95ff-a44afd8099a3 · inbound

Connecting the Dots: Benchmarking Reflective Memory in Long-Horizon Dialogue cites this paper.

Connecting the Dots: Benchmarking Reflective Memory in Long-Horizon Dialogue Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory

Reference 60

Resolution
metadata mismatch
local_arxiv, observed 2026-06-28T17:42:25.950848Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-06-28T17:41:06.216002Z digest=sha256:6ecd2135bc072ecd08a773073eef6dcffeabb2f311cfbc3228a2d7315249163b

Observation f6b69651-616d-4605-aebd-d29360283f78 · inbound

AgentCL: Toward Rigorous Evaluation of Continual Learning in Language Agents cites this paper.

AgentCL: Toward Rigorous Evaluation of Continual Learning in Language Agents Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory

Reference 20

Resolution
metadata mismatch
local_arxiv, observed 2026-07-01T22:56:20.164476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T14:52:49.244423Z digest=sha256:66fff28e778086f9fbccaa0f1915e7588a003501af8e5c5bd9527d5550eebcff

Observation 18914754-6459-4567-b3c1-236c4e069731 · inbound

M$^3$Eval: Multi-Modal Memory Evaluation through Cognitively-Grounded Video Tasks cites this paper.

M$^3$Eval: Multi-Modal Memory Evaluation through Cognitively-Grounded Video Tasks Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory

Reference 66

Resolution
verified exact
local_arxiv, observed 2026-07-02T08:16:47.771817Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T06:16:07.090870Z digest=sha256:a7e7e7cffdc92a048a64ba1d020104163135671dcfb40720dea427e83345881c

Observation b134c7f7-175b-41f6-b55b-2fe0f6311855 · inbound

EpiEvolve: Self-Evolving Agents for Streaming Pandemic Forecasting under Regime Shifts cites this paper.

EpiEvolve: Self-Evolving Agents for Streaming Pandemic Forecasting under Regime Shifts Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory

Reference 27

Resolution
metadata mismatch
local_arxiv, observed 2026-07-02T09:06:49.133300Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-06-28T05:40:18.458584Z digest=sha256:9a5200759ba2249ea0d85e1f884143a225a0653fbf907d4238205bcbd626598f

Observation bf09f45e-3139-427f-b4b9-31967b364d20 · inbound

Agent Memory: Characterization and System Implications of Stateful Long-Horizon Workloads cites this paper.

Agent Memory: Characterization and System Implications of Stateful Long-Horizon Workloads Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-07-02T13:36:59.295338Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-28T01:10:02.610738Z digest=sha256:161dd9a6286b8aadbc29d0ccb88bebbd3bd09634199c6fd4b10fee3ad840ed8e

Observation 1343f239-25ba-481f-b88f-1e4a8892f3c6 · inbound

Tree-of-Experience: A Structured Experience-Management Solution for Self-Evolving Agents under Low-Repetition and Implicit-Reward Environments cites this paper.

Tree-of-Experience: A Structured Experience-Management Solution for Self-Evolving Agents under Low-Repetition and Implicit-Reward Environments Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory

Reference 9

Resolution
metadata mismatch
local_arxiv, observed 2026-07-02T16:47:09.749147Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-27T22:24:25.502732Z digest=sha256:2cb0a37eb55f1a434ad269754370f7d17cc28fe845da16a3625afb838ae92f93

Observation 326fa9a8-b64e-4147-a12e-8da722189dbb · inbound

Experience Makes Skillful: Enabling Generalizable Medical Agent Reasoning via Self-Evolving Skill Memory cites this paper.

Experience Makes Skillful: Enabling Generalizable Medical Agent Reasoning via Self-Evolving Skill Memory Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory

Reference 72

Resolution
verified exact
local_arxiv, observed 2026-07-03T01:07:30.920151Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-27T16:45:30.431403Z digest=sha256:dba5e3a9228b0f6219e68a6f34239c9a3fac2c9d92716b6441a3d71895b26426

Observation fcd865db-df08-447b-9b56-eaabc62d6fce · inbound

Marginal Advantage Accumulation for Memory-Driven Agent Self-Evolution cites this paper.

Marginal Advantage Accumulation for Memory-Driven Agent Self-Evolution Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory

Reference 18

Resolution
metadata mismatch
local_arxiv, observed 2026-07-04T03:39:30.783334Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-06-26T17:45:49.272070Z digest=sha256:0c90278ebf6d263c213fba7755c9d9aa3405b6cb9b9259ea933a5a4f807b1195

Observation 4c24f545-ca42-47ca-8bd6-d0803abc7a13 · inbound

AlphaMemo: Structured Search-Process Memory for Self-Evolving Alpha Mining Agents cites this paper.

AlphaMemo: Structured Search-Process Memory for Self-Evolving Alpha Mining Agents Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory

Reference 2

Resolution
metadata mismatch
local_arxiv, observed 2026-06-29T17:03:41.262391Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-06-29T16:55:21.649886Z digest=sha256:5f5f17932b07a310275dad65853774d04786753889f1f9f8f96ace00e5f7a1d2

Observation d0de9262-ab7b-4fbf-ac70-6d8f9738b9e0 · inbound

MetaPS: Adaptive Programmatic Strategy Selection for Market Agents cites this paper.

MetaPS: Adaptive Programmatic Strategy Selection for Market Agents Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory

Reference 132

Resolution
metadata mismatch
local_arxiv, observed 2026-07-04T08:39:42.650845Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-06-26T11:06:28.690956Z digest=sha256:bd1cfb243cdfb4497abc29ace6fd61bc1c8ddb47d4bf2bd9791508f962a3e51a

Observation 3847d528-82cd-482e-9e54-6b382803f934 · inbound

Escaping the Self-Confirmation Trap: An Execute-Distill-Verify Paradigm for Agentic Experience Learning cites this paper.

Escaping the Self-Confirmation Trap: An Execute-Distill-Verify Paradigm for Agentic Experience Learning Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory

Reference 39

Resolution
metadata mismatch
local_arxiv, observed 2026-07-04T17:20:00.020604Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-06-25T23:49:38.932474Z digest=sha256:403f21ebb80563b35be8c19ccc1498f0d63c72b08fdf3a01a7278be19a0643e9

Observation 8ec8cff0-f92f-459e-8025-d5ac0a39c4b2 · inbound

The Past Is Prologue: A Plug-in Controller for Selective Updates in Sequentially Evolving LLM Memory cites this paper.

The Past Is Prologue: A Plug-in Controller for Selective Updates in Sequentially Evolving LLM Memory Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory

Reference 11

Resolution
metadata mismatch
local_arxiv, observed 2026-07-01T09:55:41.341437Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-07-01T06:00:11.022368Z digest=sha256:bf4f29d972277c0018f64441be8f31640210a7b5b9a097090b26e486b22f03ce

Observation e4564c04-6bfd-40cc-8caf-5df5fbee1fad · inbound

What Memory Do GUI Agents Really Need? From Passive Records to Active Task-Driving States cites this paper.

What Memory Do GUI Agents Really Need? From Passive Records to Active Task-Driving States Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory

Reference 64

Resolution
metadata mismatch
local_arxiv, observed 2026-07-01T10:35:41.276404Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-07-01T05:26:53.441833Z digest=sha256:5657ed19e54df69652d0e836aea7a80651ab6cf694b337bf0a3cc1dc83c8192e

Observation 657e666d-7df1-4c95-a97b-fc8a2eced198 · inbound

What Memory Do GUI Agents Really Need? From Passive Records to Active Task-Driving States cites this paper.

What Memory Do GUI Agents Really Need? From Passive Records to Active Task-Driving States Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory

Reference 64

Resolution
metadata mismatch
local_arxiv, observed 2026-07-03T22:08:58.667321Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-07-03T22:03:26.625209Z digest=sha256:1ee82f6c4e2869be8859597c16036f72b44db60faf73bbade8f16b13243fb3cf

Observation 252f20b3-6da8-46dd-b279-166d986c5e88 · inbound

EdgeBench: Unveiling Scaling Laws of Learning from Real-World Environments cites this paper.

EdgeBench: Unveiling Scaling Laws of Learning from Real-World Environments Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory

Reference 69

Resolution
unresolved
no resolver link, observed 2026-07-11T07:57:43.000834Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T07:57:43.000834Z digest=sha256:5d62ce9b1ad83ae357098e8c3b6353dfd11cf200c452a9db41031c0b0219e864

Observation f80d9411-59af-440a-8a82-80771dacb623 · inbound

ABot-AgentOS: A General Robotic Agent OS with Lifelong Multi-modal Memory cites this paper.

ABot-AgentOS: A General Robotic Agent OS with Lifelong Multi-modal Memory Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory

Reference 86

Resolution
unresolved
no resolver link, observed 2026-07-14T12:26:27.446079Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T12:26:27.446079Z digest=sha256:544810cb93896561468d6a23f85c2f68e7898883fcd99b6e6cd4d04ce8e49454

Observation 5fb71451-e324-4a06-aa00-80c198d4a3ea · inbound

ABot-AgentOS: A General Robotic Agent OS with Lifelong Multi-modal Memory cites this paper.

ABot-AgentOS: A General Robotic Agent OS with Lifelong Multi-modal Memory Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory

Reference 83

Resolution
unresolved
no resolver link, observed 2026-08-02T07:21:33.187454Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T07:21:33.187454Z digest=sha256:c07b4b37168224076692bb9e8d05c469afb90f56b51a8c88ccb51d780165389c

Observation 1a12b528-c9c4-43c6-8ab8-0e1627436779 · inbound

Agents Don't Just Agree, They Remember: Benchmarking Persistent Sycophancy in Stateful Personal Agents cites this paper.

Agents Don't Just Agree, They Remember: Benchmarking Persistent Sycophancy in Stateful Personal Agents Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory

Reference 35

Resolution
unresolved
no resolver link, observed 2026-07-14T11:04:44.592375Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-14T11:04:44.592375Z digest=sha256:7fde2b9ab25145b9ec6aaac82ddeb6c24fe793dd1e2810d2f1480153e7e217b9

Observation 483f00e8-5f31-4023-a97f-b7c8743336dc · inbound

Do Agents Dream of False Memories? Black-box Visual Attacks on Long-term Memory in Multimodal AI Agents cites this paper.

Do Agents Dream of False Memories? Black-box Visual Attacks on Long-term Memory in Multimodal AI Agents Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-01T22:44:00.259964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T22:44:00.259964Z digest=sha256:f97d7261572940e7c486ecd425ce780b473c13b4abfb850f3b3499369780575a

Observation 67342421-022f-45fd-859c-8120df1aba13 · inbound

Rethinking Self-Evolution: A Constrained Exploration-Exploitation Process for Mitigating Skill Overfitting cites this paper.

Rethinking Self-Evolution: A Constrained Exploration-Exploitation Process for Mitigating Skill Overfitting Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-01T11:46:02.563189Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:46:02.563189Z digest=sha256:eb096cde53ff0839dd0f9acf7603155a8fd9ea5811a82b5521d9983397f7d7dc

Observation 7e28cc05-454d-47e8-acc0-62ed074bc7e3 · inbound

RRM: Experience-Driven Reflective Retrieval Memory for Long-Horizon Multimodal Reasoning cites this paper.

RRM: Experience-Driven Reflective Retrieval Memory for Long-Horizon Multimodal Reasoning Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory

Reference 60

Resolution
unresolved
no resolver link, observed 2026-07-31T16:09:14.331637Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-31T16:09:14.331637Z digest=sha256:a1baa8d1c84b1b6282086a6aa7475de38ee3ce5f91bc3b223ce397d9c348a54c

Observation 19de0ee3-c2cd-4dac-84f9-1710b859a36c · inbound

AgentStream: How Well Do Self-Evolving LLM Agents Perform Under Streaming Tasks? cites this paper.

AgentStream: How Well Do Self-Evolving LLM Agents Perform Under Streaming Tasks? Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-04T01:12:32.201860Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:12:32.201860Z digest=sha256:53b60ec081fdd634c05bb866149c7cc7e1efa820a21d28b8402d5ccbfbdc7f75

Observation b3809490-a219-4107-ae23-10d6df1a983b · inbound

Benign Alone, Harmful Together: Exploiting Experience Composition in Self-Evolving LLM Agents cites this paper.

Benign Alone, Harmful Together: Exploiting Experience Composition in Self-Evolving LLM Agents Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-04T21:37:35.836340Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T21:37:35.836340Z digest=sha256:4b5cab562fd30ac0e067af97dbbf2fe0c6e3ee84dcf9e14f4cb1aaa9aa239aba