Pith. sign in

Paper Citation Record · LEDGER

MemoryBench: A Benchmark for Memory and Continual Learning in LLM Systems

As of 20 August 2026, this Paper Citation Record lists 22 of 22 outbound references and 27 inbound Pith citation observations for arXiv:2510.17281.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2510.17281 v7

Coverage vector

measured 22 of 22 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-04T09:08:01.244788Z

measured 49 of 49 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 27 of 27 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-11T23:27:32.610040Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-04T00:19:13.724792Z

Reference resolution

22 of 22 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved20
  • parse uncertain0
  • malformed identifier2
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 3ab0e553-cd81-4e74-9161-0297794beb04 · outbound

This paper cites If the informa- tion in the dialog history is insufficient to answer the question, you must admit that you don’t know the answer.

MemoryBench: A Benchmark for Memory and Continual Learning in LLM Systems If the informa- tion in the dialog history is insufficient to answer the question, you must admit that you don’t know the answer

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-04T09:07:59.514593Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:07:59.514593Z digest=sha256:a0ff6f6afb9b8f866ebe7105ab057e5363e31c26baed20216da419c8264b7073

Observation e5e2fc16-6b21-4907-84ab-fc2e4acdcc36 · outbound

This paper cites real” ideas in the target paper under various quality metrics. The primary metric is the “Insight Score,.

MemoryBench: A Benchmark for Memory and Continual Learning in LLM Systems real” ideas in the target paper under various quality metrics. The primary metric is the “Insight Score,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-04T09:07:59.617573Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:07:59.617573Z digest=sha256:b2fb21cf73b665dd3818cffb9a707b2ab413d4a154d3a5fd65f532cb2c0cf8d1

Observation 7269188d-64fd-49db-993d-448bbc843fbb · outbound

This paper cites - 1: Represents extremely poor quality (e.g., completely irrelevant, factually incorrect, non- sensical).

MemoryBench: A Benchmark for Memory and Continual Learning in LLM Systems - 1: Represents extremely poor quality (e.g., completely irrelevant, factually incorrect, non- sensical)

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-04T09:07:59.894635Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:07:59.894635Z digest=sha256:3d54d6c228df649004c37b9076a4f49e8258a677a1043ade264cd8284d154b8b

Observation f61a0d9e-936c-4fd1-a541-dfb43f5ca8dd · outbound

This paper cites Scores range from 0.00 to 1.00.

MemoryBench: A Benchmark for Memory and Continual Learning in LLM Systems Scores range from 0.00 to 1.00

Reference 4

Resolution
malformed identifier
no resolver link, observed 2026-08-04T09:08:00.300405Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:08:00.300405Z digest=sha256:543a3da4a23d8fb339ddf5072c4861257360ed0c562bc3eafa6dcc13166079e2

Observation e05a24e4-61e9-4e65-807b-2a9060382a20 · outbound

This paper cites an unresolved cited work.

MemoryBench: A Benchmark for Memory and Continual Learning in LLM Systems Unresolved cited work

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-04T09:07:59.707160Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:07:59.707160Z digest=sha256:6c1cda1ffc850e4c1e3c8156311abdc59113c770f1cf800bd779440f4705943c

Observation 8d5e4796-d136-446a-a46c-93cb8fb07fa7 · outbound

This paper cites an unresolved cited work.

MemoryBench: A Benchmark for Memory and Continual Learning in LLM Systems Unresolved cited work

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-04T09:07:59.808285Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:07:59.808285Z digest=sha256:c9678f5c889cc2ea1700adb354f062b80780b276e4a229a536cfc4078e13e4bd

Observation 776d9a2f-c01f-4bab-a381-0f11a9a9bdef · outbound

This paper cites Scores range from 0.00 (no similarity) to 1.00 (perfect semantic match).

MemoryBench: A Benchmark for Memory and Continual Learning in LLM Systems Scores range from 0.00 (no similarity) to 1.00 (perfect semantic match)

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-04T09:08:00.023535Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:08:00.023535Z digest=sha256:d63540f293546096f23e04c1b97bd451c78bde33f46d8a1875877843640a1f4b

Observation 418f9f05-fa4b-41e6-8f8d-b7cea85b41b3 · outbound

This paper cites Scores range from 1 (minimal overlap) to 10 (perfect overlap).

MemoryBench: A Benchmark for Memory and Continual Learning in LLM Systems Scores range from 1 (minimal overlap) to 10 (perfect overlap)

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-04T09:08:00.107843Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:08:00.107843Z digest=sha256:e05e397f1aac85ae4c44c0ded4c5e3e8f5c757bc279993f451899e177e77137e

Observation 5ba3e9e2-b4d5-45f3-aeee-58be5bd73345 · outbound

This paper cites This score is derived by ranking the generated idea(s) against the ground truth idea.

MemoryBench: A Benchmark for Memory and Continual Learning in LLM Systems This score is derived by ranking the generated idea(s) against the ground truth idea

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-04T09:08:00.206045Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:08:00.206045Z digest=sha256:a18016041e18dff52021ab95dac2d12c84d77593f7adc279c8b8928bb8149d32

Observation 3ccd316e-82ed-49b4-87bc-0748074fe99d · outbound

This paper cites CRITICAL: Focus on the initial request as the core topic that should be the primary focus throughout this entire conversation.

MemoryBench: A Benchmark for Memory and Continual Learning in LLM Systems CRITICAL: Focus on the initial request as the core topic that should be the primary focus throughout this entire conversation

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-04T09:08:00.541394Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:08:00.541394Z digest=sha256:8840fe30b25e064e20d49abeb2cec6768585e7489868b03d3f3a8c8d5cbdeae4

Observation f5d449fa-550f-44ad-889a-186f564099bb · outbound

This paper cites an unresolved cited work.

MemoryBench: A Benchmark for Memory and Continual Learning in LLM Systems Unresolved cited work

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-04T09:08:00.613993Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:08:00.613993Z digest=sha256:32aa0c4f1a8b86d5ddf9c5c0c2ba514c8b83f4ee6d42e2400a8641613f356f84

Observation af6a424f-ad9c-4c8c-8984-762f4af88499 · outbound

This paper cites an unresolved cited work.

MemoryBench: A Benchmark for Memory and Continual Learning in LLM Systems Unresolved cited work

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-04T09:08:00.678337Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:08:00.678337Z digest=sha256:ac4ffaebd1a6613a9cd46b08862241103d808d3241894ad1447f4e5a01700abe

Observation ece8b3f7-f331-4b2d-b2ed-dc8d46760a65 · outbound

This paper cites Total” denotes the full size of each dataset, while “Samples.

MemoryBench: A Benchmark for Memory and Continual Learning in LLM Systems Total” denotes the full size of each dataset, while “Samples

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-04T09:08:00.733032Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:08:00.733032Z digest=sha256:4cfbf6020a20cb45764105e59fc16bdf95952703f4308dad37598ffaad1c4433

Observation e8afffa1-e982-484e-84f4-d9bf5e42c5d9 · outbound

This paper cites an unresolved cited work.

MemoryBench: A Benchmark for Memory and Continual Learning in LLM Systems Unresolved cited work

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-04T09:08:00.793545Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:08:00.793545Z digest=sha256:08674286ea7f98b4ce68a9a5880bf7de9f8f8eb88d087bdff138593c43e95b41

Observation 91af7cb9-3f44-4355-bb62-079f8502474f · outbound

This paper cites (b) Incorporate these 100 training dialogues into the memory of LLMsys, so that it accu- mulates more experience.

MemoryBench: A Benchmark for Memory and Continual Learning in LLM Systems (b) Incorporate these 100 training dialogues into the memory of LLMsys, so that it accu- mulates more experience

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-04T09:08:00.892827Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:08:00.892827Z digest=sha256:1af1b0b993f9b9a481d54f597dfaaf12cb53af8beedf801b7c2cd79efdba4612

Observation 4d997c0c-a09d-4634-baaa-e20f46d91071 · outbound

This paper cites an unresolved cited work.

MemoryBench: A Benchmark for Memory and Continual Learning in LLM Systems Unresolved cited work

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-04T09:08:00.990195Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:08:00.990195Z digest=sha256:17a0aab105bb0421315ec5f663d3750cd3d1338ffd79e275697a066289d01c10

Observation 8a8bd273-0e6f-4eae-9976-eb45d2e451d6 · outbound

This paper cites an unresolved cited work.

MemoryBench: A Benchmark for Memory and Continual Learning in LLM Systems Unresolved cited work

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-04T09:08:01.094498Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:08:01.094498Z digest=sha256:d9f148983ad40ab6c40e1d5d5b03e4fe0d43d6518a87cadaf56dd8b7e1f48cd0

Observation baf3ba51-ee92-4054-9a64-3178264ba27c · outbound

This paper cites Speaker”, “Assistant.

MemoryBench: A Benchmark for Memory and Continual Learning in LLM Systems Speaker”, “Assistant

Reference 23

Resolution
malformed identifier
no resolver link, observed 2026-08-04T09:08:01.172466Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:08:01.172466Z digest=sha256:a0aafedc4a63710b6757ca5869dc750afd93bec072073237c492d3f8b31ca80c

Observation 80a617d5-788c-45cb-b9bb-c5e37a4e66fe · outbound

This paper cites Analyzed a case involving a judge (Wang) accused of bribery and malfea- sance to determine if there were ju- risdictional errors.

MemoryBench: A Benchmark for Memory and Continual Learning in LLM Systems Analyzed a case involving a judge (Wang) accused of bribery and malfea- sance to determine if there were ju- risdictional errors

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-04T09:08:01.244788Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:08:01.244788Z digest=sha256:f80a941066c06dcc0a85b35c400b0cf9c49cb4e8eb5dbed84549552fbd561980

Observation 8fb321e7-b7f9-4cd8-b9f5-b1aeac79b8db · outbound

This paper cites URLhttps://aclanthology.org/ W05-0909/.

MemoryBench: A Benchmark for Memory and Continual Learning in LLM Systems URLhttps://aclanthology.org/ W05-0909/

Reference 2005

Resolution
unresolved
no resolver link, observed 2026-08-04T09:07:59.232300Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:07:59.232300Z digest=sha256:41ff3d12f186f5a9b7dbf1ebe7917795cf52d8b2e8964218d25aaa8bc31d424f

Observation 9866618a-8e7d-4905-a58d-5bb6e88a0406 · outbound

This paper cites Inference Scaling Laws: An Empirical Analysis of Compute-Optimal Inference for Problem-Solving with Language Models.

MemoryBench: A Benchmark for Memory and Continual Learning in LLM Systems Inference Scaling Laws: An Empirical Analysis of Compute-Optimal Inference for Problem-Solving with Language Models

Reference 2006

Resolution
unresolved
no resolver link, observed 2026-08-04T09:07:59.456340Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:07:59.456340Z digest=sha256:a969c2631d72050f7e4d728a79466e7ddbdd29f8bdd9faad35710cb067510f0e

Observation eecf06e1-6d16-4e1f-a80c-9b0c3d1587a9 · outbound

This paper cites HelloBench: Evaluating Long Text Generation Capabilities of Large Language Models.

MemoryBench: A Benchmark for Memory and Continual Learning in LLM Systems HelloBench: Evaluating Long Text Generation Capabilities of Large Language Models

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-04T09:07:59.384379Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T09:07:59.384379Z digest=sha256:b438d16ad989452ae9f0612d0ba0c72bda8bafe3be998ba0e85fa508f07d7b34

Pith citing papers

Observation 9596df84-ce6e-417e-b647-557d28bf7026 · inbound

Improve Large Language Model Systems with User Logs cites this paper.

Improve Large Language Model Systems with User Logs MemoryBench: A Benchmark for Memory and Continual Learning in LLM Systems

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-05-21T14:34:12.942233Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-21T14:34:01.332088Z digest=sha256:74b4ba3dc0f4d5ab249a28c75a125710646e5fa5c9b29250930baf234b890ae8

Observation 96d559e0-c9f1-41c5-b1dc-db7e0077cb5d · inbound

Improve Large Language Model Systems with User Logs cites this paper.

Improve Large Language Model Systems with User Logs MemoryBench: A Benchmark for Memory and Continual Learning in LLM Systems

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-03T04:00:16.817389Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T04:00:16.817389Z digest=sha256:9a4924c9eb668aca6f8dbb1354739af899f0786d7881e7031a4e16e1c1898448

Observation 8b7d41ae-baf1-44a3-b0f9-25f6dbd3375d · inbound

PERMA: Benchmarking Personalized Memory Agents via Event-Driven Preference and Realistic Task Environments cites this paper.

PERMA: Benchmarking Personalized Memory Agents via Event-Driven Preference and Realistic Task Environments MemoryBench: A Benchmark for Memory and Continual Learning in LLM Systems

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-21T09:59:58.927980Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-21T09:55:17.236296Z digest=sha256:7b2fe50d9d5ee38efeffd7230694a98bc22987b6cb4386edddad56de9670ebc8

Observation d47c3893-1eaa-4d88-b5ad-fd667ba43488 · inbound

ATANT: An Evaluation Framework for AI Continuity cites this paper.

ATANT: An Evaluation Framework for AI Continuity MemoryBench: A Benchmark for Memory and Continual Learning in LLM Systems

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-11T00:20:55.266228Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-10T18:34:10.081438Z digest=sha256:33c7a796e0c18eecd10607840f468a4f5482490ff779a0a2227b1e15c9afff5f

Observation bd00da51-6abe-41db-bbdc-e98d3e6d9d2f · inbound

Trust Your Memory: Verifiable Control of Smart Homes through Reinforcement Learning with Multi-dimensional Rewards cites this paper.

Trust Your Memory: Verifiable Control of Smart Homes through Reinforcement Learning with Multi-dimensional Rewards MemoryBench: A Benchmark for Memory and Continual Learning in LLM Systems

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-11T08:06:01.151170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-10T16:49:00.343580Z digest=sha256:52c975e47083f968cff1276ba059b5015c696f1a8022bd0a3ac5bd6e92069d3d

Observation 46d62ddd-1b97-4fe1-bff7-8f03e9762578 · inbound

ATANT v1.1: Positioning Continuity Evaluation Against Memory, Long-Context, and Agentic-Memory Benchmarks cites this paper.

ATANT v1.1: Positioning Continuity Evaluation Against Memory, Long-Context, and Agentic-Memory Benchmarks MemoryBench: A Benchmark for Memory and Continual Learning in LLM Systems

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-11T09:06:00.716526Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-10T16:14:25.076218Z digest=sha256:a3d450b60217496be62521c18c00a323e81eb4a77086776ffc0debf432414318

Observation ba428ddb-fce7-4554-9d8c-a39a341b042c · inbound

Skill Retrieval Augmentation for Agentic AI cites this paper.

Skill Retrieval Augmentation for Agentic AI MemoryBench: A Benchmark for Memory and Continual Learning in LLM Systems

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-11T21:56:20.785231Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-08T03:44:33.631278Z digest=sha256:f40aab3b5154a5c789df40ac715d27eb6ab6d1fcd977a79fecf5ea240029ace8

Observation 8e2c0a8d-e7cd-4da8-ba60-0737e6fd5809 · inbound

Skill Retrieval Augmentation for Agentic AI cites this paper.

Skill Retrieval Augmentation for Agentic AI MemoryBench: A Benchmark for Memory and Continual Learning in LLM Systems

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-25T06:50:28.534630Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-25T06:46:49.931555Z digest=sha256:e2efdd4bd9680d1cc95936405b5784b9cdb4b0f445942d90a451e79fd33c3201

Observation ffadea85-a004-48f9-a80d-14ff147a72fd · inbound

Skill Retrieval Augmentation for Agentic AI cites this paper.

Skill Retrieval Augmentation for Agentic AI MemoryBench: A Benchmark for Memory and Continual Learning in LLM Systems

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-07-01T08:55:34.833486Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-07-01T08:52:34.254481Z digest=sha256:96cf094d01b8fd732d041e4d620dee93e3b8f3f3fc43cd75e55f2a4690720573

Observation 722587ed-42aa-40f3-9eef-486cf0f79394 · inbound

GroupMemBench: Benchmarking LLM Agent Memory in Multi-Party Conversations cites this paper.

GroupMemBench: Benchmarking LLM Agent Memory in Multi-Party Conversations MemoryBench: A Benchmark for Memory and Continual Learning in LLM Systems

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-05-15T01:58:29.254668Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-15T01:54:22.009219Z digest=sha256:198ecd4a9ac9259facd07d058b07de47a789c1cc5ab0aeba2025a44e57f169e2

Observation b84ddc13-2c94-4e5c-882f-7d091fd7c461 · inbound

GroupMemBench: Benchmarking LLM Agent Memory in Multi-Party Conversations cites this paper.

GroupMemBench: Benchmarking LLM Agent Memory in Multi-Party Conversations MemoryBench: A Benchmark for Memory and Continual Learning in LLM Systems

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-05-20T21:33:46.303938Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-20T21:31:45.854294Z digest=sha256:ec8e031db33e7e23f7b906cf9e0eee5b8dbcd485a10b72eecf2d57838df244c6

Observation fa4fe947-1753-4d1f-98e0-59a9ed946ff3 · inbound

State Contamination in Memory-Augmented LLM Agents cites this paper.

State Contamination in Memory-Augmented LLM Agents MemoryBench: A Benchmark for Memory and Continual Learning in LLM Systems

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-19T21:32:48.006718Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-19T21:28:00.370745Z digest=sha256:5a9cd137531f3577c6b6c949dd065c3419b82bbef2be1b30b04d12524c9b90b6

Observation be9ee6ec-1f07-404f-b580-e59662ca58cc · inbound

EXG: Self-Evolving Agents with Experience Graphs cites this paper.

EXG: Self-Evolving Agents with Experience Graphs MemoryBench: A Benchmark for Memory and Continual Learning in LLM Systems

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-19T22:17:49.344693Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-19T22:17:10.728289Z digest=sha256:75c3a241baa986ef30a2d180dd3605c6ac8c2c9cfb65c0e059d13931cc657c5b

Observation 391cad38-b801-40ad-82d1-72f0b7f68c57 · inbound

EvoMemBench: Benchmarking Agent Memory from a Self-Evolving Perspective cites this paper.

EvoMemBench: Benchmarking Agent Memory from a Self-Evolving Perspective MemoryBench: A Benchmark for Memory and Continual Learning in LLM Systems

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-20T11:38:14.607007Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-20T11:35:23.892275Z digest=sha256:1ae30b34af85ace702f8497e0ae8fbfd49013771dab0739348c5cf58dc155db9

Observation a89c373e-0454-4fa7-a388-af4c88a42362 · inbound

MemGym: a Long-Horizon Memory Environment for LLM Agents cites this paper.

MemGym: a Long-Horizon Memory Environment for LLM Agents MemoryBench: A Benchmark for Memory and Continual Learning in LLM Systems

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-21T04:59:36.375306Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-21T04:58:41.043782Z digest=sha256:96f931bf18786d48a42cb8e397236befb27d66c27bc26384b52644f0db234d2a

Observation 85aaf9a6-3d00-4123-845e-0578b0d44422 · inbound

Connecting the Dots: Benchmarking Reflective Memory in Long-Horizon Dialogue cites this paper.

Connecting the Dots: Benchmarking Reflective Memory in Long-Horizon Dialogue MemoryBench: A Benchmark for Memory and Continual Learning in LLM Systems

Reference 61

Resolution
metadata mismatch
local_arxiv, observed 2026-06-28T17:42:25.972078Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-06-28T17:41:06.216002Z digest=sha256:41c1637d6fe1170c19d7c667d90bc1495c087e72907eb5c5ea4551caefd5d14e

Observation 5d92bc1f-b907-4f0f-a4e2-db14557e5578 · inbound

AgentCL: Toward Rigorous Evaluation of Continual Learning in Language Agents cites this paper.

AgentCL: Toward Rigorous Evaluation of Continual Learning in Language Agents MemoryBench: A Benchmark for Memory and Continual Learning in LLM Systems

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-07-01T22:56:20.129431Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-28T14:52:49.244423Z digest=sha256:23d0e293668521d99b7552939f0b2a1c3f1347dc5b446096409dafaf55bf6825

Observation eebbd52f-e63d-4415-a886-684646ab24db · inbound

Exploring Cross-Scenario Generality of Agentic Memory Systems: Diagnostics and a Strong Baseline cites this paper.

Exploring Cross-Scenario Generality of Agentic Memory Systems: Diagnostics and a Strong Baseline MemoryBench: A Benchmark for Memory and Continual Learning in LLM Systems

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-07-02T07:36:45.398625Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-06-28T06:49:08.689588Z digest=sha256:9cbf8e205cf1612626073d38a8dbd4884cb45771917d15a484cbc0d98122aff5

Observation fe63c1bb-17cd-433f-813f-7f4ca9bdd3c5 · inbound

Continual Learning Bench: Evaluating Frontier AI Systems in Real-World Stateful Environments cites this paper.

Continual Learning Bench: Evaluating Frontier AI Systems in Real-World Stateful Environments MemoryBench: A Benchmark for Memory and Continual Learning in LLM Systems

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-07-02T12:56:56.860184Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-28T01:45:28.693098Z digest=sha256:25ee58ddf730d9d6c2326186c3fb7a8a7df6dfebf5e14f0321df20db2e485a5e

Observation 6621d3d6-2753-4515-b514-fc7ac96c1573 · inbound

MemTrace: Probing What Final Accuracy Misses in Long-Term Memory cites this paper.

MemTrace: Probing What Final Accuracy Misses in Long-Term Memory MemoryBench: A Benchmark for Memory and Continual Learning in LLM Systems

Reference 17

Resolution
metadata mismatch
local_arxiv, observed 2026-07-03T18:18:49.831306Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-06-27T03:13:52.804489Z digest=sha256:2b2c3002302c0c3a0af82fe1de5879df47e7e0e8909d9717d229fadfd6e1dc31

Observation 6f46cf93-cdcb-45c5-bff3-b4c4e5d311d1 · inbound

Beyond the Current Observation: Evaluating Multimodal Large Language Models in Controllable Non-Markov Games cites this paper.

Beyond the Current Observation: Evaluating Multimodal Large Language Models in Controllable Non-Markov Games MemoryBench: A Benchmark for Memory and Continual Learning in LLM Systems

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-07-04T00:19:13.726063Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-26T21:17:02.332687Z digest=sha256:a4740b9c77f9715333f2108a6788c051599f9dc782c6e68e845feb962ea57ee5

Observation 08e77bbe-d1f5-4005-a2c7-be3679993ef2 · inbound

MemDelta: Controlled Baselines and Hidden Confounds in Agent Memory Evaluation cites this paper.

MemDelta: Controlled Baselines and Hidden Confounds in Agent Memory Evaluation MemoryBench: A Benchmark for Memory and Continual Learning in LLM Systems

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-06-30T06:54:21.053652Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-06-30T06:45:58.680234Z digest=sha256:27ef4a9d783e2de6ea1d03c6ccc838b7bb071f29a7108bc1f5e7f3834175456b

Observation 88f60077-19fa-47e0-a6a7-30cd47874937 · inbound

When Classic Cache Policies Fail: Learning-Augmented Replacement for Semantic Retrieval Buffers cites this paper.

When Classic Cache Policies Fail: Learning-Augmented Replacement for Semantic Retrieval Buffers MemoryBench: A Benchmark for Memory and Continual Learning in LLM Systems

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-07-02T05:06:38.912256Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-07-02T05:00:49.161874Z digest=sha256:1b5a5c0d2b5d3e33733285c26f1517ad9f999efba489a4a1a22643b3fce0aeac

Observation fb077d9e-5aeb-4824-888b-2c7bd07e174c · inbound

MemSyco-Bench: Benchmarking Sycophancy in Agent Memory cites this paper.

MemSyco-Bench: Benchmarking Sycophancy in Agent Memory MemoryBench: A Benchmark for Memory and Continual Learning in LLM Systems

Reference 61

Resolution
metadata mismatch
local_arxiv, observed 2026-07-02T06:36:43.281539Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-07-02T06:29:32.975530Z digest=sha256:12757eb5c13edd373d28a4baf6a9517b2e52653c72c341631019f1887813d7bc

Observation 6348fcec-a079-42fb-a974-18f9838af2b0 · inbound

MemSyco-Bench: Benchmarking Sycophancy in Agent Memory cites this paper.

MemSyco-Bench: Benchmarking Sycophancy in Agent Memory MemoryBench: A Benchmark for Memory and Continual Learning in LLM Systems

Reference 61

Resolution
metadata mismatch
local_arxiv, observed 2026-07-03T18:58:50.794862Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-07-03T18:50:02.455542Z digest=sha256:4d9a80055fd00fd19ff4f3ac0edacdc9583854569788b84cd38d65b0234e1e1c

Observation 88b1790e-7c99-40fa-a31b-bf8d4ecbd0bd · inbound

ContextWeave: A Real-World Workflow Benchmark cites this paper.

ContextWeave: A Real-World Workflow Benchmark MemoryBench: A Benchmark for Memory and Continual Learning in LLM Systems

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T15:38:14.268235Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:38:14.268235Z digest=sha256:887bcb75850107a54b27c571b3453ea16bb7dbb0f8fb95f7e1a3a6e9c13d866f

Observation 4fcf339f-deca-450b-9596-1f0e2439ffc2 · inbound

Different Feedback, Different Updates: Selective Self-Learning from User Interactions for Large Language Models cites this paper.

Different Feedback, Different Updates: Selective Self-Learning from User Interactions for Large Language Models MemoryBench: A Benchmark for Memory and Continual Learning in LLM Systems

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-11T23:27:32.610040Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:27:32.610040Z digest=sha256:f4213c825b1d21d725fd9a6efc9132a272d75d0b99e1ebfc8e61007e05a72fee