Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T20:08:12.087697Z
Paper Citation Record · LEDGER
As of 21 August 2026, this Paper Citation Record lists 49 of 49 outbound references and 9 inbound Pith citation observations for arXiv:2506.13356.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T20:08:12.087697Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-06T19:36:05.751022Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-21T02:29:25.233098Z
49 of 49 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 69f57222-94a4-4327-b27b-0aed18bfc6a9 · outbound
StoryBench: A Dynamic Benchmark for Evaluating Long-Term Memory with Multi Turns L-eval: Instituting standardized evaluation for long context language models
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 7ed84ab1-713b-4448-b30b-14f2e79ef3aa · outbound
StoryBench: A Dynamic Benchmark for Evaluating Long-Term Memory with Multi Turns Claude 3.5 sonnet model card addendum, 2024
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 1abb18d2-cf36-4687-989d-71b09d92895e · outbound
StoryBench: A Dynamic Benchmark for Evaluating Long-Term Memory with Multi Turns Longbench: A bilingual, multitask benchmark for long context understanding
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation e4301444-94b0-4798-8966-9f82e87aea12 · outbound
StoryBench: A Dynamic Benchmark for Evaluating Long-Term Memory with Multi Turns Opportunities and challenges for ai-based analysis of rwd in pharmaceutical r&d: A practical perspective
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 52f1139f-f545-4249-be2d-6553fd5ce5d3 · outbound
StoryBench: A Dynamic Benchmark for Evaluating Long-Term Memory with Multi Turns Longformer: The Long-Document Transformer
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 64c3555e-fdfc-4410-b5b2-7320115c2b16 · outbound
StoryBench: A Dynamic Benchmark for Evaluating Long-Term Memory with Multi Turns Doubao-1.5-pro, 2025
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 76bb980a-89b2-4566-ace6-fe53e7a76961 · outbound
StoryBench: A Dynamic Benchmark for Evaluating Long-Term Memory with Multi Turns Beyond prompts: Dynamic conversational benchmarking of large language models
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 75025580-5c20-40bd-a85a-598cf664da1f · outbound
StoryBench: A Dynamic Benchmark for Evaluating Long-Term Memory with Multi Turns A survey on evaluation of large language models
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 30825cda-7ffa-477b-9936-6531273d6041 · outbound
StoryBench: A Dynamic Benchmark for Evaluating Long-Term Memory with Multi Turns HSR-Enhanced Sparse Attention Acceleration
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c8f64c8d-fabf-4ddc-8e22-cc4ad1adbe3c · outbound
StoryBench: A Dynamic Benchmark for Evaluating Long-Term Memory with Multi Turns Llf-bench: Benchmark for interactive learning from language feedback
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 1d8dc7ee-df25-400d-871d-6d960ceaa94b · outbound
StoryBench: A Dynamic Benchmark for Evaluating Long-Term Memory with Multi Turns Mem0: Building Production-Ready AI Agents with Scalable Long-Term Memory
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b7030dd1-ff99-4c8f-8cc3-f9f7b72fb9a3 · outbound
StoryBench: A Dynamic Benchmark for Evaluating Long-Term Memory with Multi Turns Generating long sequences with sparse transformers
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation fbb4c577-d707-491e-9575-d4ee2ddda1cb · outbound
StoryBench: A Dynamic Benchmark for Evaluating Long-Term Memory with Multi Turns Transformer-XL: Attentive Language Models Beyond a Fixed-Length Context
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 20c3f5dd-f178-4685-9bb3-9d23f8578e9d · outbound
StoryBench: A Dynamic Benchmark for Evaluating Long-Term Memory with Multi Turns Flashattention: Fast and memory-efficient exact attention with io-awareness
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation b2c98317-7572-46b2-aef7-3745b43a8fa7 · outbound
StoryBench: A Dynamic Benchmark for Evaluating Long-Term Memory with Multi Turns DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9f1e6352-ea59-40ed-9f65-c7d6e524c7f0 · outbound
StoryBench: A Dynamic Benchmark for Evaluating Long-Term Memory with Multi Turns Cervical cancer segmentation based on full-scale feature fusion with cascading-attention and dilated convolution
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 038ec467-1243-4d5d-a0f7-844cf694a475 · outbound
StoryBench: A Dynamic Benchmark for Evaluating Long-Term Memory with Multi Turns BAMBOO : A comprehensive benchmark for evaluating long text modeling capacities of large language models
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 33be54da-4aa4-4852-a1f0-5f1157324d84 · outbound
StoryBench: A Dynamic Benchmark for Evaluating Long-Term Memory with Multi Turns Knowledge retention: 8 main strategies to improve it, February 2024
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 622a8289-afee-4a07-a1d1-4f732e5a7fdf · outbound
StoryBench: A Dynamic Benchmark for Evaluating Long-Term Memory with Multi Turns Mamba: Linear-time sequence modeling with selective state spaces
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0549d748-8eb2-48c3-8c70-8efd52c650c9 · outbound
StoryBench: A Dynamic Benchmark for Evaluating Long-Term Memory with Multi Turns MAPLE: A Mobile Agent with Persistent Finite State Machines for Structured Task Reasoning
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation f645c181-9bad-4323-9966-509ca3f0f567 · outbound
StoryBench: A Dynamic Benchmark for Evaluating Long-Term Memory with Multi Turns Hipporag: Neurobiologically inspired long-term memory for large language models
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 4811842e-b86c-4b43-ab18-0619868e0c43 · outbound
StoryBench: A Dynamic Benchmark for Evaluating Long-Term Memory with Multi Turns Layer-adaptive low-rank adaptation of large asr model for low-resource multilingual scenarios
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 517d0326-d456-4670-902e-84f9fccaddff · outbound
StoryBench: A Dynamic Benchmark for Evaluating Long-Term Memory with Multi Turns Synthetic Data in AI: Challenges, Applications, and Ethical Implications
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5b6804b3-1f29-4533-9e90-c7150b8a4141 · outbound
StoryBench: A Dynamic Benchmark for Evaluating Long-Term Memory with Multi Turns Ruler: What's the real context size of your long-context language models? CoRR, 2024
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 53b4a1ad-deb7-4bdb-9acd-c98d0990871b · outbound
StoryBench: A Dynamic Benchmark for Evaluating Long-Term Memory with Multi Turns Needle in a haystack - pressure testing llms, 2023
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 12c40333-58d1-4706-a49c-a65dd8ae1b6e · outbound
StoryBench: A Dynamic Benchmark for Evaluating Long-Term Memory with Multi Turns Reformer: The efficient transformer
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3f79d0d3-3e9a-48ed-b760-31b958a054c8 · outbound
StoryBench: A Dynamic Benchmark for Evaluating Long-Term Memory with Multi Turns Babilong: Testing the limits of llms with long context reasoning-in-a-haystack
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 6d22d922-771d-4387-a45c-df86270cba5c · outbound
StoryBench: A Dynamic Benchmark for Evaluating Long-Term Memory with Multi Turns Efficient memory management for large language model serving with pagedattention
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 880da2a9-324a-4cd8-9e0a-d6ce8c8139e8 · outbound
StoryBench: A Dynamic Benchmark for Evaluating Long-Term Memory with Multi Turns A Human-Inspired Reading Agent with Gist Memory of Very Long Contexts
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a33f96fb-ac34-4107-8291-a5f15609a4fb · outbound
StoryBench: A Dynamic Benchmark for Evaluating Long-Term Memory with Multi Turns Loogle: Can long-context language models understand long contexts? arXiv e-prints, pages arXiv--2311, 2023
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 3867c032-6a6e-4bd8-82b1-7d5e96f31564 · outbound
StoryBench: A Dynamic Benchmark for Evaluating Long-Term Memory with Multi Turns Ringattention with blockwise transformers for near-infinite context
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 942a7a22-6870-4fd7-8cb8-6a825ee71d37 · outbound
StoryBench: A Dynamic Benchmark for Evaluating Long-Term Memory with Multi Turns Best Practices and Lessons Learned on Synthetic Data
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fa3a05e5-cda4-461e-9e73-4deb49414c07 · outbound
StoryBench: A Dynamic Benchmark for Evaluating Long-Term Memory with Multi Turns Agentbench: Evaluating llms as agents
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation e87b24f4-fa0f-449a-aec4-298bc9af453a · outbound
StoryBench: A Dynamic Benchmark for Evaluating Long-Term Memory with Multi Turns Hello gpt-4o, 2024
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation df2a3246-c07b-4e3d-925f-43b818c3d8ab · outbound
StoryBench: A Dynamic Benchmark for Evaluating Long-Term Memory with Multi Turns Memgpt: Towards llms as operating systems
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 08278005-64cb-42c2-88bc-c5ecf46618e7 · outbound
StoryBench: A Dynamic Benchmark for Evaluating Long-Term Memory with Multi Turns Faster causal attention over large sequences through sparse flash attention
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation ce6c93a0-232e-4884-a3ae-0ee440bd6132 · outbound
StoryBench: A Dynamic Benchmark for Evaluating Long-Term Memory with Multi Turns Rwkv: Reinventing rnns for the transformer era
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 3f795acf-3c7c-48f9-a558-e40fa83addef · outbound
StoryBench: A Dynamic Benchmark for Evaluating Long-Term Memory with Multi Turns Do transformers need deep long-range memory? In Proceedings of the 58th Annual Meeting of the Association for Computational Linguistics, pages 7524--7529
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation ce696ee4-53fe-4da3-82a4-6c0ad90444a4 · outbound
StoryBench: A Dynamic Benchmark for Evaluating Long-Term Memory with Multi Turns Zeroscrolls: A zero-shot benchmark for long text understanding
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 3f7e7546-8de6-4ca4-8df8-851fdb57839f · outbound
StoryBench: A Dynamic Benchmark for Evaluating Long-Term Memory with Multi Turns Cognitive Memory in Large Language Models
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ba3fd16b-0a37-4b3e-a28b-d8457f571f12 · outbound
StoryBench: A Dynamic Benchmark for Evaluating Long-Term Memory with Multi Turns Chapterbreak: A challenge dataset for long-range language models
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 1235262d-f1f6-4c76-a6ec-553d096a8e14 · outbound
StoryBench: A Dynamic Benchmark for Evaluating Long-Term Memory with Multi Turns From Human Memory to AI Memory: A Survey on Memory Mechanisms in the Era of LLMs
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e8988212-e8e1-433f-80ae-323fbf6efeb9 · outbound
StoryBench: A Dynamic Benchmark for Evaluating Long-Term Memory with Multi Turns Data and System Perspectives of Sustainable Artificial Intelligence
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 198e470d-761b-4afa-b4b1-1a2c2b384e8e · outbound
StoryBench: A Dynamic Benchmark for Evaluating Long-Term Memory with Multi Turns Long time no see! open-domain conversation with long-term persona memory
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 8a92f438-4eb9-4409-9cf8-14a5e02d4c28 · outbound
StoryBench: A Dynamic Benchmark for Evaluating Long-Term Memory with Multi Turns Memoryscope, 09 2024
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 8e3d5258-fa73-4fea-a4b5-b8d4143315d5 · outbound
StoryBench: A Dynamic Benchmark for Evaluating Long-Term Memory with Multi Turns A survey on multi-turn interaction capabilities of large language models
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation efe79906-4df8-4582-8fc7-89ca9fefdd50 · outbound
StoryBench: A Dynamic Benchmark for Evaluating Long-Term Memory with Multi Turns bench: Extending long context evaluation beyond 100k tokens
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 7b4c4dbd-2ba5-4df7-81b8-e95e555afd8c · outbound
StoryBench: A Dynamic Benchmark for Evaluating Long-Term Memory with Multi Turns Memorybank: Enhancing large language models with long-term memory
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation e16ef2a4-d38e-4a14-a8d8-322a289bf299 · outbound
StoryBench: A Dynamic Benchmark for Evaluating Long-Term Memory with Multi Turns Webarena: A realistic web environment for building autonomous agents
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation fc514041-df33-48a7-8dc8-e135856eddf9 · inbound
Evaluating Memory in LLM Agents via Incremental Multi-Turn Interactions StoryBench: A Dynamic Benchmark for Evaluating Long-Term Memory with Multi Turns
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 99af9aae-2f5e-47a5-a31e-b252d9768f55 · inbound
Evaluating Memory in LLM Agents via Incremental Multi-Turn Interactions StoryBench: A Dynamic Benchmark for Evaluating Long-Term Memory with Multi Turns
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f28ed759-8378-4802-89c6-4dda3ee488f0 · inbound
A Survey of Self-Evolving Agents: What, When, How, and Where to Evolve on the Path to Artificial Super Intelligence StoryBench: A Dynamic Benchmark for Evaluating Long-Term Memory with Multi Turns
Reference 97
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 9f37cf52-1f65-4a4c-b070-763b30f745ef · inbound
Evo-Memory: Benchmarking LLM Agent Test-time Learning with Self-Evolving Memory StoryBench: A Dynamic Benchmark for Evaluating Long-Term Memory with Multi Turns
Reference 133
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation b1c0d3c9-3af2-4d47-92cc-23664cac1401 · inbound
Cost and Accuracy of Long-Term Memory in Distributed Multi-Agent Systems Based on Large Language Models StoryBench: A Dynamic Benchmark for Evaluating Long-Term Memory with Multi Turns
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7448e09e-0062-44f7-86b9-ba38d68b59c2 · inbound
Toward Efficient Agents: Memory, Tool learning, and Planning StoryBench: A Dynamic Benchmark for Evaluating Long-Term Memory with Multi Turns
Reference 122
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eb03873c-da9d-4f31-93b8-ae173f8e4696 · inbound
From Storage to Experience: A Survey on the Evolution of LLM Agent Memory Mechanisms StoryBench: A Dynamic Benchmark for Evaluating Long-Term Memory with Multi Turns
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 532bf676-1e9a-4949-b3dd-e4ac8cf8a4b3 · inbound
EgoMemReason: A Memory-Driven Reasoning Benchmark for Long-Horizon Egocentric Video Understanding StoryBench: A Dynamic Benchmark for Evaluating Long-Term Memory with Multi Turns
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation daa066a6-8d23-49d6-95ea-e2a59ac9b31d · inbound
MemConflict: Evaluating Long-Term Memory Systems Under Memory Conflicts StoryBench: A Dynamic Benchmark for Evaluating Long-Term Memory with Multi Turns
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.