Pith. sign in

Paper Citation Record · LEDGER

Memory Reward Inflation in Self-Improving LLM Agents

As of 23 August 2026, this Paper Citation Record lists 37 of 37 outbound references and 0 inbound Pith citation observations for arXiv:2608.00017.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.00017 v1

Coverage vector

measured 37 of 37 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-04T02:16:43.153818Z

measured 37 of 37 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-23T06:30:58.430688+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

37 of 37 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved37
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation fd057453-96b2-4df1-a1ab-376a65485ac9 · outbound

This paper cites Sutton and Andrew G.

Memory Reward Inflation in Self-Improving LLM Agents Sutton and Andrew G

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-04T02:16:40.747999Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T02:16:40.747999Z digest=sha256:4e76dcd44532b350d07fb88dca48e686c212ed85d8034beca57f39c821c7232a

Observation 447acec5-3a1d-4f99-958d-76c0f90fe96f · outbound

This paper cites Reflexion: Language agents with verbal reinforcement learning.

Memory Reward Inflation in Self-Improving LLM Agents Reflexion: Language agents with verbal reinforcement learning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-04T02:16:40.856553Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T02:16:40.856553Z digest=sha256:b5f50c18ea4a21d9ecc6f931e356d54cf93e4272362455e845da4846207bd68b

Observation a6f4abcf-f05e-4d6c-bce7-1f0790d49458 · outbound

This paper cites Self-refine: Iterative refinement with self-feedback.

Memory Reward Inflation in Self-Improving LLM Agents Self-refine: Iterative refinement with self-feedback

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-04T02:16:40.977158Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T02:16:40.977158Z digest=sha256:e848564aa492fcef906db74e3a991a6913951f37da97ee67ab55451ac164f671

Observation 3eae2184-c78d-4153-b98b-ae7922e7dc3b · outbound

This paper cites Memento: Fine-tuning LLM Agents without Fine-tuning LLMs.

Memory Reward Inflation in Self-Improving LLM Agents Memento: Fine-tuning LLM Agents without Fine-tuning LLMs

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-04T02:16:41.032797Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T02:16:41.032797Z digest=sha256:fddd67b3b3961cf40b920602b4540e46c04631adef01c073bf1307f4e68dcf33

Observation 7930bc81-c268-43a0-9e0b-f5d687120db2 · outbound

This paper cites Christiano, Jan Leike, Tom B.

Memory Reward Inflation in Self-Improving LLM Agents Christiano, Jan Leike, Tom B

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-04T02:16:41.161168Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T02:16:41.161168Z digest=sha256:5669f5b2ce3c4681d436b6ce1642886a6a54feedafa3314d35c2abd32252f761

Observation 5786c519-f5b3-41e4-842d-710903a365b2 · outbound

This paper cites an unresolved cited work.

Memory Reward Inflation in Self-Improving LLM Agents Unresolved cited work

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-04T02:16:41.274957Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T02:16:41.274957Z digest=sha256:337e280b87d558d7c74a1afbdd4f4653c48db08260cc8bba5c494a053cdd85d6

Observation d6c6b4a9-8649-4919-83f0-a8c63062c1aa · outbound

This paper cites Judging LLM-as-a-Judge with MT-Bench and Chatbot Arena.

Memory Reward Inflation in Self-Improving LLM Agents Judging LLM-as-a-Judge with MT-Bench and Chatbot Arena

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-04T02:16:41.355262Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T02:16:41.355262Z digest=sha256:efca1332da157158846352ab5b0422a33237e2fb27c1686cfdef753a97065e6d

Observation 8e747de1-a639-4923-9c64-bc7ae891cec7 · outbound

This paper cites Spontaneous Reward Hacking in Iterative Self-Refinement.

Memory Reward Inflation in Self-Improving LLM Agents Spontaneous Reward Hacking in Iterative Self-Refinement

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-04T02:16:41.453612Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T02:16:41.453612Z digest=sha256:67d435f338f3ffa96ac837a3347c2d819ffe7ce2cef2d139fc78f65a937d5190

Observation 35b6a77f-a2ce-47d1-bcad-4aa26abeec89 · outbound

This paper cites ReAct: Synergizing reasoning and acting in language models.

Memory Reward Inflation in Self-Improving LLM Agents ReAct: Synergizing reasoning and acting in language models

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-04T02:16:41.531771Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T02:16:41.531771Z digest=sha256:b637d7233c3a2a98ab137bdcb41594f1595de405c5e7955494a457f563e89b8a

Observation efce4652-528d-4fbf-b651-7b5ec098f964 · outbound

This paper cites Le, and Denny Zhou.

Memory Reward Inflation in Self-Improving LLM Agents Le, and Denny Zhou

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-04T02:16:41.585883Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T02:16:41.585883Z digest=sha256:e022ee3da0d80b8b2bf3a35bc12a029fc0def4c4b1fc5308458cc4d7b272811b

Observation c9f74aae-e796-4c1a-a602-ca878731518f · outbound

This paper cites Retrieval-augmented generation for knowledge-intensive NLP tasks.

Memory Reward Inflation in Self-Improving LLM Agents Retrieval-augmented generation for knowledge-intensive NLP tasks

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-04T02:16:41.640511Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T02:16:41.640511Z digest=sha256:13f83c6ac0720505adaf5ef21849de88bc656aef91f8a84575fcab164d7c0d15

Observation e6f0ec12-58c4-472a-acd2-36a4f878530d · outbound

This paper cites O’Brien, Carrie J.

Memory Reward Inflation in Self-Improving LLM Agents O’Brien, Carrie J

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-04T02:16:41.697684Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T02:16:41.697684Z digest=sha256:428ff0b47336467ca414dc5b850d900985ad4f0d121acb56177fbcbb67726adb

Observation 6f525c53-dfb7-4ca9-bea7-94f9b8ef7b97 · outbound

This paper cites ReasoningBank: Scaling Agent Self-Evolving with Reasoning Memory.

Memory Reward Inflation in Self-Improving LLM Agents ReasoningBank: Scaling Agent Self-Evolving with Reasoning Memory

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-04T02:16:41.810687Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T02:16:41.810687Z digest=sha256:3eff94ba974a2eb008971b95197ca9f04d68ca0fafe1b8c8ef1eaa11a475b17d

Observation 3908e104-b9e0-4dac-8be4-f7b14aeb47bd · outbound

This paper cites MemRL: Self-Evolving Agents via Runtime Reinforcement Learning on Episodic Memory.

Memory Reward Inflation in Self-Improving LLM Agents MemRL: Self-Evolving Agents via Runtime Reinforcement Learning on Episodic Memory

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-04T02:16:41.870358Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T02:16:41.870358Z digest=sha256:fa0621269c274bc6a6b4c150de7aeccd196a1089fce5b33256df5dae33cefeea

Observation b3ce45e0-ef31-4735-8a5e-1e9c439c1224 · outbound

This paper cites Memory-R1: Enhancing Large Language Model Agents to Manage and Utilize Memories via Reinforcement Learning.

Memory Reward Inflation in Self-Improving LLM Agents Memory-R1: Enhancing Large Language Model Agents to Manage and Utilize Memories via Reinforcement Learning

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-04T02:16:41.921304Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T02:16:41.921304Z digest=sha256:2bcc8db728251a100537da292568c5f65cde3f2ec4af9dacdc168e333130078c

Observation 16ca1f9d-4d80-4d02-9ffe-3b8d22a1f76f · outbound

This paper cites Mem-{\alpha}: Learning Memory Construction via Reinforcement Learning.

Memory Reward Inflation in Self-Improving LLM Agents Mem-{\alpha}: Learning Memory Construction via Reinforcement Learning

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-04T02:16:41.975585Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T02:16:41.975585Z digest=sha256:5be886bebc4b51c154497d0207229abc3fd5a1f0a689634db5ddbd71e5cfe7c3

Observation c472d61b-1dc4-45c2-841a-9b957d0c94ef · outbound

This paper cites Agentic Memory: Learning Unified Long-Term and Short-Term Memory Management for Large Language Model Agents.

Memory Reward Inflation in Self-Improving LLM Agents Agentic Memory: Learning Unified Long-Term and Short-Term Memory Management for Large Language Model Agents

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-04T02:16:42.052502Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T02:16:42.052502Z digest=sha256:6bf49ef8c8f3a311f83a2d1503d0ecc60f4bf058dcedbd2dd832e7b5fca4a279

Observation f38068a7-e067-42f2-8516-ed08db3d8f06 · outbound

This paper cites Learning When to Remember: Risk-Sensitive Contextual Bandits for Abstention-Aware Memory Retrieval in LLM-Based Coding Agents.

Memory Reward Inflation in Self-Improving LLM Agents Learning When to Remember: Risk-Sensitive Contextual Bandits for Abstention-Aware Memory Retrieval in LLM-Based Coding Agents

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-04T02:16:42.111892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T02:16:42.111892Z digest=sha256:50cd5b843a477a6cb8b3357fd654d93f09e10989006c25fa8322659a55aeb23f

Observation 53f58fce-b58a-4ba1-b605-2a8acaf8246a · outbound

This paper cites Self-Consistency Improves Chain of Thought Reasoning in Language Models.

Memory Reward Inflation in Self-Improving LLM Agents Self-Consistency Improves Chain of Thought Reasoning in Language Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-04T02:16:42.170552Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T02:16:42.170552Z digest=sha256:e1c1e4630fd0e589c4b259997f178fe449b18e84104bbf760c0540c2283fdadd

Observation 82440356-8ab7-48d8-9169-ed5612ca0019 · outbound

This paper cites Process reward models that think.arXiv preprint arXiv:2504.16828, 2025.

Memory Reward Inflation in Self-Improving LLM Agents Process reward models that think.arXiv preprint arXiv:2504.16828, 2025

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-04T02:16:42.225662Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T02:16:42.225662Z digest=sha256:2e47c17182e685a3e074d1acd72284349ebf66c5f7ed0b5fcbb5932767057230

Observation 8a4080ab-d542-4623-b6fc-f8b558591d70 · outbound

This paper cites Self-Preference Bias in LLM-as-a-Judge.

Memory Reward Inflation in Self-Improving LLM Agents Self-Preference Bias in LLM-as-a-Judge

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-04T02:16:42.284063Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T02:16:42.284063Z digest=sha256:c533def5a973d7ad579f1fec55a2783480ce4f1d2b7302cabdb0eca1d8384d66

Observation a4352c5a-0e5b-440b-af5c-2a8cb2ba9922 · outbound

This paper cites Beyond the Surface: Measuring Self-Preference in LLM Judgments.

Memory Reward Inflation in Self-Improving LLM Agents Beyond the Surface: Measuring Self-Preference in LLM Judgments

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-04T02:16:42.324571Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T02:16:42.324571Z digest=sha256:f819e27d897036a36c60442b53f6086831152476d011ff4494419b7c29dc91e7

Observation dd5bacfb-663a-450b-a497-04c25024b9a9 · outbound

This paper cites Nine Judges, Two Effective Votes: Correlated Errors Undermine LLM Evaluation Panels.

Memory Reward Inflation in Self-Improving LLM Agents Nine Judges, Two Effective Votes: Correlated Errors Undermine LLM Evaluation Panels

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-04T02:16:42.387470Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T02:16:42.387470Z digest=sha256:bed515a5a61eee34c9eda35c3e58f414a7a3cf7e795bc7b113dafd76d3c1aee5

Observation 93271f3e-459c-4d8a-a4a1-31539340c8a6 · outbound

This paper cites QuickCheck: A lightweight tool for random testing of Haskell programs.

Memory Reward Inflation in Self-Improving LLM Agents QuickCheck: A lightweight tool for random testing of Haskell programs

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-04T02:16:42.432525Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T02:16:42.432525Z digest=sha256:cb30bb03b0910023b6300c3c66585771620055aa64dc871c1cf4bd18a9247e98

Observation af03401f-98e5-498a-a261-110f4ce576aa · outbound

This paper cites Finding and understanding bugs in C compilers.

Memory Reward Inflation in Self-Improving LLM Agents Finding and understanding bugs in C compilers

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-04T02:16:42.487271Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T02:16:42.487271Z digest=sha256:e3eab9bf9c607df6107fc4814ef1bc4e9057f9c2f1aff037b15e2b6892bddf27

Observation 46ef34c9-c955-4cc7-b7c5-511c74ad8216 · outbound

This paper cites Metamorphic testing: A new approach for generating next test cases.

Memory Reward Inflation in Self-Improving LLM Agents Metamorphic testing: A new approach for generating next test cases

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-04T02:16:42.553548Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T02:16:42.553548Z digest=sha256:742232c5af2c31750d6622901f573ef8a40dcde8e707d82e8a41c5dcbd5c7023

Observation 007d0614-1a62-46dc-a84f-6dcd08ccd8f5 · outbound

This paper cites Sanchez, and Antonio Ruiz-Cortés.

Memory Reward Inflation in Self-Improving LLM Agents Sanchez, and Antonio Ruiz-Cortés

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-04T02:16:42.599839Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T02:16:42.599839Z digest=sha256:f45ffe689aa479e7b8a597a42f1e6851b844877ccdad9463e9bb73269d71892f

Observation 627f7534-db2a-4faf-84b4-54f9f4c36b80 · outbound

This paper cites CodeT: Code Generation with Generated Tests.

Memory Reward Inflation in Self-Improving LLM Agents CodeT: Code Generation with Generated Tests

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-04T02:16:42.650963Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T02:16:42.650963Z digest=sha256:2a83361aa3ab07883d55c0776a9d21874713b7c75ec7993639c3635fd3d2038d

Observation 789a2fc0-204f-4784-9c26-0b7f2f84402e · outbound

This paper cites SEDM: Scalable self-evolving distributed memory for agents.arXiv preprint arXiv:2509.09498, 2025.

Memory Reward Inflation in Self-Improving LLM Agents SEDM: Scalable self-evolving distributed memory for agents.arXiv preprint arXiv:2509.09498, 2025

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-04T02:16:42.696928Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T02:16:42.696928Z digest=sha256:83d60e8e37417bd3d17f658e2653db41418b27b462cdac59d1f411ce04828411

Observation c92c542c-80b1-4a3c-93e5-ed051c3adf30 · outbound

This paper cites A-MemGuard: A proactive defense framework for LLM-based agent memory.arXiv preprint arXiv:2510.02373, 2025.

Memory Reward Inflation in Self-Improving LLM Agents A-MemGuard: A proactive defense framework for LLM-based agent memory.arXiv preprint arXiv:2510.02373, 2025

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-04T02:16:42.741781Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T02:16:42.741781Z digest=sha256:f910125964d687661026f7f2a506e2d47069cb2b5e1a4c3ed3ae3776369efcb2

Observation 6697af4e-3474-45b7-880a-d8c1346fca88 · outbound

This paper cites MemMA: Coordinating the memory cycle through multi-agent reasoning and in-situ self-evolution.arXiv preprint arXiv:2603.18718, 2026.

Memory Reward Inflation in Self-Improving LLM Agents MemMA: Coordinating the memory cycle through multi-agent reasoning and in-situ self-evolution.arXiv preprint arXiv:2603.18718, 2026

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-04T02:16:42.796735Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T02:16:42.796735Z digest=sha256:1e5b7739db765be5f1f0ce0548427336288a62a0c4060788934659ce551becac

Observation 513546c6-3ba1-4884-aca6-dd24b0455df1 · outbound

This paper cites Useful Memories Become Faulty When Continuously Updated by LLMs.

Memory Reward Inflation in Self-Improving LLM Agents Useful Memories Become Faulty When Continuously Updated by LLMs

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-04T02:16:42.847936Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T02:16:42.847936Z digest=sha256:4ace065edd964a0a4f8bc9dfb48e9887b92578801363128367f6863f1104aa3a

Observation 982f9154-3941-4ad3-9d34-9ddeabfc1b34 · outbound

This paper cites Large Language Models Cannot Self-Correct Reasoning Yet.

Memory Reward Inflation in Self-Improving LLM Agents Large Language Models Cannot Self-Correct Reasoning Yet

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-04T02:16:42.906699Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T02:16:42.906699Z digest=sha256:40399dfe600bfa82cf536aedc64372371cd4f12416cfbe7174f0fd9cc8b76bdb

Observation f5ca1904-22a6-4d15-8504-7dec3c645ce7 · outbound

This paper cites Estimating the Accuracies of Multiple Classifiers Without Labeled Data.

Memory Reward Inflation in Self-Improving LLM Agents Estimating the Accuracies of Multiple Classifiers Without Labeled Data

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-04T02:16:42.955345Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T02:16:42.955345Z digest=sha256:c9eb4bb4aa0beb3b7bfc9f0d638b83a2214245ab45cc05d8c12a1ae6298fe35d

Observation db240d47-19ca-4da0-8977-24a1072d9624 · outbound

This paper cites The logic of NTQR evaluations of noisy AI agents: Complete postulates and logically consistent error correlations.

Memory Reward Inflation in Self-Improving LLM Agents The logic of NTQR evaluations of noisy AI agents: Complete postulates and logically consistent error correlations

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-04T02:16:43.009324Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T02:16:43.009324Z digest=sha256:c4eec4c3f919f497c2a37c348d37c8e9d5edb6c2c6554624655914a3a87ae628

Observation 972e9bf5-95b7-4051-8b8c-e970ef6b557d · outbound

This paper cites Can LLM Already Serve as A Database Interface? A BIg Bench for Large-Scale Database Grounded Text-to-SQLs.

Memory Reward Inflation in Self-Improving LLM Agents Can LLM Already Serve as A Database Interface? A BIg Bench for Large-Scale Database Grounded Text-to-SQLs

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-04T02:16:43.096931Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T02:16:43.096931Z digest=sha256:985b41e71e39720c753e94d7924ef3af7e689ba7d2e981f2930d665dee558416

Observation f4f5eede-59ea-405a-adce-2633ee73e701 · outbound

This paper cites SimCSE: Simple contrastive learning of sentence embeddings.

Memory Reward Inflation in Self-Improving LLM Agents SimCSE: Simple contrastive learning of sentence embeddings

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-04T02:16:43.153818Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T02:16:43.153818Z digest=sha256:7ecf9e889dcee6dd7aef594ae124e2e4dbdf43a2eddc9dd85a078e11d0d0ef8d

Pith citing papers

No inbound Pith citation observations are available.