Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-10T16:24:28.048403Z
Paper Citation Record · LEDGER
As of 4 August 2026, this Paper Citation Record lists 44 of 44 outbound references and 1 inbound Pith citation observation for arXiv:2604.10029.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-05-10T16:24:28.048403Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-01T02:44:32.617211Z
A source-named dated measurement, never combined with another source.
Source: cited_works
44 of 44 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation d53063b9-aecc-4e31-be81-8f8160b9b60f · outbound
Self-Distilled Reinforcement Learning for Co-Evolving Agentic Recommender Systems Self-attentive sequential recommenda- tion
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 004673d4-d277-4971-b29c-fba74226ca7d · outbound
Self-Distilled Reinforcement Learning for Co-Evolving Agentic Recommender Systems Lightgcn: Simplifying and powering graph convolution network for recommenda- tion
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation e049c78f-75df-454b-b40a-8083d12d1d15 · outbound
Self-Distilled Reinforcement Learning for Co-Evolving Agentic Recommender Systems Efficient bi- level optimization for recommendation denoising
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 265aa792-383e-4cc8-8233-2462ee8e66bb · outbound
Self-Distilled Reinforcement Learning for Co-Evolving Agentic Recommender Systems Agentic feedback loop modeling improves recommendation and user simulation
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation f6a94efa-f04f-4547-884e-ed9b78454830 · outbound
Self-Distilled Reinforcement Learning for Co-Evolving Agentic Recommender Systems iagent: Llm agent as a shield between user and recommender systems
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation a19e5a89-2daa-4e0d-9bc9-d48983e9e6ad · outbound
Self-Distilled Reinforcement Learning for Co-Evolving Agentic Recommender Systems On generative agents in recommendation
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation a2014003-e152-4844-8471-6910c3a0d9d7 · outbound
Self-Distilled Reinforcement Learning for Co-Evolving Agentic Recommender Systems Recommender ai agent: Integrating large language models for interactive recommen- dations
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 424c98d2-d34e-440c-91a8-6febe9ff8d0f · outbound
Self-Distilled Reinforcement Learning for Co-Evolving Agentic Recommender Systems Re- flexion: Language agents with verbal reinforcement learning
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation bbca6b1a-1446-430e-a2b5-bf9631f2e591 · outbound
Self-Distilled Reinforcement Learning for Co-Evolving Agentic Recommender Systems Entropy guided diversification and preference elicitation in agentic recommendation systems
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 1e6a731c-e948-40f1-926a-1d0f766b3e70 · outbound
Self-Distilled Reinforcement Learning for Co-Evolving Agentic Recommender Systems Memorybank: En- hancing large language models with long-term memory
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation e05d95aa-3ffd-4d06-a64a-c59c7d9b43f1 · outbound
Self-Distilled Reinforcement Learning for Co-Evolving Agentic Recommender Systems A-MEM: Agentic Memory for LLM Agents
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 92f4dbd6-ee38-4495-ae89-95f6a3f1f76b · outbound
Self-Distilled Reinforcement Learning for Co-Evolving Agentic Recommender Systems Tallrec: An effective and efficient tuning framework to align large language model with recommendation
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation c5bd9325-2b47-4bde-89ee-3988d8ff067d · outbound
Self-Distilled Reinforcement Learning for Co-Evolving Agentic Recommender Systems DAPO: An Open-Source LLM Reinforcement Learning System at Scale
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 027e70ce-192f-4647-becc-138042797b73 · outbound
Self-Distilled Reinforcement Learning for Co-Evolving Agentic Recommender Systems Proximal Policy Optimization Algorithms
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 2035d37c-4579-4a40-a1e1-93e3616bfd03 · outbound
Self-Distilled Reinforcement Learning for Co-Evolving Agentic Recommender Systems Deepseek-r1 incentivizes reasoning in llms through reinforcement learning
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation a91d93f1-b305-491a-b8e3-0a7b3f18dc95 · outbound
Self-Distilled Reinforcement Learning for Co-Evolving Agentic Recommender Systems Amem4rec: Leveraging cross-user similarity for memory evolution in agentic llm recom- menders
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 5baec4f5-324a-4310-8ac3-69bae702ccd3 · outbound
Self-Distilled Reinforcement Learning for Co-Evolving Agentic Recommender Systems A survey on agent-as-a-judge
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 9a9f5a69-edd7-4e7b-86e0-fae84021fd1f · outbound
Self-Distilled Reinforcement Learning for Co-Evolving Agentic Recommender Systems RecoWorld: Building Simulated Environments for Agentic Recommender Systems
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 9ae63de2-cb16-4583-bad8-6ddbaa637a2f · outbound
Self-Distilled Reinforcement Learning for Co-Evolving Agentic Recommender Systems RuleAgent: Discovering Rules for Recommendation Denoising with Autonomous Language Agents
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 4f33a700-ae2c-4547-8713-6ffdaa565abe · outbound
Self-Distilled Reinforcement Learning for Co-Evolving Agentic Recommender Systems Macrec: A multi- agent collaboration framework for recommendation
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 9664e998-438a-4047-8696-33e059526f36 · outbound
Self-Distilled Reinforcement Learning for Co-Evolving Agentic Recommender Systems arXiv preprint arXiv:2602.02482 , year=
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 8ccd2e18-9a25-4da2-8a73-d799961d75c3 · outbound
Self-Distilled Reinforcement Learning for Co-Evolving Agentic Recommender Systems Treerl: Llm reinforcement learning with on-policy tree search
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation aa9dd6c5-6bd4-4fc7-855d-5344e529ec31 · outbound
Self-Distilled Reinforcement Learning for Co-Evolving Agentic Recommender Systems Supervised pretraining can learn in-context reinforcement learning
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 3be3a04a-107c-4a9a-886c-3834f5d9d2c6 · outbound
Self-Distilled Reinforcement Learning for Co-Evolving Agentic Recommender Systems Reward Is Enough: LLMs Are In-Context Reinforcement Learners
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 3348998f-c5a2-4a65-bca7-2bc632d57b7f · outbound
Self-Distilled Reinforcement Learning for Co-Evolving Agentic Recommender Systems Reinforcement Learning with Verifiable Rewards Implicitly Incentivizes Correct Reasoning in Base LLMs
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 491749b6-12b7-4845-92e4-515c786d01c9 · outbound
Self-Distilled Reinforcement Learning for Co-Evolving Agentic Recommender Systems Self-Distillation Enables Continual Learning
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation c9b4f882-5512-42be-ad90-1f98b2eec84c · outbound
Self-Distilled Reinforcement Learning for Co-Evolving Agentic Recommender Systems Reinforcement Learning via Self-Distillation
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 887d32f4-fa2f-46ef-a8cb-213c931246c8 · outbound
Self-Distilled Reinforcement Learning for Co-Evolving Agentic Recommender Systems Self-Distilled RLVR
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 20a7b5c6-b756-45ab-a252-800e25f92d69 · outbound
Self-Distilled Reinforcement Learning for Co-Evolving Agentic Recommender Systems Second workshop on infor- mation heterogeneity and fusion in recommender systems (hetrec2011)
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 94b0ae1d-bbed-426b-ac67-6784f6ac5fdb · outbound
Self-Distilled Reinforcement Learning for Co-Evolving Agentic Recommender Systems The movielens datasets: History and context
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation e6248ad7-e440-4414-bfd1-9e47d9fef17b · outbound
Self-Distilled Reinforcement Learning for Co-Evolving Agentic Recommender Systems Bridging Language and Items for Retrieval and Recommendation: Benchmarking LLMs as Semantic Encoders
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 25c7840d-4141-441b-a28f-73e075306bd7 · outbound
Self-Distilled Reinforcement Learning for Co-Evolving Agentic Recommender Systems Let me do it for you: Towards llm empowered recommendation via tool learning
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 708f2f17-0135-4b9d-8ab7-89c46c7bac65 · outbound
Self-Distilled Reinforcement Learning for Co-Evolving Agentic Recommender Systems Lora: Low-rank adaptation of large language models
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation e39cb58a-cef5-4ab6-9188-47a402cb0ef6 · outbound
Self-Distilled Reinforcement Learning for Co-Evolving Agentic Recommender Systems Self-Distilled Reasoner: On-Policy Self-Distillation for Large Language Models
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 6dd8baba-c873-45ae-bac6-a8e256312811 · outbound
Self-Distilled Reinforcement Learning for Co-Evolving Agentic Recommender Systems Negotiating the shared agency between humans & ai in the recommender system
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation bc2b7259-1584-4f12-afc5-494c5223dbe1 · outbound
Self-Distilled Reinforcement Learning for Co-Evolving Agentic Recommender Systems User behavior simulation with large lan- guage model-based agents
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 2153ce36-707b-4465-9fa8-4d405eba717b · outbound
Self-Distilled Reinforcement Learning for Co-Evolving Agentic Recommender Systems Recmind: Large language model powered agent for recommendation
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 3d8a3fc9-d1e3-44e3-8ab7-a1ebcbdf9dca · outbound
Self-Distilled Reinforcement Learning for Co-Evolving Agentic Recommender Systems Id-free not risk-free: Llm-powered agents unveil risks in id- free recommender systems
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 8b804b5b-c66a-4187-bc16-60ab154bb62a · outbound
Self-Distilled Reinforcement Learning for Co-Evolving Agentic Recommender Systems MemRec: Collaborative Memory-Augmented Agentic Recommender System
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 8bc6d331-a311-48a7-bd55-f96ea70c9114 · outbound
Self-Distilled Reinforcement Learning for Co-Evolving Agentic Recommender Systems Agentcf: Collaborative learning with autonomous language agents for recommender systems
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 67d6453c-12cb-4168-a772-dfbe4c7c8125 · outbound
Self-Distilled Reinforcement Learning for Co-Evolving Agentic Recommender Systems Agentcf++: Memory-enhanced llm-based agents for popularity-aware cross-domain recommendations
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation c6c01178-2781-4371-a4e7-aa3f7f6f528f · outbound
Self-Distilled Reinforcement Learning for Co-Evolving Agentic Recommender Systems Multi-agent collaborative filtering: Orchestrating users and items for agentic rec- ommendations
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 22b2884f-a43b-4317-89cc-049ce1ff0cfb · outbound
Self-Distilled Reinforcement Learning for Co-Evolving Agentic Recommender Systems Recnet: Self-evolving preference propagation for agentic recommender systems.arXiv preprint arXiv:2601.21609
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation aa7799e6-5e29-4d17-b41e-e2cf8e40e283 · outbound
Self-Distilled Reinforcement Learning for Co-Evolving Agentic Recommender Systems DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.
Observation 5ba483d8-be47-44fc-a746-f91a8801b4f8 · inbound
ODYSSE: Episode-wise Policy Optimization for Personalized Agentic Reasoning Self-Distilled Reinforcement Learning for Co-Evolving Agentic Recommender Systems
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.