Pith. sign in

Paper Citation Record · LEDGER

Large Language Model-Enhanced Reinforcement Learning for Diverse and Novel Recommendations

As of 8 August 2026, this Paper Citation Record lists 36 of 36 outbound references and 0 inbound Pith citation observations for arXiv:2507.21274.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.21274 v1

Coverage vector

measured 36 of 36 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T13:06:32.504261Z

measured 36 of 36 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

36 of 36 outbound references displayed

  • verified exact1
  • verified fuzzy21
  • unresolved14
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 9527e20e-14bb-4f47-890d-93febf486a30 · outbound

This paper cites Llama 3 model card.

Large Language Model-Enhanced Reinforcement Learning for Diverse and Novel Recommendations Llama 3 model card

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T13:06:32.270530Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:06:32.270530Z digest=sha256:71e94acfb502b5b8d6cbb27e0d154513567f4407f3234f776c49b062cda5f60e

Observation 31fce3d4-00b5-477b-a939-fa271ee711c5 · outbound

This paper cites The claude 3 model family: Opus, sonnet, haiku.

Large Language Model-Enhanced Reinforcement Learning for Diverse and Novel Recommendations The claude 3 model family: Opus, sonnet, haiku

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T13:06:32.283632Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:06:32.283632Z digest=sha256:31dc1a472ffc530ef65bf370d8ad3302e826f477003f0d7778f1d189be544070

Observation 6cf12dc9-72f4-42f0-b329-8c7c7fafc310 · outbound

This paper cites Tallrec: An effective and efficient tuning framework to align large language model with recommendation.

Large Language Model-Enhanced Reinforcement Learning for Diverse and Novel Recommendations Tallrec: An effective and efficient tuning framework to align large language model with recommendation

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:06:33.224678Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:06:32.289399Z digest=sha256:520640666e72254754255db432d77c30821db670852b0a994d83f03dba139c41

Observation 86ea9218-6a7d-4a89-b1ee-fe433c9d6f35 · outbound

This paper cites Adversarial model for offline reinforcement learning.

Large Language Model-Enhanced Reinforcement Learning for Diverse and Novel Recommendations Adversarial model for offline reinforcement learning

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:06:33.206275Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:06:32.298008Z digest=sha256:d233e30a0c0c357f03d891e731d3b60e90e745d2b4cb8eb095d166607c8035a1

Observation c39dd834-a376-4763-9a45-3242aab78d80 · outbound

This paper cites Stochastic approximation with two time scales.

Large Language Model-Enhanced Reinforcement Learning for Diverse and Novel Recommendations Stochastic approximation with two time scales

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:06:33.188871Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:06:32.304396Z digest=sha256:4c9e1a2c456cdf46ee79f18676a996a50dac4190588438d6e3b09fff32a3818c

Observation a7b8571b-43b6-4aff-b1c8-3a6987727fee · outbound

This paper cites When Large Language Models Meet Personalization: Perspectives of Challenges and Opportunities.

Large Language Model-Enhanced Reinforcement Learning for Diverse and Novel Recommendations When Large Language Models Meet Personalization: Perspectives of Challenges and Opportunities

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T13:06:32.314209Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:06:32.314209Z digest=sha256:43449d9e416eedb4ff6837d0f5e9aa2f8d2827784f36dd9a1e15a42c8a500443

Observation 2fe45c2c-ab35-4408-b387-13fb4afe38b6 · outbound

This paper cites Adversarially trained actor critic for offline reinforcement learning.

Large Language Model-Enhanced Reinforcement Learning for Diverse and Novel Recommendations Adversarially trained actor critic for offline reinforcement learning

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:06:33.171035Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:06:32.321400Z digest=sha256:6fa8dde7459498caf66e5b573b23f0c0dd27d5414c51261766354d821c3b135f

Observation 235ee978-ba42-4412-b5bf-4161837b600c · outbound

This paper cites On the Properties of Neural Machine Translation: Encoder-Decoder Approaches.

Large Language Model-Enhanced Reinforcement Learning for Diverse and Novel Recommendations On the Properties of Neural Machine Translation: Encoder-Decoder Approaches

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T13:06:32.326866Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:06:32.326866Z digest=sha256:07470984879131901239e5c0bd842f1f5c39f8b7460c5f0354fb090d21bfdb2b

Observation 631a51b8-7791-4f5e-b835-bf11c20585b5 · outbound

This paper cites A review of modern recommender systems using generative models (gen-recsys).

Large Language Model-Enhanced Reinforcement Learning for Diverse and Novel Recommendations A review of modern recommender systems using generative models (gen-recsys)

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:06:33.155646Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:06:32.335538Z digest=sha256:6dcb16fe70a0839241e5528125038b962eae6ead05d40d849fcb3560d4658b1e

Observation ead6fcdf-e467-46ea-ae23-9cd1aded9f74 · outbound

This paper cites BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding.

Large Language Model-Enhanced Reinforcement Learning for Diverse and Novel Recommendations BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T13:06:32.341879Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:06:32.341879Z digest=sha256:59961caac5818b932727f814263692cd78b6f5bbe83f1a741ff0882547d0bb6d

Observation 9efe23b4-678d-427b-b898-80ecc37f7feb · outbound

This paper cites Recommender Systems in the Era of Large Language Models (LLMs).

Large Language Model-Enhanced Reinforcement Learning for Diverse and Novel Recommendations Recommender Systems in the Era of Large Language Models (LLMs)

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T13:06:32.347874Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:06:32.347874Z digest=sha256:bf0f20b3a80010388fd6145554fd857280e79fe7dc1bfa069431db04b5c67c0c

Observation d4bc672e-c5d1-43b7-bafe-1f786eff552b · outbound

This paper cites Addressing function approxi- mation error in actor-critic methods.

Large Language Model-Enhanced Reinforcement Learning for Diverse and Novel Recommendations Addressing function approxi- mation error in actor-critic methods

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:06:33.140090Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:06:32.363650Z digest=sha256:5f0fc52ebdb7d526774b0b189069cf61ffeb84c060a4143607b637576dc7f1af

Observation 8510f530-5fd6-422b-8d3e-cf9fd20045bd · outbound

This paper cites Soft actor- critic: Off-policy maximum entropy deep reinforcement learning with a stochas- tic actor.

Large Language Model-Enhanced Reinforcement Learning for Diverse and Novel Recommendations Soft actor- critic: Off-policy maximum entropy deep reinforcement learning with a stochas- tic actor

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:06:33.122823Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:06:32.369523Z digest=sha256:f538371e404e47744abba8f5e4fc44835834ec535d752d26ef8abe78233e7902

Observation ab507f5d-99fd-42ba-992a-c16053591d26 · outbound

This paper cites Maxwell Harper and Joseph A.

Large Language Model-Enhanced Reinforcement Learning for Diverse and Novel Recommendations Maxwell Harper and Joseph A

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:06:33.106299Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:06:32.376998Z digest=sha256:78320046607644213b1823900caa58bd5fb9e07ee540e756b78dd2f395db1090

Observation c7d85c18-6e52-439e-b37e-2632fd564f0a · outbound

This paper cites Large lan- guage models as zero-shot conversational recommenders.

Large Language Model-Enhanced Reinforcement Learning for Diverse and Novel Recommendations Large lan- guage models as zero-shot conversational recommenders

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:06:33.088042Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:06:32.381848Z digest=sha256:eb24427ef00d219c1fe99b7d5ba2ddc58112cead34c1b51e440322d616575dd2

Observation 76edea6e-06dc-44e9-8ad0-8bf3260464d8 · outbound

This paper cites Session-based Recommendations with Recurrent Neural Networks.

Large Language Model-Enhanced Reinforcement Learning for Diverse and Novel Recommendations Session-based Recommendations with Recurrent Neural Networks

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T13:06:32.387219Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:06:32.387219Z digest=sha256:148f681e29f9d5edb4698767b72fa8b988733143d9a32f39cdf714477f8b90c4

Observation 44a34131-597b-490d-a7c3-db8d9ab19b74 · outbound

This paper cites Towards universal sequence representation learning for recommender systems.

Large Language Model-Enhanced Reinforcement Learning for Diverse and Novel Recommendations Towards universal sequence representation learning for recommender systems

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:06:33.070053Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:06:32.394208Z digest=sha256:ddadff7b97e8294f67cccad7323c46235958ce6a177bc806d7d1aa6a887d8c5c

Observation 2b46a6e7-e9f2-4988-968a-1200f7ca45dc · outbound

This paper cites LoRA: Low-Rank Adaptation of Large Language Models.

Large Language Model-Enhanced Reinforcement Learning for Diverse and Novel Recommendations LoRA: Low-Rank Adaptation of Large Language Models

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T13:06:32.399332Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:06:32.399332Z digest=sha256:e00177ce392a1a6fe5d1b134188db6d9e0bf20bbd78a4e53c6a7c94f0a101b21

Observation b2bf7143-86ce-49bc-ae97-f31aca2c4e00 · outbound

This paper cites Human-centric dialog training via offline reinforcement learning.

Large Language Model-Enhanced Reinforcement Learning for Diverse and Novel Recommendations Human-centric dialog training via offline reinforcement learning

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:06:33.049119Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:06:32.405283Z digest=sha256:bd55d5c21bba055458222eec469165cf7d765ccb59c6d0b90dd7e24b6bcad6b1

Observation b2791e8f-824f-4d4b-9c64-2545aa38da62 · outbound

This paper cites Cumulated gain-based evaluation of ir techniques.

Large Language Model-Enhanced Reinforcement Learning for Diverse and Novel Recommendations Cumulated gain-based evaluation of ir techniques

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:06:33.033601Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:06:32.413787Z digest=sha256:8efc65670d031e4d61a934b477894ca7ea501b558c9e7432de282b5e7813e577

Observation 40441afe-eff5-4cbe-9117-4ad73561fa0f · outbound

This paper cites Kingma and Jimmy Ba.

Large Language Model-Enhanced Reinforcement Learning for Diverse and Novel Recommendations Kingma and Jimmy Ba

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:06:33.016934Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:06:32.420117Z digest=sha256:6d7cd24c673fd1a9b025d1f2e0b52082f7f5f973e5e41e29c352e3267252247a

Observation d9bb156e-7077-4a25-b77e-4046b99350b9 · outbound

This paper cites How Can Recommender Systems Benefit from Large Language Models: A Survey.

Large Language Model-Enhanced Reinforcement Learning for Diverse and Novel Recommendations How Can Recommender Systems Benefit from Large Language Models: A Survey

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T13:06:32.425492Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:06:32.425492Z digest=sha256:e83e8718d5054c9a6ad1ad7a394f67f7f162809abdac78116a415fc5281323a8

Observation 0aa9d40c-b56d-47cb-b921-156523c8a3a5 · outbound

This paper cites Pre-train, prompt, and predict: A systematic survey of prompting methods in natural language processing.

Large Language Model-Enhanced Reinforcement Learning for Diverse and Novel Recommendations Pre-train, prompt, and predict: A systematic survey of prompting methods in natural language processing

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:06:32.997853Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:06:32.430643Z digest=sha256:226084c339d21a4fa1f88d2d1b8961134d664fecda216c42db007b7d2434219e

Observation 3c55693a-53a4-4bff-a68f-1f4d78cc7583 · outbound

This paper cites Diversity-promoting deep reinforcement learning for interactive recommendation.

Large Language Model-Enhanced Reinforcement Learning for Diverse and Novel Recommendations Diversity-promoting deep reinforcement learning for interactive recommendation

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:06:32.980680Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:06:32.436766Z digest=sha256:988a377d18c2e010c738961f950a466df06b773a2b554c1c55c44886c01e4530

Observation 7a19ab21-727d-43e0-912d-dc2aefe06833 · outbound

This paper cites Convergent temporal-difference learning with arbitrary smooth function approximation.

Large Language Model-Enhanced Reinforcement Learning for Diverse and Novel Recommendations Convergent temporal-difference learning with arbitrary smooth function approximation

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:06:32.962113Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:06:32.442735Z digest=sha256:e4d40d646cb7e6567cbada9c8528b20f34983fe07c91e1cd6424e6ca71b331b2

Observation 557e0e8c-9c13-45e2-bb4d-54bb07b0ebd7 · outbound

This paper cites Recent advances in natural language processing via large pre-trained language models: A survey.

Large Language Model-Enhanced Reinforcement Learning for Diverse and Novel Recommendations Recent advances in natural language processing via large pre-trained language models: A survey

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:06:32.944386Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:06:32.447719Z digest=sha256:971aa66ab20566df5c773bb77302d087a2ba1e6fe8490d338d26988c423b848a

Observation 1254bda8-8018-46c8-a2a2-f9a3e33fc42c · outbound

This paper cites Training language models to follow instructions with human feedback.Advances in Neural Information Processing Systems , 35:27730–27744, 2022.

Large Language Model-Enhanced Reinforcement Learning for Diverse and Novel Recommendations Training language models to follow instructions with human feedback.Advances in Neural Information Processing Systems , 35:27730–27744, 2022

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:06:32.926015Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:06:32.453217Z digest=sha256:58f9c1516cc288073e42cb058c1efe380b3ddd04ec3afe2e0877a7017fc83ee2

Observation c5381e4b-f6c2-4ae3-bc1a-c2fbe106b12e · outbound

This paper cites Proximal Policy Optimization Algorithms.

Large Language Model-Enhanced Reinforcement Learning for Diverse and Novel Recommendations Proximal Policy Optimization Algorithms

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T13:06:32.458374Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:06:32.458374Z digest=sha256:ced5ac159bbf71fffff15aba85ec331c90d86464f2d85749579681e8faf7e176

Observation 257ccca1-7c6c-4e00-bdda-02925ad0f94e · outbound

This paper cites Choosing the best of both worlds: Diverse and novel recom- mendations through multi-objective reinforcement learning.

Large Language Model-Enhanced Reinforcement Learning for Diverse and Novel Recommendations Choosing the best of both worlds: Diverse and novel recom- mendations through multi-objective reinforcement learning

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:06:32.909052Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:06:32.463913Z digest=sha256:13209c295f49b974092d230efcfab4c6ae91c545119b62921689f4249e913267

Observation ba59590b-e618-4bec-baca-a6ae722e81ac · outbound

This paper cites Learning to summarize with human feedback.

Large Language Model-Enhanced Reinforcement Learning for Diverse and Novel Recommendations Learning to summarize with human feedback

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:06:32.891898Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:06:32.468816Z digest=sha256:7cfaf479bf52a70f3cf279adb6c6b3d4a15df8decfb847011bb29288a1c9d204

Observation 1cfc025b-b6c0-4950-a669-f6b8aeeb1eac · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

Large Language Model-Enhanced Reinforcement Learning for Diverse and Novel Recommendations LLaMA: Open and Efficient Foundation Language Models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T13:06:32.473790Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:06:32.473790Z digest=sha256:706d875ffa28d7e2e16eabc2db99c71b7e250f4e23d5ccced5df4d02f6ebb235

Observation f60b4162-d92c-456d-a661-93715900ebbe · outbound

This paper cites an unresolved cited work.

Large Language Model-Enhanced Reinforcement Learning for Diverse and Novel Recommendations Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-06T13:06:32.874037Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:06:32.479004Z digest=sha256:9010846ac672ba2d5365c52d153159f689630ce2b19e7fbd367c1819e54d768d

Observation b5e0e357-3cea-4b17-9434-9ac56afaef47 · outbound

This paper cites Transrec: Learning transferable recommendation from mixture-of-modality feedback.

Large Language Model-Enhanced Reinforcement Learning for Diverse and Novel Recommendations Transrec: Learning transferable recommendation from mixture-of-modality feedback

Reference 33

Resolution
verified exact
raw_fallback, observed 2026-08-06T13:06:32.684421Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:06:32.486609Z digest=sha256:a63fc56d0612ae03ce9311dc1aaee44e04e450447abff402fba3359b443bb6fb

Observation d51c539b-01bf-43af-8f05-39fdd9744ee8 · outbound

This paper cites Behavior Regularized Offline Reinforcement Learning.

Large Language Model-Enhanced Reinforcement Learning for Diverse and Novel Recommendations Behavior Regularized Offline Reinforcement Learning

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T13:06:32.492596Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:06:32.492596Z digest=sha256:50d58ec88c31720df48a4cd48bd81252ea525f5f517caeb03cc4e1803b0ee397

Observation 4095c89b-7e48-4600-bee7-e7563e198836 · outbound

This paper cites A Survey of Large Language Models.

Large Language Model-Enhanced Reinforcement Learning for Diverse and Novel Recommendations A Survey of Large Language Models

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T13:06:32.497705Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T13:06:32.497705Z digest=sha256:eeda887ea691463db73533b41db99b237d5fbd9565820cce2cb9b419d85a6798

Observation 3bd8393f-bddc-4a8e-b59c-d6c18f435ca0 · outbound

This paper cites Drn: A deep reinforcement learning framework for news recommendation.

Large Language Model-Enhanced Reinforcement Learning for Diverse and Novel Recommendations Drn: A deep reinforcement learning framework for news recommendation

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T13:06:32.857292Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T13:06:32.504261Z digest=sha256:4803c1282aeb5a886a1dc4eb99e2cd0d7607b178761c9568a16c9ae8437d2c00

Pith citing papers

No inbound Pith citation observations are available.