Pith. sign in

Paper Citation Record · LEDGER

Reinforced Preference Optimization for Reasoning-Augmented Recommendations

As of 4 August 2026, this Paper Citation Record lists 42 of 42 outbound references and 0 inbound Pith citation observations for arXiv:2605.21967.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2605.21967 v1

Coverage vector

measured 42 of 42 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-22T04:34:28.214871Z

measured 42 of 42 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-03T06:30:56.289259+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

42 of 42 outbound references displayed

  • verified exact16
  • verified fuzzy4
  • unresolved22
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 5f9ba59e-0296-4b24-b18a-f15a42058393 · outbound

This paper cites an unresolved cited work.

Reinforced Preference Optimization for Reasoning-Augmented Recommendations Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-05-22T04:34:35.644582Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-22T04:34:28.214871Z digest=sha256:8bffeaa0185a18e6353155f149f8fc0fe0a803136a7e606c3a80dbd7dff92e39

Observation f81eb998-0c2d-49a0-a464-b926445ba0e2 · outbound

This paper cites an unresolved cited work.

Reinforced Preference Optimization for Reasoning-Augmented Recommendations Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-05-22T04:34:35.647847Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-22T04:34:28.214871Z digest=sha256:e75d12fc0be15a8b5c34b5190f25b59cedd5df686d3d5dcadbb8b53e0b79b5b1

Observation a43db1e8-bd24-4d81-abe4-ea6926ef74ec · outbound

This paper cites InProceedings of the Nineteenth ACM Conference on Recommender Systems.

Reinforced Preference Optimization for Reasoning-Augmented Recommendations InProceedings of the Nineteenth ACM Conference on Recommender Systems

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T04:34:35.641722Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-22T04:34:28.214871Z digest=sha256:ec1b6996214db72bf40748cee18e57e98e89fe648224469dd82cd45e98b99d07

Observation dba4d269-451a-4dc5-9fe4-d32a52165094 · outbound

This paper cites an unresolved cited work.

Reinforced Preference Optimization for Reasoning-Augmented Recommendations Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-05-22T04:34:35.650634Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-22T04:34:28.214871Z digest=sha256:ac40fd588b0b40d092dee6c8a99ca2097089cd2f9656955979ab32d5bf5bcf5b

Observation 68256a4f-22f4-424b-865b-8b34e5305d2e · outbound

This paper cites Decoding Matters: Addressing Amplification Bias and Homogeneity Issue for LLM-based Recommendation.

Reinforced Preference Optimization for Reasoning-Augmented Recommendations Decoding Matters: Addressing Amplification Bias and Homogeneity Issue for LLM-based Recommendation

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-22T04:34:35.247882Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-22T04:34:28.214871Z digest=sha256:fda62e652b9bc8ed843a73e4035b40621b4b9566b6ab146c150bcd20737d6265

Observation 901a19b3-4c19-41f0-80b9-d83604ae85d7 · outbound

This paper cites an unresolved cited work.

Reinforced Preference Optimization for Reasoning-Augmented Recommendations Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-05-22T04:34:35.642807Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-22T04:34:28.214871Z digest=sha256:cabea5f25bb048351cc7811c122babf4c390e0345d5191d3f8f6db8c403f73de

Observation 122bbed6-47e0-47fd-b3dd-992606c0360e · outbound

This paper cites an unresolved cited work.

Reinforced Preference Optimization for Reasoning-Augmented Recommendations Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-05-22T04:34:35.645331Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-22T04:34:28.214871Z digest=sha256:b4aedc9220ce1cad7111dbbf554f6ac058ff834506b6ab1b75107506f9da29f5

Observation 1ab0c19f-05e9-45a8-be37-e9f75abb17ad · outbound

This paper cites an unresolved cited work.

Reinforced Preference Optimization for Reasoning-Augmented Recommendations Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-05-22T04:34:35.653555Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-22T04:34:28.214871Z digest=sha256:797f9d35b745ce14aacc2f4e28db7c6e8a240c76954e083dca9dc99c26a93b93

Observation 6549db46-65d2-423e-81f1-298d9a728fb5 · outbound

This paper cites From LLM Reasoning to Autonomous AI Agents: A Comprehensive Review.

Reinforced Preference Optimization for Reasoning-Augmented Recommendations From LLM Reasoning to Autonomous AI Agents: A Comprehensive Review

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-05-22T04:34:35.339052Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-22T04:34:28.214871Z digest=sha256:339635ef2f717a0d68ecd26b20a9f7cb0d65977a4c635407f7fc7b1cd217355d

Observation 52b78f38-31de-40c8-9f96-fee05a2476dc · outbound

This paper cites an unresolved cited work.

Reinforced Preference Optimization for Reasoning-Augmented Recommendations Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-05-22T04:34:35.603854Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-22T04:34:28.214871Z digest=sha256:2b9e000ec783d19e7ea304237442f39dc7cec448861a16ae9bb36ecc888b20f6

Observation 56454867-4fed-4f8e-8e7f-4954db91bd15 · outbound

This paper cites InProceedings of the ACM on Web Conference 2025.

Reinforced Preference Optimization for Reasoning-Augmented Recommendations InProceedings of the ACM on Web Conference 2025

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T04:34:35.626220Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-22T04:34:28.214871Z digest=sha256:50936674c01dd27f24f45c056a29260687c3ce179d851d557a775414c7a63d2e

Observation 66db41d6-8c86-4691-ad3c-3ee1fd32cd3b · outbound

This paper cites an unresolved cited work.

Reinforced Preference Optimization for Reasoning-Augmented Recommendations Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-05-22T04:34:35.605142Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-22T04:34:28.214871Z digest=sha256:fcbf337a34b7bd96022aa2726308f2f535f737fe6a32b1e7e6647c6604825695

Observation eb953378-fc18-47b5-8b12-7f38a1f3a1ff · outbound

This paper cites Efficient Natural Language Response Suggestion for Smart Reply.

Reinforced Preference Optimization for Reasoning-Augmented Recommendations Efficient Natural Language Response Suggestion for Smart Reply

Reference 13

Resolution
verified exact
local_arxiv, observed 2026-05-22T04:34:35.312490Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-22T04:34:28.214871Z digest=sha256:8c09f5f5ca9db723b98b1e7130075271820a36599c8424ed05fe24eb47ef054f

Observation a36c88bf-eaed-48ab-b7b4-2dbf57488ca0 · outbound

This paper cites Session-based Recommendations with Recurrent Neural Networks.

Reinforced Preference Optimization for Reasoning-Augmented Recommendations Session-based Recommendations with Recurrent Neural Networks

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-05-22T04:34:35.330626Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-22T04:34:28.214871Z digest=sha256:2cf3d7f171243e4da29d975c8bbef0ca2675ac29209df821181d1fe6fdd228b7

Observation 2659d829-7f28-4399-800b-dd48ea88e6e3 · outbound

This paper cites Robust an- swers, fragile logic: Probing the decoupling hypothesis in LLM reasoning.

Reinforced Preference Optimization for Reasoning-Augmented Recommendations Robust an- swers, fragile logic: Probing the decoupling hypothesis in LLM reasoning

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-22T04:34:35.321646Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-22T04:34:28.214871Z digest=sha256:f669b4abf3849b332bab63c977820b5128b5f542ac0d640a0f1d6d7fafd38413

Observation 2fa2599a-faac-4b51-b28d-2e89a9f98ff6 · outbound

This paper cites an unresolved cited work.

Reinforced Preference Optimization for Reasoning-Augmented Recommendations Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-05-22T04:34:35.613394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-22T04:34:28.214871Z digest=sha256:d225e88e6dab3b0dcce12a0d854a5d36c61d57da6a97573196f1ba3895f6da0a

Observation 8068571f-be73-4389-b2b7-98753addfac2 · outbound

This paper cites an unresolved cited work.

Reinforced Preference Optimization for Reasoning-Augmented Recommendations Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-05-22T04:34:35.579825Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-22T04:34:28.214871Z digest=sha256:e3ab0305de283793593944399351f4bda84bb1035a4caea85893fed1c3c88386

Observation a71ae4da-ef27-490b-8b5e-5c0f728ace8c · outbound

This paper cites Matryoshka Representation Learning.

Reinforced Preference Optimization for Reasoning-Augmented Recommendations Matryoshka Representation Learning

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-22T04:34:35.334737Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-22T04:34:28.214871Z digest=sha256:d92bc1217d72510ddf3b9e2c829828aa353f97e042fc3d2c5918048bc7be753f

Observation e560c52d-3dbd-4309-8453-02eeca58a23b · outbound

This paper cites an unresolved cited work.

Reinforced Preference Optimization for Reasoning-Augmented Recommendations Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-05-22T04:34:35.610232Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-22T04:34:28.214871Z digest=sha256:27eeaf78fc696243c004246c8f160ef1252aea3b351d5dca3f0b745588324649

Observation 29a0c014-2ac9-41fe-a2fd-2fc99007a334 · outbound

This paper cites Negative Sampling in Recommendation: A Survey and Future Directions.

Reinforced Preference Optimization for Reasoning-Augmented Recommendations Negative Sampling in Recommendation: A Survey and Future Directions

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-22T04:34:35.270610Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-22T04:34:28.214871Z digest=sha256:44441e4bb63ad00aa88a162bc693aa71abae801404ecc2cf6979e51ebcd549ad

Observation 4de7736d-e55f-4b0a-99c9-23f10ffb1615 · outbound

This paper cites an unresolved cited work.

Reinforced Preference Optimization for Reasoning-Augmented Recommendations Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-05-22T04:34:35.619806Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-22T04:34:28.214871Z digest=sha256:3e0e5ee935336870ae463a4ec6e1db2683e0df6e4885e59c2adfda11e0e7c828

Observation 38491342-bdbd-4ec5-b7cd-a11b5d6d6862 · outbound

This paper cites Large Language Models: A Survey.

Reinforced Preference Optimization for Reasoning-Augmented Recommendations Large Language Models: A Survey

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-05-22T04:34:35.303437Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-22T04:34:28.214871Z digest=sha256:d3216fd2c7b819c345ea42624ab97b203677c46ed6b34a9023e99c77d5909bef

Observation 68adb89a-32dd-45a2-979e-01430f224b7e · outbound

This paper cites an unresolved cited work.

Reinforced Preference Optimization for Reasoning-Augmented Recommendations Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-05-22T04:34:35.610516Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-22T04:34:28.214871Z digest=sha256:c4fd06bafb1ba97b31d411b8167bdb5587d5923a325f176e35c369146e063977

Observation aca46138-c3d3-4fd5-961c-80b292526e87 · outbound

This paper cites an unresolved cited work.

Reinforced Preference Optimization for Reasoning-Augmented Recommendations Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-05-22T04:34:35.607134Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-22T04:34:28.214871Z digest=sha256:6d0fa5be34bf160a62fe71025258a9e19e83d6263476798a84cc99edc35b24cd

Observation 155ca587-6ddc-430f-98cf-ed38594c2518 · outbound

This paper cites an unresolved cited work.

Reinforced Preference Optimization for Reasoning-Augmented Recommendations Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-05-22T04:34:35.628940Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-22T04:34:28.214871Z digest=sha256:8c8f06c146ca536c8c8cb597aa9857d13877d2e085eee90599fda182bc903c4a

Observation 90c7103b-4535-4d6a-87b9-04d1f60b08c2 · outbound

This paper cites Sentence-BERT: Sentence Embeddings using Siamese BERT-Networks.

Reinforced Preference Optimization for Reasoning-Augmented Recommendations Sentence-BERT: Sentence Embeddings using Siamese BERT-Networks

Reference 26

Resolution
verified exact
local_arxiv, observed 2026-05-22T04:34:35.302886Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-22T04:34:28.214871Z digest=sha256:363cb5b8b9860a078eaaeed8b9dcd4ae8be71692579d94740862c5934205793f

Observation 8c929808-e972-4543-9ba4-cb94893a9c7e · outbound

This paper cites DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models.

Reinforced Preference Optimization for Reasoning-Augmented Recommendations DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models

Reference 27

Resolution
verified exact
local_arxiv, observed 2026-05-22T04:34:35.292332Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-22T04:34:28.214871Z digest=sha256:3e2d9616c7b8de5a8505e4c6843a9fb1c2df76cc907067fb5b0efdcd29befac4

Observation 826cace2-7c79-463f-9b75-2ab0d318fffb · outbound

This paper cites an unresolved cited work.

Reinforced Preference Optimization for Reasoning-Augmented Recommendations Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-05-22T04:34:35.613128Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-22T04:34:28.214871Z digest=sha256:a191a657c543dd531e38fb81b38ea8539e25684f04fe713f558707dbe6288d64

Observation 5ea01efc-5219-4958-ab84-7e729d27be15 · outbound

This paper cites Tan, J., Chen, Y ., Zhang, A., Jiang, J., Liu, B., Xu, Z., Han, Z., Xu, J., Zheng, B., and Wang, X.

Reinforced Preference Optimization for Reasoning-Augmented Recommendations Tan, J., Chen, Y ., Zhang, A., Jiang, J., Liu, B., Xu, Z., Han, Z., Xu, J., Zheng, B., and Wang, X

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-22T04:34:35.297970Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-22T04:34:28.214871Z digest=sha256:15250a724c2f30e1b44b9978ad766b144a952be3556ac8ca1b28bcbac3480698

Observation d6919245-19b2-4eb5-bc18-c96a3a34c397 · outbound

This paper cites an unresolved cited work.

Reinforced Preference Optimization for Reasoning-Augmented Recommendations Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-05-22T04:34:35.623201Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-22T04:34:28.214871Z digest=sha256:14706d891adbd288e9549d017f42b81d59795cf8d4b5146af2fd9fe39a698892

Observation a542b79f-0794-449b-9830-106597d3e4fb · outbound

This paper cites Qwen3 Technical Report.

Reinforced Preference Optimization for Reasoning-Augmented Recommendations Qwen3 Technical Report

Reference 31

Resolution
verified exact
local_arxiv, observed 2026-05-22T04:34:35.340287Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-22T04:34:28.214871Z digest=sha256:0acb0164b2ed7a8ceaf6b7663408eaa9b4488867d808ba62bf6944724b7df0ff

Observation a5b41146-aca3-4a49-9d22-b267c1e72740 · outbound

This paper cites Can ChatGPT Defend its Belief in Truth? Evaluating LLM Reasoning via Debate.

Reinforced Preference Optimization for Reasoning-Augmented Recommendations Can ChatGPT Defend its Belief in Truth? Evaluating LLM Reasoning via Debate

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-05-22T04:34:35.313310Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-22T04:34:28.214871Z digest=sha256:3529ae2a7d933b407e5b1913c70ae5d711b0386a1c11b25e833339c2f72482aa

Observation 78498029-6139-4bda-bcb5-ee393b995962 · outbound

This paper cites Beyond the 80/20 Rule: High-Entropy Minority Tokens Drive Effective Reinforcement Learning for LLM Reasoning.

Reinforced Preference Optimization for Reasoning-Augmented Recommendations Beyond the 80/20 Rule: High-Entropy Minority Tokens Drive Effective Reinforcement Learning for LLM Reasoning

Reference 33

Resolution
verified exact
local_arxiv, observed 2026-05-22T04:34:35.331165Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-22T04:34:28.214871Z digest=sha256:2ef36c9cef810905b65ad61b7a4cf2003d0b427e0f347be3b135cd9dd746b0f4

Observation 76f3fad4-d774-49aa-9fb6-b623dc151575 · outbound

This paper cites an unresolved cited work.

Reinforced Preference Optimization for Reasoning-Augmented Recommendations Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-05-22T04:34:35.590852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-22T04:34:28.214871Z digest=sha256:430a41cac2feaea209a9ad5a9e4b094d9e3eebf69b96889d3565bb11ba8d5f34

Observation 97a94677-bf49-4a83-ab85-1eb4da7538b4 · outbound

This paper cites an unresolved cited work.

Reinforced Preference Optimization for Reasoning-Augmented Recommendations Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-05-22T04:34:35.587153Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-22T04:34:28.214871Z digest=sha256:a4f7e1e7072ceaab0bf7edcf05aab4b1db1c83812e193b6f05179d18c2c41baa

Observation cb9139e4-79cf-4715-bf7f-19cb79028ab3 · outbound

This paper cites InThe Thirty-ninth Annual Conference on Neural Information Processing Systems.

Reinforced Preference Optimization for Reasoning-Augmented Recommendations InThe Thirty-ninth Annual Conference on Neural Information Processing Systems

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T04:34:35.597472Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-22T04:34:28.214871Z digest=sha256:1c0c13ef075021ef150664d7bbd4cc03c60368770dae80354ac8a2d5ebeefc35

Observation 85c38152-423f-4c36-be77-ef95c20e66a4 · outbound

This paper cites an unresolved cited work.

Reinforced Preference Optimization for Reasoning-Augmented Recommendations Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-05-22T04:34:35.602173Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-22T04:34:28.214871Z digest=sha256:b40df52b3b9b0a5b342af0d2c503c954e30041a73423f9e65d5b9d538dadaa16

Observation b346361c-fef2-41cf-a97c-79e95925b47a · outbound

This paper cites arXiv preprint arXiv:2505.19092 (2025).

Reinforced Preference Optimization for Reasoning-Augmented Recommendations arXiv preprint arXiv:2505.19092 (2025)

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-05-22T04:34:35.317052Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-22T04:34:28.214871Z digest=sha256:61f0395008664d79b96932ccd40f510dfb53534ff2aaeacbe6fe188339992452

Observation 46ce351f-fc9a-4637-bcbd-71553d42bd4a · outbound

This paper cites A Survey of Large Language Models.

Reinforced Preference Optimization for Reasoning-Augmented Recommendations A Survey of Large Language Models

Reference 39

Resolution
verified exact
local_arxiv, observed 2026-05-22T04:34:35.326114Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-22T04:34:28.214871Z digest=sha256:3edf8657495f67d0e5822d66c14b07b71ce92be3735b8426b3a48af4771a19f3

Observation 92b5a338-b141-44c4-8a12-15f2d5fea5e2 · outbound

This paper cites an unresolved cited work.

Reinforced Preference Optimization for Reasoning-Augmented Recommendations Unresolved cited work

Reference 40

Resolution
unresolved
raw_fallback, observed 2026-05-22T04:34:35.631870Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-22T04:34:28.214871Z digest=sha256:fa02a40fe10b49bfeb9face11d09bc7d1b5d706eb3cef1042b1c482b0a4c838b

Observation 103e6b65-dfc0-48a7-9cfd-2afcbcd7790c · outbound

This paper cites an unresolved cited work.

Reinforced Preference Optimization for Reasoning-Augmented Recommendations Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-05-22T04:34:35.624459Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-22T04:34:28.214871Z digest=sha256:317374ac6da859a2de8e2a0d2f194ffef2f4c6aaad4601f1ff426b76ec629bdd

Observation 27ac8ec3-dbf5-4872-8332-31da2843c194 · outbound

This paper cites decision tokens.

Reinforced Preference Optimization for Reasoning-Augmented Recommendations decision tokens

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T04:34:35.599533Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-03T06:30:56.289259+00:00.

source=pdf_text observed=2026-05-22T04:34:28.214871Z digest=sha256:589927a886a87beee9365a3c6f3d29fdbf673fb2d418985e389f456b553b10fb

Pith citing papers

No inbound Pith citation observations are available.