Pith. sign in

Paper Citation Record · LEDGER

User Simulator-Guided Multi-Turn Preference Optimization for Reasoning LLM-based Conversational Recommendation

As of 10 August 2026, this Paper Citation Record lists 51 of 51 outbound references and 0 inbound Pith citation observations for arXiv:2604.03671.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2604.03671 v1

Coverage vector

measured 51 of 51 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-13T17:24:18.413112Z

measured 51 of 51 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

51 of 51 outbound references displayed

  • verified exact18
  • verified fuzzy1
  • unresolved32
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a60fa691-d7d3-4548-bf20-25b99dd0107a · outbound

This paper cites GPT-4 Technical Report.

User Simulator-Guided Multi-Turn Preference Optimization for Reasoning LLM-based Conversational Recommendation GPT-4 Technical Report

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-13T17:28:02.569739Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T17:24:18.413112Z digest=sha256:82462ec1f45bb9795a65b6f5407d03bd234557381f624f3fa3d9dbf4698fff99

Observation 308d31df-aaf2-4d81-b378-3aaff320362c · outbound

This paper cites Large Language Monkeys: Scaling Inference Compute with Repeated Sampling.

User Simulator-Guided Multi-Turn Preference Optimization for Reasoning LLM-based Conversational Recommendation Large Language Monkeys: Scaling Inference Compute with Repeated Sampling

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-05-13T17:28:02.579228Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T17:24:18.413112Z digest=sha256:bcd51030d99b83d1abc43e91861026001037a88add8749e5e7604c5ece61a4b8

Observation 2441cb29-c57c-4546-9ce9-fe5f29bfcb97 · outbound

This paper cites an unresolved cited work.

User Simulator-Guided Multi-Turn Preference Optimization for Reasoning LLM-based Conversational Recommendation Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-05-14T04:56:43.723712Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T17:24:18.413112Z digest=sha256:f3741c390f04e40edc9a67064178b4305e298746b54c468fe452d2e86b308a13

Observation b3e2838a-8fe2-4dc6-8776-128492e4ea0a · outbound

This paper cites Towards Knowledge-Based Recommender Dialog System.

User Simulator-Guided Multi-Turn Preference Optimization for Reasoning LLM-based Conversational Recommendation Towards Knowledge-Based Recommender Dialog System

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-13T17:28:02.552716Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T17:24:18.413112Z digest=sha256:83aab8bd9e1c3059a3bd3856bb7336057ab6ee5d76fa2d77e795fad937ecd036

Observation 5d786333-b7ba-4914-a7d6-2e516a273223 · outbound

This paper cites an unresolved cited work.

User Simulator-Guided Multi-Turn Preference Optimization for Reasoning LLM-based Conversational Recommendation Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-05-14T04:56:43.727070Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T17:24:18.413112Z digest=sha256:2da91f861dd7ecbe886584c4022e761f896c8675e7f9340350307fde009179ea

Observation f15adabc-27c2-4509-af79-f27e2c7af84c · outbound

This paper cites FlashAttention-2: Faster Attention with Better Parallelism and Work Partitioning.

User Simulator-Guided Multi-Turn Preference Optimization for Reasoning LLM-based Conversational Recommendation FlashAttention-2: Faster Attention with Better Parallelism and Work Partitioning

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-05-13T17:28:02.557958Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T17:24:18.413112Z digest=sha256:d94c7c70fc608a5777592a06528014923c4ba84c08676caa6c85713530524683

Observation 29704da0-1f17-411a-b3e5-ea26e103c6dc · outbound

This paper cites an unresolved cited work.

User Simulator-Guided Multi-Turn Preference Optimization for Reasoning LLM-based Conversational Recommendation Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-05-14T04:56:43.718017Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T17:24:18.413112Z digest=sha256:6c5a83aef71f70396d73e63723a069025be90de9b824e75ab437ca1efda21d3b

Observation 5a1029c8-16fb-4d90-b428-0600ef36fe00 · outbound

This paper cites an unresolved cited work.

User Simulator-Guided Multi-Turn Preference Optimization for Reasoning LLM-based Conversational Recommendation Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-05-14T04:56:43.711573Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T17:24:18.413112Z digest=sha256:c624ca79ee1da91b961fa6c8930ad19f8a7090f02267da3d958b154a44e7a270

Observation f81f69ca-3d6b-4a5d-b60b-525a22a58f33 · outbound

This paper cites an unresolved cited work.

User Simulator-Guided Multi-Turn Preference Optimization for Reasoning LLM-based Conversational Recommendation Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-05-14T04:56:43.714590Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T17:24:18.413112Z digest=sha256:02f4e30479a18d1e29e2c23d6b3af4d491027d35116ea32e3a0dbc4aab50ab0d

Observation e13b5a21-6c3f-4a8b-b10c-a3ab2200aa21 · outbound

This paper cites an unresolved cited work.

User Simulator-Guided Multi-Turn Preference Optimization for Reasoning LLM-based Conversational Recommendation Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-05-14T04:56:43.720897Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T17:24:18.413112Z digest=sha256:a742f7df32d6e0f472edd149043ca2218fdb30f5a6736c618b0c2fdcb9bfd453

Observation 47e25540-5bab-4112-a26a-419275d6d988 · outbound

This paper cites Reason4Rec: Deliberative User Preference Alignment of Large Language Models for Recommendation.

User Simulator-Guided Multi-Turn Preference Optimization for Reasoning LLM-based Conversational Recommendation Reason4Rec: Deliberative User Preference Alignment of Large Language Models for Recommendation

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-07-29T01:24:22.003827Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T17:24:18.413112Z digest=sha256:68af45dbdb98fb5326d4f51eed8058e6385dcf830f1ccf6c2c00614eec8b2701

Observation a74804a0-4f5c-410d-b940-ffe5f5dd471a · outbound

This paper cites Expectation Confirmation Preference Optimization for Multi-Turn Conversational Recommendation Agent.

User Simulator-Guided Multi-Turn Preference Optimization for Reasoning LLM-based Conversational Recommendation Expectation Confirmation Preference Optimization for Multi-Turn Conversational Recommendation Agent

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-13T17:28:02.593141Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T17:24:18.413112Z digest=sha256:5d0b399f318ecd8b817071c60248e0712c21b109e8d07d003aec56ac9409f1f6

Observation ebf2ccc6-22f5-4fad-a687-e7dca2827336 · outbound

This paper cites DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning.

User Simulator-Guided Multi-Turn Preference Optimization for Reasoning LLM-based Conversational Recommendation DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning

Reference 13

Resolution
verified exact
local_arxiv, observed 2026-05-13T17:28:02.590918Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T17:24:18.413112Z digest=sha256:d008dea9d8d76adbf9cda52050f0b76996fbc4710c4277f0c613c9b3f6a9e0e5

Observation e8eb26db-3adb-4d28-baf9-2884686d27f7 · outbound

This paper cites INSPIRED: Toward Sociable Recommendation Dialog Systems.

User Simulator-Guided Multi-Turn Preference Optimization for Reasoning LLM-based Conversational Recommendation INSPIRED: Toward Sociable Recommendation Dialog Systems

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-13T17:28:02.572292Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T17:24:18.413112Z digest=sha256:0697c4412b07e25d1d6232e52d8719266a2550dae13b37780a1697987de16a35

Observation dfd277be-dbcf-4136-b0e7-91674b615efc · outbound

This paper cites an unresolved cited work.

User Simulator-Guided Multi-Turn Preference Optimization for Reasoning LLM-based Conversational Recommendation Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-05-14T04:56:43.730200Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T17:24:18.413112Z digest=sha256:e4fe7e1cff7288cda007476801653768088d27bb23c1f8879d33ea9c5628034f

Observation 273694fc-4406-4c8c-9ee7-080dc8823535 · outbound

This paper cites an unresolved cited work.

User Simulator-Guided Multi-Turn Preference Optimization for Reasoning LLM-based Conversational Recommendation Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-05-14T05:01:59.273800Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T17:24:18.413112Z digest=sha256:20164a2f83ed23622e38d7f37797aeeb107f484ca283e314bfe790f2ca4f12a5

Observation a68ca345-ebf7-4000-bf75-0d05b68b4b7f · outbound

This paper cites an unresolved cited work.

User Simulator-Guided Multi-Turn Preference Optimization for Reasoning LLM-based Conversational Recommendation Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-05-14T04:56:43.733060Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T17:24:18.413112Z digest=sha256:6844965acdeb8a70109fbff5e5ba161f34c1aa103a00321fdbb09c4fb8530fb1

Observation 0cb8bc96-56d7-43bb-9909-04af1a13759b · outbound

This paper cites Understanding the planning of LLM agents: A survey.

User Simulator-Guided Multi-Turn Preference Optimization for Reasoning LLM-based Conversational Recommendation Understanding the planning of LLM agents: A survey

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-13T18:12:57.765583Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T17:24:18.413112Z digest=sha256:0b0b451dc1e0e935390476d12f75cb9e8939dd7542d49ec62f7a1afd87510705

Observation 028a9c3a-c79d-42a0-9d83-ea8a096f1963 · outbound

This paper cites an unresolved cited work.

User Simulator-Guided Multi-Turn Preference Optimization for Reasoning LLM-based Conversational Recommendation Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-05-14T04:56:43.705207Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T17:24:18.413112Z digest=sha256:d8295c1b1d52b6651ebcbad5dfcfccfc1a6660630de89cf0c57cb5bbb5e6bf21

Observation 8ba94241-7c58-4434-9cae-1bec55bd9d4d · outbound

This paper cites an unresolved cited work.

User Simulator-Guided Multi-Turn Preference Optimization for Reasoning LLM-based Conversational Recommendation Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-05-14T04:56:43.676500Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T17:24:18.413112Z digest=sha256:9d5484cdadf0421947d89ba22b8451107c89ba6e74a76c6e71ca7d3b1c8dc0b4

Observation 01a2f84b-3799-4cbb-817f-7a4bccc0cee1 · outbound

This paper cites an unresolved cited work.

User Simulator-Guided Multi-Turn Preference Optimization for Reasoning LLM-based Conversational Recommendation Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-05-14T04:56:43.693326Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T17:24:18.413112Z digest=sha256:b7afe43cfad2c0706a29c2abc67f8f1f5a9e96bfbbf0f0d7d2bf9da60286e317

Observation d5aa0ac5-1723-4efb-b005-9d19a6fef7b4 · outbound

This paper cites TREA: Tree-Structure Reasoning Schema for Conversational Recommendation.

User Simulator-Guided Multi-Turn Preference Optimization for Reasoning LLM-based Conversational Recommendation TREA: Tree-Structure Reasoning Schema for Conversational Recommendation

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-13T17:28:02.574517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T17:24:18.413112Z digest=sha256:f3b80b6a37b912bc5a1a1848ca9c2d34a8703d9bf2dcc9bdc6080e37a6f0cfb1

Observation c25b82a4-19bd-4c0f-bbc1-44de6444d413 · outbound

This paper cites Interpretable User Satisfaction Estimation for Conversational Systems with Large Language Models.

User Simulator-Guided Multi-Turn Preference Optimization for Reasoning LLM-based Conversational Recommendation Interpretable User Satisfaction Estimation for Conversational Systems with Large Language Models

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-13T17:28:02.576902Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T17:24:18.413112Z digest=sha256:3a091e85ffb04b47669b54e352c44b5935f0c5f9b9e83fd85ab56a252bb1648f

Observation 6e51fa45-69be-4f1b-8f70-e528df3c5b2c · outbound

This paper cites an unresolved cited work.

User Simulator-Guided Multi-Turn Preference Optimization for Reasoning LLM-based Conversational Recommendation Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-05-14T04:56:43.689989Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T17:24:18.413112Z digest=sha256:286efaa95fcf4d13f7b733f78c457dd6baabb80eacb5a2443bcfae709bfb705a

Observation d72ba797-37f9-49cd-a004-493f21d9fa47 · outbound

This paper cites an unresolved cited work.

User Simulator-Guided Multi-Turn Preference Optimization for Reasoning LLM-based Conversational Recommendation Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-05-14T04:56:43.708652Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T17:24:18.413112Z digest=sha256:0bb5a4c5ce42b509a826acd299006a64b0d3393ec196428488f8cfac7aa19bf0

Observation ec9dcb02-e7e4-4e58-b2df-6e07fa1ba4e6 · outbound

This paper cites Kimi k1.5: Scaling Reinforcement Learning with LLMs.

User Simulator-Guided Multi-Turn Preference Optimization for Reasoning LLM-based Conversational Recommendation Kimi k1.5: Scaling Reinforcement Learning with LLMs

Reference 26

Resolution
verified exact
local_arxiv, observed 2026-05-13T17:28:02.588760Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T17:24:18.413112Z digest=sha256:e4ee6ca5e34ae6258cedc8207f1d24a69d9a2c0562f5bfa88219b03d76e70726

Observation f041ecf3-85dc-44c3-8e28-70dd0af20441 · outbound

This paper cites an unresolved cited work.

User Simulator-Guided Multi-Turn Preference Optimization for Reasoning LLM-based Conversational Recommendation Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-05-14T04:56:43.653029Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T17:24:18.413112Z digest=sha256:3cb315b5454c0e00a6360a99e8d0af5c7c5898b18a18f3dfde2328fef7fa1195

Observation d77f9783-2dc2-4850-baaf-cc2c74673aaa · outbound

This paper cites Rethinking the Evaluation for Conversational Recommendation in the Era of Large Language Models.

User Simulator-Guided Multi-Turn Preference Optimization for Reasoning LLM-based Conversational Recommendation Rethinking the Evaluation for Conversational Recommendation in the Era of Large Language Models

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-13T17:28:02.584217Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T17:24:18.413112Z digest=sha256:12e0da12276a54c01e0d6b0ae78426fcfd44b9070518c9b1ef1f348347aabb5b

Observation 1cdb594e-04ff-4cd8-afe3-15fafc75e9e2 · outbound

This paper cites Self-Consistency Improves Chain of Thought Reasoning in Language Models.

User Simulator-Guided Multi-Turn Preference Optimization for Reasoning LLM-based Conversational Recommendation Self-Consistency Improves Chain of Thought Reasoning in Language Models

Reference 29

Resolution
verified exact
local_arxiv, observed 2026-05-13T17:28:02.586506Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T17:24:18.413112Z digest=sha256:4a46024571ad53ad7cae81a4aca4aa4021a6c6940cb6a2e10c203421429a5536

Observation 599af5dd-af9b-46c8-b52c-4d045d7528ab · outbound

This paper cites an unresolved cited work.

User Simulator-Guided Multi-Turn Preference Optimization for Reasoning LLM-based Conversational Recommendation Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-05-14T04:56:43.659872Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T17:24:18.413112Z digest=sha256:ee8c0ace2dba590b76b85e0000212a285a0b2b46616fd2103ec0dfd0da8cb7be

Observation 5dcbc4ec-2c5f-4170-93fa-581ddfc3717b · outbound

This paper cites an unresolved cited work.

User Simulator-Guided Multi-Turn Preference Optimization for Reasoning LLM-based Conversational Recommendation Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-05-14T05:01:59.281505Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T17:24:18.413112Z digest=sha256:2351fea9f595256bf048c664bfdf20399d5032645feb85dc0992fb6a40ffe213

Observation e2491c63-1967-43b0-8347-be89ac3d190d · outbound

This paper cites an unresolved cited work.

User Simulator-Guided Multi-Turn Preference Optimization for Reasoning LLM-based Conversational Recommendation Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-05-14T05:01:59.284471Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T17:24:18.413112Z digest=sha256:c9393f9eaa992e96b42ae9d6eeb8c6e337d02705cace0361e97d10616ab4e487

Observation a5a9a7be-77f1-4bca-86fb-a31139d56565 · outbound

This paper cites an unresolved cited work.

User Simulator-Guided Multi-Turn Preference Optimization for Reasoning LLM-based Conversational Recommendation Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-05-14T04:56:43.662854Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T17:24:18.413112Z digest=sha256:d5a4c650b79bb06cb05f46deb11428fb503f0c04498290922fc788d5b0da823a

Observation 987cde7b-0c35-494a-b4db-787440a10770 · outbound

This paper cites an unresolved cited work.

User Simulator-Guided Multi-Turn Preference Optimization for Reasoning LLM-based Conversational Recommendation Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-05-14T05:01:59.290867Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T17:24:18.413112Z digest=sha256:3b33cdae19f808e7a2dab97fda177a6f56198cc31fdd5039c619098406dfdf03

Observation 5462c1e8-77b5-4213-a0e0-39e99fb4b8ea · outbound

This paper cites InProceedings of the 48th International ACM SIGIR Conference on Research and Development in Information Retrieval.

User Simulator-Guided Multi-Turn Preference Optimization for Reasoning LLM-based Conversational Recommendation InProceedings of the 48th International ACM SIGIR Conference on Research and Development in Information Retrieval

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T04:56:43.666184Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T17:24:18.413112Z digest=sha256:e0e5524b3b81486b90e9d64d09689778bf1e9fd8ddec7921780866b1a9c382ee

Observation 2f9f2e82-a7f4-48a7-907b-55993568d688 · outbound

This paper cites an unresolved cited work.

User Simulator-Guided Multi-Turn Preference Optimization for Reasoning LLM-based Conversational Recommendation Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-05-14T04:56:43.669291Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T17:24:18.413112Z digest=sha256:9d1ff36c5a15a3149e55abd896f9e0945b288e70e1b333d04b364c6e66a4b786

Observation 6eb9c1a2-906f-4942-9764-f8fa503dcee6 · outbound

This paper cites an unresolved cited work.

User Simulator-Guided Multi-Turn Preference Optimization for Reasoning LLM-based Conversational Recommendation Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-05-14T04:56:43.683595Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T17:24:18.413112Z digest=sha256:e95850219b12ed533565377dd2ab472e3ac0d389c3039f1fa4c72fc6e9ebb497

Observation aee1599b-a417-4a51-b6b5-aa66acf5a5d4 · outbound

This paper cites an unresolved cited work.

User Simulator-Guided Multi-Turn Preference Optimization for Reasoning LLM-based Conversational Recommendation Unresolved cited work

Reference 38

Resolution
unresolved
raw_fallback, observed 2026-05-14T04:56:43.702370Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T17:24:18.413112Z digest=sha256:e9858f84c905b5a88b2414275807db55cead4e0aed944ace4a2e3ecea6c0c529

Observation a40a7d97-b9e0-450e-8449-3258e35e200a · outbound

This paper cites an unresolved cited work.

User Simulator-Guided Multi-Turn Preference Optimization for Reasoning LLM-based Conversational Recommendation Unresolved cited work

Reference 39

Resolution
unresolved
raw_fallback, observed 2026-05-14T04:56:43.699182Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T17:24:18.413112Z digest=sha256:6fe5b9da68086a0de8b4e87278cfc9ed35a1b074c2a0502c3da4d888460c49a0

Observation 566cb346-df81-458b-8619-2ded5a40db4b · outbound

This paper cites an unresolved cited work.

User Simulator-Guided Multi-Turn Preference Optimization for Reasoning LLM-based Conversational Recommendation Unresolved cited work

Reference 40

Resolution
unresolved
raw_fallback, observed 2026-05-14T05:01:59.287256Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T17:24:18.413112Z digest=sha256:f3a098afcb796130296fb0c12a42a9e13a6cabe388b453ac2fd185d930db4233

Observation f256212c-9404-4c44-aa3a-dc9348ff74c4 · outbound

This paper cites an unresolved cited work.

User Simulator-Guided Multi-Turn Preference Optimization for Reasoning LLM-based Conversational Recommendation Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-05-14T05:01:59.278566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T17:24:18.413112Z digest=sha256:36d7f88fd51c432e02085c8cee04a91320f99f6dd7820f30c600fb76293cea69

Observation 951d67df-b2b8-4798-8972-b540732cfc64 · outbound

This paper cites Evaluating Large Language Models as Generative User Simulators for Conversational Recommendation.

User Simulator-Guided Multi-Turn Preference Optimization for Reasoning LLM-based Conversational Recommendation Evaluating Large Language Models as Generative User Simulators for Conversational Recommendation

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-13T17:28:02.562737Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T17:24:18.413112Z digest=sha256:1ac7994d75385fd9b2764fa33b324b3ea4af0787023341489a158591fa220940

Observation 99d4ef0e-2021-485e-9540-51c26da98b08 · outbound

This paper cites DAPO: An Open-Source LLM Reinforcement Learning System at Scale.

User Simulator-Guided Multi-Turn Preference Optimization for Reasoning LLM-based Conversational Recommendation DAPO: An Open-Source LLM Reinforcement Learning System at Scale

Reference 43

Resolution
verified exact
local_arxiv, observed 2026-05-13T17:28:02.560347Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T17:24:18.413112Z digest=sha256:11ee23ce73766aafd57cbca28aed154bf04deb595457c8ff42f86f38d831db5c

Observation 191891a1-b127-4f5c-8b5d-3b0e56f91998 · outbound

This paper cites VAPO: Efficient and Reliable Reinforcement Learning for Advanced Reasoning Tasks.

User Simulator-Guided Multi-Turn Preference Optimization for Reasoning LLM-based Conversational Recommendation VAPO: Efficient and Reliable Reinforcement Learning for Advanced Reasoning Tasks

Reference 44

Resolution
verified exact
local_arxiv, observed 2026-05-13T17:28:02.564997Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T17:24:18.413112Z digest=sha256:f7e45fc9b374bddcb5e7bb4f16ec556baefef046b00ed5f8a62349dc29803722

Observation 575f6653-3aac-45bc-87bc-eda8e7843a18 · outbound

This paper cites an unresolved cited work.

User Simulator-Guided Multi-Turn Preference Optimization for Reasoning LLM-based Conversational Recommendation Unresolved cited work

Reference 45

Resolution
unresolved
raw_fallback, observed 2026-05-14T04:56:43.680539Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T17:24:18.413112Z digest=sha256:955983997a7a530d85bbf2f54a6b8d5a5f9c0e5cc08e75e5757f819797566eed

Observation c4006421-143f-4249-91f7-446dfd205180 · outbound

This paper cites Reason-to-Recommend: Using Interaction-of-Thought Reasoning to Enhance LLM Recommendation.

User Simulator-Guided Multi-Turn Preference Optimization for Reasoning LLM-based Conversational Recommendation Reason-to-Recommend: Using Interaction-of-Thought Reasoning to Enhance LLM Recommendation

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-13T17:28:02.567678Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T17:24:18.413112Z digest=sha256:95c1f3fe77e4c3367a02ca6d1ceaab14c955d760b6dcb5c232c99a9e94e66173

Observation 110de21b-7898-49fc-a074-021d80da7795 · outbound

This paper cites an unresolved cited work.

User Simulator-Guided Multi-Turn Preference Optimization for Reasoning LLM-based Conversational Recommendation Unresolved cited work

Reference 47

Resolution
unresolved
raw_fallback, observed 2026-05-14T04:56:43.672432Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T17:24:18.413112Z digest=sha256:80d93fd92f3e3c305abd2501f9f90a0d5cd6e93c01c5d8a8fb1d332dbb047dd6

Observation 734828b7-084b-453c-8296-568c59153c8a · outbound

This paper cites an unresolved cited work.

User Simulator-Guided Multi-Turn Preference Optimization for Reasoning LLM-based Conversational Recommendation Unresolved cited work

Reference 48

Resolution
unresolved
raw_fallback, observed 2026-05-14T04:56:43.686825Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T17:24:18.413112Z digest=sha256:919b1b3912f3a6f353d99f09b5f7004b3e45a02a757d61ab99f1317eca3c1933

Observation bade24d0-dd13-460d-bdb9-e6ebbdbc3ffd · outbound

This paper cites an unresolved cited work.

User Simulator-Guided Multi-Turn Preference Optimization for Reasoning LLM-based Conversational Recommendation Unresolved cited work

Reference 49

Resolution
unresolved
raw_fallback, observed 2026-05-14T04:56:43.696281Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T17:24:18.413112Z digest=sha256:a5f8c3c777a2bee8f9b0ffa1769fc1cac4598bcca30cb5a6fced72ca5bba0f66

Observation 5a731c13-8716-478f-916c-04025695bcbe · outbound

This paper cites an unresolved cited work.

User Simulator-Guided Multi-Turn Preference Optimization for Reasoning LLM-based Conversational Recommendation Unresolved cited work

Reference 50

Resolution
unresolved
raw_fallback, observed 2026-05-14T04:56:43.649242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T17:24:18.413112Z digest=sha256:d86cb0752c13565cfd80bfa04f41b1367b7ab65f56dca74e39474803198f4d21

Observation 742b147e-2e1a-43b2-b6be-ec0df37d51f3 · outbound

This paper cites an unresolved cited work.

User Simulator-Guided Multi-Turn Preference Optimization for Reasoning LLM-based Conversational Recommendation Unresolved cited work

Reference 51

Resolution
unresolved
raw_fallback, observed 2026-05-14T04:56:43.656379Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T17:24:18.413112Z digest=sha256:66071643f5e6fd546e7da05f49e5bc997ee703649cef6a3ba401b9dfed3ff386

Pith citing papers

No inbound Pith citation observations are available.