Pith. sign in

Paper Citation Record · LEDGER

DialogXpert: Driving Intelligent and Emotion-Aware Conversations through Online Value-Based Reinforcement Learning with LLM Priors

As of 15 August 2026, this Paper Citation Record lists 16 of 16 outbound references and 0 inbound Pith citation observations for arXiv:2505.17795.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.17795 v1

Coverage vector

measured 16 of 16 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:45:05.543512Z

measured 16 of 16 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

16 of 16 outbound references displayed

  • verified exact0
  • verified fuzzy3
  • unresolved11
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch2

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation e4d50b19-52a5-435b-8b5f-f5d45e776197 · outbound

This paper cites an unresolved cited work.

DialogXpert: Driving Intelligent and Emotion-Aware Conversations through Online Value-Based Reinforcement Learning with LLM Priors Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:45:08.800899Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:45:04.360778Z digest=sha256:6861a003e0d462c9f0ef4e4386e8ec78fa907ff49cea75e432cea8db9a973138

Observation 0e985aac-c9b1-418d-a185-fbe73d840e2e · outbound

This paper cites an unresolved cited work.

DialogXpert: Driving Intelligent and Emotion-Aware Conversations through Online Value-Based Reinforcement Learning with LLM Priors Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:45:08.563763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:45:04.487068Z digest=sha256:cf9aca44b7973c63b12d8b9095a8e4315bee1b135c3dd3af894e44d0c03fa323

Observation 6274500f-ab8b-429d-81c3-47f04abfe6ea · outbound

This paper cites an unresolved cited work.

DialogXpert: Driving Intelligent and Emotion-Aware Conversations through Online Value-Based Reinforcement Learning with LLM Priors Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:45:08.307620Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:45:04.636802Z digest=sha256:ca7d0b7317ff559c426c53fa018ce861fb7129f50b181a8490255990bb5ccc41

Observation c05cc2ba-7616-497e-b903-44e2ed90b435 · outbound

This paper cites Reasoning with Language Model is Planning with World Model.

DialogXpert: Driving Intelligent and Emotion-Aware Conversations through Online Value-Based Reinforcement Learning with LLM Priors Reasoning with Language Model is Planning with World Model

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:03.683850Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:45:03.683850Z digest=sha256:23f521190560767971279cb4ec11b85dd7adec0e1703fd2cbf92641aea1c6583

Observation c64af7e9-f5b9-4365-945e-59f190d73473 · outbound

This paper cites Direct Preference Optimization: Your Language Model is Secretly a Reward Model.

DialogXpert: Driving Intelligent and Emotion-Aware Conversations through Online Value-Based Reinforcement Learning with LLM Priors Direct Preference Optimization: Your Language Model is Secretly a Reward Model

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:03.812358Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:45:03.812358Z digest=sha256:db0d3b38d5f90a8d6cb2c62d9fe36f2604a0f8681f1fd651665b1ee436547e22

Observation d3bc0cf5-81c3-482d-b345-59baec77a584 · outbound

This paper cites Conversational Tree Search: A New Hybrid Dialog Task.

DialogXpert: Driving Intelligent and Emotion-Aware Conversations through Online Value-Based Reinforcement Learning with LLM Priors Conversational Tree Search: A New Hybrid Dialog Task

Reference 7

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T14:45:05.841278Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:45:04.027094Z digest=sha256:946926b591b8131d4bc55015f366adafbffb71c6c25b602e3a62efd61acb9c21

Observation fdf8747c-3921-4203-8758-66c9266eb4a4 · outbound

This paper cites ProAgent: Building Proactive Cooperative Agents with Large Language Models.

DialogXpert: Driving Intelligent and Emotion-Aware Conversations through Online Value-Based Reinforcement Learning with LLM Priors ProAgent: Building Proactive Cooperative Agents with Large Language Models

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:04.170030Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:45:04.170030Z digest=sha256:f78bb0d6d60e464267fc535d7e1deb075bd2b61ee329a70b430f3e7ec030c0d9

Observation a9534df6-4938-44d9-a324-a86b09c210e0 · outbound

This paper cites Your responses are auto-saved after each item.

DialogXpert: Driving Intelligent and Emotion-Aware Conversations through Online Value-Based Reinforcement Learning with LLM Priors Your responses are auto-saved after each item

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:45:08.023757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:45:04.823665Z digest=sha256:376cc516ebc7cdb836ade6794274b215bb3f90af04b68b79f0e2320140fad10f

Observation 899b74b1-5d88-4a7c-aae6-6d65745cff9b · outbound

This paper cites an unresolved cited work.

DialogXpert: Driving Intelligent and Emotion-Aware Conversations through Online Value-Based Reinforcement Learning with LLM Priors Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:45:07.642462Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:45:04.965029Z digest=sha256:18e602865d8632ac5687132c87be955c92ff2d85bd3fa0f55693f17789bf63ae

Observation faca043b-f8b3-4c1c-a2aa-c8aaf0e8f077 · outbound

This paper cites an unresolved cited work.

DialogXpert: Driving Intelligent and Emotion-Aware Conversations through Online Value-Based Reinforcement Learning with LLM Priors Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:45:07.331720Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:45:05.127621Z digest=sha256:330cf9922fd49fd03b1807c407634f3286a9af65ef8035c745fe677eda63c92d

Observation d136a87b-5703-47cf-b4b5-a6e16fbb9c3c · outbound

This paper cites Responses are auto-saved; log back in with the sameUser IDto resume.

DialogXpert: Driving Intelligent and Emotion-Aware Conversations through Online Value-Based Reinforcement Learning with LLM Priors Responses are auto-saved; log back in with the sameUser IDto resume

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:45:06.937990Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:45:05.325415Z digest=sha256:5572d130cee9d7de3e7ed1cf144f1cc163d721dfbebe0a7800577fe396d968c1

Observation ce2da252-b646-4b40-a78f-396f6d175fa7 · outbound

This paper cites Policy LLM for {dataset}.

DialogXpert: Driving Intelligent and Emotion-Aware Conversations through Online Value-Based Reinforcement Learning with LLM Priors Policy LLM for {dataset}

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:45:06.557148Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:45:05.543512Z digest=sha256:115af6f8ed664ae9dafbea294af52921ad28c9095aea78e0fe7ecc6d3b454ed1

Observation 889ae298-7e7d-4bc6-9be1-c3038f25c6af · outbound

This paper cites Generating Emotionally Aligned Responses in Dialogues using Affect Control Theory.

DialogXpert: Driving Intelligent and Emotion-Aware Conversations through Online Value-Based Reinforcement Learning with LLM Priors Generating Emotionally Aligned Responses in Dialogues using Affect Control Theory

Reference 2017

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T14:45:06.166992Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:45:03.394545Z digest=sha256:0fa2fa430aab5b501faf96dfe2b3a36c5e2bb7b19443c42d13e6eea87fc83cde

Observation 8521ef60-4d2c-471f-876a-f7578092d8b4 · outbound

This paper cites LLaMA: Open and Efficient Foundation Language Models.

DialogXpert: Driving Intelligent and Emotion-Aware Conversations through Online Value-Based Reinforcement Learning with LLM Priors LLaMA: Open and Efficient Foundation Language Models

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:03.927445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:45:03.927445Z digest=sha256:c75e6048433cc6e5792741786d2400cc137d4fa13f14861a4111da7fa98be93f

Observation 3f8f2f33-fe71-46f3-9bfa-30c40a2c1e0d · outbound

This paper cites Improving Multi-turn Emotional Support Dialogue Generation with Lookahead Strategy Planning.

DialogXpert: Driving Intelligent and Emotion-Aware Conversations through Online Value-Based Reinforcement Learning with LLM Priors Improving Multi-turn Emotional Support Dialogue Generation with Lookahead Strategy Planning

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:03.479133Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:45:03.479133Z digest=sha256:1eaf11b09799f460fae31928bb5d700fe3e16bb26ee43ad88d1b1e5bc0142202

Observation 1cb7bb75-3b39-4219-b15f-a9e1e2969933 · outbound

This paper cites Improving Language Model Negotiation with Self-Play and In-Context Learning from AI Feedback.

DialogXpert: Driving Intelligent and Emotion-Aware Conversations through Online Value-Based Reinforcement Learning with LLM Priors Improving Language Model Negotiation with Self-Play and In-Context Learning from AI Feedback

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-07T14:45:03.591088Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:45:03.591088Z digest=sha256:eadf31ae3bd71794f11bdcd5db24c0db6b967efdca3fa380ead69ab8cb3a7901

Pith citing papers

No inbound Pith citation observations are available.