Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T17:23:08.276030Z
Paper Citation Record · LEDGER
As of 17 August 2026, this Paper Citation Record lists 44 of 44 outbound references and 2 inbound Pith citation observations for arXiv:2508.12935.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T17:23:08.276030Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-02T18:44:27.881977Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-15T15:26:10.640839Z
44 of 44 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 86142871-be70-42b5-863d-861233250486 · outbound
Towards Open-Ended Emotional Support Conversations in LLMs via Reinforcement Learning with Future-Oriented Rewards GPT-4 Technical Report
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 22133c54-a5ca-4898-8259-ed464fc5fc70 · outbound
Towards Open-Ended Emotional Support Conversations in LLMs via Reinforcement Learning with Future-Oriented Rewards Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eb9187bb-afa4-4aac-97b4-eb4a020e0984 · outbound
Towards Open-Ended Emotional Support Conversations in LLMs via Reinforcement Learning with Future-Oriented Rewards Unresolved cited work
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 6f92f9a6-7bb7-4145-8cef-ef8230d6a9fc · outbound
Towards Open-Ended Emotional Support Conversations in LLMs via Reinforcement Learning with Future-Oriented Rewards Unresolved cited work
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 799b3054-db43-459c-ba1e-3f497f9a17ac · outbound
Towards Open-Ended Emotional Support Conversations in LLMs via Reinforcement Learning with Future-Oriented Rewards Unresolved cited work
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b7474df8-fa4d-4cbd-b7e8-d30f47a87f9f · outbound
Towards Open-Ended Emotional Support Conversations in LLMs via Reinforcement Learning with Future-Oriented Rewards Unresolved cited work
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7065042c-150b-43ee-a90f-be4e7ae89df7 · outbound
Towards Open-Ended Emotional Support Conversations in LLMs via Reinforcement Learning with Future-Oriented Rewards Unresolved cited work
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 4d73e047-4165-43db-815a-10a4259521c9 · outbound
Towards Open-Ended Emotional Support Conversations in LLMs via Reinforcement Learning with Future-Oriented Rewards Unresolved cited work
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 77054278-3ed1-4475-acf6-e21c46c2e358 · outbound
Towards Open-Ended Emotional Support Conversations in LLMs via Reinforcement Learning with Future-Oriented Rewards Unresolved cited work
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 69c73f77-ab89-469f-9257-109b98d8329c · outbound
Towards Open-Ended Emotional Support Conversations in LLMs via Reinforcement Learning with Future-Oriented Rewards Unresolved cited work
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 4e3de397-e2be-4b3b-930d-230dae3b5518 · outbound
Towards Open-Ended Emotional Support Conversations in LLMs via Reinforcement Learning with Future-Oriented Rewards Knowledge-enhanced Mixed-initiative Dialogue System for Emotional Support Conversations
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fb01df84-3482-4719-a149-5e1684d0ec64 · outbound
Towards Open-Ended Emotional Support Conversations in LLMs via Reinforcement Learning with Future-Oriented Rewards Improving Language Model Negotiation with Self-Play and In-Context Learning from AI Feedback
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f6b9070e-4092-4fbe-ba4d-96e8bac795f7 · outbound
Towards Open-Ended Emotional Support Conversations in LLMs via Reinforcement Learning with Future-Oriented Rewards The Llama 3 Herd of Models
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8af4dda2-87dc-4eaa-ae11-11d2392c3da6 · outbound
Towards Open-Ended Emotional Support Conversations in LLMs via Reinforcement Learning with Future-Oriented Rewards A Survey on LLM-as-a-Judge
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ab61585c-c003-48e3-8b30-9ce5c0689137 · outbound
Towards Open-Ended Emotional Support Conversations in LLMs via Reinforcement Learning with Future-Oriented Rewards DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fc976764-be01-4850-a878-1f1dcd58c13e · outbound
Towards Open-Ended Emotional Support Conversations in LLMs via Reinforcement Learning with Future-Oriented Rewards Unresolved cited work
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0c039649-772b-4577-ad6a-d68682f169db · outbound
Towards Open-Ended Emotional Support Conversations in LLMs via Reinforcement Learning with Future-Oriented Rewards Unresolved cited work
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation b7a766de-bd57-435d-bcea-2c346f7ce583 · outbound
Towards Open-Ended Emotional Support Conversations in LLMs via Reinforcement Learning with Future-Oriented Rewards Knowledge-enhanced Memory Model for Emotional Support Conversation
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 72cbd58d-bebe-4ffc-b2fa-deb17b876f24 · outbound
Towards Open-Ended Emotional Support Conversations in LLMs via Reinforcement Learning with Future-Oriented Rewards Search-R1: Training LLMs to Reason and Leverage Search Engines with Reinforcement Learning
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7367b818-af0c-40c6-a40a-74afbe1468d8 · outbound
Towards Open-Ended Emotional Support Conversations in LLMs via Reinforcement Learning with Future-Oriented Rewards Unresolved cited work
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6f2ca1e8-b7f5-4d1e-a4d3-d76284b3bc0b · outbound
Towards Open-Ended Emotional Support Conversations in LLMs via Reinforcement Learning with Future-Oriented Rewards Unresolved cited work
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 52ceacb1-ceea-4f53-9b07-833cfc9aa726 · outbound
Towards Open-Ended Emotional Support Conversations in LLMs via Reinforcement Learning with Future-Oriented Rewards Unresolved cited work
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 4cc495fb-8272-4291-906b-d462379d8bce · outbound
Towards Open-Ended Emotional Support Conversations in LLMs via Reinforcement Learning with Future-Oriented Rewards Control Globally, Understand Locally: A Global-to-Local Hierarchical Graph Network for Emotional Support Conversation
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a05c2539-e36f-4540-8710-2ef59ed44112 · outbound
Towards Open-Ended Emotional Support Conversations in LLMs via Reinforcement Learning with Future-Oriented Rewards Unresolved cited work
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9c220708-3d9d-47d9-9a84-2c92d32cd5c6 · outbound
Towards Open-Ended Emotional Support Conversations in LLMs via Reinforcement Learning with Future-Oriented Rewards Unresolved cited work
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a78d4596-083c-4051-995f-b23dc2cc464c · outbound
Towards Open-Ended Emotional Support Conversations in LLMs via Reinforcement Learning with Future-Oriented Rewards DialogXpert: Driving Intelligent and Emotion-Aware Conversations through Online Value-Based Reinforcement Learning with LLM Priors
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 15b5f6dc-ab3f-4044-a7de-2d282c50d6fb · outbound
Towards Open-Ended Emotional Support Conversations in LLMs via Reinforcement Learning with Future-Oriented Rewards Proximal Policy Optimization Algorithms
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4661cbb7-9953-4c7a-9c3a-6eea873a13cb · outbound
Towards Open-Ended Emotional Support Conversations in LLMs via Reinforcement Learning with Future-Oriented Rewards DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cb70ae22-fbf6-4d03-b52c-4789fd24c773 · outbound
Towards Open-Ended Emotional Support Conversations in LLMs via Reinforcement Learning with Future-Oriented Rewards Unresolved cited work
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bdd8397e-9664-45ce-bf94-201cc7a15ba8 · outbound
Towards Open-Ended Emotional Support Conversations in LLMs via Reinforcement Learning with Future-Oriented Rewards Unresolved cited work
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 97a58edc-1278-49a9-af8b-0f32bc4bf717 · outbound
Towards Open-Ended Emotional Support Conversations in LLMs via Reinforcement Learning with Future-Oriented Rewards Convert Language Model into a Value-based Strategic Planner
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation cf7d4350-20c7-4229-bbee-5f0d1bb3662f · outbound
Towards Open-Ended Emotional Support Conversations in LLMs via Reinforcement Learning with Future-Oriented Rewards Logic-RL: Unleashing LLM Reasoning with Rule-Based Reinforcement Learning
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1d8e8fcd-67a8-4b44-9e1f-e6d20bf46ac6 · outbound
Towards Open-Ended Emotional Support Conversations in LLMs via Reinforcement Learning with Future-Oriented Rewards Qwen2.5-1M Technical Report
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation caf9c035-6335-4b0b-a0fe-0283d8e4c754 · outbound
Towards Open-Ended Emotional Support Conversations in LLMs via Reinforcement Learning with Future-Oriented Rewards A Survey on Recent Advances in LLM-Based Multi-turn Dialogue Systems
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a3034d0f-6934-4213-b3ec-b482a5ad0973 · outbound
Towards Open-Ended Emotional Support Conversations in LLMs via Reinforcement Learning with Future-Oriented Rewards Unresolved cited work
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation db432616-d80b-42ad-a10f-a970d1b668ed · outbound
Towards Open-Ended Emotional Support Conversations in LLMs via Reinforcement Learning with Future-Oriented Rewards Unresolved cited work
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 40f65c50-beea-4fb2-99be-3dfd909cc19f · outbound
Towards Open-Ended Emotional Support Conversations in LLMs via Reinforcement Learning with Future-Oriented Rewards Unresolved cited work
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f3b0b9a7-f180-461e-b2e5-7c1c16b0c9b4 · outbound
Towards Open-Ended Emotional Support Conversations in LLMs via Reinforcement Learning with Future-Oriented Rewards Unresolved cited work
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d158ef72-66d3-490f-8b6b-e8b1fd4206e4 · outbound
Towards Open-Ended Emotional Support Conversations in LLMs via Reinforcement Learning with Future-Oriented Rewards Unresolved cited work
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2d35f0ac-6e65-46a6-9452-277cffeea485 · outbound
Towards Open-Ended Emotional Support Conversations in LLMs via Reinforcement Learning with Future-Oriented Rewards Is ChatGPT Equipped with Emotional Dialogue Capabilities?
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 057fdc77-3d9f-4243-a36b-6ee5ce92b7ae · outbound
Towards Open-Ended Emotional Support Conversations in LLMs via Reinforcement Learning with Future-Oriented Rewards Unresolved cited work
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 83ef3e5d-136b-4901-87d2-b402e1e92c46 · outbound
Towards Open-Ended Emotional Support Conversations in LLMs via Reinforcement Learning with Future-Oriented Rewards Building Emotional Support Chatbots in the Era of LLMs
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 53ca47ee-4a2c-49e5-a426-528040c49908 · outbound
Towards Open-Ended Emotional Support Conversations in LLMs via Reinforcement Learning with Future-Oriented Rewards online" 'onlinestring :=
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a715afce-975e-46a6-8be3-788593b3d03a · outbound
Towards Open-Ended Emotional Support Conversations in LLMs via Reinforcement Learning with Future-Oriented Rewards write newline
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4e8f30e5-461e-47af-82a2-6a5025cd2432 · inbound
MICA: Multi-granularity Intertemporal Credit Assignment for Long-Horizon Emotional Support Dialogue Towards Open-Ended Emotional Support Conversations in LLMs via Reinforcement Learning with Future-Oriented Rewards
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation e9a4f98f-3c66-4700-bf45-9fc25c08dc3b · inbound
MICA: Multi-granularity Intertemporal Credit Assignment for Long-Horizon Emotional Support Dialogue Towards Open-Ended Emotional Support Conversations in LLMs via Reinforcement Learning with Future-Oriented Rewards
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.