Pith. sign in

Paper Citation Record · LEDGER

Preference Learning Unlocks LLMs' Psycho-Counseling Skills

As of 5 August 2026, this Paper Citation Record lists 46 of 46 outbound references and 2 inbound Pith citation observations for arXiv:2502.19731.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2502.19731 v2

Coverage vector

measured 46 of 46 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-23T03:08:01.049353Z

measured 48 of 48 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-04T22:24:47.389400Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T02:28:24.338817Z

Reference resolution

46 of 46 outbound references displayed

  • verified exact1
  • verified fuzzy43
  • unresolved1
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

0
pith, observed 2026-08-05T02:28:24.338817Z

Outbound references

Observation 042e1130-fb75-45e4-ad63-eb401feefb2e · outbound

This paper cites Phi-3 technical report: A highly capable language model locally on your phone.

Preference Learning Unlocks LLMs' Psycho-Counseling Skills Phi-3 technical report: A highly capable language model locally on your phone

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T04:15:24.270089Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T03:08:01.049353Z digest=sha256:69a7f3e1a9c19c55bd0f52d709eb56acaba91dd9ca06f960faab45a64aea8c1a

Observation d8e718c5-9dd2-4a40-8988-1c42749e83e4 · outbound

This paper cites Training a helpful and harmless assistant with reinforcement learning from human feedback.

Preference Learning Unlocks LLMs' Psycho-Counseling Skills Training a helpful and harmless assistant with reinforcement learning from human feedback

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T04:15:24.274122Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T03:08:01.049353Z digest=sha256:bd85a403e32ed4f66a826d86d8f5af827030ff41941d279e08a95a69f9b0edf5

Observation ae97c87d-8204-45b2-afcd-d6c383467b1d · outbound

This paper cites The use of the area under the ROC curve in the evaluation of machine learning algorithms.

Preference Learning Unlocks LLMs' Psycho-Counseling Skills The use of the area under the ROC curve in the evaluation of machine learning algorithms

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T04:15:24.283378Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T03:08:01.049353Z digest=sha256:a95b6b7ed8129e75bd2d59eadd63fbdfe8ceda20f88607b179150be94d8dcfcc

Observation f9d54bcf-ede3-439b-8926-3c77a3cdbaa9 · outbound

This paper cites Orion- 14B : Open-source multilingual large language models.

Preference Learning Unlocks LLMs' Psycho-Counseling Skills Orion- 14B : Open-source multilingual large language models

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T04:15:24.277913Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T03:08:01.049353Z digest=sha256:bde616bd0253cee7b2fcab3a86712819fea00fb2e48f825f30e967cf92f15c3d

Observation 28f9c70e-1305-499a-a4da-f44206ca06bd · outbound

This paper cites LLM -empowered chatbots for psychiatrist and patient simulation: Application and evaluation.

Preference Learning Unlocks LLMs' Psycho-Counseling Skills LLM -empowered chatbots for psychiatrist and patient simulation: Application and evaluation

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T04:15:24.259399Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T03:08:01.049353Z digest=sha256:d857b38ef6dcec422d46f363bfb64983f636d34adc62f431a2e05f3e247f4e63

Observation 10854828-c43d-413a-9941-7c02b09b1a57 · outbound

This paper cites Empowering psychotherapy with large language models: Cognitive distortion detection through diagnosis of thought prompting.

Preference Learning Unlocks LLMs' Psycho-Counseling Skills Empowering psychotherapy with large language models: Cognitive distortion detection through diagnosis of thought prompting

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T04:15:24.266734Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T03:08:01.049353Z digest=sha256:2c22c0fe86629e6ec64fa6f45c46c9452b96b7cf98da99ec7286589596707780

Observation 7b86d1e4-cf66-4f85-93d2-8eefd8676323 · outbound

This paper cites Challenges of large language models for mental health counseling.

Preference Learning Unlocks LLMs' Psycho-Counseling Skills Challenges of large language models for mental health counseling

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T04:15:24.197949Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T03:08:01.049353Z digest=sha256:ed6b2db6aaed22bbeba3fd6337ccb740db44588bfea8f1b1c6451772661b1e2a

Observation 697fa309-25ed-4695-b1a1-9a616b5e33cc · outbound

This paper cites UltraFeedback : Boosting language models with high-quality feedback.

Preference Learning Unlocks LLMs' Psycho-Counseling Skills UltraFeedback : Boosting language models with high-quality feedback

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T04:15:24.229737Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T03:08:01.049353Z digest=sha256:03c37c1da63918b61c9cf1db1c1919ad48930a1fd0b3edb0816c6e3b28b5565d

Observation 6c537233-ed67-4790-9e9d-1847da686168 · outbound

This paper cites DeepSeek LLM : Scaling open-source language models with longtermism.

Preference Learning Unlocks LLMs' Psycho-Counseling Skills DeepSeek LLM : Scaling open-source language models with longtermism

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T04:15:24.236922Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T03:08:01.049353Z digest=sha256:8f461f27b373c2b9db21ee56bf1b64d84c13f5b45e874d94031655b85593433c

Observation c2ba3234-b9fa-4e2c-85e8-c2705329f01b · outbound

This paper cites DeepSeek - R1 : Incentivizing reasoning capability in LLMs via reinforcement learning.

Preference Learning Unlocks LLMs' Psycho-Counseling Skills DeepSeek - R1 : Incentivizing reasoning capability in LLMs via reinforcement learning

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T04:15:24.240330Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T03:08:01.049353Z digest=sha256:2ee58b6d048ddff77fc596fbd3494514a05dca793ffc4fbe7dad777d9578d840

Observation 18f6abd8-a6fe-4340-89e0-2a2e92f16175 · outbound

This paper cites Gemma 2: Improving open language models at a practical size.

Preference Learning Unlocks LLMs' Psycho-Counseling Skills Gemma 2: Improving open language models at a practical size

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T04:15:24.216262Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T03:08:01.049353Z digest=sha256:0a377e74d41fb8f19aac67717161767b06476c30888fd653f7a0ebf258302908

Observation 63beb2ef-0fc0-48b0-8be8-cb476cfe4875 · outbound

This paper cites MiniCPM : Unveiling the potential of small language models with scalable training strategies.

Preference Learning Unlocks LLMs' Psycho-Counseling Skills MiniCPM : Unveiling the potential of small language models with scalable training strategies

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T04:15:24.220087Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T03:08:01.049353Z digest=sha256:07502ddbaffdd48083e158c35f54f964fc1dc07fa5ef25013927569529ad0fda

Observation a1e2150a-9757-4ea9-bf38-3a908502d1d7 · outbound

This paper cites Jamba-1.5: Hybrid transformer-mamba models at scale.

Preference Learning Unlocks LLMs' Psycho-Counseling Skills Jamba-1.5: Hybrid transformer-mamba models at scale

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T04:15:24.281244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T03:08:01.049353Z digest=sha256:21d13741505c73096deddb539515176de1e9b454a0a845d1b41540f9b279853f

Observation f859d70a-810e-4068-a082-bc804c1a3fc1 · outbound

This paper cites RewardBench : Evaluating reward models for language modeling.

Preference Learning Unlocks LLMs' Psycho-Counseling Skills RewardBench : Evaluating reward models for language modeling

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T04:15:24.205773Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T03:08:01.049353Z digest=sha256:06e4a7bebf3a87625f32892d802023da6d33621ec12bc27dd0842ab098641a4e

Observation 6f762a51-70d2-4b11-b581-c0a0873871a7 · outbound

This paper cites MentalAgora : A gateway to advanced personalized care in mental health through multi-agent debating and attribute control.

Preference Learning Unlocks LLMs' Psycho-Counseling Skills MentalAgora : A gateway to advanced personalized care in mental health through multi-agent debating and attribute control

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T04:15:24.228780Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T03:08:01.049353Z digest=sha256:cdac1883f9c0e89aa84e1eb513c346f34a143e06d65da5ddf66cb35e457e1848

Observation f8526871-8e93-460d-8e4a-b0b0a485a5d6 · outbound

This paper cites Skywork-reward: Bag of tricks for reward modeling in LLMs.

Preference Learning Unlocks LLMs' Psycho-Counseling Skills Skywork-reward: Bag of tricks for reward modeling in LLMs

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T04:15:24.212466Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T03:08:01.049353Z digest=sha256:34c9b69d86ad7c09cd141b945196d95eb2f8db1fcab9adc35d0d0d17ab9bd299

Observation 9a6f01ac-01eb-4677-a937-28044291fdb0 · outbound

This paper cites ChatCounselor : A large language models for mental health support.

Preference Learning Unlocks LLMs' Psycho-Counseling Skills ChatCounselor : A large language models for mental health support

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T04:15:24.298785Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T03:08:01.049353Z digest=sha256:f5c4940b8ae383eb64fa340b4e7bc438f770983179e6858b91713b05bf420155

Observation d395a389-9dca-468b-b3b3-b232d1317e49 · outbound

This paper cites The llama 3 herd of models.

Preference Learning Unlocks LLMs' Psycho-Counseling Skills The llama 3 herd of models

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T04:15:24.177857Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T03:08:01.049353Z digest=sha256:49d265b7ebd9f978ad6fedb45f3f6d2d6c9876e047d131c673ee9484e6871c24

Observation 448eb136-d308-4ecc-ba7b-0dc831fe2644 · outbound

This paper cites Training models to generate, recognize, and reframe unhelpful thoughts.

Preference Learning Unlocks LLMs' Psycho-Counseling Skills Training models to generate, recognize, and reframe unhelpful thoughts

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T04:15:24.224547Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T03:08:01.049353Z digest=sha256:afcb368f697e916be5d1325bc7bf385ee60881bf3c7e50704cd8bb7d1d535393

Observation 706c47ce-8539-436c-8939-8bcad6e21a51 · outbound

This paper cites Beyond training objectives: Interpreting reward model divergence in large language models.

Preference Learning Unlocks LLMs' Psycho-Counseling Skills Beyond training objectives: Interpreting reward model divergence in large language models

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T04:15:24.226461Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T03:08:01.049353Z digest=sha256:c046405ab726d0862d034e33b26a7511a5fd5d1bc63eb5aa60ddeea718668f6e

Observation cd43e6ef-b0b3-4e99-a559-d9e39f13fb86 · outbound

This paper cites Motivational interviewing: Helping people change.

Preference Learning Unlocks LLMs' Psycho-Counseling Skills Motivational interviewing: Helping people change

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T04:15:24.284676Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T03:08:01.049353Z digest=sha256:3dc33b86847fd158e0457f89b3d82e2ecf0aad158db025d31c7ac8c4bd0862af

Observation fd01dd91-0441-422e-a2f6-e4c80ac0bfbb · outbound

This paper cites OLMoE : Open mixture-of-experts language models.

Preference Learning Unlocks LLMs' Psycho-Counseling Skills OLMoE : Open mixture-of-experts language models

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T04:15:24.286938Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T03:08:01.049353Z digest=sha256:68315ea5cff1f0dc1cf99c0625fb96d9e0935b403491b47565d8f8535ed09b01

Observation 0db1079b-3904-41d6-ac6b-1d6b9a8c23e6 · outbound

This paper cites A survey of large language models in psychotherapy: Current landscape and future directions.

Preference Learning Unlocks LLMs' Psycho-Counseling Skills A survey of large language models in psychotherapy: Current landscape and future directions

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T04:15:24.175098Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T03:08:01.049353Z digest=sha256:cac275d3d68b27ab74665d8915a8842d260610d6f7f912d656115df35a41d9c7

Observation 64887c5e-bd2d-4212-99cf-bdec3e713bd3 · outbound

This paper cites Obtaining well calibrated probabilities using bayesian binning.

Preference Learning Unlocks LLMs' Psycho-Counseling Skills Obtaining well calibrated probabilities using bayesian binning

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T04:15:24.255008Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T03:08:01.049353Z digest=sha256:7906649b4cd823af1045b61ea2471fd55f6c49127b0689394538dcc4d248b613

Observation f944cad3-0f27-4717-9fe9-71f3e98a7763 · outbound

This paper cites GPT - 4o system card.

Preference Learning Unlocks LLMs' Psycho-Counseling Skills GPT - 4o system card

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T04:15:24.243854Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T03:08:01.049353Z digest=sha256:855cdd94a50c87c0a9276de636b07f41be4aaafb89a4ce412112ad2b02b4f51a

Observation 420bc5a4-0bc7-46c2-846d-bf38f4803f0f · outbound

This paper cites Training language models to follow instructions with human feedback.

Preference Learning Unlocks LLMs' Psycho-Counseling Skills Training language models to follow instructions with human feedback

Reference 26

Resolution
verified exact
local_arxiv, observed 2026-05-23T03:12:28.600337Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T03:08:01.049353Z digest=sha256:eb839522a23480ee8bc87aa5a2949d8df0a019eb8bd349cb4c65b4939bb193d2

Observation f4d69900-3290-4c80-a8a5-d0e5ac521f0c · outbound

This paper cites Iterative reasoning preference optimization.

Preference Learning Unlocks LLMs' Psycho-Counseling Skills Iterative reasoning preference optimization

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T04:15:24.178154Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T03:08:01.049353Z digest=sha256:071b2e0396270f7ad6a32798fb6a0454e89c7008de5ef0ac2921f5951ae5c386

Observation 3b858c66-4191-48ce-8fa0-2d6d0b80ce55 · outbound

This paper cites Qwen2 .5 technical report.

Preference Learning Unlocks LLMs' Psycho-Counseling Skills Qwen2 .5 technical report

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T04:15:24.237199Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T03:08:01.049353Z digest=sha256:24e7c742e64891c6b8cd2f8ad4664c829246655d2fc2bc63277af7acedc4c5c2

Observation 17b8f16f-9ad1-45e5-b9d8-98ff89f2b4e3 · outbound

This paper cites Direct preference optimization: Your language model is secretly a reward model.

Preference Learning Unlocks LLMs' Psycho-Counseling Skills Direct preference optimization: Your language model is secretly a reward model

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T04:15:24.156164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T03:08:01.049353Z digest=sha256:15c4d1892779463c8885f6ab879cb092e015b748fb4ec0c14a0deab7ce500cb8

Observation 05b7c45b-3fff-4719-8a1e-0644e3047563 · outbound

This paper cites Scaling laws for reward model overoptimization in direct alignment algorithms.

Preference Learning Unlocks LLMs' Psycho-Counseling Skills Scaling laws for reward model overoptimization in direct alignment algorithms

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T04:15:24.262934Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T03:08:01.049353Z digest=sha256:7cdad93afca49b8ba75d202964c6635e61156c6db443165d6b1a78748b288afb

Observation cdc7a873-3733-4b52-8bf8-9273dee02dcd · outbound

This paper cites Key factors in psychotherapy training: an analysis of trainers’, trainees’ and psychotherapists’ points of view.

Preference Learning Unlocks LLMs' Psycho-Counseling Skills Key factors in psychotherapy training: an analysis of trainers’, trainees’ and psychotherapists’ points of view

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T04:15:24.147614Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T03:08:01.049353Z digest=sha256:5ff62a0ebaec0221fd3b77b4ce05614e342018ca6b54d49a2a3bc07b61b1234f

Observation cd621d6b-86c4-4b6e-b21b-db7f944ee390 · outbound

This paper cites Proximal policy optimization algorithms.

Preference Learning Unlocks LLMs' Psycho-Counseling Skills Proximal policy optimization algorithms

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T04:15:24.163175Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T03:08:01.049353Z digest=sha256:b7c941e5ad788a26b4971a848fcee675c0015c5ed30f9c255005326c35a325eb

Observation d3d64ffc-6b77-4ffa-b91d-6512630dda49 · outbound

This paper cites Facilitating self-guided mental health interventions through human-language model interaction: A case study of cognitive restructuring.

Preference Learning Unlocks LLMs' Psycho-Counseling Skills Facilitating self-guided mental health interventions through human-language model interaction: A case study of cognitive restructuring

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T04:15:24.251534Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T03:08:01.049353Z digest=sha256:95bb4f238a9ac33a4e3005f19642415a1e10b4959528a45349d8f013d4bbd167

Observation 310224ef-929e-41d6-9a3f-947fa69b6903 · outbound

This paper cites Detecting cognitive distortions from patient-therapist interactions.

Preference Learning Unlocks LLMs' Psycho-Counseling Skills Detecting cognitive distortions from patient-therapist interactions

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T04:15:24.290779Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T03:08:01.049353Z digest=sha256:553c6dcdb2a1ad9b8aed2e9775031e109434d4c0b1cd8581cd7537c1ad094fee

Observation aec6346d-3834-4252-976b-306b6278bfa6 · outbound

This paper cites Large language models could change the future of behavioral healthcare: a proposal for responsible development and evaluation.

Preference Learning Unlocks LLMs' Psycho-Counseling Skills Large language models could change the future of behavioral healthcare: a proposal for responsible development and evaluation

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T04:15:24.182060Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T03:08:01.049353Z digest=sha256:5f84242b69948e22af09c82495367debb97301a4e64c278874e17fc8ff5e7f59

Observation 9599f955-06f4-4488-a4bd-31f5f840043a · outbound

This paper cites Understanding the performance gap between online and offline alignment algorithms.

Preference Learning Unlocks LLMs' Psycho-Counseling Skills Understanding the performance gap between online and offline alignment algorithms

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T04:15:24.171603Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T03:08:01.049353Z digest=sha256:b23e7d56dfc66b8c0a7a21a68c7abcdf4899d7462d6743c07c715262b7ae72ff

Observation 98caa8ac-6cb7-4cc6-af27-1aefdee794bf · outbound

This paper cites HelpSteer2 -preference: Complementing ratings with preferences.

Preference Learning Unlocks LLMs' Psycho-Counseling Skills HelpSteer2 -preference: Complementing ratings with preferences

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T04:15:24.276008Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T03:08:01.049353Z digest=sha256:190be4686abe3cb706cd51baf33d3219c7f2befb718259d21268efdd41bc5c3f

Observation 06acccde-3438-44c1-a3fb-926f8caded86 · outbound

This paper cites Is DPO superior to PPO for LLM alignment? a comprehensive study.

Preference Learning Unlocks LLMs' Psycho-Counseling Skills Is DPO superior to PPO for LLM alignment? a comprehensive study

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T04:15:24.267947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T03:08:01.049353Z digest=sha256:380fdc7bb5840672d68a3402692fe68864dd95cb7c3072d8d462276b02047695

Observation 963823fd-ef84-428d-9c97-b2458efe5d05 · outbound

This paper cites Baichuan 2: Open large-scale language models.

Preference Learning Unlocks LLMs' Psycho-Counseling Skills Baichuan 2: Open large-scale language models

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T04:15:24.271414Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T03:08:01.049353Z digest=sha256:286ca7705329c4b468766eaefca7b2058b11f159b5844a56b5ac37e860df64d8

Observation bfd644d1-cd7c-4ace-9fcd-6613689733be · outbound

This paper cites Qwen2 technical report.

Preference Learning Unlocks LLMs' Psycho-Counseling Skills Qwen2 technical report

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T04:15:24.279904Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T03:08:01.049353Z digest=sha256:dd63cdb2bd9705d70e10c7b95cb8f56c1049addebd84b657ca6c4cbbd1edfbcc

Observation 638eda66-273e-4316-a559-0faf937a6b93 · outbound

This paper cites CBT -bench: Evaluating large language models on assisting cognitive behavior therapy.

Preference Learning Unlocks LLMs' Psycho-Counseling Skills CBT -bench: Evaluating large language models on assisting cognitive behavior therapy

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T04:15:24.208968Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T03:08:01.049353Z digest=sha256:cad9c7449c3a916c0f76c438265c2e842bc590296a8f5030254b22c33c71b933

Observation cbbcb753-fb0d-41d7-ba6b-992d2e46e984 · outbound

This paper cites Judging LLM -as-a-judge with MT -bench and chatbot arena.

Preference Learning Unlocks LLMs' Psycho-Counseling Skills Judging LLM -as-a-judge with MT -bench and chatbot arena

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T04:15:24.287776Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T03:08:01.049353Z digest=sha256:f2de0b0da5b24403ff5d057ef742af51d170cf2831140e7dc4b3efae0a93a09a

Observation fac8b718-ee66-4fb8-9b58-935d42288477 · outbound

This paper cites write newline.

Preference Learning Unlocks LLMs' Psycho-Counseling Skills write newline

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T04:15:24.295014Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T03:08:01.049353Z digest=sha256:c60acfb273975abb6c54d286fd202e64ed9aeedf1c4a5f07cf92c53fe833d57e

Observation ff038a7f-254c-4618-b6d3-06690f539d18 · outbound

This paper cites @esa (Ref.

Preference Learning Unlocks LLMs' Psycho-Counseling Skills @esa (Ref

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-05-23T04:15:24.219861Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T03:08:01.049353Z digest=sha256:c6f764954daf0c362b4096f66a532e312d4380ca4952c326c02659597a2eb34d

Observation ca11f169-ac93-43e0-8040-03cfd6e468cf · outbound

This paper cites an unresolved cited work.

Preference Learning Unlocks LLMs' Psycho-Counseling Skills Unresolved cited work

Reference 45

Resolution
unresolved
raw_fallback, observed 2026-05-23T04:15:24.302195Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T03:08:01.049353Z digest=sha256:a9ac9798dd79fe12d841ba9c7e9487d11d8935c0ac2c58213ece521277b8b0e8

Observation 7a468a55-dd48-4155-ab36-616e7ec310a0 · outbound

This paper cites Victoria Beckham.

Preference Learning Unlocks LLMs' Psycho-Counseling Skills Victoria Beckham

Reference 46

Resolution
metadata mismatch
arxiv_id, observed 2026-05-23T03:12:28.595223Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-23T03:08:01.049353Z digest=sha256:87de3b29f8676e167b3b05f741c758b47eae3b1c10dfbcef49509e66c9c73524

Pith citing papers

Observation dce75e1b-e320-468c-9ab9-9ba2c7d50a9e · inbound

PersonaFuse: A Personality Activation-Driven Framework for Enhancing Human-LLM Interactions cites this paper.

PersonaFuse: A Personality Activation-Driven Framework for Enhancing Human-LLM Interactions Preference Learning Unlocks LLMs' Psycho-Counseling Skills

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-04T22:24:47.389400Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T22:24:47.389400Z digest=sha256:113629c539d4e08652072a08c50e124dcf77ce84c9ada17f2ee5e85e8879cf87

Observation cbe1f9fb-0064-43c8-a9c9-cda0b85e7488 · inbound

A clinically validated framework for auditing AI chatbot behavior in mental health interactions cites this paper.

A clinically validated framework for auditing AI chatbot behavior in mental health interactions Preference Learning Unlocks LLMs' Psycho-Counseling Skills

Reference 56

Resolution
verified exact
local_arxiv, observed 2026-08-04T06:24:08.371059Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-08-04T06:18:37.097676Z digest=sha256:e671f2d0e28547e85c6cd8f5d95f38f02769cda8b3c810c9d078c86588823e7e