Pith. sign in

REVIEW 1 cited by

Personalization in Human-Robot Interaction through Preference-based Action Representation Learning

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2409.13822 v2 pith:BQDRAK3R submitted 2024-09-20 cs.RO

classification cs.RO
keywords learningrobotactiondomainhumanpre-trainedpreference-basedrepresentation
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Preference-based reinforcement learning (PbRL) has shown significant promise for personalization in human-robot interaction (HRI) by explicitly integrating human preferences into the robot learning process. However, existing practices often require training a personalized robot policy from scratch, resulting in inefficient use of human feedback. In this paper, we propose preference-based action representation learning (PbARL), an efficient fine-tuning method that decouples common task structure from preference by leveraging pre-trained robot policies. Instead of directly fine-tuning the pre-trained policy with human preference, PbARL uses it as a reference for an action representation learning task that maximizes the mutual information between the pre-trained source domain and the target user preference-aligned domain. This approach allows the robot to personalize its behaviors while preserving original task performance and eliminates the need for extensive prior information from the source domain, thereby enhancing efficiency and practicality in real-world HRI scenarios. Empirical results on the Assistive Gym benchmark and a real-world user study (N=8) demonstrate the benefits of our method compared to state-of-the-art approaches.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Coloring Between the Lines: Personalization in the Null Space of Planning Constraints

    cs.RO 2025-05 conditional novelty 6.0 of 10

    CBTL learns parameterized personalization constraints inside the safe solution space of robot planning CSPs, using entropy-based active queries to adapt quickly.

Pith tools