Pith. sign in

REVIEW 2 cited by

CLIP-Motion: Learning Reward Functions for Robotic Actions Using Consecutive Observations

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2311.03485 v2 pith:6NEOVFNW submitted 2023-11-06 cs.RO cs.AI

classification cs.ROcs.AI
keywords learningrewardroboticconsecutivefunctionsmethodmodelmotion
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

This paper presents a novel method for learning reward functions for robotic motions by harnessing the power of a CLIP-based model. Traditional reward function design often hinges on manual feature engineering, which can struggle to generalize across an array of tasks. Our approach circumvents this challenge by capitalizing on CLIP's capability to process both state features and image inputs effectively. Given a pair of consecutive observations, our model excels in identifying the motion executed between them. We showcase results spanning various robotic activities, such as directing a gripper to a designated target and adjusting the position of a cube. Through experimental evaluations, we underline the proficiency of our method in precisely deducing motion and its promise to enhance reinforcement learning training in the realm of robotics.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. VersualRL: Closed-Loop Verbal Reinforcement Learning with Visual Execution Feedback for Task-Level Robot Planning

    cs.RO 2026-03 conditional novelty 5.5 of 10

    A critic VLM and actor LLM iteratively refine a robot's Behavior Tree from visual feedback, without gradients, improving a pick-and-place logistics task on physical hardware.

  2. STEVE-Audio: Expanding the Goal Conditioning Modalities of Embodied Agents in Minecraft

    cs.LG 2024-12 conditional novelty 4.0 of 10

    An audio-conditioned STEVE-1 agent, built with a new Minecraft audio-video CLIP model and a learned prior, matches or beats text- and video-conditioned versions on most short-horizon collection tasks.

Pith tools