Pith. sign in

REVIEW 2 cited by

Zero-shot Task Adaptation using Natural Language

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2106.02972 v1 pith:SRIN2AAM submitted 2021-06-05 cs.AI cs.CLcs.LG

classification cs.AIcs.CLcs.LG
keywords tasktargetagentlanguagedemonstrationdescriptiongivennatural
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Imitation learning and instruction-following are two common approaches to communicate a user's intent to a learning agent. However, as the complexity of tasks grows, it could be beneficial to use both demonstrations and language to communicate with an agent. In this work, we propose a novel setting where an agent is given both a demonstration and a description, and must combine information from both the modalities. Specifically, given a demonstration for a task (the source task), and a natural language description of the differences between the demonstrated task and a related but different task (the target task), our goal is to train an agent to complete the target task in a zero-shot setting, that is, without any demonstrations for the target task. To this end, we introduce Language-Aided Reward and Value Adaptation (LARVA) which, given a source demonstration and a linguistic description of how the target task differs, learns to output a reward / value function that accurately describes the target task. Our experiments show that on a diverse set of adaptations, our approach is able to complete more than 95% of target tasks when using template-based descriptions, and more than 70% when using free-form natural language.

Discussion (0). Sign in to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. A Human-in-the-loop Approach to Robot Action Replanning through LLM Common-Sense Reasoning

    cs.RO 2025-07 conditional novelty 5.0 of 10

    A human-in-the-loop system lets users refine vision-generated robot behavior trees through natural-language requests to GPT-4o, correcting errors and adapting plans before execution.

  2. Bridging Perception and Action: Spatially-Grounded Mid-Level Representations for Robot Generalization

    cs.RO 2025-06 conditional novelty 5.0 of 10

    A mixture-of-experts diffusion policy conditioned on object, pose, depth, and trajectory mid-level representations is reported to outperform language-only and representation-free baselines on bimanual dexterous tasks,...

Pith tools