REVIEW 5 cited by
TAIL: Task-specific Adapters for Imitation Learning with Large Pretrained Models
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
The full potential of large pretrained models remains largely untapped in control domains like robotics. This is mainly because of the scarcity of data and the computational challenges associated with training or fine-tuning these large models for such applications. Prior work mainly emphasizes either effective pretraining of large models for decision-making or single-task adaptation. But real-world problems will require data-efficient, continual adaptation for new control tasks. Recognizing these constraints, we introduce TAIL (Task-specific Adapters for Imitation Learning), a framework for efficient adaptation to new control tasks. Inspired by recent advancements in parameter-efficient fine-tuning in language domains, we explore efficient fine-tuning techniques -- e.g., Bottleneck Adapters, P-Tuning, and Low-Rank Adaptation (LoRA) -- in TAIL to adapt large pretrained models for new tasks with limited demonstration data. Our extensive experiments in large-scale language-conditioned manipulation tasks comparing prevalent parameter-efficient fine-tuning techniques and adaptation baselines suggest that TAIL with LoRA can achieve the best post-adaptation performance with only 1\% of the trainable parameters of full fine-tuning, while avoiding catastrophic forgetting and preserving adaptation plasticity in continual learning settings.
Forward citations
Cited by 5 Pith papers
-
Reinforced Refinement with Self-Aware Expansion for End-to-End Autonomous Driving
R2SE refines pretrained end-to-end driving policies on hard cases via residual LoRA reinforcement learning and switches between specialist and generalist policies using GPD-based uncertainty.
-
SPECI: Skill Prompts based Hierarchical Continual Imitation Learning for Robot Manipulation
A hierarchical continual imitation learning policy with an expandable skill codebook and CP-decomposed task-specific attention parameters outperforms prior CIL methods on the LIBERO robot manipulation benchmark.
-
Human-like Bots for Tactical Shooters Using Compute-Efficient Sensors
Ray-cast sensor-based imitation learning produces compute-efficient, human-like bots for a tactical shooter, with a 14.9M-parameter model running at 9.59 ms per decision on CPU.
-
Few-Shot Vision-Language Action-Incremental Policy Learning
TOPIC adds task-specific prompts and task-similarity-based weight interpolation to transformer policies, improving few-shot action-incremental learning in simulation and on a real robot.
-
Improving Vision-Language-Action Model with Online Reinforcement Learning
Alternating online RL on a frozen vision-language backbone with supervised fine-tuning on collected successes improves a VLA policy's task success and generalization.
Discussion (0). Continue with ORCID to comment.