A single PPO policy, trained with domain randomization and heuristic phase gating, performs obstacle separation, detachment, and placement, reaching 82% real-world strawberry-harvest success with zero-shot sim-to-real transfer.
An autonomous strawberry-harvesting robot: Design, development, integration, and field evaluation,
1 Pith paper cite this work, alongside 422 external citations. Polarity classification is still indexing.
1
Pith paper citing it
422
external citations · OpenAlex
fields
cs.RO 1years
2026 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
Reinforcement Learning for the Full Strawberry Harvesting Process: Obstacle Separation, Detachment, and Placement
A single PPO policy, trained with domain randomization and heuristic phase gating, performs obstacle separation, detachment, and placement, reaching 82% real-world strawberry-harvest success with zero-shot sim-to-real transfer.