OVGrasp integrates YOLO-World open-vocabulary detection, depth-based target selection, and speech-triggered release to control a cable-driven soft exoskeleton for assistive grasping.
MultiClear: Multimodal Soft Exoskeleton Glove for Transparent Object Grasping Assistance
1 Pith paper cite this work. Polarity classification is still indexing.
abstract
Grasping is a fundamental skill for interacting with the environment. However, this ability can be difficult for some (e.g. due to disability). Wearable robotic solutions can enhance or restore hand function, and recent advances have leveraged computer vision to improve grasping capabilities. However, grasping transparent objects remains challenging due to their poor visual contrast and ambiguous depth cues. Furthermore, while multimodal control strategies incorporating tactile and auditory feedback have been explored to grasp transparent objects, the integration of vision with these modalities remains underdeveloped. This paper introduces MultiClear, a multimodal framework designed to enhance grasping assistance in a wearable soft exoskeleton glove for transparent objects by fusing RGB data, depth data, and auditory signals. The exoskeleton glove integrates a tendon-driven actuator with an RGB-D camera and a built-in microphone. To achieve precise and adaptive control, a hierarchical control architecture is proposed. For the proposed hierarchical control architecture, a high-level control layer provides contextual awareness, a mid-level control layer processes multimodal sensory inputs, and a low-level control executes PID motor control for fine-tuned grasping adjustments. The challenge of transparent object segmentation was managed by introducing a vision foundation model for zero-shot segmentation. The proposed system achieves a Grasping Ability Score of 70.37%, demonstrating its effectiveness in transparent object manipulation.
citation-role summary
citation-polarity summary
fields
cs.RO 1years
2025 1verdicts
REJECT 1roles
extension 1polarities
extend 1representative citing papers
citing papers explorer
-
OVGrasp: Open-Vocabulary Grasping Assistance via Multimodal Intent Detection
OVGrasp integrates YOLO-World open-vocabulary detection, depth-based target selection, and speech-triggered release to control a cable-driven soft exoskeleton for assistive grasping.