A fine-tuned CogVLM2 with chain-of-thought is applied to autonomous driving tasks, but its claimed performance gains are not supported by the reported quantitative results.
Multi-Task Conditional Imitation Learning for Autonomous Navigation at Crowded Intersections
1 Pith paper cite this work. Polarity classification is still indexing.
abstract
In recent years, great efforts have been devoted to deep imitation learning for autonomous driving control, where raw sensory inputs are directly mapped to control actions. However, navigating through densely populated intersections remains a challenging task due to uncertainty caused by uncertain traffic participants. We focus on autonomous navigation at crowded intersections that require interaction with pedestrians. A multi-task conditional imitation learning framework is proposed to adapt both lateral and longitudinal control tasks for safe and efficient interaction. A new benchmark called IntersectNav is developed and human demonstrations are provided. Empirical results show that the proposed method can achieve a success rate gain of up to 30% compared to the state-of-the-art.
citation-role summary
citation-polarity summary
fields
cs.CL 1years
2024 1verdicts
REJECT 1roles
background 1polarities
unclear 1representative citing papers
citing papers explorer
-
Application of Multimodal Large Language Models in Autonomous Driving
A fine-tuned CogVLM2 with chain-of-thought is applied to autonomous driving tasks, but its claimed performance gains are not supported by the reported quantitative results.