Pith. sign in

REVIEW

Explanation for Trajectory Planning using Multi-modal Large Language Model for Autonomous Driving

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2411.09971 v1 pith:73ZEZAVO submitted 2024-11-15 cs.CV cs.RO

classification cs.CVcs.RO
keywords vehiclefuturemodelmodelsautonomouscaptionscontroldriving
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

End-to-end style autonomous driving models have been developed recently. These models lack interpretability of decision-making process from perception to control of the ego vehicle, resulting in anxiety for passengers. To alleviate it, it is effective to build a model which outputs captions describing future behaviors of the ego vehicle and their reason. However, the existing approaches generate reasoning text that inadequately reflects the future plans of the ego vehicle, because they train models to output captions using momentary control signals as inputs. In this study, we propose a reasoning model that takes future planning trajectories of the ego vehicle as inputs to solve this limitation with the dataset newly collected.

Discussion (0). Sign in to comment.

Pith tools