PulseExploreJournal ClubDebatesTrendingResearchersJournals
Instagram
HomeExploreJournal ClubTrending
Synapse
⌘+K
Synapse
March 29, 20240 citationsOpen Access

LeGo-Drive: Language-enhanced Goal-oriented Closed-Loop End-to-End Autonomous Driving

View Full Paper
PPPranjal PaulAGAnant GargTCTushar Choudhary

Key Points

Key points are not available for this paper at this time.

Abstract

Existing Vision-Language models (VLMs) estimate either long-term trajectory waypoints or a set of control actions as a reactive solution for closed-loop planning based on their rich scene comprehension. However, these estimations are coarse and are subjective to their "world understanding" which may generate sub-optimal decisions due to perception errors. In this paper, we introduce LeGo-Drive, which aims to address this issue by estimating a goal location based on the given language command as an intermediate representation in an end-to-end setting. The estimated goal might fall in a non-desirable region, like on top of a car for a parking-like command, leading to inadequate planning. Hence, we propose to train the architecture in an end-to-end manner, resulting in iterative refinement of both the goal and the trajectory collectively. We validate the effectiveness of our method through comprehensive experiments conducted in diverse simulated environments. We report significant improvements in standard autonomous driving metrics, with a goal reaching Success Rate of 81%. We further showcase the versatility of LeGo-Drive across different driving scenarios and linguistic inputs, underscoring its potential for practical deployment in autonomous vehicles and intelligent transportation systems.

Ask AI
Helpful
Bookmark
Share
View Full Paper

Cite This Study

Paul et al. (2024) studied this question.

synapsesocial.com/papers/68e71db5b6db643587697702https://doi.org/10.48550/arxiv.2403.20116
Ask AI
Helpful
Bookmark
Share
View Full Paper

Also Consider

Synapse has enriched 5 closely related papers on similar clinical questions. Consider them for comparative context:

  1. 1Less is More: Lean yet Powerful Vision-Language Model for Autonomous Driving2025
  2. 2DriveVLM: The Convergence of Autonomous Driving and Large Vision-Language Models2024 · 23 citations
  3. 3VLA-MP: A Vision-Language-Action Framework for Multimodal Perception and Physics-Constrained Action Generation in Autonomous Driving2025 · 1 citations
  4. 4SimpleLLM4AD: An End-to-End Vision-Language Model with Graph Visual Question Answering for Autonomous Driving2024
  5. 5LMAD: Integrated End-to-End Vision-Language Model for Explainable Autonomous Driving2025