Pith. sign in

REVIEW 3 cited by

Hybrid LLM-DDQN based Joint Optimization of V2I Communication and Autonomous Driving

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2410.08854 v3 pith:KHXWZCLI submitted 2024-10-11 cs.LG cs.AIcs.NIcs.SYeess.SY

classification cs.LGcs.AIcs.NIcs.SYeess.SY
keywords llmsdecisionsoptimizationddqnalgorithmapproachautonomousconventional
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Large language models (LLMs) have received considerable interest recently due to their outstanding reasoning and comprehension capabilities. This work explores applying LLMs to vehicular networks, aiming to jointly optimize vehicle-to-infrastructure (V2I) communications and autonomous driving (AD) policies. We deploy LLMs for AD decision-making to maximize traffic flow and avoid collisions for road safety, and a double deep Q-learning algorithm (DDQN) is used for V2I optimization to maximize the received data rate and reduce frequent handovers. In particular, for LLM-enabled AD, we employ the Euclidean distance to identify previously explored AD experiences, and then LLMs can learn from past good and bad decisions for further improvement. Then, LLM-based AD decisions will become part of states in V2I problems, and DDQN will optimize the V2I decisions accordingly. After that, the AD and V2I decisions are iteratively optimized until convergence. Such an iterative optimization approach can better explore the interactions between LLMs and conventional reinforcement learning techniques, revealing the potential of using LLMs for network optimization and management. Finally, the simulations demonstrate that our proposed hybrid LLM-DDQN approach outperforms the conventional DDQN algorithm, showing faster convergence and higher average rewards.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 3 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Deep Generative Model-Aided Power System Dynamic State Estimation and Reconstruction with Unknown Control Inputs or Data Distributions

    eess.SY 2025-01 conditional novelty 6.0 of 10

    A deep generative DSE framework jointly estimates states and unknown control inputs, robustly reconstructs corrupted latent codes via a latent diffusion model, and adapts to unseen operating conditions with a one-shot...

  2. Large Language Models (LLMs) as Traffic Control Systems at Urban Intersections: A New Paradigm

    cs.CL 2024-11 conditional novelty 5.0 of 10

    A fine-tuned GPT-4o-mini detects conflicts in synthetic four-leg intersection scenarios with 83% accuracy and produces traffic-management text with high ROUGE-L scores against the simulator's templated references.

  3. A Survey on Large Language Models for Communication, Network, and Service Management: Application Insights, Challenges, and Future Directions

    cs.NI 2024-12 conditional novelty 4.0 of 10

    A systematic survey of 108 papers classifies how large language models are used for communication network and service management across four network domains.

Pith tools