Pith. sign in

REVIEW 1 cited by

Meta Reinforcement Learning for Fast Spectrum Sharing in Vehicular Networks

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2309.17185 v1 pith:WYXW3DIE submitted 2023-09-29 cs.IT eess.SPmath.IT

classification cs.ITeess.SPmath.IT
keywords spectrumagentlearningperformancereinforcementtrainingalgorithmcommunication
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

In this paper, we investigate the problem of fast spectrum sharing in vehicle-to-everything communication. In order to improve the spectrum efficiency of the whole system, the spectrum of vehicle-to-infrastructure links is reused by vehicle-to-vehicle links. To this end, we model it as a problem of deep reinforcement learning and tackle it with proximal policy optimization. A considerable number of interactions are often required for training an agent with good performance, so simulation-based training is commonly used in communication networks. Nevertheless, severe performance degradation may occur when the agent is directly deployed in the real world, even though it can perform well on the simulator, due to the reality gap between the simulation and the real environments. To address this issue, we make preliminary efforts by proposing an algorithm based on meta reinforcement learning. This algorithm enables the agent to rapidly adapt to a new task with the knowledge extracted from similar tasks, leading to fewer interactions and less training time. Numerical results show that our method achieves near-optimal performance and exhibits rapid convergence.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Deep Reinforcement Learning-Based User Scheduling for Collaborative Perception

    cs.LG 2025-02 conditional novelty 6.0 of 10

    A DDQN-based V2X scheduler using a label-free, semantics-aware reward selects which collaborator's BEV features to transmit and outperforms simple baselines in V2X-Sim simulations.

Pith tools