A pruned multi-agent PPO algorithm with exploration incentives is designed to approximate the Stackelberg equilibrium of a vehicular AI twin migration game.
Language models meet world models: Embodied experiences enhance language models,
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.MA 1years
2025 1verdicts
REJECT 1representative citing papers
citing papers explorer
-
TinyMA-IEI-PPO: Exploration Incentive-Driven Multi-Agent DRL with Self-Adaptive Pruning for Vehicular Embodied AI Agent Twins Migration
A pruned multi-agent PPO algorithm with exploration incentives is designed to approximate the Stackelberg equilibrium of a vehicular AI twin migration game.