CT-GMARL outperforms R-MAPPO and QMIX in the new continuous-time NetForge_RL cyber-defense simulator, restoring 12x more services and transferring zero-shot to Docker exploits.
MiniLM: Deep self-attention distillation for task-agnostic compression of pre-trained transformers.Advances in Neural Information Processing Systems, 33:5776–5788
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
fields
cs.LG 1years
2026 1verdicts
CONDITIONAL 1representative citing papers
citing papers explorer
-
NetForge RL: A Multi-Agent Simulation Environment for Cyber Defense with Durative Actions
CT-GMARL outperforms R-MAPPO and QMIX in the new continuous-time NetForge_RL cyber-defense simulator, restoring 12x more services and transferring zero-shot to Docker exploits.