RL-finetuning Qwen3-8B with LLM-as-judge rewards produces a red teaming model that generates effective attacks for novel adversarial goals not seen during training.
& Ganju, S
2 Pith papers cite this work. Polarity classification is still indexing.
2
Pith papers citing it
fields
cs.CV 2years
2026 2representative citing papers
AlphaEarth land-cover priors improve SAR flood segmentation IoU over SAR-only and DEM baselines across CNN and ViT backbones on held-out events like Hurricane Florence.
citing papers explorer
-
Urban Flood Observations: A hand-labeled training and validation dataset of post-flood inundation
RL-finetuning Qwen3-8B with LLM-as-judge rewards produces a red teaming model that generates effective attacks for novel adversarial goals not seen during training.
-
Beyond Backscatter: AlphaEarth Land-Cover Priors for Rapid SAR Flood Segmentation Across Foundation Backbones
AlphaEarth land-cover priors improve SAR flood segmentation IoU over SAR-only and DEM baselines across CNN and ViT backbones on held-out events like Hurricane Florence.