A latent Q-Barrier shield learns context, dynamics, and cost critic to enforce safety constraints at action level in in-context RL, with a conditional barrier-margin guarantee and benchmark gains.
Title resolution pending
2 Pith papers cite this work. Polarity classification is still indexing.
citation-role summary
citation-polarity summary
years
2026 2verdicts
UNVERDICTED 2roles
method 1polarities
use method 1representative citing papers
MACF learns dynamic ESG costs from point-in-time multimodal evidence to impose constraints on portfolio transitions, and MACF-X adapters reduce tail ESG budget pressure across optimization interfaces while keeping financial performance competitive.
citing papers explorer
-
Latent Q-Barrier Shielding for Safe In-Context Reinforcement Learning
A latent Q-Barrier shield learns context, dynamics, and cost critic to enforce safety constraints at action level in in-context RL, with a conditional barrier-margin guarantee and benchmark gains.
-
Beyond ESG Scores: Learning Dynamic Constraints for Sequential Portfolio Optimization
MACF learns dynamic ESG costs from point-in-time multimodal evidence to impose constraints on portfolio transitions, and MACF-X adapters reduce tail ESG budget pressure across optimization interfaces while keeping financial performance competitive.