There are only finitely many k-vertex-critical graphs in the classes (P4+ℓP1, B4(m), B3(m)+)-free and (P4+ℓP1, 2P2)-free for all k, ℓ, m, with improved χ-bounds for (P4+ℓP1, Kk)-free graphs.
Align anything: Training all-modality models to follow instructions with language feedback
3 Pith papers cite this work. Polarity classification is still indexing.
citation-role summary
citation-polarity summary
roles
background 1polarities
background 1representative citing papers
SafeVLA applies constrained reinforcement learning via CMDP min-max optimization to VLAs, cutting safety violation costs by 83.58% while preserving task success on long-horizon mobile manipulation tasks.
The survey organizes the shift of LLMs toward deliberate System 2 reasoning, covering model construction techniques, performance on math and coding benchmarks, and future research directions.
citing papers explorer
-
Identifying Topological Invariants of Non-Hermitian Systems via Domain-Adaptive Multimodal Model for Mathematics
There are only finitely many k-vertex-critical graphs in the classes (P4+ℓP1, B4(m), B3(m)+)-free and (P4+ℓP1, 2P2)-free for all k, ℓ, m, with improved χ-bounds for (P4+ℓP1, Kk)-free graphs.
-
SafeVLA: Towards Safety Alignment of Vision-Language-Action Model via Constrained Learning
SafeVLA applies constrained reinforcement learning via CMDP min-max optimization to VLAs, cutting safety violation costs by 83.58% while preserving task success on long-horizon mobile manipulation tasks.
-
From System 1 to System 2: A Survey of Reasoning Large Language Models
The survey organizes the shift of LLMs toward deliberate System 2 reasoning, covering model construction techniques, performance on math and coding benchmarks, and future research directions.