Pith. sign in

Title resolution pending

1 Pith paper cite this work. Polarity classification is still indexing.

1 Pith paper citing it

fields

cs.LG 1

years

2025 1

verdicts

REJECT 1

representative citing papers

RLAE: Reinforcement Learning-Assisted Ensemble for LLMs

cs.LG · 2025-05-31 · reject · novelty 6.0

RLAE uses PPO and MAPPO policies to assign per-span ensemble weights across 7B-8B LLMs; it claims up to 3.3% accuracy improvement over prior ensemble baselines but underperforms on several tested tasks.

citing papers explorer

Showing 1 of 1 citing paper.

  • RLAE: Reinforcement Learning-Assisted Ensemble for LLMs cs.LG · 2025-05-31 · reject · none · ref 6

    RLAE uses PPO and MAPPO policies to assign per-span ensemble weights across 7B-8B LLMs; it claims up to 3.3% accuracy improvement over prior ensemble baselines but underperforms on several tested tasks.