Pith. sign in

Title resolution pending

1 Pith paper cite this work, alongside 539 external citations. Polarity classification is still indexing.

1 Pith paper citing it
539 external citations · OpenAlex

fields

cs.LG 1

years

2026 1

verdicts

CONDITIONAL 1

representative citing papers

Adaptive Multi-Horizon Reinforcement Learning

cs.LG · 2026-07-22 · conditional · novelty 5.0

A gated mixture of Q-functions with different discount factors, trained with undiscounted Bellman error, adapts its temporal horizon in small MiniGrid tasks, but its theoretical justification is circular and baseline comparisons are missing.

citing papers explorer

Showing 1 of 1 citing paper.

  • Adaptive Multi-Horizon Reinforcement Learning cs.LG · 2026-07-22 · conditional · none · ref 50

    A gated mixture of Q-functions with different discount factors, trained with undiscounted Bellman error, adapts its temporal horizon in small MiniGrid tasks, but its theoretical justification is circular and baseline comparisons are missing.