Pith. sign in

Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor

1 Pith paper cite this work. Polarity classification is still indexing.

1 Pith paper citing it

fields

cs.LG 1

years

2025 1

verdicts

REJECT 1

representative citing papers

Learning from Less: SINDy Surrogates in RL

cs.LG · 2025-04-25 · reject · novelty 3.0

SINDy-fitted surrogate environments for Mountain Car and Lunar Lander reproduce state dynamics from 75 to 1,000 samples and train RL agents with fewer steps, though policy transfer to the real environments is unquantified.

citing papers explorer

Showing 1 of 1 citing paper.

  • Learning from Less: SINDy Surrogates in RL cs.LG · 2025-04-25 · reject · none · ref 4

    SINDy-fitted surrogate environments for Mountain Car and Lunar Lander reproduce state dynamics from 75 to 1,000 samples and train RL agents with fewer steps, though policy transfer to the real environments is unquantified.