Pith. sign in

Soft actor-critic: Off-policy maximum entropy deep reinforcement learning with a stochastic actor

1 Pith paper cite this work. Polarity classification is still indexing.

1 Pith paper citing it

fields

cs.LG 1

years

2025 1

verdicts

CONDITIONAL 1

representative citing papers

Fine-Tuning without Performance Degradation

cs.LG · 2025-05-01 · conditional · novelty 6.0

Automatic Jump Start uses Fitted Q Evaluation to adapt the Jump-Start exploration schedule, reducing fine-tuning performance degradation without tuning a tolerance threshold.

citing papers explorer

Showing 1 of 1 citing paper.

  • Fine-Tuning without Performance Degradation cs.LG · 2025-05-01 · conditional · none · ref 6

    Automatic Jump Start uses Fitted Q Evaluation to adapt the Jump-Start exploration schedule, reducing fine-tuning performance degradation without tuning a tolerance threshold.