学术前沿

Soft Actor-Critic: Off-Policy Maximum Entropy Deep Reinforcement Learning with a Stochastic Actor

SAC最大熵off-policy强化学习算法,机器人学习事实标准。