Value learning from trajectory optimization and Sobolev descent A step toward reinforcement learning with superlinear convergence properties

Value learning from trajectory optimization and Sobolev descent A step toward reinforcement learning with superlinear convergence properties

Value learning from trajectory optimization and Sobolev descent A step toward reinforcement learning with superlinear convergence properties icon

This paper combines model-based trajectory optimization and Sobolev learning to accelerate value function estimation for reinforcement learning.

Videos

Video placeholder

Figures

Figure placeholder