EfficientTDMPC: Improved MPC Objectives for Sample-Efficient Continuous Control

We introduce EfficientTDMPC, a sample-efficient model-based reinforcement learning method for continuous control built on the TD-MPC family of algorithms. Central to this family is a planner that aims to find an action sequence that maximizes the estimated return. EfficientTDMPC proposes to reduce this error in two ways.