Gradient-based Planning for World Models at Longer Horizons
A world model can predict future states yet remain hard to plan through. GRASP makes long-horizon optimization less brittle by relaxing trajectories and avoiding unreliable state gradients. Read the Push-T comparison for the gain and its simulation-bound limit.