Preprint: Reinforcement learning handles larger deviations in a canal simulation
A backstepping-guided SAC controller performed best among learning-based methods when the simulated canal began far from equilibrium, although analytical backstepping failed in that test.