Tabular reinforcement learning algorithms (DP, MC, TD) and benchmarks evaluated on Gymnasium environments (FrozenLake-v1).