q-learning 4 Discounted vs. Average Reward Reinforcement Learning Jun 9, 2026 Why Do Neural TD Converge ? Jun 2, 2026 Function Approximation in RL: From Tables to Linear Models to Neural Networks May 27, 2026 Why Vanilla Q-Learning Breaks Under Corrupted Rewards Apr 12, 2026