Hyoungseo Son
← All topics

5 of 5 written

Reinforcement Learning Theory

The math spine of RL: MDPs and value functions, the Bellman operators, policy gradients, and the trust-region and natural-gradient methods that stabilize them.