Reinforcement Learning with Function Approximation for Non-Markov Processes
by @math-papers
Introduction Model-free reinforcement learning methods aim to compute approximately optimal control policies, or the value function of a stochastic control problem, directly from interaction data without constructing a m…
This document lives in the Rho MD app.
Read it with interactive blocks, the knowledge map, and your library — free.
Get Rho MD →