Home
News
Members
Projects
Publications
Contact
Light
Dark
Automatic
Human Intervention
Reinforcement Learning with a Terminator
We present the problem of reinforcement learning with exogenous termination. We define the Termination Markov Decision Process (TerMDP), an extension of the MDP framework, in which episodes may be interrupted by an external non-Markovian observer. …
Cite
×