Pseudo-MDPs and factored linear action models

2014
2014 IEEE Symposium on Adaptive Dynamic Programming and Reinforcement Learning (ADPRL)
In this paper we introduce the concept of pseudo-MDPs to develop abstractions. Pseudo-MDPs relax the requirement that the transition kernel has to be a probability kernel. We show that the new framework captures many existing abstractions. We also introduce the concept of factored linear action models; a special case. Again, the relation of factored linear action models and existing works are discussed. We use the general framework to develop a theory for bounding the suboptimality of policies

