Fr. 189.60

Sridhar Mahadaven, Sridhar Mahadevan

Learning Representation and Control in Markov Decision Processes

English · Paperback / Softback

Shipping usually within 2 to 3 weeks (title will be printed to order)

Description

Learning Representation and Control in Markov Decision Processes describes methods for automatically compressing Markov decision processes (MDPs) by learning a low-dimensional linear approximation defined by an orthogonal set of basis functions. A unique feature of the text is the use of Laplacian operators, whose matrix representations have non-positive off-diagonal elements and zero row sums. The generalized inverses of Laplacian operators, in particular the Drazin inverse, are shown to be useful in the exact and approximate solution of MDPs.

The author goes on to describe a broad framework for solving MDPs, generically referred to as representation policy iteration (RPI), where both the basis function representations for approximation of value functions as well as the optimal policy within their linear span are simultaneously learned. Basis functions are constructed by diagonalizing a Laplacian operator or by dilating the reward function or an initial set of bases by powers of the operator. The idea of decomposing an operator by finding its invariant subspaces is shown to be an important principle in constructing low-dimensional representations of MDPs. Theoretical properties of these approaches are discussed, and they are also compared experimentally on a variety of discrete and continuous MDPs. Finally, challenges for further research are briefly outlined.

Learning Representation and Control in Markov Decision Processes is a timely exposition of a topic with broad interest within machine learning and beyond.

List of contents

1: Introduction 2: Sequential Decision Problems 3: Laplacian Operators and MDPs 4: Approximating Markov Decision Processes 5: Dimensionality Reduction Principles in MDPs 6: Basis Construction: Diagonalization Methods 7: Basis Construction: Dilation Methods 8: Model-Based Representation Policy Iteration 9: Basis Construction in Continuous MDPs 10: Model-Free Representation Policy Iteration 11: Related Work and Future Challenges. References.

Product details

Authors	Sridhar Mahadaven, Sridhar Mahadevan
Publisher	Now Publishers Inc

Languages	English
Product format	Paperback / Softback
Released	02.06.2009

EAN	9781601982384
ISBN	978-1-60198-238-4
No. of pages	184
Dimensions	156 mm x 234 mm x 10 mm
Weight	289 g
Series	Foundations and Trends(r) in M Foundations and Trends (R) in Machine Learning
Subjects	Guides Natural sciences, medicine, IT, technology > IT, data processing > IT

Customer reviews

No reviews have been written for this item yet. Write the first review and be helpful to other users when they decide on a purchase.