Cantitate/Preț
Produs

Reinforcement Learning with History Lists

Autor Stephan Timmer
de Limba Germană Paperback – 21 apr 2009
A very general framework for modeling uncertainty in learning environments is given by Partially observable Markov Decision Processes (POMDPs). In a POMDP setting, the learning agent infers a policy for acting optimally in all possible states of the environment, while receiving only observations of these states. The basic idea for coping with partial observability is to include memory into the representation of the policy. Perfect memory is provided by the belief space, i.e. the space of probability distributions over environmental states. However, computing policies defined on the belief space requires a considerable amount of prior knowledge about the learning problem and is expensive in terms of computation time.The author Stephan Timmer presents a reinforcement learning algorithm for solving POMDPs based on short term memory. In contrast to belief states, short term memory is not capable of representing optimal policies, but is far more practical and requires no prior knowledge about the learning problem. It can be shown that the algorithm can also be used to solve large Markov Decision Processes (MDPs) with continuous, multi-dimensional state spaces.
Citește tot Restrânge

Preț: 44034 lei

Preț vechi: 55042 lei
-20%

Puncte Express: 661

Carte tipărită la comandă

Livrare economică 05-19 octombrie

Livrare prin curier în România Termenul estimat este afișat lângă disponibilitate.
Transport gratuit pentru acest produs Plată online sau ramburs, în funcție de opțiunile comenzii.
Retur gratuit în 14 zile Comandă securizată și suport în română.

Specificații

ISBN-13: 9783838106212
ISBN-10: 3838106210
Pagini: 160
Dimensiuni: 150 x 220 x 11 mm
Greutate: 0.26 kg
Editura: Südwestdeutscher Verlag für Hochschulschriften
Locul publicării:Germany