Sciweavers

Free Online Productivity Tools i2Speak i2Symbol i2OCR iTex2Img iWeb2Print iWeb2Shot i2Type iPdf2Split iPdf2Merge i2Bopomofo i2Arabic i2Style i2Image i2PDF iLatex2Rtf Sci2ools

33

ALT
2007
Springer

favoriteEmaildiscussreport

119views Machine Learning» more ALT 2007»

Pseudometrics for State Aggregation in Average Reward Markov Decision Processes

14 years 6 months ago

Pseudometrics for State Aggregation in Average Reward Markov Decision Processes

Download personal.unileoben.ac.at

We consider how state similarity in average reward Markov decision processes (MDPs) may be described by pseudometrics. Introducing the notion of adequate pseudometrics which are well adapted to the structure of the MDP, we show how these may be used for state aggregation. Upper bounds on the loss that may be caused by working on the aggregated instead of the original MDP are given and compared to the bounds that have been achieved for discounted reward MDPs.

Ronald Ortner

Real-time Traffic

ALT 2007 | Average Reward Markov | Machine Learning | Original Mdp | State Similarity |

claim paper

Related Content

» Adaptive Stepsize Policy Gradients with Average Reward Metric

» Complexity of Probabilistic Planning under Average Rewards

» Symblicit Calculation of LongRun Averages for Concurrent Probabilistic Systems

» Approximation Algorithms for PartialInformation Based Stochastic Control with Markovian Re...

» State Space Reduction For Hierarchical Reinforcement Learning

» On step sizes stochastic shortest paths and survival probabilities in Reinforcement Learni...

Post Info
More Details (n/a)

Added	14 Mar 2010
Updated	14 Mar 2010
Type	Conference
Year	2007
Where	ALT
Authors	Ronald Ortner

Comments (0)