Sciweavers

Free Online Productivity Tools i2Speak i2Symbol i2OCR iTex2Img iWeb2Print iWeb2Shot i2Type iPdf2Split iPdf2Merge i2Bopomofo i2Arabic i2Style i2Image i2PDF iLatex2Rtf Sci2ools

108

ICML
2003
IEEE

favoriteEmaildiscussreport

105views Machine Learning» more ICML 2003»

Principled Methods for Advising Reinforcement Learning Agents

16 years 2 months ago

Principled Methods for Advising Reinforcement Learning Agents

Download www.hpl.hp.com

An important issue in reinforcement learning is how to incorporate expert knowledge in a principled manner, especially as we scale up to real-world tasks. In this paper, we present a method for incorporating arbitrary advice into the reward structure of a reinforcement learning agent without altering the optimal policy. This method extends the potentialbased shaping method proposed by Ng et al. (1999) to the case of shaping functions based on both states and actions. This allows for much more specific information to guide the agent ? which action to choose ? without requiring the agent to discover this from the rewards on states alone. We develop two qualitatively different methods for converting a potential function into advice for the agent. We also provide theoretical and experimental justifications for choosing between these advice-giving algorithms based on the properties of the potential function.

Eric Wiewiora, Garrison W. Cottrell, Charles Elkan

Real-time Traffic

ICML 2003 | Machine Learning | Potential Function | Potentialbased Shaping Method | Reinforcement Learning Agent |

claim paper

Related Content

» Lyapunov Design for Safe Reinforcement Learning

» Preference elicitation and inverse reinforcement learning

» Learning Methods to Generate Good Plans Integrating HTN Learning and Reinforcement Learnin...

» Skill Combination for Reinforcement Learning

» Asymmetric Multiagent Reinforcement Learning

» Reinforcement learning agents with primary knowledge designed by analytic hierarchy proces...

» Reinforcement Learning via AIXI Approximation

» General Principles of LearningBased MultiAgent Systems

» Critical factors in the empirical performance of temporal difference and evolutionary meth...

Post Info
More Details (n/a)

Added	17 Nov 2009
Updated	17 Nov 2009
Type	Conference
Year	2003
Where	ICML
Authors	Eric Wiewiora, Garrison W. Cottrell, Charles Elkan

Comments (0)