Sciweavers

4544 search results - page 165 / 909
» Reinforcement Learning with Time
Sort
View
CORR
2012
Springer
216views Education» more  CORR 2012»
12 years 5 months ago
Fractional Moments on Bandit Problems
Reinforcement learning addresses the dilemma between exploration to find profitable actions and exploitation to act according to the best observations already made. Bandit proble...
Ananda Narayanan B., Balaraman Ravindran
ROBOCUP
2004
Springer
114views Robotics» more  ROBOCUP 2004»
14 years 3 months ago
Modular Learning System and Scheduling for Behavior Acquisition in Multi-agent Environment
The existing reinforcement learning approaches have been suffering from the policy alternation of others in multiagent dynamic environments such as RoboCup competitions since othe...
Yasutake Takahashi, Kazuhiro Edazawa, Minoru Asada
IJCAI
2007
13 years 11 months ago
Effective Control Knowledge Transfer through Learning Skill and Representation Hierarchies
Learning capabilities of computer systems still lag far behind biological systems. One of the reasons can be seen in the inefficient re-use of control knowledge acquired over the...
Mehran Asadi, Manfred Huber
ECML
2006
Springer
14 years 1 months ago
Approximate Policy Iteration for Closed-Loop Learning of Visual Tasks
Abstract. Approximate Policy Iteration (API) is a reinforcement learning paradigm that is able to solve high-dimensional, continuous control problems. We propose to exploit API for...
Sébastien Jodogne, Cyril Briquet, Justus H....
FLAIRS
2000
13 years 11 months ago
Resolving Conflicts Among Actions in Concurrent Behaviors
A robotic agent must coordinate its coupled concurrent behaviors to produce a coherent response to stimuli. Reinforcement learning has been used extensively in coordinating sensin...
Henry Hexmoor