計測自動制御学会論文集
Online ISSN : 1883-8189
Print ISSN : 0453-4654
ISSN-L : 0453-4654
連続値入出力を扱うファジィ内挿型Q-Learningの提案
堀内 匡藤野 昭典片井 修椹木 哲夫
著者情報
ジャーナル フリー

1999 年 35 巻 2 号 p. 271-279

詳細
抄録
Reinforcement learning is an essential class of machine learning for autonomous agents to acquire adaptive and reactive behaviors. Q-Learning is a widely-used reinforcement learning method which deals with only discretevalued inputs (states) and outputs (actions). In this paper, we propose a new method of Q-Learning where fuzzy inference is introduced to represent Q-function (action value function) that evaluates state/action pairs, in order to to deal with continuous-valued inputs and outputs. Within this method, the steepest descent method is utilized to update Q-function so that the parameters of fuzzy rules are tuned both in antecedent part and consequent part. Furthermore, we extend this method to the case of discrete actions by preparing a fuzzy inference system for each action in order to realize the speed up of learning. Finally we show the effectiveness of this method by comparing it with other methods through the applications to control problems such as the cart-pole balancing problem and the ship navigation problem.
著者関連情報
© 社団法人 計測自動制御学会
前の記事 次の記事
feedback
Top