計測自動制御学会論文集
Online ISSN : 1883-8189
Print ISSN : 0453-4654
ISSN-L : 0453-4654
Q学習に基づいた自動スリープシステムの最適制御
岡村 寛之石倉 武土肥 正
著者情報
ジャーナル フリー

2003 年 39 巻 6 号 p. 590-599

詳細
抄録
Dynamic power management (DPM) is one of the most effective techniques for reducing energy consumption. Especially, auto-sleep function is the simplest but effective way to reduce electric power consumption. In this paper, we consider a stochastic model to determine an auto-sleep timing sequentially. More precisely, we develop an optimal control scheme of auto-sleep timing based on the Q-learning, which is a part of reinforcement learning algorithms and is strictly related to the Markov decision process (MDP) and the semi-Markov decision process (SMDP). First, we reformulate the stochastic auto-sleep model under the SMDP. Second, the optimal control scheme is to determine the optimal auto-sleep timing established by applying the Q-learning algorithm. Finaly, numerical examples are presented to investigate the effectiveness of DPM with real data.
著者関連情報
© 社団法人 計測自動制御学会
前の記事 次の記事
feedback
Top