Maximum Power Point Tracking of Photovoltaic System Based on Reinforcement Learning

The maximum power point tracking (MPPT) technique is often used in photovoltaic (PV) systems to extract the maximum power in various environmental conditions. The perturbation and observation (P&O) method is one of the most well-known MPPT methods; however, it may face problems of large oscillations around maximum power point (MPP) or low-tracking efficiency. In this paper, two reinforcement learning-based maximum power point tracking (RL MPPT) methods are proposed by the use of the Q-learning algorithm. One constructs the Q-table and the other adopts the Q-network. These two proposed methods do not require the information of an actual PV module in advance and can track the MPP through offline training in two phases, the learning phase and the tracking phase. From the experimental results, both the reinforcement learning-based Q-table maximum power point tracking (RL-QT MPPT) and the reinforcement learning-based Q-network maximum power point tracking (RL-QN MPPT) methods have smaller ripples and faster tracking speeds when compared with the P&O method. In addition, for these two proposed methods, the RL-QT MPPT method performs with smaller oscillation and the RL-QN MPPT method achieves higher average power.

[1]  Rachid Chenni,et al.  Fuzzy Logic-Based Perturb and Observe Algorithm with Variable Step of a Reference Voltage for Solar Permanent Magnet Synchronous Motor Drive System Fed by Direct-Connected Photovoltaic Array , 2018 .

[2]  Sebastian Ruder,et al.  An overview of gradient descent optimization algorithms , 2016, Vestnik komp'iuternykh i informatsionnykh tekhnologii.

[3]  Hocine Belmili,et al.  A survey of the most used MPPT methods: Conventional and advanced algorithms applied for photovoltaic systems , 2015 .

[4]  Peter Dayan,et al.  Q-learning , 1992, Machine Learning.

[5]  George A. Vouros,et al.  A reinforcement learning approach for MPPT control method of photovoltaic sources , 2017 .

[6]  Wen-Yen Chen,et al.  A Reinforcement Learning-Based Maximum Power Point Tracking Method for Photovoltaic Array , 2015 .

[7]  ICME 11-RT,et al.  PERFORMANCE EVALUATION OF 1.68 kWp DC OPERATED SOLAR PUMP WITH AUTO TRACKER USING MICROCONTROLLER BASED DATA ACQUISITION SYSTEM , 2011 .

[8]  Ayman Youssef,et al.  Reinforcement Learning for Online Maximum Power Point Tracking Control , .

[9]  Ronald A. Howard,et al.  Dynamic Programming and Markov Processes , 1960 .

[10]  Shane Legg,et al.  Human-level control through deep reinforcement learning , 2015, Nature.

[11]  Andres Barrado,et al.  Review of the maximum power point tracking algorithms for stand-alone photovoltaic systems , 2006 .

[12]  N. A. Rahim,et al.  Adaptive P&O-fuzzy control MPPT for PV boost dc-dc converter , 2012, 2012 IEEE International Conference on Power and Energy (PECon).

[13]  Weidong Xiao,et al.  A modified adaptive hill climbing MPPT method for photovoltaic power systems , 2004, 2004 IEEE 35th Annual Power Electronics Specialists Conference (IEEE Cat. No.04CH37551).

[14]  Saad Mekhilef,et al.  Solar cell parameters extraction based on single and double-diode models: A review , 2016 .

[15]  Long-Ji Lin,et al.  Reinforcement learning for robots using neural networks , 1992 .

[16]  J. So,et al.  Improved perturbation and observation method (IP&O) of MPPT control for photovoltaic power systems , 2005, Conference Record of the Thirty-first IEEE Photovoltaic Specialists Conference, 2005..

[17]  Eduardo F. Morales,et al.  An Introduction to Reinforcement Learning , 2011 .

[18]  R. Bellman A Markovian Decision Process , 1957 .

[19]  Z. Salam,et al.  An improved perturb and observe (P&O) maximum power point tracking (MPPT) algorithm for higher efficiency , 2015 .

[20]  Guigang Zhang,et al.  Deep Learning , 2016, Int. J. Semantic Comput..