• DocumentCode
    585742
  • Title

    An online integral reinforcement learning algorithm to solve N-player Nash games

  • Author

    Vamvoudakis, Kyriakos G. ; Lewis, Frank L.

  • Author_Institution
    Center for Control, Dynamical-Syst. & Comput. (CCDC), Univ. of California, Santa Barbara, Santa Barbara, CA, USA
  • fYear
    2012
  • fDate
    3-5 Oct. 2012
  • Firstpage
    697
  • Lastpage
    702
  • Abstract
    In this paper we introduce an online algorithm that uses integral reinforcement knowledge for learning the continuous-time Nash game (zero-sum and non-zero-sum) solution for nonlinear systems with infinite horizon costs and partial knowledge of the system dynamics. This algorithm is a data based approach to the solution of the coupled Hamilton-Jacobi equations and it does not require explicit knowledge on the system´s drift dynamics. A novel adaptive control algorithm is given that is based on policy iteration and implemented using an actor/critic structure for every player in the game having 2N adaptive approximator structures. All 2N approximation networks are adapted simultaneously. Novel adaptive control tuning algorithms are given for the critic and actor networks. The convergence to the Nash solution of the game is proven, and stability of the system is also guaranteed. Simulation examples support the theoretical result.
  • Keywords
    adaptive control; continuous time systems; convergence; game theory; geometry; iterative methods; learning (artificial intelligence); nonlinear systems; tuning; 2N adaptive approximator structures; N-player Nash games; actor neural network structure; continuous-time Nash game solution; coupled Hamilton-Jacobi equations; critic neural networks; infinite horizon costs; nonlinear systems; novel adaptive control tuning algorithms; online integral reinforcement learning algorithm; policy iteration; reinforcement knowledge; stability; system dynamics; Approximation algorithms; Artificial neural networks; Equations; Games; Heuristic algorithms; Mathematical model; Tuning; Coupled Hamilton-Jacobi equations; Coupled Riccati equations; Nash equilibrium; Nash games; integral reinforcement learning;
  • fLanguage
    English
  • Publisher
    ieee
  • Conference_Titel
    Intelligent Control (ISIC), 2012 IEEE International Symposium on
  • Conference_Location
    Dubrovnik
  • ISSN
    2158-9860
  • Print_ISBN
    978-1-4673-4598-9
  • Electronic_ISBN
    2158-9860
  • Type

    conf

  • DOI
    10.1109/ISIC.2012.6398248
  • Filename
    6398248