Online actor–critic algorithm to solve the continuous-time infinite horizon optimal control problem

کد مقاله	کد نشریه	سال انتشار	مقاله انگلیسی	نسخه تمام متن
697563	890374	2010	11 صفحه PDF	دانلود رایگان

عنوان انگلیسی مقاله ISI

دانلود مقاله + سفارش ترجمه

دانلود مقاله ISI انگلیسی

رایگان برای ایرانیان

کلمات کلیدی

LQR Adaptive critics - منتقدان انطباقی Persistence of excitation - پایداری تحریک Adaptive control - کنترل تطبیقی

موضوعات مرتبط

مهندسی و علوم پایه سایر رشته های مهندسی کنترل و سیستم های مهندسی

پیش نمایش صفحه اول مقاله

Online actor–critic algorithm to solve the continuous-time infinite horizon optimal control problem

چکیده انگلیسی

In this paper we discuss an online algorithm based on policy iteration for learning the continuous-time (CT) optimal control solution with infinite horizon cost for nonlinear systems with known dynamics. That is, the algorithm learns online in real-time the solution to the optimal control design HJ equation. This method finds in real-time suitable approximations of both the optimal cost and the optimal control policy, while also guaranteeing closed-loop stability. We present an online adaptive algorithm implemented as an actor/critic structure which involves simultaneous continuous-time adaptation of both actor and critic neural networks. We call this ‘synchronous’ policy iteration. A persistence of excitation condition is shown to guarantee convergence of the critic to the actual optimal value function. Novel tuning algorithms are given for both critic and actor networks, with extra nonstandard terms in the actor tuning law being required to guarantee closed-loop dynamical stability. The convergence to the optimal controller is proven, and the stability of the system is also guaranteed. Simulation examples show the effectiveness of the new algorithm.

ناشر

Database: Elsevier - ScienceDirect (ساینس دایرکت)
Journal: Automatica - Volume 46, Issue 5, May 2010, Pages 878–888

نویسندگان

Kyriakos G. Vamvoudakis, Frank L. Lewis,

علوم انسانی و هنر

فنی، مهندسی و علوم پایه

پزشکی و سلامت

بیو تکنولوژی

پذیرش سفارش ترجمه

Online actor–critic algorithm to solve the continuous-time infinite horizon optimal control problem

دسترسی سریع

ارتباط

English Website