CASIA OpenIR

浏览/检索结果: 共8条,第1-8条 帮助

限定条件    
已选(0)清除 条数/页:   排序方式:
Continuous-Time Time-Varying Policy Iteration 期刊论文
IEEE TRANSACTIONS ON CYBERNETICS, 2020, 卷号: 50, 期号: 12, 页码: 4958-4971
作者:  Wei, Qinglai;  Liao, Zehua;  Yang, Zhanyu;  Li, Benkai;  Liu, Derong
Adobe PDF(3149Kb)  |  收藏  |  浏览/下载:241/49  |  提交时间:2021/03/02
Optimal control  Nonlinear systems  Time-varying systems  Mathematical model  Dynamic programming  Approximation algorithms  Iterative algorithms  Adaptive critic designs  adaptive dynamic programming (ADP)  neuro-dynamic programming  nonlinear systems  optimal control  policy iteration  
Data-Based Reinforcement Learning for Nonzero-Sum Games With Unknown Drift Dynamics 期刊论文
IEEE TRANSACTIONS ON CYBERNETICS, 2019, 卷号: 49, 期号: 8, 页码: 2874-2885
作者:  Zhang, Qichao;  Zhao, Dongbin
浏览  |  Adobe PDF(1021Kb)  |  收藏  |  浏览/下载:388/119  |  提交时间:2019/07/12
Integral reinforcement learning (IRL)  neural network (NN)  nonzero-sum (NZS) games  off-policy  single-critic  unknown drift dynamics  
Policy Iteration for H infinity Optimal Control of Polynomial Nonlinear Systems via Sum of Squares Programming 期刊论文
IEEE TRANSACTIONS ON CYBERNETICS, 2018, 卷号: 48, 期号: 2, 页码: 500-509
作者:  Zhu, Yuanheng;  Zhao, Dongbin;  Yang, Xiong;  Zhang, Qichao
Adobe PDF(892Kb)  |  收藏  |  浏览/下载:278/33  |  提交时间:2018/10/10
Adaptive Dynamic Programming (Adp)  h Infinity Optimal Control  Policy Iteration (Pi)  Polynomial Nonlinear Systems  Sum Of Squares (Sos)  
Learning A Superpixel-Driven Speed Function for Level Set Tracking 期刊论文
IEEE TRANSACTIONS ON CYBERNETICS, 2016, 卷号: 46, 期号: 7, 页码: 1498-1510
作者:  Zhou, Xue;  Li, Xi;  Hu, Weiming
浏览  |  Adobe PDF(2788Kb)  |  收藏  |  浏览/下载:368/120  |  提交时间:2016/10/20
Level Set Tracking  Metric Learning  Non-negative Matrix Factorization (Nf)  Speed Function  Superpixel (Sp)-driven  
Experience Replay for Optimal Control of Nonzero-Sum Game Systems With Unknown Dynamics 期刊论文
IEEE TRANSACTIONS ON CYBERNETICS, 2016, 卷号: 46, 期号: 3, 页码: 854-865
作者:  Zhao, Dongbin;  Zhang, Qichao;  Wang, Ding;  Zhu, Yuanheng
Adobe PDF(1769Kb)  |  收藏  |  浏览/下载:479/191  |  提交时间:2016/06/14
Adaptive Dynamic Programming (Adp)  Experience Replay  Nonzero-sum (Nzs) Games  Optimal Control  Unknown Dynamics  
Reinforcement-Learning-Based Robust Controller Design for Continuous-Time Uncertain Nonlinear Systems Subject to Input Constraints 期刊论文
IEEE TRANSACTIONS ON CYBERNETICS, 2015, 卷号: 45, 期号: 7, 页码: 1372-1385
作者:  Liu, Derong;  Yang, Xiong;  Wang, Ding;  Wei, Qinglai
Adobe PDF(1179Kb)  |  收藏  |  浏览/下载:461/242  |  提交时间:2015/09/17
Approximate Dynamic Programming (Adp)  Neural Networks (Nns)  Neuro-dynamic Programming  Nonlinear Systems  Optimal Control  Reinforcement Learning (Rl)  Robust Control  
Neural-Network-Based Online HJB Solution for Optimal Robust Guaranteed Cost Control of Continuous-Time Uncertain Nonlinear Systems 期刊论文
IEEE TRANSACTIONS ON CYBERNETICS, 2014, 卷号: 44, 期号: 12, 页码: 2834-2847
作者:  Liu, Derong;  Wang, Ding;  Wang, Fei-Yue;  Li, Hongliang;  Yang, Xiong
浏览  |  Adobe PDF(780Kb)  |  收藏  |  浏览/下载:429/203  |  提交时间:2015/08/12
Adaptive Critic Designs  Adaptive/approximate Dynamic Programming (Adp)  Hamilton-jacobi-bellman (Hjb) Equation  Neural Networks  Optimal Robust Guaranteed Cost Control  Uncertain Nonlinear Systems  
Finite-Approximation-Error-Based Discrete-Time Iterative Adaptive Dynamic Programming 期刊论文
IEEE TRANSACTIONS ON CYBERNETICS, 2014, 卷号: 44, 期号: 12, 页码: 2820-2833
作者:  Wei, Qinglai;  Wang, Fei-Yue;  Liu, Derong;  Yang, Xiong
浏览  |  Adobe PDF(1826Kb)  |  收藏  |  浏览/下载:292/109  |  提交时间:2015/08/12
Adaptive Critic Designs  Adaptive Dynamic Programming (Adp)  Approximate Dynamic Programming  Approximation Error  Neural Networks  Neuro-dynamic Programming  Nonlinear Systems  Optimal Control  Reinforcement Learning  Value Iteration