CASIA OpenIR
(本次检索基于用户作品认领结果)

浏览/检索结果: 共18条,第1-10条 帮助

限定条件        
已选(0)清除 条数/页:   排序方式:
Multi-Agent Reinforcement Learning Based on Clustering in Two-Player Games 会议论文
, Xiamen, China, 2019-12-6
作者:  Li WF(李伟凡);  Zhu YH(朱圆恒);  Zhao DB(赵冬斌)
Adobe PDF(488Kb)  |  收藏  |  浏览/下载:110/35  |  提交时间:2023/06/28
reinforcement learning  unsupervised clustering  matrix game  
A Hierarchical Deep Reinforcement Learning Framework for 6-DOF UCAV Air-to-Air Combat 期刊论文
IEEE Transactions on Systems, Man and Cybernetics: Systems, 2023, 页码: DOI: 10.1109/TSMC.2023.3270444
作者:  Jiajun Chai;  Wenzhang Chen;  Yuanheng Zhu;  Zong-xin Yao,;  Dongbin Zhao
Adobe PDF(9249Kb)  |  收藏  |  浏览/下载:201/108  |  提交时间:2023/04/26
Empirical Policy Optimization for n-Player Markov Games 期刊论文
IEEE Transactions on Cybernetics, 2022, 页码: doi={10.1109/TCYB.2022.3179775}
作者:  Yuanheng Zhu;  Weifan Li;  Mengchen Zhao;  Jianye Hao;  Dongbin Zhao
Adobe PDF(1739Kb)  |  收藏  |  浏览/下载:90/36  |  提交时间:2023/04/26
Online Minimax Q Network Learning for Two-Player Zero-Sum Markov Games 期刊论文
IEEE TRANSACTIONS ON NEURAL NETWORKS AND LEARNING SYSTEMS, 2022, 卷号: 33, 期号: 3, 页码: 1228-1241
作者:  Zhu, Yuanheng;  Zhao, Dongbin
收藏  |  浏览/下载:197/0  |  提交时间:2022/06/10
Games  Nash equilibrium  Mathematical model  Markov processes  Convergence  Dynamic programming  Training  Deep reinforcement learning (DRL)  generalized policy iteration (GPI)  Markov game (MG)  Nash equilibrium  Q network  zero sum  
Hierarchical optimal control for input-affine nonlinear systems through the formulation of Stackelberg game 期刊论文
INFORMATION SCIENCES, 2020, 卷号: 517, 页码: 1-17
作者:  Mu, Chaoxu;  Wang, Ke;  Zhang, Qichao;  Zhao, Dongbin
收藏  |  浏览/下载:207/0  |  提交时间:2020/04/07
Nonzero-sum differential game  Hierarchical optimization  Nonlinear dynamics  Stackelberg equilibrium  Neural network  
Value Iteration Algorithm for Optimal Consensus Control of Multi-agent Systems 会议论文
, Siem Reap, Cambodia, Dec.14-16
作者:  Zhang Qichao;  Zhao Dongbin
收藏  |  浏览/下载:83/0  |  提交时间:2019/10/09
Data-Based Reinforcement Learning for Nonzero-Sum Games With Unknown Drift Dynamics 期刊论文
IEEE TRANSACTIONS ON CYBERNETICS, 2019, 卷号: 49, 期号: 8, 页码: 2874-2885
作者:  Zhang, Qichao;  Zhao, Dongbin
浏览  |  Adobe PDF(1021Kb)  |  收藏  |  浏览/下载:409/120  |  提交时间:2019/07/12
Integral reinforcement learning (IRL)  neural network (NN)  nonzero-sum (NZS) games  off-policy  single-critic  unknown drift dynamics  
Event-triggered hinfinity control for continuous-time nonlinear system 会议论文
, *, 2015
作者:  Zhao,Dongbin(赵冬斌);  Zhang,Qichao;  Li,Xiangjun;  Kong,Lingda
浏览  |  Adobe PDF(365Kb)  |  收藏  |  浏览/下载:233/84  |  提交时间:2018/01/04
Event-Triggered H∞ Control for Continuous-Time Nonlinear System 会议论文
, Jeju, South Korea, October 15-18
作者:  Zhao,Dongbin;  Zhang,Qichao;  Li,Xiangjun;  Kong,Lingda
浏览  |  Adobe PDF(365Kb)  |  收藏  |  浏览/下载:173/43  |  提交时间:2017/12/28
Off-Policy Reinforcement Learning for Partially Unknown Nonzero-Sum Games 会议论文
, Guangzhou China, November 14–18
作者:  Zhang,Qichao;  Zhao,Dongbin;  Zhang,Sibo
浏览  |  Adobe PDF(119Kb)  |  收藏  |  浏览/下载:241/85  |  提交时间:2017/12/28