已选(0)清除
条数/页: 排序方式: |
| Soft Contrastive Learning with Q-irrelevance Abstraction for Reinforcement Learning 期刊论文 IEEE Transactions on Cognitive and Developmental Systems, 2023, 卷号: 15, 期号: 3, 页码: 1463 - 1473 作者: Liu MS(刘民颂) ; Li LT(李伦通); Hao S(郝帅); Zhu YH(朱圆恒) ; Zhao DB(赵冬斌)![](/image/person.jpg)
Adobe PDF(4197Kb)  |   收藏  |  浏览/下载:13/3  |  提交时间:2024/06/24 |
| FM3Q: Factorized Multi-Agent MiniMax Q-Learning for Two-Team Zero-Sum Markov Game 期刊论文 IEEE Transactions on Emerging Topics in Computational Intelligence, 2024, 页码: 1-13 作者: Guangzheng Hu; Yuanheng Zhu ; Haoran Li; Dongbin Zhao![](/image/person.jpg)
Adobe PDF(2144Kb)  |   收藏  |  浏览/下载:16/2  |  提交时间:2024/06/05 |
| A Survey on Reinforcement Learning Methods in Bionic Underwater Robots 期刊论文 BIOMIMETICS, 2023, 卷号: 8, 期号: 2, 页码: 29 作者: Tong, Ru ; Feng, Yukai ; Wang, Jian ; Wu, Zhengxing ; Tan, Min ; Yu, Junzhi![](/image/person.jpg)
Adobe PDF(1260Kb)  |   收藏  |  浏览/下载:122/12  |  提交时间:2023/11/17 bionic underwater robot reinforcement learning robotic fish intelligent control |
| NVIF: Neighboring Variational Information Flow for Cooperative Large-Scale Multiagent Reinforcement Learning 期刊论文 IEEE TRANSACTIONS ON NEURAL NETWORKS AND LEARNING SYSTEMS, 2023, 页码: 13 作者: Chai, Jiajun ; Zhu, Yuanheng ; Zhao, Dongbin![](/image/person.jpg)
Adobe PDF(2469Kb)  |   收藏  |  浏览/下载:60/2  |  提交时间:2023/11/16 Large-scale multiagent neighboring communication reinforcement learning (RL) variational information flow |
| Advantage Constrained Proximal Policy Optimization in Multi-Agent Reinforcement Learning 会议论文 , 昆士兰, 2023-6 作者: Li WF(李伟凡) ; Zhu YH(朱圆恒) ; Zhao DB(赵冬斌)![](/image/person.jpg)
Adobe PDF(4104Kb)  |   收藏  |  浏览/下载:242/79  |  提交时间:2023/06/29 multi-agent reinforcement learning policy gradient |
| A Hierarchical Deep Reinforcement Learning Framework for 6-DOF UCAV Air-to-Air Combat 期刊论文 IEEE Transactions on Systems, Man and Cybernetics: Systems, 2023, 页码: DOI: 10.1109/TSMC.2023.3270444 作者: Jiajun Chai ; Wenzhang Chen; Yuanheng Zhu ; Zong-xin Yao,; Dongbin Zhao![](/image/person.jpg)
Adobe PDF(9249Kb)  |   收藏  |  浏览/下载:266/118  |  提交时间:2023/04/26 |
| Soft Contrastive Learning with Q-irrelevance Abstraction for Reinforcement Learning 期刊论文 IEEE Transactions on Cognitive and Developmental Systems, 2022, 页码: doi={10.1109/TCDS.2022.3218940} 作者: Minsong Liu ; Luntong Li; Shuai Hao; Yuanheng Zhu ; Dongbin Zhao![](/image/person.jpg)
Adobe PDF(12013Kb)  |   收藏  |  浏览/下载:84/22  |  提交时间:2023/04/26 |
| Empirical Policy Optimization for n-Player Markov Games 期刊论文 IEEE Transactions on Cybernetics, 2022, 页码: doi={10.1109/TCYB.2022.3179775} 作者: Yuanheng Zhu ; Weifan Li ; Mengchen Zhao; Jianye Hao; Dongbin Zhao![](/image/person.jpg)
Adobe PDF(1739Kb)  |   收藏  |  浏览/下载:107/42  |  提交时间:2023/04/26 |
| A Multi-Task MRC Framework for Chinese Emotion Cause and Experiencer Extraction 会议论文 , Bratislava, Slovakia, 2021-09 作者: Haoda Qian ; Qiudan Li ; Zaichuan Tang
Adobe PDF(79001Kb)  |   收藏  |  浏览/下载:352/124  |  提交时间:2022/06/14 |
| Online Minimax Q Network Learning for Two-Player Zero-Sum Markov Games 期刊论文 IEEE TRANSACTIONS ON NEURAL NETWORKS AND LEARNING SYSTEMS, 2022, 卷号: 33, 期号: 3, 页码: 1228-1241 作者: Zhu, Yuanheng ; Zhao, Dongbin![](/image/person.jpg)
Adobe PDF(2838Kb)  |   收藏  |  浏览/下载:231/4  |  提交时间:2022/06/10 Games Nash equilibrium Mathematical model Markov processes Convergence Dynamic programming Training Deep reinforcement learning (DRL) generalized policy iteration (GPI) Markov game (MG) Nash equilibrium Q network zero sum |