已选(0)清除
条数/页: 排序方式: |
| Soft Contrastive Learning with Q-irrelevance Abstraction for Reinforcement Learning 期刊论文 IEEE Transactions on Cognitive and Developmental Systems, 2023, 卷号: 15, 期号: 3, 页码: 1463 - 1473 作者: Liu MS(刘民颂) ; Li LT(李伦通); Hao S(郝帅); Zhu YH(朱圆恒) ; Zhao DB(赵冬斌)![](/image/person.jpg)
Adobe PDF(4197Kb)  |   收藏  |  浏览/下载:34/11  |  提交时间:2024/06/24 |
| MAT: Morphological Adaptive Transformer for Universal Morphology Policy Learning 期刊论文 IEEE Transactions on Cognitive and Developmental Systems, 2024, 页码: 1-12 作者: Boyu Li; Haran Li; Yuanheng Zhu ; Dongbin Zhao![](/image/person.jpg)
Adobe PDF(9953Kb)  |   收藏  |  浏览/下载:28/9  |  提交时间:2024/06/05 |
| FM3Q: Factorized Multi-Agent MiniMax Q-Learning for Two-Team Zero-Sum Markov Game 期刊论文 IEEE Transactions on Emerging Topics in Computational Intelligence, 2024, 页码: 1-13 作者: Guangzheng Hu ; Yuanheng Zhu ; Haoran Li ; Dongbin Zhao![](/image/person.jpg)
Adobe PDF(2144Kb)  |   收藏  |  浏览/下载:37/7  |  提交时间:2024/06/05 Games Q-learning Task analysis Optimization Convergence Training Nash equilibrium Multi-agent reinforcement learning minimax-Q learning two-team zero-sum Markov games |
| Advantage Constrained Proximal Policy Optimization in Multi-Agent Reinforcement Learning 会议论文 , 昆士兰, 2023-6 作者: Li WF(李伟凡) ; Zhu YH(朱圆恒) ; Zhao DB(赵冬斌)![](/image/person.jpg)
Adobe PDF(4104Kb)  |   收藏  |  浏览/下载:254/81  |  提交时间:2023/06/29 multi-agent reinforcement learning policy gradient |
| Vision-based control in the open racing car simulator with deep and reinforcement learning 期刊论文 Journal of Ambient Intelligence and Humanized Computing, 2019, 页码: doi={10.1007/s12652-019-01503-y} 作者: Yuanheng Zhu ; Dongbin Zhao![](/image/person.jpg)
Adobe PDF(2210Kb)  |   收藏  |  浏览/下载:63/17  |  提交时间:2023/04/26 |
| Empirical Policy Optimization for n-Player Markov Games 期刊论文 IEEE Transactions on Cybernetics, 2022, 页码: doi={10.1109/TCYB.2022.3179775} 作者: Yuanheng Zhu ; Weifan Li ; Mengchen Zhao; Jianye Hao; Dongbin Zhao![](/image/person.jpg)
Adobe PDF(1739Kb)  |   收藏  |  浏览/下载:110/44  |  提交时间:2023/04/26 |
| Online Minimax Q Network Learning for Two-Player Zero-Sum Markov Games 期刊论文 IEEE TRANSACTIONS ON NEURAL NETWORKS AND LEARNING SYSTEMS, 2022, 卷号: 33, 期号: 3, 页码: 1228-1241 作者: Zhu, Yuanheng ; Zhao, Dongbin![](/image/person.jpg)
Adobe PDF(2838Kb)  |   收藏  |  浏览/下载:248/12  |  提交时间:2022/06/10 Games Nash equilibrium Mathematical model Markov processes Convergence Dynamic programming Training Deep reinforcement learning (DRL) generalized policy iteration (GPI) Markov game (MG) Nash equilibrium Q network zero sum |
| Highway Lane Change Decision-Making via Attention-Based Deep Reinforcement Learning 期刊论文 IEEE-CAA JOURNAL OF AUTOMATICA SINICA, 2022, 卷号: 9, 期号: 3, 页码: 567-569 作者: Wang, Junjie ; Zhang, Qichao ; Zhao, Dongbin![](/image/person.jpg)
Adobe PDF(803Kb)  |   收藏  |  浏览/下载:293/68  |  提交时间:2022/02/16 |
| UNMAS: Multiagent Reinforcement Learning for Unshaped Cooperative Scenarios 期刊论文 IEEE TRANSACTIONS ON NEURAL NETWORKS AND LEARNING SYSTEMS, 2021, 页码: 12 作者: Chai, Jiajun ; Li, Weifan ; Zhu, Yuanheng ; Zhao, Dongbin ; Ma, Zhe; Sun, Kewu; Ding, Jishiyu
Adobe PDF(3402Kb)  |   收藏  |  浏览/下载:287/37  |  提交时间:2022/01/27 Multi-agent systems Training Task analysis Reinforcement learning Sun Learning systems Semantics Centralized training with decentralized execution (CTDE) multiagent reinforcement learning StarCraft II |
| Enhanced Rolling Horizon Evolution Algorithm With Opponent Model Learning: Results for the Fighting Game AI Competition 期刊论文 IEEE TRANSACTIONS ON GAMES, 2023, 卷号: 5, 期号: 1, 页码: 5 - 15 作者: Zhentao Tang ; Yuanheng Zhu ; Dongbin Zhao ; Simon M. Lucas
Adobe PDF(7686Kb)  |   收藏  |  浏览/下载:344/71  |  提交时间:2021/07/05 Rolling horizon evolution opponent model reinforcement learning supervised learning fighting game |