CASIA OpenIR

浏览/检索结果: 共21条,第1-10条 帮助

限定条件    
已选(0)清除 条数/页:   排序方式:
Omnidirectional Drift Control of an Underwater Biomimetic Vehicle-Manipulator System via Reinforcement Learning 会议论文
, Suzhou, China, May 14-16, 2021
作者:  Ma, Ruichen;  Wang, Yu;  Wang, Rui;  Wang, Shuo
Adobe PDF(855Kb)  |  收藏  |  浏览/下载:82/30  |  提交时间:2023/08/02
Omnidirectional Drift Control  Undulating Fin  Underwater Biomimetic Vehicle-manipulator System (UBVMS)  Reinforcement Learning  Twin Delayed Deep Deterministic policy gradient (TD3)  
Second-Order Global Attention Networks for Graph Classification and Regression 会议论文
, Beijing, China, August 27-28, 2022
作者:  Hu Fenyu;  Cui Zeyu;  Wu Shu;  Liu Qiang;  Wu Jinlin;  Wang Liang;  Tan Tieniu
Adobe PDF(69424Kb)  |  收藏  |  浏览/下载:187/69  |  提交时间:2023/07/06
Calibration of Agent-Based Model Using Reinforcement Learning 会议论文
, Beijing, 2021
作者:  Song B(宋冰);  Xiong G(熊刚);  Yu S(于松民);  Ye P(叶佩军);  Dong X(董西松);  Lv Y(吕宜生)
Adobe PDF(437Kb)  |  收藏  |  浏览/下载:125/47  |  提交时间:2023/06/26
Learning Cooperative Policies with Graph Networks in Distributed Swarm Systems 会议论文
, Queensland, Australia, June 18-23, 2023
作者:  Zhang TL(张天乐);  Liu Z(刘振);  Pu ZQ(蒲志强);  Yi JQ(易建强);  Ai XL(艾晓琳);  Yuan GM(袁莞迈)
Adobe PDF(612Kb)  |  收藏  |  浏览/下载:153/47  |  提交时间:2023/06/12
Multi-Target Encirclement with Collision Avoidance via Deep Reinforcement Learning using Relational Graphs 会议论文
, Philadelphia, PA, USA, May 23-27, 2022
作者:  Zhang TL(张天乐);  Liu Z(刘振);  Pu ZQ(蒲志强);  Yi JQ(易建强)
Adobe PDF(4277Kb)  |  收藏  |  浏览/下载:132/33  |  提交时间:2023/06/12
Wd3: Taming the estimation bias in deep reinforcement learning 会议论文
, Baltimore, MD, USA, 2020-12
作者:  He Q(何强);  Hou XW(侯新文)
Adobe PDF(2006Kb)  |  收藏  |  浏览/下载:202/39  |  提交时间:2022/06/27
deep reinforcement learning  estimation bias  neural networks  
Trajectory-based Split Hindsight Reverse Curriculum Learning 会议论文
, Prague, Czech Republic, 2021-9
作者:  Wu, Jiaxi;  Zhang, Dianmin;  Zhong, Shanlin;  Qiao, Hong
Adobe PDF(5094Kb)  |  收藏  |  浏览/下载:218/50  |  提交时间:2022/06/14
Reinforcement Learning  Curriculum Learning  
A Policy-Based Reinforcement Learning Approach for High-Speed Railway Timetable Rescheduling 会议论文
, Indianapolis, IN, USA, 19-22 Sept. 2021
作者:  Yin Wang;  Yisheng Lv;  Jianying Zhou;  Zhiming Yuan;  Qi Zhang;  Min Zhou
Adobe PDF(1210Kb)  |  收藏  |  浏览/下载:171/48  |  提交时间:2022/04/08
知识和数据协同驱动的群体智能决策方法研究综述 期刊论文
自动化学报, 2022, 卷号: 48, 期号: 3, 页码: 1-17
作者:  蒲志强;  易建强;  刘振;  丘腾海;  孙金林;  李非墨
Adobe PDF(1352Kb)  |  收藏  |  浏览/下载:280/67  |  提交时间:2022/04/02
群体智能  知识与数据协同  多智能体  决策智能  
STGA-LSTM: A Spatial-Temporal Graph Attentional LSTM Scheme for Multi-Agent Cooperation 会议论文
, 线上, 2020-11
作者:  Huimu Wang;  Zhen Liu;  Zhiqiang Pu;  Jianqiang Yi
Adobe PDF(916Kb)  |  收藏  |  浏览/下载:94/0  |  提交时间:2021/06/24