CASIA OpenIR

浏览/检索结果: 共48条,第1-10条 帮助

限定条件                
已选(0)清除 条数/页:   排序方式:
Lazy Agents: A New Perspective on Solving Sparse Reward Problem in Multi-agent Reinforcement Learning 期刊
创刊日期: 2018,
主办者:  Liu BY(刘博寅)
Adobe PDF(5797Kb)  |  收藏  |  浏览/下载:22/5  |  提交时间:2024/07/12
MoDE-CoTD: Chain-of-Thought Distillation for Complex Reasoning Tasks with Mixture of Decoupled LoRA-Experts 会议论文
, Torino (Italia), 2024.5.20 - 2024.5.25
作者:  Xiang Li;  Shizhu He;  Jiayu Wu;  Zhao Yang;  Yao Xu;  Yang Jun;  Haifeng Liu;  Kang Liu;  Jun Zhao
Adobe PDF(1062Kb)  |  收藏  |  浏览/下载:30/6  |  提交时间:2024/06/20
M3: Modularization for Multi-task and Multi-agent Offline Pre-training 会议论文
, London, United Kingdom, 2023.5.29-2023.6.2
作者:  Meng Linghui;  Ruan Jingqing;  Xiong Xuantang;  Li Xiyun;  Zhang Xi;  Xing Dengpeng;  Xu Bo
Adobe PDF(1302Kb)  |  收藏  |  浏览/下载:30/8  |  提交时间:2024/06/11
A New Pre-Training Paradigm for Offline Multi-Agent Reinforcement Learning with Suboptimal Data 会议论文
, Seoul, Korea, 2024.4.14-2024.4.19
作者:  Meng Linghui;  Zhang Xi;  Xing Dengpeng;  Xu Bo
Adobe PDF(964Kb)  |  收藏  |  浏览/下载:44/17  |  提交时间:2024/06/11
Learn to flap: foil non-parametric path planning via deep reinforcement learning 期刊论文
Journal of Fluid Mechanics, 2024, 卷号: 984, 页码: A9
作者:  Wang, Zhipeng;  Lin, Runji;  Zhao, Zhiyu;  Chen, Xu;  Guo, Pengming;  Yang, Ning;  Wang,Zhicheng;  Fan, Dixia
Adobe PDF(1892Kb)  |  收藏  |  浏览/下载:47/11  |  提交时间:2024/06/07
Fault Diagnosis for Robotic Fish Sensors based on Spatial Domain Image Fusion and Convolution Neural Network 会议论文
, Tianjin, China, 2023-7
作者:  Xuqing Fan;  Sai Deng;  Junfeng Fan;  Chao Zhou;  Zhengxing Wu;  Yaming Ou;  Bin Zhang
Adobe PDF(1492Kb)  |  收藏  |  浏览/下载:34/9  |  提交时间:2024/06/05
Fault Diagnosis  GAF Fusion  CNN  Robotic Fish  
Enhancing efficiency and propulsion in bio-mimetic robotic fish through end-to-end deep reinforcement learning 期刊论文
Physics of Fluids, 2024, 卷号: 36, 期号: 3, 页码: 031910
作者:  Cui,Xinyu;  Sun,Boai;  Zhu,Yi;  Yang,Ning;  Zhang,Haifeng;  Cui,Weicheng;  Fan,Dixia;  Wang,Jun
Adobe PDF(4056Kb)  |  收藏  |  浏览/下载:65/26  |  提交时间:2024/06/02
bio-mimetic robotic fish  deep reinforcement learning  
Matching-based Term Semantics Pre-training for Spoken Patient Query Understanding 会议论文
, Rhodes, Greece, 2023-6-6 - 2023-6-10
作者:  Zefa Hu;  Xiuyi Chen;  Haoran Wu;  Minglun Han;  Ziyi Ni;  Jing Shi;  Shuang Xu;  Bo Xu
Adobe PDF(1049Kb)  |  收藏  |  浏览/下载:61/19  |  提交时间:2024/05/29
Spatial Domain Image Fusion with Particle Swarm Optimization and Lightweight AlexNet for Robotic Fish Sensor Fault Diagnosis 期刊论文
BIOMIMETICS, 2023, 卷号: 8, 期号: 6, 页码: 489
作者:  Fan, Xuqing;  Deng, Sai;  Wu, Zhengxing;  Fan, Junfeng;  Zhou, Chao
Adobe PDF(5062Kb)  |  收藏  |  浏览/下载:132/12  |  提交时间:2023/12/21
image fusion  lightweight AlexNet  particle swarm optimization  fault diagnosis  robotic fish  
AlphaHoldem: High-Performance Artificial Intelligence for Heads-Up No-Limit Poker via End-to-End Reinforcement Learning 会议论文
, 线上, 2022-02-22
作者:  Zhao EM(赵恩民);  Yan RY(闫仁业);  Li JQ(李金秋);  Li K(李凯);  Xing JL(兴军亮)
Adobe PDF(2593Kb)  |  收藏  |  浏览/下载:209/76  |  提交时间:2023/06/29