已选(0)清除
条数/页: 排序方式: |
| Offline Hierarchical Reinforcement Learning: Enable Large-Scale Training in HRL 会议论文 , Nanjing, 2023-11-27 作者: Yuqiao Wu; Haifeng Zhang; Jun Wang Adobe PDF(1339Kb)  |  收藏  |  浏览/下载:19/4  |  提交时间:2024/07/12 |
| Learning State-Specific Action Masks for Reinforcement Learning 期刊论文 Algorithms, 2024, 卷号: 17, 期号: 2, 页码: 60 作者: Wang ZY(王梓薏); Li XR(李欣然); Sun LY(孙罗洋); Zhang HF(张海峰); Liu HL(刘华林); Jun Wang Adobe PDF(2976Kb)  |  收藏  |  浏览/下载:30/12  |  提交时间:2024/07/05 reinforcement learning exploration efficiency space reduction |
| On the Effects of Structural Modeling for Neural Semantic Parsing 会议论文 Proceedings of the 27th Conference on Computational Natural Language Learning (CoNLL), Singapore, Singapore, 2023-12 作者: Zhang X(张翔); He SZ(何世柱); Liu K(刘康); Zhao J(赵军) Adobe PDF(730Kb)  |  收藏  |  浏览/下载:32/19  |  提交时间:2024/06/27 |
| Latent Landmark Graph for Efficient Exploration-Exploitation Balance in Hierarchical Reinforcement Learning 期刊论文 Machine Intelligence Research, 2023, 页码: 158 作者: Zhang Qingyang; Zhang Hongming; Xing Dengpeng; Bo Xu Adobe PDF(9639Kb)  |  收藏  |  浏览/下载:19/9  |  提交时间:2024/06/25 |
| A Double-Observation Policy Learning Framework for Multi-target Coverage with Connectivity Maintenance 会议论文 , online, 2022-2 作者: Xu YF(徐一凡); Pu ZQ(蒲志强); Wu SG(吴士广); Liu BY(刘博寅); Yi JQ(易建强); Geng HJ(耿虎军); Chai XH(柴兴华) Adobe PDF(9582Kb)  |  收藏  |  浏览/下载:20/5  |  提交时间:2024/06/21 |
| MoDE-CoTD: Chain-of-Thought Distillation for Complex Reasoning Tasks with Mixture of Decoupled LoRA-Experts 会议论文 , Torino (Italia), 2024.5.20 - 2024.5.25 作者: Xiang Li; Shizhu He; Jiayu Wu; Zhao Yang; Yao Xu; Yang Jun; Haifeng Liu; Kang Liu; Jun Zhao Adobe PDF(1062Kb)  |  收藏  |  浏览/下载:30/6  |  提交时间:2024/06/20 |
| Learning Robust Communication by Adversarial Training in Networked System Control 期刊论文 Lecture Notes in Electrical Engineering, 2024, 页码: Chapter 52 978-981-97-3335-4 作者: Runji, Lin; Haifeng, Zhang Adobe PDF(8334Kb)  |  收藏  |  浏览/下载:41/16  |  提交时间:2024/06/11 Networked System Control Robustness Communicative Multi-Agent Reinforcement Learning |
| Filtered Observations for Model-Based Multi-agent Reinforcement Learning 会议论文 , Turin, Italy, 2023.9.18-2023.9.22 作者: Meng Linghui; Xiong Xuantang; Zang Yifan; Zhang Xi; Li Guoqi; Xing Dengpeng; Xu Bo Adobe PDF(841Kb)  |  收藏  |  浏览/下载:42/17  |  提交时间:2024/06/11 |
| Learn to flap: foil non-parametric path planning via deep reinforcement learning 期刊论文 Journal of Fluid Mechanics, 2024, 卷号: 984, 页码: A9 作者: Wang, Zhipeng; Lin, Runji; Zhao, Zhiyu; Chen, Xu; Guo, Pengming; Yang, Ning; Wang,Zhicheng; Fan, Dixia Adobe PDF(1892Kb)  |  收藏  |  浏览/下载:47/11  |  提交时间:2024/06/07 |
| A Fish-like Binocular Vision System for Underwater Perception of Robotic Fish 期刊论文 Biomimetics, 2024, 页码: 171 作者: Tong Ru; Wu Zhengxing; Wang Jinge; Huang Yupei; Chen Di; Yu Junzhi Adobe PDF(4134Kb)  |  收藏  |  浏览/下载:39/15  |  提交时间:2024/06/06 |