CASIA OpenIR

浏览/检索结果: 共118条,第1-10条 帮助

限定条件                
已选(0)清除 条数/页:   排序方式:
Learning to Play Football from Sports Perspective: A Knowledge-embedded Deep Reinforcement Learning Framework 期刊论文
IEEE Transactions on Games, 2022, 页码: 12
作者:  Liu BY(刘博寅)
Adobe PDF(2957Kb)  |  收藏  |  浏览/下载:17/5  |  提交时间:2024/07/12
AdaNSP: Uncertainty-driven Adaptive Decoding in Neural Semantic Parsing 会议论文
Proceedings of the 57th Annual Meeting of the Association for Computational Linguistics, Florence, Italy, 2019-07
作者:  Zhang X(张翔);  He SZ(何世柱);  Liu K(刘康);  Zhao J(赵军)
Adobe PDF(400Kb)  |  收藏  |  浏览/下载:26/9  |  提交时间:2024/06/26
Balancing Exploration and Exploitation in Hierarchical Reinforcement Learning via Latent Landmark Graphs 会议论文
, 澳大利亚, 2023-6
作者:  Zhang Qingyang;  Yang Yiming;  Ruan Jingqing;  Xiong Xuantang;  Xing Dengpeng;  Xu Bo
Adobe PDF(7948Kb)  |  收藏  |  浏览/下载:30/12  |  提交时间:2024/06/25
强化学习,分层强化学习  
Latent Landmark Graph for Efficient Exploration-Exploitation Balance in Hierarchical Reinforcement Learning 期刊论文
Machine Intelligence Research, 2023, 页码: 158
作者:  Zhang Qingyang;  Zhang Hongming;  Xing Dengpeng;  Bo Xu
Adobe PDF(9639Kb)  |  收藏  |  浏览/下载:16/9  |  提交时间:2024/06/25
Global and local multi-modal feature mutual learning for retinal vessel segmentation 期刊论文
Pattern Recognition, 2024, 卷号: 151, 页码: 110376
作者:  Xin Zhao;  Zhang Jing;  Qiaozhe Li;  Tengfei Zhao;  Yi Li;  Zifeng Wu
Adobe PDF(4182Kb)  |  收藏  |  浏览/下载:33/13  |  提交时间:2024/06/21
Mutual learning  Multi-modal learning  OCTA images  Retinal vessel segmentation  
Learning to Align Question and Answer Utterances in Customer Service Conversation with Recurrent Pointer Networks 会议论文
, Honolulu, Hawaii, USA, 2019.01.27 - 2019.02.01
作者:  Shizhu HE;  Kang Liu;  Weiting An
Adobe PDF(1562Kb)  |  收藏  |  浏览/下载:42/16  |  提交时间:2024/06/20
MoDE-CoTD: Chain-of-Thought Distillation for Complex Reasoning Tasks with Mixture of Decoupled LoRA-Experts 会议论文
, Torino (Italia), 2024.5.20 - 2024.5.25
作者:  Xiang Li;  Shizhu He;  Jiayu Wu;  Zhao Yang;  Yao Xu;  Yang Jun;  Haifeng Liu;  Kang Liu;  Jun Zhao
Adobe PDF(1062Kb)  |  收藏  |  浏览/下载:25/6  |  提交时间:2024/06/20
Lead ASR Models to Generalize Better Using Approximated Bias-Variance Tradeof 会议论文
, changsha,China, 2023.11.13
作者:  Wang FY(王方圆);  Ming Hao;  Yuhai Shi;  Bo Xu
Adobe PDF(1933Kb)  |  收藏  |  浏览/下载:44/17  |  提交时间:2024/06/12
A New Pre-Training Paradigm for Offline Multi-Agent Reinforcement Learning with Suboptimal Data 会议论文
, Seoul, Korea, 2024.4.14-2024.4.19
作者:  Meng Linghui;  Zhang Xi;  Xing Dengpeng;  Xu Bo
Adobe PDF(964Kb)  |  收藏  |  浏览/下载:38/14  |  提交时间:2024/06/11
Learn to flap: foil non-parametric path planning via deep reinforcement learning 期刊论文
Journal of Fluid Mechanics, 2024, 卷号: 984, 页码: A9
作者:  Wang, Zhipeng;  Lin, Runji;  Zhao, Zhiyu;  Chen, Xu;  Guo, Pengming;  Yang, Ning;  Wang,Zhicheng;  Fan, Dixia
Adobe PDF(1892Kb)  |  收藏  |  浏览/下载:46/11  |  提交时间:2024/06/07