CASIA OpenIR

Browse/Search Results:  1-10 of 2678 Help

Selected(0)Clear Items/Page:    Sort:
Learning to Play Football from Sports Perspective: A Knowledge-embedded Deep Reinforcement Learning Framework 期刊论文
IEEE Transactions on Games, 2022, 页码: 12
Authors:  Liu BY(刘博寅)
Adobe PDF(2957Kb)  |  Favorite  |  View/Download:14/4  |  Submit date:2024/07/12
基于深度强化学习的足球智能体球员策略方法研究 学位论文
, 2024
Authors:  刘博寅
Adobe PDF(11380Kb)  |  Favorite  |  View/Download:7/0  |  Submit date:2024/07/12
足球  多智能体系统  深度强化学习  互信息  内在激励  预训练  
Offline Hierarchical Reinforcement Learning: Enable Large-Scale Training in HRL 会议论文
, Nanjing, 2023-11-27
Authors:  Yuqiao Wu;  Haifeng Zhang;  Jun Wang
Adobe PDF(1339Kb)  |  Favorite  |  View/Download:9/1  |  Submit date:2024/07/12
NExT-OOD: Overcoming Dual Multiple-Choice VQA Biases 期刊论文
IEEE Transactions on Pattern Analysis and Machine Intelligence, 2023, 页码: 1913-1931
Authors:  Zhang Xi(张熙);  Feifei Zhang;  Changsheng Xu
Adobe PDF(4719Kb)  |  Favorite  |  View/Download:20/5  |  Submit date:2024/07/08
Exploring the Interactive Dynamic Influences Between Chinese and US's Future Markets 会议论文
Proceedings of 2022 10th China Conference on Command and Control, 北京市朝阳区国家会议中心, 2023-4-21
Authors:  Huang, Haitao;  Zheng, Xiaolong;  Zeng, Dajun
Adobe PDF(908Kb)  |  Favorite  |  View/Download:17/2  |  Submit date:2024/07/08
dynamic influences  influence strength  effective transfer entropy  
A Semantic and Structural Transformer for Code Summarization Generation 会议论文
, 澳大利亚, 2023.6.8
Authors:  Ruyi Ji;  Zhenyu Tong;  Tiejian Luo;  Jing Liu;  Libo Zhang
Adobe PDF(912Kb)  |  Favorite  |  View/Download:15/5  |  Submit date:2024/07/08
基于多模态协同的驾驶行为预测 学位论文
, 2024
Authors:  董清辉
Adobe PDF(5017Kb)  |  Favorite  |  View/Download:17/0  |  Submit date:2024/07/08
人车共驾,驾驶行为预测,多模态协同,轨迹预测,多任务学习  
Learning State-Specific Action Masks for Reinforcement Learning 期刊论文
Algorithms, 2024, 卷号: 17, 期号: 2, 页码: 60
Authors:  Wang ZY(王梓薏);  Li XR(李欣然);  Sun LY(孙罗洋);  Zhang HF(张海峰);  Liu HL(刘华林);  Jun Wang
Adobe PDF(2976Kb)  |  Favorite  |  View/Download:19/8  |  Submit date:2024/07/05
reinforcement learning  exploration efficiency  space reduction  
An Improved Minimax-Q Algorithm Based on Generalized Policy Iteration to Solve a Chaser-Invader Game 会议论文
, 线上, 2020-5
Authors:  Liu MS(刘民颂);  Zhu YH(朱圆恒);  Zhao DB(赵冬斌)
Adobe PDF(727Kb)  |  Favorite  |  View/Download:15/7  |  Submit date:2024/07/04
基于强化学习动作空间精简的时序决策任务算法研究 学位论文
, 2024
Authors:  王梓薏
Adobe PDF(7273Kb)  |  Favorite  |  View/Download:31/1  |  Submit date:2024/07/04
时序决策  强化学习  动作空间约简  分层强化学习  动作掩码