CASIA OpenIR

浏览/检索结果: 共91条,第1-10条 帮助

限定条件                        
已选(0)清除 条数/页:   排序方式:
Token-level Direct Preference Optimization 会议论文
, Vienna, Austria, 2024/7/21-27
作者:  Zeng,Yongcheng;  Liu,Guoqing;  Ma,Weiyu;  Yang,Ning;  Zhang,Haifeng;  Wang,Jun
Adobe PDF(883Kb)  |  收藏  |  浏览/下载:17/4  |  提交时间:2024/06/05
Learning Superior Cooperative Policy in Competitive Multi-team Reinforcement Learning 会议论文
, Gold Coast, Australia, 2023-6
作者:  Qingxu Fu;  Tenghai Qiu;  Zhiqiang Pu;  Jianqiang Yi;  Xiaolin Ai;  Wanmai Yuan
Adobe PDF(25675Kb)  |  收藏  |  浏览/下载:10/0  |  提交时间:2024/06/05
Improve the efficiency of deep reinforcement learning through semantic exploration guided by natural language. 会议论文
, 北京华腾美居酒店, 2023-12-9
作者:  Zhourui Guo;  Meng Yao;  Yang Yu;  Qiyue Yin
Adobe PDF(2302Kb)  |  收藏  |  浏览/下载:3/1  |  提交时间:2024/06/03
Advancing Air Combat Tactics with Improved Neural Fictitious Self-Play Reinforcement Learning 会议论文
Advanced Intelligent Computing Technology and Applications, 中国郑州, 2023-8
作者:  He SQ(何少钦);  Gao Y(高阳);  Zhang BF(张保丰);  Chang H(常惠);  Zhang XC(张鑫辰)
Adobe PDF(1496Kb)  |  收藏  |  浏览/下载:13/6  |  提交时间:2024/05/31
Air Combat, Reinforcement Learning, Neural Fictitious Self-Play.  
Can We Really Trust Explanations? Evaluating the Stability of Feature Attribution Explanation Methods via Adversarial Attack 会议论文
, Nanchang, 2022-10
作者:  Zhao Yang;  Yuanzhe Zhang;  Zhongtao Jiang;  Yiming Ju
Adobe PDF(355Kb)  |  收藏  |  浏览/下载:7/3  |  提交时间:2024/05/30
Adaptive Multilingual Representations for Cross-Lingual Entity Linking with Attention on Entity Descriptions 会议论文
, Hangzhou, China, 2019-8
作者:  Wang, Chenhao;  Chen, Yubo;  Liu, Kang;  Zhao, Jun
Adobe PDF(1745Kb)  |  收藏  |  浏览/下载:8/1  |  提交时间:2024/05/30
Leveraging Explicit Lexico-logical Alignments in Text-to-SQL Parsing 会议论文
, Dublin, May 22–27, 2022
作者:  Sun, Runxin;  He, Shizhu;  Zhu, Chong;  He, Yaohan;  Li, Jinlong;  Zhao, Jun;  Liu, Kang
Adobe PDF(528Kb)  |  收藏  |  浏览/下载:17/3  |  提交时间:2024/05/28
Analysis of the Total Orientation Workspace of a Type of n-PPPS Parallel Manipulator 会议论文
, Chongqing,China, 2021-7-3至2021-7-5
作者:  Liu Zhaoyang;  Fan Junfeng;  Wang Zhe;  Jing Fengshui
Adobe PDF(4857Kb)  |  收藏  |  浏览/下载:11/2  |  提交时间:2024/05/28
Dual Self-Awareness Value Decomposition Framework without Individual Global Max for Cooperative MARL 会议论文
, New Orleans, LA, USA, December 10-16, 2023
作者:  Zhiwei Xu;  Bin Zhang;  Dapeng Li;  Guangchong Zhou;  Zeren Zhang;  Guoliang Fan
Adobe PDF(8700Kb)  |  收藏  |  浏览/下载:9/0  |  提交时间:2024/05/28
Learning to Coordinate via Multiple Graph Neural Networks 会议论文
, BALI, Indonesia, December 8-12, 2021
作者:  Zhiwei Xu;  Bin Zhang;  Yunpeng Bai;  Dapeng Li;  Guoliang Fan
Adobe PDF(2047Kb)  |  收藏  |  浏览/下载:12/5  |  提交时间:2024/05/28