CASIA OpenIR

浏览/检索结果: 共7条,第1-7条 帮助

限定条件    
已选(0)清除 条数/页:   排序方式:
Reducing Vision-Answer Biases for Multiple-Choice VQA 期刊论文
IEEE TRANSACTIONS ON IMAGE PROCESSING, 2023, 卷号: 32, 页码: 4621-4634
作者:  Zhang, Xi;  Zhang, Feifei;  Xu, Changsheng
Adobe PDF(2684Kb)  |  收藏  |  浏览/下载:91/6  |  提交时间:2023/11/17
Multiple-choice VQA  vision-answer bias  causal intervention  counterfactual interaction learning  
A Novel Biologically Inspired Structural Model for Feature Correspondence 期刊论文
IEEE TRANSACTIONS ON COGNITIVE AND DEVELOPMENTAL SYSTEMS, 2023, 卷号: 15, 期号: 2, 页码: 844-854
作者:  Lu, Yan-Feng;  Yang, Xu;  Li, Yi;  Yu, Qian;  Liu, Zhi-Yong;  Qiao, Hong
Adobe PDF(4447Kb)  |  收藏  |  浏览/下载:184/14  |  提交时间:2023/11/17
Visualization  Biological system modeling  Biology  Brain modeling  Biological information theory  Task analysis  Strain  Appearance feature descriptor  biologically inspired model  feature correspondence  feature representation  graph matching (GM)  graph structure  
GCNet: Graph Completion Network for Incomplete Multimodal Learning in Conversation 期刊论文
IEEE TRANSACTIONS ON PATTERN ANALYSIS AND MACHINE INTELLIGENCE, 2023, 卷号: 45, 期号: 7, 页码: 8419-8432
作者:  Lian, Zheng;  Chen, Lan;  Sun, Licai;  Liu, Bin;  Tao, Jianhua
Adobe PDF(3959Kb)  |  收藏  |  浏览/下载:179/7  |  提交时间:2023/11/17
Oral communication  Correlation  Data models  Task analysis  Feature extraction  Tensors  Benchmark testing  Conversational data  graph complete network (GCNet)  incomplete multimodal learning  speaker-sensitive modeling  temporal-sensitive modeling  
Frequency-based pseudo-domain generation for domain generalizable object detection 期刊论文
NEUROCOMPUTING, 2023, 卷号: 542, 页码: 12
作者:  Zhang, Siqi;  Zhang, Lu;  Liu, Zhi-Yong
Adobe PDF(3838Kb)  |  收藏  |  浏览/下载:148/14  |  提交时间:2023/11/17
Domain generalization  Object detection  Transfer learning  Self-Supervised learning  
Gated Recurrent Fusion With Joint Training Framework for Robust End-to-End Speech Recognition 期刊论文
IEEE-ACM TRANSACTIONS ON AUDIO SPEECH AND LANGUAGE PROCESSING, 2021, 期号: 29, 页码: 198-209
作者:  Fan, Cunhang;  Yi, Jiangyan;  Tao, Jianhua;  Tian, Zhengkun;  Liu, Bin;  Wen, Zhengqi
Adobe PDF(2534Kb)  |  收藏  |  浏览/下载:435/57  |  提交时间:2021/03/08
Speech enhancement  Speech recognition  Training  Noise measurement  Logic gates  Acoustic distortion  Task analysis  Gated recurrent fusion  robust end-to-end speech recognition  speech distortion  speech enhancement  speech transformer  
WAGNN: A Weighted Aggregation Graph Neural Network for robot skill learning 期刊论文
ROBOTICS AND AUTONOMOUS SYSTEMS, 2020, 卷号: 130, 页码: 9
作者:  Zhang, Fengyi;  Liu, Zhiyong;  Xiong, Fangzhou;  Su, Jianhua;  Qiao, Hong
Adobe PDF(1550Kb)  |  收藏  |  浏览/下载:372/58  |  提交时间:2020/07/20
Skill transfer learning  Serial structures  Robot skill learning  Graph Neural Network  
End-to-End Post-Filter for Speech Separation With Deep Attention Fusion Features 期刊论文
IEEE-ACM TRANSACTIONS ON AUDIO SPEECH AND LANGUAGE PROCESSING, 2020, 卷号: 28, 期号: 28, 页码: 1303-1314
作者:  Fan, Cunhang;  Tao, Jianhua;  Liu, Bin;  Yi, Jiangyan;  Wen, Zhengqi;  Liu, Xuefei
Adobe PDF(1344Kb)  |  收藏  |  浏览/下载:341/73  |  提交时间:2020/06/22
Feature extraction  Training  Interference  Speech enhancement  Clustering algorithms  Spectrogram  Speech separation  end-to-end post-filter  deep attention fusion features  deep clustering  permutation invariant training