CASIA OpenIR

浏览/检索结果: 共13条,第1-10条 帮助

限定条件                        
已选(0)清除 条数/页:   排序方式:
RTDOD: A large-scale RGB-thermal domain-incremental object detection dataset for UAVs 期刊论文
IMAGE AND VISION COMPUTING, 2023, 卷号: 140, 页码: 9
作者:  Feng, Hangtao;  Zhang, Lu;  Zhang, Siqi;  Wang, Dong;  Yang, Xu;  Liu, Zhiyong
Adobe PDF(3013Kb)  |  收藏  |  浏览/下载:104/5  |  提交时间:2024/02/22
Domain -incremental object detection  Dataset  RGB-T dataset  Object detection dataset  UAVs dataset  Object detection  
GCNet: Graph Completion Network for Incomplete Multimodal Learning in Conversation 期刊论文
IEEE TRANSACTIONS ON PATTERN ANALYSIS AND MACHINE INTELLIGENCE, 2023, 卷号: 45, 期号: 7, 页码: 8419-8432
作者:  Lian, Zheng;  Chen, Lan;  Sun, Licai;  Liu, Bin;  Tao, Jianhua
Adobe PDF(3959Kb)  |  收藏  |  浏览/下载:163/5  |  提交时间:2023/11/17
Oral communication  Correlation  Data models  Task analysis  Feature extraction  Tensors  Benchmark testing  Conversational data  graph complete network (GCNet)  incomplete multimodal learning  speaker-sensitive modeling  temporal-sensitive modeling  
Frequency-based pseudo-domain generation for domain generalizable object detection 期刊论文
NEUROCOMPUTING, 2023, 卷号: 542, 页码: 12
作者:  Zhang, Siqi;  Zhang, Lu;  Liu, Zhi-Yong
Adobe PDF(3838Kb)  |  收藏  |  浏览/下载:129/8  |  提交时间:2023/11/17
Domain generalization  Object detection  Transfer learning  Self-Supervised learning  
SMIN: Semi-Supervised Multi-Modal Interaction Network for Conversational Emotion Recognition 期刊论文
IEEE TRANSACTIONS ON AFFECTIVE COMPUTING, 2023, 卷号: 14, 期号: 3, 页码: 2415-2429
作者:  Lian, Zheng;  Liu, Bin;  Tao, Jianhua
Adobe PDF(2103Kb)  |  收藏  |  浏览/下载:136/5  |  提交时间:2023/11/15
Emotion recognition  Feature extraction  Training  Acoustics  Semisupervised learning  Benchmark testing  Hidden Markov models  Semi-supervised multi-modal interaction network (SMIN)  conversational emotion recognition  semi-supervised learning  intra-modal interaction  cross-modal interaction  
Cross stage partial connections based weighted Bi-directional feature pyramid and enhanced spatial transformation network for robust object detection 期刊论文
NEUROCOMPUTING, 2022, 卷号: 513, 页码: 70-82
作者:  Lu, Yan-Feng;  Yu, Qian;  Gao, Jing-Wen;  Li, Yi;  Zou, Jun-Cheng;  Qiao, Hong
Adobe PDF(3025Kb)  |  收藏  |  浏览/下载:246/7  |  提交时间:2022/11/14
Robust object detection  Structural deformation  Image detection  Spatial transformation  
Incremental few-shot object detection via knowledge transfer 期刊论文
PATTERN RECOGNITION LETTERS, 2022, 卷号: 156, 页码: 67-73
作者:  Feng, Hangtao;  Zhang, Lu;  Yang, Xu;  Liu, Zhiyong
Adobe PDF(1094Kb)  |  收藏  |  浏览/下载:206/10  |  提交时间:2022/07/25
Machine learning  Convolutional neural networks  Transfer learning  Incremental few-shot object detection  
Weakly Aligned Feature Fusion for Multimodal Object Detection 期刊论文
IEEE TRANSACTIONS ON NEURAL NETWORKS AND LEARNING SYSTEMS, 2021, 页码: 15
作者:  Zhang, Lu;  Liu, Zhiyong;  Zhu, Xiangyu;  Song, Zhan;  Yang, Xu;  Lei, Zhen;  Qiao, Hong
Adobe PDF(19222Kb)  |  收藏  |  浏览/下载:229/3  |  提交时间:2022/01/27
Object detection  Feature extraction  Detectors  Robustness  Cameras  Automation  Training  Deep learning  feature fusion  multimodal object detection  pedestrian detection  
CTNet: Conversational Transformer Network for Emotion Recognition 期刊论文
IEEE-ACM TRANSACTIONS ON AUDIO SPEECH AND LANGUAGE PROCESSING, 2021, 期号: 29, 页码: 985-1000
作者:  Lian, Zheng;  Liu, Bin;  Tao, Jianhua
Adobe PDF(2230Kb)  |  收藏  |  浏览/下载:381/61  |  提交时间:2021/05/06
Emotion recognition  Context modeling  Feature extraction  Fuses  Speech processing  Data models  Bidirectional control  Context-sensitive modeling  conversational transformer network (CTNet)  conversational emotion recognition  multimodal fusion  speaker-sensitive modeling  
Gated Recurrent Fusion With Joint Training Framework for Robust End-to-End Speech Recognition 期刊论文
IEEE-ACM TRANSACTIONS ON AUDIO SPEECH AND LANGUAGE PROCESSING, 2021, 期号: 29, 页码: 198-209
作者:  Fan, Cunhang;  Yi, Jiangyan;  Tao, Jianhua;  Tian, Zhengkun;  Liu, Bin;  Wen, Zhengqi
Adobe PDF(2534Kb)  |  收藏  |  浏览/下载:419/54  |  提交时间:2021/03/08
Speech enhancement  Speech recognition  Training  Noise measurement  Logic gates  Acoustic distortion  Task analysis  Gated recurrent fusion  robust end-to-end speech recognition  speech distortion  speech enhancement  speech transformer  
WAGNN: A Weighted Aggregation Graph Neural Network for robot skill learning 期刊论文
ROBOTICS AND AUTONOMOUS SYSTEMS, 2020, 卷号: 130, 页码: 9
作者:  Zhang, Fengyi;  Liu, Zhiyong;  Xiong, Fangzhou;  Su, Jianhua;  Qiao, Hong
Adobe PDF(1550Kb)  |  收藏  |  浏览/下载:351/52  |  提交时间:2020/07/20
Skill transfer learning  Serial structures  Robot skill learning  Graph Neural Network