CASIA OpenIR

浏览/检索结果: 共912条,第1-10条 帮助

已选(0)清除 条数/页:   排序方式:
Image captioning: Semantic selection unit with stacked residual attention 期刊论文
IMAGE AND VISION COMPUTING, 2024, 卷号: 144, 页码: 12
作者:  Song, Lifei;  Li, Fei;  Wang, Ying;  Liu, Yu;  Wang, Yuanhua;  Xiang, Shiming
收藏  |  浏览/下载:0/0  |  提交时间:2024/07/03
Image captioning  Semantic attributes  Semantic selection unit  Transformer  Stacked residual attention  
Comprehensive Attribute Prediction Learning for Person Search by Language 期刊论文
IEEE TRANSACTIONS ON IMAGE PROCESSING, 2024, 卷号: 33, 页码: 1990-2003
作者:  Niu, Kai;  Huang, Linjiang;  Long, Yuzhou;  Huang, Yan;  Wang, Liang;  Zhang, Yanning
收藏  |  浏览/下载:0/0  |  提交时间:2024/07/03
Person search by language  cross-modal retrieval  smart video surveillance  attribute prediction  
几何驱动的三维场景检测与分割 学位论文
, 2024
作者:  关赫
Adobe PDF(31711Kb)  |  收藏  |  浏览/下载:22/0  |  提交时间:2024/06/27
几何驱动  单目三维检测  多维场景分割  数据增强  实用性  特征交互  
基于多尺度特征融合的图像语义分割方法研究 学位论文
, 2024
作者:  朱袁兵
Adobe PDF(29615Kb)  |  收藏  |  浏览/下载:22/1  |  提交时间:2024/06/27
图像语义分割  实时语义分割  开放词汇语义分割  视觉语言模型  
MapGuide: A Simple yet Effective Method to Reconstruct Continuous Language from Brain Activities 会议论文
Proceedings of the 2024 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies (Volume 1: Long Papers), Mexico City, Mexico, 2024-6
作者:  Xinpei, Zhao;  Jingyuan, Sun;  Shaonan, Wang;  Jing, Ye;  Xiaohan, Zhang;  Chengqing, Zong
Adobe PDF(843Kb)  |  收藏  |  浏览/下载:11/4  |  提交时间:2024/06/27
neural decoding  
Born a BabyNet with Hierarchical Parental Supervision for End-to-End Text Image Machine Translation 会议论文
Proceedings of the 2024 Joint International Conference on Computational Linguistics, Language Resources and Evaluation (LREC-COLING 2024), Torino, Italia, 20-25 May, 2024
作者:  Ma, Cong;  Zhang, Yaping;  Zhang, Zhiyang;  Liang, Yupu;  Zhao, Yang;  Zhou, Yu;  Zong, Chengqing
Adobe PDF(891Kb)  |  收藏  |  浏览/下载:10/6  |  提交时间:2024/06/27
VECTOR QUANTIZATION KNOWLEDGE TRANSFER FOR END-TO-END TEXT IMAGE MACHINE TRANSLATION 会议论文
Proceedings of 2024 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP), Seoul, Korea, 14-19 April 2024
作者:  Ma, Cong;  Zhang, Yaping;  Zhao, Yang;  Zhou, Yu;  Zong, Chengqing
Adobe PDF(1090Kb)  |  收藏  |  浏览/下载:17/8  |  提交时间:2024/06/26
E2TIMT: Efficient and Effective Modal Adapter for Text Image Machine Translation 会议论文
Proceedings of the 17th Document Analysis and Recognition (ICDAR 2023), San José, California, USA, August 21-26, 2023
作者:  Ma, Cong;  Zhang, Yaping;  Tu, Mei;  Zhao, Yang;  Zhou, Yu;  Zong, Chengqing
Adobe PDF(1430Kb)  |  收藏  |  浏览/下载:11/3  |  提交时间:2024/06/26
Multi-teacher Knowledge Distillation for End-to-End Text Image Machine Translation 会议论文
Proceedings of the 17th Document Analysis and Recognition (ICDAR 2023), San José, California, USA, August 21-26, 2023
作者:  Ma, Cong;  Zhang, Yaping;  Tu, Mei;  Zhao, Yang;  Zhou, Yu;  Zong, Chengqing
Adobe PDF(1478Kb)  |  收藏  |  浏览/下载:16/9  |  提交时间:2024/06/26
跨模态信息融合的文本图像翻译方法研究 学位论文
, 2024
作者:  马聪
Adobe PDF(11285Kb)  |  收藏  |  浏览/下载:26/4  |  提交时间:2024/06/26
文本图像翻译  跨模态信息融合  多任务学习  跨模态对比学习  参数高效微调