CASIA OpenIR

浏览/检索结果: 共413条,第1-10条 帮助

限定条件    
已选(0)清除 条数/页:   排序方式:
AnyFace++: A Unified Framework for Free-style Text-to-Face Synthesis and Manipulation 期刊论文
IEEE Transactions on Pattern Analysis and Machine Intelligence, 2024, 页码: 1-15
作者:  Sun, Jianxin;  Deng, Qiyao;  Li, Qi;  Sun, Muyi;  Liu, Yunfan;  Sun, Zhenan
Adobe PDF(16839Kb)  |  收藏  |  浏览/下载:39/7  |  提交时间:2024/02/23
Understanding and Mitigating Overfitting in Prompt Tuning for Vision-Language Models 期刊论文
IEEE TRANSACTIONS ON CIRCUITS AND SYSTEMS FOR VIDEO TECHNOLOGY, 2023, 卷号: 33, 期号: 9, 页码: 4616-4629
作者:  Ma, Chengcheng;  Liu, Yang;  Deng, Jiankang;  Xie, Lingxi;  Dong, Weiming;  Xu, Changsheng
Adobe PDF(1644Kb)  |  收藏  |  浏览/下载:81/11  |  提交时间:2023/11/16
Vision-language model  prompt tuning  over-fitting  subspace learning  gradient projection  
Latent Structure Mining With Contrastive Modality Fusion for Multimedia Recommendation 期刊论文
IEEE TRANSACTIONS ON KNOWLEDGE AND DATA ENGINEERING, 2023, 卷号: 35, 期号: 9, 页码: 9154-9167
作者:  Zhang, Jinghao;  Zhu, Yanqiao;  Liu, Qiang;  Zhang, Mengqi;  Wu, Shu;  Wang, Liang
收藏  |  浏览/下载:81/0  |  提交时间:2023/11/17
Multimedia recommendation  graph structure learning  contrastive learning  
Medical visual question answering with symmetric interaction attention and cross-modal gating 期刊论文
BIOMEDICAL SIGNAL PROCESSING AND CONTROL, 2023, 卷号: 85, 页码: 10
作者:  Chen, Zhi;  Zou, Beiji;  Dai, Yulan;  Zhu, Chengzhang;  Kong, Guilan;  Zhang, Wensheng
收藏  |  浏览/下载:63/0  |  提交时间:2023/11/17
Medical visual question answering  Self-attention  Information interaction  Cross-modal gating  
Multi-scale Dual Domain Network for Nonlinear Magnetization Signal Filtering in Magnetic Particle Imaging 期刊
创刊日期: 2023, 收录类别: SCI,
主办者:  IEEE Engineering in Medicine and Biology Society
Adobe PDF(1549Kb)  |  收藏  |  浏览/下载:208/88  |  提交时间:2023/06/15
A Multi-Modal Neural Geometric Solver with Textual Clauses Parsed from Diagram 会议论文
, 中国 澳门, 2023-7-19
作者:  Zhang Ming-Liang;  Yin Fei;  Liu Cheng-Lin
Adobe PDF(1110Kb)  |  收藏  |  浏览/下载:31/10  |  提交时间:2024/04/03
Recovering Generalization via Pre-training-like Knowledge Distillation for Out-of-Distribution Visual Question Answering 期刊论文
IEEE Transactions on Multimedia, 2023, 页码: 1-15
作者:  Song, Yaguang;  Yang, Xiaoshan;  Wang, Yaowei;  Xu, Changsheng
Adobe PDF(2397Kb)  |  收藏  |  浏览/下载:152/41  |  提交时间:2023/06/12
Multi-modal Foundation Model  Out-of-Distribution Generalization  Visual Question Answering  Knowledge Distillation  
3D Semantic Segmentation of Aerial Photogrammetry Models Based on Orthographic Projection 期刊论文
IEEE Transactions on Circuits and Systems for Video Technology, 2023, 页码: early-access
作者:  Mengqi Rong;  Shuhan Shen
Adobe PDF(5811Kb)  |  收藏  |  浏览/下载:102/35  |  提交时间:2023/09/25
Generating Emotion Descriptions for Fine Art Paintings via Multiple Painting Representations 期刊论文
IEEE Intelligent Systems, 2023, 卷号: 38, 期号: 3, 页码: 31-40
作者:  Lu, Yue;  Guo, Chao;  Dai, Xingyuan;  Wang, Fei-Yue
Adobe PDF(1045Kb)  |  收藏  |  浏览/下载:87/13  |  提交时间:2023/06/25
painting captioning  
Multimodal motor imagery decoding method based on temporal spatial feature alignment and fusion 期刊论文
Journal of Neural Engineering, 2023, 卷号: 20, 期号: 2, 页码: 026009
作者:  Zhang,Yukun;  Qiu,Shuang;  He,Huiguang
Adobe PDF(16882Kb)  |  收藏  |  浏览/下载:114/20  |  提交时间:2023/05/31
brain-computer interface  motor imagery  multimodal  EEG-fNIRS  center loss