CASIA OpenIR
(本次检索基于用户作品认领结果)

浏览/检索结果: 共21条,第1-10条 帮助

限定条件        
已选(0)清除 条数/页:   排序方式:
CrossRectify: Leveraging disagreement for semi-supervised object detection 期刊论文
Pattern recognition, 2022, 卷号: 137, 页码: 109280
作者:  Ma CC(马成丞);  Pan XJ(潘兴甲);  Ye QX(叶齐祥);  Tang F(唐帆);  Dong WM(董未名);  Xu CS(徐常胜)
Adobe PDF(1969Kb)  |  收藏  |  浏览/下载:103/10  |  提交时间:2024/01/29
Reducing Vision-Answer Biases for Multiple-Choice VQA 期刊论文
IEEE TRANSACTIONS ON IMAGE PROCESSING, 2023, 卷号: 32, 页码: 4621-4634
作者:  Zhang, Xi;  Zhang, Feifei;  Xu, Changsheng
Adobe PDF(2684Kb)  |  收藏  |  浏览/下载:84/2  |  提交时间:2023/11/17
Multiple-choice VQA  vision-answer bias  causal intervention  counterfactual interaction learning  
Understanding and Mitigating Overfitting in Prompt Tuning for Vision-Language Models 期刊论文
IEEE TRANSACTIONS ON CIRCUITS AND SYSTEMS FOR VIDEO TECHNOLOGY, 2023, 卷号: 33, 期号: 9, 页码: 4616-4629
作者:  Ma, Chengcheng;  Liu, Yang;  Deng, Jiankang;  Xie, Lingxi;  Dong, Weiming;  Xu, Changsheng
Adobe PDF(1644Kb)  |  收藏  |  浏览/下载:152/21  |  提交时间:2023/11/16
Vision-language model  prompt tuning  over-fitting  subspace learning  gradient projection  
The Model May Fit You: User-Generalized Cross-Modal Retrieval 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2021, 卷号: 24, 页码: 2998-3012
作者:  Ma, Xinhong;  Yang, Xiaoshan;  Gao, Junyu;  Xu, Changsheng
Adobe PDF(6549Kb)  |  收藏  |  浏览/下载:281/54  |  提交时间:2022/06/17
cross-modal retrieval  domain generalization  meta-learning  
Evo-ViT: Slow-Fast Token Evolution for Dynamic Vision Transformer 会议论文
, Online, 2022-3-1
作者:  Xu, Yifan;  Zhang, Zhijie;  Zhang, Mengdan;  Sheng, Kekai;  Li, Ke;  Dong, Weiming;  Zhang, Liqing;  Xu, Changsheng;  Sun, Xing
Adobe PDF(636Kb)  |  收藏  |  浏览/下载:285/69  |  提交时间:2022/06/14
Transformers in computational visual media: A survey 期刊论文
Computational Visual Media, 2021, 卷号: 8, 期号: 1, 页码: 33-62
作者:  Xu,Yifan;  Wei,Huapeng;  Lin,Minxuan;  Deng,Yingying;  Sheng,Kekai;  Zhang,Mengdan;  Tang,Fan;  Dong,Weiming;  Huang,Feiyue;  Xu,Changsheng
Adobe PDF(5366Kb)  |  收藏  |  浏览/下载:320/44  |  提交时间:2021/12/28
visual transformer  computational visual media (CVM)  high-level vision  low-level vision  image generation  multi-modal learning  
Multi-modal Knowledge-aware Event Memory Network for Social Media Rumor Detection 会议论文
MM, Nice, France, October 21 - 25, 2019
作者:  Huaiwen Zhang;  Quan Fang;  Shengsheng Qian;  Changsheng Xu
Adobe PDF(2626Kb)  |  收藏  |  浏览/下载:204/67  |  提交时间:2021/06/30
Social Media  Rumor Detection  Multi-Modal  Knowledge Graph  Memory Network  
Intra-domain Consistency Enhancement for Unsupervised Person Re-identification 期刊论文
IEEE Transactions on Multimedia, 2021, 卷号: 0, 期号: 0, 页码: 0-0
作者:  Li, Yaoyu;  Yao, Hantao;  Xu, Changsheng
Adobe PDF(2167Kb)  |  收藏  |  浏览/下载:217/61  |  提交时间:2021/06/22
Person Re-identification  unsupervised domain adaptation  representation learning  
A Unified Generative Adversarial Framework for Image Generation and Person Re-identification 会议论文
ACM Multimedia Conference (MM 18), Seoul, Republic of Korea, October 22–26, 2018
作者:  Li, Yaoyu;  Zhang, Tianzhu;  Duan, Lingyu;  Xu, Changsheng
Adobe PDF(3314Kb)  |  收藏  |  浏览/下载:202/59  |  提交时间:2021/06/22
Person Re-identification  Multimedia System  GAN  
Multi-Level Correlation Adversarial Hashing for Cross-Modal Retrieval 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2020, 卷号: 22, 期号: 12, 页码: 3101-3114
作者:  Ma, Xinhong;  Zhang, Tianzhu;  Xu, Changsheng
Adobe PDF(4322Kb)  |  收藏  |  浏览/下载:339/66  |  提交时间:2021/03/01
Semantics  Correlation  Aircraft propulsion  Deep learning  Bridges  Aircraft  Task analysis  Cross-modal retrieval  adversarial hashing  multi-level correlation