CASIA OpenIR
(本次检索基于用户作品认领结果)

浏览/检索结果: 共41条,第1-10条 帮助

限定条件        
已选(0)清除 条数/页:   排序方式:
VQACL: A Novel Visual Question Answering Continual Learning Setting 会议论文
, Canada, 2023
作者:  Zhang X(张熙);  Feifei Zhang;  Changsheng Xu
Adobe PDF(1199Kb)  |  收藏  |  浏览/下载:21/6  |  提交时间:2024/07/08
Part-aware Prompt Tuning For Weakly Supervised Referring Expression Grounding 会议论文
, Amsterdam, 2024-1-29
作者:  Chenlin, Zhao;  Jiabo, Ye;  Yaguang, Song;  Ming, Yan;  Xiaoshan, Yang;  Changsheng, Xu
Adobe PDF(6114Kb)  |  收藏  |  浏览/下载:23/8  |  提交时间:2024/06/21
Multimodal Global Relation Knowledge Distillation for Egocentric Action Anticipation 会议论文
MM '21: Proceedings of the 29th ACM International Conference on Multimedia, Chengdu, China, 2021.10.20—2021.10.24
作者:  Huang Yi;  Yang Xiaoshan;  Xu Changsheng
Adobe PDF(1162Kb)  |  收藏  |  浏览/下载:187/72  |  提交时间:2023/06/21
Self-supervised Calorie-aware Heterogeneous Graph Networks for Food Recommendation 期刊论文
ACM Transactions on Multimedia Computing, Communications, and Applications, 2023, 卷号: 19, 期号: 1s, 页码: 1-23
作者:  Song, Yaguang;  Yang, Xiaoshan;  Xu, Changsheng
Adobe PDF(1381Kb)  |  收藏  |  浏览/下载:211/66  |  提交时间:2023/06/12
Food recommendation  recipe calories  heterogeneous graph  selfsupervised learning  
Learning Hierarchical Video Graph Networks for One-Stop Video Delivery 期刊论文
ACM Transactions on Multimedia Computing, Communications, and Applications, 2022, 卷号: 18, 期号: 1, 页码: 1-23
作者:  Song, Yaguang;  Gao, Junyu;  Yang, Xiaoshan;  Xu, Changsheng
Adobe PDF(7608Kb)  |  收藏  |  浏览/下载:174/52  |  提交时间:2023/04/25
Cross modal  video retrieval  deep learning  graph neural networks  
Many Hands Make Light Work: Transferring Knowledge from Auxiliary Tasks for Video-Text Retrieval 期刊论文
IEEE Transactions on Multimedia, 2022, 页码: 1-15
作者:  Wang, Wei;  Gao, Junyu;  Yang, Xiaoshan;  Xu, Changsheng
Adobe PDF(3679Kb)  |  收藏  |  浏览/下载:142/31  |  提交时间:2023/04/25
Explicit Cross-Modal Representation Learning for Visual Commonsense Reasoning 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2022, 卷号: 24, 页码: 2986-2997
作者:  Zhang, Xi;  Zhang, Feifei;  Xu, Changsheng
Adobe PDF(5681Kb)  |  收藏  |  浏览/下载:407/3  |  提交时间:2022/07/25
Cognition  Video recording  Syntactics  Visualization  Task analysis  Semantics  Linguistics  Visual Commonsense Reasoning  explicit reasoning  syntactic structure  interpretability  
The Model May Fit You: User-Generalized Cross-Modal Retrieval 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2021, 卷号: 24, 页码: 2998-3012
作者:  Ma, Xinhong;  Yang, Xiaoshan;  Gao, Junyu;  Xu, Changsheng
Adobe PDF(6549Kb)  |  收藏  |  浏览/下载:282/54  |  提交时间:2022/06/17
cross-modal retrieval  domain generalization  meta-learning  
StyTr2: Image Style Transfer with Transformers 会议论文
, New Orleans, Louisiana, 2022-6
作者:  Deng, Yingying;  Tang, Fan;  Dong, Weiming;  Ma, Chongyang;  Pan, Xingjia;  Wang, Lei;  Xu, Changsheng
Adobe PDF(10025Kb)  |  收藏  |  浏览/下载:269/44  |  提交时间:2022/06/14
Evo-ViT: Slow-Fast Token Evolution for Dynamic Vision Transformer 会议论文
, Online, 2022-3-1
作者:  Xu, Yifan;  Zhang, Zhijie;  Zhang, Mengdan;  Sheng, Kekai;  Li, Ke;  Dong, Weiming;  Zhang, Liqing;  Xu, Changsheng;  Sun, Xing
Adobe PDF(636Kb)  |  收藏  |  浏览/下载:285/69  |  提交时间:2022/06/14