CASIA OpenIR

浏览/检索结果: 共9条,第1-9条 帮助

限定条件    
已选(0)清除 条数/页:   排序方式:
Understanding and Mitigating Overfitting in Prompt Tuning for Vision-Language Models 期刊论文
IEEE TRANSACTIONS ON CIRCUITS AND SYSTEMS FOR VIDEO TECHNOLOGY, 2023, 卷号: 33, 期号: 9, 页码: 4616-4629
作者:  Ma, Chengcheng;  Liu, Yang;  Deng, Jiankang;  Xie, Lingxi;  Dong, Weiming;  Xu, Changsheng
Adobe PDF(1644Kb)  |  收藏  |  浏览/下载:80/11  |  提交时间:2023/11/16
Vision-language model  prompt tuning  over-fitting  subspace learning  gradient projection  
Holographic Feature Learning of Egocentric-Exocentric Videos for Multi-Domain Action Recognition 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2022, 卷号: 24, 页码: 2273-2286
作者:  Huang, Yi;  Yang, Xiaoshan;  Gao, Junyun;  Xu, Changsheng
Adobe PDF(2409Kb)  |  收藏  |  浏览/下载:288/61  |  提交时间:2022/07/25
Videos  Feature extraction  Visualization  Task analysis  Computational modeling  Target recognition  Prototypes  Egocentric videos  exocentric videos  holographic feature  multi-domain  action recognition  
Weakly-Supervised Video Object Grounding Via Learning Uni-Modal Associations 期刊论文
IEEE Transactions on Multimedia, 2022, 卷号: 25, 页码: 1-12
作者:  Wang, Wei;  Gao, Junyu;  Xu, Changsheng
Adobe PDF(5406Kb)  |  收藏  |  浏览/下载:78/21  |  提交时间:2023/04/25
Visualization  Grounding  Task analysis  Prototypes  Annotations  Uncertainty  Proposals  Cross-modal retrieval  weakly-supervised learning  video object grounding  uni-modal association  
Visual affordance detection using an efficient attention convolutional neural network 期刊论文
NEUROCOMPUTING, 2021, 卷号: 440, 期号: 2021, 页码: 36-44
作者:  Gu, Qipeng;  Su, Jianhua;  Yuan, Lei
Adobe PDF(1561Kb)  |  收藏  |  浏览/下载:290/62  |  提交时间:2021/06/15
Affordance detection  Attention mechanism  Up-sampling layer  
AutoDet: Pyramid Network Architecture Search for Object Detection 期刊论文
INTERNATIONAL JOURNAL OF COMPUTER VISION, 2021, 期号: 4, 页码: 1087-1105
作者:  Li, Zhihang;  Xi, Teng;  Zhang, Gang;  Liu, Jingtuo;  He, Ran
Adobe PDF(2604Kb)  |  收藏  |  浏览/下载:244/38  |  提交时间:2021/03/01
Object detection  Neural architecture search  Feature pyramids  
Visual Question Answering With Dense Inter- and Intra-Modality Interactions 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2021, 卷号: 23, 页码: 3518-3529
作者:  Liu, Fei;  Liu, Jing;  Fang, Zhiwei;  Hong, Richang;  Lu, Hanqing
Adobe PDF(2891Kb)  |  收藏  |  浏览/下载:247/55  |  提交时间:2021/12/28
Visualization  Knowledge discovery  Connectors  Encoding  Task analysis  Image coding  Stacking  Visual question answering  attention  dense interactions  
Show, Tell, and Polish: Ruminant Decoding for Image Captioning 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2020, 卷号: 22, 期号: 8, 页码: 2149-2162
作者:  Guo, Longteng;  Liu, Jing;  Lu, Shichen;  Lu, Hanqing
Adobe PDF(4378Kb)  |  收藏  |  浏览/下载:197/28  |  提交时间:2020/08/31
Image captioning  Multi-pass decoding  Rumination  
Deep Hierarchical Encoder-Decoder Network for Image Captioning 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2019, 卷号: 21, 期号: 11, 页码: 2942-2956
作者:  Xiao, Xinyu;  Wang, Lingfeng;  Ding, Kun;  Xiang, Shiming;  Pan, Chunhong
收藏  |  浏览/下载:285/0  |  提交时间:2020/03/30
Visualization  Semantics  Hidden Markov models  Decoding  Logic gates  Training  Computer architecture  Deep hierarchical structure  encoder-decoder  LSTM  image captioning  retrieval  vision-sentence  
Dense semantic embedding network for image captioning 期刊论文
PATTERN RECOGNITION, 2019, 卷号: 90, 页码: 285-296
作者:  Xiao, Xinyu;  Wang, Lingfeng;  Ding, Kun;  Xiang, Shiming;  Pan, Chunhong
收藏  |  浏览/下载:340/0  |  提交时间:2019/04/23
Image captioning  Retrieval  High-level semantic information  Visual concept  Densely embedding  Long short-term memory