CASIA OpenIR

浏览/检索结果: 共4条,第1-4条 帮助

限定条件                            
已选(0)清除 条数/页:   排序方式:
Explicit Cross-Modal Representation Learning for Visual Commonsense Reasoning 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2022, 卷号: 24, 页码: 2986-2997
作者:  Zhang, Xi;  Zhang, Feifei;  Xu, Changsheng
Adobe PDF(5681Kb)  |  收藏  |  浏览/下载:399/1  |  提交时间:2022/07/25
Cognition  Video recording  Syntactics  Visualization  Task analysis  Semantics  Linguistics  Visual Commonsense Reasoning  explicit reasoning  syntactic structure  interpretability  
Holographic Feature Learning of Egocentric-Exocentric Videos for Multi-Domain Action Recognition 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2022, 卷号: 24, 页码: 2273-2286
作者:  Huang, Yi;  Yang, Xiaoshan;  Gao, Junyun;  Xu, Changsheng
Adobe PDF(2409Kb)  |  收藏  |  浏览/下载:366/74  |  提交时间:2022/07/25
Videos  Feature extraction  Visualization  Task analysis  Computational modeling  Target recognition  Prototypes  Egocentric videos  exocentric videos  holographic feature  multi-domain  action recognition  
Deep Hierarchical Encoder-Decoder Network for Image Captioning 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2019, 卷号: 21, 期号: 11, 页码: 2942-2956
作者:  Xiao, Xinyu;  Wang, Lingfeng;  Ding, Kun;  Xiang, Shiming;  Pan, Chunhong
收藏  |  浏览/下载:334/0  |  提交时间:2020/03/30
Visualization  Semantics  Hidden Markov models  Decoding  Logic gates  Training  Computer architecture  Deep hierarchical structure  encoder-decoder  LSTM  image captioning  retrieval  vision-sentence  
Deep-Structured Event Modeling for User-Generated Photos 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2018, 卷号: 20, 期号: 8, 页码: 2100-2113
作者:  Yang, Xiaoshan;  Zhang, Tianzhu;  Xu, Changsheng
收藏  |  浏览/下载:330/0  |  提交时间:2019/12/16
Event analysis  unusual event detection  deep learning