CASIA OpenIR

浏览/检索结果: 共8条,第1-8条 帮助

限定条件        
已选(0)清除 条数/页:   排序方式:
Explicit Cross-Modal Representation Learning for Visual Commonsense Reasoning 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2022, 卷号: 24, 页码: 2986-2997
作者:  Zhang, Xi;  Zhang, Feifei;  Xu, Changsheng
Adobe PDF(5681Kb)  |  收藏  |  浏览/下载:404/2  |  提交时间:2022/07/25
Cognition  Video recording  Syntactics  Visualization  Task analysis  Semantics  Linguistics  Visual Commonsense Reasoning  explicit reasoning  syntactic structure  interpretability  
Unsupervised Video Summarization via Relation-Aware Assignment Learning 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2021, 卷号: 23, 页码: 3203-3214
作者:  Gao, Junyu;  Yang, Xiaoshan;  Zhang, Yingying;  Xu, Changsheng
Adobe PDF(3649Kb)  |  收藏  |  浏览/下载:342/67  |  提交时间:2021/11/03
Feature extraction  Training  Optimization  Semantics  Recurrent neural networks  Task analysis  Graph neural network  unsupervised learning  video summarization  
ScleraSegNet: An Attention Assisted U-Net Model for Accurate Sclera Segmentation 期刊论文
IEEE Transactions on Biometrics, Behavior, and Identity Science, 2020, 卷号: 2, 期号: 1, 页码: 40-54
作者:  Wang, Caiyong;  Wang, Yunlong;  Liu, Yunfan;  He, Zhaofeng;  He, Ran;  Sun, Zhenan
浏览  |  Adobe PDF(5290Kb)  |  收藏  |  浏览/下载:425/129  |  提交时间:2020/06/10
Sclera segmentation  sclera recognition  U-net  attention mechanism  SSBC  
Long video question answering: A Matching-guided Attention Model 期刊论文
PATTERN RECOGNITION, 2020, 卷号: 102, 期号: 1, 页码: 11
作者:  Wang, Weining;  Huang, Yan;  Wang, Liang
浏览  |  Adobe PDF(1963Kb)  |  收藏  |  浏览/下载:393/79  |  提交时间:2020/06/02
Long video QA  Matching-guided attention  
Recurrent Prediction with Spatio-temporal Attention for Crowd Attribute Recognition 期刊论文
IEEE Transactions on Circuits and Systems for Video Technology, 2019, 卷号: 30, 期号: Early Access, 页码: 1 - 1
作者:  Li, Qiaozhe;  Zhao, Xin;  He, Ran;  Huang, Kaiqi
浏览  |  Adobe PDF(2648Kb)  |  收藏  |  浏览/下载:418/107  |  提交时间:2020/01/14
Crowd video understanding , Attribute recognition , Attention mechanism , Multi-label classification  
Deep Multi-Modality Adversarial Networks for Unsupervised Domain Adaptation 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2019, 卷号: 21, 期号: 9, 页码: 2419-2431
作者:  Ma, Xinhong;  Zhang, Tianzhu;  Xu, Changsheng
Adobe PDF(2142Kb)  |  收藏  |  浏览/下载:374/48  |  提交时间:2019/12/16
Unsupervised domain adaptation  triplet loss  stacked attention  multi-modality  social event recognition  
GaitNet: An end-to-end network for gait based human identification 期刊论文
PATTERN RECOGNITION, 2019, 卷号: 96, 期号: 106988, 页码: 11
作者:  Song, Chunfeng;  Huang, Yongzhen;  Huang, Yan;  Jia, Ning;  Wang, Liang
浏览  |  Adobe PDF(3015Kb)  |  收藏  |  浏览/下载:635/166  |  提交时间:2019/12/16
Gait recognition  Video-based human identification  End-to-end CNN  Joint learning  
Focal Boundary Guided Salient Object Detection 期刊论文
IEEE TRANSACTIONS ON IMAGE PROCESSING, 2019, 卷号: 28, 期号: 6, 页码: 2813-2824
作者:  Wang, Yupei;  Zhao, Xin;  Hu, Xuecai;  Li, Yin;  Huang, Kaiqi
浏览  |  Adobe PDF(3275Kb)  |  收藏  |  浏览/下载:610/273  |  提交时间:2019/04/19
Visual saliency detection  salient object segmentation  boundary detection  deep learning