CASIA OpenIR

浏览/检索结果: 共7条,第1-7条 帮助

限定条件    
已选(0)清除 条数/页:   排序方式:
Weakly-Supervised Video Object Grounding Via Learning Uni-Modal Associations 期刊论文
IEEE Transactions on Multimedia, 2022, 卷号: 25, 页码: 1-12
作者:  Wang, Wei;  Gao, Junyu;  Xu, Changsheng
Adobe PDF(5406Kb)  |  收藏  |  浏览/下载:84/23  |  提交时间:2023/04/25
Visualization  Grounding  Task analysis  Prototypes  Annotations  Uncertainty  Proposals  Cross-modal retrieval  weakly-supervised learning  video object grounding  uni-modal association  
End -to -end video text detection with online tracking 期刊论文
PATTERN RECOGNITION, 2021, 卷号: 113, 页码: 12
作者:  Yu, Hongyuan;  Huang, Yan;  Pi, Lihong;  Zhang, Chengquan;  Li, Xuan;  Wang, Liang
Adobe PDF(4997Kb)  |  收藏  |  浏览/下载:301/56  |  提交时间:2021/05/06
End-to-end  Video text detection  Online tracking  
Learning Aligned Image-Text Representations Using Graph Attentive Relational Network 期刊论文
IEEE TRANSACTIONS ON IMAGE PROCESSING, 2021, 期号: 30, 页码: 1840-1852
作者:  Jing, Ya;  Wang, Wei;  Wang, Liang;  Tan, Tieniu
Adobe PDF(4532Kb)  |  收藏  |  浏览/下载:312/50  |  提交时间:2021/03/08
Graph neural networks  Visualization  Semantics  Task analysis  Feature extraction  Annotations  Recurrent neural networks  Image-text matching  cross-modal retrieval  person search  graph neural network  
Long video question answering: A Matching-guided Attention Model 期刊论文
PATTERN RECOGNITION, 2020, 卷号: 102, 期号: 1, 页码: 11
作者:  Wang, Weining;  Huang, Yan;  Wang, Liang
浏览  |  Adobe PDF(1963Kb)  |  收藏  |  浏览/下载:349/68  |  提交时间:2020/06/02
Long video QA  Matching-guided attention  
Fast A3RL: Aesthetics-Aware Adversarial Reinforcement Learning for Image Cropping 期刊论文
IEEE TRANSACTIONS ON IMAGE PROCESSING, 2019, 卷号: 28, 期号: 10, 页码: 5105-5120
作者:  Li, Debang;  Wu, Huikai;  Zhang, Junge;  Huang, Kaiqi
Adobe PDF(6588Kb)  |  收藏  |  浏览/下载:367/41  |  提交时间:2019/12/16
Reinforcement learning  adversarial learning  image cropping  
Deep Multi-Modality Adversarial Networks for Unsupervised Domain Adaptation 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2019, 卷号: 21, 期号: 9, 页码: 2419-2431
作者:  Ma, Xinhong;  Zhang, Tianzhu;  Xu, Changsheng
Adobe PDF(2142Kb)  |  收藏  |  浏览/下载:326/41  |  提交时间:2019/12/16
Unsupervised domain adaptation  triplet loss  stacked attention  multi-modality  social event recognition  
Recurrent Prediction with Spatio-temporal Attention for Crowd Attribute Recognition 期刊论文
IEEE Transactions on Circuits and Systems for Video Technology, 2019, 卷号: 30, 期号: Early Access, 页码: 1 - 1
作者:  Li, Qiaozhe;  Zhao, Xin;  He, Ran;  Huang, Kaiqi
浏览  |  Adobe PDF(2648Kb)  |  收藏  |  浏览/下载:382/103  |  提交时间:2020/01/14
Crowd video understanding , Attribute recognition , Attention mechanism , Multi-label classification