CASIA OpenIR

浏览/检索结果: 共11条,第1-10条 帮助

限定条件        
已选(0)清除 条数/页:   排序方式:
Weakly-Supervised Video Object Grounding Via Learning Uni-Modal Associations 期刊论文
IEEE Transactions on Multimedia, 2022, 卷号: 25, 页码: 1-12
作者:  Wang, Wei;  Gao, Junyu;  Xu, Changsheng
Adobe PDF(5406Kb)  |  收藏  |  浏览/下载:140/41  |  提交时间:2023/04/25
Visualization  Grounding  Task analysis  Prototypes  Annotations  Uncertainty  Proposals  Cross-modal retrieval  weakly-supervised learning  video object grounding  uni-modal association  
Global Instance Tracking: Locating Target More Like Humans 期刊论文
IEEE TRANSACTIONS ON PATTERN ANALYSIS AND MACHINE INTELLIGENCE, 2023, 卷号: 45, 期号: 1, 页码: 576-592
作者:  Hu, Shiyu;  Zhao, Xin;  Huang, Lianghua;  Huang, Kaiqi
Adobe PDF(15055Kb)  |  收藏  |  浏览/下载:253/56  |  提交时间:2023/02/22
Global instance tracking  single object tracking  benchmark dataset  performance evaluation  human tracking ability  
PDNet: Toward Better One-Stage Object Detection With Prediction Decoupling 期刊论文
IEEE Transactions on Image Processing, 2022, 卷号: 31, 页码: 5121-5133
作者:  Yang, Li;  Xu, Yan;  Wang, Shaoru;  Yuan, Chunfeng;  Zhang, Ziqi;  Li, Bing;  Hu, Weiming
Adobe PDF(3190Kb)  |  收藏  |  浏览/下载:323/41  |  提交时间:2022/09/19
Object detection  prediction decoupling  convolutional neural network  
Learning Coarse-to-Fine Graph Neural Networks for Video-Text Retrieval 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2021, 卷号: 23, 页码: 2386-2397
作者:  Wang, Wei;  Gao, Junyu;  Yang, Xiaoshan;  Xu, Changsheng
Adobe PDF(2165Kb)  |  收藏  |  浏览/下载:356/47  |  提交时间:2021/11/02
Feature extraction  Encoding  Task analysis  Semantics  Data models  Cognition  Focusing  Video-text retrieval  graph neural network  coarse-to-fine strategy  
Richly Activated Graph Convolutional Network for Robust Skeleton-Based Action Recognition 期刊论文
IEEE TRANSACTIONS ON CIRCUITS AND SYSTEMS FOR VIDEO TECHNOLOGY, 2021, 卷号: 31, 期号: 5, 页码: 1915-1925
作者:  Song, Yi-Fan;  Zhang, Zhang;  Shan, Caifeng;  Wang, Liang
Adobe PDF(3381Kb)  |  收藏  |  浏览/下载:419/67  |  提交时间:2021/06/15
Skeleton  Robustness  Noise measurement  Three-dimensional displays  Degradation  Standards  Feature extraction  Action recognition  skeleton  activation map  graph convolutional network  occlusion  jittering  
Learning Aligned Image-Text Representations Using Graph Attentive Relational Network 期刊论文
IEEE TRANSACTIONS ON IMAGE PROCESSING, 2021, 期号: 30, 页码: 1840-1852
作者:  Jing, Ya;  Wang, Wei;  Wang, Liang;  Tan, Tieniu
Adobe PDF(4532Kb)  |  收藏  |  浏览/下载:367/62  |  提交时间:2021/03/08
Graph neural networks  Visualization  Semantics  Task analysis  Feature extraction  Annotations  Recurrent neural networks  Image-text matching  cross-modal retrieval  person search  graph neural network  
Part-based Structured Representation Learning for Person Re-identification 期刊论文
ACM TRANSACTIONS ON MULTIMEDIA COMPUTING COMMUNICATIONS AND APPLICATIONS, 2020, 卷号: 16, 期号: 4, 页码: 22
作者:  Li, Yaoyu;  Yao, Hantao;  Zhang, Tianzhu;  Xu, Changsheng
Adobe PDF(19052Kb)  |  收藏  |  浏览/下载:325/48  |  提交时间:2021/03/08
Person re-identification  representation learning  graph convolutional network  
Recurrent Prediction with Spatio-temporal Attention for Crowd Attribute Recognition 期刊论文
IEEE Transactions on Circuits and Systems for Video Technology, 2019, 卷号: 30, 期号: Early Access, 页码: 1 - 1
作者:  Li, Qiaozhe;  Zhao, Xin;  He, Ran;  Huang, Kaiqi
浏览  |  Adobe PDF(2648Kb)  |  收藏  |  浏览/下载:423/107  |  提交时间:2020/01/14
Crowd video understanding , Attribute recognition , Attention mechanism , Multi-label classification  
Fast A3RL: Aesthetics-Aware Adversarial Reinforcement Learning for Image Cropping 期刊论文
IEEE TRANSACTIONS ON IMAGE PROCESSING, 2019, 卷号: 28, 期号: 10, 页码: 5105-5120
作者:  Li, Debang;  Wu, Huikai;  Zhang, Junge;  Huang, Kaiqi
Adobe PDF(6588Kb)  |  收藏  |  浏览/下载:403/46  |  提交时间:2019/12/16
Reinforcement learning  adversarial learning  image cropping  
Deep Multi-Modality Adversarial Networks for Unsupervised Domain Adaptation 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2019, 卷号: 21, 期号: 9, 页码: 2419-2431
作者:  Ma, Xinhong;  Zhang, Tianzhu;  Xu, Changsheng
Adobe PDF(2142Kb)  |  收藏  |  浏览/下载:381/49  |  提交时间:2019/12/16
Unsupervised domain adaptation  triplet loss  stacked attention  multi-modality  social event recognition