CASIA OpenIR

浏览/检索结果: 共36条,第1-10条 帮助

限定条件    
已选(0)清除 条数/页:   排序方式:
Multi-Source Knowledge Reasoning Graph Network for Multi-Modal Commonsense Inference 期刊论文
ACM TRANSACTIONS ON MULTIMEDIA COMPUTING COMMUNICATIONS AND APPLICATIONS, 2023, 卷号: 19, 期号: 4, 页码: 17
作者:  Ma, Xuan;  Yang, Xiaoshan;  Xu, Changsheng
收藏  |  浏览/下载:57/0  |  提交时间:2023/11/17
Knowledge reasoning  multi-modal commonsense inference  graph neural network  
Zero-Shot Predicate Prediction for Scene Graph Parsing 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2023, 卷号: 25, 页码: 3140-3153
作者:  Li, Yiming;  Yang, Xiaoshan;  Huang, Xuhui;  Ma, Zhe;  Xu, Changsheng
收藏  |  浏览/下载:111/0  |  提交时间:2023/11/17
Deep learning  zero-shot  scene graph  
Global Instance Tracking: Locating Target More Like Humans 期刊论文
IEEE TRANSACTIONS ON PATTERN ANALYSIS AND MACHINE INTELLIGENCE, 2023, 卷号: 45, 期号: 1, 页码: 576-592
作者:  Hu, Shiyu;  Zhao, Xin;  Huang, Lianghua;  Huang, Kaiqi
Adobe PDF(15055Kb)  |  收藏  |  浏览/下载:211/54  |  提交时间:2023/02/22
Global instance tracking  single object tracking  benchmark dataset  performance evaluation  human tracking ability  
Temporal-adaptive sparse feature aggregation for video object detection 期刊论文
PATTERN RECOGNITION, 2022, 卷号: 127, 页码: 108587
作者:  He, Fei;  Li, Qiaozhe;  Zhao, Xin;  Huang, Kaiqi
Adobe PDF(1549Kb)  |  收藏  |  浏览/下载:287/38  |  提交时间:2022/06/10
Video object detection  Temporal-adaptive sparse sampling  Pixel-adaptive aggregation  Object-relational aggregation  
PDNet: Toward Better One-Stage Object Detection With Prediction Decoupling 期刊论文
IEEE Transactions on Image Processing, 2022, 卷号: 31, 页码: 5121-5133
作者:  Yang, Li;  Xu, Yan;  Wang, Shaoru;  Yuan, Chunfeng;  Zhang, Ziqi;  Li, Bing;  Hu, Weiming
Adobe PDF(3190Kb)  |  收藏  |  浏览/下载:222/32  |  提交时间:2022/09/19
Object detection  prediction decoupling  convolutional neural network  
Holographic Feature Learning of Egocentric-Exocentric Videos for Multi-Domain Action Recognition 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2022, 卷号: 24, 页码: 2273-2286
作者:  Huang, Yi;  Yang, Xiaoshan;  Gao, Junyun;  Xu, Changsheng
Adobe PDF(2409Kb)  |  收藏  |  浏览/下载:301/61  |  提交时间:2022/07/25
Videos  Feature extraction  Visualization  Task analysis  Computational modeling  Target recognition  Prototypes  Egocentric videos  exocentric videos  holographic feature  multi-domain  action recognition  
Explicit Cross-Modal Representation Learning for Visual Commonsense Reasoning 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2022, 卷号: 24, 页码: 2986-2997
作者:  Zhang, Xi;  Zhang, Feifei;  Xu, Changsheng
收藏  |  浏览/下载:278/0  |  提交时间:2022/07/25
Cognition  Video recording  Syntactics  Visualization  Task analysis  Semantics  Linguistics  Visual Commonsense Reasoning  explicit reasoning  syntactic structure  interpretability  
Attribute-Induced Bias Eliminating for Transductive Zero-Shot Learning 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2022, 卷号: 24, 页码: 1933-1942
作者:  Yao, Hantao;  Min, Shaobo;  Zhang, Yongdong;  Xu, Changsheng
收藏  |  浏览/下载:187/0  |  提交时间:2022/06/10
Semantics  Visualization  Bridges  Training  Knowledge transfer  Image recognition  Topology  Transductive Zero-Shot Learning  Graph Attribute Embedding  Attribute-Induced Bias Eliminating  Semantic-Visual Alignment  
Weakly-Supervised Video Object Grounding Via Learning Uni-Modal Associations 期刊论文
IEEE Transactions on Multimedia, 2022, 卷号: 25, 页码: 1-12
作者:  Wang, Wei;  Gao, Junyu;  Xu, Changsheng
Adobe PDF(5406Kb)  |  收藏  |  浏览/下载:80/22  |  提交时间:2023/04/25
Visualization  Grounding  Task analysis  Prototypes  Annotations  Uncertainty  Proposals  Cross-modal retrieval  weakly-supervised learning  video object grounding  uni-modal association  
Practical Face Swapping Detection Based on Identity Spatial Constraints 会议论文
, 中国广东省深圳市+线上, 2021-8-4至2021-8-7
作者:  Jiang J(姜君);  Wang B(王博);  Li B(李兵);  Hu WM(胡卫明)
Adobe PDF(2496Kb)  |  收藏  |  浏览/下载:183/39  |  提交时间:2022/01/10