CASIA OpenIR

浏览/检索结果: 共6条,第1-6条 帮助

限定条件            
已选(0)清除 条数/页:   排序方式:
Human Parsing With Part-Aware Relation Modeling 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2022, 卷号: 25, 页码: 2601-2612
作者:  Zhang, Xiaomei;  Chen, Yingying;  Tang, Ming;  Wang, Jinqiao;  Zhu, Xiangyu;  Lei, Zhen
Adobe PDF(6053Kb)  |  收藏  |  浏览/下载:172/24  |  提交时间:2023/11/17
Human parsing  modeling  part-aware relation  
Explicit Cross-Modal Representation Learning for Visual Commonsense Reasoning 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2022, 卷号: 24, 页码: 2986-2997
作者:  Zhang, Xi;  Zhang, Feifei;  Xu, Changsheng
Adobe PDF(5681Kb)  |  收藏  |  浏览/下载:416/4  |  提交时间:2022/07/25
Cognition  Video recording  Syntactics  Visualization  Task analysis  Semantics  Linguistics  Visual Commonsense Reasoning  explicit reasoning  syntactic structure  interpretability  
Learning Coarse-to-Fine Graph Neural Networks for Video-Text Retrieval 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2021, 卷号: 23, 页码: 2386-2397
作者:  Wang, Wei;  Gao, Junyu;  Yang, Xiaoshan;  Xu, Changsheng
Adobe PDF(2165Kb)  |  收藏  |  浏览/下载:356/47  |  提交时间:2021/11/02
Feature extraction  Encoding  Task analysis  Semantics  Data models  Cognition  Focusing  Video-text retrieval  graph neural network  coarse-to-fine strategy  
Joint Learning in the Spatio-Temporal and Frequency Domains for Skeleton-Based Action Recognition 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2020, 卷号: 22, 期号: 9, 页码: 2207-2220
作者:  Guyue, Hu;  Bo, Cui;  Shan, Yu
Adobe PDF(4803Kb)  |  收藏  |  浏览/下载:335/67  |  提交时间:2020/09/28
Skeleton-based Action Recognition  Frequency Attention  Synchronous Local and Non-local Learning  Soft-margin Focal Loss  Pesudo Multi-task Learning  
Knowledge-Based Topic Model for Multi-Modal Social Event Analysis 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2020, 卷号: 22, 期号: 8, 页码: 2098-2110
作者:  Xue, Feng;  Hong, Richang;  He, Xiangnan;  Wang, Jianwei;  Qian, Shengsheng;  Xu, Changsheng
收藏  |  浏览/下载:247/0  |  提交时间:2020/08/31
Analytical models  Knowledge based systems  Social networking (online)  Data mining  Data models  Internet  Knowledge engineering  Knowledge embedding  multi-modal  topic coherence  event classification  
Bidirectional Attention-Recognition Model for Fine-Grained Object Classification 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2020, 卷号: 22, 期号: 7, 页码: 1785-1795
作者:  Liu, Chuanbin;  Xie, Hongtao;  Zha, Zhengjun;  Yu, Lingyun;  Chen, Zhineng;  Zhang, Yongdong
收藏  |  浏览/下载:211/0  |  提交时间:2020/08/03
Fine-grained object classification  interpretable machine learning  visual attention  pattern recognition  data augmentation