CASIA OpenIR

浏览/检索结果: 共44条,第1-10条 帮助

限定条件    
已选(0)清除 条数/页:   排序方式:
AnANet: Association and Alignment Network for Modeling Implicit Relevance in Cross-Modal Correlation Classification 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2023, 卷号: 25, 页码: 7867-7880
作者:  Xu, Nan;  Wang, Junyan;  Tian, Yuan;  Zhang, Ruike;  Mao, Wenji
收藏  |  浏览/下载:6/0  |  提交时间:2024/03/26
Association and alignment network  classification scheme  cross-modal correlation  implicit relevance  
The Model May Fit You: User-Generalized Cross-Modal Retrieval 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2021, 卷号: 24, 页码: 2998-3012
作者:  Ma, Xinhong;  Yang, Xiaoshan;  Gao, Junyu;  Xu, Changsheng
Adobe PDF(6549Kb)  |  收藏  |  浏览/下载:236/46  |  提交时间:2022/06/17
cross-modal retrieval  domain generalization  meta-learning  
Visual Question Answering With Dense Inter- and Intra-Modality Interactions 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2021, 卷号: 23, 页码: 3518-3529
作者:  Liu, Fei;  Liu, Jing;  Fang, Zhiwei;  Hong, Richang;  Lu, Hanqing
Adobe PDF(2891Kb)  |  收藏  |  浏览/下载:261/58  |  提交时间:2021/12/28
Visualization  Knowledge discovery  Connectors  Encoding  Task analysis  Image coding  Stacking  Visual question answering  attention  dense interactions  
Unsupervised Video Summarization via Relation-Aware Assignment Learning 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2021, 卷号: 23, 页码: 3203-3214
作者:  Gao, Junyu;  Yang, Xiaoshan;  Zhang, Yingying;  Xu, Changsheng
Adobe PDF(3649Kb)  |  收藏  |  浏览/下载:287/59  |  提交时间:2021/11/03
Feature extraction  Training  Optimization  Semantics  Recurrent neural networks  Task analysis  Graph neural network  unsupervised learning  video summarization  
Learning Coarse-to-Fine Graph Neural Networks for Video-Text Retrieval 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2021, 卷号: 23, 页码: 2386-2397
作者:  Wang, Wei;  Gao, Junyu;  Yang, Xiaoshan;  Xu, Changsheng
Adobe PDF(2165Kb)  |  收藏  |  浏览/下载:304/42  |  提交时间:2021/11/02
Feature extraction  Encoding  Task analysis  Semantics  Data models  Cognition  Focusing  Video-text retrieval  graph neural network  coarse-to-fine strategy  
Multi-Level Correlation Adversarial Hashing for Cross-Modal Retrieval 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2020, 卷号: 22, 期号: 12, 页码: 3101-3114
作者:  Ma, Xinhong;  Zhang, Tianzhu;  Xu, Changsheng
Adobe PDF(4322Kb)  |  收藏  |  浏览/下载:272/51  |  提交时间:2021/03/01
Semantics  Correlation  Aircraft propulsion  Deep learning  Bridges  Aircraft  Task analysis  Cross-modal retrieval  adversarial hashing  multi-level correlation  
Joint Learning in the Spatio-Temporal and Frequency Domains for Skeleton-Based Action Recognition 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2020, 卷号: 22, 期号: 9, 页码: 2207-2220
作者:  Guyue, Hu;  Bo, Cui;  Shan, Yu
Adobe PDF(4803Kb)  |  收藏  |  浏览/下载:274/54  |  提交时间:2020/09/28
Skeleton-based Action Recognition  Frequency Attention  Synchronous Local and Non-local Learning  Soft-margin Focal Loss  Pesudo Multi-task Learning  
Knowledge-Based Topic Model for Multi-Modal Social Event Analysis 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2020, 卷号: 22, 期号: 8, 页码: 2098-2110
作者:  Xue, Feng;  Hong, Richang;  He, Xiangnan;  Wang, Jianwei;  Qian, Shengsheng;  Xu, Changsheng
收藏  |  浏览/下载:214/0  |  提交时间:2020/08/31
Analytical models  Knowledge based systems  Social networking (online)  Data mining  Data models  Internet  Knowledge engineering  Knowledge embedding  multi-modal  topic coherence  event classification  
Bidirectional Attention-Recognition Model for Fine-Grained Object Classification 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2020, 卷号: 22, 期号: 7, 页码: 1785-1795
作者:  Liu, Chuanbin;  Xie, Hongtao;  Zha, Zhengjun;  Yu, Lingyun;  Chen, Zhineng;  Zhang, Yongdong
收藏  |  浏览/下载:176/0  |  提交时间:2020/08/03
Fine-grained object classification  interpretable machine learning  visual attention  pattern recognition  data augmentation  
Weakly Semantic Guided Action Recognition 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2019, 卷号: 21, 期号: 10, 页码: 2504-2517
作者:  Yu, Tingzhao;  Wang, Lingfeng;  Da, Cheng;  Gu, Huxiang;  Xiang, Shiming;  Pan, Chunhong
浏览  |  Adobe PDF(18774Kb)  |  收藏  |  浏览/下载:417/108  |  提交时间:2019/05/15
Semantic guided module  action recognition  cross domain  3D convolution  attention model