CASIA OpenIR

浏览/检索结果: 共11条,第1-10条 帮助

限定条件    
已选(0)清除 条数/页:   排序方式:
Visual Question Answering With Dense Inter- and Intra-Modality Interactions 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2021, 卷号: 23, 页码: 3518-3529
作者:  Liu, Fei;  Liu, Jing;  Fang, Zhiwei;  Hong, Richang;  Lu, Hanqing
Adobe PDF(2891Kb)  |  收藏  |  浏览/下载:285/60  |  提交时间:2021/12/28
Visualization  Knowledge discovery  Connectors  Encoding  Task analysis  Image coding  Stacking  Visual question answering  attention  dense interactions  
Unsupervised Video Summarization via Relation-Aware Assignment Learning 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2021, 卷号: 23, 页码: 3203-3214
作者:  Gao, Junyu;  Yang, Xiaoshan;  Zhang, Yingying;  Xu, Changsheng
Adobe PDF(3649Kb)  |  收藏  |  浏览/下载:312/62  |  提交时间:2021/11/03
Feature extraction  Training  Optimization  Semantics  Recurrent neural networks  Task analysis  Graph neural network  unsupervised learning  video summarization  
Learning Coarse-to-Fine Graph Neural Networks for Video-Text Retrieval 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2021, 卷号: 23, 页码: 2386-2397
作者:  Wang, Wei;  Gao, Junyu;  Yang, Xiaoshan;  Xu, Changsheng
Adobe PDF(2165Kb)  |  收藏  |  浏览/下载:321/44  |  提交时间:2021/11/02
Feature extraction  Encoding  Task analysis  Semantics  Data models  Cognition  Focusing  Video-text retrieval  graph neural network  coarse-to-fine strategy  
Multi-Level Correlation Adversarial Hashing for Cross-Modal Retrieval 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2020, 卷号: 22, 期号: 12, 页码: 3101-3114
作者:  Ma, Xinhong;  Zhang, Tianzhu;  Xu, Changsheng
Adobe PDF(4322Kb)  |  收藏  |  浏览/下载:291/53  |  提交时间:2021/03/01
Semantics  Correlation  Aircraft propulsion  Deep learning  Bridges  Aircraft  Task analysis  Cross-modal retrieval  adversarial hashing  multi-level correlation  
Three-Dimensional Attention-Based Deep Ranking Model for Video Highlight Detection 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2018, 卷号: 20, 期号: 10, 页码: 2693-2705
作者:  Jiao,Yifan;  Li,Zhetao;  Huang,Shucheng;  Yang,Xiaoshan;  Liu,Bin;  Zhang,Tianzhu
浏览  |  Adobe PDF(4692Kb)  |  收藏  |  浏览/下载:516/184  |  提交时间:2018/10/10
Video Highlight Detection  Attention Model  Deep Ranking  
EgoGesture: A New Dataset and Benchmark for Egocentric Hand Gesture Recognition 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2018, 卷号: 20, 期号: 5, 页码: 1038-1050
作者:  Zhang, Yifan;  Cao, Congqi;  Cheng, Jian;  Lu, Hanqing
浏览  |  Adobe PDF(1260Kb)  |  收藏  |  浏览/下载:927/357  |  提交时间:2018/05/05
Benchmark  Dataset  Egocentric Vision  Gesture Recognition  First-person View  
Cross-Modal Hashing via Rank-Order Preserving 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2017, 卷号: 19, 期号: 3, 页码: 571-585
作者:  Kun Ding;  Bin Fan;  Chunlei Huo;  Shiming Xiang;  Chunhong Pan;  Huo CL(霍春雷)
Adobe PDF(1451Kb)  |  收藏  |  浏览/下载:828/327  |  提交时间:2016/10/24
Cross-modal Similarity Search  Cross-modal Hashing (Cmh)  Rank-order Preserving  
Multi-Instance Multi-Label Learning Combining Hierarchical Context and its Application to Image Annotation 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2016, 卷号: 18, 期号: 8, 页码: 1616-1627
作者:  Ding, Xinmiao;  Li, Bing;  Xiong, Weihua;  Guo, Wen;  Hu, Weiming;  Wang, Bo
Adobe PDF(549Kb)  |  收藏  |  浏览/下载:437/144  |  提交时间:2016/10/20
Image Annotation  Instance Context  Label Context  Multi-instance  Multi-label  
Cross-Domain Feature Learning in Multimedia 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2015, 卷号: 17, 期号: 1, 页码: 64-78
作者:  Yang, Xiaoshan;  Zhang, Tianzhu;  Xu, Changsheng;  Xu CS(徐常胜)
浏览  |  Adobe PDF(3097Kb)  |  收藏  |  浏览/下载:364/94  |  提交时间:2016/06/27
Cross-domain  Deep Learning  Feature Learning  Multi-modal  
Multi-Perspective Cost-Sensitive Context-Aware Multi-Instance Sparse Coding and Its Application to Sensitive Video Recognition 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2016, 卷号: 18, 期号: 1, 页码: 76-89
作者:  Hu, Weiming;  Ding, Xinmiao;  Li, Bing;  Wang, Jianchao;  Gao, Yan;  Wang, Fangshi;  Maybank, Stephen
Adobe PDF(2469Kb)  |  收藏  |  浏览/下载:411/93  |  提交时间:2016/03/19
Cost-sensitive Context-aware Multi-instance Sparse Coding (Mi-sc)  Horror Video Recognition  Multi-perspective Multi-instance Joint Sparse Coding (Mi-j-sc)  Video Emotional Feature Extraction  Violent Video Recognition