CASIA OpenIR

浏览/检索结果: 共26条,第1-10条 帮助

限定条件    
已选(0)清除 条数/页:   排序方式:
Explicit Cross-Modal Representation Learning for Visual Commonsense Reasoning 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2022, 卷号: 24, 页码: 2986-2997
作者:  Zhang, Xi;  Zhang, Feifei;  Xu, Changsheng
Adobe PDF(5681Kb)  |  收藏  |  浏览/下载:433/8  |  提交时间:2022/07/25
Cognition  Video recording  Syntactics  Visualization  Task analysis  Semantics  Linguistics  Visual Commonsense Reasoning  explicit reasoning  syntactic structure  interpretability  
Instance GNN: A Learning Framework for Joint Symbol Segmentation and Recognition in Online Handwritten Diagrams 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2022, 卷号: 24, 页码: 2580-2594
作者:  Yun, Xiao-Long;  Zhang, Yan-Ming;  Yin, Fei;  Liu, Cheng-Lin
Adobe PDF(3236Kb)  |  收藏  |  浏览/下载:345/6  |  提交时间:2022/07/25
Handwriting recognition  Task analysis  Grammar  Semantics  Image segmentation  Trajectory  Text recognition  Online handwritten diagram recognition  symbol segmentation  symbol recognition  freehand sketch analysis  graph neural networks  
The Model May Fit You: User-Generalized Cross-Modal Retrieval 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2021, 卷号: 24, 页码: 2998-3012
作者:  Ma, Xinhong;  Yang, Xiaoshan;  Gao, Junyu;  Xu, Changsheng
Adobe PDF(6549Kb)  |  收藏  |  浏览/下载:298/59  |  提交时间:2022/06/17
cross-modal retrieval  domain generalization  meta-learning  
Exploring the Representativity of Art Paintings 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2021, 卷号: 23, 页码: 2794-2805
作者:  Deng, Yingying;  Tang, Fan;  Dong, Weiming;  Ma, Chongyang;  Huang, Feiyue;  Deussen, Oliver;  Xu, Changsheng
Adobe PDF(5313Kb)  |  收藏  |  浏览/下载:301/49  |  提交时间:2021/11/03
Painting  Art  Image color analysis  Feature extraction  Task analysis  Engineering profession  Electronic mail  Representativity  style enhancement  feature representation  artwork evaluation  
Unsupervised Video Summarization via Relation-Aware Assignment Learning 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2021, 卷号: 23, 页码: 3203-3214
作者:  Gao, Junyu;  Yang, Xiaoshan;  Zhang, Yingying;  Xu, Changsheng
Adobe PDF(3649Kb)  |  收藏  |  浏览/下载:362/72  |  提交时间:2021/11/03
Feature extraction  Training  Optimization  Semantics  Recurrent neural networks  Task analysis  Graph neural network  unsupervised learning  video summarization  
Multi-Level Correlation Adversarial Hashing for Cross-Modal Retrieval 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2020, 卷号: 22, 期号: 12, 页码: 3101-3114
作者:  Ma, Xinhong;  Zhang, Tianzhu;  Xu, Changsheng
Adobe PDF(4322Kb)  |  收藏  |  浏览/下载:366/72  |  提交时间:2021/03/01
Semantics  Correlation  Aircraft propulsion  Deep learning  Bridges  Aircraft  Task analysis  Cross-modal retrieval  adversarial hashing  multi-level correlation  
Joint Learning in the Spatio-Temporal and Frequency Domains for Skeleton-Based Action Recognition 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2020, 卷号: 22, 期号: 9, 页码: 2207-2220
作者:  Guyue, Hu;  Bo, Cui;  Shan, Yu
Adobe PDF(4803Kb)  |  收藏  |  浏览/下载:348/72  |  提交时间:2020/09/28
Skeleton-based Action Recognition  Frequency Attention  Synchronous Local and Non-local Learning  Soft-margin Focal Loss  Pesudo Multi-task Learning  
Weakly Semantic Guided Action Recognition 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2019, 卷号: 21, 期号: 10, 页码: 2504-2517
作者:  Yu, Tingzhao;  Wang, Lingfeng;  Da, Cheng;  Gu, Huxiang;  Xiang, Shiming;  Pan, Chunhong
Adobe PDF(18774Kb)  |  收藏  |  浏览/下载:470/120  |  提交时间:2019/05/15
Semantic guided module  action recognition  cross domain  3D convolution  attention model  
Multiview Label Sharing for Visual Representations and Classifications 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2018, 卷号: 20, 期号: 4, 页码: 903-913
作者:  Zhang, Chunjie;  Cheng, Jian;  Tian, Qi
Adobe PDF(615Kb)  |  收藏  |  浏览/下载:394/110  |  提交时间:2018/10/10
Multi-view Learning  Linear Transformation  Shared Space  Image Representation  Visual Classification  
EgoGesture: A New Dataset and Benchmark for Egocentric Hand Gesture Recognition 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2018, 卷号: 20, 期号: 5, 页码: 1038-1050
作者:  Zhang, Yifan;  Cao, Congqi;  Cheng, Jian;  Lu, Hanqing
浏览  |  Adobe PDF(1260Kb)  |  收藏  |  浏览/下载:995/373  |  提交时间:2018/05/05
Benchmark  Dataset  Egocentric Vision  Gesture Recognition  First-person View