CASIA OpenIR

浏览/检索结果: 共15条,第1-10条 帮助

限定条件        
已选(0)清除 条数/页:   排序方式:
Weakly-Supervised Video Object Grounding Via Learning Uni-Modal Associations 期刊论文
IEEE Transactions on Multimedia, 2022, 卷号: 25, 页码: 1-12
作者:  Wang, Wei;  Gao, Junyu;  Xu, Changsheng
Adobe PDF(5406Kb)  |  收藏  |  浏览/下载:102/30  |  提交时间:2023/04/25
Visualization  Grounding  Task analysis  Prototypes  Annotations  Uncertainty  Proposals  Cross-modal retrieval  weakly-supervised learning  video object grounding  uni-modal association  
Learning Semantic-Aware Spatial-Temporal Attention for Interpretable Action Recognition 期刊论文
IEEE TRANSACTIONS ON CIRCUITS AND SYSTEMS FOR VIDEO TECHNOLOGY, 2022, 卷号: 32, 期号: 8, 页码: 5213-5224
作者:  Fu, Jie;  Gao, Junyu;  Xu, Changsheng
收藏  |  浏览/下载:299/0  |  提交时间:2022/09/19
Visualization  Semantics  Task analysis  Three-dimensional displays  Feature extraction  Solid modeling  Predictive models  Semantic-aware  spatial-temporal attention  interpretable  action recognition  
Explicit Cross-Modal Representation Learning for Visual Commonsense Reasoning 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2022, 卷号: 24, 页码: 2986-2997
作者:  Zhang, Xi;  Zhang, Feifei;  Xu, Changsheng
收藏  |  浏览/下载:324/0  |  提交时间:2022/07/25
Cognition  Video recording  Syntactics  Visualization  Task analysis  Semantics  Linguistics  Visual Commonsense Reasoning  explicit reasoning  syntactic structure  interpretability  
Multi-Modal Meta Multi-Task Learning for Social Media Rumor Detection 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2022, 卷号: 24, 页码: 1449-1459
作者:  Zhang, Huaiwen;  Qian, Shengsheng;  Fang, Quan;  Xu, Changsheng
收藏  |  浏览/下载:259/0  |  提交时间:2022/06/10
Task analysis  Social networking (online)  Feature extraction  Learning systems  Semantics  Media  Blogs  Meta learning  multi-modal  multi-task learning  rumor detection  social media  
Learning Video Moment Retrieval Without a Single Annotated Video 期刊论文
IEEE TRANSACTIONS ON CIRCUITS AND SYSTEMS FOR VIDEO TECHNOLOGY, 2022, 卷号: 32, 期号: 3, 页码: 1646-1657
作者:  Gao, Junyu;  Xu, Changsheng
收藏  |  浏览/下载:214/0  |  提交时间:2022/06/06
Visualization  Task analysis  Generators  Training  Graph neural networks  Semantics  Detectors  Video moment retrieval  graph neural network  unpaired learning  
Heterogeneous Hierarchical Feature Aggregation Network for Personalized Micro-Video Recommendation 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2022, 卷号: 24, 页码: 805-818
作者:  Cai, Desheng;  Qian, Shengsheng;  Fang, Quan;  Xu, Changsheng
收藏  |  浏览/下载:271/0  |  提交时间:2022/06/06
Graph neural networks  Task analysis  Semantics  Aggregates  Data structures  Collaboration  Visualization  Heterogeneous graph  micro-video recommendation  multi-modal  
Geometry Sensitive Cross-Modal Reasoning for Composed Query Based Image Retrieval 期刊论文
IEEE TRANSACTIONS ON IMAGE PROCESSING, 2022, 卷号: 31, 页码: 1000-1011
作者:  Zhang, Feifei;  Xu, Mingliang;  Xu, Changsheng
收藏  |  浏览/下载:242/0  |  提交时间:2022/02/16
Visualization  Image retrieval  Semantics  Cognition  Geometry  Task analysis  Electronic mail  Composed query based image retrieval  semantic gap  spatial structure  inter-modal attention  text-guided visual reasoning  
Unsupervised Video Summarization via Relation-Aware Assignment Learning 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2021, 卷号: 23, 页码: 3203-3214
作者:  Gao, Junyu;  Yang, Xiaoshan;  Zhang, Yingying;  Xu, Changsheng
Adobe PDF(3649Kb)  |  收藏  |  浏览/下载:315/62  |  提交时间:2021/11/03
Feature extraction  Training  Optimization  Semantics  Recurrent neural networks  Task analysis  Graph neural network  unsupervised learning  video summarization  
Learning Coarse-to-Fine Graph Neural Networks for Video-Text Retrieval 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2021, 卷号: 23, 页码: 2386-2397
作者:  Wang, Wei;  Gao, Junyu;  Yang, Xiaoshan;  Xu, Changsheng
Adobe PDF(2165Kb)  |  收藏  |  浏览/下载:326/45  |  提交时间:2021/11/02
Feature extraction  Encoding  Task analysis  Semantics  Data models  Cognition  Focusing  Video-text retrieval  graph neural network  coarse-to-fine strategy  
Heterogeneous Community Question Answering via Social-Aware Multi-Modal Co-Attention Convolutional Matching 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2021, 卷号: 23, 页码: 2321-2334
作者:  Hu, Jun;  Qian, Shengsheng;  Fang, Quan;  Xu, Changsheng
收藏  |  浏览/下载:237/0  |  提交时间:2021/11/02
Visualization  Semantics  Knowledge discovery  Context modeling  Portable computers  Task analysis  Object detection  Question-answering  attention  multi-modal  social multimedia