CASIA OpenIR

浏览/检索结果: 共11条,第1-10条 帮助

限定条件    
已选(0)清除 条数/页:   排序方式:
Instance GNN: A Learning Framework for Joint Symbol Segmentation and Recognition in Online Handwritten Diagrams 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2022, 卷号: 24, 页码: 2580-2594
作者:  Yun, Xiao-Long;  Zhang, Yan-Ming;  Yin, Fei;  Liu, Cheng-Lin
Adobe PDF(3236Kb)  |  收藏  |  浏览/下载:300/1  |  提交时间:2022/07/25
Handwriting recognition  Task analysis  Grammar  Semantics  Image segmentation  Trajectory  Text recognition  Online handwritten diagram recognition  symbol segmentation  symbol recognition  freehand sketch analysis  graph neural networks  
The Model May Fit You: User-Generalized Cross-Modal Retrieval 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2021, 卷号: 24, 页码: 2998-3012
作者:  Ma, Xinhong;  Yang, Xiaoshan;  Gao, Junyu;  Xu, Changsheng
Adobe PDF(6549Kb)  |  收藏  |  浏览/下载:277/53  |  提交时间:2022/06/17
cross-modal retrieval  domain generalization  meta-learning  
Visual Question Answering With Dense Inter- and Intra-Modality Interactions 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2021, 卷号: 23, 页码: 3518-3529
作者:  Liu, Fei;  Liu, Jing;  Fang, Zhiwei;  Hong, Richang;  Lu, Hanqing
Adobe PDF(2891Kb)  |  收藏  |  浏览/下载:322/71  |  提交时间:2021/12/28
Visualization  Knowledge discovery  Connectors  Encoding  Task analysis  Image coding  Stacking  Visual question answering  attention  dense interactions  
Unsupervised Video Summarization via Relation-Aware Assignment Learning 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2021, 卷号: 23, 页码: 3203-3214
作者:  Gao, Junyu;  Yang, Xiaoshan;  Zhang, Yingying;  Xu, Changsheng
Adobe PDF(3649Kb)  |  收藏  |  浏览/下载:336/65  |  提交时间:2021/11/03
Feature extraction  Training  Optimization  Semantics  Recurrent neural networks  Task analysis  Graph neural network  unsupervised learning  video summarization  
Show, Tell, and Polish: Ruminant Decoding for Image Captioning 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2020, 卷号: 22, 期号: 8, 页码: 2149-2162
作者:  Guo, Longteng;  Liu, Jing;  Lu, Shichen;  Lu, Hanqing
Adobe PDF(4378Kb)  |  收藏  |  浏览/下载:227/35  |  提交时间:2020/08/31
Image captioning  Multi-pass decoding  Rumination  
Recall What You See Continually Using GridLSTM in Image Captioning 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2020, 卷号: 22, 期号: 3, 页码: 808-818
作者:  Wu, Lingxiang;  Xu, Min;  Wang, Jinqiao;  Perry, Stuart
收藏  |  浏览/下载:169/0  |  提交时间:2020/06/02
Visualization  Decoding  Task analysis  Neural networks  Training  Computational modeling  Logic gates  Image captioning  GridLSTM  recurrent neural network  
Three-Dimensional Attention-Based Deep Ranking Model for Video Highlight Detection 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2018, 卷号: 20, 期号: 10, 页码: 2693-2705
作者:  Jiao,Yifan;  Li,Zhetao;  Huang,Shucheng;  Yang,Xiaoshan;  Liu,Bin;  Zhang,Tianzhu
浏览  |  Adobe PDF(4692Kb)  |  收藏  |  浏览/下载:537/188  |  提交时间:2018/10/10
Video Highlight Detection  Attention Model  Deep Ranking  
EgoGesture: A New Dataset and Benchmark for Egocentric Hand Gesture Recognition 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2018, 卷号: 20, 期号: 5, 页码: 1038-1050
作者:  Zhang, Yifan;  Cao, Congqi;  Cheng, Jian;  Lu, Hanqing
浏览  |  Adobe PDF(1260Kb)  |  收藏  |  浏览/下载:961/364  |  提交时间:2018/05/05
Benchmark  Dataset  Egocentric Vision  Gesture Recognition  First-person View  
Text2Video: An End-to-end Learning Framework for Expressing Text With Videos 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2018, 卷号: 20, 期号: 9, 页码: 2360-2370
作者:  Yang, Xiaoshan;  Zhang, Tianzhu;  Xu, Changsheng
浏览  |  Adobe PDF(2281Kb)  |  收藏  |  浏览/下载:535/146  |  提交时间:2018/02/07
Multimedia Storytelling  Video Analysis  Deep Learning  
Deep Structured Event Modeling for User-Generated Photos 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2018, 卷号: 20, 期号: 8, 页码: 2100-2113
作者:  Yang, Xiaoshan;  Zhang, Tianzhu;  Xu, Changsheng
浏览  |  Adobe PDF(1164Kb)  |  收藏  |  浏览/下载:390/88  |  提交时间:2018/01/03
Event Analysis  Unusual Event Detection  Deep Learning