CASIA OpenIR

浏览/检索结果: 共13条,第1-10条 帮助

限定条件    
已选(0)清除 条数/页:   排序方式:
Objformer: Boosting 3D object detection via instance-wise interaction 期刊论文
PATTERN RECOGNITION, 2024, 卷号: 146, 页码: 9
作者:  Tao, Manli;  Zhao, Chaoyang;  Tang, Ming;  Wang, Jinqiao
Adobe PDF(3261Kb)  |  收藏  |  浏览/下载:138/7  |  提交时间:2024/02/22
3D object detection  Point clouds  Incompletion and occlusion  Instance-wise interaction  
Transformer-based stroke relation encoding for online handwriting and sketches 期刊论文
PATTERN RECOGNITION, 2024, 卷号: 148, 页码: 13
作者:  Liu, Jing-Yu;  Zhang, Yan-Ming;  Yin, Fei;  Liu, Cheng-Lin
Adobe PDF(1999Kb)  |  收藏  |  浏览/下载:83/1  |  提交时间:2024/02/22
Online stroke classification  Handwritten document analysis  Diagram recognition  Sketch semantic segmentation  Position encoding in transformer  
Reparameterizing and dynamically quantizing image features for image generation 期刊论文
PATTERN RECOGNITION, 2024, 卷号: 146, 页码: 11
作者:  Sun, Mingzhen;  Wang, Weining;  Zhu, Xinxin;  Liu, Jing
Adobe PDF(3612Kb)  |  收藏  |  浏览/下载:163/22  |  提交时间:2023/12/21
Vector quantization  Variational auto-encoder  Unconditional image generation  Text-to-image generation  Autoregressive generation  
Table Structure Recognition and Form Parsing by End-to-End Object Detection and Relation Parsing 期刊论文
PATTERN RECOGNITION, 2022, 卷号: 132, 页码: 14
作者:  Li, Xiao-Hui;  Yin, Fei;  Dai, He-Sen;  Liu, Cheng-Lin
Adobe PDF(4030Kb)  |  收藏  |  浏览/下载:271/1  |  提交时间:2022/11/14
Table detection  Table structure recognition  Template -free form parsing  Graph neural network  End -to -end training  
End -to -end video text detection with online tracking 期刊论文
PATTERN RECOGNITION, 2021, 卷号: 113, 页码: 12
作者:  Yu, Hongyuan;  Huang, Yan;  Pi, Lihong;  Zhang, Chengquan;  Li, Xuan;  Wang, Liang
Adobe PDF(4997Kb)  |  收藏  |  浏览/下载:357/70  |  提交时间:2021/05/06
End-to-end  Video text detection  Online tracking  
Re-ranking Image-text Matching by Adaptive Metric Fusion 期刊论文
PATTERN RECOGNITION, 2020, 卷号: 104, 期号: 1, 页码: 13
作者:  Niu, Kai;  Huang, Yan;  Wang, Liang
浏览  |  Adobe PDF(2236Kb)  |  收藏  |  浏览/下载:489/112  |  提交时间:2020/06/22
Image-text matching  Re-ranking method  Adaptive metric fusion  
Long video question answering: A Matching-guided Attention Model 期刊论文
PATTERN RECOGNITION, 2020, 卷号: 102, 期号: 1, 页码: 11
作者:  Wang, Weining;  Huang, Yan;  Wang, Liang
浏览  |  Adobe PDF(1963Kb)  |  收藏  |  浏览/下载:390/77  |  提交时间:2020/06/02
Long video QA  Matching-guided attention  
Dense semantic embedding network for image captioning 期刊论文
PATTERN RECOGNITION, 2019, 卷号: 90, 页码: 285-296
作者:  Xiao, Xinyu;  Wang, Lingfeng;  Ding, Kun;  Xiang, Shiming;  Pan, Chunhong
收藏  |  浏览/下载:390/0  |  提交时间:2019/04/23
Image captioning  Retrieval  High-level semantic information  Visual concept  Densely embedding  Long short-term memory  
Improving visual question answering using dropout and enhanced question encoder 期刊论文
PATTERN RECOGNITION, 2019, 卷号: 90, 期号: 1, 页码: 404-414
作者:  Fang, Zhiwei;  Liu, Jing;  Li, Yong;  Qiao, Yanyuan;  Lu, Hanqing
浏览  |  Adobe PDF(1624Kb)  |  收藏  |  浏览/下载:496/130  |  提交时间:2019/04/23
Visual question answering  Coherent dropout  Siamese dropout  Enhanced question encoder  
MAPNet: Multi-modal attentive pooling network for RGB-D indoor scene classification 期刊论文
PATTERN RECOGNITION, 2019, 期号: 90, 页码: 436-449
作者:  Li, Yabei;  Zhang, Zhang;  Cheng, Yanhua;  Wang, Liang;  Tan, Tieniu
浏览  |  Adobe PDF(5797Kb)  |  收藏  |  浏览/下载:419/44  |  提交时间:2019/04/23
Indoor scene classification  Multi-modal fusion  RGB-D  Attentive pooling