CASIA OpenIR

浏览/检索结果: 共135条,第1-20条 帮助

限定条件    
已选(0)清除 条数/页:   排序方式:
Positive Unlabeled Fake News Detection via Multi-Modal Masked Transformer Network 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2024, 卷号: 26, 页码: 234-244
作者:  Wang, Jinguang;  Qian, Shengsheng;  Hu, Jun;  Hong, Richang
收藏  |  浏览/下载:6/0  |  提交时间:2024/03/26
Fake news detection  multi-modal learning  social media  
AnANet: Association and Alignment Network for Modeling Implicit Relevance in Cross-Modal Correlation Classification 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2023, 卷号: 25, 页码: 7867-7880
作者:  Xu, Nan;  Wang, Junyan;  Tian, Yuan;  Zhang, Ruike;  Mao, Wenji
收藏  |  浏览/下载:6/0  |  提交时间:2024/03/26
Association and alignment network  classification scheme  cross-modal correlation  implicit relevance  
Cross-Lingual Text Image Recognition via Multi-Hierarchy Cross-Modal Mimic 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2023, 卷号: 25, 页码: 4830-4841
作者:  Chen, Zhuo;  Yin, Fei;  Yang, Qing;  Liu, Cheng-Lin
收藏  |  浏览/下载:18/0  |  提交时间:2024/02/22
Cross-lingual text image recognition  cross-modal mimic  multihierarchy mimic  
Illumination Guided Attentive Wavelet Network for Low-Light Image Enhancement 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2023, 卷号: 25, 页码: 6258-6271
作者:  Xu, Jingzhao;  Yuan, Mengke;  Yan, Dong-Ming;  Wu, Tieru
收藏  |  浏览/下载:14/0  |  提交时间:2024/02/22
Lighting  Wavelet transforms  Image enhancement  Frequency modulation  Wavelet coefficients  Noise reduction  Discrete wavelet transforms  Attention mechanism  illumination guidance  low-light image enhancement  wavelet transform  
Quality-Aware Network for Human Parsing 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2023, 卷号: 25, 页码: 7128-7138
作者:  Yang, Lu;  Song, Qing;  Wang, Zhihui;  Liu, Zhiwei;  Xu, Songcen;  Li, Zhihao
收藏  |  浏览/下载:12/0  |  提交时间:2024/02/22
Computer vision  image segmentation  multi-media computing  
Adversarial Learning Guided Task Relatedness Refinement for Multi-Task Deep Learning 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2023, 卷号: 25, 页码: 6946-6957
作者:  Fang, Yuchun;  Cai, Sirui;  Cao, Yiting;  Li, Zhengchen;  Zhang, Zhaoxiang
收藏  |  浏览/下载:14/0  |  提交时间:2024/02/22
Index Terms-Multi-task learning  deep learning  task relatedness  
Robust Video-Text Retrieval Via Noisy Pair Calibration 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2023, 卷号: 25, 页码: 8632-8645
作者:  Zhang, Huaiwen;  Yang, Yang;  Qi, Fan;  Qian, Shengsheng;  Xu, Changsheng
收藏  |  浏览/下载:16/0  |  提交时间:2024/02/22
Noise calibration  uncertainty  video text retrieval  
SMNet: Synchronous Multi-Scale Low Light Enhancement Network With Local and Global Concern 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2023, 卷号: 25, 页码: 9506-9517
作者:  Lin, Shideng;  Tang, Fan;  Dong, Weiming;  Pan, Xingjia;  Xu, Changsheng
收藏  |  浏览/下载:17/0  |  提交时间:2024/02/21
Low-light image enhancement  multi-scale feature learning  deep-learning  
Dual Structural Knowledge Interaction for Domain Adaptation 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2023, 卷号: 25, 页码: 9057-9070
作者:  Zuo, Yukun;  Yao, Hantao;  Zhuang, Liansheng;  Xu, Changsheng
收藏  |  浏览/下载:13/0  |  提交时间:2024/02/21
Manifolds  Feature extraction  Adaptation models  Representation learning  Task analysis  Semisupervised learning  Pattern recognition  Domain adaptation  structural knowledge  dual structural knowledge interaction  
Zero-Shot Predicate Prediction for Scene Graph Parsing 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2023, 卷号: 25, 页码: 3140-3153
作者:  Li, Yiming;  Yang, Xiaoshan;  Huang, Xuhui;  Ma, Zhe;  Xu, Changsheng
收藏  |  浏览/下载:101/0  |  提交时间:2023/11/17
Deep learning  zero-shot  scene graph  
Human Parsing With Part-Aware Relation Modeling 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2022, 卷号: 25, 页码: 2601-2612
作者:  Zhang, Xiaomei;  Chen, Yingying;  Tang, Ming;  Wang, Jinqiao;  Zhu, Xiangyu;  Lei, Zhen
Adobe PDF(6053Kb)  |  收藏  |  浏览/下载:93/4  |  提交时间:2023/11/17
Human parsing  modeling  part-aware relation  
Seeing Through Darkness: Visual Localization at Night via Weakly Supervised Learning of Domain Invariant Features 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2023, 卷号: 25, 页码: 1713-1726
作者:  Fan, Bin;  Yang, Yuzhu;  Feng, Wensen;  Wu, Fuchao;  Lu, Jiwen;  Liu, Hongmin
收藏  |  浏览/下载:66/0  |  提交时间:2023/11/17
Domain invariant local features  image matching  long-term visual localization  weakly supervised learning  
ATF: An Alternating Training Framework for Weakly Supervised Face Alignment 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2023, 卷号: 25, 页码: 1798-1809
作者:  Lan, Xing;  Hu, Qinghao;  Cheng, Jian
收藏  |  浏览/下载:50/0  |  提交时间:2023/11/17
Face alignment  multi-task learning  weakly supervised  
A Two-Level Rectification Attention Network for Scene Text Recognition 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2023, 卷号: 25, 页码: 2404-2414
作者:  Wu, Lintai;  Xu, Yong;  Hou, Junhui;  Chen, C. L. Philip;  Liu, Cheng-Lin
收藏  |  浏览/下载:38/0  |  提交时间:2023/11/17
Scene text recognition  text rectification  spatial transformer network  optical character recognition  
Depth-Aware Multi-Person 3D Pose Estimation With Multi-Scale Waterfall Representations 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2023, 卷号: 25, 页码: 1439-1451
作者:  Shen, Tianyu;  Li, Deqi;  Wang, Fei-Yue;  Huang, Hua
收藏  |  浏览/下载:22/0  |  提交时间:2023/11/17
Three-dimensional displays  Pose estimation  Feature extraction  Location awareness  Cameras  Semantics  Solid modeling  Human depth perceiving  multi-person 3d pose estimation  multi-scale representation  occlusion handling  
Progressive Context-Aware Graph Feature Learning for Target Re-Identification 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2023, 卷号: 25, 页码: 1230-1242
作者:  Cao, Min;  Ding, Cong;  Chen, Chen;  Dou, Hao;  Hu, Xiyuan;  Yan, Junchi
收藏  |  浏览/下载:93/0  |  提交时间:2023/11/16
Target re-identification  graph convolutional network  feature learning  contextual information  graph feature learning  
Explicit Cross-Modal Representation Learning for Visual Commonsense Reasoning 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2022, 卷号: 24, 页码: 2986-2997
作者:  Zhang, Xi;  Zhang, Feifei;  Xu, Changsheng
收藏  |  浏览/下载:265/0  |  提交时间:2022/07/25
Cognition  Video recording  Syntactics  Visualization  Task analysis  Semantics  Linguistics  Visual Commonsense Reasoning  explicit reasoning  syntactic structure  interpretability  
Holographic Feature Learning of Egocentric-Exocentric Videos for Multi-Domain Action Recognition 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2022, 卷号: 24, 页码: 2273-2286
作者:  Huang, Yi;  Yang, Xiaoshan;  Gao, Junyun;  Xu, Changsheng
Adobe PDF(2409Kb)  |  收藏  |  浏览/下载:288/61  |  提交时间:2022/07/25
Videos  Feature extraction  Visualization  Task analysis  Computational modeling  Target recognition  Prototypes  Egocentric videos  exocentric videos  holographic feature  multi-domain  action recognition  
Instance GNN: A Learning Framework for Joint Symbol Segmentation and Recognition in Online Handwritten Diagrams 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2022, 卷号: 24, 页码: 2580-2594
作者:  Yun, Xiao-Long;  Zhang, Yan-Ming;  Yin, Fei;  Liu, Cheng-Lin
收藏  |  浏览/下载:198/0  |  提交时间:2022/07/25
Handwriting recognition  Task analysis  Grammar  Semantics  Image segmentation  Trajectory  Text recognition  Online handwritten diagram recognition  symbol segmentation  symbol recognition  freehand sketch analysis  graph neural networks  
The Model May Fit You: User-Generalized Cross-Modal Retrieval 期刊论文
IEEE TRANSACTIONS ON MULTIMEDIA, 2021, 卷号: 24, 页码: 2998-3012
作者:  Ma, Xinhong;  Yang, Xiaoshan;  Gao, Junyu;  Xu, Changsheng
Adobe PDF(6549Kb)  |  收藏  |  浏览/下载:227/46  |  提交时间:2022/06/17
cross-modal retrieval  domain generalization  meta-learning