CASIA OpenIR
(本次检索基于用户作品认领结果)

浏览/检索结果: 共267条,第1-10条 帮助

限定条件                        
已选(0)清除 条数/页:   排序方式:
Spatial reconstructed local attention Res2Net with F0 subband for fake speech detection 期刊论文
NEURAL NETWORKS, 2024, 卷号: 175, 页码: 11
作者:  Fan, Cunhang;  Xue, Jun;  Tao, Jianhua;  Yi, Jiangyan;  Wang, Chenglong;  Zheng, Chengshi;  Lv, Zhao
收藏  |  浏览/下载:21/0  |  提交时间:2024/07/04
ASVspoof  Fake speech detection  Fundamental frequency  Res2Net  
SceneFake: An initial dataset and benchmarks for scene fake audio detection 期刊论文
PATTERN RECOGNITION, 2024, 卷号: 152, 页码: 12
作者:  Yi, Jiangyan;  Wang, Chenglong;  Tao, Jianhua;  Zhang, Chu Yuan;  Fan, Cunhang;  Tian, Zhengkun;  Ma, Haoxin;  Fu, Ruibo
收藏  |  浏览/下载:16/0  |  提交时间:2024/07/04
Scene manipulation  Fake audio detection  Speech enhancement  SceneFake dateset  
WavDepressionNet: Automatic Depression Level Prediction via Raw Speech Signals 期刊论文
IEEE TRANSACTIONS ON AFFECTIVE COMPUTING, 2024, 卷号: 15, 期号: 1, 页码: 285-296
作者:  Niu, Mingyue;  Tao, Jianhua;  Li, Yongwei;  Qin, Yong;  Li, Ya
收藏  |  浏览/下载:5/0  |  提交时间:2024/07/03
Assessment block  depression level prediction  representation block  speech signals  WavDepressionNet  
Emotion selectable end-to-end text-based speech editing 期刊论文
ARTIFICIAL INTELLIGENCE, 2024, 卷号: 329, 页码: 16
作者:  Wang, Tao;  Yi, Jiangyan;  Fu, Ruibo;  Tao, Jianhua;  Wen, Zhengqi;  Zhang, Chu Yuan
收藏  |  浏览/下载:11/0  |  提交时间:2024/07/03
Emotion selectable  Text-based speech editing  Emotion decoupling  Mask prediction  Few-shot learning  Text-to-speech  
Distinguishing Neural Speech Synthesis Models Through Fingerprints in Speech Waveforms 会议论文
, Taiyuan, Shanxi, China, 2024-07-27
作者:  Zhang, Chu Yuan;  Yi, Jiangyan;  Tao, Jianhua;  Wang, Chenglong;  Yan, Xinrui
Adobe PDF(2254Kb)  |  收藏  |  浏览/下载:25/12  |  提交时间:2024/06/26
Multi-Scale Permutation Entropy for Audio Deepfake Detection 会议论文
, 韩国首尔, 2024-4-14
作者:  Chenglong Wang;  He JY(何佳毅);  Jiangyan Yi;  Jianhua Tao;  Chu Yuan Zhang;  Xiaohui Zhang
Adobe PDF(997Kb)  |  收藏  |  浏览/下载:51/17  |  提交时间:2024/06/13
End-to-End Network Based on Transformer for Automatic Detection of Covid-19 会议论文
, Singapore, 22-27 May 2022
作者:  Cong Cai;  Bin Liu;  Jianhua Tao;  Zhengkun Tian;  Jiahao Lu;  Kexin Wang
Adobe PDF(1210Kb)  |  收藏  |  浏览/下载:47/12  |  提交时间:2024/06/11
GCC-Speaker: Target Speaker Localization with Optimal Speaker-Dependent Weighting in Multi-Speaker Scenarios 会议论文
, 希腊罗得岛, 2023年6月
作者:  Li GJ(李冠君);  Liu WJ(刘文举);  Yi JY(易江燕);  Tao JH(陶建华)
Adobe PDF(3463Kb)  |  收藏  |  浏览/下载:36/12  |  提交时间:2024/06/06
MULTIMODAL CROSS- AND SELF-ATTENTION NETWORK FOR SPEECH EMOTION RECOGNITION 会议论文
, Toronto, Canada, 6-12 June 2021
作者:  Licai Sun;  Bin Liu;  Jianhua Tao;  Zheng Lian
Adobe PDF(1078Kb)  |  收藏  |  浏览/下载:34/8  |  提交时间:2024/06/03
EmotionNAS: Two-stream Neural Architecture Search for Speech Emotion Recognition 会议论文
, Dublin, Ireland, 20-24 August 2023
作者:  Haiyang Sun;  Zheng Lian;  Bin Liu;  Ying Li;  Licai Sun;  Cong Cai;  Jianhua Tao;  Meng Wang;  Yuan Cheng
Adobe PDF(826Kb)  |  收藏  |  浏览/下载:39/10  |  提交时间:2024/05/31