World's Best Scientists 2026 revealed!

D-Index & Metrics

Computer Science

D-Index
32
Citations
3823
World Ranking
13242
National Ranking
1622

Best Publications

  • Perturbed Masking: Parameter-free Probing for Analyzing and Interpreting BERT

    Zhiyong Wu;Yun Chen;Ben Kao;Qun Liu

  • A Review of Deep Learning Based Speech Synthesis

    Unknown

  • Speech Emotion Recognition Using Capsule Networks

    Xixin Wu;Songxiang Liu;Yuewen Cao;Xu Li

  • FullSubNet+: Channel Attention Fullsubnet with Complex Spectrograms for Speech Enhancement

    Unknown

  • Emotion Recognition from Variable-Length Speech Segments Using Deep Learning on Spectrograms.

    Xi Ma;Zhiyong Wu;Jia Jia;Mingxing Xu

  • A deep recurrent approach for acoustic-to-articulatory inversion

    Peng Liu;Quanjie Yu;Zhiyong Wu;Shiyin Kang

  • Modality-specific and shared generative adversarial network for cross-modal retrieval

    Fei Wu;Xiao-Yuan Jing;Zhiyong Wu;Yimu Ji

  • Question detection from acoustic features using recurrent neural network with gated recurrent unit

    Yaodong Tang;Yuchen Huang;Zhiyong Wu;Helen Meng

  • Towards Multi-Scale Style Control for Expressive Speech Synthesis

    Unknown

  • Dilated Residual Network with Multi-head Self-attention for Speech Emotion Recognition

    Runnan Li;Zhiyong Wu;Jia Jia;Sheng Zhao

  • Emotion Controllable Speech Synthesis Using Emotion-Unlabeled Dataset with the Assistance of Cross-Domain Speech Emotion Recognition

    Xiong Cai;Dongyang Dai;Zhiyong Wu;Xiang Li

  • Speech-XLNet: Unsupervised Acoustic Model Pretraining For Self-Attention Networks

    Xingcheng Song;Guangsen Wang;Zhiyong Wu;Yiheng Huang

  • Real-time synthesis of Chinese visual speech and facial expressions using MPEG-4 FAP features in a three-dimensional avatar.

    Zhiyong Wu;Shen Zhang;Lianhong Cai;Helen M. Meng

  • PATA: Fuzzing with Path Aware Taint Analysis

    Unknown

  • Good for Misconceived Reasons: An Empirical Revisiting on the Need for Visual Context in Multimodal Machine Translation

    Zhiyong Wu;Lingpeng Kong;Wei Bi;Xiang Li

  • FullSubNet+: Channel Attention Fullsubnet with Complex Spectrograms for Speech Enhancement

    Unknown

  • Automatic lexical stress and pitch accent detection for L2 English speech using multi-distribution deep neural networks

    Kun Li;Shaoguang Mao;Xu Li;Zhiyong Wu

  • Adversarially learning disentangled speech representations for robust multi-factor voice conversion

    Unknown

  • Towards Discriminative Representation Learning for Speech Emotion Recognition.

    Runnan Li;Zhiyong Wu;Jia Jia;Yaohua Bu

  • Multi-level fusion of audio and visual features for speaker identification

    Zhiyong Wu;Lianhong Cai;Helen Meng

If you think any of the details on this page are incorrect, let us know.

Report an issue

We appreciate your kind effort to assist us to improve this page, it would be helpful providing us with as much detail as possible in the text box below: